This model is a fine-tuned version of fpadovani/eng-latn-10mb-ppt-Dp-100mb-packedseed3407. It has been trained using TRL. This model was trained with SFT.
Open-weight model · Text generation
eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407
by Francesca Padovani fpadovani/eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407
This model is a fine-tuned version of fpadovani/eng-latn-10mb-ppt-Dp-10mb-packedseed3407. It has been trained using TRL. This model was trained with SFT.
Runs On
What it takes to serve eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407 (39M parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.
| Precision | Weights | Memory needed | Cheapest setup | Per hour | Also fits |
|---|---|---|---|---|---|
| 16-bit | 0.1 GB | 0.1 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
| 8-bit | 0.0 GB | 0.0 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
| 4-bit | 0.0 GB | 0.0 GB | 1x MI300X (192 GB) Vultr |
$1.85 | 1x H100 $1.99 · 1x MI325X $2.00 |
Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 18, 2026.
Model Card
This model is a fine-tuned version of fpadovani/eng-latn-10mb-ppt-Dp-10mb-packedseed3407. It has been trained using TRL. This model was trained with SFT.
Excerpt from the card by Francesca Padovani.
Configuration
- Architecture
- GPT2LMHeadModel
- Vocabulary size
- 51,200
- Model type
- gpt2
Identity and Version
- Repository
- fpadovani/eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407
- Publisher
- Francesca Padovani
- Task
- Text generation
- Modality
- Text
- Library
- transformers
- Parameters
- 39M parameters
- Languages
- trl, sft
- Revision
- 2fc701c1a636945a98288c11f8da15cdaf833f8b
- First published
- 2026-09-18
- Last updated
- 2026-09-18
Files and Weights
200 files, 1.6 GB in total. The weights are 59 files totalling 1.6 GB in bin, pth, safetensors.
Every file
| File | Type | Size | SHA-256 |
|---|---|---|---|
| checkpoint-1000/model.safetensors | Weights | 78.2 MB | f8f6c0c52a10 |
| checkpoint-1000/rng_state.pth | Weights | 14.6 KB | c23194f8722b |
| checkpoint-1000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-1500/model.safetensors | Weights | 78.2 MB | e64cb994ce3f |
| checkpoint-1500/rng_state.pth | Weights | 14.6 KB | 7d326e95f81b |
| checkpoint-1500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-2000/model.safetensors | Weights | 78.2 MB | d105bf862e35 |
| checkpoint-2000/rng_state.pth | Weights | 14.6 KB | bcfae038aeee |
| checkpoint-2000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-2500/model.safetensors | Weights | 78.2 MB | 75cd5acfdf3b |
| checkpoint-2500/rng_state.pth | Weights | 14.6 KB | a8598ff184c9 |
| checkpoint-2500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-3000/model.safetensors | Weights | 78.2 MB | 08e35a535c61 |
| checkpoint-3000/rng_state.pth | Weights | 14.6 KB | 3097834e9193 |
| checkpoint-3000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-3500/model.safetensors | Weights | 78.2 MB | 71d8d58f6487 |
| checkpoint-3500/rng_state.pth | Weights | 14.6 KB | 46cda71e4728 |
| checkpoint-3500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-4000/model.safetensors | Weights | 78.2 MB | a5975ffd0eed |
| checkpoint-4000/rng_state.pth | Weights | 14.6 KB | 2c28e3c1f356 |
| checkpoint-4000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-4500/model.safetensors | Weights | 78.2 MB | a0a98b2474a2 |
| checkpoint-4500/rng_state.pth | Weights | 14.6 KB | 6839f735c672 |
| checkpoint-4500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-500/model.safetensors | Weights | 78.2 MB | 1527086775db |
| checkpoint-500/rng_state.pth | Weights | 14.6 KB | c5d03ea2e051 |
| checkpoint-500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-5000/model.safetensors | Weights | 78.2 MB | 5fde81dfb2b4 |
| checkpoint-5000/rng_state.pth | Weights | 14.6 KB | d3cfa223222d |
| checkpoint-5000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-5500/model.safetensors | Weights | 78.2 MB | 9a13cbc1529b |
| checkpoint-5500/rng_state.pth | Weights | 14.6 KB | 69de6c56f00f |
| checkpoint-5500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-6000/model.safetensors | Weights | 78.2 MB | 3752127f5deb |
| checkpoint-6000/rng_state.pth | Weights | 14.6 KB | d357116556b1 |
| checkpoint-6000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-6500/model.safetensors | Weights | 78.2 MB | ec076a20de6a |
| checkpoint-6500/rng_state.pth | Weights | 14.6 KB | 3091256c4ef4 |
| checkpoint-6500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-7000/model.safetensors | Weights | 78.2 MB | 2b99c3a44eec |
| checkpoint-7000/rng_state.pth | Weights | 14.6 KB | 61c4a81b46ac |
| checkpoint-7000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-7500/model.safetensors | Weights | 78.2 MB | 1e5de7b71402 |
| checkpoint-7500/rng_state.pth | Weights | 14.6 KB | d329c3269910 |
| checkpoint-7500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-8000/model.safetensors | Weights | 78.2 MB | 410f2d663756 |
| checkpoint-8000/rng_state.pth | Weights | 14.6 KB | 7b61de60f287 |
| checkpoint-8000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-8500/model.safetensors | Weights | 78.2 MB | 88e46536b7fe |
| checkpoint-8500/rng_state.pth | Weights | 14.6 KB | ffbd1461dd36 |
| checkpoint-8500/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-9000/model.safetensors | Weights | 78.2 MB | 8ffbaf413494 |
| checkpoint-9000/rng_state.pth | Weights | 14.6 KB | e5a5105a17e6 |
| checkpoint-9000/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| checkpoint-9240/model.safetensors | Weights | 78.2 MB | e9fd91aa2769 |
| checkpoint-9240/rng_state.pth | Weights | 14.6 KB | 2ab4a7185b7d |
| checkpoint-9240/training_args.bin | Weights | 6.4 KB | f130d354c699 |
| model.safetensors | Weights | 78.2 MB | e9fd91aa2769 |
| training_args.bin | Weights | 6.4 KB | f130d354c699 |
| added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-1000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-1000/config.json | Configuration | 801 B | — |
| checkpoint-1000/generation_config.json | Configuration | 154 B | — |
| checkpoint-1000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-1000/trainer_state.json | Configuration | 55.5 KB | — |
| checkpoint-1500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-1500/config.json | Configuration | 801 B | — |
| checkpoint-1500/generation_config.json | Configuration | 154 B | — |
| checkpoint-1500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-1500/trainer_state.json | Configuration | 83.9 KB | — |
| checkpoint-2000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-2000/config.json | Configuration | 801 B | — |
| checkpoint-2000/generation_config.json | Configuration | 154 B | — |
| checkpoint-2000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-2000/trainer_state.json | Configuration | 112.3 KB | — |
| checkpoint-2500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-2500/config.json | Configuration | 801 B | — |
| checkpoint-2500/generation_config.json | Configuration | 154 B | — |
| checkpoint-2500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-2500/trainer_state.json | Configuration | 140.7 KB | — |
| checkpoint-3000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-3000/config.json | Configuration | 801 B | — |
| checkpoint-3000/generation_config.json | Configuration | 154 B | — |
| checkpoint-3000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-3000/trainer_state.json | Configuration | 169.1 KB | — |
| checkpoint-3500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-3500/config.json | Configuration | 801 B | — |
| checkpoint-3500/generation_config.json | Configuration | 154 B | — |
| checkpoint-3500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-3500/trainer_state.json | Configuration | 197.4 KB | — |
| checkpoint-4000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-4000/config.json | Configuration | 801 B | — |
| checkpoint-4000/generation_config.json | Configuration | 154 B | — |
| checkpoint-4000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-4000/trainer_state.json | Configuration | 225.8 KB | — |
| checkpoint-4500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-4500/config.json | Configuration | 801 B | — |
| checkpoint-4500/generation_config.json | Configuration | 154 B | — |
| checkpoint-4500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-4500/trainer_state.json | Configuration | 254.2 KB | — |
| checkpoint-500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-500/config.json | Configuration | 801 B | — |
| checkpoint-500/generation_config.json | Configuration | 154 B | — |
| checkpoint-500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-500/trainer_state.json | Configuration | 28.2 KB | — |
| checkpoint-5000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-5000/config.json | Configuration | 801 B | — |
| checkpoint-5000/generation_config.json | Configuration | 154 B | — |
| checkpoint-5000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-5000/trainer_state.json | Configuration | 282.6 KB | — |
| checkpoint-5500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-5500/config.json | Configuration | 801 B | — |
| checkpoint-5500/generation_config.json | Configuration | 154 B | — |
| checkpoint-5500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-5500/trainer_state.json | Configuration | 311.1 KB | — |
| checkpoint-6000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-6000/config.json | Configuration | 801 B | — |
| checkpoint-6000/generation_config.json | Configuration | 154 B | — |
| checkpoint-6000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-6000/trainer_state.json | Configuration | 339.5 KB | — |
| checkpoint-6500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-6500/config.json | Configuration | 801 B | — |
| checkpoint-6500/generation_config.json | Configuration | 154 B | — |
| checkpoint-6500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-6500/trainer_state.json | Configuration | 367.9 KB | — |
| checkpoint-7000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-7000/config.json | Configuration | 801 B | — |
| checkpoint-7000/generation_config.json | Configuration | 154 B | — |
| checkpoint-7000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-7000/trainer_state.json | Configuration | 396.4 KB | — |
| checkpoint-7500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-7500/config.json | Configuration | 801 B | — |
| checkpoint-7500/generation_config.json | Configuration | 154 B | — |
| checkpoint-7500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-7500/trainer_state.json | Configuration | 424.8 KB | — |
| checkpoint-8000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-8000/config.json | Configuration | 801 B | — |
| checkpoint-8000/generation_config.json | Configuration | 154 B | — |
| checkpoint-8000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-8000/trainer_state.json | Configuration | 453.1 KB | — |
| checkpoint-8500/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-8500/config.json | Configuration | 801 B | — |
| checkpoint-8500/generation_config.json | Configuration | 154 B | — |
| checkpoint-8500/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-8500/trainer_state.json | Configuration | 481.5 KB | — |
| checkpoint-9000/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-9000/config.json | Configuration | 801 B | — |
| checkpoint-9000/generation_config.json | Configuration | 154 B | — |
| checkpoint-9000/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-9000/trainer_state.json | Configuration | 509.9 KB | — |
| checkpoint-9240/added_tokens.json | Configuration | 27.7 KB | — |
| checkpoint-9240/config.json | Configuration | 801 B | — |
| checkpoint-9240/generation_config.json | Configuration | 154 B | — |
| checkpoint-9240/special_tokens_map.json | Configuration | 22.6 KB | — |
| checkpoint-9240/trainer_state.json | Configuration | 523.3 KB | — |
| config.json | Configuration | 801 B | — |
| generation_config.json | Configuration | 154 B | — |
| special_tokens_map.json | Configuration | 22.6 KB | — |
| README.md | Documentation | 1.9 KB | — |
| .gitattributes | Repository | 1.5 KB | — |
| checkpoint-1000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-1000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-1500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-1500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-2000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-2000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-2500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-2500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-3000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-3000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-3500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-3500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-4000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-4000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-4500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-4500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-5000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-5000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-5500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-5500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-6000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-6000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-6500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-6500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-7000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-7000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-7500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-7500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-8000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-8000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-8500/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-8500/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-9000/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-9000/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| checkpoint-9240/spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| checkpoint-9240/tokenizer_config.json | Tokenizer | 233.6 KB | — |
| spiece.model | Tokenizer | 1.1 MB | f0cc7d2d7e69 |
| tokenizer_config.json | Tokenizer | 233.6 KB | — |
License and Download
- License
- Not stated by the source
- Access
- Open weights, no gate
- Download size
- 1.6 GB
Released by Francesca Padovani through its official repository on Hugging Face.
Built From
- Derived from fpadovani/eng-latn-10mb-ppt-Dp-10mb-packed_seed3407
Memory Requirements
| Precision | Weights in memory |
|---|---|
| As published | 1.6 GB |
| 16-bit | 0.1 GB |
| 8-bit | 0.0 GB |
| 4-bit | 0.0 GB |
Weights only, from the published parameter count; the key-value cache and runtime add to this.
Questions About eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407
How much GPU memory does eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407 need?
About 0.1 GB at 16-bit and 0 GB at 4-bit: the weights (39M parameters) plus a working margin. A long context needs more.
What is the cheapest GPU to run eng-latn-10mb-after-ppt-Dp-10mb-packed-ckpt500_seed3407 on?
At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.
Similar Models
SAGI is a novel causal language model that integrates swarm intelligence dynamics with transformer architecture. The model treats cognition as a dynamic, adaptive system where multiple internal "agents" collaborate through differentiable routing, trust mechanisms, and shared memory. - Episodic + Semantic Memory: Dual memory system with trainable retrieval utility The swarm processes observations derived from token embeddings, updating its internal state S. This state conditions the transformer's attention patterns and feed-forward activations via learned projections, creating bidirectional information flow between symbolic (tokens) and subsymbolic (swarm dynamics) processing. - Educational…
A 69M parameter causal language model built on the Mixture-of-Attentions (MoA) architecture — distance-based metric attention that respects the triangle inequality by construction, not approximation. Every attention head operates in a proper metric space. The geometry is enforced, not hoped for. Standard transformers compute attention as a dot product: Q·Kᵀ. This has no geometric meaning — it's a bilinear form, not a distance. Two tokens can be "close" by dot product while violating basic metric properties. MoA replaces this with negative squared distance under a learned diagonal Mahalanobis metric, then enforces the triangle inequality through a regularizer over random triples sampled…
A 70M parameter causal language model built on the Mixture-of-Attentions (MoA) architecture — distance-based metric attention that respects the triangle inequality by construction, not approximation. Every attention head operates in a proper metric space. The geometry is enforced, not hoped for. Standard transformers compute attention as a dot product: Q·Kᵀ. This has no geometric meaning — it's a bilinear form, not a distance. Two tokens can be "close" by dot product while violating basic metric properties. MoA replaces this with negative squared distance under a learned diagonal Mahalanobis metric, then enforces the triangle inequality through a regularizer over random triples sampled…
Aloha! Today, we are releasing Ornith-1.0, a self-improving family of open-source models for agentic coding. This model card documents Ornith-1.0-9B, the most lightweight member of the Ornith family, designed for efficient single-GPU deployment. Ornith-1.0-9B is a dense ~9B model (≈19 GB in bf16), so it serves comfortably on a single 80GB GPU. The recipes below stand up an OpenAI-compatible server; add --tensor-parallel-size / --tp if you want to shard across more GPUs. For a quick local test (or to script offline generation), load the model directly with Transformers. Make sure you have a recent release installed — see the Transformers installation guide; Ornith-1.0-9B requires…
Aloha! Today, we are releasing Ornith-1.0, a self-improving family of open-source models for agentic coding. This model card documents Ornith-1.0-35B, the lightweight member of the Ornith family, designed for efficient single-GPU deployment. The two recipes below stand up an OpenAI-compatible server on a single 8×80GB GPU node (tensor-parallel 8). Adjust --tensor-parallel-size / --tp to the number of GPUs you have. For a quick local test (or to script offline generation), load the model directly with Transformers. Make sure you have a recent release installed — see the Transformers installation guide; Ornith-1.0-35B requires transformers >= 5.8.1. To split the reasoning trace from the final…