mid-training
Here are 8 public repositories matching this topic...
Revisiting Mid-training in the Era of Reinforcement Learning Scaling
-
Updated
Jul 23, 2025 - Jupyter Notebook
[𝗜𝗖𝗠𝗟 𝟮𝟬𝟮𝟲] Dispersion loss counteracts embedding condensation and improves generalization in small language models
-
Updated
Jun 21, 2026 - Python
Official code, models, and dataset for "Evolution Fine-Tuning (EFT): Learning to Discover Across 371 Optimization Tasks"
-
Updated
Jun 30, 2026 - Python
How Post-Training Shapes Biological Reasoning Models
-
Updated
Jun 16, 2026 - Jupyter Notebook
Open catalog of datasets used to train and align LLMs across pretraining, mid-training, and post-training.
-
Updated
Jan 6, 2026 - Python
Experiments with agentic mid-training and agentic capabilities
-
Updated
Feb 2, 2026 - Python
Unofficial reproduction of Nemotron-CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training (NeurIPS 2025, arXiv:2504.13161) — search-found mixtures beat uniform baselines +0.014–0.031 STEM at d28; novel finding: selection-mechanism winner's curse; Ascend NPU backend.
-
Updated
Sep 24, 2026 - Python
Add this topic to your repo
To associate your repository with the mid-training topic, visit your repo's landing page and select "manage topics."