A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.
-
Updated
Sep 4, 2026 - Python
A high-performance framework for training LLMs, VLMs, diffusion, and embodied models on NVIDIA GPUs and Kunlun XPUs.
Revisiting Mid-training in the Era of Reinforcement Learning Scaling
[๐๐๐ ๐ ๐ฎ๐ฌ๐ฎ๐ฒ] Dispersion loss counteracts embedding condensation and improves generalization in small language models
Official code, models, and dataset for "Evolution Fine-Tuning (EFT): Learning to Discover Across 371 Optimization Tasks"
How Post-Training Shapes Biological Reasoning Models
Open catalog of datasets used to train and align LLMs across pretraining, mid-training, and post-training.
To associate your repository with the mid-training topic, visit your repo's landing page and select "manage topics."