#post-training
Post-Training: 6 AI articles covering post-training news, analysis, and research
Articles
Fireworks AI Releases Ember-1: A Post-Trained Kimi K3 That UsesNEW⭐8
Fireworks AI's Ember-1 post-trains Kimi K3 to cut reasoning tokens by ~40% while preserving accuracy, available only via serverless API research preview.
Perplexity Trains Its Computer Agent on Real Mistakes With Hint⭐9
Perplexity trains its computer agent on real mistakes using hint-guided self-distillation, cutting tool-call failures 21.2% in live testing.
Cognition Releases SWE-2: A Kimi K3 Post-Trained Coding Model⭐7
Cognition's SWE-2 coding model, post-trained from Kimi K3, hits 50.0% on FrontierCode 1.1, matching Fable 5.1 at 64% lower cost and running only inside Devin.
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better at⭐8
Z.ai's GLM-5.3 boosts coding and cybersecurity via post-training, no retraining. Scores surge on complex benchmarks; open weights release in two weeks.
SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned⭐7
SpaceXAI unveils Grok 4.6 with 500K context, boosting agentic coding and knowledge work with improved scores and new xhigh reasoning tier.
AllenAI Open Instruct Tulu 3: Building a Compact Post-Training⭐9
Learn to build a compact post-training pipeline with AllenAI's Tulu 3: SFT, DPO, and RLVR using GRPO in a single-GPU environment.
