Article // Intelligence
TheSequence · Sep 22, 2026
The Sequence Knowledge - Issue 937: RSI in Post-Training: The Loop That Already Shipped
Inside the RSI post-training pipeline
What's interesting
Engagement
Likes: 12
Comments: 0
Restacks: 1
Entities
Sponsorship
No sponsorship detected
Related articles
13 theses on agentic AI and regulation
Dan Davies - "Back of Mind" · Frontier labs
I keep hearing the same advice about agents
Gradient Flow · model evaluation
TAI #220: The Next Models Will Change How We Work…Again! Take AI Agent Swarms Seriously
Towards AI Newsletter · model evaluation
Ornith-1.5, LFM2.5 QAD, and a 10 GB Qwen3.8 27B
The Kaitchup – AI on a Budget · model evaluation
Anthropic Looks At Some Of Its Alignment Problems
Don't Worry About the Vase · model evaluation
Jev introduces a new shape of LLM - System One, aka Decision Models
Simon Willison’s Newsletter · model evaluation
Everything You Need to Prepare for a Hugging Face Interview
AI Engineering Insider · model evaluation
🗞️ OpenAI stops reinforcement learning training for 2 weeks after Astra model reached “Critical” cybersecurity capabilities.
Rohan's Bytes · model evaluation