Skip to content
Topic

#Reinforcement Learning

12 articles on Reinforcement Learning — news, releases, guides and analysis from the SourceFeed engine.

Fine-Tuning Can't Teach What Pretraining Never Saw
Article 1w ago 0

Fine-Tuning Can't Teach What Pretraining Never Saw

A 5B model raised strictly on grade-school text shrugs off every trick meant to push it further.

Rachel Goldstein
Pretraining Sets a Ceiling Post-Training Can't Break

Pretraining Sets a Ceiling Post-Training Can't Break

Article · 2w ago3
That 100x-Cheaper Retrieval Claim Is Half Right

That 100x-Cheaper Retrieval Claim Is Half Right

Article · 3w ago0
The $500 Fine-Tune Is Real, but the Eval Is the Moat

The $500 Fine-Tune Is Real, but the Eval Is the Moat

Article · 1mo ago1
A $500 RL Fine-Tune Beat the Frontier. Sort Of.

A $500 RL Fine-Tune Beat the Frontier. Sort Of.

Article · 1mo ago1
Why Every LLM Vendor Killed the Thinking-Token Budget

Why Every LLM Vendor Killed the Thinking-Token Budget

Article · 1mo ago0
Why a Mario JEPA Predicts Well and Plans Poorly

Why a Mario JEPA Predicts Well and Plans Poorly

Article · 1mo ago1
Nested RL Agents That Write Real Training Jobs

Nested RL Agents That Write Real Training Jobs

Article · 1mo ago5
SWE-1.7 and the Myth of the Post-Training Ceiling

SWE-1.7 and the Myth of the Post-Training Ceiling

Article · 1mo ago5
Ornith-1.0: Coding Models That Train Their Own Agent Scaffolds

Ornith-1.0: Coding Models That Train Their Own Agent Scaffolds

Article · 2mos ago0
Simulating the World Inside the LLM

Simulating the World Inside the LLM

Article · 2mos ago2
When a 3B Model Out-Reasons Opus 4.5, Read the Fine Print

When a 3B Model Out-Reasons Opus 4.5, Read the Fine Print

Article · 2mos ago3