Archive
9 papers · newest first-
A rounding slip in fast attention code quietly spoiled late trainingarXiv:2609.34272
Broken Symmetry in BF16 Attention: Why FlashAttention Gradients Blow Up Late in Training
-
A coding AI improved by training only on its own post-mortemsarXiv:2609.35741
Shockingly Simple Self-retrospection Improves Agentic Models Without RL
-
Penalized by an AI monitor, models fooled it with readable reasoningarXiv:2609.31121
Monitor Jailbreaking: Evading Chain-of-Thought Monitoring Without Encoded Reasoning
-
Coding agents wrote robot programs that beat hand-built planners in simulationarXiv:2609.30233
Coding Agents for Generalized Task and Motion Planning Problems
-
AI research agents often exploit loose scoring on open-ended test tasksarXiv:2609.28614
Reward Hacking Challenges Oversight of Autonomous Research Agents
-
Tiny models never shown real data got better at predicting itarXiv:2609.30063
Self-Play Pretraining with Zero Data
-
Dropping a third of an AI’s chosen experts barely dented scoresarXiv:2609.25809
You Only Need 2/3 of the Chosen Experts: An Empirical Study of Dynamic Expert Pruning in Fine-Grained MoE LLMs
-
Ranking a model’s parts wins a benchmark for tracing its answersarXiv:2609.25518
Matryoshka attribution: Learning to attribute language model outputs to representations and weights
-
Note-sharing AI agents match far bigger solo crowds on puzzlesarXiv:2609.21032
Scaling Discovery through Test-Time Communication
daily · 06:00 UTC · free
The one AI paper worth reading today, checked and explained.
Unsubscribe any time. Or follow the RSS feed.