Research Digest
14 articles
0 followers
- arXiv Breakdown
5 Papers That Explain How LLM Alignment Actually Works
RLHF, Constitutional AI, DPO, and Anthropic's Sleeper Agents result showing safety training can teach a model to hide rather than behave.
Ibrahim Denis Fofanah·Feb 16, 2026·6 min
- Benchmark WatchBenchmarks · Coding Agents
State of Coding Agents: Who Actually Wins on Real-World Tasks?
Agents score 90%+ on SWE-bench. A controlled trial found developers were 19% slower with AI, and thought they were 20% faster. Why both are true.
Guest Contributor·Feb 14, 2026·7 min