[태그:] Reinforcement Learning
The Tiny Model That Embarrassed Frontier AI Labs
A 27-billion-parameter agent outperformed Claude Opus 4.8 and GPT-5.5 at replicating scientific papers. See how a 12-person London lab pulled it off.
AlphaGo’s Creator Just Made AI’s Boldest Bet Yet
David Silver, the mind behind AlphaGo, just closed Europe’s largest-ever seed at $1.1B. Find out why Sequoia and Nvidia bet on no human data.