Videos
| Thumb | Title | Channel | Status | Published |
|---|---|---|---|---|
|
|
Taking Reinforcement Learning Cross Datacenter — Nan Jiang, Modal
Reinforcement learning post-training can be decoupled from a single, tightly-coupled GPU cluster by exploiting the... |
AI Engineer | summarized | 2026-08-10 17:30 |
|
|
Teaching AI to Find Real Vulnerabilities — David Brumley, Bugcrowd
Teaching AI to hack effectively requires designing reinforcement learning environments with deterministic grading... |
AI Engineer | summarized | 2026-08-01 00:30 |
|
|
Everything Is a Rollout — Alex Shaw + Ryan Marten, Terminal-Bench, Harbor, Laude Institute
Agent development is more like machine learning than traditional software engineering, requiring empirical... |
AI Engineer | summarized | 2026-07-24 16:00 |
|
|
Poolside’s Model Factory, Laguna S, Open Models, and the Race to AGI — Eiso Kant, Poolside AI
Poolside's co-founder Eiso Kant argues that model building is primarily an engineering discipline, and that their... |
Latent Space | summarized | 2026-07-22 20:46 |
|
|
🔬 RL with Verifiable Rewards, but the Verifier is a Lab — Lila Sciences
Lila Sciences is building AI science factories that treat the physical lab as a verifier for reinforcement learning,... |
Latent Space | summarized | 2026-07-16 13:30 |
|
|
Recursive Model Improvement — Lee Robinson, Cursor, SpaceXAI
Cursor trains AI models for code generation using a recursive self-improvement loop. The outer loop gathers user... |
AI Engineer | summarized | 2026-07-15 20:13 |
|
|
Cooking with OpenAI’s Research Chief: AGI, o1, Evals, and Scaling Laws — Mark Chen
OpenAI's Chief Research Officer Mark Chen discusses the enduring power of scaling laws, the importance of... |
Latent Space | summarized | 2026-06-25 21:32 |
|
|
Qwen-AgentWorld The World Model for RL Environments
Qwen-AgentWorld introduces a world model that hallucinates environments to train agents more effectively than... |
Sam Witteveen | summarized | 2026-06-25 13:00 |
|
|
VibeThinker 3B - Taking on Giant Models
VibeThinker 3B, a small model from Weibo AI Lab, beats giant models like Gemini 3 Pro and Claude Opus on math and... |
Sam Witteveen | summarized | 2026-06-19 13:15 |
Frontier News · by Hyperjump Technology