Videos
| Thumb | Title | Channel | Status | Published |
|---|---|---|---|---|
|
|
Scale Your API to 1 million requests Per Second
The gap between benchmark throughput (3M req/s on a single box) and real-world API performance is due to five... |
Cloud Codes | summarized | 2026-08-21 16:30 |
|
|
The Next Medium: Why Real-Time Interactive Video Changes Everything — Ahmed Ahres, Reactor
Reactor argues that real-time interactive video generation, which it calls world models, will fundamentally change... |
AI Engineer | summarized | 2026-08-18 17:30 |
|
|
Infra behind Krea 2: How to train and serve at scale — Gabriel Jorge Menezes, Krea.ai
Krea.ai trained K2, a diffusion transformer from scratch on thousands of GPUs, and built infrastructure to... |
AI Engineer | summarized | 2026-08-18 17:00 |
|
|
MCP Goes Stateless | John Dellenbaugh & Pankaj Kumar | MCP Release Party - Seattle
The Model Context Protocol (MCP) has gone stateless, removing the need for sticky sessions and session stores when... |
MLOps.community | summarized | 2026-08-17 22:26 |
|
|
xAI's Real Plan to Win Isn't Grok 4.6
Grok 4.6 ties OpenAI's top model on a key benchmark at a fifth of the output price, but the model is the least... |
Cloud Codes | summarized | 2026-08-14 04:05 |
|
|
Ex-Uber dev explains his Multi-Agent Workflow
The future of AI is multiplayer: agents as full team members with Slack accounts, email addresses, and shared... |
David Ondrej | summarized | 2026-08-10 21:12 |
|
|
Building Turbopuffer: Gergely Orosz (@pragmaticengineer ) × Simon Eskildsen (CEO)
Simon Eskildsen, CEO of Turbopuffer, shares his journey from a self-taught programmer at Shopify to building a... |
AI Engineer | summarized | 2026-08-03 18:26 |
|
|
Serving 2 Million Models Without Melting: Scaling the Hugging Face Hub — Arek Borucki, Hugging Face
Hugging Face scaled its infrastructure to serve 3 million models and 14 million users by decoupling metadata storage... |
AI Engineer | summarized | 2026-07-28 13:41 |
|
|
Agents are slower than LLMs?
Agents are slower than LLMs primarily because they rely on external tool calls (e.g., fetching web pages, making API... |
Caleb Writes Code | summarized | 2026-07-07 21:50 |
|
|
Agents & the $40M Bet on Multiplayer AI
Dust is building a multiplayer AI platform where humans and agents collaborate around shared state called 'pods',... |
MLOps.community | summarized | 2026-06-15 14:00 |
Frontier News · by Hyperjump Technology