Videos
| Thumb | Title | Channel | Status | Published |
|---|---|---|---|---|
|
|
Turn Hermes Agent Into Your Chief of Staff In 18 Mins
Hermes Agent's Quicksilver update introduces smart approvals, durable background jobs, delivery ledgers, profile... |
Jack Roberts | summarized | 2026-07-29 18:45 |
|
|
Skills are new features: Building Skill-Centric Harness — Yogendra Miraje, FactSet
Skills are the new features in agentic products, shifting the role of engineers from shipping features to building... |
AI Engineer | summarized | 2026-07-29 18:00 |
|
|
This letter could change EVERYTHING
A debate is intensifying between proponents of open-source AI, led by Nvidia's Jensen Huang, and closed-source... |
Matthew Berman | summarized | 2026-07-29 17:52 |
|
|
Morgan Stanley's ALPHALAB: Multi-Agent Research Across Optimization Domains — Brendan Rappazzo
Morgan Stanley's AlphaLab is an agentic harness for automating quantitative research, using a multi-agent system... |
AI Engineer | summarized | 2026-07-29 17:06 |
|
|
Your Agent Didn't Fail. Your Harness Did. — Vinoth Govindarajan, OpenAI
Most production agent failures are not model failures but harness failures—the system that owns state, orders... |
AI Engineer | summarized | 2026-07-29 16:00 |
|
|
Paste This Into Claude, Never Get A Generic Response Again
Generic responses from AI can be eliminated by providing five layers of context: voice, knowledge, collaborative,... |
Austin Marchese | summarized | 2026-07-29 13:45 |
|
|
How Forward Deployed Engineering is done at Factory — Eno Reyes
Factory uses deployed engineers as the tip of the spear to help enterprise customers build autonomous software... |
AI Engineer | summarized | 2026-07-28 22:00 |
|
|
Steal This AI Follow-Up System Every Business Needs (But Nobody Has)
A business follow-up system that captures call conversations, drafts personalized messages using Claude, and designs... |
AI Founders | summarized | 2026-07-28 21:48 |
|
|
AI tools for Forward Deployed Engineering — Vasuman Moza, Varick Agents
The next bottleneck in AI adoption is not execution but understanding and re-engineering business processes around... |
AI Engineer | summarized | 2026-07-28 21:00 |
|
|
How Forward Deployed Engineering is done at Cognition — Jia Wu
Forward deployed engineers at Cognition maximize the overlap between their AI coding agent Devin and enterprise... |
AI Engineer | summarized | 2026-07-28 20:00 |
|
|
How Forward Deployed Engineering is done at Ramp — Leo Mehr
Forward Deployed Engineering at Ramp focuses on winning upmarket by making core product and agentic features work... |
AI Engineer | summarized | 2026-07-28 19:00 |
|
|
The Dirty Secret of Forward Deployed Engineering — Natalie Meurer, Sierra
Forward deployed engineering (FDE) is an ill-defined but increasingly hot role in AI, evolving from Palantir's... |
AI Engineer | summarized | 2026-07-28 18:00 |
|
|
ChatGPT Voice 2.0 Just Dropped, and…
OpenAI's ChatGPT Voice 2.0 introduces a real-time conversational agent that can multitask, control desktop apps, and... |
Jack Roberts | summarized | 2026-07-28 17:49 |
|
|
How Forward Deployed Engineering is done at Decagon — Sunny Rekhi
Decagon's forward deployed engineering is identical to product engineering, with engineers acting as both executors... |
AI Engineer | summarized | 2026-07-28 17:00 |
|
|
How Forward Deployed Engineering is done at Kepler — Vinoo Ganesh
Forward Deployed Engineering (FDE) is a product strategy, not a go-to-market role, where engineers embed with... |
AI Engineer | summarized | 2026-07-28 16:00 |
|
|
OpenAI’s Plan to Make ChatGPT the Everything App — Akshay Nathan, OpenAI
OpenAI's core product engineering lead Akshay Nathan discusses the launch of ChatGPT Work as the company's strategy... |
Latent Space | summarized | 2026-07-28 14:47 |
|
|
Serving 2 Million Models Without Melting: Scaling the Hugging Face Hub — Arek Borucki, Hugging Face
Hugging Face scaled its infrastructure to serve 3 million models and 14 million users by decoupling metadata storage... |
AI Engineer | summarized | 2026-07-28 13:41 |
|
|
Ling 3.0 Flash First Test – A Surprisingly GOOD Coding Model!
Ling 3.0 Flash from Ant is a surprisingly competent coding model at 124B total/5.1B active parameters, outperforming... |
Bijan Bowen | summarized | 2026-07-28 13:18 |
|
|
Forward Deployed Engineering 101 — Kevin Bai, Anthropic
Forward Deployed Engineering (FDE) is a go-to-market model for selling complex technical platforms to non-technical... |
AI Engineer | summarized | 2026-07-28 09:04 |
|
|
AI Agents for Performance: Ship Faster, Pay Less — Rajat Shah, Netflix
Rajat Shah from Netflix presents a playbook for using AI agents to automate performance engineering, reducing the... |
AI Engineer | summarized | 2026-07-28 00:59 |
|
|
How I Tricked AI Into Leaking Your Darkest Secrets
The Memory Heist attack exploits AI agents that have web fetch and access to private information by embedding... |
Goda Go | summarized | 2026-07-27 22:17 |
|
|
Kimi K3 Builds $20,000 Websites in 19 Mins
Kimi K3 is the current top AI model for front-end design, capable of building stunning, interactive websites with 3D... |
Jack Roberts | summarized | 2026-07-27 19:02 |
|
|
Agentic Engineering, explained by a 10x developer
Forston Ball, founding engineer at AMP, argues that the era of hand-written code is ending and that software... |
David Ondrej | summarized | 2026-07-27 15:22 |
|
|
DeepSWE: A Contamination-Resistant Coding Benchmark — James Shi, Datacurve
DeepSWE is a contamination-resistant coding benchmark from Datacurve, composed of 113 original software engineering... |
AI Engineer | summarized | 2026-07-26 18:10 |
|
|
State of Data — Sean Cai, Independent / State of Data
Data markets are undergoing a fundamental shift from type two (contrived) to type one (real workflow capture) data,... |
AI Engineer | summarized | 2026-07-26 17:00 |
|
|
Claude Opus 5: CoWork Just Got Cheaper & More Powerful!
Claude Opus 5 is now available at half the price of Fable while outperforming it on knowledge work, business... |
Systems Made Better | summarized | 2026-07-26 16:07 |
|
|
ChatGPT Voice 2.0 + Codex is INSANE (endless possibilities)
OpenAI's new ChatGPT Voice 2.0 mode allows real-time interruptible conversation, parallel task execution, and... |
Brock Mesarich | AI for Non Techies | summarized | 2026-07-26 16:03 |
|
|
You're Using 10% of Claude. This Only Takes 13 Minutes To Learn
Most Claude users only utilize about 10% of its capabilities because they haven't configured the built-in settings.... |
AI Founders | summarized | 2026-07-26 16:01 |
|
|
This AI Technology Will Replace Millions (Here's How to Prepare)
Agentic AI, exemplified by tools like Claude Code, is poised to automate a significant portion of desk jobs by... |
Nate Herk | summarized | 2026-07-26 14:52 |
|
|
The Messy Reality of Scale: Synthetic Data and Pre-Training — Marah Abdin & Robert McHardy, poolside
Poolside's Marah Abdin and Robert McHardy share the messy reality of scaling large language models, focusing on... |
AI Engineer | summarized | 2026-07-26 01:00 |
|
|
Evals-Driven Development for a Mental Health AI Coach — Akele Reed & Dave Revere, SonderMind
SonderMind's AI coach Sonder uses evals-driven development with modular input/output guardrails powered by separate... |
AI Engineer | summarized | 2026-07-25 23:00 |
|
|
Loop Engineering from First Principles — Kyle Mistele, HumanLayer
Building effective AI coding loops for real-world, team-based software requires applying control theory... |
AI Engineer | summarized | 2026-07-25 20:41 |
|
|
Why Large? Tiny LMs & Agents on Edge/Robotics — Cormac Brick, Google
Tiny models (50M-500M parameters) are now viable for edge devices and robotics, enabling voice-to-function calling... |
AI Engineer | summarized | 2026-07-25 17:00 |
|
|
The 5 BIGGEST Lies You've Been Told About Claude
The video exposes five common lies about Claude and AI productivity, arguing that staying updated with every new... |
Austin Marchese | summarized | 2026-07-25 15:15 |
|
|
Claude Opus 5 Builds $15,000+ Animated Websites (FOR CHEAP)
Claude Opus 5 is half the price of Fable 5 and can generate stunning animated websites using the Higsfield MCP and a... |
Brock Mesarich | AI for Non Techies | summarized | 2026-07-25 14:25 |
|
|
From Agent Traces to Agent Simulations — Rustem Feyzkhanov, Snorkel AI
Every company needs a private benchmark to reliably evaluate, release, and improve AI agents, moving beyond... |
AI Engineer | summarized | 2026-07-25 01:00 |
|
|
Claude Opus 5 Is INSANE – Is This the BEST Model Yet?
Claude Opus 5 delivers exceptional performance in creative coding tests, outperforming previous models and even... |
Bijan Bowen | summarized | 2026-07-25 00:23 |
|
|
Evaling Video Slop — Maor Bril, Character.ai
Evaluating AI-generated video quality is harder than generating the video itself. Character.ai built a fast, small... |
AI Engineer | summarized | 2026-07-25 00:00 |
|
|
I Tested Opus 5 vs. Fable 5. What You Need to Know.
Claude Opus 5 is often cheaper than Fable 5 and can outperform it on coding and verification tasks, but Fable 5... |
Nate Herk | summarized | 2026-07-24 23:38 |
|
|
Building Closed-Loop Evals for a Multimodal Agent at Scale — Soumya Gupta & Jai Chopra, Uber
Uber's computer vision team built a closed-loop evaluation system for a multimodal agent that enhances food photos... |
AI Engineer | summarized | 2026-07-24 22:00 |
|
|
Opus 5 is FINALLY here! (WOAH)
Anthropic's Claude Opus 5 was released, outperforming the larger Fable 5 on most benchmarks while costing about half... |
Matthew Berman | summarized | 2026-07-24 21:47 |
|
|
Model Whisperers How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, Google
Building reliable agents requires a systematic evaluation approach that starts with intuitive checks and scales to... |
AI Engineer | summarized | 2026-07-24 21:00 |
|
|
From Signal to PR: Anatomy of a Self-Improving Agent — Jason Lopatecki, Arize
Arize AI's Signal agent transforms observability data into automated pull requests by combining telemetry traces,... |
AI Engineer | summarized | 2026-07-24 20:15 |
|
|
The Future of Evals: From LLM as a Judge to Agent as a Judge — Aparna Dhinakaran, Arize AI
Evals for AI agents need to evolve from static LLM-as-a-judge to dynamic agent-as-a-judge because traditional evals... |
AI Engineer | summarized | 2026-07-24 20:00 |
|
|
OPUS 5 CLICK NOW
Jensen Huang and other AI leaders signed a letter supporting open-weight AI models, arguing open source drives... |
Matthew Berman | summarized | 2026-07-24 17:56 |
|
|
Claude Opus 5 is Going to Save You Money
Claude Opus 5 has been released with state-of-the-art performance on coding and knowledge work benchmarks, often... |
Nate Herk | summarized | 2026-07-24 17:45 |
|
|
Everything Is a Rollout — Alex Shaw + Ryan Marten, Terminal-Bench, Harbor, Laude Institute
Agent development is more like machine learning than traditional software engineering, requiring empirical... |
AI Engineer | summarized | 2026-07-24 16:00 |
|
|
Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs
Andon Labs created Vending-Bench, a long-horizon evaluation benchmark where AI agents run a simulated vending... |
AI Engineer | summarized | 2026-07-24 15:00 |
|
|
Full Workshop: Setting Yourself Up for Success —Jason Liu, OpenAI Codex | AI Engineer | no captions | 2026-07-24 15:00 |
|
|
Is Kimi K3 Really That Good?! (Don't Just Believe The Hype)
Kimi K3 is the most powerful open-weight model released, but it suffers from reliability issues that public... |
Cole Medin | summarized | 2026-07-24 14:00 |
Frontier News · by Hyperjump Technology