Videos list
Thumb Title Channel Status Published
Speech-to-Speech Model Research at Google DeepMind — Valeria Wu Fon & Tom Ouyang, Google DeepMind

Google DeepMind is building a single native speech-to-speech model for Gemini that handles streaming translation,...

AI Engineer summarized 2026-09-15 13:00
Agents Without Code: Skills, YAML, and Filesystems Replaced Python — Philipp Schmid, Google DeepMind

Agent orchestration code is being replaced by files: markdown instructions, YAML configurations, and bash skills are...

AI Engineer summarized 2026-09-14 17:30
SOTA Generative Media Panel — Dumitru Erhan, Shane Gu & Nicole Brichtova, Google DeepMind

Google DeepMind launched Gemini Omni Flash APIs for video generation and editing, along with NanoBanana 2 Light, a...

AI Engineer summarized 2026-08-30 14:00

Frontier News · by Hyperjump Technology