Videos
3 total
| Thumb | Title | Channel | Status | Published |
|---|---|---|---|---|
|
|
Speech-to-Speech Model Research at Google DeepMind — Valeria Wu Fon & Tom Ouyang, Google DeepMind
Google DeepMind is building a single native speech-to-speech model for Gemini that handles streaming translation,... |
AI Engineer | summarized | 2026-09-15 13:00 |
|
|
Agents Without Code: Skills, YAML, and Filesystems Replaced Python — Philipp Schmid, Google DeepMind
Agent orchestration code is being replaced by files: markdown instructions, YAML configurations, and bash skills are... |
AI Engineer | summarized | 2026-09-14 17:30 |
|
|
SOTA Generative Media Panel — Dumitru Erhan, Shane Gu & Nicole Brichtova, Google DeepMind
Google DeepMind launched Gemini Omni Flash APIs for video generation and editing, along with NanoBanana 2 Light, a... |
AI Engineer | summarized | 2026-08-30 14:00 |
Frontier News · by Hyperjump Technology