Google launches agentic video understanding in Gemini

Google's agentic video understanding is now available across Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite, cutting token consumption by up to 88% and analysis costs by up to 66% while improving accuracy by up to 7%. It's available via the Gemini API in Google AI Studio and the Gemini Enterprise Agent Platform.
How this story unfolded
same day · 2 reports · 4 community posts · 6 of 7 shown
More stories today
US AI data center boom faces hidden China supply chain risks
US AI data centers depend on China-linked supplies of transformers, batteries, and optical gear. Washington is moving to tighten restrictions on foreign equipment.
CNBC Technology·2 hours ago

China's new micro-drama rules bring AI-generated shows under regulation
China's first micro-drama regulation took effect September 1, requiring AI-generated micro-dramas to carry AI disclosure labels in every episode and banning recommendation algorithms that encourage addiction. In H1 2026, AI-generated productions accounted for over 74% of 367,000 micro-dramas released.
TechNode·2 hours ago

Tesla Cybercabs spotted across Austin ahead of Sept 3 launch
Tesla Cybercabs are now operating across downtown Austin with no steering wheel, pedals, or human driver, two years after a closed-track demo. The fleet is expanding ahead of a September 3rd launch.
Vaibhav Sisinty·3 hours ago
Google DeepMind by email
Get an email when Google DeepMind has news
No news that day, no email.
Training a coding model to paint watercolours with TRL and OpenEnv
Sergio Paniego's open-source recipe trains Qwen3.5-35B-A3B via GRPO to write p5.brush JavaScript, reproducing Surya Narreddi's viral watercolour model. All artifacts—environment, hand-rated pool, 3 runs, and paintings—are published on Hugging Face.
Sergio Paniego Blog·3 hours ago

Fréchet Audio Distance explained with math and code
Video deep-dive explains FAD, a key objective metric for evaluating generative audio models like Suno, building intuition from math to code.
YouTube·3 hours ago
Hermes Agent sheds 375,000 lines in 15-hour autonomous repo cleanup
Teknium·4 hours agoAto: $99 AI device for seniors who can't use smartphones
Vaibhav Sisinty·5 hours ago
Reddit user shares rule of thumb for choosing LLMs
A Reddit user in r/LocalLLaMA shares a personal heuristic for selecting models, noting that with Qwen 27B, a task that would take 15 hours of active programming without an LLM takes only 4 hours. The post is anecdotal and community-focused.
r/LocalLLaMA·5 hours ago