AZ Labs

News & Insights

The latest in artificial intelligence

Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

8 results for AI Research + Capability: video-understanding

Filtered from the full AZ Labs news archive.

Clear filters
Google artwork for Gemini 3.8 Flash and Gemini 3.8 Flash Cyber
AI Research
headsetAudio02 Sept 2026

Google releases Gemini 3.8 Flash for long-horizon coding

Google has released Gemini 3.8 Flash as a generally available model for long-horizon coding, autonomous agents and enterprise workflows. The API keeps the introductory price of Gemini 3.7 Flash, with a one-million-token input limit and 65,536-token output limit.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
5 min read4 takeawaysReadarrow_forward
Qwen3.8 Flash-Next announcement artwork from the official Qwen research page
AI Research
headsetAudio26 Aug 2026

Qwen3.8 Flash Reaches OpenRouter as Qwen Publishes Flash-Next

OpenRouter has added Qwen3.8 Flash, the production model Qwen says is served through QwenCloud with a one-million-token context and built-in tools. Qwen published the related Flash-Next open-weight preview on the same day.

smart_toyQwen1.0M ctx
IN:descriptionText+visibilityVision+videocamVideo
OUT:chatText
visionimage understanding$0.1/M
4 min read3 takeawaysReadarrow_forward
AZ Labs GLM-5.3 release and provider-route timeline, with later route expansion tracked in the article
AI Research
headsetAudio26 Aug 2026

GLM-5.3 Flash Reaches Four Provider Routes

AIMI saw GLM-5.3 Flash reach OpenCode Go, OpenRouter, Cloudflare Workers AI and Ollama Cloud on 26 August. The route wave expands access to Z.ai's model, but Z.ai has not announced a separate Flash launch.

smart_toyZ.ai200K ctx
IN:descriptionText+visibilityVision+videocamVideo
OUT:chatText
visionimage understandingboltFree
5 min read4 takeawaysReadarrow_forward
Official Google DeepMind artwork for Gemini 3.7 Flash
AI Research
headsetAudio13 Aug 2026

Gemini 3.7 Flash Is Now Generally Available

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
4 min read3 takeawaysReadarrow_forward
Official Google artwork introducing Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber
AI Research
21 Jul 2026

Google Releases Gemini 3.6 Flash: High-Efficiency Multimodal Intelligence for Fast Agent Loops

Google launched Gemini 3.6 Flash on 21 July 2026 alongside 3.5 Flash-Lite and 3.5 Flash Cyber, cutting output token usage by 17% against 3.5 Flash at $0.75 per million input tokens.

smart_toyGoogle DeepMind1.0M ctx
IN:descriptionText+visibilityVision+videocamVideo+attach_fileDocs+micAudio
OUT:chatText
visionimage understanding$0.75/M
6 min read4 takeawaysReadarrow_forward