AZ Labs

News & Insights

The latest in artificial intelligence

Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

9 results for Capability: reasoning + Agents

Filtered from the full AZ Labs news archive.

Clear filters
Official DeepSeek V4.1 Flash launch graphic introducing the multimodal architecture
AI Research
headsetAudio10 Sept 2026

DeepSeek releases V4.1 Flash with native multimodal vision, 552B MoE and lower API rates

DeepSeek has launched DeepSeek-V4.1-Flash, a 552B-parameter mixture-of-experts model featuring a novel Causal Encoder–Decoder architecture with just 8B active input and 16B active output parameters. The release brings native vision, compresses KV cache storage by up to 8x, and slashes off-peak API pricing to $0.15 per million input tokens.

smart_toyDeepSeek1.0M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$0.15/M
6 min read6 takeawaysReadarrow_forward
Google artwork for Gemini 3.8 Flash and Gemini 3.8 Flash Cyber
AI Research
headsetAudio02 Sept 2026

Google releases Gemini 3.8 Flash for long-horizon coding

Google has released Gemini 3.8 Flash as a generally available model for long-horizon coding, autonomous agents and enterprise workflows. The API keeps the introductory price of Gemini 3.7 Flash, with a one-million-token input limit and 65,536-token output limit.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
5 min read4 takeawaysReadarrow_forward
Tencent Hy4 preview announcement artwork with the Hy4 model name
AI Research
headsetAudio28 Aug 2026

Tencent releases Hy4 preview with 770B parameters and 1M context

Tencent has released and open-sourced Hy4 preview, a 770-billion-parameter mixture-of-experts model with 49 billion active parameters and a context window above one million tokens. AIMI saw it reach OpenRouter and OpenCode Go later the same day.

smart_toyTencent1M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$0.8/M
5 min read4 takeawaysReadarrow_forward
AZ Labs artwork showing the dots3-note preview name and its core model specifications
AI Research
headsetAudio14 Aug 2026

Dots3-Note Preview Opens a 280B Multimodal Agent Model

Dots Studio has released dots3-note preview, its first open-weight dots3 model. The multimodal MoE has 280 billion total parameters, 16 billion active parameters, a 512K context window, and an Apache 2.0 release.

smart_toyRedNote524K ctx
IN:descriptionText+visibilityVision+videocamVideo+micAudio
OUT:chatText
visionimage understandingboltFree
4 min read3 takeawaysReadarrow_forward
Official Google DeepMind artwork for Gemini 3.7 Flash
AI Research
headsetAudio13 Aug 2026

Gemini 3.7 Flash Is Now Generally Available

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
4 min read3 takeawaysReadarrow_forward
Grok 4.6 announcement artwork from the xAI news page
AI Research
headsetAudio12 Aug 2026

Grok 4.6: xAI Ships a Long-Running Agent Model to OpenRouter

xAI announced Grok 4.6 on 12 August 2026, tuned for long-running agents and ambitious visual work. It is live on OpenRouter with a 500,000-token context at $2 per million input tokens, and the release date is confirmed by the official announcement.

smart_toyxAI2M ctx
IN:descriptionText+visibilityVision+attach_fileDocs
OUT:chatText
visionimage understanding$2/M
4 min read3 takeawaysReadarrow_forward