AZ Labs

News & Insights

The latest in artificial intelligence

Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

23 results for AI Research + Capability: tools

Filtered from the full AZ Labs news archive.

Clear filters
Official DeepSeek V4.1 Flash launch graphic introducing the multimodal architecture
AI Research
headsetAudio10 Sept 2026

DeepSeek releases V4.1 Flash with native multimodal vision, 552B MoE and lower API rates

DeepSeek has launched DeepSeek-V4.1-Flash, a 552B-parameter mixture-of-experts model featuring a novel Causal Encoder–Decoder architecture with just 8B active input and 16B active output parameters. The release brings native vision, compresses KV cache storage by up to 8x, and slashes off-peak API pricing to $0.15 per million input tokens.

smart_toyDeepSeek1.0M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$0.15/M
6 min read6 takeawaysReadarrow_forward
OpenAI GPT-6 Astra official release artwork displaying flagship capabilities
AI Research
04 Sept 2026

OpenAI Releases GPT-6 Astra: Next-Generation Flagship with 1.05M Context and Deep Multimodal Reasoning

OpenAI has officially launched GPT-6 Astra, its frontier flagship model featuring a 1,050,000-token context window, 128,000 max output tokens, and native tool-use for autonomous agent workflows.

smart_toyOpenAI1.1M ctx
IN:descriptionText+visibilityVision+attach_fileDocs
OUT:chatText
visionimage understanding$10/M
7 min read5 takeawaysReadarrow_forward
Google artwork for Gemini 3.8 Flash and Gemini 3.8 Flash Cyber
AI Research
headsetAudio02 Sept 2026

Google releases Gemini 3.8 Flash for long-horizon coding

Google has released Gemini 3.8 Flash as a generally available model for long-horizon coding, autonomous agents and enterprise workflows. The API keeps the introductory price of Gemini 3.7 Flash, with a one-million-token input limit and 65,536-token output limit.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
5 min read4 takeawaysReadarrow_forward
Tencent Hy4 preview announcement artwork with the Hy4 model name
AI Research
headsetAudio28 Aug 2026

Tencent releases Hy4 preview with 770B parameters and 1M context

Tencent has released and open-sourced Hy4 preview, a 770-billion-parameter mixture-of-experts model with 49 billion active parameters and a context window above one million tokens. AIMI saw it reach OpenRouter and OpenCode Go later the same day.

smart_toyTencent1M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$0.8/M
5 min read4 takeawaysReadarrow_forward
Qwen3.8 Flash-Next announcement artwork from the official Qwen research page
AI Research
headsetAudio26 Aug 2026

Qwen3.8 Flash Reaches OpenRouter as Qwen Publishes Flash-Next

OpenRouter has added Qwen3.8 Flash, the production model Qwen says is served through QwenCloud with a one-million-token context and built-in tools. Qwen published the related Flash-Next open-weight preview on the same day.

smart_toyQwen1.0M ctx
IN:descriptionText+visibilityVision+videocamVideo
OUT:chatText
visionimage understanding$0.1/M
4 min read3 takeawaysReadarrow_forward
AZ Labs GLM-5.3 release and provider-route timeline, with later route expansion tracked in the article
AI Research
headsetAudio26 Aug 2026

GLM-5.3 Flash Reaches Four Provider Routes

AIMI saw GLM-5.3 Flash reach OpenCode Go, OpenRouter, Cloudflare Workers AI and Ollama Cloud on 26 August. The route wave expands access to Z.ai's model, but Z.ai has not announced a separate Flash launch.

smart_toyZ.ai200K ctx
IN:descriptionText+visibilityVision+videocamVideo
OUT:chatText
visionimage understandingboltFree
5 min read4 takeawaysReadarrow_forward
OpenRouter artwork for the Qwen3.8-27B model route
AI Research
headsetAudio14 Aug 2026

Qwen3.8-27B Ships Open Weights with 262K Context

Qwen has released Qwen3.8-27B under Apache 2.0. The dense vision-language model has a native 262,144-token context, adjustable reasoning, image and video input, and a live OpenRouter route while Qwen Cloud hosting remains pending.

smart_toyQwen262K ctx
IN:descriptionText
OUT:chatText
reasoningtools$0.18/M
4 min read3 takeawaysReadarrow_forward
AZ Labs artwork showing the dots3-note preview name and its core model specifications
AI Research
headsetAudio14 Aug 2026

Dots3-Note Preview Opens a 280B Multimodal Agent Model

Dots Studio has released dots3-note preview, its first open-weight dots3 model. The multimodal MoE has 280 billion total parameters, 16 billion active parameters, a 512K context window, and an Apache 2.0 release.

smart_toyRedNote524K ctx
IN:descriptionText+visibilityVision+videocamVideo+micAudio
OUT:chatText
visionimage understandingboltFree
4 min read3 takeawaysReadarrow_forward
Official Google DeepMind artwork for Gemini 3.7 Flash
AI Research
headsetAudio13 Aug 2026

Gemini 3.7 Flash Is Now Generally Available

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

smart_toyGoogle1.0M ctx
IN:descriptionText+visibilityVision+micAudio+videocamVideo+attach_fileDocs
OUT:chatText
visionimage understandingboltFree
4 min read3 takeawaysReadarrow_forward
Grok 4.6 announcement artwork from the xAI news page
AI Research
headsetAudio12 Aug 2026

Grok 4.6: xAI Ships a Long-Running Agent Model to OpenRouter

xAI announced Grok 4.6 on 12 August 2026, tuned for long-running agents and ambitious visual work. It is live on OpenRouter with a 500,000-token context at $2 per million input tokens, and the release date is confirmed by the official announcement.

smart_toyxAI2M ctx
IN:descriptionText+visibilityVision+attach_fileDocs
OUT:chatText
visionimage understanding$2/M
4 min read3 takeawaysReadarrow_forward
ByteDance Seed 2.1 Turbo listing artwork from OpenRouter
AI Research
headsetAudio12 Aug 2026

ByteDance Seed 2.1 Turbo Reaches OpenRouter After Its June Release

ByteDance released Seed 2.1 Turbo on 23 June 2026. Its OpenRouter route reached AZ Labs monitoring on 12 August with text, image and video input, a 262,144-token context, and pricing from $0.50 per million input tokens.

smart_toyByteDance262K ctx
IN:descriptionText+visibilityVision+videocamVideo
OUT:chatText
reasoningtools$0.2/M
3 min read3 takeawaysReadarrow_forward
Meta Muse Glimmer 30B model artwork from its Hugging Face model card
AI Research
headsetAudio10 Aug 2026

Meta Muse Glimmer 30B: An Open-Weight Agentic Model Built to Run Locally

Meta Superintelligence Labs has released Muse Glimmer 30B, an Apache 2.0 open-weight model distilled from Muse Spark and tuned for local agent workflows on consumer GPUs. It covers tool use, long-context reasoning and multimodal input in a package that fits under 20 GB with 4-bit quantisation.

smart_toyMeta131K ctx
IN:descriptionText+visibilityVision
OUT:chatText
reasoningtoolsboltFree
6 min read3 takeawaysReadarrow_forward
Official Google artwork introducing Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber
AI Research
21 Jul 2026

Google Releases Gemini 3.6 Flash: High-Efficiency Multimodal Intelligence for Fast Agent Loops

Google launched Gemini 3.6 Flash on 21 July 2026 alongside 3.5 Flash-Lite and 3.5 Flash Cyber, cutting output token usage by 17% against 3.5 Flash at $0.75 per million input tokens.

smart_toyGoogle DeepMind1.0M ctx
IN:descriptionText+visibilityVision+videocamVideo+attach_fileDocs+micAudio
OUT:chatText
visionimage understanding$0.75/M
6 min read4 takeawaysReadarrow_forward
OpenAI GPT-5 announcement artwork showing the GPT-5 flagship model
AI Research
headsetAudio05 Feb 2026

OpenAI Releases GPT-5 with Enhanced Reasoning Capabilities

OpenAI has unveiled GPT-5, its most advanced language model to date, featuring breakthrough reasoning capabilities that bring AI closer to human-level problem solving.

smart_toyOpenAI1.1M ctx
IN:descriptionText+visibilityVision
OUT:chatText
visionimage understanding$2.5/M
6 min read3 takeawaysReadarrow_forward
Anthropic illustration for Claude 4 showing Claude balancing multiple tasks
AI Research
headsetAudio25 Jan 2026

Anthropic's Claude 4 Sets New Benchmarks

Anthropic's latest model, Claude 4, pushes the frontier of responsible AI development with state-of-the-art performance on safety and helpfulness benchmarks.

smart_toyAnthropic200K ctx
IN:descriptionText+visibilityVision+attach_fileDocs
OUT:chatText
visionimage understanding$3/M
5 min read3 takeawaysReadarrow_forward