AZ Labs

News & Insights

The latest in artificial intelligence

Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

17 results for Model release

Filtered from the full AZ Labs news archive.

Clear filters
OpenAI GPT-5 announcement artwork showing the GPT-5 flagship model
AI Research05 Feb 2026

OpenAI Releases GPT-5 with Enhanced Reasoning Capabilities

OpenAI has unveiled GPT-5, its most advanced language model to date, featuring breakthrough reasoning capabilities that bring AI closer to human-level problem solving.

6 min read3 takeaways
Read articlearrow_forward
Tencent Hy4 preview announcement artwork with the Hy4 model name
AI Research28 Aug 2026

Tencent releases Hy4 preview with 770B parameters and 1M context

Tencent has released and open-sourced Hy4 preview, a 770-billion-parameter mixture-of-experts model with 49 billion active parameters and a context window above one million tokens. AIMI saw it reach OpenRouter and OpenCode Go later the same day.

5 min read4 takeaways
Read articlearrow_forward
Qwen3.8 Flash-Next announcement artwork from the official Qwen research page
AI Research26 Aug 2026

Qwen3.8 Flash Reaches OpenRouter as Qwen Publishes Flash-Next

OpenRouter has added Qwen3.8 Flash, the production model Qwen says is served through QwenCloud with a one-million-token context and built-in tools. Qwen published the related Flash-Next open-weight preview on the same day.

4 min read3 takeaways
Read articlearrow_forward
OpenRouter artwork for the Qwen3.8-27B model route
AI Research14 Aug 2026

Qwen3.8-27B Ships Open Weights with 262K Context

Qwen has released Qwen3.8-27B under Apache 2.0. The dense vision-language model has a native 262,144-token context, adjustable reasoning, image and video input, and a live OpenRouter route while Qwen Cloud hosting remains pending.

4 min read3 takeaways
Read articlearrow_forward
AZ Labs GLM-5.3 release and provider-route timeline, with later route expansion tracked in the article
AI Research26 Aug 2026

GLM-5.3 Flash Reaches Four Provider Routes

AIMI saw GLM-5.3 Flash reach OpenCode Go, OpenRouter, Cloudflare Workers AI and Ollama Cloud on 26 August. The route wave expands access to Z.ai's model, but Z.ai has not announced a separate Flash launch.

5 min read4 takeaways
Read articlearrow_forward
AZ Labs artwork showing the dots3-note preview name and its core model specifications
AI Research14 Aug 2026

Dots3-Note Preview Opens a 280B Multimodal Agent Model

Dots Studio has released dots3-note preview, its first open-weight dots3 model. The multimodal MoE has 280 billion total parameters, 16 billion active parameters, a 512K context window, and an Apache 2.0 release.

4 min read3 takeaways
Read articlearrow_forward
DeepSeek API change log showing the V4 Flash Vision Exp release dated 21 August 2026
AI Research21 Aug 2026

DeepSeek Releases V4 Flash Vision Exp for Multimodal Agent Work

DeepSeek has released V4 Flash Vision Exp, an experimental API model that adds image understanding to the V4 Flash family while keeping its text capabilities.

4 min read4 takeaways
Read articlearrow_forward
DeepSeek benchmark table comparing the GA version of DeepSeek V4 Pro with other AI models
AI Research13 Aug 2026

DeepSeek V4 Pro Reaches GA with Adjustable Reasoning and Responses API Support

DeepSeek has released the GA version of V4 Pro for its app, web service and API. The 0813 model adds adjustable reasoning, native Responses API support, and a later hosted route on NVIDIA NIM.

4 min read3 takeaways
Read articlearrow_forward
Official Google DeepMind artwork for Gemini 3.7 Flash
AI Research13 Aug 2026

Gemini 3.7 Flash Is Now Generally Available

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

4 min read3 takeaways
Read articlearrow_forward
Grok 4.6 announcement artwork from the xAI news page
AI Research12 Aug 2026

Grok 4.6: xAI Ships a Long-Running Agent Model to OpenRouter

xAI announced Grok 4.6 on 12 August 2026, tuned for long-running agents and ambitious visual work. It is live on OpenRouter with a 500,000-token context at $2 per million input tokens, and the release date is confirmed by the official announcement.

4 min read3 takeaways
Read articlearrow_forward
ByteDance Seed 2.1 Turbo listing artwork from OpenRouter
AI Research12 Aug 2026

ByteDance Seed 2.1 Turbo Reaches OpenRouter After Its June Release

ByteDance released Seed 2.1 Turbo on 23 June 2026. Its OpenRouter route reached AZ Labs monitoring on 12 August with text, image and video input, a 262,144-token context, and pricing from $0.50 per million input tokens.

3 min read3 takeaways
Read articlearrow_forward
OpenRouter model artwork for InclusionAI Ling 3.0 Flash Fin
Industry News27 Aug 2026

OpenRouter Lists InclusionAI's Ling 3.0 Flash Fin as a Free Route

OpenRouter has added InclusionAI's Ling 3.0 Flash Fin as a zero-price route with a 262,144-token context window, 32,768-token maximum output, reasoning and tool support.

5 min read4 takeaways
Read articlearrow_forward
OpenRouter model card for the free Ling 3.0 Tiny route
Industry News06 Aug 2026

OpenRouter Lists InclusionAI's Ling 3.0 Tiny as a Free Route

OpenRouter now lists InclusionAI's Ling 3.0 Tiny as a zero-price route with a 262K-token context window, switchable thinking and instant modes, and a 32K maximum output.

6 min read5 takeaways
Read articlearrow_forward
Gemini 2.5 announcement artwork from Google
Industry News04 Feb 2026

Google DeepMind Announces Gemini 2.5 Pro

Google DeepMind launched Gemini 2.5 Pro with native multimodal reasoning, setting new benchmarks across coding, math, and scientific analysis tasks.

5 min read3 takeaways
Read articlearrow_forward
Anthropic illustration for Claude 4 showing Claude balancing multiple tasks
AI Research25 Jan 2026

Anthropic's Claude 4 Sets New Benchmarks

Anthropic's latest model, Claude 4, pushes the frontier of responsible AI development with state-of-the-art performance on safety and helpfulness benchmarks.

5 min read3 takeaways
Read articlearrow_forward
NVIDIA illustration of AI model infrastructure and accelerated inference
Industry News07 Aug 2026

NVIDIA NIM Removes DeepSeek V4 and Mistral Medium Models from Free API

NVIDIA removed DeepSeek V4 Flash, DeepSeek V4 Pro and Mistral Medium 3.5 128B from the NIM serverless API on 7 August 2026. We confirmed the removals against the live endpoint and tracked where these models still run.

3 min read3 takeaways
Read articlearrow_forward
Meta Muse Glimmer 30B model artwork from its Hugging Face model card
AI Research10 Aug 2026

Meta Muse Glimmer 30B: An Open-Weight Agentic Model Built to Run Locally

Meta Superintelligence Labs has released Muse Glimmer 30B, an Apache 2.0 open-weight model distilled from Muse Spark and tuned for local agent workflows on consumer GPUs. It covers tool use, long-context reasoning and multimodal input in a package that fits under 20 GB with 4-bit quantisation.

6 min read3 takeaways
Read articlearrow_forward
mail

Stay ahead of the curve

Get the latest AI news, tutorials, and product updates delivered straight to your inbox. No spam, unsubscribe anytime.

Join operators, founders, and delivery teams who want applied AI insight, not hype.