AZ Labs

News & Insights

The latest in artificial intelligence

Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

OpenRouter artwork for the Qwen3.8-27B model route
Featured
AI Research14 Aug 2026

Qwen3.8-27B Ships Open Weights with 262K Context

Qwen has released Qwen3.8-27B under Apache 2.0. The dense vision-language model has a native 262,144-token context, adjustable reasoning, image and video input, and a live OpenRouter route while Qwen Cloud hosting remains pending.

4 min read3 key takeaways
OpenAI GPT-5 announcement artwork showing the GPT-5 flagship model
AI Research05 Feb 2026

OpenAI Releases GPT-5 with Enhanced Reasoning Capabilities

OpenAI has unveiled GPT-5, its most advanced language model to date, featuring breakthrough reasoning capabilities that bring AI closer to human-level problem solving.

6 min read3 takeaways
Read articlearrow_forward
AZ Labs GLM-5.3 release timeline showing the maker announcement and provider route timings
AI Research14 Aug 2026

GLM-5.3 Arrives for Coding and Cyber Work, with Weights Still Pending

Z.ai has released GLM-5.3 through its Coding Plan and ZCode. The model uses the same 743B base as GLM-5.2, adds low, high and max thinking effort, and is already listed by OpenCode Go and ClinePass. Maker-hosted API access and open weights remain staged.

5 min read3 takeaways
Read articlearrow_forward
AZ Labs artwork showing the dots3-note preview name and its core model specifications
AI Research14 Aug 2026

Dots3-Note Preview Opens a 280B Multimodal Agent Model

Dots Studio has released dots3-note preview, its first open-weight dots3 model. The multimodal MoE has 280 billion total parameters, 16 billion active parameters, a 512K context window, and an Apache 2.0 release.

4 min read3 takeaways
Read articlearrow_forward
DeepSeek benchmark table comparing the GA version of DeepSeek V4 Pro with other AI models
AI Research13 Aug 2026

DeepSeek V4 Pro Reaches GA with Adjustable Reasoning and Responses API Support

DeepSeek has released the GA version of V4 Pro for its app, web service and API. The 0813 model adds low, high and max reasoning effort, native Responses API support, and a new peak and off-peak pricing schedule from 16 August.

4 min read3 takeaways
Read articlearrow_forward
Official Google DeepMind artwork for Gemini 3.7 Flash
AI Research13 Aug 2026

Gemini 3.7 Flash Is Now Generally Available

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

4 min read3 takeaways
Read articlearrow_forward
Grok 4.6 announcement artwork from the xAI news page
AI Research12 Aug 2026

Grok 4.6: xAI Ships a Long-Running Agent Model to OpenRouter

xAI announced Grok 4.6 on 12 August 2026, tuned for long-running agents and ambitious visual work. It is live on OpenRouter with a 500,000-token context at $2 per million input tokens, and the release date is confirmed by the official announcement.

4 min read3 takeaways
Read articlearrow_forward
ByteDance Seed 2.1 Turbo listing artwork from OpenRouter
AI Research12 Aug 2026

ByteDance Seed 2.1 Turbo Reaches OpenRouter After Its June Release

ByteDance released Seed 2.1 Turbo on 23 June 2026. Its OpenRouter route reached AZ Labs monitoring on 12 August with text, image and video input, a 262,144-token context, and pricing from $0.50 per million input tokens.

3 min read3 takeaways
Read articlearrow_forward
InclusionAI Ling 3.0 Flash model artwork from its Hugging Face model page
Industry News25 Jul 2026

Ling 3.0 Flash Runs in Grok CLI Through OpenRouter

A practical model-routing experiment puts InclusionAI's Ling 3.0 Flash inside Grok CLI, with the request verified through OpenRouter and the same route working in Pi.

5 min read3 takeaways
Read articlearrow_forward
OpenRouter model card for the free Ling 3.0 Tiny route
Industry News06 Aug 2026

OpenRouter Lists InclusionAI's Ling 3.0 Tiny as a Free Route

OpenRouter now lists InclusionAI's Ling 3.0 Tiny as a zero-price route with a 262K-token context window, switchable thinking and instant modes, and a 32K maximum output.

6 min read5 takeaways
Read articlearrow_forward
Gemini 2.5 announcement artwork from Google
Industry News04 Feb 2026

Google DeepMind Announces Gemini 2.5 Pro

Google DeepMind launched Gemini 2.5 Pro with native multimodal reasoning, setting new benchmarks across coding, math, and scientific analysis tasks.

5 min read3 takeaways
Read articlearrow_forward
The Human Edge visual about artificial intelligence and South African business
Industry News03 Feb 2026

How South African Businesses Are Adopting AI in 2026

From fintech to healthcare, South African companies are rapidly integrating AI into their operations. We explore the trends driving adoption across the continent.

7 min read3 takeaways
Read articlearrow_forward
AI-assisted scheduling interface showing a calendar and appointment times
Company News30 Jan 2026

AZ Labs Launches Aura: AI Booking Assistant

Introducing Aura, our AI-powered booking assistant that handles scheduling, reminders, and client communications autonomously for service businesses.

4 min read3 takeaways
Read articlearrow_forward
Illustration of a person using an AI voice agent through a headset
Tutorials28 Jan 2026

Building Effective AI Voice Agents: A Complete Guide

Learn how to design, build, and deploy AI voice agents that deliver natural conversations and measurable business results.

8 min read3 takeaways
Read articlearrow_forward
Anthropic illustration for Claude 4 showing Claude balancing multiple tasks
AI Research25 Jan 2026

Anthropic's Claude 4 Sets New Benchmarks

Anthropic's latest model, Claude 4, pushes the frontier of responsible AI development with state-of-the-art performance on safety and helpfulness benchmarks.

5 min read3 takeaways
Read articlearrow_forward
African Union AI policy event about building AI for Africa’s security and collaboration
Industry News22 Jan 2026

The Future of AI Regulation in Africa

As AI adoption accelerates across Africa, governments are developing regulatory frameworks. We examine the policies and what they mean for businesses.

6 min read3 takeaways
Read articlearrow_forward
Diagram showing AI-assisted contract review from author to client
Tutorials18 Jan 2026

5 Ways AI Is Transforming Legal Services

From contract analysis to legal research, AI is reshaping how law firms operate. Discover the practical applications driving efficiency in the legal sector.

7 min read3 takeaways
Read articlearrow_forward
Students using laptops and tablets in an AI-supported classroom
Company News15 Jan 2026

AZ Labs Expands Into Education and EdTech

We're bringing our AI expertise to the education sector, partnering with institutions to build intelligent tutoring systems and administrative automation.

4 min read3 takeaways
Read articlearrow_forward
Protection of Personal Information Act graphic for South Africa
Tutorials14 Jul 2026

Deploying AI Voice Agents Under POPIA: What SA Businesses Need to Know

A practical guide to running AI voice agents in South Africa while staying compliant with POPIA. Covers data residency, consent, storage, and how to design voice workflows that respect privacy obligations from day one.

7 min read3 takeaways
Read articlearrow_forward
NVIDIA illustration of AI model infrastructure and accelerated inference
Industry News07 Aug 2026

NVIDIA NIM Removes DeepSeek V4 and Mistral Medium Models from Free API

NVIDIA removed DeepSeek V4 Flash, DeepSeek V4 Pro and Mistral Medium 3.5 128B from the NIM serverless API on 7 August 2026. We confirmed the removals against the live endpoint and tracked where these models still run.

3 min read3 takeaways
Read articlearrow_forward
Meta Muse Glimmer 30B model artwork from its Hugging Face model card
AI Research10 Aug 2026

Meta Muse Glimmer 30B: An Open-Weight Agentic Model Built to Run Locally

Meta Superintelligence Labs has released Muse Glimmer 30B, an Apache 2.0 open-weight model distilled from Muse Spark and tuned for local agent workflows on consumer GPUs. It covers tool use, long-context reasoning and multimodal input in a package that fits under 20 GB with 4-bit quantisation.

6 min read3 takeaways
Read articlearrow_forward
mail

Stay ahead of the curve

Get the latest AI news, tutorials, and product updates delivered straight to your inbox. No spam, unsubscribe anytime.

Join operators, founders, and delivery teams who want applied AI insight, not hype.