
OpenAI Releases GPT-5 with Enhanced Reasoning Capabilities
OpenAI has unveiled GPT-5, its most advanced language model to date, featuring breakthrough reasoning capabilities that bring AI closer to human-level problem solving.
News & Insights
Reporting, explainers, and implementation-focused analysis on AI systems, automation, voice agents, and business delivery.

Qwen has released Qwen3.8-27B under Apache 2.0. The dense vision-language model has a native 262,144-token context, adjustable reasoning, image and video input, and a live OpenRouter route while Qwen Cloud hosting remains pending.

OpenAI has unveiled GPT-5, its most advanced language model to date, featuring breakthrough reasoning capabilities that bring AI closer to human-level problem solving.

Z.ai has released GLM-5.3 through its Coding Plan and ZCode. The model uses the same 743B base as GLM-5.2, adds low, high and max thinking effort, and is already listed by OpenCode Go and ClinePass. Maker-hosted API access and open weights remain staged.

Dots Studio has released dots3-note preview, its first open-weight dots3 model. The multimodal MoE has 280 billion total parameters, 16 billion active parameters, a 512K context window, and an Apache 2.0 release.

DeepSeek has released the GA version of V4 Pro for its app, web service and API. The 0813 model adds low, high and max reasoning effort, native Responses API support, and a new peak and off-peak pricing schedule from 16 August.

Google has released Gemini 3.7 Flash as a stable model for coding, agent workflows and multimodal reasoning. It has a 1,048,576-token input limit, 65,536-token output, and direct API pricing from $0.75 per million input tokens through 2026.

xAI announced Grok 4.6 on 12 August 2026, tuned for long-running agents and ambitious visual work. It is live on OpenRouter with a 500,000-token context at $2 per million input tokens, and the release date is confirmed by the official announcement.

ByteDance released Seed 2.1 Turbo on 23 June 2026. Its OpenRouter route reached AZ Labs monitoring on 12 August with text, image and video input, a 262,144-token context, and pricing from $0.50 per million input tokens.

A practical model-routing experiment puts InclusionAI's Ling 3.0 Flash inside Grok CLI, with the request verified through OpenRouter and the same route working in Pi.

OpenRouter now lists InclusionAI's Ling 3.0 Tiny as a zero-price route with a 262K-token context window, switchable thinking and instant modes, and a 32K maximum output.

Google DeepMind launched Gemini 2.5 Pro with native multimodal reasoning, setting new benchmarks across coding, math, and scientific analysis tasks.

From fintech to healthcare, South African companies are rapidly integrating AI into their operations. We explore the trends driving adoption across the continent.

Introducing Aura, our AI-powered booking assistant that handles scheduling, reminders, and client communications autonomously for service businesses.

Learn how to design, build, and deploy AI voice agents that deliver natural conversations and measurable business results.

Anthropic's latest model, Claude 4, pushes the frontier of responsible AI development with state-of-the-art performance on safety and helpfulness benchmarks.

As AI adoption accelerates across Africa, governments are developing regulatory frameworks. We examine the policies and what they mean for businesses.

From contract analysis to legal research, AI is reshaping how law firms operate. Discover the practical applications driving efficiency in the legal sector.

We're bringing our AI expertise to the education sector, partnering with institutions to build intelligent tutoring systems and administrative automation.

A practical guide to running AI voice agents in South Africa while staying compliant with POPIA. Covers data residency, consent, storage, and how to design voice workflows that respect privacy obligations from day one.

NVIDIA removed DeepSeek V4 Flash, DeepSeek V4 Pro and Mistral Medium 3.5 128B from the NIM serverless API on 7 August 2026. We confirmed the removals against the live endpoint and tracked where these models still run.

Meta Superintelligence Labs has released Muse Glimmer 30B, an Apache 2.0 open-weight model distilled from Muse Spark and tuned for local agent workflows on consumer GPUs. It covers tool use, long-context reasoning and multimodal input in a package that fits under 20 GB with 4-bit quantisation.
Get the latest AI news, tutorials, and product updates delivered straight to your inbox. No spam, unsubscribe anytime.
Join operators, founders, and delivery teams who want applied AI insight, not hype.