Skip to content

Changelog

Product updates, model releases, and platform improvements.

August 2026

Add Grok 4.6, DeepSeek V4 Pro 0813 & Qwen3.8 2.4T A95B

Grok 4.6

Grok 4.6 is SpaceXAI's smartest frontier model, delivering top-tier performance across coding, knowledge work, and STEM reasoning. It is designed for demanding technical and professional workloads that require strong problem solving, accurate instruction following, and reliable execution.

Optimized for software engineering, scientific analysis, and complex knowledge tasks, Grok 4.6 is well suited for advanced coding, research, and agentic workflows where high capability and reasoning quality are critical.

DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813 is DeepSeek's large-scale Mixture-of-Experts (MoE) model and the general availability (GA) release of DeepSeek V4 Pro. It is designed for high-capability workloads requiring advanced reasoning, coding, and agentic task execution.

As the production-ready V4 Pro release, it is well suited for complex software engineering, long-horizon agent workflows, and demanding reasoning tasks where reliability and model capability are critical.

Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is Qwen's open-weight sparse Mixture-of-Experts (MoE) model and the open-weight counterpart to Qwen3.8 Max. It features 2.4T total parameters with 95B activated per token, combining frontier-scale capacity with efficient sparse inference.

Designed for coding, research, complex reasoning, and agentic workflows, the model is well suited for demanding long-horizon tasks and advanced autonomous systems while providing the flexibility and customization benefits of open weights.

Read update

Add Nemotron 3.5 Lightning

Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning is an open Mixture-of-Experts (MoE) model with 30B total parameters and 3B active per token, optimized for high-throughput agentic workloads and efficient inference.

Its lightweight active compute and open design make it well suited for specialized agents, domain-specific customization, and scalable production deployments where speed, cost efficiency, and adaptability are key.

Read update

Add Muse Glimmer 30B

Muse Glimmer 30B

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It combines strong multi-step reasoning, reliable tool use, failure recovery, image understanding, and multilingual support across 100+ languages.

Designed for long-horizon agentic and coding workflows, Muse Glimmer 30B offers a practical balance of capability and deployment efficiency, making it well suited for local coding assistants, multimodal agents, and production workflows that require sustained autonomous execution.

Read update

Add Muse Spark 1.2

Muse Spark 1.2

Muse Spark 1.2 is Meta's multimodal reasoning model designed for complex agentic and software engineering workflows. It supports text, image, video, audio, and PDF inputs with text output, and features a 1M-token context window for sustained reasoning across large, multi-stage tasks.

Built for flexible multi-agent execution, Muse Spark 1.2 can serve as either a coordinating main agent or a parallel task-focused subagent. With configurable reasoning effort, structured outputs, parallel function calling, and broad coding-harness compatibility, it is well suited for multi-file refactoring, extended debugging, whole-repository generation, and long-horizon development workflows.

Enjoy.

Read update

Add Qwen3.8 Max

Qwen3.8 Max

Qwen3.8 Max is the flagship model in Alibaba’s Qwen3.8 series and the general-availability successor to Qwen3.8 Max Preview. It is a multimodal reasoning model designed for complex tasks across reasoning, visual understanding, coding, and agentic workflows.

As the production-ready top tier of the Qwen3.8 family, it is well suited for advanced problem solving, multimodal analysis, software engineering, and long-running tool-driven applications.

Enjoy.

Read update