Skip to content

Changelog

Product updates, model releases, and platform improvements.

July 2026

Price Cut for OpenAI GPT-5.6 Terra & Luna

Price Cut on GPT-5.6 Terra & Luna

Today, OpenAI made GPT-5.6 more affordable and faster. GPT-5.6 Luna now costs 80% less, GPT-5.6 Terra costs 20% less.

Therefore, we also lower the prices for GPT-5.6 Terra & Luna models.

Enjoy.

Read update

Add Qwen3.7 Flash

Qwen3.7 Flash

Qwen3.7 Flash is Alibaba's vision-language reasoning model, designed for multimodal agents, visual coding, search, and computer interaction. It combines fast inference with strong visual understanding, including object recognition, spatial reasoning, and real-world scene perception.

Optimized for interactive and agentic workflows, Qwen3.7 Flash is well suited for GUI understanding, visual question answering, multimodal search, and computer-use applications that require responsive reasoning across text and images.

Enjoy.

Read update

Add Claude Opus 5

Claude Opus 5

Claude Opus 5 is Anthropic's flagship model for advanced reasoning, coding, and long-horizon agentic workflows. It excels at end-to-end software engineering, code review, bug detection, visual analysis of charts and documents, complex office deliverables, and parallel subagent coordination.

The model maintains reliable instruction following and tool use across extended tasks, while remaining effective at lower reasoning-effort settings for workloads that prioritize latency and token efficiency.

Enjoy.

Read update

Add Gemini 3.6 Flash & Gemini 3.5 Flash-Lite

Gemini 3.6 Flash

Gemini 3.6 Flash is Google's high-efficiency model for coding, agentic workflows, and web and application development. It is optimized to produce polished, production-ready outputs with fewer unnecessary revisions, less hedging, and more direct task execution.

By reducing both token usage and the number of model calls required to complete complex tasks, Gemini 3.6 Flash is well suited for high-throughput development, scalable agent systems, and cost-sensitive production workflows.

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is Google's high-efficiency model with enhanced agentic capabilities, optimized for fast, cost-effective inference. It is designed to handle focused tasks with low latency while maintaining strong reasoning and execution quality.

Well suited for subagents in complex multi-agent systems, Gemini 3.5 Flash-Lite excels at executing specialized tasks within larger workflows, making it ideal for scalable agent orchestration and high-throughput production environments.

Enjoy them.

Read update