Add Qwen 3.5 Full Series & Seed-2.0-Mini
The full Qwen 3.5 series is provided at **Apertis Coding Plan** as well, Enjoy it.
Product updates, model releases, and platform improvements.
Add Qwen 3.5 Full Series & Seed-2.0-Mini
The full Qwen 3.5 series is provided at **Apertis Coding Plan** as well, Enjoy it.
Add Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Gemini 3.1 Flash Image Preview (also known as "Nano Banana 2") is Google's latest state-of-the-art image generation and editing model, delivering Pro-level visual quality at Flash-level speed. It combines strong contextual understanding with fast, cost-efficient inference, enabling high-quality image generation and seamless iterative editing. Optimized for both performance and accessibility, it makes advanced visual creation workflows faster and more scalable.
Cached responses now support streaming (SSE) delivery, covering ~80% of API traffic that uses stream: true.
New feature: Cached responses now support streaming (SSE) delivery, covering ~80% of API traffic that uses stream: true.
Cache Correctness Hardening
cacheable — providers default to ~1.0 for omitted values
corruption
Cache TTL & Infrastructure
Enjoy it.
✨ New Feature: Monthly Budget Controls
You can now set a monthly spending cap on your API usage. Once enabled, usage is tracked against your limit and automatically resets on your chosen billing cycle date.
Add MiniMax M2.5 (Lightning)
MiniMax-M2.5-Lightning is the high-speed variant of the M2.5 series, optimized for low latency, real-time responsiveness, and high-frequency workloads. It retains the core planning and execution strengths of M2.5 while further improving inference efficiency and response speed, making it ideal for interactive applications, rapid coding assistance, and workflow automation. With enhanced cost efficiency and reduced latency, M2.5-Lightning is particularly well suited for high-throughput, always-on deployments and production environments where speed and scalability are critical.