Skip to content

Changelog

Product updates, model releases, and platform improvements.

April 2026

Quota multiplier adjustments for Lite, Pro, and Plus plans — effective 2026-04-24 00:00 UTC

Effective 2026-04-24 at 00:00 UTC, quota multipliers across Lite, Pro, and Plus plans will be adjusted in response to recent upstream AI provider pricing changes.

Why

  • Anthropic Claude (upstream routing channel) base rates have trended upward across multiple providers
  • Z.AI GLM-5.1 base rates have trended upward across multiple providers
  • OpenAI GPT-5.4 serving costs have increased on several routes
  • Additionally, claude-opus-4-6 on the Lite plan is being lowered because our review showed the previous multiplier was set higher than current real cost justifies

Changes

Lite Plan ($12/month, 600 quota per cycle)

  • glm-5.1: 0.51 → 1.5
  • code:claude-opus-4-6: 5.00 → 10.0
  • claude-opus-4-6: 5.00 → 2.0 (decrease)
  • gemini-3-flash-preview: 0.10 → 0.2
  • gemini-3.1-pro-preview: 0.43 → 0.77

Pro Plan ($25/month, 900 quota per cycle)

  • glm-5.1: 0.37 → 1.5
  • claude-opus-4-6: 1.00 → 1.5
  • gpt-5.4: 0.62 → 1.0
  • code:claude-opus-4-6: 0.75 → 1.5

Plus Plan ($60/month, 1,500 quota per cycle)

  • claude-opus-4-6: 0.75 → 1.5
  • glm-5.1: 0.25 → 1.0
  • gpt-5.4: 0.45 → 0.8
  • claude-opus-4-7: 3.00 → 4.0

What stays the same

  • Monthly subscription fee
  • Billing cycle and quota allowance per plan
  • Model access and plan tiers
  • Pay-As-You-Go (PAYG) fallback behavior

Your options

If you do not agree with the changes, you may cancel your subscription at any time before 2026-04-24 00:00 UTC at https://apertis.ai/setting. Affected users will also be notified by email.

For questions, contact us at hi@apertis.ai.

Read update

Add Claude Opus 4.7

Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, designed for long-running, asynchronous agent workflows. Building on Opus 4.6, it delivers stronger performance on complex, multi-step tasks and more reliable execution across extended pipelines such as large codebases, multi-stage debugging, and end-to-end project orchestration. Beyond coding, Opus 4.7 enhances knowledge work capabilities, including document drafting, presentation creation, and data analysis. With strong coherence over long outputs and sessions, it is well suited for tasks requiring persistence, judgment, and sustained execution.

Enjoy it.

Read update

Fallback Timeout Setting for Coding Plan Users

What's new

When Apertis routes your request to an upstream provider, it waits a set amount of time before switching to the next available channel. Previously this was fixed at 30 seconds — fine for most models, but too short for preview and reasoning models processing large context windows.

You can now adjust this in Settings → Subscription Keys → Fallback Timeout (range: 5s–300s).

Who should change this

  • Using gemini-3-flash-preview, claude-opus-4-thinking, or other preview/reasoning models with large prompts? Increase to 120s+
  • Using standard models like gpt-4o, claude-sonnet-4? Default 30s is fine

How it works

1. Go to Settings → Subscription Keys 2. Find Fallback Timeout in the metadata section 3. Enter your preferred value in milliseconds (e.g., 120000 for 120s) 4. Click Save

Changes take effect immediately.

Read update

Add Claude Opus 4.6 (Fast)

Claude Opus 4.6 (Fast)

Opus 4.6 is Anthropic's more faster version of Opus 4.6 model for coding and long-running professional workflows, designed for agents that operate across entire workflows rather than single prompts. It demonstrates strong performance on large codebases, complex refactoring, and multi-step debugging, with improved contextual understanding, deeper problem decomposition, and higher reliability on challenging engineering tasks compared to earlier generations. Beyond software development, Opus 4.6 excels at sustained knowledge work, producing near production-ready documents, technical plans, and analyses in a single pass while maintaining coherence across long outputs and extended sessions. Its strength in persistence, judgment, and structured execution makes it well suited for technical design, migration planning, and end-to-end project execution.

Enjoy it.

Read update

Add GLM 5.1

GLM 5.1

GLM-5.1 delivers a major advancement in coding capability, with significant improvements in handling long-horizon tasks. It is designed to operate beyond short interactions, enabling continuous, autonomous execution over extended periods. The model can work independently on a single task for 8+ hours, performing planning, execution, and iterative self-improvement to produce complete, engineering-grade results, making it well suited for complex development workflows and autonomous agent systems.

Enjoy it.

Read update