💸 The
SD2Mini / SD2Fast discount groups now run through 7 October 2026 23:59 (UTC+8), instead of ending on 7 September —— Volcengine has extended its limited-time promotion on Seedance 2.0 mini (60% off) and fast (25% off) to 7 October, and APIYI has followed suit: both groups keep their 0.10x / 0.15x rates, their validity in the console is now 7 October 23:59 (UTC+8), and existing Tokens carry over with no code changes. Any further extension depends on the provider’s announcements.📖 View details✅
gpt-image-2-vip saw about 15 minutes of concurrency saturation from 13:46 to 14:00 (UTC+8) and has been back to normal since 14:00 —— Root cause was insufficient concurrency in the account pool behind it: the reverse channel is an ongoing contest with the provider’s platform risk controls, so pool capacity fluctuates. No code changes are needed; we suggest wiring the official gpt-image-2 as a fallback, since both share the images API with compatible parameters.📖 View details🚀
gpt-6-astra has joined the half-price Codex_Reverse group, billed at 0.5x the official rate — the group is the Codex reverse-engineered economy channel; with the 0.5x discount Astra bills at $5 in / $25 out per 1M tokens and $0.50 cached reads. Existing Codex_Reverse tokens can call it directly. Suited to Codex coding, Cherry Studio chat, and agent setups such as OpenClaw; the default / svip official-relay groups remain available for production and stability-sensitive work.📖 View details🚀
gpt-6-astra is live in both the default and svip groups — OpenAI’s next-generation flagship, released on 3 September and built for computer use and long-horizon agent work: 1.05M context, reasoning_effort gains xhigh / max for five levels. Priced at $10 in / $50 out per 1M tokens with $1 cached reads, all four billing items match OpenAI’s list price; both /v1/responses and /v1/chat/completions are open. Note the standard public release refuses vulnerability-discovery tasks; ordinary development is unaffected.📖 View details🚀 New
Gemini_Reverse group: a Gemini reverse-engineered economy channel at a 0.5x discount — backed by reverse-engineered channels from Gemini CLI and Antigravity, billed at half the official price; select the group when creating a token in the console, with model names and endpoints unchanged. The default group remains the AI Studio and Vertex official relay at official parity. The reverse group is cheaper and suits personal learning and coding; production workloads and stability-sensitive tasks should use the default official-relay group.📖 View details💸 The
Codex_Reverse group discount drops from 0.7x to 0.5x, half the official price — for gpt-5.6-sol, the official $4 input / $20 output now bills at $2 / $10 on this group. Existing Codex_Reverse tokens pick up the new rate automatically; model names and endpoints are unchanged. The group runs on Codex reverse-engineered resources and suits personal learning and coding; use the default official-relay group for production.📖 View details✅ Claude models are back to normal, and the earlier unavailability notice is withdrawn —— Anthropic’s status page
status.claude.com posted Resolved at 16:23 UTC on 3 September: the provider-side incident affecting claude-mythos-5-1, claude-fable-5-1, and claude-opus-5 ended at 00:16 on 4 September (UTC+8). The APIYI gateway was not changed and all claude-* models are serving normally again; Grok status is being tracked separately.📖 View details⚠️ Claude models are temporarily unavailable due to a provider-side incident, and Grok models are down at the same time —— Anthropic’s service is currently experiencing an incident, so requests to all
claude-* models fail or time out; follow the official status page status.claude.com for progress and the estimated recovery time. grok-* models are also interrupted and we are following up. Other routes: OpenAI models are running normally, with the OpenAI API official route largely normal and the Azure official route available as backup; Gemini models are unaffected. We will post here once service is restored.📖 View details💸
gpt-5.6-sol price cut synced: input $5 → $4, output $30 → $20, so the flagship tier is now cheaper than gpt-5.5 —— OpenAI lowered GPT-5.6 Sol pricing on 3 September (official $4 / $20, promotional rate guaranteed at least through 21 November 2026), and APIYI has synced: within 272K, $4 input / $20 output / $0.40 cached read, with cache writes at 1.25× the input rate ($5); above 272K the whole request is billed at 2× input and 1.5× output. Against the previous-generation gpt-5.5 ($5 / $30), Sol is 20% cheaper on input and a third cheaper on output; existing code migrates by changing only the model field.📖 View details📊 Google has published the official
gemini-3.8-flash spec sheet, confirming the 1M-token context —— The model page is up: a 1,048,576-token input limit, 65,536-token output limit, and text / image / video / audio / PDF input. Thinking offers only low / medium / high; minimal returns an error, matching our pre-launch tests. APIYI went live with the model on the evening of 2 September across both groups and both endpoints, so the earlier advice to wait for official specs before a full cutover no longer applies.📖 View details📖 For earlier updates, visit the Live Updates Archive — browse history by month, category, or vendor.
Deep Dive
AI Radar
Telegram
Global users
Media Models
New additions