Bookmark this page: press Ctrl+D to keep it handy — new model launches and promotions are announced here first.
🔥 Latest Updates
New Model · OpenAI
GPT-6 Astra Is Live
OpenAI’s New Flagship for Computer Use and Long-Horizon Agents, at Official PriceOpenAI’s next-gen flagship, released 3 September: 1.05M context, five reasoning-effort levels with newxhigh / max, Terminal-Bench 4.0 up from 37.3% to 57.9%. Priced at $10 in / $50 out per 1M tokens with $1 cached reads, matching OpenAI’s list price item for item, in the default / svip official-relay groups, with the Codex_Reverse group at a 0.5x discount.📖 Read details | 🔗 PromotionsPrice Update · OpenAI
GPT-5.6 Sol Price Cut Synced
Input Down 20%, Output Down 33%, Now Cheaper Than gpt-5.5OpenAI cut GPT-5.6 Sol pricing on 3 September and APIYI has synced:gpt-5.6-sol now costs $4 input (was $5) and $20 output (was $30) per 1M tokens, cached reads $0.40, with the limited-time promotional rate guaranteed at least through 21 November 2026. The flagship tier is now cheaper than the previous-generation gpt-5.5 ($5 / $30); existing code migrates by changing only the model field.📖 Read details | 🔗 Model detailsNew Model · Google
Gemini 3.8 Flash Is Live
Ahead of the Docs, Same Price as 3.7Google’s newest Flash, out 2 September. Google’s own model docs and launch blog have not listed it yet; APIYI has it open for calls. Priced line-for-line withgemini-3.7-flash at $0.75 in / $3.75 out per 1M tokens, so migrating costs nothing. 150 paired test cases across both protocols found capability parity with 3.7 and no regression unique to it.📖 Read details | 🔗 PromotionsNew Model · Anthropic
Claude Fable 5.1 Is Live
Same Price, Cache Reads Cut to a QuarterAnthropic’s new Mythos-class flagship, shipped 1 September. The headline change is cache reads falling from $1.00 to $0.25 per 1M tokens, with input at $10 and output at $50 unchanged. APIYI has matched the cut — every line item is priced in line with the provider. Three breaking changes: forced tool use now 400s, thinking blocks are model-bound, and editing earlier turns invalidates them.📖 Read details | 🔗 PromotionsNew Model · ByteDance
Seedance 2.5 Is Live
30-Second Clips, 30 Reference Images, A Broad Capability JumpThe model isdoubao-seedance-2-5-260628, with an endpoint and request shape identical to 2.0 — changing the model field is the whole migration. The duration cap rises from 15 to 30 seconds and reference images from 9 to 30, audio can stand alone as a reference, and it adds mov output plus explicit video edit/extend task types. For current group and pricing, see the 31 August entry above.📖 Read details | 🔗 Capabilities and pricingNew Model · DeepSeek
DeepSeek Vision Model Is Live
No Vision Premium, Priced Like Text-Only V4 Flashdeepseek-v4-flash-vision-exp is DeepSeek’s first vision model, adding image input on top of the V4 Flash base while keeping the 1M context, thinking, function calling and caching. Images convert to input tokens by size and are capped at 384 each; pricing matches the text-only version at $0.44 input and $1.32 output per 1M tokens. Use a default group token for the OpenAI format and a ClaudeCode one for the Anthropic format.📖 Read details | 🔗 Integration docsFeature Update · OpenAI
Transparent Backgrounds on gpt-image-2
One Parameter, Real Alpha-Channel PNGsOpenAI opened up thetransparent value of background for GPT-Image-2, and APIYI has verified it. Pass background: "transparent" with png or webp to get a genuinely transparent image; text-to-image, image editing, and the Responses image tool all support it, at no extra cost. jpeg has no alpha channel and is mutually exclusive with transparency.📖 Read details | 🔗 Transparency FAQPrice Update · DeepSeek
DeepSeek Price Change
Peak/Off-Peak Upstream, Peak Tier at APIYIDeepSeek switches to two-tier billing at 00:00 on 17 August (UTC+8), with off-peak at half the peak rate. APIYI follows fordeepseek-v4-flash and deepseek-v4-pro but bills at the peak tier at all times: $0.44/$1.32 and $1.32/$3.96 per 1M tokens. The reason is supply — official capacity shortfalls are covered by the pricier BytePlus and Alibaba Cloud routes, at no margin. Recharge bonuses still stack.📖 Read details | 🔗 PromotionsNew Model · Google
Gemini 3.7 Flash Is Live
Official Price, Half of What the Previous Gen CostsGoogle’s next-gen Flash workhorse, shipped 13 August. Coding and agents lead the upgrade: DeepSWE v1.1 climbs 48.6% → 65.3%, business-process automation 17.0% → 30.4%. 1M context, three thinking tiers, priced at $0.75/$3.75 per 1M tokens matching Google — a limited-time promotional rate that reverts to $1.50/$7.50 after 31 December.📖 Read details | 🔗 Top-up promotions📖 These are the 10 most recent announcements. For everything earlier, see the announcement archive, browsable by month, category, or vendor.
Deep dives
Browse the AI Radar section
Live updates
Model status and service notices
Promotions
See current deposit bonuses