31 related articles

GitHub Copilot switches to usage-based billing. AI coding tools move from subscriptions to compute consumption. Learn the industry logic and how AI Cost Engineering helps developers control spending.

Moonshot AI launches Kimi K3 with 2.8 trillion parameters and 1M token context. Google delays Gemini 3.5 Pro, AI coding tools upgrade collectively as competition shifts to coding and Agent capabilities.

Anthropic Claude Max subscribers find Claude Code only deducts Extra Usage Credits, sparking debate over subscription benefit boundaries and billing transparency.

Anthropic is migrating Claude from subscriptions to pay-as-you-go credits. This deep dive explains the mechanics, rationale, industry impact, and user strategies for this billing revolution.

DeepSeek seeks $7B for custom AI inference chips; Zhipu AI explores ASIC. Deep dive into China's AI compute independence strategy, multimodal generation, agents, and hardware trends.

A Cursor user's Auto+Composer usage dropped from 16% to 8% overnight. Explore 4 possible causes: billing resets, silent rule changes, refunds, and UI bugs.

Are third-party ChatGPT top-up services really safe? This deep dive unpacks how they work, the ban risks, and financial dangers — plus the right way to subscribe officially.

A Cursor user found Composer 2.5 running in Fast Mode despite disabling it, causing 6x usage overruns. We analyze the causes, cost impact, and how to protect your budget.

Pylva is an open-source, self-hosted AI Agent billing engine with full usage tracking, flexible per-customer billing rules, and automated invoicing. A deep dive into its features and the economics of the Agent era.

AI coding startup Lovable is reportedly in talks to raise $300M led by Menlo Ventures, potentially doubling its valuation to $13.2B. A deep dive into the business logic and risks.

Muse Spark 1.1 launches with an ultra-low cost focus. We break down the pricing strategy, technical approaches behind it, and its real value for developers and small teams.

An in-depth look at the Skills paradigm in AI programming: through intent routing and script encapsulation, let AI agents auto-manage multi-channel LLM APIs on a One API gateway for one-click distribution, health checks, and auto-degradation.

Fable 5, an AI storytelling platform, opens to all paid users and sparks debate on Hacker News. We analyze AI creation tools' practicalization trend across product positioning, competition, and access strategy.

OpenAI's top flagship model integrates with Codex, hitting 750 tokens/sec on Cerebras wafer chips. We break down MoE architecture, subscription changes, and open-source advances from Hunyuan and Longcat 2.0.

Struggling to subscribe to ChatGPT Plus in China? This guide explains how third-party top-up services work, the real security and ban risks involved, and safer alternatives including official subscriptions, API access, and Chinese LLMs.

App Builder generates single-file runnable apps from natural language, with real-time sandbox preview and conversational revision. Deep analysis of its workflow, architecture, limitations, and costs.

GitHub Copilot shifts from flat-rate to per-token billing, sending dev costs from $29/mo to $1,000+. Uber burns its annual AI budget in months. A deep dive into Token Doomsday.

Cursor built Composer 2.5 on Kimi K2 open-source model, ranking 3rd on coding benchmarks and surpassing K2.6. Deep dive into Cursor's data flywheel, product architecture, and pricing.

Microsoft Copilot Cowork launches with multi-model architecture, considering DeepSeek V4 as a low-cost option. Deep dive into usage-based pricing, WebIQ search, and Microsoft's enterprise AI agent strategy.

Hands-on test of OpenAI Codex free quota reset cards, verifying whether both weekly and 5-hour rolling quotas can be fully restored. Includes usage methods and invite-a-friend rules.