GPU Prices in 2026: Why They Rise, and Whether to Buy
Nvidia RTX cards are up 20 to 30 percent in a third 2026 hike. Here is the memory-cost chain behind it, and whether to buy a GPU now or wait.
The price of intelligence
Token prices are the sticker. Context, retries, subscriptions and unfinished work decide the bill.
Topic archive · page 2
Analysis of AI pricing, API bills, subscriptions and the real cost of completing useful work.
Nvidia RTX cards are up 20 to 30 percent in a third 2026 hike. Here is the memory-cost chain behind it, and whether to buy a GPU now or wait.
Claude Opus 5 ships at $5/$25 per million tokens with an effort dial, matching Fable 5 intelligence at half the price. Benchmarks, specs, and cost per task.
A practitioner guide to LLM evaluation metrics: reference-based scores, LLM-as-judge, and human review, ranked by accuracy and cost per 10,000 responses.
SK hynix says 2027 will be the worst supply year ever for memory and ADATA expects a decade-long shortage. Here is how long the RAM squeeze may last.
An AI data center explained: how it differs from a traditional data center, what it costs per megawatt, and why power is the binding constraint.
Nearly 200 startups urged Trump not to restrict Chinese open-weight AI. Here is what a ban would target, why enforcement is hard, and the stakes.
Qwen3.8 Max Preview is live through the Alibaba Token Plan and Qoder. Here is what is confirmed about access, benchmarks, open weights, and the 2.4T claim.
Chinese AI models already lead on price and open weights. A six-part scorecard shows why broad leadership is plausible by 2028–2030, but not inevitable.
Current 2026 API and subscription prices for DeepSeek, Qwen, Kimi, GLM, and MiniMax, compared per million tokens and against US models.