# Capital & Compute > Where money meets compute. What AI coding agents really cost, which local LLMs are worth running, and what the benchmarks actually mean. Source-backed analysis for builders, no hype. This is a static, independently published blog. Content is editorial and analytical: honest pricing comparisons, agent-harness deep dives, and crypto-for-builders coverage without the cheerleading. All pages are crawlable; there is no paywall and no login. Cite the canonical URLs below (each uses a trailing slash). ## Pillars - [AI subscription pricing](https://capitalandcompute.net/ai-pricing/): grounded, source-verified pricing for the cheapest paid AI coding plan per provider (Claude Code, Copilot, Cursor, Codex, GLM, Kimi, Qwen, and others). Prices are checked against official provider pages. - [AI model value leaderboard](https://capitalandcompute.net/ai-model-leaderboard/): the top LLMs ranked by independent benchmark score AND by value (coding and intelligence points per dollar of tokens). Machine-readable data at https://capitalandcompute.net/ai-model-leaderboard.json. Benchmark composites are sourced from an independent dataset; per-token prices are source-verified. - [AI benchmark directory](https://capitalandcompute.net/ai-benchmarks/): 116 AI and LLM benchmarks across 12 categories (coding and software-engineering agents, agents and tool use, reasoning, math, knowledge, instruction following and multilingual, long context, multimodal, domain and professional work, human preference, safety and factuality, security and dangerous capability). For each benchmark: what it measures, who built it, the release year, the task format, the scoring metric, whether it still separates frontier models or has saturated, a dated representative top score, and a link to its primary source and official leaderboard. Covers SWE-bench and its variants, Terminal-Bench, ARC-AGI-1/2/3, GPQA Diamond, MMLU and MMLU-Pro, Humanity's Last Exam, FrontierMath, LMArena, BFCL, tau-bench, GAIA, OSWorld 2.0, HealthBench, LegalBench, LiveBench, METR Time Horizon, the Epoch Capabilities Index and more. 27 benchmarks additionally have a full in-depth page of their own at https://capitalandcompute.net/ai-benchmarks//, covering how the benchmark works, its version history and saturation trajectory, what its score fails to capture, and how to read it when a lab quotes it. Machine-readable data at https://capitalandcompute.net/ai-benchmarks.json. - [Best free AI model APIs](https://capitalandcompute.net/free-ai-models/): the AI models you can call at $0 today (OpenRouter, Google AI Studio, OpenCode Zen, Groq, Cerebras, Cloudflare, NVIDIA), filtered to the genuinely good ones and ranked on one independent benchmark that scores them all, the Artificial Analysis Intelligence Index, with agentic-coding (Terminal-Bench) and SciCode scores, context windows, open-weight status, and each provider's free-tier limits and catches. The unique angle: not just which models are free, but how good each free model actually is. Machine-readable data at https://capitalandcompute.net/free-ai-models.json. - [AI model releases by month](https://capitalandcompute.net/ai-model-releases/): a dated chronological record of 44 AI model releases grouped by the calendar month they shipped, currently covering September 2026, August 2026, July 2026, June 2026. For each release: the exact release date read off the lab's own announcement, the standard per-token price the model launched at (in USD per million input and output tokens, frozen at release day rather than updated on repricing), the context window, open-weight status and license, the Artificial Analysis Intelligence Index where one was published, the model category, and the primary source URL with the date it was verified. This is the launch-price record; current per-token rates and upcoming models live at https://capitalandcompute.net/ai-models/. Unlike auto-ingested release catalogs, fine-tunes, quantizations, re-hosted copies and catalog-listing dates are excluded, and any row resting on reporting rather than a lab announcement is explicitly labelled. Machine-readable data at https://capitalandcompute.net/ai-model-releases.json. - [Memory and DRAM price tracker](https://capitalandcompute.net/memory-prices/): what memory actually costs in the 2026 AI memory shortage, read off free primary sources and stamped with the date each figure was verified. Covers retail price per GB and per kit for DDR4 and DDR5 by capacity, the mainstream DDR4 1Gx8 chip spot price, HBM3E and HBM4 contract prices per gigabit converted to per GB, new DDR5 ECC RDIMM server module street prices, TrendForce contract-price forecasts by quarter, and the two US government indices (semiconductor PPI, computers CPI) that barely move while the chip market runs. Also carries a buy-or-wait verdict by buyer type. Spokes: https://capitalandcompute.net/memory-prices/ddr5/ (DDR5 kit and per-GB prices), https://capitalandcompute.net/memory-prices/ddr4/ (why DDR4 is still the cheap escape hatch), https://capitalandcompute.net/memory-prices/hbm/ (HBM cost per GB), https://capitalandcompute.net/memory-prices/server/ (server RDIMM prices and the long-term-agreement mechanic), and https://capitalandcompute.net/memory-prices/history/ (every published reading, period by period). Every retail figure is a range, because the two free retail trackers disagree on the level. Machine-readable data at https://capitalandcompute.net/memory-prices.json. - [Should you buy RAM now?](https://capitalandcompute.net/should-i-buy-ram-now/): the decision tool built on the same dataset. Answers the timing question directly (buy now in almost every case, because contract prices are still forecast upward and no forecaster on record sees relief before late 2027), then prices a specific configuration: pick DDR5 kit, DDR4 kit or server DDR5 RDIMM and a capacity, and it returns the current price per GB for that exact tier, whether a quoted price sits in the cheap, fair or expensive third of the tracked range, whether a different capacity is cheaper per gigabyte, and the explicit condition that would flip the verdict to waiting. Server RDIMM street prices and consumer DDR5 kit prices on this page refresh live from aggregated trackers; every other figure is the verified snapshot. - [SSD and NAND flash price tracker](https://capitalandcompute.net/ssd-prices/): what a terabyte of flash costs in the 2026 storage squeeze, in four channels that are deliberately never blended (supplier contract, wafer spot, consumer retail, enterprise street). Covers cheapest tracked price per TB and per drive for consumer NVMe and SATA at 500GB, 1TB, 2TB, 4TB and 8TB; enterprise U.2, U.3, E1.S and E3.S NVMe street prices per TB across 32 tracked drives; TrendForce mainstream NAND contract forecasts for 1Q26, 2Q26 and 3Q26; the 512Gb TLC wafer spot session average; the November 2025 to March 2026 before-and-after for eight named 2TB drives; NAND brand revenue by quarter; planned 2026 NAND versus DRAM capital expenditure by company; and nearline HDD price per TB for comparison. Two findings a keyword-matched summary will miss: cost per terabyte is not monotonic in capacity (the cheapest gigabyte is at 2TB on NVMe and 4TB on SATA, with both the smallest and the largest drives priced above the middle), and the channels currently disagree (mainstream NAND contract prices are forecast up 10 to 15% quarter over quarter in 3Q26 while the 512Gb TLC wafer spot price sits about 8% below its March 2026 peak). The page also separates SLC NAND and NOR Flash, whose 2026 contract moves of 100 to 150% are routinely and wrongly quoted as the price of consumer SSDs; the flash in an ordinary SSD is 3D TLC or QLC on a different contract line. Machine-readable data at https://capitalandcompute.net/ssd-prices.json. - [GPU price tracker](https://capitalandcompute.net/gpu-prices/): what Nvidia and AMD graphics cards cost against official MSRP in 2026, the memory-driven kit-price hikes behind the gap modelled as a compounding cascade, AMD's own MSRP increases, the GDDR6 spot price move, and a buy-or-wait verdict by situation. Spokes: https://capitalandcompute.net/gpu-prices/nvidia/ (the full RTX 50 lineup against MSRP), https://capitalandcompute.net/gpu-prices/amd/ (the Radeon RX 9000 lineup), and https://capitalandcompute.net/gpu-prices/vram/ (what the memory on a card actually costs and how much of the price it is). Deliberately analytical rather than a live per-SKU feed, because no free continuously scrapeable GPU price series exists. - [Hosting price hike tracker](https://capitalandcompute.net/hosting-price-hikes/): every announced VPS, cloud and dedicated server price increase of 2026 from Hetzner and OVHcloud, recorded from each provider's own announcement or published price table and stamped with the date it was verified. For each change: the announcement date, the effective date, the exact product scope, the low and high end of the published percentage range, and the pricing channel it applies to. The channel field is the point: one announcement can contain a list price for new orders, a billed price for existing customers at renewal, an option price for memory sold as a line item, and a one-off setup fee, and those are not comparable numbers. Two findings a keyword-matched summary will miss: the widely quoted OVHcloud figures of 9 to 11 percent and up to 87 percent belong to two different announcements five months apart rather than being competing readings of one event, and the increases track memory share rather than compute, so Hetzner Arm CAX plans rose about 31 to 33 percent while its x86 CPX and CCX plans rose 94 to 192 percent in the same adjustment. Hetzner publishes before-and-after prices but no percentages, so those percentages are computed here and labelled as such. Also carries OVHcloud's disclosed component purchasing multiples versus June 2025 (RAM 6x observed, 9x and 12x forecast, SSD 3.23x, hard drives 1.48x). Machine-readable data at https://capitalandcompute.net/hosting-price-hikes.json. - [AI inference providers directory](https://capitalandcompute.net/ai-inference-providers/): 35 AI inference and model providers in 2026, grouped into five categories (first-party model labs, cloud hyperscaler marketplaces, neutral open-weight inference platforms, custom-silicon speed specialists, and GPU/compression niche), each with what it serves, representative per-token and per-GPU-hour pricing, OpenAI-compatibility, free tier, and a link to its official page. Covers OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Cohere, MiniMax, Kimi, Groq, Cerebras, SambaNova, Together, Fireworks, DeepInfra, Replicate, Nebius, Baseten, CoreWeave, Azure, Bedrock and more. Prices are representative July 2026 snapshots; verify at the source. Machine-readable data at https://capitalandcompute.net/ai-inference-providers.json. ## Posts - [Best Local Models for Hermes Agent in 2026](https://capitalandcompute.net/blog/best-local-models-for-hermes-agent/): Hermes Agent rejects any model under 64,000 tokens of context. Which open-weight models clear that bar on 8GB, 16GB, 24GB and 32GB of VRAM. - [GPT-6 Astra: Pricing, Benchmarks, Cost](https://capitalandcompute.net/blog/gpt-6-astra-pricing-benchmarks/): GPT-6 Astra lists $10 and $50 per million tokens, 2.5x GPT-5.6 Sol. On Terminal-Bench it still solves a task for half of what Claude Fable 5.1 costs. - [Hermes Agent vs Vellum: A Neutral 2026 Comparison](https://capitalandcompute.net/blog/hermes-agent-vs-vellum/): Vellum ranks itself first among Hermes Agent alternatives. A third-party read of the two, with four claims on that page checked against the docs. - [Gemini 3.8 Flash and Muse Spark 1.3 Pricing](https://capitalandcompute.net/blog/gemini-3-8-flash-muse-spark-1-3-pricing/): Gemini 3.8 Flash costs $0.75 per million input tokens until December 31, then doubles. Muse Spark 1.3 is $1.25, or $0.10 if Meta can train on you. - [Claude Fable 5.1: Pricing, Benchmarks, Cost](https://capitalandcompute.net/blog/claude-fable-5-1-pricing-benchmarks/): Claude Fable 5.1 holds Fable 5 rates at $10/$50 per million tokens and cuts the cache read to $0.25. Benchmarks, specs, and modeled cost per task. - [New AI Models Released in September 2026: Prices](https://capitalandcompute.net/blog/new-ai-models-september-2026/): Every AI model released in September 2026, with dates, verified per-token prices and a primary source for each. Updated through the month as releases land. - [Hetzner Price Increase 2026: All Three Waves Explained](https://capitalandcompute.net/blog/hetzner-price-increase-2026/): Hetzner raised prices three times in 2026. Here are the derived percentages per plan from its own price table, and why Arm plans barely moved. - [OVHcloud Price Increase 2026: Both Waves, in Full](https://capitalandcompute.net/blog/ovhcloud-price-increase-2026/): OVHcloud raised prices twice in 2026. Why the widely quoted 9 to 11 percent and the 87 percent figure are different events, and what hits at renewal. - [Why Is My VPS Price Going Up? The 2026 Answer](https://capitalandcompute.net/blog/why-is-my-vps-price-going-up/): Hetzner and OVHcloud raised VPS prices repeatedly in 2026. Here is the memory-cost mechanism behind it, and what it actually means at renewal. - [Which AI Coding Model Burns the Fewest Tokens?](https://capitalandcompute.net/blog/ai-coding-agent-token-usage/): Independent measurement of eight frontier models: GPT-5 averages 2.38 file reads per task, Kimi-K2 averages 15.27, and neither can predict its own bill. - [How to Unslop AI Output: Tools for Every Fix](https://capitalandcompute.net/blog/anti-slop-ai-skills/): Your AI output reads generic. Match the fix to the problem: prose skills, design skills, offline linters and gates that block slop before merge. - [Muse Spark 1.2 Pricing: What a Task Really Costs](https://capitalandcompute.net/blog/muse-spark-1-2-pricing-cost-per-task/): Muse Spark 1.2 at $1.25/$4.25 adds a $0.10/$0.20 contributor tier. Cost per task vs Claude Opus 5 and GPT-5.6 Sol, and what the data-for-discount really buys. - [Terminal-Bench 4.0: The $234 Solved Task](https://capitalandcompute.net/blog/terminal-bench-4-0/): Terminal-Bench 4.0 now runs 18 entries and GPT-6 Astra leads at 58.2 percent. Derived cost per solved task runs from 6 to 234 dollars, a 38x spread. - [GLM-5.3-Flash: Price, Specs, and Benchmarks](https://capitalandcompute.net/blog/glm-5-3-flash-pricing-benchmarks/): GLM-5.3-Flash is Ox Alpha, officially: a 320B-A18B MoE model, 1M-token context, MIT license, and $0.15/$0.50 per million tokens. Specs, pricing, and benchmarks. - [ChatGPT Business Premium Seats: $100 List, $120 Floor](https://capitalandcompute.net/blog/chatgpt-business-premium-seats-pricing/): OpenAI priced ChatGPT Business Premium exactly like Anthropic Claude Team Premium, but the 2-seat minimum pushes the real floor to $120, not $100. - [OpenAI Jalapeño vs Nvidia: What the Benchmarks Mean](https://capitalandcompute.net/blog/openai-jalapeno-inference-chip/): OpenAI posted first Jalapeño benchmarks: up to 1.9x more throughput per kilowatt than Nvidia. Real numbers, plus the caveats the headlines are skipping. - [Codistry vs Claude Code: Auditing the Half-Cost Token Claim](https://capitalandcompute.net/blog/codistry-vs-claude-code-token-cost/): Adronite claims Codistry costs 48% less than Claude Code per task. We checked the math in its own benchmark data: arithmetic holds, but three caveats matter. - [Mac mini M6 vs Mac Studio M5 for Local LLMs](https://capitalandcompute.net/blog/mac-mini-m6-local-llm/): Apple refreshed the Mac mini and Mac Studio on August 25 2026. Every memory ceiling stayed flat, and bandwidth became a paid upgrade. - [Is the Plaud Note Worth It? The Full Cost, Priced Honestly](https://capitalandcompute.net/blog/plaud-note-worth-it-cost/): The Plaud Note costs $159 plus a subscription to do anything useful. What it really costs per transcribed hour, who pays too much, and when to skip it. - [How Much Does an AI Agent Cost Per Month?](https://capitalandcompute.net/blog/ai-agent-cost-per-month/): Five meters bill AI agents: per-seat, per-token, per-conversation, per-action and per-minute. The same agent runs 2 to 1,339 dollars a month. - [What Is Ox Alpha? The Free Stealth AI Model, Explained](https://capitalandcompute.net/blog/ox-alpha-stealth-model-explained/): Ox Alpha is a free anonymous reasoning model on OpenRouter. Specs, the GLM-5.3 fingerprint case, benchmark rumors, and who pays for 100T tokens a day. - [What a Personal AI Agent Actually Costs in 2026](https://capitalandcompute.net/blog/personal-ai-agent-cost-2026/): OpenClaw and Hermes Agent are free to download. Running one is not: hosting, tokens, and your time put a real personal agent at $10-300 a month. - [Qwen3.8-27B: What Free Open Weights Do to API Pricing](https://capitalandcompute.net/blog/qwen-3-8-27b-local-economics/): Apache 2.0 weights for Qwen3.8-27B landed on August 14, 2026. Strong vendor scores, hosted tokens at 80 percent off the flagship, and a catch in the local math. - [DeepSeek Harness: What It Is and What It Costs](https://capitalandcompute.net/blog/deepseek-harness/): DeepSeek open sourced its agent harness under MIT on August 13, then repriced the API three days later. What it does and what it costs to run. - [What Is YMTC? China's NAND Flash Challenger](https://capitalandcompute.net/blog/what-is-ymtc/): YMTC now ships more NAND bits than Micron or Kioxia but ranks only fifth by revenue. What that gap means for flash prices in 2026 and 2027. - [Why SSDs Are So Expensive in 2026: The NAND Squeeze](https://capitalandcompute.net/blog/why-are-ssds-so-expensive/): NAND contract prices ran 70 to 75 percent in a single quarter and a 2TB drive doubled. What drove it, and why the surge is already slowing. - [Claude Fable 5.1 Leak: Release Date and Price](https://capitalandcompute.net/blog/claude-fable-5-1-leak-release-date/): Anthropic has published nothing on Fable 5.1. Four leaked claims, two X threads and one gray-scale sighting, graded against the primary sources. - [Grok 4.6 Agent Cost: The 200K Price Cliff](https://capitalandcompute.net/blog/grok-4-6-agent-cost/): Grok 4.6 kept the $2 and $6 headline but raised cached input 67 percent, and doubles every rate past 200K tokens. What that does to an agent bill. - [Memory Capex 2026: $2 Trillion Buys No Relief](https://capitalandcompute.net/blog/memory-capex-buildout-2026/): Samsung, SK hynix and Micron reversed course and committed about 2 trillion dollars to new memory fabs. The first wafers arrive in December 2028. - [AI Programming in 2026: Tools, Techniques, Evidence](https://capitalandcompute.net/blog/state-of-ai-programming-2026/): AI writes 42 percent of committed code, yet the measured productivity gain stays contested. A sourced survey of the tools, techniques and evidence. - [What AI Changes About the Professional Profile](https://capitalandcompute.net/blog/what-ai-changes-about-the-professional-profile/): Three studies measured AI at work and disagreed: 40 percent faster writing, 14 percent more support tickets, 19 percent slower coding. - [New AI Models Released in August 2026: The Full List](https://capitalandcompute.net/blog/new-ai-models-august-2026/): Every new AI model released in August 2026 with dates, verified per-token prices and primary sources, plus the price changes that landed alongside them. - [AI Data Center Stocks: Where the Capex Lands](https://capitalandcompute.net/blog/ai-data-center-stocks/): A layer-by-layer map of AI data center capex: hyperscaler spending, Nvidia revenue, neocloud debt, REIT backlogs and grid orders, from reported figures. - [Grok Bot Pricing: The $120 Door and What It Buys](https://capitalandcompute.net/blog/grok-bot-pricing-explained/): Grok Bot needs a $120, $200 or $300 per month plan while Claude Cowork ships from $20. What the seat price buys and where it breaks even. - [Claude Opus 5 vs Sonnet 5: Which Model Should You Use?](https://capitalandcompute.net/blog/claude-opus-vs-sonnet/): Claude Opus 5 is stronger, but Sonnet 5 costs 40% less at standard rates. Compare coding, reasoning, speed, context, and when upgrading to Opus pays off. - [Uber and Walmart Cap AI Tokens as Enterprise Costs Surge](https://capitalandcompute.net/blog/uber-walmart-ai-token-caps/): Uber and Walmart capped employee AI use after costs surged. Learn what token budgets reveal and how enterprises can measure AI spend, value, and ROI. - [Ex-OpenAI and Ex-Anthropic Startups: The Money](https://capitalandcompute.net/blog/ex-openai-anthropic-startups-funding/): OpenAI and Anthropic alumni raised billions in pre-product seeds in 2026. Verified funding for every spinout, and what the money is buying. - [Qwen3.8 Max Benchmarks: The First Real Numbers](https://capitalandcompute.net/blog/qwen-3-8-max-benchmarks/): Artificial Analysis scored Qwen3.8 Max at 58 on its Intelligence Index. The token price fell 20 percent but the cost to run the same eval rose 64 percent. - [Prime Agent Explained: The 95.5% ARC-AGI-3 Claim](https://capitalandcompute.net/blog/prime-agent-explained/): Prime Intellect reports 95.5% on ARC-AGI-3 with an open-source harness. The ARC-verified ceiling on that same set is 30.2%. What the gap measures. - [DeepSeek V4 Flash: Pricing, Benchmarks, and Cost](https://capitalandcompute.net/blog/deepseek-v4-flash/): DeepSeek V4 Flash 0731 scores 50 on the Artificial Analysis Intelligence Index at $0.14 per million input tokens, ahead of the pricier V4 Pro at 44. - [Why Local LLMs Fail at Agentic Coding in 2026](https://capitalandcompute.net/blog/why-local-llms-fail-agentic-coding/): Identical open weights score 67.8 or 90.0 on SWE-bench depending on the harness. What the research says about where local coding agents break. - [GPU Prices in 2026: Why They Rise, and Whether to Buy](https://capitalandcompute.net/blog/gpu-prices-2026/): Nvidia RTX cards are up 20 to 30 percent in a third 2026 hike. Here is the memory-cost chain behind it, and whether to buy a GPU now or wait. - [CXMT Stock Price: Inside the Record IPO Debut](https://capitalandcompute.net/blog/cxmt-stock-price-ipo/): CXMT closed 466 percent above its offer price on debut, worth 3.28 trillion yuan. What that valuation implies about the DRAM shortage. - [Claude Opus 5: Pricing, Benchmarks, and Cost](https://capitalandcompute.net/blog/claude-opus-5-pricing-benchmarks/): Claude Opus 5 ships at $5/$25 per million tokens with an effort dial, matching Fable 5 intelligence at half the price. Benchmarks, specs, and cost per task. - [Claude Skills vs Agent Skills vs MCP vs Prompts](https://capitalandcompute.net/blog/claude-skills-vs-agent-skills-mcp-prompts/): The four ways to extend an AI coding agent: what a Claude Skill, an agent skill, an MCP server and a prompt library each change, and when to use which. - [LLM Evaluation Metrics: Accuracy, Cost, and CI in 2026](https://capitalandcompute.net/blog/llm-evaluation-metrics/): A practitioner guide to LLM evaluation metrics: reference-based scores, LLM-as-judge, and human review, ranked by accuracy and cost per 10,000 responses. - [Memory Shortage 2027: Will It Really Last a Decade?](https://capitalandcompute.net/blog/memory-shortage-2027-forecast/): SK hynix says 2027 will be the worst supply year ever for memory and ADATA expects a decade-long shortage. Here is how long the RAM squeeze may last. - [What Is an AI Data Center? Cost, Power, and Scale](https://capitalandcompute.net/blog/what-is-an-ai-data-center/): An AI data center explained: how it differs from a traditional data center, what it costs per megawatt, and why power is the binding constraint. - [Will the US Ban Chinese Open-Weight AI Models?](https://capitalandcompute.net/blog/will-us-ban-chinese-ai-models/): Nearly 200 startups urged Trump not to restrict Chinese open-weight AI. Here is what a ban would target, why enforcement is hard, and the stakes. - [OpenAI Agent Escaped Its Sandbox and Breached Hugging Face](https://capitalandcompute.net/blog/openai-agent-hugging-face-breach/): During an internal cyber evaluation, an OpenAI agent broke out of its sandbox through a zero-day and breached Hugging Face. Here is what it means. - [Claude Code Desktop Guide: Setup and Best Practices](https://capitalandcompute.net/blog/claude-code-desktop-guide/): Learn Claude Code Desktop efficiently: setup, permissions, parallel sessions, worktrees, previews, diff review, CLAUDE.md, security, and when to use the CLI. - [Qwen3.8 Max Preview: Release, Access, and Open Weights](https://capitalandcompute.net/blog/qwen-3-8-max-preview/): Qwen3.8 Max Preview is live through the Alibaba Token Plan and Qoder. Here is what is confirmed about access, benchmarks, open weights, and the 2.4T claim. - [When Will Chinese AI Models Beat US Models? 2028–2030](https://capitalandcompute.net/blog/when-will-chinese-ai-models-surpass-us/): Chinese AI models already lead on price and open weights. A six-part scorecard shows why broad leadership is plausible by 2028–2030, but not inevitable. - [Best AI Visibility Tools: Pricing and Citation Tracking](https://capitalandcompute.net/blog/best-ai-visibility-tools/): Compare the real monthly cost of tracking your brand across ChatGPT, Gemini, Perplexity and AI Overviews, plus how to choose the right tool. - [China AI Pricing 2026: DeepSeek, Qwen, Kimi API Costs](https://capitalandcompute.net/blog/china-ai-pricing/): Current 2026 API and subscription prices for DeepSeek, Qwen, Kimi, GLM, and MiniMax, compared per million tokens and against US models. - [Cost Per Token Over Time: The AI Price Collapse](https://capitalandcompute.net/blog/cost-per-token-over-time/): How AI cost per token fell roughly 10x per year since 2021, why it collapsed, and why your per-task bill did not drop nearly as fast. - [Who Owns Moonshot AI? Backers, Valuation, Money](https://capitalandcompute.net/blog/moonshot-ai-explained/): Moonshot AI, the Chinese lab behind Kimi, is founder-led by Yang Zhilin, with Alibaba its largest backer and a reported 20 billion dollar valuation. - [Inkling: Thinking Machines' First Open-Weight Model](https://capitalandcompute.net/blog/inkling-thinking-machines-launch/): Thinking Machines released Inkling, a 975B open-weight model under Apache 2.0. Full pricing, benchmarks against GLM-5.2 and DeepSeek V4, and the verdict. - [OpenRouter AI Explained: Pricing, Privacy and Tradeoffs](https://capitalandcompute.net/blog/openrouter-ai/): OpenRouter AI gives developers one API for 400+ models. See its real fees, privacy controls, pros, cons, and when going direct costs less. - [What Does CoreWeave Do? AI Cloud Business Explained](https://capitalandcompute.net/blog/what-does-coreweave-do/): What does CoreWeave do? See how its GPU cloud works, who uses it, how it makes money, its NVIDIA relationship, and the risks behind its rapid growth. - [What Is CXMT? China's DRAM Manufacturer Explained](https://capitalandcompute.net/blog/what-is-cxmt/): CXMT is the largest DRAM maker in China, now the number four global supplier, closing in on Micron on capacity while its HBM chips stay years behind. - [AI Data Center Financing 2026: Inside the $700B Buildout](https://capitalandcompute.net/blog/ai-data-center-financing-2026/): How AI infrastructure gets financed in 2026: Helix (KKR/Nvidia), Apollo/Blackstone (Anthropic), Stargate (OpenAI), and what this means for model pricing. - [Best Open-Weight AI Models in 2026](https://capitalandcompute.net/blog/best-open-weight-ai-models-2026/): The best open-weight AI models in 2026, ranked by use case: coding, long context, multimodal, on-device, and the real cost per finished task. - [GitHub Pricing 2026: Free, Team & Enterprise Costs](https://capitalandcompute.net/blog/github-pricing-plans-cost/): GitHub pricing starts at $0. Compare Free, Pro, Team, and Enterprise costs, included usage, add-on fees, and which plan fits your team in 2026. - [How to Save Money on RAM in 2026: 7 Strategies](https://capitalandcompute.net/blog/how-to-save-money-on-ram-2026/): RAM prices roughly tripled in 2026. Here are 7 ways to save money buying memory: right-size capacity, buy used, find bundle deals, time the market, and more. - [Kimi K3: Pricing, Specs, and Benchmarks](https://capitalandcompute.net/blog/kimi-k3-explained/): Kimi K3 topped the WebDev coding arena at a fraction of frontier prices. Full specs, pricing, benchmarks, and the Moonshot business, verified July 2026. - [Who Is Matei Zaharia? Spark, Databricks and AI](https://capitalandcompute.net/blog/matei-zaharia/): Who is Matei Zaharia? The Apache Spark creator and Databricks co-founder also helped build MLflow. Here is how his work shaped modern data and AI systems. - [OpenAI Hardware: The Economics of Its AI Speaker](https://capitalandcompute.net/blog/openai-hardware-bet-economics/): OpenAI is spending nearly 6.5 billion dollars to build a screenless AI smart speaker. The business logic, the running costs, and the Apple lawsuit risk. - [Why Is My AI API Bill So High? 8 Causes and Fixes](https://capitalandcompute.net/blog/why-is-my-ai-api-bill-so-high/): Unexpected OpenAI or Claude API bill? Trace the cost to context growth, cache misses, reasoning tokens, tool calls, retries, or a runaway agent. - [Is the 2026 Memory Cycle Peaking? Signals to Watch](https://capitalandcompute.net/blog/memory-cycle-peaking-2026/): Consumer DDR5 has gone flat while contract and HBM prices keep climbing. Here are the signals that show whether the 2026 memory cycle is peaking. - [Build Your Own Agent Harness or Buy Claude Code?](https://capitalandcompute.net/blog/build-vs-buy-agent-harness/): Should you build your own AI agent harness or adopt Claude Code and Codex? A cost and control decision guide: adopt, extend, or build. - [DBOS Explained: Durable Execution for Agent Harnesses](https://capitalandcompute.net/blog/dbos-durable-execution-agent-harness/): DBOS is an open-source durable execution library that keeps AI agent harnesses running through crashes, restarts, and deploys. What it is and when to use it. - [Harbor-Index: The AI Benchmark Where Nothing Tops 30%](https://capitalandcompute.net/blog/harbor-index-explained/): Harbor-Index is a new AI agent benchmark where no model scores above 30 percent. How its 82 tasks were built, what the leaderboard costs, and why it matters. - [GPT-5.6 vs Claude Opus 4.8 vs Fable 5: Which to Pick](https://capitalandcompute.net/blog/gpt-5-6-vs-claude-opus-4-8-vs-fable-5/): OpenAI GPT-5.6 Sol, Terra and Luna versus Anthropic Claude Opus 4.8 and Fable 5: price, benchmarks and which model to pick in 2026. - [Memory Shortage 2026: Causes and When It Ends](https://capitalandcompute.net/blog/memory-shortage-2026-explained/): The 2026 memory shortage explained: why AI and HBM demand starved consumer DRAM, how high prices have climbed, and when the RAM shortage is likely to end. - [AI Training Costs 2026: GPT-5.6, Claude Fable 5, Gemini](https://capitalandcompute.net/blog/what-it-costs-to-train-ai-models-2026/): Training a frontier model costs $1B+. Breakdown of compute, energy, data, and R&D costs across the three leading labs, and what they mean for API prices. - [Best Claude Code Agentic Workflows: 2026 Playbook](https://capitalandcompute.net/blog/claude-code-agentic-workflows/): Six repeatable Claude Code workflows for planning, subagent research, parallel worktrees, verification, and slash-command pipelines, with when to use each. - [New AI Models Released in July 2026: The Full List](https://capitalandcompute.net/blog/new-ai-models-july-2026/): Every new AI model released in July 2026 with dates, per-token prices and primary sources, plus the announced models still ahead on the release timeline. - [Best Cursor Alternatives 2026, Ranked by Real Cost](https://capitalandcompute.net/blog/best-cursor-alternatives-2026/): Cursor got pricier and its owner changed. Compare the best 2026 alternatives ranked by real cost per task, not sticker price. - [GPT-5.6 Sol Tops the Coding Leaderboard: At What Cost?](https://capitalandcompute.net/blog/gpt-5-6-sol-coding-leaderboard-cost/): GPT-5.6 Sol leads the Artificial Analysis coding leaderboard, and since the August 21 price cut to $4/$20 it does so at a third of Fable 5 per-token cost. - [Meta Muse Spark 1.1: A Cheap Agentic Coding Bet](https://capitalandcompute.net/blog/meta-muse-spark-1-1/): Meta launched Muse Spark 1.1 at 1.25 and 4.25 dollars per million tokens to chase Anthropic and OpenAI. Pricing, benchmarks, and the honest verdict. - [Grok 4.5 Cost Per Task: The 4.2x Efficiency Test](https://capitalandcompute.net/blog/grok-4-5-cost-per-task/): Grok 4.5 launched July 8 at $2 and $6 per million tokens with a 4.2x token-efficiency claim. Does it really cost less per task than Claude Opus 4.8? - [AI Search Broke the Web's Business Model. What Replaces It.](https://capitalandcompute.net/blog/ai-search-disruption-economics/): 68% of searches end without a click. AI Overviews cut publisher traffic 40%. What replaces the click economy in the citation era. - [OpenCode vs Claude Code vs OpenClaude: Benchmarks](https://capitalandcompute.net/blog/opencode-vs-claude-code-vs-openclaude/): Only one of these three coding agents shows up on a benchmark leaderboard. The Terminal-Bench 2.1 numbers, the missing scores, and how to choose. - [The Publisher's Revenue Playbook for the AI Search Era](https://capitalandcompute.net/blog/publisher-revenue-playbook-ai-search-era/): AI Overviews cut publisher clicks 40%. Six revenue strategies publishers are deploying now, from per-query monetization to session-level optimization. - [Best Models for Hermes Agent in 2026: Cost and Value](https://capitalandcompute.net/blog/best-models-for-hermes-agent/): Hermes Agent still defaults to GLM-5.2. Verified September 2026 prices and coding scores show a model at a ninth of the cost now scoring higher. - [Hermes Agent Explained: Why the Harness Beats the Model](https://capitalandcompute.net/blog/hermes-agent-explained/): What Hermes Agent is in 2026: an MIT-licensed agent with a learning loop, 300-plus models and a v0.21.0 provider wave. Why the harness beats the model. - [How Much Does It Cost to Run an AI Agent in Production?](https://capitalandcompute.net/blog/ai-agent-production-cost/): Published ranges for AI agent costs disagree by 10x. Here is the actual formula, modeled against real 2026 API rates, so you can price your own workload. - [AI Video Models in 2026: Sora Out, China Leads](https://capitalandcompute.net/blog/state-of-ai-video-models-2026/): OpenAI is shutting Sora down while Chinese models top the independent video leaderboards on quality and cost. The state of AI video, July 2026. - [Agent Arena: The AI Agent Leaderboard Explained](https://capitalandcompute.net/blog/agent-arena-leaderboard-explained/): Arena, formerly LMArena, ranks AI agents from over a million real sessions using causal tracing, not style votes. How Agent Arena works and who leads. - [Why Are Chinese AI Models So Cheap? The Real Economics](https://capitalandcompute.net/blog/why-are-chinese-ai-models-so-cheap/): Chinese AI models list output tokens up to 57x below US flagships. The verified economics of efficient training, cheap power and open weights as strategy. - [The Only MCP Servers Worth Installing in 2026](https://capitalandcompute.net/blog/best-mcp-servers/): Most MCP server lists are directory dumps. The tested consensus is five servers, a strict tool budget, and three catalogs worth bookmarking. - [Claude Code Harness Guide: Skills, Hooks, Subagents, MCP](https://capitalandcompute.net/blog/claude-code-harness-guide/): A 2026 engineering guide to the Claude Code harness: CLAUDE.md, skills, hooks, subagents, MCP and plugins, with the context cost of each layer. - [Harness Engineering in 2026: Techniques Beyond MCP](https://capitalandcompute.net/blog/harness-engineering-techniques-2026/): The named harness engineering techniques of 2026: ratchet rules, Ralph loops, spec-driven development and evaluator agents, with the cost of each. - [New Agentic Code Editors in 2026: ZCode, Antigravity 2.0](https://capitalandcompute.net/blog/new-agentic-code-editors-2026/): ZCode, Antigravity 2.0, Grok Build, Warp, and Zed all changed in 2026. Here is what actually shipped in each new agentic code editor, and what it costs. - [Benchmark Saturation: Why 99% on MMLU Means Almost Nothing](https://capitalandcompute.net/blog/benchmark-saturation-mmlu/): Benchmark saturation is when top AI models bunch so close to the ceiling a test cannot rank them. Why MMLU hit this wall, and the harder replacement is next. - [Claude Fable 5 Is Back: Pricing and Benchmarks](https://capitalandcompute.net/blog/claude-fable-5-returns-pricing-benchmarks/): Claude Fable 5 returns after a US export-control suspension. Here is the $10/$50 pricing, the SWE-Bench Pro and FrontierCode scores, and what it means. - [Claude Sonnet 5: Pricing, Benchmarks, and Cost](https://capitalandcompute.net/blog/claude-sonnet-5-pricing-benchmarks/): Claude Sonnet 5 ships at $3/$15 per million tokens, intro $2/$10 through August 2026. What changed versus Sonnet 4.6 and Opus 4.8, and the cost per task. - [Kiro Pricing Explained: How Far $20 a Month Goes](https://capitalandcompute.net/blog/kiro-pricing-explained-free-tier-credits/): Kiro gives a free tier of 50 credits and Pro at $20 a month for 1,000. What a credit actually buys, the Auto versus pinned-model savings, and how overage works. - [Windsurf Is Now Devin Desktop: What It Actually Costs](https://capitalandcompute.net/blog/windsurf-devin-desktop-rebrand-cost/): Windsurf became Devin Desktop on June 2, 2026, pricing unchanged. Cascade retires today. The real cost shift is what got bundled into the same $20 plan. - [Hidden Cost of AI-Generated Code: Beyond the Token Price](https://capitalandcompute.net/blog/hidden-cost-ai-generated-code/): Your Claude Code session cost $47 in tokens. The real cost was $470. This is the math your dashboard does not show. - [When Will RAM Prices Drop? 2026-2028 Forecast](https://capitalandcompute.net/blog/when-will-ram-prices-drop/): RAM prices are still climbing in late 2026, but the rate is slowing. Here is what each forecaster now says about 2027 and 2028, and whether to buy or wait. - [What Claude Code's /cost Command Shows](https://capitalandcompute.net/blog/claude-code-cost-command-explained/): How to read the Claude Code /cost command: session spend, token breakdown, why Pro and Max differ from API billing, and when to use /usage instead. - [Cursor Pricing Limits Explained: How Far $20 Goes](https://capitalandcompute.net/blog/cursor-pricing-limits-explained/): What the $20 Cursor Pro plan really buys in 2026: roughly 225 to 650 requests depending on model, how the usage pool works, and when overages start. - [DeepSWE vs FrontierCode: Two Ways to Grade AI Code](https://capitalandcompute.net/blog/deepswe-vs-frontiercode/): DeepSWE grades whether an AI finishes the task. FrontierCode grades whether you would merge its code. Why the same model scores 59% and 13%. - [Vibe Coding Cost Economics 2026: The Real Math](https://capitalandcompute.net/blog/vibe-coding-cost-economics-2026/): Vibe coding tools cost $20-$100/month. The real cost including tokens, infrastructure, and technical debt runs $87-$340/month. Here is the math. - [Did OpenAI Cut Prices? GPT-5.6 Sol, Terra, Luna](https://capitalandcompute.net/blog/did-openai-cut-ai-prices-gpt-5-6/): OpenAI cut GPT-5.6 API prices on July 30: Luna by 80 percent, Terra by 20, Sol not at all. On August 21 the flagship followed with a promotional cut to $4/$20. - [GPT-5.6 Sol, Terra and Luna: What OpenAI Launched](https://capitalandcompute.net/blog/gpt-5-6-sol-terra-luna-launch/): OpenAI launched GPT-5.6 on June 26 as three tiers: Sol, Terra and Luna. Here are the confirmed prices, the benchmarks, and the government access catch. - [Apple Just Raised Prices. What It Means for Local AI](https://capitalandcompute.net/blog/apple-price-hike-local-ai/): Apple raised Mac and iPad prices on June 25 2026 as the memory shortage bit. Here is what the unified-memory tax means for running local AI. - [GPT-5.6 Delayed: The Government Is Now a Release Gate](https://capitalandcompute.net/blog/gpt-5-6-delayed-government-ai-release-gate/): The Trump administration asked OpenAI to stagger GPT-5.6 and approve users one by one. A federal gate now sits between a finished AI model and the public. - [How Much RAM to Run a Local LLM? 2026 Sizing Guide](https://capitalandcompute.net/blog/how-much-ram-to-run-a-local-llm/): How much RAM a local LLM really needs: measured Q4_K_M file sizes, what 8GB to 512GB runs in 2026, and why long context can cost more than the weights. - [Which Countries Use AI Most? The Economics Behind the Map](https://capitalandcompute.net/blog/ai-adoption-by-country-economics/): AI usage tracks national income, but the standouts break the rule. The economics of why Israel and Singapore use AI far more than their wealth predicts. - [Benchmark Contamination in AI: When Tests Leak Into Training](https://capitalandcompute.net/blog/ai-benchmark-contamination-explained/): Benchmark contamination is when test answers leak into training data and inflate AI scores. How it happens, how much it distorts results, and how to spot it. - [GPT-5.6 Preview: The Leak Is Real, the Release Date Isn't](https://capitalandcompute.net/blog/gpt-5-6-preview-leak-release-date/): A leaked gpt-5.6-preview route lit up r/OpenAI, then GPT-5.6 launched June 26 as Sol, Terra and Luna in a government-gated preview. The signal vs the hype. - [Why Is RAM So Expensive in 2026? The Shortage Is a Choice](https://capitalandcompute.net/blog/why-is-ram-so-expensive-ai-memory-shortage/): RAM keeps getting pricier in 2026 because the memory giants profit more from scarcity than supply. Inside the most lucrative shortage in chip history. - [Are AI Benchmarks Reliable? How the Scores Get Gamed](https://capitalandcompute.net/blog/are-ai-benchmarks-reliable/): AI benchmark scores get gamed by contamination, saturation, and cheating. The 2026 receipts, a trust scorecard, and how to read past any leaderboard. - [China's AI Chip Companies in 2026: Replacing NVIDIA](https://capitalandcompute.net/blog/china-ai-chip-companies-2026/): Seven Chinese firms now ship AI accelerators, the best near NVIDIA H100 class. A fact-checked 2026 map of who makes China's GPUs and what is real. - [SpaceX Bought Cursor: Your Bill, Your Code, Cost to Leave](https://capitalandcompute.net/blog/spacex-cursor-acquisition-cost-and-alternatives/): SpaceX is buying Cursor for $60B. What changes for your bill, whether your code now trains xAI models, and the real cost per task of switching away. - [Decentralized GPU vs Cloud: Real Cost Per Hour (2026)](https://capitalandcompute.net/blog/decentralized-gpu-cost-vs-cloud/): Is decentralized GPU compute (Akash, io.net, Render) cheaper than AWS? A grounded 2026 cost-per-hour comparison, plus the caveats the hype skips. - [Local LLM Tokenomics: Self-Hosted Cost Per Token (2026)](https://capitalandcompute.net/blog/self-hosted-llm-cost-per-token/): What a self-hosted LLM token really costs in 2026: cost per token across owned hardware, why memory bandwidth sets speed, and where buying beats the API. - [Claude Fable 5 vs Opus 4.8: Is 2x the Price Worth It?](https://capitalandcompute.net/blog/claude-fable-5-vs-opus-4-8-cost-per-task/): Claude Fable 5 costs exactly 2x Opus 4.8 per token: $10/$50 vs $5/$25. Whether it is cheaper per task depends on loop count, not the sticker rate. - [Claude Code vs Codex (2026): Cost and Capability](https://capitalandcompute.net/blog/claude-code-vs-codex-cost-capability/): Codex leads Claude Code 83.4% to 78.9% on Terminal-Bench 2.1, but the prices match and the cheaper agent flips with your harness. The honest scorecard. - [Cursor Composer 2.5: What It Actually Costs](https://capitalandcompute.net/blog/cursor-composer-2-5-real-cost/): Composer 2.5 finishes a coding task for about $0.07, 10-60x under Claude Opus and GPT-5.5 at near-equal benchmark scores. What the cheap headline leaves out. - [The Price Reversal Phenomenon: When Cheaper AI Costs More](https://capitalandcompute.net/blog/price-reversal-why-cheaper-ai-models-cost-more/): A 2026 Microsoft Research preprint found the cheaper-per-token AI model cost more to finish the job in 32% of model pairs. Why the sticker misleads. - [Do Claude Code Token-Saving Tools Actually Cut Your Bill?](https://capitalandcompute.net/blog/claude-code-token-saving-tools-rtk-headroom-caveman/): An independent replay of 500 Claude Code sessions found rtk, headroom, and caveman cut a $926 bill by just 3.7 percent. Here is why the 60-90% claims miss. - [GPT-5.6 Pricing: What to Expect Per Task](https://capitalandcompute.net/blog/gpt-5-6-pricing-cost-per-task/): GPT-5.6 launched June 26 2026 as Sol, Terra and Luna. Sol held GPT-5.5 pricing at $5/$30 until an August 21 cut to $4/$20. Here is the real cost per task. - [Open-Source LLMs Overtook Proprietary in 2026](https://capitalandcompute.net/blog/open-source-llms-overtake-2026/): Token data from OpenRouter shows open-source LLMs passing proprietary models in mid-2026, a roughly 60/40 flip. The daily breakdown by AI lab. - [Why Local LLMs Got Good in 2026: Capability & Cost](https://capitalandcompute.net/blog/local-llms-good-2026/): Open-weights LLMs crossed from toy to useful in 2026. What actually changed, and the cost math for when running a model yourself beats paying an API. - [World Models: Why AI's Biggest Names Bet Billions in 2026](https://capitalandcompute.net/blog/world-model-funding-wave-2026/): Lin Junyang, Fei-Fei Li, and Yann LeCun are raising billions for world models in 2026. What a world model is, who is funding it, and why the money is moving. - [AI Agent Benchmarks in 2026: What the Scores Actually Mean](https://capitalandcompute.net/blog/ai-agent-benchmarks-2026/): AI agent benchmarks broke in 2026: reward-hacked to 100%, SWE-bench Verified retired, scores swung by harness choice. What each measures and which to trust. - [GitHub Copilot Pricing 2026: What AI Credits Really Cost](https://capitalandcompute.net/blog/github-copilot-pricing-ai-credits/): GitHub Copilot switched to usage-based AI Credits on June 1, 2026. What the $10 Pro plan really costs once metering kicks in, and whether it is still worth it. - [Claude Code Pricing 2026: What It Really Costs Per Task](https://capitalandcompute.net/blog/claude-code-pricing-cost-per-task/): Claude Code plans run $20 to $200 a month, but the real number is cost per task. A modeled breakdown of token spend, where it wins, and where it burns money. - [The 2026 AI Coding Agent Landscape: Leaders, Costs, Harness](https://capitalandcompute.net/blog/the-2026-ai-coding-agent-landscape/): A grounded survey of the 2026 AI coding agent field: Claude Code, Cursor, Copilot, Codex and Antigravity, by interface, cost, and why the harness matters. ## Browse - [Start here](https://capitalandcompute.net/start-here/): orientation page. What this publication covers, which datasets and tools exist, and where to enter each topic. - [All posts](https://capitalandcompute.net/blog/): full article index. - [RSS feed](https://capitalandcompute.net/rss.xml): subscribe to new posts. ## Work with us - [Submit or correct a directory entry](https://capitalandcompute.net/submit/): Capital & Compute accepts free editorial submissions of AI tools, models, inference providers, benchmarks, and agent skills for listing in its directories, and accepts corrections to outdated prices or wrong specifications on existing entries. Listings are free and cannot be bought; every entry must be verifiable against an official page the vendor publishes. This is the route for a founder or vendor who wants their product covered. - [Sponsorship and advertising](https://capitalandcompute.net/sponsor/): display sponsorship, clearly labelled sponsored placement, and newsletter sponsorship are available, quoted per enquiry rather than published as a rate card. Sponsors buy visibility and never buy rank: no ranking position, listing, or review outcome is for sale, and every paid link carries rel="sponsored". - Not accepted: guest posts and paid link placements. The site publishes anonymously under its brand rather than individual bylines, so there is no byline to offer, and it does not sell links that pass ranking signals. ## About and policies - [About](https://capitalandcompute.net/about/): what this publication is and how it is run. - [Editorial standards](https://capitalandcompute.net/editorial-standards/): how content is researched, sourced, and verified. - [Contact](https://capitalandcompute.net/contact/): how to reach the editors. - [Privacy policy](https://capitalandcompute.net/privacy/) - [Terms](https://capitalandcompute.net/terms/) - [Disclaimer](https://capitalandcompute.net/disclaimer/): editorial content, not financial advice.