# Capital & Compute > Where money meets compute. What AI coding agents really cost, which local LLMs are worth running, and what the benchmarks actually mean. Source-backed analysis for builders, no hype. This is a static, independently published blog. Content is editorial and analytical: honest pricing comparisons, agent-harness deep dives, and crypto-for-builders coverage without the cheerleading. All pages are crawlable; there is no paywall and no login. Cite the canonical URLs below (each uses a trailing slash). ## Pillars - [AI subscription pricing](https://capitalandcompute.net/ai-pricing/): grounded, source-verified pricing for the cheapest paid AI coding plan per provider (Claude Code, Copilot, Cursor, Codex, GLM, Kimi, Qwen, and others). Prices are checked against official provider pages. - [AI model value leaderboard](https://capitalandcompute.net/ai-model-leaderboard/): the top LLMs ranked by independent benchmark score AND by value (coding and intelligence points per dollar of tokens). Machine-readable data at https://capitalandcompute.net/ai-model-leaderboard.json. Benchmark composites are sourced from an independent dataset; per-token prices are source-verified. - [AI benchmark directory](https://capitalandcompute.net/ai-benchmarks/): 116 AI and LLM benchmarks across 12 categories (coding and software-engineering agents, agents and tool use, reasoning, math, knowledge, instruction following and multilingual, long context, multimodal, domain and professional work, human preference, safety and factuality, security and dangerous capability). For each benchmark: what it measures, who built it, the release year, the task format, the scoring metric, whether it still separates frontier models or has saturated, a dated representative top score, and a link to its primary source and official leaderboard. Covers SWE-bench and its variants, Terminal-Bench, ARC-AGI-1/2/3, GPQA Diamond, MMLU and MMLU-Pro, Humanity's Last Exam, FrontierMath, LMArena, BFCL, tau-bench, GAIA, OSWorld 2.0, HealthBench, LegalBench, LiveBench, METR Time Horizon, the Epoch Capabilities Index and more. 27 benchmarks additionally have a full in-depth page of their own at https://capitalandcompute.net/ai-benchmarks//, covering how the benchmark works, its version history and saturation trajectory, what its score fails to capture, and how to read it when a lab quotes it. Machine-readable data at https://capitalandcompute.net/ai-benchmarks.json. - [Best free AI model APIs](https://capitalandcompute.net/free-ai-models/): the AI models you can call at $0 today (OpenRouter, Google AI Studio, OpenCode Zen, Groq, Cerebras, Cloudflare, NVIDIA), filtered to the genuinely good ones and ranked on one independent benchmark that scores them all, the Artificial Analysis Intelligence Index, with agentic-coding (Terminal-Bench) and SciCode scores, context windows, open-weight status, and each provider's free-tier limits and catches. The unique angle: not just which models are free, but how good each free model actually is. Machine-readable data at https://capitalandcompute.net/free-ai-models.json. - [AI model releases by month](https://capitalandcompute.net/ai-model-releases/): a dated chronological record of 20 AI model releases grouped by the calendar month they shipped, currently covering July 2026. For each release: the exact release date read off the lab's own announcement, the standard per-token price the model launched at (in USD per million input and output tokens, frozen at release day rather than updated on repricing), the context window, open-weight status and license, the Artificial Analysis Intelligence Index where one was published, the model category, and the primary source URL with the date it was verified. This is the launch-price record; current per-token rates and upcoming models live at https://capitalandcompute.net/ai-models/. Unlike auto-ingested release catalogs, fine-tunes, quantizations, re-hosted copies and catalog-listing dates are excluded, and any row resting on reporting rather than a lab announcement is explicitly labelled. Machine-readable data at https://capitalandcompute.net/ai-model-releases.json. - [AI inference providers directory](https://capitalandcompute.net/ai-inference-providers/): 35 AI inference and model providers in 2026, grouped into five categories (first-party model labs, cloud hyperscaler marketplaces, neutral open-weight inference platforms, custom-silicon speed specialists, and GPU/compression niche), each with what it serves, representative per-token and per-GPU-hour pricing, OpenAI-compatibility, free tier, and a link to its official page. Covers OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Cohere, MiniMax, Kimi, Groq, Cerebras, SambaNova, Together, Fireworks, DeepInfra, Replicate, Nebius, Baseten, CoreWeave, Azure, Bedrock and more. Prices are representative July 2026 snapshots; verify at the source. Machine-readable data at https://capitalandcompute.net/ai-inference-providers.json. ## Posts - [GPU Prices in 2026: Why They Rise, and Whether to Buy](https://capitalandcompute.net/blog/gpu-prices-2026/): Nvidia RTX cards are up 20 to 30 percent in a third 2026 hike. Here is the memory-cost chain behind it, and whether to buy a GPU now or wait. - [CXMT Stock Price: Inside the Record IPO Debut](https://capitalandcompute.net/blog/cxmt-stock-price-ipo/): CXMT closed 466 percent above its offer price on debut, worth 3.28 trillion yuan. What that valuation implies about the DRAM shortage. - [Claude Opus 5: Pricing, Benchmarks, and Cost](https://capitalandcompute.net/blog/claude-opus-5-pricing-benchmarks/): Claude Opus 5 ships at $5/$25 per million tokens with an effort dial, matching Fable 5 intelligence at half the price. Benchmarks, specs, and cost per task. - [Claude Skills vs Agent Skills vs MCP vs Prompts](https://capitalandcompute.net/blog/claude-skills-vs-agent-skills-mcp-prompts/): The four ways to extend an AI coding agent: what a Claude Skill, an agent skill, an MCP server and a prompt library each change, and when to use which. - [LLM Evaluation Metrics: Accuracy, Cost, and CI in 2026](https://capitalandcompute.net/blog/llm-evaluation-metrics/): A practitioner guide to LLM evaluation metrics: reference-based scores, LLM-as-judge, and human review, ranked by accuracy and cost per 10,000 responses. - [Memory Shortage 2027: Will It Really Last a Decade?](https://capitalandcompute.net/blog/memory-shortage-2027-forecast/): SK hynix says 2027 will be the worst supply year ever for memory and ADATA expects a decade-long shortage. Here is how long the RAM squeeze may last. - [What Is an AI Data Center? Cost, Power, and Scale](https://capitalandcompute.net/blog/what-is-an-ai-data-center/): An AI data center explained: how it differs from a traditional data center, what it costs per megawatt, and why power is the binding constraint. - [Will the US Ban Chinese Open-Weight AI Models?](https://capitalandcompute.net/blog/will-us-ban-chinese-ai-models/): Nearly 200 startups urged Trump not to restrict Chinese open-weight AI. Here is what a ban would target, why enforcement is hard, and the stakes. - [OpenAI Agent Escaped Its Sandbox and Breached Hugging Face](https://capitalandcompute.net/blog/openai-agent-hugging-face-breach/): During an internal cyber evaluation, an OpenAI agent broke out of its sandbox through a zero-day and breached Hugging Face. Here is what it means. - [Claude Code Desktop Guide: Setup and Best Practices](https://capitalandcompute.net/blog/claude-code-desktop-guide/): Learn Claude Code Desktop efficiently: setup, permissions, parallel sessions, worktrees, previews, diff review, CLAUDE.md, security, and when to use the CLI. - [Qwen3.8 Max Preview: Release, Access, and Open Weights](https://capitalandcompute.net/blog/qwen-3-8-max-preview/): Qwen3.8 Max Preview is live through the Alibaba Token Plan and Qoder. Here is what is confirmed about access, benchmarks, open weights, and the 2.4T claim. - [When Will Chinese AI Models Beat US Models? 2028–2030](https://capitalandcompute.net/blog/when-will-chinese-ai-models-surpass-us/): Chinese AI models already lead on price and open weights. A six-part scorecard shows why broad leadership is plausible by 2028–2030, but not inevitable. - [Best AI Visibility Tools: Pricing and Citation Tracking](https://capitalandcompute.net/blog/best-ai-visibility-tools/): Compare the real monthly cost of tracking your brand across ChatGPT, Gemini, Perplexity and AI Overviews, plus how to choose the right tool. - [China AI Pricing 2026: DeepSeek, Qwen, Kimi API Costs](https://capitalandcompute.net/blog/china-ai-pricing/): Current 2026 API and subscription prices for DeepSeek, Qwen, Kimi, GLM, and MiniMax, compared per million tokens and against US models. - [Cost Per Token Over Time: The AI Price Collapse](https://capitalandcompute.net/blog/cost-per-token-over-time/): How AI cost per token fell roughly 10x per year since 2021, why it collapsed, and why your per-task bill did not drop nearly as fast. - [How to Use the Kimi K3 API: Access, Code, Limits](https://capitalandcompute.net/blog/how-to-use-kimi-k3-api/): Call Kimi K3 through Moonshot OpenAI-compatible endpoint or OpenRouter. Get a key, set the base URL, use model kimi-k3, and mind the max reasoning setting. - [Kimi K3 vs DeepSeek V4: Price, Specs, Benchmarks](https://capitalandcompute.net/blog/kimi-k3-vs-deepseek-v4/): Kimi K3 and DeepSeek V4 are both open-weight Chinese models. K3 edges the AA intelligence index; DeepSeek V4 costs a fraction and leads on SWE-bench. - [Who Owns Moonshot AI? Backers, Valuation, Money](https://capitalandcompute.net/blog/moonshot-ai-explained/): Moonshot AI, the Chinese lab behind Kimi, is founder-led by Yang Zhilin, with Alibaba its largest backer and a reported 20 billion dollar valuation. - [Inkling: Thinking Machines' First Open-Weight Model](https://capitalandcompute.net/blog/inkling-thinking-machines-launch/): Thinking Machines released Inkling, a 975B open-weight model under Apache 2.0. Full pricing, benchmarks against GLM-5.2 and DeepSeek V4, and the verdict. - [OpenRouter AI Explained: Pricing, Privacy and Tradeoffs](https://capitalandcompute.net/blog/openrouter-ai/): OpenRouter AI gives developers one API for 400+ models. See its real fees, privacy controls, pros, cons, and when going direct costs less. - [What Does CoreWeave Do? AI Cloud Business Explained](https://capitalandcompute.net/blog/what-does-coreweave-do/): What does CoreWeave do? See how its GPU cloud works, who uses it, how it makes money, its NVIDIA relationship, and the risks behind its rapid growth. - [What Is CXMT? China's DRAM Manufacturer Explained](https://capitalandcompute.net/blog/what-is-cxmt/): CXMT is the largest DRAM maker in China, now the number four global supplier, closing in on Micron on capacity while its HBM chips stay years behind. - [AI Data Center Financing 2026: Inside the $700B Buildout](https://capitalandcompute.net/blog/ai-data-center-financing-2026/): How AI infrastructure gets financed in 2026: Helix (KKR/Nvidia), Apollo/Blackstone (Anthropic), Stargate (OpenAI), and what this means for model pricing. - [Best Open-Weight AI Models in 2026](https://capitalandcompute.net/blog/best-open-weight-ai-models-2026/): The best open-weight AI models in 2026, ranked by use case: coding, long context, multimodal, on-device, and the real cost per finished task. - [GitHub Pricing 2026: Free, Team & Enterprise Costs](https://capitalandcompute.net/blog/github-pricing-plans-cost/): GitHub pricing starts at $0. Compare Free, Pro, Team, and Enterprise costs, included usage, add-on fees, and which plan fits your team in 2026. - [How to Save Money on RAM in 2026: 7 Strategies](https://capitalandcompute.net/blog/how-to-save-money-on-ram-2026/): RAM prices roughly tripled in 2026. Here are 7 ways to save money buying memory: right-size capacity, buy used, find bundle deals, time the market, and more. - [Kimi K3: Pricing, Specs, and Benchmarks](https://capitalandcompute.net/blog/kimi-k3-explained/): Kimi K3 topped the WebDev coding arena at a fraction of frontier prices. Full specs, pricing, benchmarks, and the Moonshot business, verified July 2026. - [Who Is Matei Zaharia? Spark, Databricks and AI](https://capitalandcompute.net/blog/matei-zaharia/): Who is Matei Zaharia? The Apache Spark creator and Databricks co-founder also helped build MLflow. Here is how his work shaped modern data and AI systems. - [OpenAI Hardware: The Economics of Its AI Speaker](https://capitalandcompute.net/blog/openai-hardware-bet-economics/): OpenAI is spending nearly 6.5 billion dollars to build a screenless AI smart speaker. The business logic, the running costs, and the Apple lawsuit risk. - [Why Is My AI API Bill So High? 8 Causes and Fixes](https://capitalandcompute.net/blog/why-is-my-ai-api-bill-so-high/): Unexpected OpenAI or Claude API bill? Trace the cost to context growth, cache misses, reasoning tokens, tool calls, retries, or a runaway agent. - [Is the 2026 Memory Cycle Peaking? Signals to Watch](https://capitalandcompute.net/blog/memory-cycle-peaking-2026/): Consumer DDR5 has gone flat while contract and HBM prices keep climbing. Here are the signals that show whether the 2026 memory cycle is peaking. - [Build Your Own Agent Harness or Buy Claude Code?](https://capitalandcompute.net/blog/build-vs-buy-agent-harness/): Should you build your own AI agent harness or adopt Claude Code and Codex? A cost and control decision guide: adopt, extend, or build. - [DBOS Explained: Durable Execution for Agent Harnesses](https://capitalandcompute.net/blog/dbos-durable-execution-agent-harness/): DBOS is an open-source durable execution library that keeps AI agent harnesses running through crashes, restarts, and deploys. What it is and when to use it. - [Harbor-Index: The AI Benchmark Where Nothing Tops 30%](https://capitalandcompute.net/blog/harbor-index-explained/): Harbor-Index is a new AI agent benchmark where no model scores above 30 percent. How its 82 tasks were built, what the leaderboard costs, and why it matters. - [GPT-5.6 vs Claude Opus 4.8 vs Fable 5: Which to Pick](https://capitalandcompute.net/blog/gpt-5-6-vs-claude-opus-4-8-vs-fable-5/): OpenAI GPT-5.6 Sol, Terra and Luna versus Anthropic Claude Opus 4.8 and Fable 5: price, benchmarks and which model to pick in 2026. - [Memory Shortage 2026: Causes and When It Ends](https://capitalandcompute.net/blog/memory-shortage-2026-explained/): The 2026 memory shortage explained: why AI and HBM demand starved consumer DRAM, how high prices have climbed, and when the RAM shortage is likely to end. - [AI Training Costs 2026: GPT-5.6, Claude Fable 5, Gemini](https://capitalandcompute.net/blog/what-it-costs-to-train-ai-models-2026/): Training a frontier model costs $1B+. Breakdown of compute, energy, data, and R&D costs across the three leading labs, and what they mean for API prices. - [Best Claude Code Agentic Workflows: 2026 Playbook](https://capitalandcompute.net/blog/claude-code-agentic-workflows/): Six repeatable Claude Code workflows for planning, subagent research, parallel worktrees, verification, and slash-command pipelines, with when to use each. - [New AI Models Released in July 2026: The Full List](https://capitalandcompute.net/blog/new-ai-models-july-2026/): Every new AI model released in July 2026 with dates, per-token prices and primary sources, plus the announced models still ahead on the release timeline. - [Best Cursor Alternatives 2026, Ranked by Real Cost](https://capitalandcompute.net/blog/best-cursor-alternatives-2026/): Cursor got pricier and its owner changed. Compare the best 2026 alternatives ranked by real cost per task, not sticker price. - [GPT-5.6 Sol Tops the Coding Leaderboard: At What Cost?](https://capitalandcompute.net/blog/gpt-5-6-sol-coding-leaderboard-cost/): GPT-5.6 Sol went GA on July 9 and now leads the Artificial Analysis coding-agents leaderboard, but Claude Fable 5 still edges it on general intelligence. - [Meta Muse Spark 1.1: A Cheap Agentic Coding Bet](https://capitalandcompute.net/blog/meta-muse-spark-1-1/): Meta launched Muse Spark 1.1 at 1.25 and 4.25 dollars per million tokens to chase Anthropic and OpenAI. Pricing, benchmarks, and the honest verdict. - [Grok 4.5 Cost Per Task: The 4.2x Efficiency Test](https://capitalandcompute.net/blog/grok-4-5-cost-per-task/): Grok 4.5 launched July 8 at $2 and $6 per million tokens with a 4.2x token-efficiency claim. Does it really cost less per task than Claude Opus 4.8? - [AI Search Broke the Web's Business Model. What Replaces It.](https://capitalandcompute.net/blog/ai-search-disruption-economics/): 68% of searches end without a click. AI Overviews cut publisher traffic 40%. What replaces the click economy in the citation era. - [OpenCode vs Claude Code vs OpenClaude: Benchmarks](https://capitalandcompute.net/blog/opencode-vs-claude-code-vs-openclaude/): Only one of these three coding agents shows up on a benchmark leaderboard. The Terminal-Bench 2.1 numbers, the missing scores, and how to choose. - [The Publisher's Revenue Playbook for the AI Search Era](https://capitalandcompute.net/blog/publisher-revenue-playbook-ai-search-era/): AI Overviews cut publisher clicks 40%. Six revenue strategies publishers are deploying now, from per-query monetization to session-level optimization. - [Best Models for Hermes Agent: Speed, Cost, Value](https://capitalandcompute.net/blog/best-models-for-hermes-agent/): Hermes Agent runs 300-plus models, so which one should you actually run? A grounded 2026 guide to the best pick for performance, cost, and value. - [Hermes Agent: Why the Harness Beats the Model](https://capitalandcompute.net/blog/hermes-agent-explained/): Nous Research shipped Hermes Agent, an open-source AI agent that learns as it works. Here is what it does and why the harness now beats the model. - [How Much Does It Cost to Run an AI Agent in Production?](https://capitalandcompute.net/blog/ai-agent-production-cost/): Published ranges for AI agent costs disagree by 10x. Here is the actual formula, modeled against real 2026 API rates, so you can price your own workload. - [AI Video Models in 2026: Sora Out, China Leads](https://capitalandcompute.net/blog/state-of-ai-video-models-2026/): OpenAI is shutting Sora down while Chinese models top the independent video leaderboards on quality and cost. The state of AI video, July 2026. - [Agent Arena: The AI Agent Leaderboard Explained](https://capitalandcompute.net/blog/agent-arena-leaderboard-explained/): Arena, formerly LMArena, ranks AI agents from over a million real sessions using causal tracing, not style votes. How Agent Arena works and who leads. - [Why Are Chinese AI Models So Cheap? The Real Economics](https://capitalandcompute.net/blog/why-are-chinese-ai-models-so-cheap/): Chinese AI models list output tokens up to 57x below US flagships. The verified economics of efficient training, cheap power and open weights as strategy. - [The Only MCP Servers Worth Installing in 2026](https://capitalandcompute.net/blog/best-mcp-servers/): Most MCP server lists are directory dumps. The tested consensus is five servers, a strict tool budget, and three catalogs worth bookmarking. - [Claude Code Harness Guide: Skills, Hooks, Subagents, MCP](https://capitalandcompute.net/blog/claude-code-harness-guide/): A 2026 engineering guide to the Claude Code harness: CLAUDE.md, skills, hooks, subagents, MCP and plugins, with the context cost of each layer. - [Harness Engineering in 2026: Techniques Beyond MCP](https://capitalandcompute.net/blog/harness-engineering-techniques-2026/): The named harness engineering techniques of 2026: ratchet rules, Ralph loops, spec-driven development and evaluator agents, with the cost of each. - [New Agentic Code Editors in 2026: ZCode, Antigravity 2.0](https://capitalandcompute.net/blog/new-agentic-code-editors-2026/): ZCode, Antigravity 2.0, Grok Build, Warp, and Zed all changed in 2026. Here is what actually shipped in each new agentic code editor, and what it costs. - [Benchmark Saturation: Why 99% on MMLU Means Almost Nothing](https://capitalandcompute.net/blog/benchmark-saturation-mmlu/): Benchmark saturation is when top AI models bunch so close to the ceiling a test cannot rank them. Why MMLU hit this wall, and the harder replacement is next. - [Claude Fable 5 Is Back: Pricing and Benchmarks](https://capitalandcompute.net/blog/claude-fable-5-returns-pricing-benchmarks/): Claude Fable 5 returns after a US export-control suspension. Here is the $10/$50 pricing, the SWE-Bench Pro and FrontierCode scores, and what it means. - [Claude Sonnet 5: Pricing, Benchmarks, and Cost](https://capitalandcompute.net/blog/claude-sonnet-5-pricing-benchmarks/): Claude Sonnet 5 ships at $3/$15 per million tokens, intro $2/$10 through August 2026. What changed versus Sonnet 4.6 and Opus 4.8, and the cost per task. - [Kiro Pricing Explained: How Far $20 a Month Goes](https://capitalandcompute.net/blog/kiro-pricing-explained-free-tier-credits/): Kiro gives a free tier of 50 credits and Pro at $20 a month for 1,000. What a credit actually buys, the Auto versus pinned-model savings, and how overage works. - [Windsurf Is Now Devin Desktop: What It Actually Costs](https://capitalandcompute.net/blog/windsurf-devin-desktop-rebrand-cost/): Windsurf became Devin Desktop on June 2, 2026, pricing unchanged. Cascade retires today. The real cost shift is what got bundled into the same $20 plan. - [Hidden Cost of AI-Generated Code: Beyond the Token Price](https://capitalandcompute.net/blog/hidden-cost-ai-generated-code/): Your Claude Code session cost $47 in tokens. The real cost was $470. This is the math your dashboard does not show. - [When Will RAM Prices Drop? 2026 Forecast](https://capitalandcompute.net/blog/when-will-ram-prices-drop/): RAM prices roughly tripled in 2026 as AI memory demand starved consumer DRAM. Here is where the trend is heading and whether to buy now or wait. - [What Claude Code's /cost Command Shows](https://capitalandcompute.net/blog/claude-code-cost-command-explained/): How to read the Claude Code /cost command: session spend, token breakdown, why Pro and Max differ from API billing, and when to use /usage instead. - [Cursor Pricing Limits Explained: How Far $20 Goes](https://capitalandcompute.net/blog/cursor-pricing-limits-explained/): What the $20 Cursor Pro plan really buys in 2026: roughly 225 to 650 requests depending on model, how the usage pool works, and when overages start. - [DeepSWE vs FrontierCode: Two Ways to Grade AI Code](https://capitalandcompute.net/blog/deepswe-vs-frontiercode/): DeepSWE grades whether an AI finishes the task. FrontierCode grades whether you would merge its code. Why the same model scores 59% and 13%. - [Vibe Coding Cost Economics 2026: The Real Math](https://capitalandcompute.net/blog/vibe-coding-cost-economics-2026/): Vibe coding tools cost $20-$100/month. The real cost including tokens, infrastructure, and technical debt runs $87-$340/month. Here is the math. - [Did OpenAI Cut Prices? GPT-5.6 Sol, Terra, Luna](https://capitalandcompute.net/blog/did-openai-cut-ai-prices-gpt-5-6/): OpenAI cut GPT-5.6 API prices on July 30, 2026: Luna by 80 percent, Terra by 20 percent, Sol not at all. The cheap floor deleted in June is back. - [GPT-5.6 Sol on Cerebras: 750 Tokens per Second](https://capitalandcompute.net/blog/gpt-5-6-sol-cerebras-speed-tokens-per-second/): GPT-5.6 Sol goes live on Cerebras at up to 750 tokens per second in July. Here is what that speed actually means, and why it is a chip story. - [GPT-5.6 Sol, Terra and Luna: What OpenAI Launched](https://capitalandcompute.net/blog/gpt-5-6-sol-terra-luna-launch/): OpenAI launched GPT-5.6 on June 26 as three tiers: Sol, Terra and Luna. Here are the confirmed prices, the benchmarks, and the government access catch. - [Apple Just Raised Prices. What It Means for Local AI](https://capitalandcompute.net/blog/apple-price-hike-local-ai/): Apple raised Mac and iPad prices on June 25 2026 as the memory shortage bit. Here is what the unified-memory tax means for running local AI. - [GPT-5.6 Delayed: The Government Is Now a Release Gate](https://capitalandcompute.net/blog/gpt-5-6-delayed-government-ai-release-gate/): The Trump administration asked OpenAI to stagger GPT-5.6 and approve users one by one. A federal gate now sits between a finished AI model and the public. - [How Much RAM Do You Need to Run a Local LLM?](https://capitalandcompute.net/blog/how-much-ram-to-run-a-local-llm/): How much RAM you need to run a local LLM in 2026: what models 8GB to 512GB can run, the per-billion-parameter math, and the device for each tier. - [Which Countries Use AI Most? The Economics Behind the Map](https://capitalandcompute.net/blog/ai-adoption-by-country-economics/): AI usage tracks national income, but the standouts break the rule. The economics of why Israel and Singapore use AI far more than their wealth predicts. - [Benchmark Contamination in AI: When Tests Leak Into Training](https://capitalandcompute.net/blog/ai-benchmark-contamination-explained/): Benchmark contamination is when test answers leak into training data and inflate AI scores. How it happens, how much it distorts results, and how to spot it. - [The AI Productivity Paradox: A Central Bank Measured It](https://capitalandcompute.net/blog/ai-productivity-paradox/): The Bank of Korea surveyed 5,512 workers: AI saves 1.5 hours a week, but the correlation with actual output is near zero. Why time saved is not productivity. - [GPT-5.6 Preview: The Leak Is Real, the Release Date Isn't](https://capitalandcompute.net/blog/gpt-5-6-preview-leak-release-date/): A leaked gpt-5.6-preview route lit up r/OpenAI, then GPT-5.6 launched June 26 as Sol, Terra and Luna in a government-gated preview. The signal vs the hype. - [Why Is RAM So Expensive in 2026? The Shortage Is a Choice](https://capitalandcompute.net/blog/why-is-ram-so-expensive-ai-memory-shortage/): RAM keeps getting pricier in 2026 because the memory giants profit more from scarcity than supply. Inside the most lucrative shortage in chip history. - [Are AI Benchmarks Reliable? How the Scores Get Gamed](https://capitalandcompute.net/blog/are-ai-benchmarks-reliable/): AI benchmark scores get gamed by contamination, saturation, and cheating. The 2026 receipts, a trust scorecard, and how to read past any leaderboard. - [China's AI Chip Companies in 2026: Replacing NVIDIA](https://capitalandcompute.net/blog/china-ai-chip-companies-2026/): Seven Chinese firms now ship AI accelerators, the best near NVIDIA H100 class. A fact-checked 2026 map of who makes China's GPUs and what is real. - [Cohere North Mini Code: What a Free Coding Model Costs](https://capitalandcompute.net/blog/cohere-north-mini-code-cost-per-task/): Cohere North Mini Code is free on the API and open-weight. Here is what it really costs per task once you self-host it on a single H100. - [SpaceX Bought Cursor: Your Bill, Your Code, Cost to Leave](https://capitalandcompute.net/blog/spacex-cursor-acquisition-cost-and-alternatives/): SpaceX is buying Cursor for $60B. What changes for your bill, whether your code now trains xAI models, and the real cost per task of switching away. - [GPT-5.6 vs Claude Fable 5: Which Costs Less Per Task?](https://capitalandcompute.net/blog/gpt-5-6-vs-claude-fable-5-cost-per-task/): GPT-5.6 launched June 26 as Sol at $5/$30, half Claude Fable 5 at $10/$50. Which flagship costs less per task, now that Fable 5 is back from suspension. - [Decentralized GPU vs Cloud: Real Cost Per Hour (2026)](https://capitalandcompute.net/blog/decentralized-gpu-cost-vs-cloud/): Is decentralized GPU compute (Akash, io.net, Render) cheaper than AWS? A grounded 2026 cost-per-hour comparison, plus the caveats the hype skips. - [Local LLM Tokenomics: Self-Hosted Cost Per Token (2026)](https://capitalandcompute.net/blog/self-hosted-llm-cost-per-token/): What a self-hosted LLM token really costs in 2026: cost per token across owned hardware, why memory bandwidth sets speed, and where buying beats the API. - [Best Gemini CLI Alternatives 2026: Real Cost Per Task](https://capitalandcompute.net/blog/best-gemini-cli-alternatives-cost-per-task/): Google ended free Gemini CLI access on June 18, 2026. The best alternatives (Claude Code, Aider, OpenCode, Antigravity) ranked by real cost per task. - [Claude Fable 5 vs Opus 4.8: Is 2x the Price Worth It?](https://capitalandcompute.net/blog/claude-fable-5-vs-opus-4-8-cost-per-task/): Claude Fable 5 costs exactly 2x Opus 4.8 per token: $10/$50 vs $5/$25. Whether it is cheaper per task depends on loop count, not the sticker rate. - [Gemini 3.5 Flash: Cheapest Coding Model Per Task?](https://capitalandcompute.net/blog/gemini-3-5-flash-coding-cost-per-task/): Gemini 3.5 Flash lists at $1.50/$9.00 per million tokens: about 2x Haiku per token, yet roughly 3x cheaper per task than GPT-5.5 on agentic coding work. - [How to Scaffold an AI Agent Harness Without Wasting Tokens](https://capitalandcompute.net/blog/how-to-scaffold-ai-agent-harness/): A cost-first guide to building an AI agent harness: loop, context, tools, verification, and guardrails, and the token cost each layer controls. - [MAI-Code-1-Flash Cost Per Task: Cheapest Coding Model?](https://capitalandcompute.net/blog/microsoft-mai-code-1-flash-cost-per-task/): Microsoft calls MAI-Code-1-Flash its cheapest coding model. In Copilot it bills like Claude Haiku 4.5 (0.33x); its token edge holds only on easy benchmarks. - [Qwen 3.7 Max vs Claude for Coding: The Real Cost Per Task](https://capitalandcompute.net/blog/qwen-3-7-max-vs-claude-coding-cost/): Qwen 3.7 Max lists at half Claude Opus 4.8 and runs the same eval suite for a third of the cost. The catch is not hidden cost. It is what the price buys. - [Claude Code vs Codex (2026): Cost and Capability](https://capitalandcompute.net/blog/claude-code-vs-codex-cost-capability/): Codex leads Claude Code 83.4% to 78.9% on Terminal-Bench 2.1, but the prices match and the cheaper agent flips with your harness. The honest scorecard. - [Cursor Composer 2.5: What It Actually Costs](https://capitalandcompute.net/blog/cursor-composer-2-5-real-cost/): Composer 2.5 finishes a coding task for about $0.07, 10-60x under Claude Opus and GPT-5.5 at near-equal benchmark scores. What the cheap headline leaves out. - [The Price Reversal Phenomenon: When Cheaper AI Costs More](https://capitalandcompute.net/blog/price-reversal-why-cheaper-ai-models-cost-more/): A 2026 Microsoft Research preprint found the cheaper-per-token AI model cost more to finish the job in 32% of model pairs. Why the sticker misleads. - [Do Claude Code Token-Saving Tools Actually Cut Your Bill?](https://capitalandcompute.net/blog/claude-code-token-saving-tools-rtk-headroom-caveman/): An independent replay of 500 Claude Code sessions found rtk, headroom, and caveman cut a $926 bill by just 3.7 percent. Here is why the 60-90% claims miss. - [GPT-5.6 Pricing: What to Expect Per Task](https://capitalandcompute.net/blog/gpt-5-6-pricing-cost-per-task/): GPT-5.6 launched June 26 2026 as Sol, Terra and Luna. The flagship Sol holds GPT-5.5 pricing at $5/$30 per million tokens. Here is the real cost per task. - [Open-Source LLMs Overtook Proprietary in 2026](https://capitalandcompute.net/blog/open-source-llms-overtake-2026/): Token data from OpenRouter shows open-source LLMs passing proprietary models in mid-2026, a roughly 60/40 flip. The daily breakdown by AI lab. - [Why Local LLMs Got Good in 2026: Capability & Cost](https://capitalandcompute.net/blog/local-llms-good-2026/): Open-weights LLMs crossed from toy to useful in 2026. What actually changed, and the cost math for when running a model yourself beats paying an API. - [World Models: Why AI's Biggest Names Bet Billions in 2026](https://capitalandcompute.net/blog/world-model-funding-wave-2026/): Lin Junyang, Fei-Fei Li, and Yann LeCun are raising billions for world models in 2026. What a world model is, who is funding it, and why the money is moving. - [AI Agent Benchmarks in 2026: What the Scores Actually Mean](https://capitalandcompute.net/blog/ai-agent-benchmarks-2026/): AI agent benchmarks broke in 2026: reward-hacked to 100%, SWE-bench Verified retired, scores swung by harness choice. What each measures and which to trust. - [GitHub Copilot Pricing 2026: What AI Credits Really Cost](https://capitalandcompute.net/blog/github-copilot-pricing-ai-credits/): GitHub Copilot switched to usage-based AI Credits on June 1, 2026. What the $10 Pro plan really costs once metering kicks in, and whether it is still worth it. - [Claude Code Pricing 2026: What It Really Costs Per Task](https://capitalandcompute.net/blog/claude-code-pricing-cost-per-task/): Claude Code plans run $20 to $200 a month, but the real number is cost per task. A modeled breakdown of token spend, where it wins, and where it burns money. - [The 2026 AI Coding Agent Landscape: Leaders, Costs, Harness](https://capitalandcompute.net/blog/the-2026-ai-coding-agent-landscape/): A grounded survey of the 2026 AI coding agent field: Claude Code, Cursor, Copilot, Codex and Antigravity, by interface, cost, and why the harness matters. ## Browse - [Start here](https://capitalandcompute.net/start-here/): orientation page. What this publication covers, which datasets and tools exist, and where to enter each topic. - [All posts](https://capitalandcompute.net/blog/): full article index. - [RSS feed](https://capitalandcompute.net/rss.xml): subscribe to new posts. ## About and policies - [About](https://capitalandcompute.net/about/): what this publication is and how it is run. - [Editorial standards](https://capitalandcompute.net/editorial-standards/): how content is researched, sourced, and verified. - [Contact](https://capitalandcompute.net/contact/): how to reach the editors. - [Privacy policy](https://capitalandcompute.net/privacy-policy/) - [Terms](https://capitalandcompute.net/terms/) - [Disclaimer](https://capitalandcompute.net/disclaimer/): editorial content, not financial advice.