<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0"><channel><title>gonioAI — Business &amp; funding</title><link>https://gonioai.pages.dev/topics/business/</link><description>Business &amp; funding stories from gonioAI.</description><language>en</language><lastBuildDate>Tue, 11 Aug 2026 10:45:13 +0000</lastBuildDate><item><title>Nvidia guarantees its own chips' resale value to unlock $500B in AI debt</title><link>https://the-decoder.com/nvidia-guarantees-its-own-chips-value-to-unlock-500-billion-in-ai-infrastructure-financing</link><guid isPermaLink="false">2026-08-11:business:https://the-decoder.com/nvidia-guarantees-its-own-chips-value-to-unlock-500-billion-in-ai-infrastructure-financing</guid><pubDate>Tue, 11 Aug 2026 07:00:00 +0000</pubDate><description>Nvidia signed letters of intent with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion in third-party capital for data centers, fabs, and power plants. To make the financing pencil out, Nvidia will backstop up to 25% of the residual value of its own installed GPUs on a per-project basis, effectively absorbing part of the depreciation risk. Jensen Huang argues the hardware lasts far longer than critics claim, citing A100s still earning revenue six years on and H100 rental rates rising from $1.70 to $2.35 per GPU-hour. The move reads as a direct rebuttal to Michael Burry's warning that GPU depreciation is understated by ~$176B through 2028.

Why it matters: The whole AI buildout rests on how long a GPU stays economically useful. Nvidia putting its balance sheet behind that number, rather than just selling chips, is a tell about how circular the financing has become, and how much rides on utilization staying high.</description></item><item><title>KPMG: nearly half of executives dialed back AI agents over cost</title><link>https://www.reddit.com/r/LocalLLaMA/comments/1vk60uz/kpmg_says_nearly_half_of_executives_pulled_back</link><guid isPermaLink="false">2026-08-10:business:https://www.reddit.com/r/LocalLLaMA/comments/1vk60uz/kpmg_says_nearly_half_of_executives_pulled_back</guid><pubDate>Mon, 10 Aug 2026 07:00:00 +0000</pubDate><description>A KPMG survey reported by Forbes finds nearly half of surveyed executives have pulled back AI agent deployments because of cost. It lands amid mounting evidence that agentic token consumption is punishing—alongside this week's GitHub Models shutdown and recent accounts of individual developers burning billions of tokens in weeks.

Why it matters: The gap between agent demos and unit economics is now showing up in boardroom decisions. For the near term, budget rather than capability may be the ceiling on agent rollouts.</description></item><item><title>DeepMind loses its independence; Hassabis reportedly on the way out</title><link>https://the-decoder.com/google-dismantles-deepmind-and-bets-on-a-fresh-start-as-hassabis-heads-for-the-exit</link><guid isPermaLink="false">2026-08-09:business:https://the-decoder.com/google-dismantles-deepmind-and-bets-on-a-fresh-start-as-hassabis-heads-for-the-exit</guid><pubDate>Sun, 09 Aug 2026 07:00:00 +0000</pubDate><description>Following Jeff Dean's departure, reports say Google DeepMind is being downgraded to a subdivision: day-to-day operations pass to Koray Kavukcuoglu (without a CEO title), all Gemini work moves to the Bay Area, and Sergey Brin takes a larger role. Demis Hassabis was 'promoted' to chairman and could leave in the coming months to focus on Isomorphic Labs. SemiAnalysis reads the shakeup as Google conceding the frontier-model race and leaning into cloud and TPU revenue ($73B+ projected AI infra), while defenders frame it as a deliberate infrastructure play.

Why it matters: The lab that produced the Transformer's successors and Gemini is being reorganized around cloud margins, not model leadership. If you build on Gemini, the roadmap signals matter: 3.1 Pro is still preview and 3.5 Pro appears shelved.</description></item><item><title>ByteDance pre-trains a 10-trillion-parameter model to chase Mythos</title><link>https://arstechnica.com/ai/2026/08/bytedance-trains-massive-ai-model-in-bid-to-rival-anthropic</link><guid isPermaLink="false">2026-08-08:business:https://arstechnica.com/ai/2026/08/bytedance-trains-massive-ai-model-in-bid-to-rival-anthropic</guid><pubDate>Sat, 08 Aug 2026 07:00:00 +0000</pubDate><description>Per the Financial Times, ByteDance is early in pre-training a model with as many as 10 trillion parameters — three times Moonshot's Kimi K3 and in the range of estimates for Anthropic's ~8T Mythos 5. Sources say ByteDance has avoided distillation from rival model outputs for over a year, and founder Zhang Yiming has told the 2,000-person Seed team to aim for world-leading capability. xAI is reportedly training 6T and 10T Grok variants on its Colossus 2 cluster.

Why it matters: The parameter gap between Chinese labs and the US frontier is closing fast, and raw scale is back in fashion at the very moment everyone else is preaching the efficiency frontier.</description></item><item><title>Databricks: chase the efficiency frontier, not the intelligence frontier</title><link>https://www.databricks.com/blog/managing-ai-coding-costs-scale</link><guid isPermaLink="false">2026-08-08:business:https://www.databricks.com/blog/managing-ai-coding-costs-scale</guid><pubDate>Sat, 08 Aug 2026 07:00:00 +0000</pubDate><description>Databricks, with input from Stripe, Coinbase, Uber and Ramp, details how it cut internal AI coding spend by up to 90% while usage grew: aggressively adopt cheaper models that clear the quality bar, use a meta-harness (its open-sourced Omnigent) and an AI gateway for model flexibility, route work to the cheapest capable model, and cut context bloat — harness and cache tuning alone dropped generated tokens ~50%. Notably, Stripe found Opus 4.7 didn't beat 4.6, and Databricks saw regressions from Opus 5.0 versus 4.8. A leaked Accenture meeting separately fingers PDF-to-markdown conversion as a top token burner.

Why it matters: For teams, the 'best model' is usually the best routing plus harness plus budget policy, not the flagship checkpoint — and non-engineers converting PDFs are a real line item on the bill.</description></item><item><title>AMD buys Taalas to etch whole models into silicon</title><link>https://www.theregister.com/systems/2026/08/06/amd-acquires-ai-chip-startup-taalas-to-boost-inference-performance-by-etching-models-into-silicon/5284344</link><guid isPermaLink="false">2026-08-07:business:https://www.theregister.com/systems/2026/08/06/amd-acquires-ai-chip-startup-taalas-to-boost-inference-performance-by-etching-models-into-silicon/5284344</guid><pubDate>Fri, 07 Aug 2026 07:00:00 +0000</pubDate><description>AMD acquired chip startup Taalas, which builds model-specific integrated circuits that hard-wire a model's weights into silicon rather than loading them onto general-purpose GPUs. Early demos claim up to 17,000 tokens per second on these etched-model chips. AMD is framing it as an enterprise inference play, betting the market goes vertical as serving costs dominate.

Why it matters: If per-model ASICs deliver order-of-magnitude throughput, the economics of inference shift away from flexible GPU fleets toward fixed silicon per model, changing how anyone plans a serving stack for the next few years.</description></item><item><title>Alibaba floats revenue-sharing for the next open-weight Qwen</title><link>https://www.artificialintelligence-news.com/news/alibaba-qwen-open-source-ai-revenue-sharing</link><guid isPermaLink="false">2026-08-07:business:https://www.artificialintelligence-news.com/news/alibaba-qwen-open-source-ai-revenue-sharing</guid><pubDate>Fri, 07 Aug 2026 07:00:00 +0000</pubDate><description>Reuters reports Alibaba plans to require large companies that resell its next Qwen open-weight model as a service to strike a commercial agreement, with a revenue-sharing rate still unset. That breaks from the current Apache 2.0 Qwen3 terms and mirrors Moonshot's Kimi K3 license, which triggers a separate deal above $20M in annual MaaS revenue and reportedly can take up to 30% of revenue. The next model, Qwen3.8-Max, is a 2.4T-parameter MoE activating about 95B parameters per request.

Why it matters: The open-weight discount war has a catch: 'open weights' increasingly means 'free to download, pay if you make money,' so teams building on Chinese models need to read the license, not just the benchmark.</description></item><item><title>Jeff Dean and three Google legends quit to build an autoresearch startup</title><link>https://blog.google/company-news/inside-google/message-ceo/next-chapter-ai-momentum</link><guid isPermaLink="false">2026-08-06:business:https://blog.google/company-news/inside-google/message-ceo/next-chapter-ai-momentum</guid><pubDate>Thu, 06 Aug 2026 07:00:00 +0000</pubDate><description>Jeff Dean, Sanjay Ghemawat, Oriol Vinyals and Quoc Le are leaving Google DeepMind to co-found Discovery Loop, a public benefit corporation aimed at automating ML, science and engineering experiments at massive scale, with Alphabet as a founding investor and cloud partner alongside Radical and Khosla. In the same reshuffle Demis Hassabis moves from CEO to Chair of GDM and Chief Scientist of Alphabet, leaning into Isomorphic Labs, while CTO Koray Kavukcuoglu steps up to SVP running Gemini and frontier research. The exits follow Noam Shazeer, John Jumper and David Silver out the door, and land six months into a Gemini Pro update drought.

Why it matters: The people most associated with Google's infra, model-building and research stack are now chasing recursive self-improvement outside the company — a loud signal that AI-for-science is the next frontier and that Google's talent moat is leaking.</description></item><item><title>Eisman warns cheap Chinese open models could ignite an AI price war before the IPOs</title><link>https://247wallst.com/investing/2026/08/04/id-be-petrified-steve-eisman-says-cheap-chinese-ai-models-could-wreck-openai-and-anthropics-valuations</link><guid isPermaLink="false">2026-08-04:business:https://247wallst.com/investing/2026/08/04/id-be-petrified-steve-eisman-says-cheap-chinese-ai-models-could-wreck-openai-and-anthropics-valuations</guid><pubDate>Tue, 04 Aug 2026 07:00:00 +0000</pubDate><description>On his show, 'Big Short' investor Steve Eisman said that if he ran OpenAI or Anthropic he'd be 'petrified' of a price war. His specific example: Moonshot's open-weight Kimi K3 at $3/M input tokens versus $5 for GPT-5.6 Sol and $10 for Claude Fable 5, with open weights removing the switching cost premium subscriptions depend on. Both labs have filed confidentially with the SEC targeting ~$1T listings. Bloomberg Intelligence cited 988 approved Chinese LLMs, DeepSeek cutting API prices up to 50%, and Baidu cutting 99% earlier this year.

Why it matters: The moat debate now has an IPO clock on it: the pricing power a trillion-dollar valuation assumes is exactly what an open-weight price war erodes, and public investors will price it directly.</description></item><item><title>OpenAI answers Apple's trade-secret suit with the chat logs</title><link>https://openai.com/index/apple-is-getting-this-wrong</link><guid isPermaLink="false">2026-08-04:business:https://openai.com/index/apple-is-getting-this-wrong</guid><pubDate>Tue, 04 Aug 2026 07:00:00 +0000</pubDate><description>OpenAI published emails and iMessages to rebut Apple's July complaint, which alleges former Apple engineer Chang Liu improperly accessed confidential files after joining OpenAI. The receipts show Apple's outside counsel emailed the wrong person after confusing two Asian last names and claimed a phone call that OpenAI says never happened, and that Apple employees kept texting Liu for internal files after his January 22 departure. As critics note, the messages don't refute Apple's core claim that OpenAI encouraged new hires to bring proprietary information. The case ties to OpenAI's Jony Ive-led io Products hardware push and 400+ ex-Apple staff.

Why it matters: Good theater, but the document dump sidesteps the central allegation; the real fight is over OpenAI poaching Apple hardware talent for its consumer-device ambitions.</description></item><item><title>Alibaba ships Qwen3.8-Max at 2.4T params, claims Fable 5 parity</title><link>https://qwen.ai/blog?id=qwen3.8</link><guid isPermaLink="false">2026-08-03:business:https://qwen.ai/blog?id=qwen3.8</guid><pubDate>Mon, 03 Aug 2026 07:00:00 +0000</pubDate><description>Alibaba released Qwen3.8-Max, its largest model yet at 2.4 trillion parameters, sharing benchmark results that rank it above Moonshot's Kimi K3 and comparable to or better than Anthropic's Fable 5 on several tests. A smaller Qwen3.8-27B was announced alongside it; Unsloth's Daniel Han says the 27B fits in about 17GB of VRAM. The Max numbers are Alibaba's own, so treat the Fable 5 comparison as a vendor claim until third parties replicate it.

Why it matters: Another Chinese lab is claiming frontier-parity within weeks of Kimi K3, and the paired 27B means the same generation is usable on a single consumer GPU, not just via API.</description></item><item><title>OpenAI's super PAC linked to an AI-generated fake news site</title><link>https://www.modelrepublic.org/articles/the-reporters-at-this-news-site-are-ai-bots.-openai%E2%80%99s-super-pac-appears-to-be-using-it-to-advance-its-political-agenda</link><guid isPermaLink="false">2026-08-03:business:https://www.modelrepublic.org/articles/the-reporters-at-this-news-site-are-ai-bots.-openai%E2%80%99s-super-pac-appears-to-be-using-it-to-advance-its-political-agenda</guid><pubDate>Mon, 03 Aug 2026 07:00:00 +0000</pubDate><description>An investigation by Model Republic found that Acutus, an anonymous 'news' site publishing 94 articles since December, is almost entirely AI-generated: 69% of pieces flagged as fully AI-written, an exposed /api/wire endpoint leaks its automated editorial pipeline, and a bot named 'Michael Chen' emails critics posing as a reporter. Its AI-policy coverage mirrors Leading The Future, the $125M super PAC funded by OpenAI president Greg Brockman and a16z, with a funding trail running through PR firm Novus and GOP consultancy Targeted Victory. The site attacks Anthropic and AI-safety advocates while calling itself 'independent journalism.'

Why it matters: This is the AI-driven political influence campaign OpenAI's own usage policy once flagged as a top risk category, now apparently deployed on its behalf.</description></item><item><title>OpenAI cuts GPT-5.6 by up to 80% and credits its own model for the savings</title><link>https://www.latent.space/p/ainews-gpt-56-price-cut-by-20-80</link><guid isPermaLink="false">2026-07-31:business:https://www.latent.space/p/ainews-gpt-56-price-cut-by-20-80</guid><pubDate>Fri, 31 Jul 2026 07:00:00 +0000</pubDate><description>OpenAI dropped GPT-5.6 Luna 80% (now $0.20/$1.20 per million in/out tokens) and Terra 20% ($2/$12), and added a Sol Fast tier running up to 2.5x lower latency at 2x price with no claimed intelligence change. The company attributes the cuts to systems work partly done by GPT-5.6 Sol itself, which it says analyzed production traffic and autonomously rewrote Triton and Gluon serving kernels to cut end-to-end costs ~20%, plus a &gt;15% speculative-decoding gain. Swyx's analysis notes GPT-5.4's full flagship intelligence (AA index 51) now sells at roughly one-thirteenth of March's token price via Luna, and OpenAI is moving Codex and ChatGPT auto-review off GPT-5.4 onto Luna for ~10x lower cost.

Why it matters: Constant-level intelligence is getting an order of magnitude cheaper every few months, and OpenAI now undercuts several open models on cost-per-task. For anyone budgeting agent workloads, re-pricing your stack quarterly is no longer optional.</description></item><item><title>Amodei denies pushing an open-weights ban as NVIDIA's alliance goes live</title><link>https://www.cnbc.com/2026/07/27/anthropic-ceo-dario-amodei-isnt-advocating-open-weight-model-ban.html</link><guid isPermaLink="false">2026-07-28:business:https://www.cnbc.com/2026/07/27/anthropic-ceo-dario-amodei-isnt-advocating-open-weight-model-ban.html</guid><pubDate>Tue, 28 Jul 2026 07:00:00 +0000</pubDate><description>After days of criticism for skipping the Nvidia-led open-weights letter, Dario Amodei published a post saying Anthropic 'never advocated for a ban on open-weights models as a category,' instead backing chip export controls, anti-distillation rules, and mandatory safety testing for any sufficiently capable model. He explicitly rejected the letter's claim that open weights favor defenders over attackers. Meanwhile Jensen Huang formally launched the Open Secure AI Alliance (Hugging Face, IBM, Cloudflare, Cisco and others), and OpenAI management reportedly decided not to join, drawing internal backlash.

Why it matters: The people who actually make the models and chips are now split into rival camps, and the framing they win with will shape whether Chinese open-weight models like Kimi and Qwen get regulated out of the US market.</description></item><item><title>SK Hynix's $476K bonus is bleeding Samsung's chip engineers dry</title><link>https://www.technologyreview.com/2026/07/28/1140853/samsung-chip-workers-exodus-sk-hynix</link><guid isPermaLink="false">2026-07-28:business:https://www.technologyreview.com/2026/07/28/1140853/samsung-chip-workers-exodus-sk-hynix</guid><pubDate>Tue, 28 Jul 2026 07:00:00 +0000</pubDate><description>SK Hynix's record HBM profits translated into a roughly $476,000 per-employee cash bonus this year, versus about $135,000 for Samsung's loss-making foundry division, and Samsung engineers are defecting en masse. A union survey found 81.5% of foundry staff want out within two years; Samsung won an 18-month injunction blocking two former workers from joining its rival. The exodus threatens Samsung's one structural edge in HBM4: being the only memory maker that also runs its own advanced logic foundry.

Why it matters: The AI boom's constraint is shifting from GPUs to the HBM stacked on them, and whoever retains the memory-and-logic talent controls the supply that feeds every Nvidia accelerator.</description></item><item><title>Robotics gets its bitter-lesson moment as Enigma raises $71M</title><link>https://importai.substack.com/p/import-ai-466-the-bitter-lesson-for</link><guid isPermaLink="false">2026-07-28:business:https://importai.substack.com/p/import-ai-466-the-bitter-lesson-for</guid><pubDate>Tue, 28 Jul 2026 07:00:00 +0000</pubDate><description>Import AI rounds up evidence that scaling general models is starting to pay off in robotics: Anthropic's Project Fetch had Opus 4.7 autonomously complete quadruped tasks in ~9 minutes that a human record set at 181, purely as a byproduct of general scaling, while startup Sunday's ACT-2 hit a 99.1% garment-folding success rate via a strong base model plus minimal in-house data. Separately, Enigma emerged from stealth with a $71M seed (Index, Ribbit, Conviction) betting instead on studying how humans want to interact with robots, opening 100+ of its own arms to online public control. Epoch and METR also released MirrorCode, a long-horizon coding benchmark where Opus 4.7 reimplemented a 61k-line program.

Why it matters: If robot generalization really is now a base-model problem rather than a bespoke-data problem, the field could inherit the same scaling curve that transformed language, and the money is already moving on that thesis.</description></item><item><title>Chinese DRAM maker CXMT surpasses Intel's market cap on a 500% debut</title><link>https://www.reddit.com/r/LocalLLaMA/comments/1v7vdvg/chinese_chipmaker_cxmts_market_capitalization</link><guid isPermaLink="false">2026-07-27:business:https://www.reddit.com/r/LocalLLaMA/comments/1v7vdvg/chinese_chipmaker_cxmts_market_capitalization</guid><pubDate>Mon, 27 Jul 2026 07:00:00 +0000</pubDate><description>CXMT, mainland China's only integrated device manufacturer mass-producing general-purpose DRAM, surged nearly 500% on its first trading day to roughly RMB 3.28 trillion, the largest company by value on China's A-share market. That edges past Intel, which closed the prior day at about $465.6 billion (~RMB 3.15 trillion). The Hefei-based firm is central to China's push for domestic memory supply.

Why it matters: Memory is the bottleneck for AI accelerators; a well-capitalized domestic DRAM champion signals China intends to close the HBM and DRAM gap that export controls were meant to hold open.</description></item><item><title>Inside the gray market reselling LLM tokens at a discount</title><link>https://simonwillison.net/2026/Jul/26/relay-market</link><guid isPermaLink="false">2026-07-27:business:https://simonwillison.net/2026/Jul/26/relay-market</guid><pubDate>Mon, 27 Jul 2026 07:00:00 +0000</pubDate><description>Simon Willison flags Matt Lenhard's investigation into a mostly-Chinese marketplace that resells API tokens below cost by pooling keys — abusing free trials, proxying through unprotected support bots, and sometimes using stolen cards. The plumbing is open source: the one-api proxy and its more active fork new-api load-balance requests across a pool of credentials. Buyers want cheap tokens, geo-bypass, and distillation data.

Why it matters: If you expose an LLM-backed endpoint, there is now an ecosystem hunting for it to monetize your token budget — a hard argument for strict per-key spend caps that vendors still mostly don't offer.</description></item><item><title>Open-weights letter doubles to 50 names; Anthropic and Amazon hold out</title><link>https://www.forbes.com/sites/sandycarter/2026/07/25/huangs-open-weights-letter-doubled-to-50-without-amazon-and-anthropic</link><guid isPermaLink="false">2026-07-26:business:https://www.forbes.com/sites/sandycarter/2026/07/25/huangs-open-weights-letter-doubled-to-50-without-amazon-and-anthropic</guid><pubDate>Sun, 26 Jul 2026 07:00:00 +0000</pubDate><description>Jensen Huang's 'Open Weights and American AI Leadership' letter went from 25 to 50 signatories in a single day, adding OpenAI, Google, AMD, Cisco, GitHub, Cloudflare, Block and Ollama. Anthropic and Amazon are the conspicuous absences, even though Google, another Anthropic backer, signed. Meanwhile the NYT reports the White House leans toward targeted bans on specific Chinese models rather than a blanket ban, and that Anthropic and OpenAI are privately lobbying to restrict Chinese open weights, even as OpenAI publicly signs the pro-openness letter.

Why it matters: The model layer is the one place almost every signatory keeps no moat, so watch who lobbies privately versus who signs publicly. Nvidia asks for openness in everyone's yard but CUDA.</description></item><item><title>Anthropic asks SK Hynix for supplies to build its own chips</title><link>https://fortune.com/2026/07/25/sk-chair-chey-tae-won-anthropic-chip-supplies-skhynix</link><guid isPermaLink="false">2026-07-26:business:https://fortune.com/2026/07/25/sk-chair-chey-tae-won-anthropic-chip-supplies-skhynix</guid><pubDate>Sun, 26 Jul 2026 07:00:00 +0000</pubDate><description>SK Group chair Chey Tae-won said Anthropic approached SK Hynix, one of the largest memory makers, for supplies to make its own semiconductors, speaking on stage alongside Dario Amodei at a San Francisco AI event. Chey called it remarkable for an AI developer to pursue its own silicon. The visit coincided with South Korea's president convening an AI summit, where Nvidia also announced partnerships with Naver and SK Group.

Why it matters: After committing to 2GW of AMD MI450s last week, Anthropic sniffing at custom silicon signals it wants leverage over both the Nvidia and AMD supply queues.</description></item><item><title>Nvidia, Microsoft, Meta rally 20+ firms against open-weight curbs</title><link>https://www.cnbc.com/2026/07/24/nvidia-microsoft-meta-open-weight-ai-models.html</link><guid isPermaLink="false">2026-07-25:business:https://www.cnbc.com/2026/07/24/nvidia-microsoft-meta-open-weight-ai-models.html</guid><pubDate>Sat, 25 Jul 2026 07:00:00 +0000</pubDate><description>A Microsoft-initiated open letter, 'Open Weights and American AI Leadership,' was signed by more than 20 companies including Nvidia, Meta, Palantir, Hugging Face and Mistral, urging policymakers to avoid 'premature restrictions' on open-weight models and to treat distillation as legitimate rather than theft. It lands as the Trump administration weighs sanctions on Chinese labs like Moonshot (Kimi K3) over alleged distillation of Anthropic. Notably absent: OpenAI, Anthropic and Google — though Microsoft's own site briefly listed OpenAI as a signatory. The Decoder argues the campaign is transparently an Azure play, since more models on Azure and cheaper in-house MAI models improve Microsoft's margins.

Why it matters: The policy fight now pits closed-model incumbents against their own customers; developers' access to cheap, high-performing open weights is the stake, and the industry is lining up heavily on the open side.</description></item><item><title>Cognition buys Poke to give Devin a personality</title><link>https://techcrunch.com/2026/07/24/why-cognition-bought-poke-ai-personality-is-becoming-a-competitive-advantage</link><guid isPermaLink="false">2026-07-25:business:https://techcrunch.com/2026/07/24/why-cognition-bought-poke-ai-personality-is-becoming-a-competitive-advantage</guid><pubDate>Sat, 25 Jul 2026 07:00:00 +0000</pubDate><description>Coding startup Cognition acquired The Interaction Company, maker of the text-a-friend assistant Poke, for a price in the 'low nine figures.' The plan is to graft Poke's proactive, chatty interaction model onto the Devin coding agent while Poke gains Cognition's models and infrastructure, routing some tasks to the new SWE-1.7 model. Poke users exchanged over 100M messages in three months but the product was expensive to run and unprofitable.

Why it matters: A bet that agent UX and personality — not just raw model quality — are becoming the differentiator, and that a Poke-style orchestrator could manage multiple parallel Devin sessions.</description></item><item><title>Stripe in talks to buy model router OpenRouter for $10B</title><link>https://www.reddit.com/r/LocalLLaMA/comments/1v5l9m6/stripe_eyes_10_billion_deal_for_ai_model</link><guid isPermaLink="false">2026-07-25:business:https://www.reddit.com/r/LocalLLaMA/comments/1v5l9m6/stripe_eyes_10_billion_deal_for_ai_model</guid><pubDate>Sat, 25 Jul 2026 07:00:00 +0000</pubDate><description>Stripe is reportedly in talks to acquire OpenRouter, the model-routing marketplace that aggregates access to hundreds of LLMs, for around $10 billion. OpenRouter has been a prime beneficiary of the surge in cheap Chinese open-weight models, alongside inference providers like Baseten and Fireworks.

Why it matters: A payments giant paying eleven figures for a router underlines how much value is accruing to the routing/aggregation layer as model choice explodes and prices fall.</description></item><item><title>Etched raises $300M at $10.3B to build transformer-inference systems</title><link>https://techcrunch.com/2026/07/23/ai-chip-startup-etched-defies-skeptics-hits-10-3b-valuation-from-big-name-investors</link><guid isPermaLink="false">2026-07-24:business:https://techcrunch.com/2026/07/23/ai-chip-startup-etched-defies-skeptics-hits-10-3b-valuation-from-big-name-investors</guid><pubDate>Fri, 24 Jul 2026 07:00:00 +0000</pubDate><description>Etched closed a $300M Series C at a $10.3B valuation led by Sequoia, with a16z, SK Hynix, and Jane Street participating, doubling its December valuation in seven months. The company says it has already booked $1B in orders and is shipping full rack systems, not just chips, with a low-voltage prefill chip and a 'cluster-scale memory' interconnect for the decode phase. It pushes back on the perception that its silicon runs only specific LLMs, claiming support for MoE models and non-transformer designs like Mamba. Etched also opened an 80,000 sq ft, 10 MW facility in Milpitas, framing its pitch as 'run the world's inference.'

Why it matters: Inference-specialized silicon is graduating from thesis to booked revenue, and the more credible these alternatives get, the more pricing pressure Nvidia faces on the serving side.</description></item><item><title>Alphabet posts its first-ever negative cash flow as AI capex bites</title><link>https://www.reuters.com/business/retail-consumer/alphabets-cash-burn-raises-alarm-big-tech-ai-spending-climbs-2026-07-23</link><guid isPermaLink="false">2026-07-24:business:https://www.reuters.com/business/retail-consumer/alphabets-cash-burn-raises-alarm-big-tech-ai-spending-climbs-2026-07-23</guid><pubDate>Fri, 24 Jul 2026 07:00:00 +0000</pubDate><description>Alphabet burned $5.9B in Q2, its first cash burn on record, despite $119.8B in revenue and Google Cloud growing 23.8% quarter-over-quarter to $24.8B. The company raised its 2026 capex outlook by roughly $15B and expects to spend more next year, with Big Tech capex on track to top $700B in 2026. Shares fell about 6%, and analysts expect Amazon to burn cash too while Meta's free cash flow is projected to shrink 95.7%. Microsoft, Meta, and Amazon all report next week, sharpening scrutiny of whether AI revenue can outrun capex, depreciation, and operating costs.

Why it matters: The infrastructure bill behind every API you call is now large enough to push the most profitable companies into the red, and next week's earnings will show whether the payoff is keeping pace.</description></item><item><title>Treasury puts Chinese model distillation on the sanctions table</title><link>https://techcrunch.com/2026/07/22/treasury-threatens-sanctions-after-white-house-claims-moonshot-distilled-anthropics-fable</link><guid isPermaLink="false">2026-07-23:business:https://techcrunch.com/2026/07/22/treasury-threatens-sanctions-after-white-house-claims-moonshot-distilled-anthropics-fable</guid><pubDate>Thu, 23 Jul 2026 07:00:00 +0000</pubDate><description>Treasury Secretary Scott Bessent said sanctions and Entity List designations are "on the table" after White House science chief Michael Kratsios accused Moonshot of "large-scale, covert industrial distillation" of Anthropic's Fable to build Kimi K3, and alleged it accessed export-banned Nvidia GB300 servers in Thailand. Critics flag the timeline: Fable only became public July 1, and K3 shipped roughly two weeks later, making a distillation-only leap hard to square. Separately, a group of startup founders urged the Trump administration not to ban Chinese open-weight models outright.

Why it matters: If "distillation equals IP theft" becomes enforceable policy, training on another model's outputs — something every lab does, including on their own prior generations — enters legal gray territory, and downloadable Chinese weights that many defenders now rely on could be restricted.</description></item><item><title>Anthropic commits to 2GW of AMD MI450 GPUs; AMD invests up to $5B</title><link>https://the-decoder.com/anthropic-will-deploy-2-gigawatts-of-amd-gpus-for-claude-in-a-deal-worth-up-to-5-billion</link><guid isPermaLink="false">2026-07-23:business:https://the-decoder.com/anthropic-will-deploy-2-gigawatts-of-amd-gpus-for-claude-in-a-deal-worth-up-to-5-billion</guid><pubDate>Thu, 23 Jul 2026 07:00:00 +0000</pubDate><description>AMD will invest up to $5 billion in Anthropic, which in turn will deploy up to 2 gigawatts of Instinct MI450-series accelerators in Helios rack systems — MI455X GPUs paired with EPYC "Venice" CPUs, Pensando networking and ROCm — with the first gigawatt landing in H1 2027. AMD's stake is milestone-gated on deployment, echoing its 6GW OpenAI and 6GW Meta arrangements. A multi-year engineering program will use Claude to improve AMD's ROCm software, and AMD will run Claude internally across its dev teams.

Why it matters: It's another circular chip-lab financing loop, but it gives Anthropic a real second GPU source alongside Nvidia, Amazon Trainium and Google TPUs — and puts Claude to work hardening the weakest part of AMD's stack, its software.</description></item><item><title>Judge signs off on Anthropic's $1.5B book-piracy settlement</title><link>https://techcrunch.com/2026/07/20/anthropics-landmark-1-5b-copyright-settlement-is-approved</link><guid isPermaLink="false">2026-07-21:business:https://techcrunch.com/2026/07/20/anthropics-landmark-1-5b-copyright-settlement-is-approved</guid><pubDate>Tue, 21 Jul 2026 07:00:00 +0000</pubDate><description>US District Judge Araceli Martinez-Olguin granted final approval to Anthropic's $1.5 billion class-action settlement, paying roughly $3,000 per work across about 500,000 titles it downloaded from pirate libraries like Library Genesis to train Claude. The late Judge Alsup's underlying ruling stands: training on copyrighted text is fair use, but obtaining it via piracy is not, and Anthropic must now destroy the pirated copies. Because Anthropic settled rather than appealed, none of this becomes binding precedent, and parallel suits against Google, Meta, OpenAI and Midjourney roll on.

Why it matters: Fair-use-for-training survives as the industry's working assumption, but provenance is now a nine-to-ten-figure liability: where you sourced the data matters as much as what you did with it.</description></item><item><title>Kimi K3 freezes new subscriptions 48 hours in as demand outruns GPUs</title><link>https://www.reuters.com/legal/transactional/chinas-moonshot-pauses-kimi-subscriptions-amid-hot-demand-ipo-push-2026-07-20</link><guid isPermaLink="false">2026-07-20:business:https://www.reuters.com/legal/transactional/chinas-moonshot-pauses-kimi-subscriptions-amid-hot-demand-ipo-push-2026-07-20</guid><pubDate>Mon, 20 Jul 2026 07:00:00 +0000</pubDate><description>Moonshot paused new Kimi K3 consumer subscriptions after requests 'pushed close to the limits of our current capacity,' prioritizing existing paid users and splitting plans into a general 'Kimi Membership' and a separate 'Kimi Code Membership' to ration compute. Reuters reports the crunch coincides with a fresh $2B raise at a $30B valuation and preparations for a Hong Kong IPO. Analysts note K3's 2.8T size and agentic, multi-call workloads make it expensive to serve — and impractical for most to self-host despite the open weights.

Why it matters: So much for open weights cutting compute needs: the largest open model to date is capacity-constrained days after launch, a reminder that 'open' doesn't mean 'runnable' at 2.8T and that hosted access, not the download, is where the business lives.</description></item><item><title>Musk v. Altman exposes 2022 email: OpenAI's open-source plan was to freeze out rivals</title><link>https://simonwillison.net/2026/Jul/20/sam-altman</link><guid isPermaLink="false">2026-07-20:business:https://simonwillison.net/2026/Jul/20/sam-altman</guid><pubDate>Mon, 20 Jul 2026 07:00:00 +0000</pubDate><description>A newly surfaced October 2022 email from Sam Altman to OpenAI's board, exposed in the Musk v. Altman litigation, proposes releasing a locally-runnable GPT-3-class model — explicitly to 'discourage others from releasing similarly-powerful models' and make it 'harder for new efforts to get funded.' Simon Willison flagged the quote as a candid window into how open releases were pitched internally as a competitive moat rather than a gift.

Why it matters: Against a backdrop of OpenAI execs now warning about Chinese open weights, the 2022 framing lands differently: openness was a strategic lever the whole time, useful context for reading today's 'open-source is dangerous' arguments.</description></item></channel></rss>
