Tencent's Hy4 balloons to 770B open weights
Open weights kept scaling: Tencent's Hy4 Preview jumped to 770B total parameters barely a month after Hy3, and vLLM shipped a 584-commit release chasing DeepSeek V4 and Kimi-K3. On the business side, Sony and Warner opened a fresh copyright front against Anthropic, and Anthropic quietly trimmed effective Claude Code limits under the cover of a "raise."
Tencent's Hy4 Preview more than doubles Hy3 to 770B open weights
Tencent released Hy4 Preview, an open-weight text-only (no vision) LLM with 770B total and 49B active parameters, a 1M-token context window, and a 1.56TB footprint on Hugging Face. That is a large jump from July's Hy3 at 295B total, 21B active and 256k context. Simon Willison notes the chat template exposes only two reasoning modes, 'high' (default) and 'no_think'. Separately, r/LocalLLaMA posters report Tencent shipped a low-bit quant (labeled Q1, actually ~2.38 bpw) compressing the model to roughly 200GB while, per Tencent's own posted numbers, moving benchmarks like SWE-Bench multi from 82.9 to 81.3.
Why it matters: A 770B open-weight model with a 1M context is a serious artifact to watch, but the 1.56TB download and text-only scope mean most developers will be waiting on community quants before they can touch it.
- Introducing Hy4 Preview (Simon Willison)
- An official 1-bit quant for Hy4??? (r/LocalLLaMA)
- Tencent compressed Hy4-preview from 1.5TB to about 200GB GGUF and kept about 98% performance (r/LocalLLaMA)
Sony and Warner sue Anthropic, naming Amodei and Mann personally
Sony Music Publishing, Warner Chappell and other publishers sued Anthropic in the Northern District of California, accusing it of a 'brazen campaign' of torrenting, scraping and downloading copyrighted works to train Claude. The complaint names CEO Dario Amodei and co-founder Benjamin Mann as individual defendants and seeks up to $150,000 per infringed work, focusing on how the training data was acquired rather than only how it was used. It builds directly on the Bartz case, where Anthropic agreed to a $1.5B settlement after a judge ruled pirating the source material was illegal even if training on it was fair use. Anthropic says it disagrees and will defend itself.
Why it matters: The suit reuses the exact acquisition-not-use theory that already cost Anthropic $1.5B, and naming the founders personally raises the stakes for every lab that quietly torrented its pretraining corpus.
- Sony Music, Warner sue Anthropic, alleging a 'brazen campaign' of intellectual property theft (TechCrunch AI)
- Sony and Warner sue Anthropic over 'one of the largest and most blatant ongoing thefts of intellectual property in history' (The Decoder)
- Music publishers sue Anthropic, allege 'blatant theft' of copyrighted music (Axios)
Anthropic's Claude Code 'limit raise' is a 17% cut from today
Anthropic said that starting September 14 it will permanently raise standard weekly Claude Code limits by 25% for Pro, Max, Team and seat-based Enterprise plans. But a temporary 50% boost currently in place expires the same day, so relative to what users have now the change is a 17% reduction, which Anthropic acknowledged after deleting its original X thread. The company says more usage changes are coming to give users 'more visibility and control.'
Why it matters: If you budget agent runs against your weekly Claude Code allowance, plan for less headroom after September 14, not more, regardless of how the announcement is framed.
vLLM 0.28.0 lands with DeepSeek V4, Kimi-K3 and tiered KV offload
vLLM cut v0.28.0, a release of 584 commits from 270 contributors. Highlights include end-to-end sparse MLA for DeepSeek V4 (plain decode, MTP and speculative decoding), a broad Kimi-K3 performance push across the stack, new speculative-decoding methods (DFlash2, DSpark confidence-scheduled verification), and tiered KV cache offloading that now supports spilling to disk. Defaults changed too: max_num_batched_tokens rose from 8192 to 16384 and prefix caching is on by default for Mamba models. Breaking changes include bitsandbytes moving to an out-of-tree plugin and a bump to Transformers 5.15.0.
Why it matters: vLLM remains the reference serving engine, so its default and breaking-change list is effectively a migration checklist for anyone self-hosting these models.
- vLLM v0.28.0 (GitHub)
Rockstar says GTA 6 ships with no generative AI and no microtransactions
Rockstar confirmed on the record that Grand Theft Auto 6 will launch on November 19 with no generative AI and no microtransactions in its single-player game. Co-studio head Rob Nelson gave a flat 'no' on both, echoing Take-Two CEO Strauss Zelnick's earlier line that generative AI has 'zero part' in the game and that its worlds are 'handcrafted.' Both promises are scoped to the single-player campaign; Rockstar declined to discuss the next iteration of GTA Online, whose Shark Card model remains a core Take-Two revenue stream.
Why it matters: The industry's biggest launch explicitly rejecting generative AI is a marketing data point about how toxic 'gen AI' has become as a label, even as studios quietly use the same tools elsewhere.
- GTA 6 Will Have No Microtransactions or Generative AI at Launch, Rockstar Says (IGN)
- GTA 6 Ships Without Microtransactions or Generative AI, Rockstar Confirms (GamesReviews.com)
- Rockstar Games Issues New Official Statement on GTA 6 Generative AI Use (OpenCritic)
- GTA 6 does not use generative AI and will not offer microtransactions (Neowin)
OpenAI expands ChatGPT for Teachers to 100,000 more educators
OpenAI expanded its free ChatGPT for Teachers program by 55 school systems across 20 states, adding more than 100,000 educators and staff; it says it now works with 100-plus K-12 organizations across 30 states, covering about 340,000 educators serving over 2 million students. The managed workspace offers admin controls and role-based access, and OpenAI says workspace data is not used to train its models by default. It also announced a 16-state data-privacy agreement under the Student Data Privacy Consortium framework, with the tool free for verified U.S. K-12 educators through June 2028.
Why it matters: OpenAI is locking in institutional distribution and a standardized privacy contract, the unglamorous plumbing that turns a chatbot into default infrastructure for a sector.
- OpenAI partners with Kamehameha Schools to expand AI access for educators (The Garden Island)
Also worth a look
- Humaneval benchmark for Deepseek V4 Flash 0731 vs GLM5.3 Flash on 2x DGX Spark setup (r/LocalLLaMA)
- Qwen3.8-Flash-Next optimised for Macs (r/LocalLLaMA)
- Qwen3.8-Flash-Next NVFP4 Day-3 support for 4xV100 (r/LocalLLaMA)
- This finance-model benchmark card is more useful for what it discloses than for who 'wins' (r/LocalLLaMA)
- LLMs are making me lose my savviness (Hacker News)
- OpenAI and Anthropic are battling Big Tech for talent. We asked workers who's winning them over (Business Insider)
- Not a token effort: beer meets AI at bar in China's capital (South China Morning Post)
- Koboldcpp v1.120 released (r/LocalLLaMA)
- It's official! 192GB Framework (r/LocalLLaMA)