Monday, August 31, 2026
Coverage: 66 videos reviewed (0 partial or status unknown); videos under 45 s were not reviewed. Dates are UTC upload dates. Vendor claims are labelled as such.
New today
Anthropic to raise standard Claude Code weekly limits 25% from Sept. 14, ending the 50% boost
Anthropic said standard weekly Claude Code limits rise permanently 25% for Pro, Max, Teams and seat-based enterprise plans from Sept. 14, 2026, and the current 50% increase stays until then, according to a reworded post that Theo read on screen. Theo and AI Code King both calculated that moving from a 50% boost to a 25% boost is about a 17% cut from current limits (1.25 divided by 1.5, or 150 to 125 units); the arithmetic is theirs. AI Code King said Anthropic reposted a clarified announcement conceding the 17% reduction, while Theo said Anthropic's post does not state the net change and that the first version was deleted and reposted. The 5-hour limit doubling stays.
- Evidence: 0 first-party, 0 hands-on, 2 relaying
- Disagreements: AI Code King says Anthropic's reposted announcement conceded a 17% reduction versus current limits; Theo says the post does not state the net change and derives the 17% himself.
- Watch: Theo - t3.gg: Anthropic Is "Increasing" Your Limits (high hype)
Continuing stories
- Tencent's HY4 preview, released Aug. 28, 2026, is a 770B MoE with 49B active parameters and 1M context - Tencent released HY4 preview on Aug. [0 first-party, 0 hands-on, 2 relaying] Watch: Bijan Bowen: Tencent HY4 Is INSANE– Is THIS Tencent’s Next Frontier Model?
Also notable
- Theo says Claude Fable 5 use is capped at about half of weekly limits on Anthropic plans - Theo said that since Fable 5 returned, it no longer counts against the whole weekly limit: in his hypothetical of $1,000 of inference, Fable stops after about $500 and users must switch to Opus or Sonnet. [0 first-party, 0 hands-on, 1 relaying] Watch: Theo - t3.gg: Anthropic Is "Increasing" Your Limits (high hype)
- Theo says OpenAI models are being banned in Cursor while Anthropic pledges more compute for Cursor - Theo described an OpenAI and SpaceX breakup with Cursor in which OpenAI models are being banned. [0 first-party, 0 hands-on, 1 relaying] Watch: Theo - t3.gg: Anthropic Is "Increasing" Your Limits (high hype)
- Claude Opus 5 is the default Opus in Claude Code with 1M-token context and $10/$50 fast mode, per AI Code King - AI Code King said Claude Opus 5 rolled out in late July as the default Opus in Claude Code, with 1M-token context on the API and Max, Team and Enterprise plans, and a fast mode on Opus 5 priced at $10 input and $50 output per million tokens. [0 first-party, 0 hands-on, 1 relaying] Watch: AI Code King: Claude Code 3.0 (All Upgrades Explained): You don't KNOW about THESE C
- Claude Code adds default auto mode, session messaging, /design preview and desktop simulator pane, per AI Code King - AI Code King said Claude Code's auto mode became the default permission mode for new Pro, Max and Team sessions from Aug. [0 first-party, 0 hands-on, 1 relaying] Watch: AI Code King: Claude Code 3.0 (All Upgrades Explained): You don't KNOW about THESE C
- OpenClaw 2.0 released after seven weeks without an update; Alex Finn's upgrade and sub-agent tests stalled - OpenClaw 2.0 was released, Julian Goldie said, with almost 1,000 contributors and over 16,000 changes, adding shared cloud sessions, grounded dreaming memory, SQLite-backed sessions, dashboards and widgets, and an experimental swarm; breaking changes include a removed plugin and renamed model routes. [0 first-party, 1 hands-on, 1 relaying] Watch: Alex Finn: OpenClaw 2.0 just dropped. It's officially over...
- Kimi K3 quantization on four 512GB Mac Studios reached 14.7 tok/s generation in Alex Ziskind's test - In a sponsor-funded video, Alex Ziskind ran an unpruned quantization of Kimi K3 (all 896 experts, 817GB; the model is 2.8T parameters, 1.56TB at original precision) across four 512GB Mac Studios over Thunderbolt 5 with RDMA and MLX, measuring about 238 tok/s prompt processing and 14.7 tok/s generation. [0 first-party, 1 hands-on, 0 relaying] Watch: Alex Ziskind: I Gave Local AI and the Cloud the Exact Same Job
- Alibaba says Wan 3.0 launched Aug. 24 with 30-second clips, up to 1080p and native audio - An Alibaba Cloud host said Wan 3.0 launched Aug. [1 first-party, 0 hands-on, 0 relaying] Watch: Alibaba Cloud: Wan3.0 livestream client sharing clip - Picsart.
- Google releases Gemini Omni Flash 1.1 video model with longer extension and 4K upscale - Julian Goldie said Google released Gemini Omni Flash 1.1, whose extension analyzes up to 10 seconds of prior footage (versus 1 second previously) and extends in 10-second steps to 40 seconds, with first and last frame control, and drafts at 360p up to 4K, available in AI Studio, Flow, the Gemini app and ComfyUI. [0 first-party, 0 hands-on, 2 relaying] Watch: Julian Goldie: NEW Gemini Omni Flash 1.1 is WILD (high hype)
- Tencent's internal blind evaluation scores HY4 preview 2.99 of 4 versus 2.94 for Kimi K3 - Julian Goldie relayed Tencent's own evaluation in which 163 internal experts judged 203 engineering tasks: HY4 preview averaged 2.99, GLM 5.3 2.92 and Kimi K3 2.94. [0 first-party, 0 hands-on, 1 relaying] Watch: Julian Goldie: NEW Tencent Hy4 Got Upgraded! 🤯
- Tencent Angel Slim releases HY4 preview GGUFs, with a 213.66 GB STQ1_0 build losing 0.2 to 1.6 points - Julian Goldie said Tencent released two GGUF builds on Aug. [0 first-party, 0 hands-on, 1 relaying] Watch: Julian Goldie: NEW Tencent Hy4 Got Upgraded! 🤯
Models & learning
- Bijan Bowen's hands-on tests of HY4 Preview produced working apps with fixes and some failures - Bijan Bowen ran HY4 Preview through OpenCode, pi and Blender MCP. [0 first-party, 1 hands-on, 0 relaying] Watch: Bijan Bowen: Tencent HY4 Is INSANE– Is THIS Tencent’s Next Frontier Model?
- BreezeBlue releases BreezeTTS2, a 3B open-weights TTS under a non-commercial license - Sam Witteveen said Chinese startup BreezeBlue released BreezeTTS2, a 3B open-weights model with voice design, cloning, emotion steering, vocal events and 50 languages with streaming. [0 first-party, 1 hands-on, 1 relaying] Watch: Sam Witteveen: BreezeTTS2 - 100% Local Real-Time Voice
- Daily releases PhoneLLM Alpha 1, an open-weights voice-agent fine-tune of Nemotron 3 Nano - Daily said PhoneLLM Alpha 1 is an open-weights fine-tune of Nemotron 3 Nano, described as a 30B mixture-of-experts with 3B active parameters, thinking off, for customer-support voice agents. [1 first-party, 0 hands-on, 0 relaying] Watch: Daily: Pipecat TV - Episode 6 - PhoneLLM
- Hands-on tests of BreezeTTS2: about 7.5 GB VRAM, real-time streaming, weaker German and Hindi - Fahd Mirza's run on a 48GB GPU used about 7.5 to 7.6 GB VRAM; voice design and cloning were mostly good, with a clone missing some tone and a plasticky male voice, and his German and Hindi samples sounded poor by his own ear, single samples each. [0 first-party, 2 hands-on, 0 relaying] Watch: Sam Witteveen: BreezeTTS2 - 100% Local Real-Time Voice
- Fahd Mirza's test: Thomson-1.0-Small flagged seven planted NDA issues plus an eighth, using 87 GB VRAM - Fahd Mirza tested Thomson Reuters' Thomson-1.0-Small, a continued-learning model built on Cohere's open 35B mixture-of-experts. [0 first-party, 1 hands-on, 0 relaying] Watch: Fahd Mirza: Thomson Reuters Built Their Own AI Lawyer: Run Thomson-1 Locally
- Julian Goldie's GoldyBench gives Claude Opus 5 an 8.27 of 10 average over 50 one-shot tasks - Julian Goldie reported that Claude Opus 5 averaged 8.27 out of 10 across 50 one-shot tasks with the same prompt for every model on his own bench. [0 first-party, 1 hands-on, 0 relaying] Watch: Julian Goldie: Claude Memory Just Got a HUGE Upgrade
- Nate Herk's Grokbot tests: false completion on a spreadsheet task, key.ai connector failure, one-prompt Slack routine - In Nate Herk's tests, a Grokbot agent claimed spreadsheet formatting was done, its own check found it had not landed, and the fix took about 7 minutes; a key.ai connector via Composio did not let agents generate images, so he used a browser workaround; and a Slack-triggered routine was created from one prompt. [0 first-party, 1 hands-on, 0 relaying] Watch: Nate Herk: Build & Sell Grok Bots (2 Hour Course)
- Blum describes Claude Cowork workflows at Melio, claiming a week of PM work in a day - On How I AI, Blum said his Cowork setup lets him do a week of PM work in a day, a self-reported claim he conceded sounds like hype, and showed a weekly self-improvement loop with masked data. [0 first-party, 0 hands-on, 1 relaying] Watch: How I AI: I built a Claude Cowork system that does a week of PM work in a day