Stay ahead on AI.
All in ONE place.
The latest AI news, expert commentary, top tools, and exclusive interviews — curated for you by Matt Wolfe.
Join 250,000+ readers from companies like
a16z
Sequoia CapitalAI News
The biggest AI headlines, updated throughout the day.
Google Launches Gemini 3.7 Flash With Major Coding Gains at Half the Price of 3.6 Flash
Google has launched Gemini 3.7 Flash, its most capable workhorse model for coding and agents, just three weeks after Gemini 3.6 Flash. The model shows major benchmark gains, including FrontierCode 1.1 (43.6% vs 34.4%) and DeepSWE v1.1 (65.3% vs 49.0%). It also outperforms 3.6 Flash on GDP.pdf document reasoning (34.0% vs 22.0%) and AutomationBench (30.4% vs 17.0%). Priced at $0.75 per million input tokens, it costs half of 3.6 Flash. Gemini Spark is also being upgraded to use 3.7 Flash starting today.
OpenAI Launches GPT-5.6 Sol Ultrafast Mode: 14x Faster, 750 Tokens/Sec via Cerebras
OpenAI has launched Ultrafast mode for GPT-5.6 Sol, a new API service tier powered by Cerebras that generates up to 750 output tokens per second, making it 14 times faster than standard processing. Currently in limited preview, Ultrafast targets time-sensitive business workflows including incident response, financial research, customer support, and live experimentation. Early customers include Jane Street, Podium, Basis, and Rogo. OpenAI says the tier delivers frontier intelligence without sacrificing speed for a smaller model.
Anthropic's Claude Tag Gains Channel-Wide Context, 30% Better at Knowing When to Respond
Anthropic has updated Claude Tag, its Slack integration, to use channel-wide context when deciding whether to proactively respond. Previously, a lightweight classifier evaluated each message individually; now Claude reads across the full channel, its memory, and standing instructions to choose one of four actions: reply inline, start a thread, route to an existing workstream, or stay silent. The update makes Claude roughly 30% better at knowing when to respond. Available now for Teams and Enterprise customers at no additional cost.
DeepSeek Launches V4-Pro with Agent Upgrades, Flexible Reasoning & New API Pricing
DeepSeek has launched DeepSeek-V4-Pro into general availability, featuring major agent upgrades with strong production gains. The model introduces flexible reasoning effort levels — low for simple tasks, high for daily agent workflows, and max for complex tasks — across both V4-Pro and V4-Flash. It also adds native OpenAI Responses API support optimized for Codex with one-click setup. New API pricing introduces peak and off-peak rates, with off-peak 50% cheaper, effective August 16, 2026.
Alibaba Launches Wan3.0: AI Video Generation in 30 Seconds on Cloud Model Studio
Alibaba has launched Wan3.0, the latest generation of its Wan video generation model family, now available on Alibaba Cloud Model Studio. The model generates up to 30 seconds of video in a single pass from text, images, audio, video, or documents including PPT, PDF, and XLS files — a first for the family. Wan3.0 features lifelike diverse human faces, reference-to-video consistency, and built-in video editing. API pricing starts at $0.05 per second for 480P, $0.10 for 720P, and $0.20 for 1080P.
Claude Cowork Comes to Chrome Side Panel, Syncing Sessions Across Desktop and Mobile
Anthropic's Claude in Chrome side panel has been upgraded to Claude Cowork, syncing browser sessions with the desktop, web, and mobile apps. Conversations are saved to account history, and tasks started in a browser tab can be continued on other devices. Skills and connectors work in the browser, letting Claude navigate tabs, fill forms, and interact with sites like vendor portals using existing logins. Available now on Max and Team plans, rolling out to Pro users soon.
Twitch Uses Streamer Content for Amazon AI Training by Default; Opt-Out Available
Twitch now defaults to allowing Amazon to use streamers' broadcasts, clips, VODs, highlights, chat, and images to train generative AI models. An opt-out was announced August 12, 2026, and can be found under Security and Privacy in account settings, labeled Training for Generative AI. The opt-out only covers future training, and Twitch has not confirmed whether Amazon has already used content. Opting out does not disable AI for captions, recommendations, AutoMod, or other platform features. Viewers cannot control how their chat messages are used in others' streams.
OpenAI: Frontier Firms Generate 8.3x More AI Output as Agentic Adoption Widens
OpenAI reports that frontier firms—the top 10% of enterprise AI users—now generate 8.3 times more output tokens per active user than typical firms, up from 2.6 times in January. Codex accounted for 64% of combined enterprise output tokens as of June, reflecting a shift toward agentic, multi-step work. Agentic adoption is spreading beyond engineering, with legal growing 108 times and sales 41 times since February. Early-career employees send 13 more messages weekly than executives.
xAI Launches Grok 4.6 with Enhanced Long-Running Agents and Visual Capabilities
xAI has released Grok 4.6, an upgrade to Grok 4.5 focused on long-running agents and visual work. The model matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index with a score of 61, and outperforms it on Harvey LAB and CursorBench benchmarks. Grok 4.6 underwent a longer supplemental training run using curated model-generated data and an improved optimizer. It is available now in Cursor, Grok Build, and via API, priced at $2 per million input tokens and $6 per million output tokens.
Google DeepMind's SL2T Model Brings ASL-to-Text to Gboard and Live Transcribe on Pixel 11
Google DeepMind has introduced SL2T, a massively multilingual sign-language-to-text translation model trained on over 100,000 hours of data across more than 50 sign languages. Starting with ASL-to-English, SL2T powers sign-to-text dictation in Gboard and Live Transcribe on Pixel 11, letting Deaf users sign anywhere they'd normally type, including web searches, messages, and Gemini queries. The model achieves a zero-shot score of 70 BLEURT on the FLEURS-ASL benchmark, surpassing all previously reported scores.
AI News with Matt
The latest breakdowns from the YouTube channel.
Commentary & Analysis
View all articles
What We Should Keep Human
By Matt Wolfe · Aug 11, 2026

Why Everyone Should Have a Local AI Model (Even If You Love ChatGPT)
By Roya Lotfi · Jul 7, 2026
AI Agents, 6G, and the Future of Computing
By Matt Wolfe & Cristiano Amon, CEO of Qualcomm · Jul 1, 2026
Agents, Humanist Superintelligence, and Healthcare
By Matt Wolfe & Mustafa Suleyman, CEO of Microsoft AI · Jun 2, 2026
Free subscriber bonus
Get the AI Income Database, free when you join
38+ real ways people are making side-hustle money with AI right now. Each one comes with the playbook and the exact tools to pull it off. No fluff and no course to buy. It's the bonus every new subscriber gets on day one.
- 38+ proven AI side-hustles, from first dollar to scale
- The exact tools for each one, picked from the 4,500+ we track
- Updated as new opportunities appear (subscribers hear first)
Joins the twice-weekly AI briefing read by 250,000+ people. Free forever, unsubscribe anytime.











