Daily AI News Brief

Today on the Brief: Watermarks Go Optional

Google makes the visible watermark on Gemini media optional, Alibaba open-sources a 2.4-trillion-parameter model, Google halves the price of its best coding model, OpenAI makes GPT-5.6 Sol run 14× faster, Apple gets cleared to run its own AI inside China — and DeepSeek prices jump up to 1,100% tomorrow.

▶ Watch Today's Brief on YouTube
Today's Stories
STORY 01

Alibaba Open-Sources Qwen 3.8

On August 14 the Qwen team released open weights for both Qwen3.8-Max and Qwen3.8-27B on Hugging Face and ModelScope. Max is the headline: a 2.4-trillion-parameter mixture-of-experts model activating roughly 95 billion parameters per query, natively multimodal across text, image, and video, with a 1-million-token context window.

The more consequential release is the smaller one. Qwen3.8-27B is a dense 27.78-billion-parameter multimodal model with a 262,144-token native context window (extendable to 1M via YaRN scaling), shipped under Apache 2.0 — a permissive license with no revenue clause and no gatekeeper. Bloomberg also reported today that the Qwen family has passed 3 billion cumulative downloads across 460+ open models — more than Google's and Meta's open models combined, per Hugging Face data.

Do this: Got a 24GB GPU or a 32GB+ Mac? Grab a quantized Qwen3.8-27B build from Hugging Face, load it in LM Studio or Ollama, and run 10 real samples of your highest-volume API task through it this weekend. If the output holds up, that line on your API bill goes to zero.
STORY 02

Google Makes Visible Watermarks Optional on Gemini

Rolling out over the next few days: a new Settings > Media Watermark toggle in the Gemini app and the Flow editor lets you turn off the visible "sparkle" watermark on AI generations — images from Nano Banana, video from Omni, and music from Lyria. The invisible SynthID watermark and C2PA provenance metadata always stay embedded and cannot be removed.

The toggle isn't universal: work and school accounts don't get it, and in countries that require visible labels — India, South Korea, Vietnam — it's available only on the paid AI Ultra plan. Google also open-sourced Credentio, a C++ library for validating C2PA content credentials locally, with no cloud round-trip. Meanwhile Anthropic went the other way: all new Claude models now watermark their text worldwide (a statistical pattern in word choice) plus C2PA on files, under the EU AI Act's transparency rules.

Do this: (1) Open Gemini → Settings → Media Watermark and set it the way your work needs — 30 seconds. (2) If you deliver AI-generated work to clients, add one line to your contract or delivery email saying how AI content is labeled. Provenance stays in the file permanently — make disclosure your policy before a client asks.
STORY 03

Gemini 3.7 Flash Lands at Half Price

Google shipped Gemini 3.7 Flash on August 13, just three weeks after 3.6 Flash, calling it its most intelligent workhorse model yet for coding and agents. On Google's own DeepSWE v1.1 evaluation it scores 65.3% against 49.0% for the model it replaces.

Pricing is the aggressive part: $0.75 per million input tokens and $3.75 per million output, with the same ~1,048,576-token context window. That rate is introductory and expires December 31, 2026 — on January 1 it becomes $1.50 and $7.50. Access is live through the Gemini API, Google AI Studio, Android Studio, Google Antigravity, and the Gemini Enterprise Agent Platform.

Do this: Write down the one internal tool or agent you keep wishing existed and scope it this weekend while the discount is live. You have roughly four and a half months of half-price, frontier-grade coding — and the window closes December 31.
STORY 04

OpenAI Previews Ultrafast Mode with Cerebras

Also on August 13, OpenAI previewed Ultrafast mode — a new API service tier running GPT-5.6 Sol at up to 750 output tokens per second, up to 14× faster than Standard processing. It runs on Cerebras wafer-scale hardware, the follow-through on a multi-year agreement the two companies signed in January to deploy up to 750 megawatts of Cerebras inference systems through 2028.

Same model, same intelligence — just dramatically faster. Two caveats: it's a limited preview open to a small set of customers, and OpenAI published neither pricing nor a general-availability date.

Do this: Make a list of every place in your product or workflow where someone stares at a spinner waiting on a model. That's your upgrade list for the day Ultrafast (or a competitor) goes GA — the people with the list ready ship the fast version first.
STORY 05

Apple Builds Its Own Model for China

Reuters reported on August 14 that Apple has trained its own large language model for the Chinese market with support from Alibaba — a reversal of its plan to license a domestic Chinese model to power Apple Intelligence there.

The notable detail: Apple is now the first foreign company approved by the Chinese government to offer its own proprietary AI model inside the country. Apple Intelligence is expected to launch in China in the coming months after years of regulatory delay, with Huawei as the competitive target.

Do this: Open your analytics and look at your actual geography split. If a meaningful slice of your audience or customers sits in a market with its own AI rules, "one AI product, one global market" is already over for you — plan for it now instead of discovering it at launch.
QUICK HITS

DeepSeek's 1,100% Price Hike · Anthropic's First Profit · Prompt Injection in Court

DeepSeek launched V4-Pro — stronger agent and cybersecurity performance, plus native OpenAI Responses API support — and raises V4 API prices by up to 1,100% starting August 16 (16:00 UTC), moving to peak/off-peak billing. V4-Flash output goes from a flat $0.28/M to $1.32/M at peak ($0.66 off-peak); V4-Pro output goes from $0.87/M to $3.96/M at peak. Re-check your batch jobs.

Anthropic reportedly cleared $11.5B+ in Q2 revenue — a 14-fold jump year over year and its first profitable quarter on an adjusted basis — ahead of a fall IPO. It also raised its catastrophic-misalignment risk rating from "very low" to "low" (citing broader uncertainty, not new evidence of danger) and disclosed it is shelving a more capable internal "Model 2" with no release plans. OpenAI told investors enterprise revenue has overtaken consumer ChatGPT for the first time, driven heavily by Codex; its public S-1 is expected within weeks at a reported $852B–$1T valuation.

Grok 4.6 is now live in GitHub Copilot and Cursor. And a Connecticut judge sanctioned a self-represented litigant for hiding white-on-white prompt-injection text in court filings — instructions telling any AI reading the document to rule in his favor. It's the first documented prompt-injection attempt aimed at a U.S. court. It did not work.

Do this: If you run DeepSeek batch jobs, re-quote them at the new rates tonight and shift anything that can wait into the off-peak window — off-peak is half the peak price, so that one scheduling change cuts the hike in half.
✅ DO THIS TODAY

Your Moves, Smallest First

1. Thirty seconds: Gemini → Settings → Media Watermark. Set it the way your work needs, and add one line to your client contracts saying how AI content is labeled.

2. Tonight (deadline is real): Using DeepSeek? Re-quote your batch jobs at the new rates and move what you can to off-peak before the Aug 16 (16:00 UTC) hike.

3. This weekend: Run 10 samples of your most expensive recurring API task through Qwen 3.8 27B locally (LM Studio or Ollama). Find out whether that bill is still worth paying.

Ongoing: Scope the internal tool you keep wishing existed while Gemini 3.7 Flash is half price (through Dec 31), and keep a list of every "spinner moment" in your workflow for the day ultra-fast inference goes GA.

Master AI. Build Your Empire.

Open weights, half-price frontier models, and 14× faster inference all landed in the same week. The tools have never been cheaper — the hard part is knowing which ones to actually build on. That's exactly why I built the AI Creators Roundtable.

Vetted prompts & workflows Creator network & feedback Weekly live AMAs Job board & client leads Founding rate — locked in
Join The Roundtable →
7-day money-back guarantee · Cancel anytime