Renting The Expert | Daily AI News Brief (Aug 24)
Friday was about agents moving into the software you already use. Today is about something narrower and stranger โ renting the expert. A model that turns your deck into a finished video. A professor who's an avatar. A research agent that reproduces papers. And in the middle of all that, the thing most of us build on went dark for a few hours.
Alibaba's Model Turns Your PDF Into A Video
Let's be straight about the date first: this is not from today. Public beta opened August 6, Alibaba showcased it in Hangzhou on the 10th, and their own blog post went up on the 13th. It's back in circulation because Alibaba just raised $10.2B in Hong Kong to fund exactly this kind of work. I'm covering it because most of you haven't seen it yet โ not because it's new. That distinction matters and I'd rather say it than let you assume.
Wan3.0 generates up to 30 seconds of video in a single pass, with audio made in that same pass, from text, images, audio, video โ and actual documents. A slide deck, a Word doc, a spreadsheet: pdf/doc/xls/ppt/txt/md, one file or link, up to 100 MB and 50 pages. You hand it a pitch deck, it hands you back a video with a soundtrack. That's a workflow a lot of people are currently paying a freelancer for.
Pricing is per second of output โ $0.05 at 480p, $0.10 at 720p, $0.20 at 1080p. So a full 30-second 1080p clip runs $6. Reference images, audio, documents and web pages aren't billed at all; reference video seconds are. Rate limit is 30 requests/min, 2 concurrent.
Now let me slow you down, because there's a lot of confidently wrong information about this model. Two things people are repeating that aren't true. It does not do native 4K โ the pricing table has three tiers and 4K isn't one of them. And the weights are not open. The Apache-2.0 weights being cited in 1.3B/14B variants actually belong to Wan 2.1, from February 2025. Recycled specs, repeated until they sounded like fact. Alibaba's last open-weight video flagship was Wan 2.2 (July 2025); everything from 2.5 onward is API-only.
One more thing, and it cuts against my own enthusiasm. Wan3.0 isn't on the Artificial Analysis Video Arena, so there's no independent measure of how good it is. Every quality claim right now is Alibaba's own.
Creator takeaway: I wouldn't call this an open model. I'd call it a rented one. That's not a reason to skip it โ $6 is cheap enough to find out for yourself whether document-to-video actually works for your workflow. Just go in knowing nobody neutral has graded it.
Sources: Read the announcement (Alibaba Cloud) ยท Document-to-video breakdown
Claude Went Down โ And That's Not Really The Story
Elevated errors flagged at 05:06 UTC, root cause identified around 05:27, remediation running past 06:42, everything resolved by roughly 08:30. It affected Mythos 5, Fable 5, Opus 5 and Opus 4.8 โ and it hit claude.ai, the API, Claude Code and Cowork. Every surface at once.
I should tell you I'm not neutral here. This show runs on Anthropic tooling. When Claude's down, my day gets worse too.
But the outage isn't the story. One bad morning is just infrastructure. Here's the story: this is the seventh disruption this month โ the 5th, 12th, 13th, 16th, 18th, 20th, and today. That's seven in twenty days, and that's a pattern.
I wouldn't call that a reliability problem, exactly. I'd call it a dependency problem โ and it's yours, not theirs. If your entire client workflow sits on one provider's API and that provider has a bad Tuesday, you don't have a technical issue. You have a business continuity issue, and your client doesn't care whose fault it was.
Creator takeaway: The practical move is boring and it works. Know which one thing in your stack, if it died for three hours, would stop you delivering. Then find out today what your fallback is. Not build it โ just know it. Twenty minutes of thinking that saves you a very bad afternoon.
Sources: Anthropic status page ยท Read more (Android Authority)
Harvard Is Selling You An AI Professor For $699
Harvard Business School has an eight-week online bootcamp called Foundry. $699. You practice your pitch, and an AI avatar of the faculty critiques it โ simulated board meetings, simulated sales calls. There are live human sessions weekly too, but the feedback on your pitch comes from the avatar. One of them represents senior lecturer Jeff Bussgang.
I have to disclose something, and it's the whole reason I'm covering this. Those avatars were built by HeyGen โ the same platform I use to deliver this show. I'm not observing this from outside. I'm in the same business. It's also not new; it rolled out in April 2026, and it's trending because TechCrunch wrote it up on August 22.
Here's what I actually think. An avatar critiquing your pitch at two in the morning is genuinely useful, and cheaper than any equivalent access to that faculty. But be precise about what's being sold. You're not buying access to those professors. You're buying a model trained on their material, wearing their face. Those are different products, and the second shouldn't be priced like the first.
Creator takeaway: The thing I'd watch is whether the face is doing persuasion work that the content can't. If the same feedback in plain text would feel thin, the avatar isn't adding teaching โ it's adding authority. That's the line, and I don't think anyone's drawn it clearly yet, including the people selling it.
Sources: Read more (TechCrunch) ยท Read more (Inc.)
NVIDIA's Groq Chip, And A 27B Agent That Isn't What It Looks Like
NVIDIA's Groq 3 LPX went into full production today โ the inference chip from the roughly $20B Groq deal. 3,400 output tokens/sec on Gemma 4 31B in Artificial Analysis benchmarking, 256 LPU accelerators per rack, each carrying 500 MB SRAM and 150 TB/s SRAM bandwidth. It extends Vera Rubin NVL72, and Nebius is the first AI cloud to adopt it, with racks online later this year. You won't touch this directly; you'll feel it as agents that stop feeling laggy.
A London lab called Inherent has Faraday, a 27B agent that reproduces scientific papers and reportedly beats Claude Opus 4.8 and GPT-5.5 at it โ evaluated on Replica, 310 tasks drawn from ~100 papers, built on a Qwen base, with a ~$50M seed led by Index Ventures. Two corrections. It launched on August 14 โ ten days old, not new. And Faraday calls GPT-5.5 Codex to do its actual coding. So it's not a small model beating big ones. It's a small model directing a big one, and beating that big one used alone.
And briefly, reported but not confirmed to primary sources: Hugging Face is said to be exploring a sale around $13B, and Apple reportedly cut 200+ roles across Siri, Vision Pro and AI teams.
Sources: Groq 3 LPX (NVIDIA) ยท Read more (CNBC) ยท Faraday research (Inherent)
Actionable Takeaways for Creators & Solos
- If you make video, spend $6 on Wan3.0 and feed it a real deck. Cheapest way to find out whether document-to-video works for you.
- Write down the one dependency that would stop you delivering if it died for three hours โ and know your fallback. Just know it.
- If you sell expertise, look hard at Harvard's price point, because the floor under "access to an expert" just moved.
- Treat "open weights" claims as unverified until you've seen the actual repo. Today's Wan3.0 story is the case study.
- Latency is about to stop being your excuse โ so make sure the workflow is actually right.