ax@ax-radar:~/podcasts/dwarkesh $ ls -t podcasts/
33 srcsignal 72%cycle 04:32

podcasts

11 episodes · updated 3m ago
6 channels tracked
tierfeaturedallcurated only
Dwarkesh Patel11 episodes
2026-09-08 · Tue
2026-08-07 · Fri
17:17
46d ago
● P1Dwarkesh Patel· rssEN17:17 · 08·07
The Era of Continual Learning: AI Models Update Weights After Deployment
Dwarkesh Patel argues that once models can update weights continuously from deployment, the whole AI landscape shifts. Instead of train-then-deploy, models will learn from every interaction like a human practicing saxophone—notes alone can't transfer the skill. This breaks the current regulatory assumption of pre-deployment checks; monthly or quarterly risk inspections make more sense. Alignment research must pivot from controlling frozen weights to preventing jailbreaks or backdoors during constant updates. Commercially, the leading lab's advantage compounds: more usage yields more feedback, making the model smarter and pushing labs to ship their best models earlier. Switching costs become massive—ditching a model that has learned your org's context for months is like firing a veteran employee for a clueless intern, creating durable high margins. Enterprises will face a trade-off: accept lock-in for a model that improves with use, or lose access to top-tier AI. Labs may subsidize users who allow training on their sessions. Continual learning also increases AI mind diversity, breaking today's monoculture of a few similar base models. On the inference side, per-company full weight updates create huge batching economies; for a sparse model like DeepSeek v3, optimal batch size exceeds 2,400 concurrent sequences.
#Inference-opt#Dwarkesh Patel#Anthropic#DeepSeek
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
This isn't news — it's Dwarkesh's 8 predictions on models updating weights post-deployment. Both sources are his own blog and YouTube, with zero external cross-coverage, so read it as an opinion pi...
sharp
Dwarkesh skipped the interview format and wrote a long-read himself, laying out what changes if continual learning — models updating weights from live usage — actually ships. Both sources are identical content across his blog and YouTube, with no independent outlets picking it up, so don't mistake this for industry consensus. His core bets: post-deployment learning breaks the 'evaluate before release' regulatory model, pushing toward monthly or quarterly audits instead. Alignment research would need to shift from locking down frozen weights to preventing backdoors in constantly updating ones. First-mover advantage compounds because more usage makes the model smarter, and switching costs become real — like firing an employee who's accumulated months of organizational context. The logic holds together, but there's zero external confirmation. No lab has said they're doing this, and he doesn't name a technical path. I'd treat it as a thought experiment — the direction is interesting, but it's one person drawing the map without ground truth yet.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
2026-08-03 · Mon
2026-06-08 · Mon
2026-06-04 · Thu
2026-05-22 · Fri
2026-05-16 · Sat
2026-04-29 · Wed
2026-04-27 · Mon

more

feeds

admin