ax@ax-radar:~/podcasts/latent-space $ ls -t podcasts/
40 srcsignal 72%cycle 04:32

podcasts

50 episodes · updated 3m ago
6 channels tracked
tierfeaturedallincludes low-score
Latent Space50 episodes
2026-07-16 · Thu
2026-07-14 · Tue
2026-07-08 · Wed
2026-07-03 · Fri
2026-07-02 · Thu
07:10
26d ago
Latent Space· rssEN07:10 · 07·02
Fable 5 returns with safety guardrails, pushing devs toward multi-model orchestration
Anthropic re-enabled Claude Fable 5 with updated safety guardrails that may route some requests to Opus 4.8; biology/chemistry classifiers remain overly broad. Cursor reports Fable 5 leads its evals but is the most expensive per task; Devin and Perplexity have restored support. Developers are adopting multi-model orchestration, using Fable only for high-value reasoning and delegating execution to other models. On the open-source side, Z.ai launched ZCode, an official IDE for GLM-5.2, which leads open models on APEX-SWE Integration with 55.3% Pass@1. Inference optimizations include vLLM's DSpark speculative decoding for DeepSeek (~250 tok/s on 8×B300) and a GLM-5.2 DSpark preview claiming ~1.5× faster decode. Agent infrastructure sees 'wiki memory' as a new pattern: LangChain released OpenWiki, and Weaviate's Engram resolves contradictions before committing memories. The post does not disclose Fable 5's specific pricing or Opus 4.8 trigger conditions.
#Code#Anthropic#Claude Fable 5#Opus 4.8
editor take
Fable 5 is back but some requests get routed to Opus 4.8; safety guardrails remain overly broad.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
2026-07-01 · Wed
23:52
26d ago
Latent Space· rssEN23:52 · 07·01
Autoresearch: The feedback loop behind self-improving agents
Introspection CEO Roland Gavrilescu explains autoresearch at AIEWF: an outer loop where agents maintain and improve the primary system. Three patterns emerge—treat the loop as the product, package human expertise and evals into portable 'recipes,' and optimize for cheaper, better systems over time. Gavrilescu previously worked on agent infra at xAI. He compares the open-source Pi framework to Linux and positions Introspection as its Red Hat.
#Agent#Benchmarking#Reasoning#Introspection
editor take
Introspection sells the feedback loop as the product, open-sources Pi as the Linux of agent infra, and wants to be its Red Hat.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
00:20
27d ago
Latent Space· rssEN00:20 · 07·01
Sierra's Natalie Meurer: Forward deployed engineering is about customer accountability, not a fixed skill set
At the AI Engineer World's Fair, Sierra's Head of Agent Engineering Natalie Meurer said forward deployed engineering lacks a consistent definition but is unified by accountability to customers. Sierra calls the role 'agent engineer'—a 120+ person team building custom conversational AI agents for enterprise customer service. Most customer-specific work happens at the orchestration layer above the models. Voice agent design also requires 'taste' for what sounds human. She sees product and customer-facing engineering roles starting to converge.
#Sierra#Natalie Meurer#Palantir
editor take
Sierra's 120+ agent engineers are defined by customer accountability, not a fixed skill set.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
2026-06-30 · Tue
2026-06-26 · Fri
2026-06-24 · Wed
2026-06-22 · Mon
2026-06-18 · Thu
2026-06-16 · Tue
2026-06-12 · Fri
05:34
46d ago
Latent Space· rssEN05:34 · 06·12
Stop prompting, start stacking loops to let AI run itself
Peter Steinberger, Boris Cherny, and Andrej Karpathy all land on the same point: stop being the human in the loop—you're the bottleneck. Karpathy, on Autoresearch, says refactor everything so you hit go once and the system runs fully autonomous. The post calls this 'stacking loops' and shows two diagrams of loops we're already inside. The salty lesson: don't fix things yourself; build goals and orchestration that scale with more agents. Separately, Anthropic silently degraded Claude Fable 5 for some AI-research use cases, reversed it within a day after backlash. Simon Willison welcomed the rollback; Ryan Greenblatt and Natasha/Lambert argued the real error was opaque model-layer sabotage, not the safeguards themselves. Fable 5 hit 87.8% on WeirdML and #1 on FrontierSWE, but one dev spent ~$250 on a PR and found it not worth it; Cline noted cheaper models plus adversarial review loops often match it on cost/perf.
#Agent#Code#Anthropic#Claude Fable 5
editor take
Karpathy, Steinberger, and Cherny all say the same thing: stop being the human in the loop—stack loops so you hit go once and the system runs fully autonomous.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
2026-06-11 · Thu
2026-06-09 · Tue
2026-06-06 · Sat
04:34
52d ago
Latent Space· rssEN04:34 · 06·06
[AINews] Not Much Happened Today
AINews checked 12 subreddits and 544 Twitter sources for June 4–5, 2026, summarizing model, agent-evaluation, and open-release updates from Anthropic, Sakana AI, Google, Ideogram, and NVIDIA.
#Agent#Benchmarking#Inference-opt#Anthropic
editor take
AINews scanned 12 subreddits and 544 Twitter sources; ignore the sleepy title, agent evals and open weights carry the issue.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
2026-06-05 · Fri
06:44
53d ago
Latent Space· rssEN06:44 · 06·05
[AINews] Not Much Happened Today
AINews summarized June 3-4, 2026 updates, covering NVIDIA Nemotron 3 Ultra, Anthropic’s recursive self-improvement framing, ChatGPT crossing 1B MAU with improved memory, and Cloudflare’s acquisition of VoidZero.
#Agent#Memory#Benchmarking#NVIDIA
editor take
AINews scanned 12 subreddits and 544 Twitters; NVIDIA’s 550B open MoE lands harder than the RSI narrative.
HKR breakdown
hook knowledge resonance
open source
58
SCORE
H0·K1·R0
2026-06-04 · Thu
03:24
54d ago
Latent Space· rssEN03:24 · 06·04
[AINews] Reve 2 and Ideogram 4: Layouts in Image Generation
Latent Space summarized AI News for June 2-3, 2026 after checking 12 subreddits and 544 Twitter accounts, covering MAI-Thinking-1 with 97% on AIME 2025, Ideogram 4.0’s open weights, and Google’s Gemma 4 12B on-device multimodal release.
#Multimodal#Reasoning#Agent#Latent Space
editor take
Ideogram 4.0 ranks #1 open in Arena; GPT-Image-2 still leads, so open image models win distribution before parity.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1
2026-06-03 · Wed
2026-06-02 · Tue
2026-06-01 · Mon
2026-05-30 · Sat
01:57
59d ago
Latent Space· rssEN01:57 · 05·30
[AINews] Founders and Forward Deployed Engineers
Latent Space published its May 28–29, 2026 AINews issue after checking 12 subreddits and 544 Twitter accounts. The post covers Claude Opus 4.8 benchmark friction, multi-turn RL tokenization bugs, open-weight model adoption, managed agents in Gemini API, and OpenAI Codex Windows control.
#Agent#Code#Benchmarking#Latent Space
editor take
AINews checked 12 subreddits and 544 accounts; I’d chase Token-In Token-Out bugs before another Opus 4.8 benchmark fight.
HKR breakdown
hook knowledge resonance
open source
58
SCORE
H0·K1·R0
2026-05-28 · Thu
2026-05-27 · Wed
2026-05-23 · Sat
04:21
66d ago
Latent Space· rssEN04:21 · 05·23
[AINews] All Model Labs Are Now Agent Labs
Latent Space summarized AI News for May 4–5 after checking 12 subreddits and 544 Twitter accounts, arguing that OpenAI, AI21, DeepSeek and other model labs are moving product focus from standalone models to agents, harnesses, workflows, UI, memory and cost structure.
#Agent#Tools#Code#Latent Space
editor take
Latent Space checked 12 subreddits and 544 accounts; model labs are adding agent shells, and closed harnesses can choke API competition.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1
2026-05-22 · Fri
2026-05-21 · Thu
2026-05-20 · Wed
2026-05-19 · Tue
2026-05-18 · Mon
2026-05-15 · Fri
00:30
74d ago
Latent Space· rssEN00:30 · 05·15
[AINews] Everything is Conductor
Latent Space summarized AI News for May 13-14, 2026 after checking 12 subreddits and 544 Twitter accounts, covering Codex mobile workflows, the GitHub Copilot App preview, Anthropic Claude Code restrictions, and Figure’s 24/7 autonomous package-sorting livestream.
#Agent#Code#Robotics#Latent Space
editor take
Latent Space checked 12 subreddits and 544 Twitter accounts; agent-first IDEs are crowded, while Claude Code throttling exposes the pricing wall.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R1
2026-05-14 · Thu
2026-05-13 · Wed
02:47
76d ago
Latent Space· rssEN02:47 · 05·13
[AINews] The End of Finetuning
Latent Space frames OpenAI’s deprecation of finetuning APIs as the lead item in its May 11–12, 2026 AI News issue, which aggregates signals from 12 subreddits and 544 Twitter accounts across benchmarks, agent systems, inference stacks, multimodal releases, and training efficiency work.
#Fine-tuning#Benchmarking#Inference-opt#OpenAI
editor take
OpenAI deprecated finetuning APIs; RSS gives snippets only. I don't buy the death claim—Cursor and Cognition are increasing open-model RLFT.
HKR breakdown
hook knowledge resonance
open source
71
SCORE
H1·K1·R1
2026-05-12 · Tue
04:33
77d ago
● P1Latent Space· rssEN04:33 · 05·12
Thinking Machines' Native Interaction Models: TML-Interaction-Small 276B-A12B Advances Realtime Voice
Thinking Machines released TML-Interaction-Small, a 276B-parameter MoE model with 12B active parameters, and the post says it advances realtime voice through 200ms time-aligned microturns, encoder-free early fusion for audio and images under 200ms, and benchmark wins over GPT-Realtime-2 and Gemini 3.1-Flash.
#Multimodal#Audio#Agent#Thinking Machines
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
Thinking Machines moved realtime voice inside the model loop: 276B MoE, 12B active, 200ms microturns. That hits harder than another chat leaderboard.
sharp
Thinking Machines is betting on the interaction clock, not a speech wrapper. TML-Interaction-Small is a 276B MoE with 12B active parameters, encoder-free early fusion for audio and images, and 200ms time-aligned microturns. That attacks the hand-coded turn logic sitting between VAD, ASR, LLM, and TTS stacks. I’d discount the official leaderboard for now: wins over GPT-Realtime-2 and Gemini 3.1-Flash on BigBench Audio, IFEval, and FD-bench lack reproducibility details in the snippet. The stronger signal is the new task shape: TimeSpeak, CueSpeak, RepCount-A, and ProactiveVideoQA test when to talk, when to stay silent, and when visual evidence becomes available. OpenAI’s 4o “Her” demo sold presence; Thinking Machines is trying to own timing.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
2026-05-09 · Sat
2026-05-05 · Tue
2026-05-04 · Mon
23:29
84d ago
Latent Space· rssEN23:29 · 05·04
[AINews] The Other vs The Utility
Latent Space summarized AI News for May 1-4, 2026, covering 12 subreddits and 544 Twitter accounts, with focus on Claude as “the Other,” GPT as a utility, Sierra’s roughly $1B raise, and concrete threads on agent harnesses, Codex token costs, and benchmark design.
#Agent#Code#Benchmarking#Latent Space
editor take
AINews scanned 12 subreddits and 544 Twitter accounts; I trust the 52.8%-to-66.5% harness gain over Claude worship discourse.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1

more

feeds

admin