ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
33 srcsignal 72%cycle 04:32

posts · 2026-07-01

5 items · updated 3m ago
RSS live
2026-07-01 · Wed
00:20
84d ago
Latent Space· rssEN00:20 · 07·01
Sierra's Natalie Meurer: Forward deployed engineering is about customer accountability, not a fixed skill set
At the AI Engineer World's Fair, Sierra's Head of Agent Engineering Natalie Meurer said forward deployed engineering lacks a consistent definition but is unified by accountability to customers. Sierra calls the role 'agent engineer'—a 120+ person team building custom conversational AI agents for enterprise customer service. Most customer-specific work happens at the orchestration layer above the models. Voice agent design also requires 'taste' for what sounds human. She sees product and customer-facing engineering roles starting to converge.
#Sierra#Natalie Meurer#Palantir
editor take
Sierra's 120+ agent engineers are defined by customer accountability, not a fixed skill set.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
00:00
84d ago
● P1AI HOT (Curated Pool)· aihot-apiZH00:00 · 07·01
xAI launches Voice Agent Builder beta for no-code voice agents
xAI packaged Grok Voice into a no-code platform, now in beta as of July 1. You describe the call flow in plain language, upload docs as a knowledge base, and connect tools like calendars or ticketing systems—then you get a working voice agent. It uses a speech-to-speech path instead of chaining ASR→LLM→TTS, which xAI claims cuts latency and failure points. Pricing is $0.05/min of audio plus $0.01/min for a platform-provided number. xAI also published τ-voice Bench scores: Grok Voice Think Fast 1.0 hit 67.3% overall, versus 43.8% for Gemini 3.1 Flash Live and 35.3% for GPT Realtime 1.5. Take the benchmark with a grain of salt—it's xAI's own test, and third-party results aren't out yet.
#xAI#Grok Voice#Gemini 3.1 Flash Live
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
xAI's no-code voice agent at $0.05/min, but the benchmark is their own—discount the scores for now.
sharp
The draw here is how low xAI set the barrier: describe the call flow in plain language, upload docs, connect a calendar or ticketing tool, and you've got a working voice agent in about two minutes. It's speech-to-speech, not the usual ASR-LLM-TTS chain, so latency and failure points should drop. At $0.05/min of audio plus $0.01/min for a platform number, the pricing is competitive. But I'd discount that τ-voice Bench score. Grok Voice Think Fast 1.0 at 67.3% versus Gemini 3.1 Flash Live at 43.8% and GPT Realtime 1.5 at 35.3%—those gaps are suspiciously wide. It's xAI's own benchmark, and they haven't released the test set, scoring rubric, or exact model configs. No third-party replication yet either. It's not that the model is bad; it's that this number reads more like marketing than an engineering comparison right now. The real gaps are what the post doesn't cover: concurrency limits, regional availability, and SLAs. Those are what matter in production. For a few test calls, two-minute setup is genuinely fast. For core business lines, I'd wait for more detail.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
00:00
84d ago
Hugging Face Blog· rssEN00:00 · 07·01
Hugging Face and Cerebras bring Gemma 4 to real-time voice AI
Only the title is disclosed; the post does not spell out technical details. Hugging Face and Cerebras are collaborating to deploy Google's Gemma 4 model for real-time voice AI. Cerebras chips excel at low-latency inference, likely serving as an acceleration layer for voice interaction. The post does not clarify whether this is an on-device or cloud solution, nor which voice tasks are supported.
#Hugging Face#Cerebras#Google
editor take
Hugging Face + Cerebras are putting Gemma 4 on low-latency chips for voice AI, but the post doesn't say if it's on-device or cloud, or which tasks.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0

more

feeds

admin