FEATUREDLatent Space· rssEN17:14 · 04·07
→Extreme Harness Engineering for Token Billionaires: 1M LOC, 1B toks/day, 0% human code, 0% human review
OpenAI Frontier says it built an internal beta over five months with a repo above 1M LOC, over 1B tokens per day, and 0% human-written or human-reviewed code before merge. The post says the team treated failures as missing capability, context, or structure, then used Symphony orchestration, specs, tests, observability, and sub-1-minute build loops to constrain Codex. The shift to watch is from humans reviewing code to humans designing the harness; the $2k-$3k/day cost is cited secondhand in the post.
#Agent#Code#Tools#OpenAI
why featured
Featured · importance 82 · hook + knowledge + resonance
editor take
OpenAI Frontier’s 1M LOC, 1B-token/day claim is not a coding-agent victory lap; it’s a harness discipline story most teams cannot fake.
sharp
OpenAI Frontier’s sharp claim is not 0% human-written code; it is 0% human review before merge at a repo above 1M LOC. Ryan Lopopolo gives concrete machinery: more than 1B tokens per day, five months on an internal beta, Symphony orchestration, PRD-like specs, tests, observability, and sub-one-minute build loops. Failures get mapped to missing capability, context, or structure, not another prompt tweak.
I don’t buy the “negligent if you aren’t using 1B tokens/day” posture. The $2k-$3k/day spend is secondhand and depends on market-rate plus caching assumptions. The system is also an OpenAI internal beta, not a public SLA surface. Cursor and Devin have spent the last year selling agent throughput; this is the colder version. Humans move from reviewers to harness designers. Without that harness, zero-review code is just incident debt with better branding.
HKR breakdown
hook ✓knowledge ✓resonance ✓