FEATUREDAI Chat-Group Daily (群聊日报)· atomZH10:23 · 07·17
→Kimi K3 tops Frontend Code Arena, weights to open-source, early tests show brilliance and burnout
Kimi K3 hit #1 on Frontend Code Arena with 1679 points, beating Claude Fable 5's 1631 and taking six of seven frontend domains. It packs 2.8T params, 1M context, $3/$15 per million tokens, with full weights opening by July 27. Early testers got mixed results: one user's 199-yuan monthly plan produced stunning particle VJ effects from chat history, while another burned through a $40 coding plan in five hours as the model looped on a domain spelling error. Benchmark trust is shaky—GLM-5.2 scored well on paper but felt worse than 5.5 in practice. Writing style drew split reactions: less AI flavor but forced casual tone, nowhere near the natural Chinese of the old Opus 4.6. Same day, GPT-5.6's frontend taste was called 'very Claude-like,' Sol traced a deadlock only reproducible on Ubuntu, Linus told kernel devs AI is here to stay, and Schema harness pushed ARC-AGI-3 efficiency to 98.98% by making models think like physicists.
#Code#Agent#Reasoning#Kimi
why featured
Featured · importance 84 · hook + knowledge + resonance
editor take
Kimi K3 tops Frontend Code Arena, weights open July 27, but early testers split: stunning output vs. $40 burned on a spelling loop.
sharp
Three things happened at once here: K3 beat Claude Fable 5 on Frontend Code Arena, costs about a third as much, and full weights drop by July 27. 2.8T params, 1M context, $3/$15 per million tokens—on paper, it's a strong open-source coding option.
But early testers tell a messier story. One user's 199-yuan monthly plan produced stunning particle VJ effects from chat history. Another burned through a $40 coding plan in five hours because the model kept looping on a domain spelling error. That's a real reliability gap.
Benchmark trust is also shaky. GLM-5.2 scored well on paper but felt worse than 5.5 in practice, so the group is skeptical about K3's scores until they test it themselves. Writing style got mixed reviews too: less AI flavor, but forced casual tone that doesn't match the natural Chinese of the old Opus 4.6.
I'd treat K3 as a strong frontend coding option with open weights, not a blanket closed-source killer. The post doesn't cover non-frontend tasks systematically, and the distillation question is still open.
HKR breakdown
hook ✓knowledge ✓resonance ✓