ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

all posts

50 items · updated 3m ago
RSS live
2026-07-05 · Sun
21:00
23d ago
Financial Times · Technology· rssEN21:00 · 07·05
UK regulator warns of AI arms race in financial services
The UK's Financial Conduct Authority (FCA) warns that financial firms are racing to deploy AI, creating an 'arms race' that regulators must match. It fears rapid AI use in trading, risk management, and customer service could pose systemic risks. The post does not disclose specific cases or timelines.
#Financial Conduct Authority (FCA)
editor take
UK regulator FCA warns financial firms' AI race is an arms race regulators can't keep up with.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
18:47
23d ago
Hacker News Frontpage· rssEN18:47 · 07·05
Dartmouth stats course pilot: AI textbook lifts final exam scores by 0.71–1.30 SD
Dartmouth deployed Phosphor, an AI learning platform, in an intro stats course with 151 students. It was optional and ungraded, yet 90.2% of students used it. Full dosage was linked to a 0.71 SD final exam gain after controlling for prior scores, and 1.30 SD without controls. The platform embeds AI-graded quizzes into readings; Claude Sonnet 4.6 grades short-answer questions against rubrics. In Module 2, students complained the auto-grader was too rigid, so quizzes switched to multiple-choice only—the paper hints this may have hurt outcomes. The post does not report inter-rater reliability or how far Claude's grading diverged from human graders.
#RAG#Dartmouth College#Phosphor#Claude Sonnet 4.6
editor take
Dartmouth's Phosphor platform got 90% voluntary usage and a 0.71 SD exam gain, but the paper doesn't report inter-rater reliability for Claude's grading.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
17:43
23d ago
AI HOT (Curated Pool)· aihot-apiZH17:43 · 07·05
I Accidentally Started a Small Business Three Weeks Ago
A father built a communication app for his non-verbal autistic son. In the speech therapy waiting room, it made every mom and the therapist sob. He accidentally found product-market fit and decided to scale it despite his busy life. The post also details the long, frustrating journey of realizing his child's speech delay and dodging pseudoscience.
editor take
A dad built a communication app for his non-verbal autistic son; it made every mom and the therapist sob in the waiting room.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H1·K0·R1
17:05
23d ago
● P1Hacker News Frontpage· rssEN17:05 · 07·05
Zuckerberg admits AI agent development slower than expected
At an internal town hall, Mark Zuckerberg said AI agent development hasn't accelerated as executives expected. Meta cut ~8,000 jobs in May and reassigned 7,000 to AI groups, but he admitted the layoffs weren't 'clean' and the new AI-focused structure hasn't paid off yet. He expects returns from AI investments in 3–6 months. Meta plans to spend up to $145B on AI infrastructure this year.
#Agent#Meta#Mark Zuckerberg
why featured
Featured · importance 96 · hook + knowledge + resonance
editor take
Zuckerberg told staff AI agents aren't progressing as fast as hoped. Both sources cite the same internal town hall — consistent but no public Meta comment yet.
sharp
This comes from Meta's internal town hall on Thursday. Both TechCrunch and aihot are relaying a Reuters report, so we're looking at one original source, not multiple independent confirmations. Zuckerberg said AI agent development hasn't "accelerated in the way" executives expected, and he admitted the May layoffs of 8,000 people plus reassigning 7,000 into AI teams wasn't as "clean" as it should have been. The new structure's upside hasn't materialized yet. I'd read this as internal pressure management rather than a product roadmap shift. He gave a 3-to-6-month window for seeing improvements from AI investments — that's a concrete timeline to bookmark and check back on. What's missing: Meta hasn't issued a public statement, and none of the coverage specifies which agent capabilities are lagging. Is it code generation, conversation quality, task completion rate? We don't know yet.
HKR breakdown
hook knowledge resonance
open source
96
SCORE
H1·K1·R1
15:00
23d ago
Financial Times · Technology· rssEN15:00 · 07·05
Data centres are a crucial test of US industrial resolve
FT argues that building data centres at scale is not just an AI infrastructure challenge but a political test of US manufacturing resolve. Permitting, power grids, and supply chains all expose weaknesses. If the US can't build data centres smoothly, other advanced manufacturing will struggle.
#Financial Times
editor take
FT frames data centre buildout as a political test of US industrial resolve, flagging permitting and grid bottlenecks.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R1
13:39
23d ago
Product Hunt · AI· rssEN13:39 · 07·05
CodeMote: Control Claude Code, Codex, and any CLI agent from your iPhone
CodeMote is an iOS app that lets you remotely control CLI agents on your machine or VPS from your iPhone. It offers a live terminal on the lock screen, push notifications when an agent needs approval, full diffs, and complete Git flow. The connection is directly encrypted, and your code never touches their servers. It supports Claude Code, Codex, and any CLI tool. The post does not disclose pricing details, only mentioning free options and a 1-month free trial.
#CodeMote#Claude Code#Codex
editor take
CodeMote puts a live CLI agent terminal on your iPhone lock screen with push approvals, but pricing details are missing.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
13:33
23d ago
Product Hunt · AI· rssEN13:33 · 07·05
Nixmac: Describe your Mac setup in plain English, Nix writes the config
Nixmac lets you describe your Mac setup in plain English, then auto-generates Nix code and applies it safely. It targets developers who want reproducible, version-controlled systems via Nix-darwin without writing Nix by hand. Just launched on Product Hunt, free and open-source. The post doesn't specify which model handles the NL-to-Nix translation, nor the accuracy rate or rollback mechanism.
#Nixmac#Nix#Nix-darwin#Open source
editor take
Nixmac auto-generates Nix config from plain English for reproducible Mac setups. Open-source, but no info on the NL model or accuracy.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
12:31
23d ago
Product Hunt · AI· rssEN12:31 · 07·05
Mozaik: TypeScript runtime for self-organizing AI agent teams
Mozaik is a TypeScript runtime for self-organizing AI agent teams. It enables concurrent work, event-driven reactions, intelligent communication, and autonomous collaboration decisions during execution. The post doesn't spell out technical implementation details or benchmarks, but positions itself as a tool for developers building complex multi-agent systems.
#Mozaik#JigJoy
editor take
Mozaik lets agents self-organize without manual orchestration, but the post skips benchmarks and implementation details.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R0
07:59
23d ago
Hacker News Frontpage· rssEN07:59 · 07·05
Knowledge Should Not Be Gated: Google OKF lets LLMs read plain Markdown
Formaly argues that RAG gated knowledge behind vector databases and SDKs, making it unreadable to humans. Meanwhile, tools like Claude Code and Codex already proved LLMs prefer plain Markdown files like CLAUDE.md. Andrej Karpathy's LLM Wiki pattern formalizes this: a three-layer file structure (sources/, wiki/, schema) where the model maintains its own knowledge base, avoiding the 'retrieval tax' on every query. Google's Open Knowledge Format (OKF) v0.1, released in June, standardizes this as a vendor-neutral directory of Markdown files. The post doesn't disclose benchmarks or adoption cases—its core claim is that format walls, not paywalls, are the real gate.
#Google#Andrej Karpathy#Formaly
editor take
RAG gated knowledge behind vector DBs; Karpathy's LLM Wiki uses plain Markdown files the model maintains itself.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
07:50
23d ago
AI HOT (Curated Pool)· aihot-apiZH07:50 · 07·05
LlamaIndex releases legal-kb: agentic retrieval for legal docs
LlamaIndex open-sourced legal-kb, a reference app built on Index v2. It gives the model four tools—retrieve, find, read, grep—to search legal documents like a paralegal. The post doesn't disclose performance numbers or real-world use cases, but the idea is solid: put agent workflow into a professional domain, not just chatbots.
#LlamaIndex#Open source
editor take
LlamaIndex open-sourced legal-kb: four tools (retrieve, find, read, grep) to turn an LLM into a paralegal for document search.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
07:25
23d ago
Hacker News Frontpage· rssEN07:25 · 07·05
Fast Software, the Best Software
Craig Mod argues that software speed is a proxy for engineering quality and trust. He cites nvALT and Sublime Text as examples where millisecond responsiveness makes tools feel integrated, while Ulysses' occasional lag erodes confidence. Adobe Lightroom and Photoshop have slowed over time, leading him to pay for Affinity Photo and Figma—the latter, despite being browser-based, is so fast it delights him. Speed is a commercial asset: Sketch won market share from Adobe by being faster.
#Craig Mod#nvALT#Sublime Text
editor take
Craig Mod argues software speed is a proxy for engineering quality—Lightroom's bloat drove him to pay for Affinity Photo.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
05:01
23d ago
Hacker News Frontpage· rssEN05:01 · 07·05
HIC Mouse: Precision Editing Tools for AI Coding Agents
HIC Mouse is a file-editing tool for AI coding agents, offering coordinate-based editing, staged changes, and atomic rollback. It replaces simple string replacement with six declarative operations like INSERT and DELETE for surgical accuracy. Edits are staged for approval, inspection, or refinement before saving. Tool responses include contextual guidance and risk assessment. Free 14-day trial, no credit card required. The post doesn't specify supported models or IDE versions beyond VS Code Marketplace availability.
#Code#HIC AI#HIC Mouse
editor take
HIC Mouse replaces AI coders' sloppy find-and-replace with six coordinate-based edit commands, staging every change for approval before save.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
02:54
23d ago
Hacker News Frontpage· rssEN02:54 · 07·05
Backon: Python retry library with zero deps, circuit breaker, async native
Backon is a Python retry library with zero dependencies, native async support, and a built-in circuit breaker. It's useful for adding fault tolerance to microservices or API calls. Currently 5 points and 0 comments on HN—low traction but solid design. The post doesn't include benchmarks or comparisons with tenacity, so you'll need to test yourself.
#Backon#GitHub#Open source
editor take
Backon is a zero-dependency Python retry lib with native async and circuit breaker—no benchmarks vs tenacity yet, so test it yourself.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
00:00
24d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 07·05
ResearchStudio-Reel: Auto-Generate Posters, Videos, and Blogs from One Paper
Microsoft team open-sources a pipeline that turns a paper into editable posters, talk videos, and bilingual blogs. The key idea: extract once, reuse everywhere. Posters beat prior automated systems and single-shot LLMs on 84%–93% of papers. The post doesn't disclose runtime cost or latency.
#Microsoft#Claude Code#Codex
editor take
Microsoft open-sources a pipeline that extracts a paper once and auto-generates editable posters, videos, and blogs.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
2026-07-04 · Sat
23:51
24d ago
Hacker News Frontpage· rssEN23:51 · 07·04
Zo Computer: a personal cloud computer you control with natural language
Zo Computer is a cloud-based personal computer you control via chat—build websites, run a business, and call AI models. It integrates OpenAI, Anthropic, Google, DeepSeek, and 1000+ tools like Slack, Discord, Gmail. Users report replacing 12 tools and building a site in minutes. Free tier has daily limits; paid unlocks all models. Zo claims to be the "original OpenClaw" but requires no terminal or Mac Mini. The post doesn't disclose pricing tiers, model versions, or latency.
#Agent#Zo Computer#OpenAI#Anthropic
editor take
Zo Computer turns chat into a cloud desktop with 1000+ integrations and multiple models, but pricing and latency aren't disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
23:25
24d ago
Hacker News Frontpage· rssEN23:25 · 07·04
RidgeText uses in-memory layer queues for map composition so the LLM never touches GeoJSON
RidgeText keeps map data in server memory so the LLM only sees lightweight acknowledgments like '847 features queued' before calling generate_map to composite an image. A single wildfire dataset can hit 125K tokens—expensive and error-prone to pass through context. Their approach mirrors Mapbox's layer model: each retrieve_* tool appends a layer to an in-memory queue, and layers are composited in call order at render time. The queue expires after 30 minutes. The renderer currently uses a Mapbox Static API base plus canvas compositing, and can swap to headless Mapbox GL JS later without changing the LLM interface. The post does not disclose cost or latency figures.
#RidgeText#Mapbox
editor take
RidgeText keeps GeoJSON in server memory so the LLM only sees tiny acknowledgments—no 125K-token wildfire datasets clogging context.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
21:51
24d ago
Hacker News Frontpage· rssEN21:51 · 07·04
GPT-5.5 Codex reasoning-token clustering at 516/1034/1552 may degrade complex-task performance
A GitHub issue on the OpenAI Codex repo reports that GPT-5.5 reasoning tokens cluster around positions 516, 1034, and 1552. The reporter suspects this pattern hurts code generation quality on complex tasks. The post does not include benchmark numbers, reproduction steps, or a response from OpenAI—it's a community report for now.
#Code#Reasoning#OpenAI#GPT-5.5
editor take
GPT-5.5 reasoning tokens cluster at positions 516, 1034, 1552—community suspects it degrades complex code output.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
19:50
24d ago
● P1Hacker News Frontpage· rssEN19:50 · 07·04
AI reduces junior programmer jobs while total developer employment grows
Stanford ADP payroll data shows US software developers aged 22-25 fell 19% from their late-2022 peak, while ages 41-49 rose 14%. After controlling for firm-level shocks, young workers in AI-automatable occupations still saw a 16% relative decline. Entry-level postings dropped 28%, and CS grads hit 6.1% unemployment—higher than liberal arts majors. Yet total developer employment rose 4.4% over the same period because juniors are only ~8% of the workforce. The BLS category 'computer programmer' (coding to spec) fell 16% in one year; data scientists grew 12%. Meanwhile, GitHub added 36M new accounts and 121M repos in a year, 80% of newcomers used Copilot in their first week, and iOS App Store submissions reversed an eight-year decline with 24% growth in 2025. The author argues the long tail of new developers arrived—they just don't use the job title. The post does not provide data beyond early 2026.
#Code#Stanford Digital Economy Lab#ADP#Bureau of Labor Statistics
why featured
Featured · importance 92 · hook + knowledge + resonance
editor take
Junior dev jobs really are collapsing, but more people are writing code than ever — they just don't call themselves programmers.
sharp
Two sources are on this, both pointing to the same Stanford ADP payroll chart: devs aged 22–25 dropped 19% from their late-2022 peak, while every cohort over 30 grew. The Stanford team controlled for firm-level shocks and interest rate exposure, and the damage still concentrates in AI-automatable roles — that's what makes this more than a post-ZIRP hangover. The twist is that total developer employment rose 4.4% over the same period. Juniors are a small slice of the workforce, so their collapse barely moves the average, which is why aggregate studies keep finding nothing. I'd flag one caveat: using age as a proxy for experience is messy — a 23-year-old could be a senior, a 45-year-old could be a career switcher. But the direction holds. GitHub added 36M new accounts in a year, 80% used Copilot in week one, and iOS App Store submissions reversed an eight-year decline with 24% growth in 2025. The long tail of new builders showed up — they just don't have the job title. What I haven't seen yet: income data for these new builders. Are they making money, or just shipping side projects?
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
18:00
24d ago
TechCrunch AI· rssEN18:00 · 07·04
Midjourney wants Hollywood studios to reveal the details of their AI usage
Midjourney is pushing Disney, Universal, and Warner Bros. to disclose their own AI usage in an ongoing copyright lawsuit. The studios sued last year over Midjourney's ability to generate copyrighted characters; Midjourney claims fair use. The current fight is over discovery scope—a judge already ruled the studios must hand over some info, but the exact documents are still disputed.
#Vision#Midjourney#Disney#Universal
editor take
Midjourney flipped the script in its copyright fight, demanding Disney, Universal, and Warner Bros. disclose their own AI use—and a judge already ordered partial compliance.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R1
16:55
24d ago
Hacker News Frontpage· rssEN16:55 · 07·04
Plein Air: A painting matched to your current weather, right now
Plein Air picks a public-domain painting from the Met, Art Institute of Chicago, or Cleveland Museum based on your current weather and season. It uses free Open-Meteo data and rotates sources randomly. Tap the title to see why that painting was chosen. The post doesn't spell out mobile support or latency, but the idea is simple: let a Monet haystack sit through the same rain you're in.
#The Metropolitan Museum of Art#Art Institute of Chicago#Cleveland Museum of Art
editor take
Uses your live weather to pick a matching public-domain painting from museum collections — rain gets you a Monet haystack.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
16:30
24d ago
Hacker News Frontpage· rssEN16:30 · 07·04
Zig moves all package management from compiler to build system, shrinks binary 4%
Zig moved all package management logic from the compiler into the build system process. Commands like zig build and zig fetch now run in the maker process, and large parts—package fetching, HTTP client, TLS, Git protocol, compression libraries—are shipped as source code instead of being baked into the compiler binary. The compiler shrank 4%, from 14.1 to 13.5 MiB. Networking code now runs in ReleaseSafe mode, enabling better safety checks and CPU-specific optimizations. The change unblocks a build server protocol needed by ZLS. Four blocking issues remain; the author expects to finish by early August.
#Zig#Andrew Kelley#ZLS
editor take
Zig moved all package management out of the compiler into the build process — compiler binary shrinks 4%, and users can patch networking/crypto code without rebuilding the compiler.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
15:51
24d ago
TechCrunch AI· rssEN15:51 · 07·04
What is Mistral AI? Everything to know about the OpenAI competitor
TechCrunch profiles French AI lab Mistral, arguing that judging it as 'the OpenAI of Europe' sets it up for disappointment. Its chat product Vibe (formerly Le Chat) has far less brand recognition than ChatGPT, and Claude is more popular even among founders at Paris' Station F campus. The post notes Mistral has raised significant funding since its 2023 founding with the goal of 'putting frontier AI in everyone's hands,' but does not disclose specific funding amounts, valuation, revenue, or user numbers.
#Mistral AI#OpenAI#Anthropic#Open source
editor take
TechCrunch profiles Mistral: don't judge it as 'the OpenAI of Europe' — even Paris founders prefer Claude over its chat product Vibe.
HKR breakdown
hook knowledge resonance
open source
45
SCORE
H0·K0·R0
14:17
24d ago
Product Hunt · AI· rssEN14:17 · 07·04
Scarlett. puts an AI co-worker inside Slack and iMessage
Scarlett. launched today on Product Hunt as an 'AI co-worker,' not just another bot. It lives inside Slack and iMessage, automates workflows, and claims to run your company on autopilot. Built by Ben Lang and team, powered by Anthropic Claude. Free tier offers 2x credits ($200 value). The post doesn't spell out which workflows it supports or latency. I'd stay cautious—Slack bots are everywhere; real co-worker value depends on integration depth.
#Agent#Scarlett.#Ben Lang#Anthropic Claude
editor take
Scarlett. launched today as an 'AI co-worker' inside Slack and iMessage, but the post doesn't spell out which workflows it supports.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
05:37
24d ago
Hacker News Frontpage· rssEN05:37 · 07·04
2026 Unslop AI-Written Fiction Contest Results Announced, $10K Grand Prize
Hyperstition's Unslop contest results are out. A. Best won the $10,000 grand prize for "The June." The process: ~120 applicants each got a 1-month Claude Code subscription or cash to write a prompt generating one short story. Judges picked 6 finalists from ~15 semi-finalists, who each submitted a second story. The post doesn't name the judges or detail the judging criteria.
#Hyperstition#A. Best#Aaron Silverbook
editor take
Unslop contest winner: A. Best's "The June" took the $10k prize, but judges and criteria aren't disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
04:00
24d ago
Financial Times · Technology· rssEN04:00 · 07·04
Who really designed that dress? Fashion reckons with AI
The FT piece covers fashion's growing unease over generative AI and copyright. Brands and designers worry that AI-generated patterns, silhouettes, and even full designs lack legal protection and blur the line between human and machine authorship. It cites cases like a label accused of copying an independent designer after using Midjourney, and how EU rules on training-data disclosure could affect fashion weeks. The article doesn't offer a unified fix—just notes that brands are experimenting with watermarks, blockchain provenance, and internal ethics guidelines on their own.
#Financial Times#Midjourney#European Union
editor take
FT on fashion's AI copyright anxiety: good case studies, no clear fix yet.
HKR breakdown
hook knowledge resonance
open source
50
SCORE
H0·K0·R0
01:30
24d ago
Hacker News Frontpage· rssEN01:30 · 07·04
CueBench launches to score how well developers drive coding agents
CueBench launched a tool for developers: upload your AI coding session logs, get scored on four AI fluency skills (0-100), and receive coaching. Your scores are private; session files are deleted after scoring. The post doesn't specify the four skills or which AI tools are supported.
#CueBench#Benchmark
editor take
Upload your AI coding session logs to CueBench, get scored on four fluency skills (0-100), and receive coaching — scores are private, logs deleted after scoring.
HKR breakdown
hook knowledge resonance
open source
45
SCORE
H1·K0·R0
00:00
25d ago
Computing Life · Share (鸭哥 research reports)· rssZH00:00 · 07·04
AI benchmarks are like EV 0–60 times: they win attention, not pricing power
Yage uses car culture to explain AI benchmarking: 0–60 times are the easiest metric for EVs to win, just like MMLU or SWE-bench scores for models. The Dodge Demon 170 hits ~1.7 s 0–60 but can't price like a Ferrari. MIT research tracking five months of OpenRouter data shows open models catch up to closed benchmarks within 13 weeks, yet closed models still capture 80% of usage and 96% of revenue. The Ferrari 12Cilindri Manuale costs 50% more than the automatic and has a lower top speed—it sells an irreplicable narrative. The piece argues benchmarks are a ticket to enter, not pricing power; what anchors price is context infrastructure, workflow depth, and user trust, none of which a competitor can simply copy.
#Benchmarking#MIT#OpenRouter#Ferrari
editor take
0–60 times explain AI benchmarking: scores are a ticket to enter, not pricing power. MIT data backs it—open models catch up in 13 weeks, closed models still take 96% of revenue.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
2026-07-03 · Fri
22:34
25d ago
Hacker News Frontpage· rssEN22:34 · 07·03
GitFut: Turn your GitHub stats into a World-Cup-style player card
GitFut turns your GitHub stats into a World-Cup-style player card rated out of 99. Enter a username to generate a card — torvalds gets a 96 with attributes like PAC, SHO, DEF. The project has 921 GitHub stars and has rated 150,806 cards. The post doesn't explain the scoring algorithm or attribute weights, only shows the front-end output.
#GitHub#Younes#Mawsis
editor take
GitHub stats turned into FIFA-style player cards — torvalds gets a 96, but the scoring formula is a black box.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
22:31
25d ago
Dwarkesh Patel· atomEN22:31 · 07·03
Mathematicians will become art curators – Grant Sanderson
Only the title is available; the post does not elaborate. Grant Sanderson suggests mathematicians will shift to curating mathematical art, implying discovery may be automated while humans select and interpret beauty. No further context is given.
#Grant Sanderson
editor take
Grant Sanderson: mathematicians become art curators as AI automates discovery. The post doesn't elaborate — interesting direction, thin on details.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
21:49
25d ago
Hacker News Frontpage· rssEN21:49 · 07·03
Wafer serves GLM5.2 on AMD MI355X at 2626 tok/s/node, over 2x cheaper than Blackwell
Wafer ran GLM-5.2 on AMD MI355X, hitting 2626 tok/s/node aggregate and 213 tok/s single-stream decode, at less than half the cost of B200. They quantized the model to MXFP4 with AMD Quark with negligible accuracy loss, used sglang as the inference engine, and fixed two MTP bugs to enable speculative decode—yielding a ~3x single-stream gain. Manual MoE kernel tuning lifted prefill-bound throughput from 1944 to 2626 tok/s. Overall performance is ~80% of B200, but per-dollar performance is clearly ahead. The post doesn't disclose pricing or regional availability.
#Inference-opt#Wafer#AMD#MI355X
editor take
Wafer hit 2626 tok/s/node on AMD MI355X with GLM-5.2 at <50% B200 cost, but the post doesn't disclose pricing or regional availability.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
21:28
25d ago
Hacker News Frontpage· rssEN21:28 · 07·03
Software, from First Principles: A Computer Science Primer for Non-Programmers
A 54-minute read that walks from mechanical calculators to operating systems and networking, aiming to demystify computers for non-CS readers. The author uses a gear simulator to explain carry mechanisms and die-shot photos to show memory vs. logic. No specific models, frameworks, or company products are discussed—pure conceptual primer.
editor take
A 54-minute primer from mechanical calculators to die-shot photos—no CS degree needed to see how computers actually work.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
21:14
25d ago
Product Hunt · AI· rssEN21:14 · 07·03
Termi Protocol: Watch your AI coding agents build, live in 3D
Termi Protocol is a 3D simulation layer for AI coding agents. It gives each agent a face, a desk, and a room, then visualizes their read/write/run actions in real time like a game. The post doesn't disclose which agent frameworks it supports, whether it's open-source, or the performance overhead. You still run the agents; Termi just visualizes the process.
#Termi Protocol
editor take
A 3D layer that gives coding agents a face, desk, and room, visualizing their actions like a game. The post doesn't disclose framework support, open-source status, or performance overhead—keep expe...
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
21:03
25d ago
Hacker News Frontpage· rssEN21:03 · 07·03
ContextCodeCache in Rust: cut token costs by caching LLM context
ContextCodeCache is an open-source Rust tool that caches LLM context windows to avoid recomputing repeated inputs. It reuses previously computed context across calls, which helps cut token costs and latency in high-frequency scenarios like code completion or chat history. The post doesn't disclose exact savings or benchmarks.
#ContextCodeCache#Rust
editor take
Open-source Rust tool caches LLM context to avoid recompute on repeated inputs, saving tokens. No benchmarks disclosed yet, so temper expectations.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0

more

feeds

admin