ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
33 srcsignal 72%cycle 04:32

posts · 2026-07-30

20 items · updated 3m ago
RSS live
2026-07-30 · Thu
10:00
54d ago
Financial Times · Technology· rssEN10:00 · 07·30
Amazon finds cases of AI causing runaway spending on tech projects
The FT article is behind a hard paywall; only the headline and site navigation are visible. The headline states Amazon has identified cases where AI is causing runaway spending on tech projects. The post does not disclose which projects, how much overspend, or whether the cost driver is inference, engineering, or something else. Treat this as a signal that big tech is starting to audit AI project ROI seriously, not just another vague 'AI is expensive' complaint.
#Amazon#Financial Times
editor take
FT paywall hides details, but the headline signals big tech is now auditing AI project ROI for real.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
09:34
54d ago
Hacker News Frontpage· rssEN09:34 · 07·30
Agent-Manager: A Tmux TUI for Running Claude Code, Codex and OpenCode
Agent-Manager is a terminal UI that runs multiple AI coding agents (Claude Code, Codex, OpenCode, Grok Build) inside tmux. It shows live status, group tree, pane preview, and resource gauges for each session. Useful if you juggle several agent sessions in the terminal. The post doesn't disclose installation steps or performance benchmarks.
#Claude Code#Codex#OpenCode
editor take
A tmux TUI that runs Claude Code, Codex, and other coding agents in one window—handy if you juggle multiple sessions.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
07:43
54d ago
AI HOT (Curated Pool)· aihot-apiZH07:43 · 07·30
Token Saver: An open-source MCP extension that cuts Claude PDF token costs by 92–99% with local hybrid RAG
Marktechpost released Token Saver, an open-source MCP extension that stops Claude from re-ingesting entire PDFs on every chat turn. It uses local hybrid RAG to retrieve only relevant chunks, cutting token usage by 92–99% in their tests. The project is MIT-licensed, currently at v1.0, and available on GitHub. The post doesn't disclose the PDF sizes or exact Claude model version used for benchmarking, so I'd take the 99% figure with a grain of salt until more data surfaces.
#RAG#Marktechpost#Claude#Anthropic
editor take
Local hybrid RAG for Claude PDFs, feeding only relevant chunks and claiming 92–99% token savings—but the post doesn't disclose PDF sizes or model version, so I'd discount the 99% figure.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
06:46
54d ago
r/LocalLLaMA· rssEN06:46 · 07·30
4090 + 5060 Ti + 64GB RAM: 206 t/s on a 35B-A3B, 122B at 37 t/s
A Reddit user posted local inference benchmarks: RTX 4090 + RTX 5060 Ti + 64GB system RAM hitting 206 tokens/sec on a 35B-A3B model and 37 t/s on a 122B model. The post body is blocked by Reddit, so it doesn't disclose the exact model name, quantization, or inference framework. These speeds are impressive for local LLM deployment—especially 37 t/s on a 122B—showing heterogeneous multi-GPU setups can deliver.
#NVIDIA RTX 4090#NVIDIA RTX 5060 Ti#Benchmark
editor take
4090 + 5060 Ti hits 37 t/s on a 122B model, but Reddit blocked the post body—no model name or quantization disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
04:38
54d ago
AI Chat-Group Daily (群聊日报)· atomZH04:38 · 07·30
After K3 goes open-weight, a law firm's $30K/mo API bill makes self-hosting math work
The K3 open-weight release is forcing real cost calculations. A US law firm spending nearly $30K/month on Claude API is testing K3 internally; if it passes, they'll buy an 8-card AMD MI355X server for about $400K, bringing monthly costs down to ~$18K—20–30% cheaper with data staying on-prem. The breakeven point cited: $20K/month in API spend. Unsloth released a 1-bit quantized K3 at 594GB retaining ~78.9% top-1 accuracy, but community members warn that Mac mini hardware is too slow to run it practically. OpenAI launched ChatGPT for Academic Researchers, giving 100K scientists free access to GPT-5.6 models; the first 10K start this summer, and people are already exploiting SheerID verification for free 12-month Pro accounts. On the tools side, a community member got an M5Stack board running as a voice input device overnight, Grok was found to crawl X directly for stock analysis with cheap fast STT, and Cursor gave a legacy user $20 in free credits to finish interrupted work.
#Code#Kimi K3#Anthropic Claude#AMD MI355X
editor take
K3 open-weight release triggers real cost math: self-hosting beats API at $20K/month spend.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
04:00
54d ago
Financial Times · Technology· rssEN04:00 · 07·30
The tech wreck roiling Wall Street
Tech stocks are plunging on Wall Street, triggered by SK Hynix's disappointing profits. Meta shares tumbled as Zuckerberg's AI agent pitch failed to calm investors. Microsoft signed $130bn in data center leases, but the market is focused on near-term earnings pressure.
#SK Hynix#Meta#Microsoft#Funding
editor take
SK Hynix's profit miss triggered a tech sell-off; Meta and Microsoft couldn't calm the market.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
04:00
54d ago
Financial Times · Technology· rssEN04:00 · 07·30
AI investment concentration risk is not just in equities
FT warns AI practitioners that investment concentration risk goes beyond equities—it also lies in infrastructure, talent, and supply chains. The post doesn't name specific firms or figures, but the headline makes the core point: don't just watch stock volatility; the entire AI dependency chain is a risk source.
#Financial Times
editor take
FT warns AI concentration risk isn't just stocks—compute, talent, and chip supply chains are all bottlenecked.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
04:00
54d ago
Financial Times · Technology· rssEN04:00 · 07·30
Why the UK has worse mobile coverage than Romania
The UK lags behind Romania in mobile coverage. FT blames slow planning approvals, low carrier investment, and inefficient spectrum allocation. No hard data in the post, but it notes the UK trails several Eastern European countries on 5G. For AI practitioners, poor connectivity limits edge inference, connected vehicles, and real-time agent deployment—infrastructure can't keep up with model capability.
#Financial Times
editor take
UK 5G coverage lags Romania—slow planning approvals and low carrier investment are the culprits.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
04:00
54d ago
Financial Times · Technology· rssEN04:00 · 07·30
CuspAI's Max Welling: AI-designed molecules to capture forever chemicals from water
CuspAI founder Max Welling tells FT his startup uses generative AI to design molecules that capture PFAS (forever chemicals) from water. The pipeline generates candidate structures, simulates them, then validates in the lab. Welling, a student of Geoffrey Hinton and a Dutch ML professor, didn't disclose specific molecule performance, cost, or timeline.
#CuspAI#Max Welling#Financial Times
editor take
CuspAI uses generative AI to design molecules that capture PFAS from water, but the article doesn't disclose performance or cost.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
04:00
54d ago
Financial Times · Technology· rssEN04:00 · 07·30
ByteDance’s plan to dominate AI, powered by TikTok and its recommendation engine
The FT maps out ByteDance’s AI strategy: it embeds AI into TikTok, CapCut, and Douyin, leveraging its recommendation-engine expertise for inference optimization. The piece covers in-house chip efforts and data center builds, plus overseas expansion. It reads more like a business strategy analysis—model specs and chip timelines are thin.
#Inference-opt#ByteDance#TikTok#Douyin
editor take
FT maps ByteDance's AI play: embed models into TikTok/CapCut, optimize inference like recommendations, build own chips.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
03:37
55d ago
New York Times Chinese· rssZH03:37 · 07·30
The Hidden Cost of China's Free AI: Built-in Censorship and Geopolitical Bias
This NYT op-ed argues that China's free open-source models (DeepSeek, Qwen) are widely adopted but trained under Chinese law to avoid criticism and echo official narratives. NIST tests found DeepSeek repeating Beijing's line on Xinjiang; Swedish researchers saw Qwen's internal chain-of-thought checklist favoring positive framing. CrowdStrike found DeepSeek generated code with security flaws for certain groups. The author proposes mandatory 'Made for China' labeling and US subsidies to compete on price, though no specific funding or timeline is disclosed.
#DeepSeek#Alibaba (Tongyi Qianwen / Qwen)#Moonshot AI (Kimi)
editor take
This NYT op-ed ties together NIST tests, Qwen's internal checklist, and CrowdStrike's code-vuln findings on Chinese models—useful as a single thread, but it's an opinion piece, not a technical audit.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R1
03:30
55d ago
Financial Times · Technology· rssEN03:30 · 07·30
In-house legal teams get creative with AI tools
FT reports corporate legal teams are using AI for contract review and due diligence. One firm cut review time from 3 hours to 15 minutes. Lawyers still need to check AI output. The post doesn't name specific models or tools, just stresses creative use.
#Financial Times
editor take
FT: corporate legal teams cut contract review from 3 hours to 15 min with AI, but lawyers still check output.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
02:49
55d ago
AI HOT (Curated Pool)· aihot-apiZH02:49 · 07·30
OpenAI President admits new ChatGPT desktop app is 'a bit messy', aims for 'zero tabs' by year-end
OpenAI President Greg Brockman admitted in an interview that the new ChatGPT desktop app is 'a bit messy' after merging with Codex, confusing users. He called it a transitional phase, with a goal to eliminate Work tabs by year-end. Codex users doubled from 5M to 10M in days.
#OpenAI#Greg Brockman#Codex#Product update
editor take
OpenAI's president admits the new ChatGPT desktop app is messy; plans to kill Work tabs by year-end.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
00:51
55d ago
Hacker News Frontpage· rssEN00:51 · 07·30
Logic for Programmers: A practical book on math, software, and fixing one with the other
Hillel Wayne's new book teaches working programmers how to use logic (Boolean math) to design, verify, and improve software. 227 pages, no math background required, but expects intermediate-to-advanced programming skills. Covers simplifying conditionals, safe API changes, property testing, formal verification with Dafny, TLA+ for distributed systems, and more across 11 chapters. Author previously wrote Practical TLA+ and trained NASA, Meta, and others. Available as DRM-free ebook (PDF/EPUB) on Leanpub and print on Amazon; free sample chapter on the site. Price not disclosed in the post.
#Hillel Wayne#NASA#Meta
editor take
Hillel Wayne's new book teaches working programmers to use logic (Boolean math) to design, verify, and fix code—no math background needed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
00:21
55d ago
● P1TechCrunch AI· rssEN00:21 · 07·30
Microsoft pitches own AI models and tools to compete directly with OpenAI and Anthropic
Microsoft pitched its own AI models, toolchains, and a Mythos competitor to Wall Street during its earnings call. CEO Nadella made it clear he won't let OpenAI and Anthropic own customer relationships through apps and agent infrastructure. The company just posted $331.8B in annual revenue and $133.7B in net income, giving it plenty of leverage to compete directly.
#Agent#Microsoft#OpenAI#Anthropic
why featured
Featured · importance 92 · hook + knowledge + resonance
editor take
Microsoft's earnings reveal a $3.2B gain on Anthropic and a $600M write-down on OpenAI. The bigger signal: on the earnings call, Microsoft pitched its own models directly against both labs.
sharp
Microsoft's Q4 earnings (ending June 30) dropped two numbers that tell a clear story: a $3.2 billion gain on its Anthropic investment, and a roughly $600 million write-down on OpenAI. Both TechCrunch pieces draw from the same earnings release and call transcript, so the facts are solid. The angle split is useful — one article focuses on the accounting, the other on the competitive shift. I'd discount the $3.2B figure a bit. It's an unrealized gain tied to Anthropic's valuation, not cash Microsoft pocketed. The real signal is what happened on the earnings call: Satya Nadella pitched Microsoft's own AI models and toolchain directly to customers, positioning them against OpenAI and Anthropic. That's a sharp turn from the "we're the best partners" tone of a year ago. What's missing: actual performance benchmarks and pricing for Microsoft's in-house models. The call mentioned the direction but didn't share specs.
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
00:06
55d ago
Bloomberg Technology· rssEN00:06 · 07·30
Samsung chip profit soars over 250-fold on AI memory shortages
Samsung's Q2 chip profit hit 9.2 trillion won (~$6.9 billion), up from just 36 billion won a year ago. HBM supply is fully booked by AI server demand, and even standard DRAM and NAND prices are rising. Samsung expects HBM shipments to double again in H2, though the article doesn't disclose capacity ceilings or customer allocation details.
#Samsung
editor take
Samsung's chip profit jumped 250x YoY on AI memory demand; HBM is fully booked and shipments are set to double in H2.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
00:00
55d ago
OpenAI Blog· rssEN00:00 · 07·30
Japan retailer Yamada Denki built a 24/7 voice shopping agent on GPT-Realtime, 30k users in 2 weeks
Japan's Yamada Denki partnered with AI customer service startup avatarin to build a 24/7 multilingual voice shopping agent on OpenAI's GPT-Realtime. In a two-week public trial, ~30,000 shoppers used it; 92% of survey responses were positive. Unlike traditional chatbots, the agent understands context and asks follow-up questions—e.g., 'My kitchen is small for a family of four, which fridge should I pick?' It uses RAG for accurate product info and embeds Yamada's sales expertise into conversation flows. OpenAI helped optimize prompt structure and API costs for always-on voice. The post doesn't disclose exact latency or cost figures but says low latency was a key reason for choosing GPT-Realtime.
#Audio#OpenAI#avatarin#Yamada Denki
editor take
Yamada Denki's GPT-Realtime voice shopping agent: 30K users in 2 weeks, 92% positive feedback.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0

more

feeds

admin