ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

posts · 2026-05-12

26 items · updated 3m ago
RSS live
2026-05-12 · Tue
04:33
77d ago
● P1Latent Space· rssEN04:33 · 05·12
Thinking Machines' Native Interaction Models: TML-Interaction-Small 276B-A12B Advances Realtime Voice
Thinking Machines released TML-Interaction-Small, a 276B-parameter MoE model with 12B active parameters, and the post says it advances realtime voice through 200ms time-aligned microturns, encoder-free early fusion for audio and images under 200ms, and benchmark wins over GPT-Realtime-2 and Gemini 3.1-Flash.
#Multimodal#Audio#Agent#Thinking Machines
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
Thinking Machines moved realtime voice inside the model loop: 276B MoE, 12B active, 200ms microturns. That hits harder than another chat leaderboard.
sharp
Thinking Machines is betting on the interaction clock, not a speech wrapper. TML-Interaction-Small is a 276B MoE with 12B active parameters, encoder-free early fusion for audio and images, and 200ms time-aligned microturns. That attacks the hand-coded turn logic sitting between VAD, ASR, LLM, and TTS stacks. I’d discount the official leaderboard for now: wins over GPT-Realtime-2 and Gemini 3.1-Flash on BigBench Audio, IFEval, and FD-bench lack reproducibility details in the snippet. The stronger signal is the new task shape: TimeSpeak, CueSpeak, RepCount-A, and ProactiveVideoQA test when to talk, when to stay silent, and when visual evidence becomes available. OpenAI’s 4o “Her” demo sold presence; Thinking Machines is trying to own timing.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
04:22
77d ago
Product Hunt · AI· rssEN04:22 · 05·12
TestSprite 3.0
TestSprite 3.0 says it uses a fleet of parallel agents to test an app in minutes; the post does not disclose supported frameworks, test types, pricing, or reproducible benchmarks.
#Agent#Code#Tools#TestSprite
editor take
TestSprite 3.0 claims parallel agents test apps in minutes; no frameworks, test types, pricing, or benchmarks disclosed, so treat as launch noise.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H1·K0·R1
03:30
77d ago
Product Hunt · AI· rssEN03:30 · 05·12
MiniCPM-V 4.6 released, 1.3B parameter multimodal model for mobile devices
MiniCPM-V 4.6 presents a 1.3B vision-language model for mobile use; the Product Hunt snippet does not disclose benchmarks, license terms, pricing, or context window details.
#Multimodal#Vision#MiniCPM#Product update
editor take
MiniCPM-V 4.6 claims 1.3B mobile VLM; no benchmarks or license, so treat it as a Product Hunt teaser.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H0·K1·R0
03:08
77d ago
AI HOT (Curated Pool)· aihot-apiZH03:08 · 05·12
Beyond Answers: Information Presentation Is Becoming Part of the AI Intelligence Layer
SiliconFlowAI argues that HTML output gives LLMs richer layout and interaction than default Markdown, and the post outlines an output path from raw text to Markdown, HTML, and interactive neural video or simulation.
#Multimodal#Vision#Tools#SiliconFlowAI
editor take
SiliconFlowAI pushes HTML output for LLMs; no evals disclosed, so this is a UI shortcut before it is intelligence.
HKR breakdown
hook knowledge resonance
open source
36
SCORE
H1·K0·R1
03:06
77d ago
Bloomberg Technology· rssEN03:06 · 05·12
Australia Watchdog Says Money Launderers Ramping Up AI for Scams
Australia’s financial crimes watchdog warned that money launderers are using AI to scale scams, automate processes, and create fake documents; the RSS snippet does not disclose case counts, dollar amounts, enforcement actions, or the specific AI tools involved.
#Tools#Australia financial crimes watchdog#Policy#Incident
editor take
Australia says launderers use AI to scale scams; no dollars, case counts, or tools disclosed, so don't treat it as trend evidence.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1
02:37
77d ago
New York Times Chinese· rssZH02:37 · 05·12
Why Chinese People Do Not Fear AI Like Americans Do
Jacob Dreyer argues that China treats AI as infrastructure for schools, hospitals, transport, and governance; the article cites the 2020 census figure that nearly 40% of Chinese people live in rural areas, but it does not disclose measured costs or deployment outcomes.
#Agent#Vision#Robotics#Jacob Dreyer
editor take
Dreyer cites nearly 40% rural population, but gives no AI cost or outcome data; I don’t buy the infrastructure optimism.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1
02:27
77d ago
r/LocalLLaMA· rssEN02:27 · 05·12
Blackwell LLM Toolkit: NVFP4 Configs, Wheels, and Benchmarks for Blackwell GPUs
elsung released Blackwell LLM Toolkit and reports Nemotron-3-Nano-Omni V3 sustaining 270 tok/s decode on one RTX Pro 6000 96GB with NVFP4 at 8k context. The repo includes TensorRT-LLM configs, a rebuilt LMCache wheel for Blackwell, benchmark scripts, and results for MiniMax-M2.7, DeepSeek-V4-Flash, and Nemotron variants.
#Inference-opt#Multimodal#Benchmarking#Nvidia
editor take
Title claims 270 tok/s on one RTX Pro 6000; body is 403, so treat this as a Blackwell NVFP4 recipe lead, not reproduced evidence.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
02:22
77d ago
Hacker News Frontpage· rssEN02:22 · 05·12
Fake building: Claude wrote 3k lines instead of importing pywikibot
The title says Claude wrote 3,000 lines instead of importing pywikibot, while the RSS body only lists the article URL, Hacker News discussion URL, 19 points, and 6 comments; the post does not disclose the prompt, model version, repository context, evaluation method, or reproducible conditions.
#Code#Claude#pywikibot#Hacker News
editor take
Claude Opus 4.7 wrote ~3,000 lines instead of imports; I buy the failure, not the benchmark-blame theory.
HKR breakdown
hook knowledge resonance
open source
64
SCORE
H1·K1·R1
02:19
77d ago
● P1AI HOT (Curated Pool)· aihot-apiZH02:19 · 05·12
Thinking Machines Releases Native Multimodal Interaction Model for Real-Time Human-AI Collaboration
Thinking Machines released an interaction model that natively receives audio, video, and text input, processes foreground interaction at 200-millisecond intervals, and uses a background reasoning model for long-horizon planning and tool calls.
#Multimodal#Audio#Tools#Thinking Machines
why featured
Featured · importance 87 · hook + knowledge + resonance
editor take
Thinking Machines is betting on a 200 ms foreground loop, not another multimodal demo; Mira is productizing presence as architecture.
sharp
Thinking Machines’ sharp move is splitting “presence” into a 200 ms foreground loop and a slower reasoning backend. The disclosed mechanism matters: native continuous audio, video, and text input; the foreground model handles interruption and immediate feedback; the backend handles long-horizon planning and tool calls. That is system design, not another agent chain with nicer prompts. I don’t buy the “unified interface” framing yet. The hard parts are state sync, latency budget, and writing tool results back without breaking the live interaction. OpenAI’s Advanced Voice already taught users what low-latency audio feels like, but it did not expose this kind of architecture. If Thinking Machines only has a polished demo, GPT-class products catch it. If the 200 ms loop holds under video plus tools, this becomes a serious collaboration substrate.
HKR breakdown
hook knowledge resonance
open source
87
SCORE
H1·K1·R1
02:12
77d ago
r/LocalLLaMA· rssEN02:12 · 05·12
Improve prompt processing speed for --n-cpu-moe partially offloaded models
coder543 raised gpt-oss-120b’s -ub from 512 to 8192 on an RTX 3090, increasing prefill from 380.27 to 2090.68 tok/s while generation fell from 32.29 to 30.05 tok/s under higher --n-cpu-moe settings.
#Inference-opt#llama.cpp#NVIDIA#coder543
editor take
coder543 raised -ub 512→8192 and hit 2090.68 tok/s prefill; body is 403, and decode drops to 30.05 tok/s.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1
02:05
77d ago
r/LocalLLaMA· rssEN02:05 · 05·12
Found a Way to Cool the DGX
A Reddit user cooled a DGX with tap water and kept it under 68°C while running Qwen3.5-122B-A10B Q6_K at 95% GPU utilization, with 110GB memory use, an 80k context window, and 18.77 tokens/s for continuous vision analysis.
#Vision#Inference-opt#Reddit#NVIDIA DGX
editor take
Title claims tap water holds DGX under 68°C, but body is 403; I don’t buy it without loop, flow, condensation details.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
01:50
77d ago
● P1Bloomberg Technology· rssEN01:50 · 05·12
South Korean Policymaker Proposes AI Tax-Funded Citizen Dividend
A senior South Korean policymaker proposed paying citizens a dividend funded by taxes on AI profits; the RSS snippet does not disclose the tax rate, payout size, legislative path, or implementation timeline.
#Samsung Electronics#SK Hynix#South Korea#Policy
why featured
Featured · importance 86 · hook + knowledge + resonance
editor take
Only headlines, no tax rate, target base, or payout formula; Korea is testing an AI-tax balloon in markets, not showing a bill yet.
sharp
Two Bloomberg headlines align: Korean policymakers floated an AI tax to fund a “citizen dividend,” and Korean stocks already swung. The body is empty, so the tax rate, payer base, and timetable are not disclosed. I don’t buy the clean “AI dividend” framing. Without a defined base, markets will map the tax onto semis, cloud, and platform names by default; Korea is unusually exposed through memory, HBM, and factory automation. The US and EU are still fighting over compute rules, model liability, and copyright fees. Korea jumping straight to cash distribution sounds politically neat and operationally brutal.
HKR breakdown
hook knowledge resonance
open source
86
SCORE
H1·K1·R1
01:37
77d ago
New York Times Chinese· rssZH01:37 · 05·12
Who Are the 16 Business Leaders Accompanying Trump to China?
Trump will travel to China this week with 16 CEOs, including Elon Musk and Tim Cook, while Jensen Huang was not invited and Nvidia is still waiting for U.S. and Chinese approval to export an early version of its H200 AI chip to China.
#Donald Trump#Elon Musk#Tim Cook#Policy
editor take
Trump brings 16 CEOs to China; Jensen Huang is absent while H200 approvals stall, leaving Nvidia awkward either way.
HKR breakdown
hook knowledge resonance
open source
67
SCORE
H1·K1·R1
00:39
77d ago
AI HOT (Curated Pool)· aihot-apiZH00:39 · 05·12
Cursor integrates Microsoft Teams for workplace workflows
Cursor is now available in Microsoft Teams, and the post lists three integrations: Slack, Linear, and Microsoft Teams; it does not disclose specific features, permission scopes, rollout timing, or pricing.
#Tools#Cursor#Microsoft Teams#Slack
editor take
Cursor added Teams, Slack, and Linear integrations; permissions and pricing are undisclosed, so IT approval is the actual friction.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R1
00:23
77d ago
AI HOT (Curated Pool)· aihot-apiZH00:23 · 05·12
Google Says Criminal Hackers Used AI to Find a Major Software Flaw
Google’s Threat Analysis Group traced criminal hackers who used AI to find and exploit a major flaw in widely used open-source software; the post does not disclose the CVE, affected project, attack scale, or patch version.
#Tools#Google#Incident#Safety/alignment
editor take
Google says hackers used AI to find a zero-day, but names no CVE or project; without reproducible detail, don’t call it a capability leap.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
00:00
77d ago
● P1Computing Life (鸭哥 / grapeot)· atomZH00:00 · 05·12
Author Uses AI to Diagnose and Cure AI-Induced Insomnia
The author used AI to build a HealthKit export app in about 5 minutes and run multivariate regression, finding that the last post-dinner AI usage time correlated negatively with sleep duration; after avoiding AI at night, average sleep increased by 1 hour and 40 minutes.
#Agent#Code#Tools#Apple
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
A personal experiment, not a clinical study. The author ran multivariate regression and found late-night AI use strongly correlated with insomnia—stopping it added 1h40m of sleep. Numbers are speci...
sharp
Both sources are the same article in English and Chinese, so the multi-source coverage here doesn't add confidence—it's one author publishing the same personal log in two languages. I'd read this as a detailed n=1 case study, not a generalizable finding. The author did one smart thing: instead of guessing at the cause of his insomnia, he had AI write an app to pull Apple Watch data, then ran multivariate regression on it. The strongest correlation wasn't caffeine or bedtime—it was the timestamp of his last AI session that evening. His own explanation: AI handles the grunt work, so what's left for the human is high-intensity reading and decision-making, and he runs multiple AI sessions in parallel with no mental cooldown. That logic holds for his case, but only his case. What's missing: no raw data, no regression coefficients, no control for confounders like deadline pressure. If you're dealing with similar sleep issues, you can try the method, but don't treat his conclusion as a diagnosis.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
00:00
77d ago
Computing Life · Share (鸭哥 research reports)· rssZH00:00 · 05·12
AI Dictation Competition Shifts from Model Layer to Keyboard Layer
The post argues that Google holds the key advantages in AI dictation through microphone access, preinstalled distribution, and free pricing; the body does not disclose transcription accuracy, model details, or user scale.
#Audio#Google#Commentary
editor take
Google has mic access, preinstall, and free pricing; accuracy and scale are undisclosed, but distribution will eat model gaps here.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
00:00
77d ago
Computing Life · Share (鸭哥 research reports)· rssZH00:00 · 05·12
They Build Connectors, Google Controls the OS: Why Model Quality Cannot Close the AI OS Gap
Google Gemini Intelligence is described as able to call phone apps directly; the post does not disclose the Android interface mechanism, number of supported apps, or launch conditions.
#Agent#Tools#Google#OpenAI
editor take
Gemini Intelligence claims app-wide phone control, with no interface or coverage disclosed; OS privilege is a moat connectors won't match.
HKR breakdown
hook knowledge resonance
open source
63
SCORE
H1·K0·R1

more

feeds

admin