ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

posts · 2026-05-16

50 items · updated 3m ago
RSS live
2026-05-16 · Sat
23:39
72d ago
r/LocalLLaMA· rssEN23:39 · 05·16
Anyone else running pre-release MTP branches to maintain higher speeds?
A Reddit user says a pre-release MTP branch runs about 20% faster on Dual Xeon 8268 CPUs with a Tesla T4, reaching about 38 output tokens per second; the release branch reaches about 30 tokens per second and crashed llama.cpp during light coding.
#Inference-opt#Vision#Code#Reddit
editor take
MTP pre-release hits 38 t/s on a T4; I trust the throughput claim before I trust the stability story.
HKR breakdown
hook knowledge resonance
open source
56
SCORE
H1·K1·R1
23:04
72d ago
AI HOT (Curated Pool)· aihot-apiZH23:04 · 05·16
Figure humanoid robot runs autonomously for four consecutive days, moving toward practical use
Figure’s F.03 humanoid robot entered its fourth day of 24/7 autonomous testing in a real warehouse, performing grasping, carrying, and sorting tasks; the post does not disclose failure counts or maintenance intervals.
#Robotics#Agent#Figure#Benchmark
editor take
Figure F.03 ran warehouse tasks for four days; without failures or maintenance intervals, don't call it practical yet.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
22:23
72d ago
Hacker News Frontpage· rssEN22:23 · 05·16
Zerostack – A Unix-inspired coding agent written in pure Rust
Zerostack published a 1.0.0 package on crates.io, and the title describes it as a Unix-inspired coding agent written in pure Rust; the post does not disclose its architecture, tool interface, or benchmark results.
#Agent#Code#Tools#Zerostack
editor take
Zerostack shipped crates.io 1.0.0; only the title is disclosed, with no architecture, tool API, or benchmarks.
HKR breakdown
hook knowledge resonance
open source
64
SCORE
H1·K0·R1
22:19
72d ago
r/LocalLLaMA· rssEN22:19 · 05·16
Now that MTP is merged, what are the best Qwen 3.6 35B outputs on 2×3090s?
A Reddit user asks for Qwen 3.6 35B results on dual RTX 3090s after llama.cpp merged MTP; their split-layer setup previously reached 1500 p/p and 120 t/g, MTP testing fell to 80 t/g, and their CPU overflow fallback reports 3500 p/p and 80 t/g.
#Inference-opt#Qwen#llama.cpp#NVIDIA
editor take
Qwen 3.6 35B on 2x3090 drops to 80 t/s with MTP. Honestly, one Reddit rig is not a win signal.
HKR breakdown
hook knowledge resonance
open source
52
SCORE
H1·K1·R1
21:54
72d ago
r/LocalLLaMA· rssEN21:54 · 05·16
Qwen3.5-122B-Q5-MTP and Qwen3.5-122B-Q6-MTP
A Reddit user tested two Qwen3.5-122B MTP quantized models under llama.cpp server-rocm-mtp with --spec-type draft-mtp and --spec-draft-n-max 3; Qwen3.5-122B-Q5-MTP-General reached 20.24 t/s over 4,200 eval tokens, while Qwen3.5-122B-Q6-MTP-General reached 17.17 t/s over 3,283 eval tokens.
#Inference-opt#Benchmarking#Qwen#Unsloth
editor take
Qwen3.5-122B MTP shows 20.24 t/s, but the body is 403; treat this as one Reddit rig's number.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H0·K1·R1
21:34
72d ago
r/LocalLLaMA· rssEN21:34 · 05·16
I fitted the new δ-mem research for Apple Silicon using MLX and OpenClaw integration
A Reddit user adapted δ-mem to MLX on a 64GB Apple Silicon Mac mini and tested Qwen3-4B-Instruct with OpenClaw history. LoCoMo-10 mini rose from 0.0500 to 0.1833, while OpenClaw replay improved from 6/8 to 7/8 passed probes with about 1.30x latency.
#Memory#Agent#Benchmarking#Apple
editor take
Summary says δ-mem lifts LoCoMo-10 from 0.0500 to 0.1833; body is 403, so distrust the 1.30x tradeoff.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
19:58
73d ago
AI HOT (Curated Pool)· aihot-apiZH19:58 · 05·16
US Starts Seeing Heavy Job Losses in Roles Exposed to AI
Bloomberg says US roles exposed to AI are starting to see heavy job losses; the post does not disclose layoff counts, affected industries, or the measurement method.
#Bloomberg#Commentary
editor take
Bloomberg flags heavy AI-exposed US job losses, but gives no counts or method here; don’t weaponize this headline in planning.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K0·R1
19:51
73d ago
r/LocalLLaMA· rssEN19:51 · 05·16
Local Qwen 3.6 vs Frontier Models on a Single-File HTML Canvas Driving Animation
A Reddit user tested 11 models with the same single-file HTML Canvas driving-animation prompt, and local Qwen3.6-27B Q4_K_M ranked second subjectively at 2.70 tok/s, behind Kimi k2.6 Thinking and ahead of the Claude-opus-reasoning-distilled 27B quant.
#Code#Benchmarking#Qwen#Claude
editor take
Title says Qwen3.6-27B Q4_K_M ranked 2nd among 11 models; body is 403, so scoring and GIFs are unverified.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
19:43
73d ago
AI HOT (Curated Pool)· aihot-apiZH19:43 · 05·16
Codex Adds Custom Keyboard Shortcuts
Codex added custom keyboard shortcuts, letting users adjust key bindings in settings; the post does not disclose a version number, supported platforms, or rollout schedule.
#Code#Tools#Product update
editor take
Codex now supports custom shortcuts in settings. No version, platforms, or rollout disclosed; this is editor-table-stakes catch-up.
HKR breakdown
hook knowledge resonance
open source
58
SCORE
H0·K1·R0
18:58
73d ago
r/LocalLLaMA· rssEN18:58 · 05·16
How I Started Programming Differently Over the Last Year. What About You?
Reddit user /u/ievkz says they stopped using LLM autocomplete in the IDE, now use a CLI coding agent with @-referenced files, and keep the IDE mainly for Git diffs, debugging, and navigation that they estimate covers 5-10% of their work.
#Agent#Code#Tools#JetBrains
editor take
The poster says IDE navigation/debugging is 5-10% of work. CLI agents replacing autocomplete tracks my experience.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1
18:31
73d ago
AI HOT (Curated Pool)· aihot-apiZH18:31 · 05·16
Customize Keyboard Shortcuts to Fit Your Workflow
OpenAI Devs says Codex now supports custom keyboard shortcuts through settings. Users can map shortcuts around their workflow, but the post does not disclose platform coverage, rollout timing, or version requirements.
#Code#Tools#OpenAI#Product update
editor take
Codex now supports custom shortcuts; platform and version are undisclosed. Small fix, but default keymaps finally stop dictating flow.
HKR breakdown
hook knowledge resonance
open source
63
SCORE
H0·K1·R1
18:12
73d ago
r/LocalLLaMA· rssEN18:12 · 05·16
OpenReader: Open-source read-along document reader with TTS and audiobook export
OpenReader v3.0.0 ships an open-source TTS document reader for EPUB, PDF, DOCX, TXT, and Markdown, with OpenAI, Replicate, Deepinfra, or self-hosted OpenAI-compatible APIs, plus m4b/mp3 audiobook export with chapter metadata through ffmpeg.
#Audio#Tools#OpenReader#OpenAI
editor take
OpenReader v3.0.0 covers 5 formats to m4b/mp3; the body is 403-blocked, so I’d treat it as handy tooling.
HKR breakdown
hook knowledge resonance
open source
65
SCORE
H1·K1·R0
17:43
73d ago
Product Hunt · AI· rssEN17:43 · 05·16
CtrlOps
CtrlOps says it uses AI to deploy, debug, and manage Linux servers; the post does not disclose pricing, permission controls, supported distributions, or operational safeguards.
#Agent#Code#Tools#CtrlOps
editor take
CtrlOps claims AI-managed Linux servers, but discloses no permission model; before prod, ask where the audit log lives.
HKR breakdown
hook knowledge resonance
open source
48
SCORE
H1·K0·R1
17:19
73d ago
r/LocalLLaMA· rssEN17:19 · 05·16
Corsair desktop PC with Ryzen AI Max 395 and 128GB unified RAM: has anyone tested it for LLM?
A Reddit user posted a Corsair AI Workstation 300 listing with Ryzen AI Max 395, 128GB LPDDR5X memory, up to 96GB VRAM, and a 1TB SSD; the post does not disclose LLM throughput, tested model sizes, or the actual price.
#Inference-opt#Corsair#AMD#Reddit
editor take
Title says Ryzen AI Max 395 and 128GB; Reddit 403 hides tokens/s and price, so skip the value hype.
HKR breakdown
hook knowledge resonance
open source
46
SCORE
H1·K1·R1
17:02
73d ago
r/LocalLLaMA· rssEN17:02 · 05·16
LLM Phone Home: Reliable Apps That Can Deliver Inference from a Local Backend
A Reddit user asks for an iOS app that can serve an OpenAI-compatible endpoint from a local backend and has tested Apollo, Locally AI, Noema, and 3 Sparks. The post says 3 Sparks works for endpoint use but lacks MCP and web search, while Noema fails to complete DeepSeek V4 Flash requests from a Mac Studio.
#Agent#Tools#Inference-opt#3 Sparks
editor take
Body is only a 403; four iOS clients are named, and local OpenAI endpoints still smell like tinkering, not dependable UX.
HKR breakdown
hook knowledge resonance
open source
46
SCORE
H0·K1·R1
16:41
73d ago
r/LocalLLaMA· rssEN16:41 · 05·16
Strix Halo Llama.cpp MTP Benchmarks: 27B Gets Much Faster, 35B Is Mixed
Qwen3.6-27B-MTP reduced llama.cpp wall time from 258.65s to 200.55s in a 5-turn test reaching about 28.5k context, while Qwen3.6-35B-MTP increased wall time from 58.86s to 60.24s under the same setup.
#Inference-opt#Benchmarking#Qwen#Unsloth
editor take
Qwen3.6-27B-MTP hit 200.55s; body is 403, and 35B slowing to 60.24s kills blind MTP toggles.
HKR breakdown
hook knowledge resonance
open source
67
SCORE
H1·K1·R1
16:38
73d ago
AI HOT (Curated Pool)· aihot-apiZH16:38 · 05·16
vLLM Adds Support for Trillion-Parameter Models
The title says vLLM supports trillion-parameter models, while the body only mentions Day 0 community collaboration and does not disclose the model name, exact parameter count, implementation details, or reproducible conditions.
#Inference-opt#vLLM#Product update#Open source
editor take
vLLM claims trillion-scale support, but gives no model name, size, or repro path; don’t treat Day 0 coordination as a perf win.
HKR breakdown
hook knowledge resonance
open source
63
SCORE
H1·K0·R1
15:37
73d ago
The Verge · AI· rssEN15:37 · 05·16
Sony tries to explain that its AI Camera Assistant doesn’t suck
Sony says the Xperia 1 XIII AI Camera Assistant does not edit photos; it gives four suggestions for exposure, color, and background blur based on lighting, depth, and subject.
#Vision#Sony#The Verge#Product update
editor take
Sony’s AI Camera Assistant gives four shooting suggestions; the “photogenic angle” demo only shows zoom, so the AI label feels padded.
HKR breakdown
hook knowledge resonance
open source
61
SCORE
H1·K1·R0
15:28
73d ago
r/LocalLLaMA· rssEN15:28 · 05·16
Local speech to text for iOS using Apple Watch
The author released Dictawiz for Apple Watch recording and local iPhone transcription, citing Parakeet and Whisper support plus integrations with Notion, Obsidian, custom webhooks, and a Cloudflare memory layer; the post does not disclose latency, pricing, model sizes, or accuracy metrics.
#Audio#Tools#Memory#Apple
editor take
Dictawiz records on Apple Watch and transcribes locally on iPhone; no latency, pricing, or accuracy, so I don't buy the productivity pitch yet.
HKR breakdown
hook knowledge resonance
open source
64
SCORE
H1·K1·R1
15:18
73d ago
r/LocalLLaMA· rssEN15:18 · 05·16
Extension idea: llama-server with custom samplers
DeProgrammer99 proposed a llama-server custom sampler extension prototype, with one short C++ loop-detector example that breaks repeated 1-3 token loops seen in heavily quantized models. The branch targets llama.cpp master after MTP was merged, works with speculative decoding, and includes a Windows x64 Vulkan release plus an example command using Qwen3.6-27B with 32,768 context.
#Inference-opt#Code#Tools#DeProgrammer99
editor take
Title says llama-server custom samplers; body is 403, no patch details disclosed, so wait for a reproducible branch.
HKR breakdown
hook knowledge resonance
open source
63
SCORE
H0·K1·R1
14:54
73d ago
AI HOT (Curated Pool)· aihot-apiZH14:54 · 05·16
Show HN: Burn, Baby, Burn (Those Tokens)
A developer open-sourced “Burn, Baby, Burn” on GitHub, providing a tool for users to burn their own tokens to reduce total supply; the Hacker News post reached 100 points.
#GitHub#Hacker News#Open source
editor take
GitHub body only shows chrome, HN has 100 points; a token-burn tool smells like a gag, not an AI signal.
HKR breakdown
hook knowledge resonance
open source
28
SCORE
H0·K0·R0
14:40
73d ago
r/LocalLLaMA· rssEN14:40 · 05·16
macOS support in Lemonade has graduated out of beta
Lemonade moved macOS support out of beta and says five capability areas are available: OmniRouter, coding, image generation, speech generation, and transcription; the post also states the local AI tool uses a 3 MB portable binary across Linux, Windows, and macOS.
#Multimodal#Code#Audio#Lemonade
editor take
Lemonade says macOS is stable with 5 capability areas; Reddit 403s, so I won't endorse the 3 MB binary claim.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H0·K1·R1
14:15
73d ago
r/LocalLLaMA· rssEN14:15 · 05·16
Same double-pendulum prompt, same renderer, two models picked opposite θ conventions
The author tested Claude 3.5 Sonnet and DeepSeek V3 with the same double-pendulum contract, using θ1=π/2, θ2=π/2, and zero angular velocities; under one host renderer, the two outputs showed mirror-image behavior within one second.
#Code#Reasoning#Benchmarking#Claude 3.5 Sonnet
editor take
Same pendulum prompt split Claude 3.5 Sonnet and DeepSeek V3 within 1s; Reddit 403s, so don't benchmark from screenshots.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1
13:46
73d ago
AI HOT (Curated Pool)· aihot-apiZH13:46 · 05·16
Hangzhou Base Opens as a National Vocational Skills Training Site for Robots
The National AI Application Pilot Base for Embodied Intelligence opened in Hangzhou on May 16, and Hangzhou has gathered more than 700 robotics-related companies, with its embodied intelligence industrial cluster reaching 106.8 billion yuan in output value in 2025.
#Robotics#Hangzhou#国家人工智能应用中试基地#Policy
editor take
Hangzhou opened an embodied-AI pilot base with 700+ robotics firms; without open data and eval protocols, it's a policy showroom.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R0
12:49
73d ago
r/LocalLLaMA· rssEN12:49 · 05·16
Built a 6x Cheaper CodeRabbit Alternative Using Open Source Models
Reddit user Axintwo says PrixAI uses open source models for PR review and detected 10 of 10 planted issues in a test PR, while costing 6x less than CodeRabbit’s stated $60 per month plan.
#Code#Agent#CodeRabbit#PrixAI
editor take
PrixAI claims 10/10 detections at 6x lower cost; the body is 403, with no model, repo, or repro script.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
12:11
73d ago
Product Hunt · AI· rssEN12:11 · 05·16
pixserp
pixserp offers a live-web LLM endpoint with ten answer shapes, but the RSS post does not disclose pricing, supported models, latency, or API details.
#RAG#Tools#pixserp#Product update
editor take
pixserp discloses one endpoint and ten answer shapes; no models, latency, or pricing, so I’m filing this as a wrapper.
HKR breakdown
hook knowledge resonance
open source
42
SCORE
H0·K1·R0
11:34
73d ago
Hacker News Frontpage· rssEN11:34 · 05·16
OpenClaw Creator Spent $1.3M on OpenAI Tokens in 30 Days
The title says the OpenClaw creator spent $1.3 million on OpenAI tokens in 30 days; the post does not disclose usage volume, model mix, pricing structure, or billing evidence.
#OpenClaw#OpenAI#Commentary
editor take
OpenClaw’s creator claims $1.3M in OpenAI tokens over 30 days; without bills or model mix, I treat it as spend-bragging.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
11:03
73d ago
r/LocalLLaMA· rssEN11:03 · 05·16
Reduce Your GPU Power Limit
Reddit user NotArticuno tested GPU power-limit changes against TG128 generation and PP512 processing, likely using qwen3.5:9b; the post does not disclose the exact GPU model or numeric results in the RSS body.
#Inference-opt#NotArticuno#Qwen#Commentary
editor take
Title says lower GPU power limits; body is 403. No GPU model or tok/s, so don't call this inference optimization yet.
HKR breakdown
hook knowledge resonance
open source
52
SCORE
H1·K0·R1
10:22
73d ago
Synced (机器之心) · WeChat· rssZH10:22 · 05·16
Anthropic Brings Claude Code to a Card-Sized Computer
Anthropic gave developers a Cardputer at its Code With Claude event, and the post says the ESP32-S3 handheld development board can run the full Claude Code.
#Code#Tools#Anthropic#Claude
editor take
Cardputer running Claude Code cites a GitHub link, with no local inference disclosed; this smells like terminal-wrapper demo art.
HKR breakdown
hook knowledge resonance
open source
69
SCORE
H1·K1·R1
10:22
73d ago
Synced (机器之心) · WeChat· rssZH10:22 · 05·16
This Time, Robots Compete on Work, Not Flashy Demos
The 2026 Hangzhou International Embodied Robot Scenario Application Competition set three tracks and tested more than 200 teams in real scenarios including fire rescue, power inspection, data centers, underwater rescue, and warehouse logistics.
#Robotics#Agent#Multimodal#机器之心
editor take
Hangzhou tested 200+ robot teams in field-like tasks; useful, but no completion rates, failure rates, or procurement data yet.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R1
09:30
73d ago
Hacker News Frontpage· rssEN09:30 · 05·16
Δ-Mem: Efficient Online Memory for Large Language Models
The title presents Δ-Mem as an efficient online memory method for large language models; the post only discloses an arXiv URL, 36 Hacker News points, and 8 comments, and does not disclose the mechanism, benchmark results, model scale, latency, memory cost, or code availability.
#Memory#Research release
editor take
δ-mem claims 1.10× average gain with an 8×8 state; I buy the lightweight-memory angle, not agent longevity without code.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R1
07:28
73d ago
AI Chat-Group Daily (群聊日报)· atomZH07:28 · 05·16
2026-05-15 Chat Group Daily
The chat-group daily summarizes 5 AI discussion areas: Bloomberg reported a 0.2% employment drop across 18 BLS-labeled AI-exposed occupations, while Anthropic reset Claude Code 5-hour and weekly rate limits without changing the original reset schedule.
#Agent#Code#Tools#Bloomberg
editor take
Bloomberg says 18 AI-exposed jobs fell 0.2%; technical writers dropped 18.1%, while Claude Code's reset is just candy.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1

more

feeds

admin