ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

posts · 2026-05-09

24 items · updated 3m ago
RSS live
2026-05-09 · Sat
07:09
80d ago
● P1AI HOT (Curated Pool)· aihot-apiZH07:09 · 05·09
ERNIE 5.1 Released With Pretraining Cost at 6% of Comparable Models
Baidu released ERNIE 5.1, saying it builds on ERNIE 5.0 pretraining and improves search, reasoning, knowledge QA, creative writing, and agent capabilities, with pretraining cost at about 6% of comparable models.
#Reasoning#Agent#Baidu#ERNIE
why featured
Featured · importance 87 · hook + knowledge + resonance
editor take
Baidu leads ERNIE 5.1 with “6% pretraining cost,” which reads like a DeepSeek-era efficiency counter, not a capability break.
sharp
Baidu is selling ERNIE 5.1 as a cost story, and I don’t buy the clean headline yet. The hard number is “about 6% of comparable models” for pretraining, built on ERNIE 5.0. The missing parts are the comparator model, token count, compute accounting, and whether this means full pretraining or incremental continued training. Without that, 6% is a denominator game. The capability list covers search, reasoning, knowledge QA, creative writing, and agents, which is basically the standard frontier-model menu. After DeepSeek, every Chinese lab knows efficiency is the winning narrative. DeepSeek-R1 paired the cost story with open weights and visible reasoning UX. ERNIE 5.1, from this RSS snippet, gives no benchmark, pricing, latency, API detail, or context window. Baidu has to show that cheaper training becomes cheaper inference or better developer throughput, not just a slide-friendly ratio.
HKR breakdown
hook knowledge resonance
open source
87
SCORE
H1·K1·R1
05:29
80d ago
r/LocalLLaMA· rssEN05:29 · 05·09
Caliby Open-Sourced: Embedded High-Performance Vector Database for AI Agents
Sea-Land AI and Michael Stonebraker’s team open-sourced Caliby, an embedded vector retrieval library with HNSW, DiskANN, and IVF+PQ; the title claims 4x pgvector performance and stronger disk-storage results than FAISS, while the post does not disclose full benchmark methodology in the snippet.
#Agent#RAG#Embedding#Sea-Land AI
editor take
Caliby claims 4x pgvector speed; the body is 403-blocked, so benchmark conditions are undisclosed and I don't buy it yet.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
05:24
80d ago
AI HOT (Curated Pool)· aihot-apiZH05:24 · 05·09
The Dumbest Way to Raise Lobsters Is Repeating the Same Line Every Time
Garry Tan published the OpenClaw prompt, which tells AI agents to avoid one-off tasks and use a six-step workflow to retain repeatable skills for daily reports, emails, and similar recurring work.
#Agent#Tools#Memory#Garry Tan
editor take
Garry Tan published OpenClaw’s prompt and six-step workflow. Treating repeated asks as failure is product discipline, not agent magic.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
05:21
80d ago
r/LocalLLaMA· rssEN05:21 · 05·09
What llama.cpp's WebUI Has and What It Lacks
A Reddit user compared five development chat UIs and preferred llama.cpp WebUI for its context token counter; the cited gaps are conversation loss after failed tool calls, no project-level system prompts, and no built-in MCP tool hiding controls.
#Tools#Memory#llama.cpp#Jan.ai
editor take
Reddit body is 403, so only the 5-UI summary stands; llama.cpp WebUI wins token counting, then loses chats on tool failure.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H0·K1·R1
04:53
80d ago
Hacker News Frontpage· rssEN04:53 · 05·09
Using Claude Code: The Unreasonable Effectiveness of HTML
The HN item links to a Claude Code and HTML case post with 38 points and 14 comments; the RSS snippet only provides example links and does not disclose the method, task setup, or evaluation conditions.
#Code#Anthropic#Claude#Commentary
editor take
Claude Code case has 38 points and 14 comments. No task setup disclosed; don’t canonize HTML yet.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K0·R1
04:19
80d ago
AI HOT (Curated Pool)· aihot-apiZH04:19 · 05·09
Hermes Agent Tops OpenRouter Global Token Ranking
NousResearch says Hermes Agent ranks No. 1 in OpenRouter’s global token ranking; the post does not disclose the measurement period, token volume, or model version.
#Agent#NousResearch#OpenRouter#Benchmark
editor take
Hermes Agent tops OpenRouter’s token chart, but period and volume are undisclosed; treat it as usage heat, not capability proof.
HKR breakdown
hook knowledge resonance
open source
63
SCORE
H1·K0·R1
04:05
80d ago
AI HOT (Curated Pool)· aihot-apiZH04:05 · 05·09
StepAudio 2.5 TTS ranks top three globally in voice arena blind test
StepFun’s StepAudio 2.5 TTS ranked third on the Artificial Analysis voice arena blind-test leaderboard with an Elo score of 1187, priced at $85 per million characters and generating 37.6 characters per second.
#Audio#StepFun#Artificial Analysis#Google
editor take
StepAudio 2.5 TTS ranks third at Elo 1187; $85/M chars is pricey, but StepFun is now biting Google in TTS.
HKR breakdown
hook knowledge resonance
open source
70
SCORE
H1·K1·R1
03:47
80d ago
r/LocalLLaMA· rssEN03:47 · 05·09
Is Qwen3-coder the best kept secret out there?
A Reddit user says Qwen3-coder-next for MLX uses about 80GB of memory on an M2 Ultra 192GB Mac and runs faster than Qwen 3.5-35B-a3B; the post does not disclose its parameter count.
#Code#Fine-tuning#Inference-opt#Qwen
editor take
Qwen3-coder-next is claimed at 80GB RAM, but Reddit 403s; no params or benchmarks, so I don't buy the “secret” hype.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1
03:27
80d ago
AI HOT (Curated Pool)· aihot-apiZH03:27 · 05·09
Codex Chrome Extension Installation and Usage Notes
The user completed one shopping task with the Codex Chrome extension; installation requires the latest Codex version and official subscription login, while third-party API mode is not supported.
#Agent#Tools#Codex#Chrome
editor take
Codex Chrome completed 1 shopping task; third-party APIs and Hong Kong nodes are blocked, so this smells subscription-gated.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1
03:06
80d ago
AI HOT (Curated Pool)· aihot-apiZH03:06 · 05·09
GPT Image 2 Prompt: Ink-Wash Style Slides/PPT
The post introduces an ink-wash slide prompt template with six structural parts: title, key points, visual elements, layout preferences, text hierarchy, and continuity notes, while the body does not disclose model settings, pricing, or reproducible generation parameters.
#Multimodal#Vision#GPT Image 2#Codex
editor take
GPT Image 2 template lists six fields; no settings are disclosed, so I don’t buy this non-reproducible PPT aesthetics hack.
HKR breakdown
hook knowledge resonance
open source
64
SCORE
H1·K1·R0
02:44
80d ago
AI HOT (Curated Pool)· aihot-apiZH02:44 · 05·09
GPT Image 2 Prompt: Chinese Tech News Viral Cover Generator
The prompt framework asks AI to generate 16:9 Chinese tech news cover images from article content, using sections for news context, oversized headline, main visual, data cards, and bottom summary while adapting colors, fonts, background, brand cues, and industry sentiment.
#Multimodal#Vision#GPT Image 2#Product update
editor take
GPT Image 2 prompt targets 16:9 news covers; only a snippet, no samples or consistency tests—smells like thumbnail SOP.
HKR breakdown
hook knowledge resonance
open source
61
SCORE
H0·K1·R0
02:32
80d ago
Bloomberg Technology· rssEN02:32 · 05·09
China’s Top Economic Planner Urges Stronger Coordination on AI
The title says China’s top economic planner urged stronger AI coordination and oversight, with publication time listed as 2026-05-09T02:32:17.908Z. The body provided is Bloomberg page boilerplate and does not disclose the specific agency name, coordination mechanism, regulatory measures, affected companies, implementation timeline, or enforcement conditions.
#Bloomberg#Policy
editor take
The title only says China’s planner wants stronger AI coordination; no mechanism is disclosed, so don’t price it as regulation yet.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R1
01:49
80d ago
r/LocalLLaMA· rssEN01:49 · 05·09
Those Who Like Gemma4 Models: How Are You Using Them?
A Reddit user tested Gemma4 31B Q5 and 27B Q8 for Windows coding and tool use; the post says Gemma4 still struggles after 3-4 prompts to distinguish a pi harness skill from a tool call.
#Code#Tools#Vision#Gemma
editor take
Reddit returns 403; Gemma4 31B Q5 tool-call failure lacks prompts and reproducible conditions.
HKR breakdown
hook knowledge resonance
open source
58
SCORE
H0·K1·R1
01:06
80d ago
r/LocalLLaMA· rssEN01:06 · 05·09
Qwen3.6 35B A3B uncensored heretic Native MTP Preserved is out
LLMFan46 released Qwen3.6-35B-A3B uncensored heretic in Safetensors, GGUF, NVFP4, and GPTQ-Int4 formats. The post says all releases preserve the full MTP tensors, counted as 19 entries in Safetensors and 20 in GGUF because a fused gate_up_proj tensor is split.
#Inference-opt#Benchmarking#Qwen#LLMFan46
editor take
Title claims Qwen3.6-35B-A3B keeps 19 MTP tensors; body is 403, so I don’t buy the uncensored-quality pitch yet.
HKR breakdown
hook knowledge resonance
open source
66
SCORE
H1·K1·R1

more

feeds

admin