ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
33 srcsignal 72%cycle 04:32

posts · 2026-07-23

25 items · updated 3m ago
RSS live
2026-07-23 · Thu
08:13
61d ago
r/LocalLLaMA· rssEN08:13 · 07·23
SLAI T-Rex: Full-Parameter Post-Training of DeepSeek-V4 on Ascend SuperPOD
This paper reports end-to-end full-parameter post-training of DeepSeek-V4 on Huawei Ascend NPU clusters. A hierarchical optimization framework pushes MFU to 34.22%, a 2.93× gain over the open-source baseline. Using DeepSeek-V4-Flash, the team built a 10K-sample SFT dataset for operations research and fine-tuned a specialist model. It hits 71.81% zero-shot Pass@1, beating GPT-5.4-Mini by 3.98 points and the base V4-Flash by 11.27 points. Weights are on ModelScope; the post doesn't mention a HuggingFace mirror.
#DeepSeek#DeepSeek-V4#DeepSeek-V4-Flash
editor take
34% MFU on Ascend NPU for full-parameter DeepSeek-V4 post-training is a 2.93x gain over baseline, but that's just normal by Nvidia standards.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H0·K1·R0
07:17
61d ago
Synced (机器之心) · WeChat· rssZH07:17 · 07·23
Painting Actions into Video World Models for Cross-Embodiment Bidirectional Reasoning, with Fei-Fei Li
The article body is blocked by WeChat, only the title remains. Fei-Fei Li is involved in work that paints actions into video world models for cross-embodiment bidirectional reasoning. The post does not disclose specific methods, dataset size, or experimental results.
#Fei-Fei Li
editor take
Fei-Fei Li paints actions into video world models for cross-embodiment reasoning. Body blocked by WeChat, wait for paper.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
06:29
61d ago
Product Hunt · AI· rssEN06:29 · 07·23
Rehello: A people-memory companion for introverts to reconnect naturally
Rehello is a people-memory companion for introverts. Jot down messy notes after meeting someone; GPT-5.6 turns them into structured recall cards with who they are, where you met, what matters to them, and a conversation opener. It's not a CRM—it reduces the mental load of remembering people so reconnecting feels natural. The post doesn't disclose pricing or privacy policy.
#Rehello#GPT-5.6
editor take
Rehello uses GPT-5.6 to turn messy post-meeting notes into recall cards with a conversation opener for next time.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R1
06:17
61d ago
r/LocalLLaMA· rssEN06:17 · 07·23
A 'caveman' Qwen3.6 27B claims 90% fewer tokens
A Reddit user spotted grug-27b on Hugging Face, a fine-tune of Qwen3.6 27B that rewrites outputs in a 'caveman' style. The model card claims over 90% fewer reasoning tokens and better benchmarks. If true, a 27B running at 3 tps on an old laptop could feel like 30 tps for the thinking part. The post doesn't disclose training details or evaluation methodology.
#Fine-tuning#Qwen#Hugging Face#Reddit
editor take
A Reddit user spotted a Qwen3.6 27B fine-tune that rewrites outputs in caveman style, claiming 90% fewer reasoning tokens—but no training details or eval methodology disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
05:29
61d ago
Product Hunt · AI· rssEN05:29 · 07·23
OpenCode Superapp: Codex-level agents on local models, with voice
OpenCode Superapp is a new AI coding tool that lets you bring Codex-level agentic power to your own models—cloud, local, or self-hosted. Agents understand project context, work with files, Git, and terminals, support voice input, and can operate Mac apps via supervised Computer Use. Built on Codex and GPT-5.6, it's privacy-first and extensible via skills and MCPs. The post doesn't disclose pricing, benchmarks, or a supported model list.
#Code#OpenCode#Codex#GPT-5.6
editor take
OpenCode Superapp lets you run Codex-level coding agents on your own models (local/cloud), but pricing and supported model list are missing.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
05:12
61d ago
r/LocalLLaMA· rssEN05:12 · 07·23
Distillation accusations are overblown: API outputs ≠ stealing a model
A Reddit user argues that every strong open model release gets accused of being 'just distilled from GPT-4/Claude.' Real distillation requires access to logits (full probability distribution over vocabulary), not just API text outputs—that's synthetic data generation, not distillation. Many accused models perform well in domains where API outputs are filtered, suggesting they aren't simple copies. Identity confusion (a model claiming to be Claude) only proves data contamination, not wholesale distillation. The user notes these accusations land disproportionately on Chinese labs, looking more like a reflex dismissal than a technical assessment.
#Reddit#LocalLLaMA#GPT-4
editor take
A Reddit post draws a clear line between real distillation (needs logits) and training on API outputs, arguing the accusation is overused and lands disproportionately on Chinese labs.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
04:00
61d ago
Financial Times · Technology· rssEN04:00 · 07·23
Buyout groups hunt for software bargains after ‘SaaS-pocalypse’
Private equity firms are hunting for bargains in the SaaS sector after a valuation crash they call 'SaaS-pocalypse.' They target stable-revenue software companies whose share prices have fallen. The post does not disclose specific deals or target names.
editor take
PE firms are scooping up SaaS bargains after the valuation crash they call 'SaaS-pocalypse.' No specific targets named.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
03:16
62d ago
Product Hunt · AI· rssEN03:16 · 07·23
Basedash AI Kit turns its AI data analyst into an API for product-embedded BI
Basedash launched AI Kit, an API that exposes its AI-native BI platform. Developers can embed natural-language querying, auto-generated charts, and customer-scoped data access into their own products. It runs on GPT-5.6 and claims the #1 spot on BI Bench. The post doesn't disclose pricing tiers or rate limits, but mentions free options.
#Basedash#GPT-5.6#Y Combinator
editor take
Basedash turns its AI-native BI into an API, powered by GPT-5.6 and claiming #1 on BI Bench. Good for embedding natural-language analytics, but pricing and rate limits aren't disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
02:50
62d ago
r/LocalLLaMA· rssEN02:50 · 07·23
MoE Models Around 2B Active Parameters: The Middle Ground
A Reddit user compiled a list of MoE models with ~2B active parameters, filling the gap between 1B and 3B+ tiers. These models target CPU inference or low-end GPUs with 4-12GB VRAM, potentially outperforming dense models of similar size. Listed models include Liquid LFM2 24B A2B, JetBrains Mellum 2 12B A2.5B, Moondream 3.1 9B A2B, DeepSeek V2 Lite 16B A2.4B, and others. However, community discussion is sparse, and real-world benchmarks are lacking.
#Liquid AI#JetBrains#Moondream
editor take
A Reddit list of ~2B active-param MoE models for 4-12GB VRAM, but real-world benchmarks are missing.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
02:26
62d ago
Product Hunt · AI· rssEN02:26 · 07·23
HOL Guard: a firewall that sits between AI agents and your systems to block high-risk actions
HOL Guard positions itself as the first firewall for AI agents. It sits between agents and your systems, blocking high-risk actions like deleting production data or exposing secrets before they execute. Built by HOL, it's free, open source, and already has 400K+ downloads. The post doesn't spell out which agent frameworks it supports, latency overhead, or how rules are customized—hold off on judging production readiness until those details surface.
#HOL
editor take
HOL Guard is a free, open-source firewall for AI agents with 400K+ downloads, but it doesn't specify framework support or latency overhead.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
01:35
62d ago
Product Hunt · AI· rssEN01:35 · 07·23
Plow Mac App: Run GPT-5.6 on OpenClaw & Hermes safely on your Mac
Plow is a Mac app that installs OpenClaw and Hermes models with one click, compatible with GPT-5.6-sol. It integrates iMessage out of the box, plus Gmail and Slack, with usage monitoring. You can grant it file access. The post doesn't disclose pricing or model capabilities beyond the Product Hunt listing.
#Plow#OpenClaw#Hermes
editor take
Plow one-click installs OpenClaw and Hermes locally with iMessage integration, but no pricing or model specs disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
01:33
62d ago
Hacker News Frontpage· rssEN01:33 · 07·23
Petals: Run LLMs at home, BitTorrent-style
Petals lets you run Llama 3.1 (up to 405B), Mixtral 8x22B, and other LLMs on a consumer GPU or Google Colab. You load a part of the model; others serve the rest. Single-batch inference hits ~6 tokens/sec for Llama 2 70B and ~4 tokens/sec for Falcon 180B—good enough for chatbots. You get PyTorch-level flexibility for fine-tuning and hidden states. The post doesn't disclose active node count or latency variance, so real-world speed may differ.
#Fine-tuning#Petals#BigScience#Llama 3.1
editor take
Petals runs LLMs like BitTorrent—your GPU loads part of the model, others serve the rest. Works for Llama 405B on one card, but speed depends on who's online.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H1·K1·R0
01:23
62d ago
r/LocalLLaMA· rssEN01:23 · 07·23
Make a 3B model act like a 30B+ with structural harness, not blind prompt engineering
Reddit user Foxtor shares a method to drastically improve output from 3B-8B local models: instead of open-ended goals, hardcode the cognitive steps (pain point, cost of inaction, solution, CTA) into a Markdown 'structural harness.' The model only fills slots with raw variables, not inventing narrative arcs. Comments mention similar work like tiny-coder and argue that harness matters more than one-shot prompting once a model reaches a certain intelligence threshold. The post doesn't specify which models were tested or quantify the improvement, but the approach is practical for local inference.
#Foxtor#Reddit#LocalLLaMA
editor take
Hardcode cognitive steps into a Markdown template; the small model just fills slots instead of inventing narrative arcs.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
01:10
62d ago
Bloomberg Technology· rssEN01:10 · 07·23
Khosla Ventures in Talks to Raise $5.5 Billion in New Funds
Khosla Ventures is in talks to raise $5.5 billion in new funds. That's a huge sum, signaling the veteran VC is doubling down on the AI boom. The post doesn't disclose the fund's focus, stage, or LP lineup—only the fundraising intent is confirmed.
#Khosla Ventures#Funding
editor take
Khosla Ventures is raising $5.5B—huge sum, but the post doesn't say stage or focus.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
00:00
62d ago
● P1OpenAI Blog· rssEN00:00 · 07·23
OpenAI launches ChatGPT Health feature connecting Apple Health and medical records
OpenAI rolled out Health in ChatGPT to U.S. users. You can connect Apple Health and supported medical records so ChatGPT can compare lab results, summarize changes since your last visit, and factor in sleep or activity data. Connected health data won't train foundation models or target ads. It's live on web and iOS for Free, Go, Plus, and Pro plans; not yet in Codex.
#OpenAI#ChatGPT#Apple Health
why featured
Featured · importance 92 · hook + knowledge + resonance
editor take
OpenAI plugged Apple Health and medical records directly into ChatGPT, US-only for now. All three sources align on the same official announcement — the facts are solid.
sharp
OpenAI rolled out ChatGPT Health to all US users today. You can connect Apple Health and supported medical records, and the model will pull from your sleep data, activity, lab results, and medications during conversations. All three outlets are working off the same official blog post — no independent testing or third-party commentary yet, so we're only seeing OpenAI's own framing. Two things I'd flag. First, they say health data won't be used for training or ads, but there's no mention of external audits or compliance certifications. The privacy promise is self-declared for now. Second, on the model side: GPT-5.5 Instant is already powering this for free users, and they claim GPT-5.6 Sol is stronger on complex questions, but they didn't publish any benchmark numbers or say when Sol hits the free tier. Early testing showed over 70% of health conversations happened outside the dedicated Health tab, so they're now weaving health context into regular chats instead of forcing a separate space. That's the right direction, but whether the permission prompts are clear enough to prevent accidental health disclosures in casual conversations — we won't know until people actually use it. What's missing: external reviews and real error-rate data.
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
00:00
62d ago
Hugging Face Blog· rssEN00:00 · 07·23
Nunchaku 4-bit Diffusion Inference Now in Diffusers
The post does not disclose details. Hugging Face blog announces Nunchaku 4-bit diffusion inference is now integrated into Diffusers, making low-bit quantized model inference easier to use.
#Hugging Face#Nunchaku#Diffusers
editor take
Nunchaku 4-bit diffusion inference lands in Diffusers, but the post skips speedup and VRAM savings — you'll have to benchmark it yourself.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
00:00
62d ago
Computing Life · Share (鸭哥 research reports)· rssZH00:00 · 07·23
How Turbo and EVs Cross the Tipping Point: AI Competition Isn't Just About Tech
This article uses the cyclical shift between turbocharged and naturally aspirated engines in the auto industry to draw parallels with AI technology competition. The core argument: technical specs alone don't determine which route wins. External constraints (like carbon emission fines, policy subsidies) and scale effects (path dependency, learning curves) often rewrite the cost-benefit ledger, turning initially less-optimal solutions into mainstream choices. Examples include turbo engines boosted by EU emission regulations, CMOS sensors overtaking CCD via the mobile phone market, and EVs jumpstarted by policy. The return of naturally aspirated engines in hybrids shows that when system division of labor changes—electric motors fill the low-end torque gap—the complexity cost of turbo becomes a liability. For AI practitioners: don't just stare at benchmark scores; watch how policy, market, and supply chain reshape the real-world cost-performance of technical routes.
#Volkswagen#Toyota#Sony
editor take
Uses the turbo vs. naturally aspirated engine cycle to show how external constraints (regulation, scale) rewrite a tech route's cost-benefit ledger—more useful than just staring at benchmarks.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0

more

feeds

admin