ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

all posts

50 items · updated 3m ago
RSS live
2026-07-23 · Thu
18:35
5d ago
AI HOT (Curated Pool)· aihot-apiZH18:35 · 07·23
TheNumbers.com collapsed under AI crawlers and attacks, forced to rebuild
TheNumbers.com, the film industry's definitive data source, vanished for a week in March 2026 and returned as a shell. Founder Bruce Nash blames two waves of AI crawlers: training bots from 2024, then agentic AI from late 2025. Combined with security attacks, the site had to be rebuilt from scratch. The post doesn't disclose rebuild cost or timeline, but confirms 78,000+ films' historical data is gone for now.
#TheNumbers.com#Bruce Nash
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
18:29
5d ago
Hacker News Frontpage· rssEN18:29 · 07·23
Mozilla AI at ACM FAccT 2026: Guardrails Need the Same Scrutiny as Models
At ACM FAccT 2026 in Montreal, Mozilla AI argued that guardrails need the same rigorous evaluation as the models they govern. They tested 120 refugee-asylum scenario pairs across five languages (English, Farsi, Arabic, Kurdish-Sorani, Pashto) and found that text-only guardrails often miss factual errors—like whether an NGO exists or a law is current. So they built an agentic guardrail with web search. 35 attendees ran the demo: 90% of verdicts matched the tool-less version, but the tool-equipped one corrected factual mistakes. Performance depended heavily on the judge LLM—Claude Sonnet 4.6 used search on every run (4.1 calls/run), GPT-5 Nano almost never (0.2 calls/run). The post doesn't disclose latency or cost, but the takeaway is clear: reliable guardrails need tool access.
#Mozilla AI#ACM FAccT#Claude Sonnet 4.6#Benchmark
HKR breakdown
hook knowledge resonance
open source
65
SCORE
H0·K1·R0
18:23
5d ago
Hacker News Frontpage· rssEN18:23 · 07·23
ATProto wants to be the app layer protocol, but privacy isn't there yet
Luke Kanies wants to build review apps on ATProto to replace Yelp and GoodReads, with user-owned data and public/private sharing. After the Local First Conference, he finds ATProto's identity system ready but the protocol still public-only. The community is designing "permissioned data" but the post doesn't spell out when or how it will work.
#ATProto#Bluesky#Luke Kanies
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
18:11
5d ago
Hacker News Frontpage· rssEN18:11 · 07·23
Geekbench 7 adds AV1 encoding, Whisper captions, Jolt physics, and smarter multi-core scoring
Primate Labs ships Geekbench 7, a major cross-platform benchmark update. New media workloads: encode screen-sharing video with AV1, compress audio with Opus, and generate live captions via Whisper. Multi-core tests now only run workloads that are actually multi-threaded in real apps—HTML5 browsing is excluded because browsers are single-threaded. GPU benchmark adds ML tasks: face tracking filters, AI upscaling, and background blur, plus CUDA support for the first time. Datasets are larger: compression tests include more source code and documents; PDF tests add technical papers and park maps. Free for personal use; Pro 20% off until August 6.
#Benchmarking#Primate Labs#Geekbench#Jolt Physics
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
17:06
5d ago
Hacker News Frontpage· rssEN17:06 · 07·23
Claude-thermos: keep your Claude session warm
Claude-thermos is an open-source tool that keeps your Claude session alive by preventing idle timeouts. Useful for long conversations or background tasks. The post does not disclose implementation details or performance numbers.
#Claude#izeigerman#Open source
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
16:28
5d ago
r/LocalLLaMA· rssEN16:28 · 07·23
Apple M5's matmul cores are still underutilized by inference backends
M5 hardware supports INT8 activations (w4a8), but MLX and llama.cpp still run everything in 16-bit. The author wrote custom w8a8 kernels and got Gemma4 prefill on an M5 MacBook Air from 2,193 tps to 3,029 tps—nearly 10k tps at small context lengths. The code isn't one-click ready yet. Commenters note INT8 activations can hurt accuracy, but the author saw semantically identical decode output with a 4-bit QAT checkpoint. No full quality evaluation is provided, so take that with a grain of salt.
#Apple#MLX#llama.cpp
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
15:42
5d ago
Hacker News Frontpage· rssEN15:42 · 07·23
OneCLI: an open-source credential gateway that keeps secrets out of AI agents
OneCLI is an open-source credential gateway with a built-in vault. AI agents call external services through its CLI, and the gateway injects secrets without exposing plaintext keys. The repo has 2.6k stars with active issues and PRs. The post doesn't spell out which services are supported, what access control granularity looks like, or the latency overhead.
#OneCLI
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
15:18
5d ago
Hacker News Frontpage· rssEN15:18 · 07·23
Why Software Factories Fail: Harness Engineering Is Not Enough
This post from HumanLayer argues that pure engineering skill—writing code and setting up frameworks—isn't enough to make AI coding agents work in real business contexts. The author claims many 'software factory' projects fail because they neglect context engineering: precisely feeding business logic, constraints, and past decisions to the model. The post doesn't provide specific cases or data, but highlights the core tension: models can generate code, but generating the right code requires finer context management.
#Code#HumanLayer
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
15:11
5d ago
Hacker News Frontpage· rssEN15:11 · 07·23
Palmier Pro: open-source macOS video editor built for AI workflows
Palmier Pro is an open-source macOS video editor built for AI integration. It has 11.1k stars on GitHub and supports AI-driven editing features. The post does not disclose specific supported models, APIs, or performance benchmarks.
#Palmier Pro#GitHub
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R0
14:40
5d ago
Hacker News Frontpage· rssEN14:40 · 07·23
Data centers used 1.5% of global electricity in 2025; AI's share was 0.5%
Our World in Data breaks down IEA figures: data centers drew roughly 485 TWh in 2025, about 1.5% of global electricity and equal to Germany's annual generation. AI-focused facilities accounted for 155 TWh, or 0.5% of the global total. Non-AI workloads—email, streaming, banking—still made up two-thirds. The IEA's base-case projection sees data center demand nearly doubling to 945 TWh (3% of global electricity) by 2030, with AI driving most of the growth. Estimates vary widely: the Energy Institute's S&P Global figures are about 60% higher and include crypto mining. The post does not provide per-query energy numbers or a training-vs-inference split.
#International Energy Agency#IEA#Our World in Data
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
14:08
5d ago
Hacker News Frontpage· rssEN14:08 · 07·23
PullRun: Run same OCI images as containers or Firecracker microVMs
PullRun is a new open-source container runtime that runs the same OCI image as a Linux container, Firecracker microVM, or Apple Silicon VM. It uses zero-copy DAG storage and P2P image sync for faster startup and distribution. For AI inference, microVM isolation is stronger than plain containers, but the post doesn't disclose specific performance numbers or production use cases.
#PullRun#Firecracker#Apple Silicon#Open source
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
13:51
5d ago
Hacker News Frontpage· rssEN13:51 · 07·23
DARPA and U.S. Air Force fly AI-controlled F-16
A modified F-16 is flying under AI control with a safety pilot monitoring. The VENOM Autonomy Kit interfaces with flight controls without altering the jet's core software, letting a pilot toggle between human and AI control. This follows the X-62A dogfight demo and moves the capability onto a standard fleet aircraft. The next phase under DARPA's AIR program will test multi-agent teaming for uncrewed wingmen. The post does not disclose test duration, model architecture, or failure rates.
#DARPA#U.S. Air Force#VENOM program
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R0
11:20
5d ago
AI HOT (Curated Pool)· aihot-apiZH11:20 · 07·23
Kunlun CEO: Tokens alone won't build an AI-native org; models are the foundation
Kunlun CEO Fang Han said at WAIC that token consumption alone can't measure AI value—model capability needs engineering frameworks built by coding agents like Claude Code to become productive. He disclosed Kunlun is still training models and will release music, embodied world, and game world models, arguing models and compute are the long-term foundation for AI companies. He also warned that technical debt from AI coding could multiply production incidents, so code review and accountability must keep pace.
#昆仑万维#方汉#Claude Code
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
08:13
5d ago
r/LocalLLaMA· rssEN08:13 · 07·23
SLAI T-Rex: Full-Parameter Post-Training of DeepSeek-V4 on Ascend SuperPOD
This paper reports end-to-end full-parameter post-training of DeepSeek-V4 on Huawei Ascend NPU clusters. A hierarchical optimization framework pushes MFU to 34.22%, a 2.93× gain over the open-source baseline. Using DeepSeek-V4-Flash, the team built a 10K-sample SFT dataset for operations research and fine-tuned a specialist model. It hits 71.81% zero-shot Pass@1, beating GPT-5.4-Mini by 3.98 points and the base V4-Flash by 11.27 points. Weights are on ModelScope; the post doesn't mention a HuggingFace mirror.
#DeepSeek#DeepSeek-V4#DeepSeek-V4-Flash
editor take
34% MFU on Ascend NPU for full-parameter DeepSeek-V4 post-training is a 2.93x gain over baseline, but that's just normal by Nvidia standards.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H0·K1·R0
06:17
5d ago
r/LocalLLaMA· rssEN06:17 · 07·23
A 'caveman' Qwen3.6 27B claims 90% fewer tokens
A Reddit user spotted grug-27b on Hugging Face, a fine-tune of Qwen3.6 27B that rewrites outputs in a 'caveman' style. The model card claims over 90% fewer reasoning tokens and better benchmarks. If true, a 27B running at 3 tps on an old laptop could feel like 30 tps for the thinking part. The post doesn't disclose training details or evaluation methodology.
#Fine-tuning#Qwen#Hugging Face#Reddit
editor take
A Reddit user spotted a Qwen3.6 27B fine-tune that rewrites outputs in caveman style, claiming 90% fewer reasoning tokens—but no training details or eval methodology disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
05:12
5d ago
r/LocalLLaMA· rssEN05:12 · 07·23
Distillation accusations are overblown: API outputs ≠ stealing a model
A Reddit user argues that every strong open model release gets accused of being 'just distilled from GPT-4/Claude.' Real distillation requires access to logits (full probability distribution over vocabulary), not just API text outputs—that's synthetic data generation, not distillation. Many accused models perform well in domains where API outputs are filtered, suggesting they aren't simple copies. Identity confusion (a model claiming to be Claude) only proves data contamination, not wholesale distillation. The user notes these accusations land disproportionately on Chinese labs, looking more like a reflex dismissal than a technical assessment.
#Reddit#LocalLLaMA#GPT-4
editor take
A Reddit post draws a clear line between real distillation (needs logits) and training on API outputs, arguing the accusation is overused and lands disproportionately on Chinese labs.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
04:00
5d ago
Financial Times · Technology· rssEN04:00 · 07·23
Buyout groups hunt for software bargains after ‘SaaS-pocalypse’
Private equity firms are hunting for bargains in the SaaS sector after a valuation crash they call 'SaaS-pocalypse.' They target stable-revenue software companies whose share prices have fallen. The post does not disclose specific deals or target names.
editor take
PE firms are scooping up SaaS bargains after the valuation crash they call 'SaaS-pocalypse.' No specific targets named.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
02:50
5d ago
r/LocalLLaMA· rssEN02:50 · 07·23
MoE Models Around 2B Active Parameters: The Middle Ground
A Reddit user compiled a list of MoE models with ~2B active parameters, filling the gap between 1B and 3B+ tiers. These models target CPU inference or low-end GPUs with 4-12GB VRAM, potentially outperforming dense models of similar size. Listed models include Liquid LFM2 24B A2B, JetBrains Mellum 2 12B A2.5B, Moondream 3.1 9B A2B, DeepSeek V2 Lite 16B A2.4B, and others. However, community discussion is sparse, and real-world benchmarks are lacking.
#Liquid AI#JetBrains#Moondream
editor take
A Reddit list of ~2B active-param MoE models for 4-12GB VRAM, but real-world benchmarks are missing.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
01:33
5d ago
Hacker News Frontpage· rssEN01:33 · 07·23
Petals: Run LLMs at home, BitTorrent-style
Petals lets you run Llama 3.1 (up to 405B), Mixtral 8x22B, and other LLMs on a consumer GPU or Google Colab. You load a part of the model; others serve the rest. Single-batch inference hits ~6 tokens/sec for Llama 2 70B and ~4 tokens/sec for Falcon 180B—good enough for chatbots. You get PyTorch-level flexibility for fine-tuning and hidden states. The post doesn't disclose active node count or latency variance, so real-world speed may differ.
#Fine-tuning#Petals#BigScience#Llama 3.1
editor take
Petals runs LLMs like BitTorrent—your GPU loads part of the model, others serve the rest. Works for Llama 405B on one card, but speed depends on who's online.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H1·K1·R0
01:23
5d ago
r/LocalLLaMA· rssEN01:23 · 07·23
Make a 3B model act like a 30B+ with structural harness, not blind prompt engineering
Reddit user Foxtor shares a method to drastically improve output from 3B-8B local models: instead of open-ended goals, hardcode the cognitive steps (pain point, cost of inaction, solution, CTA) into a Markdown 'structural harness.' The model only fills slots with raw variables, not inventing narrative arcs. Comments mention similar work like tiny-coder and argue that harness matters more than one-shot prompting once a model reaches a certain intelligence threshold. The post doesn't specify which models were tested or quantify the improvement, but the approach is practical for local inference.
#Foxtor#Reddit#LocalLLaMA
editor take
Hardcode cognitive steps into a Markdown template; the small model just fills slots instead of inventing narrative arcs.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
01:10
5d ago
Bloomberg Technology· rssEN01:10 · 07·23
Khosla Ventures in Talks to Raise $5.5 Billion in New Funds
Khosla Ventures is in talks to raise $5.5 billion in new funds. That's a huge sum, signaling the veteran VC is doubling down on the AI boom. The post doesn't disclose the fund's focus, stage, or LP lineup—only the fundraising intent is confirmed.
#Khosla Ventures#Funding
editor take
Khosla Ventures is raising $5.5B—huge sum, but the post doesn't say stage or focus.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
2026-07-22 · Wed
23:34
5d ago
r/LocalLLaMA· rssEN23:34 · 07·22
Poolside Laguna S 2.1 hands-on: good for coding agents, weak on knowledge vs Qwen 3.6
Reddit users testing Poolside's Laguna S 2.1 report it's decent for coding agent tasks—slightly better tool calling than Qwen 3.6 27B—but significantly worse on knowledge and reasoning. Some users had to add flags to stop looping; in chat mode it barely thinks, and in the pi agent framework it can overthink, spending 20k tokens on a single refactor. Low-VRAM users (40GB) got only 2.5 t/s and found it not worth the setup hassle. Most still prefer Qwen 3.6 35B as the best all-rounder. The post does not disclose model size, training data, or official benchmarks.
#Code#Reasoning#Poolside#Qwen
editor take
Reddit users say Laguna S 2.1 is decent for coding agents but worse on knowledge and reasoning than Qwen 3.6.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
21:54
5d ago
Hacker News Frontpage· rssEN21:54 · 07·22
Real-world text-to-SQL is far from ready
Michael Stonebraker and Peter Baile Chen argue that current text-to-SQL benchmarks like Spider 1.0 (80%+ accuracy) and Bird-SQL fail to capture real-world data warehouse complexity. Production schemas use cryptic table and column names, business logic spans dozens of tables, and user phrasing varies wildly. The post catalogs the gaps—dirty data, missing metadata, complex joins—without offering a new solution. The takeaway: don't trust leaderboards; natural-language querying for non-programmers is still a long way off.
#Benchmarking#Michael Stonebraker#Peter Baile Chen#Communications of the ACM
editor take
Stonebraker says don't trust Spider's 80% accuracy—real schemas use cryptic names and 30-table joins, and no benchmark tests that.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
21:15
5d ago
Hacker News Frontpage· rssEN21:15 · 07·22
400 lines of Elisp turns GitHub issues into Org tasks
The author got tired of manually copying GitHub issues into Org Agenda and built a package called fj in one day. It delegates auth to the gh CLI, parses JSON responses, and presents issues in a vtable with a Transient menu. The whole thing is 392 lines of Elisp; basic behavior took 2.5 hours. The post highlights Emacs's malleable computing: dynamic code evaluation without restart, unlike the edit-compile-debug cycle of static languages. The post doesn't disclose whether fj is on MELPA or supports GitHub Enterprise.
#Code#Emacs#GitHub#Charles Choi
editor take
392 lines of Elisp to pull GitHub issues into Org Agenda by delegating auth to the gh CLI. A neat demo of Emacs as a malleable frontend, not a full client.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
20:54
5d ago
Bloomberg Technology· rssEN20:54 · 07·22
Google Cloud Backlog Hits $514 Billion
Google reported a cloud services backlog of $514 billion, a big jump from last quarter. This is the total value of signed but not yet recognized contracts, showing enterprise customers are making longer commitments to Google Cloud. For AI practitioners, it signals Google's infrastructure and AI platform (Vertex AI) are winning more long-term deals, shifting the competitive landscape.
#Google#Google Cloud
editor take
Google Cloud backlog hit $514B — enterprises are signing longer deals, and Vertex AI is landing them.
HKR breakdown
hook knowledge resonance
open source
65
SCORE
H0·K1·R0

more

feeds

admin