ax@ax-radar:~/all $ grep -v 'tier=excluded' stream.log
40 srcsignal 72%cycle 04:32

all posts

50 items · updated 3m ago
RSS live
2026-07-09 · Thu
18:19
19d ago
Product Hunt · AI· rssEN18:19 · 07·09
Notion launches Ship OS: an agent-native way to ship software
Notion launched Ship OS today, an agent-native tool that runs the entire product development cycle inside Notion. Agents handle triaging, routing, and summarizing from customer feedback to a merged PR; the team only makes judgment calls. The post doesn't disclose which model powers the agents, whether it supports on-prem deployment, or pricing.
#Notion
editor take
Notion turns product dev into a doc workflow—agents triage and route, humans just decide.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R1
17:36
19d ago
Product Hunt · AI· rssEN17:36 · 07·09
Kastra: runtime authorization layer for Claude, Cursor, Codex and OpenClaw
Kastra is a runtime authorization layer that intercepts AI agent actions before execution and enforces policies in under 1 ms. It covers tools, prompts, inputs, and outputs across Claude Code, Cursor, Codex, OpenClaw, and the Anthropic/OpenAI SDKs. The pitch is preventing unauthorized tool use, prompt injection, and sensitive data exposure. The post does not disclose pricing, deployment details, or real-world customer stories—treat it as an early-stage Product Hunt launch for now.
#Agent#Kastra#Anthropic#OpenAI
editor take
Kastra intercepts AI agent actions before execution and enforces policies in under 1 ms.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
17:17
19d ago
TechCrunch AI· rssEN17:17 · 07·09
Meta's new AI chips will begin production in September
Meta will start producing its latest in-house AI chips in September to cut GPU spending. The chips are part of the MTIA family and use a modular design so components can be swapped as needs change. At least one chip passed testing in about six weeks. Meta is working with Broadcom on design, TSMC on manufacturing, and sourcing RAM from Samsung, storage from Sandisk, and fiber-optic gear from Sumitomo Electric.
#Meta#Broadcom#TSMC
editor take
Meta's next MTIA chips hit production in September, using a modular design with Broadcom, TSMC, and Samsung.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
17:10
19d ago
AI HOT (Curated Pool)· aihot-apiZH17:10 · 07·09
OpenAI CEO calls new model 'best ever' in short post
Sam Altman tweeted that the new model is OpenAI's best ever, alongside what he calls their best blog post. The post doesn't disclose the model name, capabilities, or release timeline—only a link to openai.com.
#OpenAI#Sam Altman
editor take
Sam Altman says the new model is OpenAI's best ever, with their best blog post—but no name, capabilities, or timeline disclosed.
HKR breakdown
hook knowledge resonance
open source
25
SCORE
H0·K0·R0
16:47
19d ago
Hacker News Frontpage· rssEN16:47 · 07·09
Devthropology: A GitHub Repo Census Tool
Devthropology is a GitHub repo analytics tool that surfaces contributor activity, merge time, and language breakdown. Using the Sentry repo as a demo, it shows 1,222 total contributors, 211 active in 3 months, 89% merged PR rate, and a median author tenure of 2.4 years. The post doesn't disclose pricing or whether it's open-source—only a demo page is available.
#Devthropology#Sentry
editor take
Devthropology runs a census on GitHub repos—Sentry demo shows median author tenure 2.4y and 89% merged PR rate, but no pricing or open-source info yet.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
15:26
19d ago
Hacker News Frontpage· rssEN15:26 · 07·09
Kastra: sub-millisecond authorization for AI actions, covering Claude Code, Cursor, and Codex
Kastra is a runtime authorization layer that checks every AI action—prompts, tool calls, shell commands, API requests—before execution, with p99 latency of 0.8ms. It covers Claude Code, Cursor, Codex CLI, and OpenClaw browser agents, blocking risky operations like rm -rf, force-push, or writing secrets to disk. Recon scans AI history to surface past dangerous actions and drafts policies for them. Deployment options include cloud, self-hosted, and air-gapped; audit trails are signed and append-only, streamable to SIEM or S3. The post says it's used by banks, federal agencies, and frontier AI labs, with SOC 2 Type II in audit and ISO 27001 Stage 2 targeted for 2026.
#Kastra#Claude Code#Cursor
editor take
A 0.8ms authorization layer that blocks rm -rf and secret writes for Claude Code, Cursor, and Codex — already in use at banks and federal agencies.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
15:19
19d ago
Hacker News Frontpage· rssEN15:19 · 07·09
LazyPi: One-command setup for Pi coding agent with 60+ community skills
LazyPi is a one-command installer that adds 60+ community skills, 67 themes, MCP support, sub-agents, and persistent memory to the Pi coding agent. Pi ships minimal by design; LazyPi bundles the best community packages so you don't have to hunt them down. The post doesn't discuss performance overhead or stability impact, but the time saved is real.
#Code#Pi#Earendil#Mario Zechner
editor take
One command adds 60+ skills, 67 themes, MCP, and persistent memory to Pi. Saves hours of config.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K1·R0
15:01
19d ago
AI HOT (Curated Pool)· aihot-apiZH15:01 · 07·09
OpenAI Seeks Design Partners for GPT-Live API
OpenAI is recruiting design partners to test the GPT-Live API. Developers can apply to build new apps or integrate it into existing products. The post doesn't spell out the API's capabilities, pricing, or release timeline.
#OpenAI
editor take
OpenAI is recruiting design partners for the GPT-Live API, but the post doesn't spell out capabilities, pricing, or timeline — I'd hold off on excitement.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
14:50
19d ago
Hacker News Frontpage· rssEN14:50 · 07·09
Mozilla.ai: The next AI battleground is infrastructure, not models
Mozilla.ai argues that model capability is no longer the bottleneck—infrastructure is. In 2026, production AI budgets blow out in months, costs are opaque, teams juggle dozens of models, and governance tooling is missing. They pitch their own product Otari as the control layer that routes requests by cost, capability, latency, and compliance. The post doesn't disclose Otari's pricing or customer traction.
#Mozilla.ai#Otari
editor take
Mozilla.ai argues models are no longer the bottleneck—infrastructure is. They pitch Otari as the control layer.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
14:45
19d ago
Hacker News Frontpage· rssEN14:45 · 07·09
Wire moves its AI context containers off Cloudflare Durable Objects
Wire builds context containers for AI agents, originally all on Cloudflare Durable Objects. The team moved off due to four structural limits: the vector index lived in a separate service, adding a network hop on the hot retrieval path; compute couldn't sit next to data, so multi-stage retrieval pipelines suffered latency variance; placement was fixed at creation with no dedicated capacity tier; and self-hosting was impossible. The new runtime runs on Bun on Fly Machines, with one SQLite file per container and sqlite-vec embedded in-process. Warm tool calls dropped from ~0.4s to ~0.3s, cold start from 3.7s to 1.4s. Durability is bought back via continuous WAL shipping to object storage, acking writes in ~100ms. Recall@5 rose from 78.1% to 89.1%, though the team notes a newer embedding model also contributed. The new runtime is in beta, with plans to open-source it.
#Wire#Cloudflare#Cloudflare Durable Objects
editor take
Wire moved its AI agent context containers off Cloudflare Durable Objects because vector search required an external hop. Warm calls dropped to ~0.3s and Recall@5 jumped from 78.1% to 89.1%, though...
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
14:33
19d ago
AI HOT (Curated Pool)· aihot-apiZH14:33 · 07·09
Mistral adds version control for prompts and skills
Mistral introduces a system of record for prompts and skills in Studio. You can version, rollback, and collaborate on prompts like code. The post doesn't disclose supported models or pricing, but the idea is to stop managing prompts via copy-paste.
#Mistral
editor take
Mistral adds version control for prompts in Studio — treat them like code. No pricing or model details yet.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
14:10
19d ago
● P1Hacker News Frontpage· rssEN14:10 · 07·09
Meta releases Muse Spark 1.1 multimodal reasoning model for agentic tasks
Meta Superintelligence Labs released Muse Spark 1.1, a multimodal reasoning model with major gains in tool use, computer use, and coding. It zero-shot generalizes to new tools and MCP servers, manages a 1M-token context window, and compacts memory to keep critical steps. The model orchestrates multi-agent systems, delegating tasks to parallel subagents to cut end-to-end latency. Coding improvements cover bug fixes, feature additions, and large code migrations in complex codebases. It is live in Meta AI's Thinking mode and in the new Meta Model API public preview.
#Reasoning#Code#Meta#Meta Superintelligence Labs
why featured
Featured · importance 94 · hook + knowledge + resonance
editor take
Meta dropped Muse Spark 1.1, a multimodal reasoning model that can use tools, write code, and control a computer, with a 1M-token context window — but no pricing yet.
sharp
Meta released Muse Spark 1.1 today from their Superintelligence Labs. Two sources picked it up, with HN linking straight to the official blog — the community is clearly watching Meta's agent play closely. The model pushes three things: multimodal reasoning, tool use, and computer control. It has a 1M-token context window and can remember actions from way earlier in a session, compacting what it needs to keep. Meta claims it's much faster than the original Muse Spark on complex codebases, multi-app desktop workflows, and multi-agent orchestration, with zero-shot generalization to new MCP servers and custom skills. I'd take the benchmark charts with a grain of salt — the images in the blog are too low-res to read actual numbers, and I haven't seen third-party evals yet. No pricing has been disclosed either. If you're thinking about building on this, watch the Meta Model API public preview for real latency and cost data before committing.
HKR breakdown
hook knowledge resonance
open source
94
SCORE
H1·K1·R1
14:08
19d ago
Financial Times · Technology· rssEN14:08 · 07·09
Computacenter shares rise as FTSE 100 new entrant taps into AI boom
Computacenter, a new FTSE 100 member, saw its shares rise as it benefits from businesses investing in AI infrastructure. The company provides hardware and networking for AI deployment. The article doesn't specify the exact share price increase.
#Computacenter#FTSE 100
editor take
Computacenter, a new FTSE 100 member, got a share bump from selling hardware for AI deployment.
HKR breakdown
hook knowledge resonance
open source
45
SCORE
H0·K0·R0
14:01
19d ago
AI HOT (Curated Pool)· aihot-apiZH14:01 · 07·09
TeXada: Local math agent turns handwriting to LaTeX offline
The community built TeXada, a local math agent on MiniCPM5-1B and MiniCPM-V 4.6. It converts natural language to LaTeX, OCRs handwriting or image formulas into editable LaTeX, and fixes LaTeX errors. All inference runs locally with no cloud dependency, ensuring privacy. Targets students, researchers, and developers. Open-sourced on GitHub; models on HuggingFace. The post doesn't spell out OCR accuracy or latency, so take that with a grain of salt.
#Code#OpenBMB#MiniCPM#TeXada
editor take
Community-built TeXada turns MiniCPM into a local math agent: OCR formulas, convert to LaTeX, fix errors, all offline.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
13:30
19d ago
● P1Hacker News Frontpage· rssEN13:30 · 07·09
Anthropic launches Claude Reflect dashboard to visualize user AI usage patterns
Anthropic launched a beta feature called Reflect inside Claude’s web and desktop settings. It visualizes your chat activity over the past 1–12 months: when you use Claude most, which topics dominate, and what task patterns emerge. The report also maps your usage to Anthropic’s 4D AI Fluency Framework—Delegation, Description, Discernment, Diligence—and offers practical tips, like starting a Project instead of re-explaining context. Incognito chats, health-integration conversations, and source files from connected tools are excluded; sensitive topics appear only at a high level. Available now for Free, Pro, and Max users with Memory turned on; Cowork conversation support is coming soon.
#Anthropic#Claude#MIT Media Lab Advancing Humans with AI (AHA)
why featured
Featured · importance 86 · hook + knowledge + resonance
editor take
Anthropic added a usage dashboard to Claude that summarizes your chat history, sets break nudges, and ships with a 4D AI fluency framework. Sensitive topics still appear at a high level despite pri...
sharp
This is Anthropic's own announcement, picked up by HN and AIhot with no third-party testing or user reports yet, so everything we know comes straight from the official post. The dashboard itself is straightforward: pick a 1/3/6/12-month window and get a summary of your top topics, peak usage times, and soon, total time spent. The more interesting layer is the 4D AI Fluency Framework baked into it—Delegation, Description, Discernment, Diligence. It doesn't just count chats; it tries to characterize how you work with Claude, like whether you nail down strategy yourself before delegating, or whether you tend to rework drafts in your own voice. It'll also nudge you toward practical moves, like starting a Project instead of re-explaining context every time. On privacy: incognito chats are excluded, source files from connected tools aren't pulled in, and health integrations are fully carved out. But sensitive conversations can still show up at a high level in your summary. Anthropic brought in advisors from MIT Media Lab and Boston Children's Hospital to shape this, which signals they know the territory is tricky. What I'd wait on: there are no real user screenshots yet, just polished product images. No word on whether summaries hold up equally well across languages. And the whole thing requires Memory to be on—the more you let Claude remember, the richer your reflection gets, which is a tradeoff worth thinking through before you flip the switch.
HKR breakdown
hook knowledge resonance
open source
86
SCORE
H1·K1·R1
13:23
19d ago
Hacker News Frontpage· rssEN13:23 · 07·09
FableCut: A browser video editor AI agents can drive, zero deps
FableCut is a zero-dependency browser video editor that AI agents can drive via JSON timeline, MCP, and REST APIs. It features live-reloading UI, letting models edit video like using a tool. The post doesn't disclose supported video formats or performance benchmarks yet.
#FableCut#ronak-create#Open source
editor take
FableCut lets AI agents drive video editing via JSON timeline in-browser with zero deps, but the post doesn't disclose supported formats or benchmarks.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
13:19
19d ago
Hacker News Frontpage· rssEN13:19 · 07·09
Knock built its AI agent with a virtual filesystem and bash instead of piling on API tools
Knock shipped the Knock Agent in March 2026 to manage messaging workflows, templates, and audiences from the dashboard, Slack, or API. Their first prototype used one tool per API endpoint, which would have bloated the context window. Inspired by Vercel's approach, they switched to giving the agent a virtual filesystem and bash. The agent explores account data with commands like ls and cat, edits files locally, and calls back to persist changes. They ported Vercel's just-bash TypeScript library to Elixir, reusing its test suite. The pattern is closer to Claude Code than a traditional in-product assistant.
#Knock#Vercel#just-bash
editor take
Knock gave their agent a virtual filesystem and bash instead of one tool per API endpoint—ls/cat to explore account data, inspired by Vercel's approach. Closer to Claude Code than a typical in-prod...
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
13:02
19d ago
● P1Ben's Bites· rssEN13:02 · 07·09
Cursor and SpaceXAI release general-purpose model Grok 4.5
SpaceXAI and Cursor jointly trained Grok 4.5, landing between Opus 4.7 and 4.8 in performance but 6x cheaper than Opus and 3x cheaper than GPT-5.5 on a per-token basis. OpenAI rolled out GPT-5.6 (Sol, Terra, Luna) to all users; early testers say Sol is less smart than Fable but far more reliable. ChatGPT Voice got new GPT-Live-1 and Live-1-mini models that can talk while you speak and use GPT-5.5 in the background. Anthropic extended Fable 5 access for Claude subscribers to July 12—the post doesn't explain the repeated delays. Meta introduced Muse Image and Muse Video; image editing and text rendering look solid, but images still have an AI look, and the video model is in preview.
#Code#Audio#Vision#SpaceXAI
why featured
Featured · importance 92 · hook + knowledge + resonance
editor take
Cursor and SpaceXAI co-trained Grok 4.5, slotting between Opus 4.7 and 4.8 but at 1/6 the cost. Both sources point to the same xAI announcement, so the numbers are likely solid.
sharp
The headline here isn't model capability, it's pricing. xAI positions Grok 4.5 as Opus-class but openly says it lands between Opus 4.7 and 4.8 — so it's half a step behind Anthropic's best. The real punch is cost: 6x cheaper than Opus, 3x cheaper than GPT-5.5. Cursor baked it in with higher usage limits, which tells me they're betting on undercutting the competition for high-frequency coding users. Both sources are working off the same xAI blog post, no independent evals yet. I'd wait for third-party benchmarks before believing the performance claims. Also, the SpaceXAI + Cursor co-training arrangement is odd — a rocket company and an IDE maker jointly training a general-purpose model, and neither source explains why these two specifically teamed up.
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
13:00
19d ago
● P1OpenAI Blog· rssEN13:00 · 07·09
GPT-5.6 becomes preferred model for Microsoft 365 Copilot
OpenAI made GPT-5.6 the default model inside Microsoft 365 Copilot—Word, Excel, PowerPoint, Chat, and Cowork. The pitch is more useful work per token and fewer rounds of prompting. The post doesn't share performance benchmarks, latency figures, or whether Copilot pricing changes.
#OpenAI#Microsoft#Nitin Agrawal
why featured
Featured · importance 94 · hook + resonance
editor take
OpenAI rushed to call GPT-5.6 the 'preferred model' for Copilot on launch day, but didn't define what 'preferred' means or deny that Microsoft is using its own MAI models to cut costs.
sharp
Here's the context: a few days ago Bloomberg reported that Microsoft is increasingly using its own MAI models in Word and Excel to reduce costs. On Thursday, during the GPT-5.6 launch, OpenAI published a blog post calling itself the 'preferred model' for Microsoft 365 Copilot. Both sources covering this are citing the same OpenAI blog, so the messaging is entirely one-sided — Microsoft hasn't echoed it. The phrase 'preferred model' is slippery. It doesn't promise exclusivity or disclose traffic share. TechCrunch flagged this directly: nobody ever said OpenAI models would be fully removed from Copilot, just that Microsoft was adding its own models to the mix. OpenAI's blog doesn't refute that reporting; it reads more like a public posture move to calm breakup chatter. I'd take this with a grain of salt. What's missing: confirmation from Microsoft, actual traffic split numbers, and what 'preferred' means contractually. If it's just a blog post line, it's PR, not a product roadmap shift.
HKR breakdown
hook knowledge resonance
open source
94
SCORE
H1·K0·R1
13:00
19d ago
● P1TechCrunch AI· rssEN13:00 · 07·09
Ollama raises $65M Series B, reaches nearly 9 million users
Ollama, the tool that lets devs run open-weight models locally on their PCs, raised a $65M Series B led by Theory Ventures. That follows a $15M Series A led by Benchmark, bringing total funding to $88M. Launched in 2023, it has 176K GitHub stars, nearly 17K forks, and close to 9M users. The post doesn't disclose valuation, revenue, or commercialization plans.
#Ollama#Theory Ventures#Benchmark
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
Ollama raised $65M Series B with nearly 9M users, but both sources are repeating the company's own numbers — no independent verification of active usage or revenue yet.
sharp
Ollama just closed a $65M Series B led by Theory Ventures, following Benchmark's $15M Series A — $88M total raised. User count is nearly 9 million, with 176K GitHub stars. Both sources are running the same company-provided numbers, so there's no independent usage data to cross-check. I'd take the user figure with a grain of salt. Ollama is genuinely the easiest way to run open-weight models locally, and downloads are massive, but "user" could mean anything — installs, monthly actives, or just people who tried it once. The GitHub stars are the harder signal here: 176K puts it in the top tier of dev tools. The real question isn't the raise size, it's the business model. Ollama is free, and the company hasn't said how it plans to make money. $65M buys runway for infra and hiring, but open-source dev tools have a rough track record converting to paid. LM Studio and Jan are chasing the same audience. If the next round comes with a valuation jump and still zero revenue, that's a different story.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
13:00
19d ago
The Verge · AI· rssEN13:00 · 07·09
FL Studio 2026 turns its AI chatbot into your assistant engineer
FL Studio 2026 upgrades its built-in AI chatbot Gopher from an interactive manual to an assistant engineer. Gopher can now perform mixing tasks like burying an instrument in the mix, but won't play piano for you. The post doesn't spell out which specific operations Gopher supports, whether it relies on cloud models, or if it costs extra.
#Image Line#FL Studio
editor take
FL Studio's Gopher chatbot graduates from manual to mixing assistant, but the post skips pricing, cloud model, and supported operations.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
12:20
19d ago
Hacker News Frontpage· rssEN12:20 · 07·09
Entire CEO Thomas Dohmke on how Git hosting must evolve for the agent coding boom
Thomas Dohmke, former GitHub CEO now at Entire, argues that as AI agents generate most code, Git repos must store not just diffs but agent session logs—prompts, tool calls, checkpoints. He calls this a semantic memory layer that helps agents avoid repeated mistakes, saves tokens, and lets humans review agent-built code faster. He also pushes for Git hosting to finally deliver on its decentralized promise: centralized platforms will hit rate limits under agent-scale load, so repos should be mirrored across many hosts for resilience and data sovereignty. The post is a vision piece; no implementation details or benchmarks are provided.
#Code#Entire#Thomas Dohmke
editor take
Ex-GitHub CEO argues Git repos should store agent session logs as semantic memory, but it's a vision piece with no implementation details.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
12:10
19d ago
MIT Technology Review· rssEN12:10 · 07·09
China plans to let Alibaba, ByteDance, DeepSeek buy Nvidia H200 chips
China plans to let top AI firms Alibaba, ByteDance, and DeepSeek buy Nvidia H200 chips. The US had authorized the sale, but China withheld approval until now. The post doesn't disclose quantities or timeline.
#Alibaba#ByteDance#DeepSeek
editor take
China finally lets Alibaba, ByteDance, DeepSeek buy H200s, but the post doesn't say how many or when.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H0·K1·R0
10:43
19d ago
Product Hunt · AI· rssEN10:43 · 07·09
StoryChief Connect: Let Claude publish and schedule content directly
StoryChief Connect plugs Claude into a full marketing workflow. Instead of just generating text, Claude can now pull data from HubSpot, Notion, Google Drive, etc., research, plan, create multi-channel content, manage approvals, schedule publishing, and learn from performance. The founder says the shift is 'from giving text to participating in the workflow.' Free to use, but the post doesn't specify which Claude models are supported or if there are usage limits.
#StoryChief#Claude#HubSpot
editor take
StoryChief plugs Claude into HubSpot, Notion, and more so it can research, plan, publish, and review—not just generate text.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
10:00
19d ago
● P1OpenAI Blog· rssEN10:00 · 07·09
OpenAI releases GPT-5.6 family with flagship Sol outperforming Claude Fable 5
OpenAI released the GPT-5.6 family: flagship Sol, balanced Terra, and low-cost Luna. Sol scores 53.6 on Agents' Last Exam, 13.1 points above Claude Fable 5 at roughly one-quarter the estimated cost. On the Coding Agent Index, Sol hits 80, 2.8 points ahead of Fable 5 while using less than half the output tokens, taking under half the time, and costing about one-third less. A new ultra mode coordinates four agents in parallel by default, trading higher token usage for faster results on demanding tasks. The post does not disclose exact pricing or regional availability.
#Code#OpenAI#Anthropic#Claude Fable 5
why featured
Featured · importance 100 · hook + knowledge + resonance
editor take
GPT-5.6 drops Thursday with Sol, Terra, and Luna. Commerce Dept approval is the real signal here, but both sources are repackaging the same Axios report — no spec sheet from OpenAI yet.
sharp
OpenAI is pushing GPT-5.6 live this Thursday with three variants — Sol, Terra, and Luna. Two sources are covering this, but honestly they're both running off the same Axios scoop. IT之家 added some background for Chinese readers; the HN post is just a headline. So the multi-source coverage here doesn't mean independent confirmation — it's one exclusive getting amplified fast. The part that makes this real is the Commerce Department sign-off. GPT-5.6 was stuck in a phased rollout, only available to government-approved entities, and OpenAI made it clear they weren't happy about it. Now the restriction is lifted after testing by the department's AI standards center, with OpenAI engineers camped out in DC to answer questions. Same playbook as Anthropic's Mythos and Fable models — the government is turning advanced model release reviews into a routine process. What I'd hold back on: nobody has specs, pricing, or benchmarks for Sol, Terra, or Luna. We know the launch date, but we don't know how these three are positioned against each other. If you need to make a call today, wait for OpenAI's own technical blog on Thursday. Don't treat the Axios exclusive as the official announcement.
HKR breakdown
hook knowledge resonance
open source
100
SCORE
H1·K1·R1
10:00
19d ago
● P1OpenAI Blog· rssEN10:00 · 07·09
OpenAI launches ChatGPT Work agent that operates across apps and runs for hours
OpenAI introduced ChatGPT Work on July 9, an agent that operates across apps and files, powered by the new GPT‑5.6 model. It breaks goals into multi-step tasks and independently produces sheets, slides, docs, or web apps, and can run scheduled jobs in the background. Zapier used it to trace customer touchpoints across CRM and email, generating a weekly exec dashboard that surfaced seven-figure potential sales. RingCentral turned manual launch checks into a repeatable workflow, letting one person support roughly 50 PMs. The post does not disclose pricing details; the desktop app is available now and enterprise sales inquiries are open.
#Agent#OpenAI#ChatGPT Work#GPT-5.6
why featured
Featured · importance 100 · hook + knowledge + resonance
editor take
OpenAI pushed its desktop agent from dev tool to all-office play, bundling GPT-5.6 and Codex — but pricing and real cross-app compatibility are still missing.
sharp
OpenAI dropped ChatGPT Work today — a desktop agent that can operate across your apps and files. Three sources covered it, but they're all working off the same official blog post. One headline plays up the "partner for ambitious work" angle, another calls out the Codex + GPT-5.6 integration, and HN just has the bare product name with no discussion yet. That pattern tells me this is a clean PR push, not a leak or independent scoop. I'd hold off on the full picture until we see pricing and compatibility details. The blog says it works with Google Workspace, Teams, Slack, Jira, and others, but doesn't say whether it's screen-based automation or direct API connections. No error rates, no latency numbers. Pricing is enterprise-only for now — "contact sales" — so individual users are in the dark. GPT-5.6 launched alongside it, and OpenAI claims it's state-of-the-art at multi-step reasoning and template-following, but there are no benchmark tables in the post. The four customer stories from Zapier, RingCentral, Virgin Atlantic, and NVIDIA are specific and credible — real names, real workflows — but these are early testers. I'd want to see what breaks when thousands of people throw messy real-world tasks at it.
HKR breakdown
hook knowledge resonance
open source
100
SCORE
H1·K1·R1
09:47
19d ago
Hacker News Frontpage· rssEN09:47 · 07·09
Zig creator on Bun's Rust rewrite: bad code, worse management, and a relief to part ways
Andrew Kelley says Bun's Zig code was 'hacks on top of hacks,' management was a mess, and the Zig community wanted distance long ago. After taking VC, Jarred raced to ship features, ignored bugs, and demanded brutal hours from staff. When Anthropic acquired Bun, donations stopped and meetings went silent—the relationship was already over. The Zig team was relieved by the Rust rewrite announcement. Kelley also calls out the blog post for blaming language features instead of admitting they just didn't invest in bug fixing.
#Andrew Kelley#Jarred Sumner#Bun#Open source
editor take
Andrew Kelley says Bun's Zig code was hacks on hacks, management was a mess, and the Zig community was relieved by the Rust rewrite.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
09:36
19d ago
Product Hunt · AI· rssEN09:36 · 07·09
Mispher: On-device dictation, rewrite, translate, and an agent on Mac
Mispher is a local macOS app that combines dictation, rewriting, translation, and an agent. It supports multiple speech models: Parakeet EOU for English, Parakeet TDT for 25 European languages, CTC for Chinese, Nemotron for ~40 languages, and Qwen3-ASR for ASR. The agent uses a local LFM2.5 model to plan and call tools like Apple Notes, clipboard, files, and MCP. Users can remap the dial, hotkeys, choose from three HUD styles, set tool approvals, and swap models or prompts. Upcoming support includes Gemma 4, Qwen3.6, and Ornith 1.0. It runs fully offline, no cloud or account needed, free and MIT-licensed, only for Apple Silicon and macOS 26. The post does not disclose latency, accuracy, or the full list of supported MCP tools.
#Mispher#Parakeet#Nemotron#Open source
editor take
Mispher packs dictation, rewrite, translation, and an agent into one free local app—but the post doesn't disclose latency or accuracy.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
09:00
19d ago
● P1最佳拍档 (BestPartners)· atomZH09:00 · 07·09
Lilian Weng argues harness engineering is key to AI self-improvement over model design
The post does not disclose details. The title says AI self-improvement via recursion starts with harness engineering, and Lilian Weng's latest long-form post covers feedback loops and three design patterns: ACE, MCE, Meta-Harness. Core intelligence and STOP are key terms, but specifics require watching the video.
#Lilian Weng
why featured
Featured · importance 88 · hook
editor take
Lilian Weng's survey of 35 papers shifts the RSI conversation from model weights to engineering harnesses. Both sources agree because they're reading the same original blog post — the signal is solid.
sharp
Lilian Weng dropped a long survey covering 35 papers on recursive self-improvement, and her core argument is blunt: the future of AI self-improvement isn't about models rewriting their own weights — it's about harness engineering. That means the scaffolding, feedback loops, goal specification, and context management wrapped around the model. Both sources covering this (Latent Space and BestPartners) are reading the same original blog post, so the agreement is real but narrow — no independent reporting or new facts beyond what Weng published. She breaks out three design patterns and highlights two papers in particular: ACE and Meta-Harness. The Meta-Harness thread is the wild one — using AI to automatically optimize the harness that optimizes AI. Latent Space also notes this probably hints at what Thinky, her new startup, is building. I'd read this as a research roadmap, not a product signal. No pricing, no benchmarks, no Thinky product details yet. If you're building agent products or long-running task systems, the paper list here is worth working through.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K0·R0
08:47
19d ago
AI HOT (Curated Pool)· aihot-apiZH08:47 · 07·09
NVIDIA Releases Nemotron-Labs-3-Puzzle-75B-A9B: A Compressed Hybrid MoE LLM Delivering 2.03x Server Throughput at Matched User Throughput
NVIDIA's new Nemotron-Labs-3-Puzzle-75B-A9B is a compressed hybrid MoE model. It packs 75B total parameters but activates only 9B per inference, delivering 2.03x server throughput at matched user throughput. The post doesn't disclose training data, benchmarks, or open-source plans—only the throughput figure.
#Inference-opt#NVIDIA
editor take
NVIDIA's new MoE model activates only 9B of 75B params, claiming 2x server throughput at matched user throughput.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
08:05
19d ago
● P1Hacker News Frontpage· rssEN08:05 · 07·09
Colibrì pure C engine runs 744B parameter GLM 5.2 on laptop
Colibrì is a single-file C engine that runs the 744B-parameter GLM 5.2 model on a 32 GB laptop with no GPU. The dense part stays in RAM at int4 (~9.9 GB), while 21,504 routed experts stream from disk on demand with an LRU cache. Cold-start speed is 0.1 tok/s. The author built and tested it on a 12-core, 25 GB machine. The post does not include quality benchmarks against the full-precision model.
#Inference-opt#GLM 5.2#Colibrì#JustVugg
why featured
Featured · importance 88 · hook + knowledge + resonance
editor take
A pure-C engine claims to run GLM 5.2's 744B MoE model on a 25GB laptop, but so far it's just a Show HN post and a GitHub README — no independent reproduction reports yet.
sharp
This popped up on both HN and an AI news aggregator today, both pointing to the same GitHub repo: JustVugg/colibri. The author claims a pure-C inference engine with zero dependencies that runs GLM 5.2 — a 744B-parameter MoE model — on a consumer machine with 25GB RAM. The trick: expert weights stay on disk and get streamed in on demand. I'd take this with a grain of salt for now. GLM 5.2, released by Zhipu in June 2026, is a mixture-of-experts model — 744B total params but only a fraction activate per token, so the actual compute load is way smaller than the headline number suggests. Colibrì exploits that sparsity aggressively, treating disk as swap space for expert weights. The idea isn't new — llama.cpp and Ollama have been doing similar things — but a pure-C, zero-dep build tuned specifically for GLM's MoE architecture could be genuinely lighter than general-purpose alternatives. What's missing matters: no token-per-second numbers, no quantization details, and zero independent reproduction reports. I haven't seen anyone in the HN thread post their own benchmarks yet. Treat this as an interesting engineering demo, not proof that a 744B model runs locally in any practical sense.
HKR breakdown
hook knowledge resonance
open source
88
SCORE
H1·K1·R1
06:50
19d ago
Product Hunt · AI· rssEN06:50 · 07·09
ChatCut: An AI video editor that lives inside ChatGPT
ChatCut is a lightweight AI video editor that works inside ChatGPT, on desktop, and on the web. It combines trimming, captions, music, B-roll, motion graphics, and AI-generated video on an editable timeline, with XML export for other tools. Launched today on Product Hunt, ranked #2 of the day with 529 upvotes. The team claims 'no editing experience needed,' but the post doesn't disclose pricing, max video length, or model details.
#ChatCut#Product Hunt
editor take
ChatCut runs a full video editor inside ChatGPT—trim, captions, music, B-roll. #2 on Product Hunt today, but no pricing or max video length disclosed.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H1·K0·R0
05:46
19d ago
Hacker News Frontpage· rssEN05:46 · 07·09
AI changes the economics of software rewrites
The author argues AI shifts the economics of rewrites: codebases with clear, common patterns get higher-quality AI output because models have seen millions of examples. Proprietary languages and inconsistent legacy systems force the model to spend context tokens learning your stack first, raising cost and variance. A rewrite is a chance to rebuild around patterns that play to AI's strengths—otherwise competitors gain an edge in both speed and output quality.
#Code
editor take
AI code quality depends on your codebase. Popular stacks get leverage; proprietary or messy ones burn context tokens and produce worse output.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0

more

feeds

admin