ax@ax-radar:~/curated $ grep -l 'curated=true' sources/
33 srcsignal 72%cycle 04:32

ax curated

50 items · updated 3m ago
2026-09-22 · Tue
23:46
2h ago
● P1AI HOT (Curated Pool)· aihot-apiZH23:46 · 09·22
Claude Opus 5.5 and GPT-6 Sol/Luna launch on the same day, kicking off a new price war
Simon Willison compares three models launched on the same day. GPT-6 Luna drops to $0.10/M input tokens—half the price of GPT-5.6 Luna and one of OpenAI's cheapest models ever. GPT-6 Sol also halves its predecessor's price. Claude Opus 5.5 gets a 20% cut but still costs twice as much as GPT-6 Sol. In testing, Opus 5.5 at max thinking level over-thinks to the point of hitting its 128k output limit, failing to produce even a simple pelican SVG. Each failed attempt cost $2.56 and took nearly 20 minutes. Willison calls the max mode effectively useless.
#Reasoning#Code#Anthropic#OpenAI
why featured
Featured · importance 92 · hook + knowledge + resonance
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
18:25
7h ago
STILL DEVELOPING · 1d● P1AI HOT (Curated Pool)· aihot-apiZH18:25 · 09·22
OpenAI releases GPT-6 Sol and GPT-6 Luna models with API pricing fifty percent below promo rates
OpenAI released GPT-6 Sol and GPT-6 Luna, both built on GPT-6 Astra tech and aimed at cheaper, faster high-volume workloads. API pricing is 50% lower than GPT-5.6 promotional pricing, driven by more efficient caching and inference. Sam Altman reposted the announcement and called the character designs cute. The post doesn't disclose benchmark scores, latency figures, or regional availability.
#OpenAI#Sam Altman
why featured
Featured · importance 100 · hook + knowledge + resonance
editor take
GPT-6 ships as two models with API pricing at half of GPT-5.6's promo rate — this isn't a tweak, it's a repricing.
sharp
OpenAI dropped GPT-6 in two flavors: Sol and Luna. All three sources agree on the headline number — API pricing at 50% below GPT-5.6's already-discounted promo rate — and both models are live on Arena for testing. I'd hold off on the full picture though: we're working off titles and summaries, no official blog post yet, no context window specs, no benchmark scores, and no breakdown of what separates Sol from Luna. The pricing move is the real signal here. GPT-5.6's promo rate was already a cut, and halving it again for a new generation means OpenAI is forcing competitors to match or lose on cost. The dual-model naming suggests a heavy/lite split — think GPT-4 vs GPT-4-mini — but I can't confirm that without the announcement. What I'm waiting for: Arena scores to show actual capability gaps between Sol and Luna, and whether the listed price includes volume discounts or is the raw per-token rate.
HKR breakdown
hook knowledge resonance
open source
100
SCORE
H1·K1·R1
18:00
8h ago
● P1AI HOT (Curated Pool)· aihot-apiZH18:00 · 09·22
OpenAI launches GPT-6 Sol and Luna, API pricing cut 50% vs GPT-5.6
OpenAI added two cheaper models to the GPT-6 family: Sol and Luna, with API prices halved across input and output. Sol costs $2/$10 per 1M tokens, Luna $0.10/$0.50. Sol scored 33.2% on AutomationBench at xhigh effort at 9% of Claude Opus 5's cost per task, and 56.4% on Agents' Last Exam at max effort at 60% lower cost. On internal factuality evals, Sol makes about half as many mistakes as its predecessor. The post does not specify a launch date beyond 'available now.'
#Code#Agent#OpenAI#GPT-6 Sol
why featured
Featured · importance 97 · hook + knowledge + resonance
editor take
GPT-6 Sol and Luna halve API prices; Sol runs AutomationBench at 9% of Claude Opus 5's cost per task.
sharp
The reason to click: OpenAI filled out the GPT-6 family with two cheaper models, and API prices are literally cut in half. Sol costs $2/$10 per 1M tokens, Luna $0.10/$0.50. Sol scored 33.2% on AutomationBench at xhigh effort at 9% of Claude Opus 5's cost per task, and 56.4% on Agents' Last Exam at 60% lower cost than its predecessor. On internal factuality evals, Sol makes about half as many mistakes as the previous generation. I'd discount the benchmarks a bit—AutomationBench is Zapier's cross-app workflow test, not a universal agent metric. But the cost drop is real. If you're already running GPT-5.6 Sol for batch tasks, switching saves you half the bill. Luna's pricing is approaching near-free tier territory, good for high-throughput, latency-tolerant workloads. The post doesn't disclose parameter counts or inference latency for either model, which is a notable gap.
HKR breakdown
hook knowledge resonance
open source
97
SCORE
H1·K1·R1
14:38
11h ago
AI HOT (Curated Pool)· aihot-apiZH14:38 · 09·22
LiteParse September update: PDFium 20-25% faster, plus visual grounding and is-complex routing
LiteParse shipped four updates. First, a fork of PDFium with surgical optimizations cuts text extraction time by 20-25%. With OCR off, it averages 2.8ms/page for text and 3.9ms/page for full markdown rendering—the fastest open parser they've tested. Second, markdown heuristics accuracy improved, though the post doesn't share specific metrics. Third, visual grounding now maps parsed elements back to PDF page coordinates. Fourth, a new is-complex API lets callers route documents by complexity before choosing a parsing pipeline. LiteParse currently sees 300k+ weekly downloads and 12k+ GitHub stars.
#LiteParse#LlamaIndex#PDFium
editor take
LiteParse ships PDFium fork with 20-25% faster text extraction, visual grounding, and a complexity-based routing API.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
14:35
11h ago
AI HOT (Curated Pool)· aihot-apiZH14:35 · 09·22
Meta's AI assistant Muse has a serious 0-day that lets local apps steal account tokens
A serious 0-day in Meta's AI assistant Muse allows attackers to fully hijack the agent and steal account tokens via a ClickFix attack. CEO Zuckerberg had touted Muse as 'built from the ground up for privacy and security.' The post does not disclose whether the vulnerability has been patched or the scope of affected users.
#Meta#Mark Zuckerberg#Muse
editor take
Meta's Muse AI assistant has a serious 0-day that lets attackers steal account tokens—right after Zuckerberg touted its privacy and security.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H1·K0·R1
12:59
13h ago
AI HOT (Curated Pool)· aihot-apiZH12:59 · 09·22
New Mac mini and Mac Studio are available today
Apple today launched the new Mac mini and Mac Studio. The Mac mini offers M6 or M5 Pro chips, while the Mac Studio comes with M5 Max or M5 Ultra. The post does not disclose performance benchmarks, pricing, or shipping timelines.
#Apple
editor take
Apple announced M6 Mac mini and M5 Ultra Mac Studio are shipping, but no benchmarks or prices yet.
HKR breakdown
hook knowledge resonance
open source
15
SCORE
H0·K0·R0
11:00
15h ago
AI HOT (Curated Pool)· aihot-apiZH11:00 · 09·22
Kimi launches browser extension that fills forms and replays recorded tasks
Kimi renamed its WebBridge to a browser extension that lives in the sidebar. It can navigate pages, fill forms, and record a task sequence as a reusable skill. Available on Chrome Web Store and kimi.com. The post doesn't specify browser support, pricing, or skill complexity limits.
#Kimi#Moonshot AI
editor take
Kimi renamed WebBridge to a sidebar browser extension that records steps as reusable skills. No word on browser support or pricing yet — useful for simple form fills, but I'd wait on complex workfl...
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
02:02
1d ago
AI HOT (Curated Pool)· aihot-apiZH02:02 · 09·22
Kazike tests Grok 4.7 vs Xiaomi MiMo V2.6: the latter is the answer to the impossible triangle
The body does not disclose any test details. The title says Kazike compared Grok 4.7 with Xiaomi MiMo V2.6 and concluded that MiMo V2.6 is the answer to the 'impossible triangle'. However, the article was blocked by WeChat, showing only an environment anomaly and verification page, with no model parameters, test methodology, or specific results.
#Grok#Xiaomi#MiMo#Benchmark
editor take
WeChat blocked the article body. Only the title claims MiMo V2.6 solves the impossible triangle — no test details, so take it with a grain of salt.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
01:49
1d ago
AI HOT (Curated Pool)· aihot-apiZH01:49 · 09·22
Step 5 Preview scored 44 on Intelligence Index at roughly 1/2.8 the cost of peers
Artificial Analysis rated Step 5 Preview at 44 on its Intelligence Index, tying Kimi K3 (max) and trailing GLM-5.3 (max) and Qwen3.8 Max by 1 point. Cost per task is ~$0.72 vs. ~$2.00 for peers, roughly 1/2.8 the price. The post doesn't disclose evaluation dimensions, latency, or context window.
#阶跃星辰#Step 5 Preview#Artificial Analysis
editor take
Step 5 Preview ties Kimi K3 on the IQ index at $0.72 per task vs. ~$2 for peers. The post doesn't disclose what's tested, latency, or context window, so I'd discount it for now.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
00:00
1d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 09·22
OpenRouter launches Batch API with 50% off for bundled inference
OpenRouter's new Batch API lets you bundle requests so providers can process them within a 24-hour window, cutting per-token price by 50% or more. Across 230k+ batches during a two-week beta, the median finished in 7 minutes and 90% within an hour. Submission time matters more than batch size: batches sent 5am–noon Pacific are slowest, with the worst tenth taking 2–4.5 hours; after 6pm Pacific, 90% finish under 50 minutes. Over 70 models are supported for chat completions, messages, and embeddings—good for labeling, back-filling vectors, eval scoring, or summarizing ticket backlogs.
#OpenRouter
editor take
OpenRouter's Batch API halves inference price; across 230k batches the median was 7 min—good for labeling, back-filling, or overnight eval runs.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
2026-09-21 · Mon
20:18
1d ago
AI HOT (Curated Pool)· aihot-apiZH20:18 · 09·21
Xiaomi open-sources MiMo-V2.6, using RL to let models teach themselves
Xiaomi released and open-sourced the MiMo-V2.6 series, focusing on scaling reinforcement learning for self-improvement. The post is blocked by WeChat and does not disclose specific parameters, performance, or repo links.
#Xiaomi
editor take
Xiaomi open-sourced MiMo-V2.6, claiming RL-driven self-improvement, but the post is blocked by WeChat with no params, benchmarks, or repo link — I'd hold off.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
17:06
1d ago
AI HOT (Curated Pool)· aihot-apiZH17:06 · 09·21
Musk says Grok 4.7 puts xAI third in agentic coding
Elon Musk cites Artificial Analysis to claim Grok 4.7 ranks xAI third in agentic coding, behind only Anthropic and OpenAI. The post doesn't disclose the benchmark's metrics, scores, or version comparisons—only the ranking and competitors.
#Code#Agent#xAI#Elon Musk
editor take
Musk cites Artificial Analysis to claim Grok 4.7 ranks third in agentic coding, but the post doesn't disclose metrics or scores—I'd discount this.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R1
07:54
1d ago
AI HOT (Curated Pool)· aihot-apiZH07:54 · 09·21
Kimi launches Code Desktop 1.0 for macOS and Windows simultaneously
Kimi released Code Desktop 1.0, available on both macOS and Windows. The post does not disclose specific features, pricing, or technical details—only the launch itself is confirmed.
#Kimi
editor take
Kimi launched Code Desktop 1.0 for Mac and Windows, but the post is behind a CAPTCHA—no features or pricing disclosed.
HKR breakdown
hook knowledge resonance
open source
30
SCORE
H0·K0·R0
02:51
1d ago
AI HOT (Curated Pool)· aihot-apiZH02:51 · 09·21
Tencent Hunyuan Releases Hy Image3.5 Preview Image Generation Model
Tencent Hunyuan released Hy Image3.5 preview, an image generation model. The body only contains the title and navigation bar, with no details on capabilities, parameters, or release timeline. Wait for official disclosure.
#Vision#Tencent#Hunyuan
editor take
Tencent Hunyuan dropped Hy Image3.5 preview, but the page is just a title and nav bar — no specs, no timeline. Wait for the real post.
HKR breakdown
hook knowledge resonance
open source
25
SCORE
H0·K0·R0
01:52
2d ago
AI HOT (Curated Pool)· aihot-apiZH01:52 · 09·21
Interview with Fansub Groups: The Real Situation of Subtitle and Manga Teams in the AI Era
The article body is blocked by WeChat, showing only a CAPTCHA page. The title indicates an interview about how AI is affecting the real situation of fansub and manga translation groups. No details are available, but the topic is relevant: AI translation tools are disrupting traditional volunteer-based localization teams.
#数字生命卡兹克
editor take
WeChat blocked the article body; only the title about AI's impact on fansub groups is visible, no details.
HKR breakdown
hook knowledge resonance
open source
15
SCORE
H0·K0·R0
2026-09-20 · Sun
02:00
3d ago
AI HOT (Curated Pool)· aihot-apiZH02:00 · 09·20
StepFun Releases Step 5 Preview, Open-Source Weights on Oct 15
StepFun released its flagship model Step 5 Preview and plans to open-source its weights on October 15. The page encountered an error, so no details on model specs, performance, or open-source scope are available.
#阶跃星辰#StepFun#Open source
editor take
Step 5 Preview open-sources weights Oct 15, but the page is blocked — no specs, no benchmarks, no scope yet.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
2026-09-19 · Sat
02:48
3d ago
AI HOT (Curated Pool)· aihot-apiZH02:48 · 09·19
Alexandr Wang shares Muse prompt: AI agent auto-adds travel time to calendar
Alexandr Wang shared a Muse prompt that scans the next 14 days of calendar and auto-adds travel time blocks for off-site meetings. The prompt also works with Instinct and Grok @bot. The post does not disclose the prompt format or usage limits.
#Alexandr Wang#Muse#Instinct
editor take
Alexandr Wang shared a Muse prompt that auto-adds travel buffers for off-site meetings—but the post doesn't include the prompt format.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
2026-09-18 · Fri
04:45
4d ago
AI HOT (Curated Pool)· aihot-apiZH04:45 · 09·18
Hacktron chains libheif overflow and OpenAI SSO flaw to compromise employee accounts
On July 25, 2026, Hacktron chained two critical vulnerabilities to breach OpenAI's internal repos. They exploited a heap buffer overflow in the libheif image decoder to get RCE on community.openai.com, then abused an OpenAI SSO identity flaw to take over multiple employees' ChatGPT and Codex accounts. They opened a harmless PR in OpenAI's internal monorepo as proof. The entire chain took under 72 hours and earned a $6,500 bounty. The post does not spell out the SSO flaw's technical details.
#OpenAI#Hacktron#Discourse
editor take
Hacktron published a detailed write-up on chaining a libheif heap overflow with an OpenAI forum SSO flaw to take over employee ChatGPT accounts and access internal repos. Two outlets are covering i...
HKR breakdown
hook knowledge resonance
open source
49
SCORE
H0·K0·R0
00:00
5d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 09·18
xAI launches Grok Voice Transcribe 2.0, doubling accuracy at the same price
xAI released Grok Voice Transcribe 2.0 on Sep 18, claiming it's one of the most accurate speech-to-text models in real-world evals and twice as accurate as v1.0. Pricing stays at $0.10/hr for batch and $0.20/hr for streaming, with diarization, timestamps, and key terms included. It handles hard cases like noisy phone calls and short multilingual commands—word error rate on short phrases dropped from 20.6% to 6.8%. Atlassian Loom already swapped it in and pipes transcripts into Cursor for code updates. The post doesn't disclose parameter count or training details.
#xAI#SpaceXAI#Grok
editor take
2x accuracy at same price, short-phrase WER dropped from 20.6% to 6.8%, but no param count or training details disclosed.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
2026-09-17 · Thu
23:36
5d ago
AI HOT (Curated Pool)· aihot-apiZH23:36 · 09·17
ChatGPT lands in Word; OpenAI says Excel and PowerPoint usage has surged recently
ChatGPT is now built into Word: it can turn rough notes into a draft, rephrase paragraphs, proofread, suggest edits, and catch formatting issues. OpenAI's Sherwin Wu says Excel and PowerPoint usage has spiked recently, and adding Word completes the Office suite integration. The post doesn't disclose launch date, pricing, or feature limits.
#OpenAI#Microsoft#Sherwin Wu
editor take
ChatGPT now lives inside Word, completing the Office suite—but no launch date or pricing yet.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
15:38
5d ago
● P1AI HOT (Curated Pool)· aihot-apiZH15:38 · 09·17
Noam Brown on 10,000-agent swarms solving math problems and recursive self-improvement
Noam Brown, a core contributor to OpenAI's o1 reasoning models, now works on multi-agent systems. His team just solved a Millennium Prize Problem using 10,000 agents, 130 billion tokens, and 88 hours of compute. Brown frames multi-agent as parallel test-time compute: a single agent hits a latency wall, so you throw more agents at the problem to go faster, at the cost of some efficiency. In the 5.6 release's Ultra Mode, 4 agents cut solve time in half; 16 agents push it further, especially on parallel-friendly tasks like math. The conversation also covers what math progress signals for recursive self-improvement, degrading chain-of-thought quality, and how to verify alignment before kicking off RSI.
#Reasoning#Agent#Noam Brown#OpenAI
why featured
Featured · importance 98 · hook + knowledge + resonance
editor take
Noam Brown reveals a 10,000-agent system for math, but the model isn't public and details are all from a podcast — treat this as a directional signal, not a product launch.
sharp
Two sources covered this, but both trace back to a single Dwarkesh podcast episode — no blog post, no paper, no public demo. Noam Brown says OpenAI used 10,000 agents running for 88 hours and burning 130 billion tokens to solve a Millennium Prize math problem. All numbers come from his spoken remarks, so there's no way to cross-check. The logic he lays out: reasoning models get better the longer they think, but serial latency becomes unbearable. Parallelizing across many agents trades some efficiency for speed, and math problems happen to be highly parallelizable. The idea isn't new, but the scale is — this is the first time anyone from a major lab has talked about running 10,000 agents on a single hard problem. I'd discount this a bit for now. No pricing was mentioned, and it's unclear whether 88 hours is wall-clock time or GPU time. He didn't specify which Millennium Problem was solved or what verification looked like. What's solid: OpenAI is betting heavily on multi-agent as the next scaling axis. What's missing: any signal on when this becomes a product rather than a research flex.
HKR breakdown
hook knowledge resonance
open source
98
SCORE
H1·K1·R1
15:03
5d ago
AI HOT (Curated Pool)· aihot-apiZH15:03 · 09·17
Unsloth ships Docker image and desktop app to train & run 500+ models locally
Unsloth released a Docker image and Unsloth Desktop to train and run 500+ models locally with zero setup. It includes a new GUI and notebook workflows, supporting both NVIDIA and AMD GPUs. The post doesn't disclose specific performance numbers or the full model list, but the install guide is live.
#Unsloth
editor take
Unsloth shipped a Docker image and desktop app for zero-setup local training of 500+ models, but no performance numbers yet — I'd hold off on the hype.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K0·R1

more

feeds

admin