ax@ax-radar:~/curated $ grep -l 'curated=true' sources/
33 srcsignal 72%cycle 04:32

ax curated

38 items · updated 3m ago
2026-09-02 · Wed
00:30
21d ago
AI HOT (Curated Pool)· aihot-apiZH00:30 · 09·02
Anthropic releases Claude Fable 5.1 and Mythos 5.1, with hands-on tips from a tester
Anthropic dropped two new models, pitched as its most capable for coding and knowledge work. Tester Thariq says they're solid and a full review is coming. Two practical notes: use low effort for tasks that need less verification or have fewer edge cases, and switching effort no longer breaks the prompt cache.
#Code#Anthropic#Thariq
editor take
Thariq's hands-on with Claude Fable/Mythos 5.1: use low effort for low-verification tasks to save compute, and switching effort no longer breaks the prompt cache.
HKR breakdown
hook knowledge resonance
open source
72
SCORE
H1·K1·R0
2026-09-01 · Tue
2026-08-31 · Mon
17:03
22d ago
AI HOT (Curated Pool)· aihot-apiZH17:03 · 08·31
Runway introduces Solaris, a world model that generates OS-level interfaces in real time
Runway unveiled Solaris, a world model that generates full OS interfaces in real time from text prompts. It outputs interactive desktops, windows, and controls—not static mockups, but a live, navigable interface. Runway calls it an 'interface world model.' So far there's only a demo video and a blog post; no technical details, model parameters, or public access have been shared. I'd hold off on full excitement: the video looks smooth, but the post doesn't clarify whether it's real-time inference or pre-rendered, nor does it mention latency, supported apps, or hardware requirements.
#Runway
editor take
Runway's Solaris generates interactive OS interfaces from text prompts, but it's just a demo video with no public access or technical details yet.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K0·R1
00:36
23d ago
AI HOT (Curated Pool)· aihot-apiZH00:36 · 08·31
A 24/7 AI live streaming site built on MiniMax H3 Max is now live
Someone built a 24/7 AI live streaming site using MiniMax H3 Max. The post is blocked by WeChat, so no details on content, cost, or latency are available.
#MiniMax#H3 Max
editor take
WeChat blocked the post. Title says someone built a 24/7 AI live stream with MiniMax H3 Max — no cost, latency, or content details. I'd hold off.
HKR breakdown
hook knowledge resonance
open source
25
SCORE
H0·K0·R0
2026-08-30 · Sun
2026-08-29 · Sat
04:31
24d ago
● P1AI HOT (Curated Pool)· aihot-apiZH04:31 · 08·29
Zhipu open-sources GLM-5.3 model weights for agentic coding and cybersecurity
Zhipu released GLM-5.3 weights for local deployment and commercial use. It scores 60 on the AA Intelligence Index, matching closed-source flagships like Claude Fable 5 and GPT-5.6 Sol, and ties with Kimi K3 for top open-source model. The model excels at complex coding, cybersecurity, and long-horizon tasks. Zhipu added two extra weeks of safety review before release due to its advanced cyber capabilities. Organizations with over $10B annual revenue need a security audit before offering it as an external model service.
#Agent#Code#Zhipu#GLM-5.3
why featured
Featured · importance 96 · hook + knowledge + resonance
editor take
Zhipu released GLM-5.3 weights, targeting coding and cybersecurity. The commercial license only triggers a security review for companies with over $10B annual revenue — small teams can just use it.
sharp
Zhipu open-sourced GLM-5.3 weights last night, available on HuggingFace and ModelScope. Two tech outlets covered this, both pulling from the same IT Home report — so we're looking at a single official source, not independent verification. The positioning is clear: complex coding, defensive cybersecurity, and long-horizon tasks. Zhipu's Z.ai lead Li Zixuan confirmed local deployment, fine-tuning, and commercial use are all allowed. The only catch: companies with over $10B annual revenue need a security review before offering GLM-5.3 as an external model service. That threshold is deliberately high — it basically targets Google, Microsoft, and Amazon-tier cloud providers while leaving startups untouched. On performance, Zhipu cites the Artificial Analysis Intelligence Index, where GLM-5.3 scored 60 — same tier as Claude Fable 5 and GPT-5.6 Sol, tied with Kimi K3 for top open-source model. I'd take this with a grain of salt though: AA is a composite score, and we don't have breakdowns showing how much better it actually is on coding or security specifically versus alternatives. The detail I find most credible: Zhipu ran an extra two weeks of safety evaluation before release because of the cybersecurity capabilities. That's a concrete acknowledgment of what this model can do and a real effort to prevent misuse. What's missing: parameter count, inference cost comparisons, and actual SWE-bench or coding benchmark splits. Wait for the community to run those before drawing conclusions.
HKR breakdown
hook knowledge resonance
open source
96
SCORE
H1·K1·R1
2026-08-28 · Fri
00:00
26d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 08·28
Open ASR Leaderboard Adds Hindi and Indian English Benchmarks
Voice Arena and Hugging Face added two new benchmarks to the Open ASR Leaderboard: Monsoon en-IN (Indian English) and Monsoon hi-IN (Hindi). Hindi is the first non-European language in the multilingual section. The dataset includes public and private splits, 4,888 speakers, and 12 speaker attributes to expose uneven ASR errors across region, age, and gender. The post doesn't disclose benchmark size or model performance.
#Benchmarking#Hugging Face#Voice Arena#Benchmark
editor take
Hindi is the first non-European language on the Open ASR Leaderboard. The dataset logs 12 speaker attributes to expose where ASR fails unevenly.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
2026-08-27 · Thu
23:32
26d ago
AI HOT (Curated Pool)· aihot-apiZH23:32 · 08·27
Midjourney Opens Testing for V8.2 Image Edit Model
Midjourney is rolling out its first V8.2 image edit model for public testing. It supports instruction-based editing, generating images from up to 4 references, inpainting, outpainting, and personalization with moodboards. The website and alpha UI have been updated, but the team warns of many edge cases and asks users to report bugs. The post doesn't specify what editing improvements V8.2 brings over V8 or V8.1, nor does it disclose model parameters or inference speed.
#Vision#Midjourney
editor take
Midjourney is testing its V8.2 edit model with text-based editing, up to 4 image references, inpainting, and outpainting, but the team warns of many edge cases and doesn't specify what's improved o...
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H1·K1·R0
20:52
26d ago
AI HOT (Curated Pool)· aihot-apiZH20:52 · 08·27
Lawsuit alleges xAI trained Grok on child sexual abuse material
A new lawsuit accuses Elon Musk's xAI of using real and AI-generated child sexual abuse material to train its Grok models. The complaint alleges such content was included in training data, potentially causing the model to generate or propagate harmful outputs. The post does not disclose specific dataset sources, model versions, or training timelines.
#xAI#Elon Musk#Grok
editor take
Lawsuit claims xAI used child sexual abuse material to train Grok, but the post doesn't name the dataset or model version.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
16:11
26d ago
AI HOT (Curated Pool)· aihot-apiZH16:11 · 08·27
Gemini Omni 1.1 Flash gives developers more control over video generation
Google DeepMind released Gemini Omni 1.1 Flash, focused on giving developers finer control over generative video. The post body only contains the title and site navigation right now—it doesn't spell out which control parameters are new, whether latency improved, or how pricing works. I'd hold off until the full details land.
#Google DeepMind#Gemini Omni 1.1 Flash
editor take
The post body is just a title and nav bar—no control params, latency, or pricing yet. Hold off.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
09:05
26d ago
AI HOT (Curated Pool)· aihot-apiZH09:05 · 08·27
China's daily token calls top 500 trillion; model cycles shrink to 4–6 weeks
CCTV reports China's daily token calls exceeded 500 trillion as of June 2026. MiniMax's Feng Wen says release cycles compressed from quarterly to every 4–6 weeks. Tencent's Liu Feng notes Hunyuan 3's first-week token volume jumped 68× over Hunyuan 2. Competition is shifting from benchmark scores to agent deployment and locking in compute capacity. The post doesn't spell out the methodology or scope behind the 500 trillion figure.
#Agent#Reasoning#MiniMax#Tencent
editor take
CCTV says daily token calls hit 500T in China, but no methodology — treat as a trend signal, not a hard number.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
00:00
27d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 08·27
Claude Console adds personal keys and service account keys
Anthropic added personal keys and service account keys to the Claude Console. Personal keys act as you, service account keys act as a service account, and both stop working when the linked account leaves the organization. Admins can track usage per account and scope keys to a workspace or admin endpoints. Workspace API keys remain supported as legacy.
#Anthropic#Claude Console
editor take
Anthropic added API keys that auto-expire when a user leaves—good for team security hygiene.
HKR breakdown
hook knowledge resonance
open source
62
SCORE
H0·K1·R0
2026-08-26 · Wed
20:10
27d ago
AI HOT (Curated Pool)· aihot-apiZH20:10 · 08·26
Google unveils GlucoFM, a foundation model for continuous glucose monitoring
Google Research released GlucoFM, a foundation model built for continuous glucose monitoring (CGM) data. It learns general representations from glucose time series and can be fine-tuned for tasks like predicting fluctuations or detecting anomalies. The model is pre-trained on large-scale real-world CGM data and aims to reduce the need for patient-specific training. The post does not disclose model size, training data volume, or benchmark comparisons.
#Google Research
editor take
Google released GlucoFM, a foundation model for CGM data, but the post doesn't disclose model size, training data volume, or benchmarks.
HKR breakdown
hook knowledge resonance
open source
35
SCORE
H0·K0·R0
17:02
27d ago
AI HOT (Curated Pool)· aihot-apiZH17:02 · 08·26
Warp builds self-improving agents on Claude with a reusable pattern
Warp devised a simple development pattern on Claude that lets agents improve themselves. The post doesn't spell out the implementation details, but says anyone can use the pattern directly.
#Warp#Claude#Anthropic
editor take
Warp built a pattern on Claude that lets agents improve themselves. The post says anyone can use it but doesn't share the implementation.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K0·R0
14:00
27d ago
● P1AI HOT (Curated Pool)· aihot-apiZH14:00 · 08·26
Zhipu open-sources GLM-5.3-Flash multimodal model matching Claude Opus performance
Zhipu released and open-sourced GLM-5.3-Flash, a 320B-parameter native multimodal model with 18B active parameters. It scores 57 on the Artificial Analysis Intelligence Index, matching Anthropic Claude Opus 4.8, and delivers comparable coding performance at 1/40 the API price. The model uses a hybrid sparse-and-linear attention architecture, cutting attention compute by over 3x versus GLM-5.3 on long contexts. It can use visual feedback in coding loops to self-correct—it once ran autonomously for 16 hours to build a 400 m² kitchen scene in Blender. All public test traffic last week ran on a domestic chip cluster; the team used EPD disaggregated serving and aggressive memory optimizations to achieve 3x end-to-end speedup, bringing per-token cost on par with mainstream NVIDIA GPU setups. Weights are open on HuggingFace, with API access via ZCode and the BigModel platform.
#Code#Zhipu AI#Anthropic#Claude Opus 4.8
why featured
Featured · importance 100 · hook + knowledge + resonance
editor take
Zhipu open-sourced GLM-5.3-Flash, a 320B multimodal model scoring 57 on the AA Index—matching Claude Opus 4.8 at 1/40th the price, served entirely on domestic chips.
sharp
I'd take this with a grain of salt since both sources are repackaging Zhipu's own announcement—no third-party benchmarks yet. But the numbers are specific: 320B total params, 18B active, AA Index score of 57, directly matching Claude Opus 4.8. Pricing at 1/40th of Opus and 90% cheaper than their own GLM-5.2. The architecture choice is the interesting part. They're using a hybrid of sparse and linear attention, which cuts KV cache and compute by 3-4x for long-context workloads. The other signal: the entire service runs on a domestic chip cluster, and they claim end-to-end performance improved 3x with per-token costs now matching NVIDIA GPUs. If reproducible, that's a real milestone for domestic silicon in production inference. What's missing: independent benchmarks beyond the AA Index, real-world user feedback, and clarity on whether that 1/40 price is a limited-time discount or the standard rate.
HKR breakdown
hook knowledge resonance
open source
100
SCORE
H1·K1·R1
08:27
27d ago
AI HOT (Curated Pool)· aihot-apiZH08:27 · 08·26
First Agent After Feishu and Doubao Merge: 8 Tips
The article body is blocked by WeChat, only the title remains. It mentions the first Agent after Feishu and Doubao merged, called 'Doubao Work', with 8 usage tips. The post does not disclose what the Agent does or what the tips are.
#Feishu#Doubao
HKR breakdown
hook knowledge resonance
open source
25
SCORE
H0·K0·R0
08:04
27d ago
AI HOT (Curated Pool)· aihot-apiZH08:04 · 08·26
Tencent Hunyuan shrinks on-device translation model to 440MB, deployed for Bilibili live danmaku
Tencent Hunyuan compresses its on-device translation model Hy-MT2-1.8B to 440MB and deploys it for Bilibili live danmaku translation. The post does not disclose compression techniques, inference latency, or accuracy trade-offs.
#Tencent Hunyuan#Bilibili
editor take
Tencent Hunyuan shrinks a 1.8B translation model to 440MB for Bilibili live danmaku, but the post skips compression method and latency.
HKR breakdown
hook knowledge resonance
open source
39
SCORE
H0·K0·R0
2026-08-25 · Tue
21:12
28d ago
AI HOT (Curated Pool)· aihot-apiZH21:12 · 08·25
LangChain and Airbyte Team Up to Make Data Ingestion Production-Ready
LangChain and Airbyte launched a new integration that flips the direction: Airbyte now has a LangChain destination to pipe data directly into vector stores. The post argues production apps need scheduled re-indexing, not one-time loads, and Airbyte's orchestration fills that gap. It doesn't spell out which vector stores are supported or how to configure refresh schedules.
#LangChain#Airbyte
editor take
LangChain flips Airbyte's direction to pipe data into vector stores, solving scheduled re-indexing for production RAG.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H0·K1·R0
18:16
28d ago
AI HOT (Curated Pool)· aihot-apiZH18:16 · 08·25
Andrew Ng's OpenWorker adds built-in cybersecurity agents, fully auditable harness
OpenWorker, Andrew Ng's open-source agent project, now ships with three built-in cybersecurity agents: code vulnerability scanning, dependency supply-chain injection detection, and cloud security posture checks. Its harness is fully open-source so security teams can audit for backdoors. It also supports running open-weight models locally to keep sensitive code on-prem. The post doesn't name specific models or benchmarks.
#Andrew Ng#OpenWorker#Open source
editor take
OpenWorker ships three built-in cybersecurity agents; harness is fully open-source and local model support keeps code on-prem.
HKR breakdown
hook knowledge resonance
open source
68
SCORE
H1·K1·R0
15:32
28d ago
● P1AI HOT (Curated Pool)· aihot-apiZH15:32 · 08·25
Dylan Patel: Anthropic and OpenAI will control most global compute by 2028
SemiAnalysis founder Dylan Patel laid out the numbers: Anthropic and OpenAI took ~30% of new global compute this year, will take 40–50% next year, and could control most usable FLOPs by 2028. The driver is unit economics—Anthropic is already generating up to $50M per megawatt in inference revenue against a ~$10–15M cost, and plowing the surplus into training. Anthropic turned profitable in Q2; OpenAI is expected to follow in Q3 with Codex and GPT-5.6. Total AI infrastructure capex has passed $1T this year and is on track to exceed $2T by 2028, with SpaceX entering as a new compute builder next year. The conversation also flagged a tail risk: >$10T in cumulative AI capex by 2030 could push up interest rates and trigger a sovereign debt crisis for non-AI-exposed countries.
#Anthropic#OpenAI#SemiAnalysis
why featured
Featured · importance 92 · hook + knowledge + resonance
editor take
Dylan Patel's core claim: Anthropic and OpenAI are using inference profits to outbid everyone for compute, putting them on track to control most of the world's usable FLOPs by 2028.
sharp
Two sources picked up this podcast episode, which tells me Patel's centralization thesis is hitting a nerve. He lays out specific numbers: Anthropic and OpenAI took about 30% of new compute this year, that jumps to 40-50% next year, and by 2028 they could control most of the world's usable FLOPs. The mechanism is straightforward—Anthropic is generating $50M per megawatt in inference revenue against $10-15M in costs, and all that profit gets funneled back into training. It's a flywheel that's hard for anyone else to match. Both sources framed the story identically around centralization, which makes sense since that's the headline claim from the episode. I'd take the 2028 projection with a grain of salt though. This is a podcast conversation, not a SemiAnalysis research note—Patel is sketching a trajectory, not publishing verified forecasts. The specific market-share numbers for 2028 aren't in the transcript, so the "most of the world's compute" claim is directional rather than pinned to a concrete figure. He also touched on China getting under 10% of new compute and the possibility of AI capex triggering a sovereign debt crisis, but neither outlet led with those angles. If you're using these numbers for anything serious, wait for the written SemiAnalysis piece—podcast estimates tend to be looser than their published research.
HKR breakdown
hook knowledge resonance
open source
92
SCORE
H1·K1·R1
03:19
29d ago
AI HOT (Curated Pool)· aihot-apiZH03:19 · 08·25
Feishu and Doubao launch first joint Agent product 'Doubao Work'
Feishu and Doubao released their first joint Agent product 'Doubao Work' after merging. The article body is inaccessible due to an environment error, so no details on features, pricing, or launch timeline are available. The title confirms this is the first Agent product post-merger.
#飞书#豆包
HKR breakdown
hook knowledge resonance
open source
25
SCORE
H0·K0·R0
00:00
29d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 08·25
OpenRouter Video API: One endpoint, swap models without rewriting code
OpenRouter wraps video generation into one async API. Submit a prompt, get a job ID, poll until done, download the MP4. Models like Seedance, Veo, and Wan all use the same POST /api/v1/videos endpoint—switching models means changing one line. The post doesn't disclose pricing or generation speed, but argues the async design beats per-provider integrations or local GPU setups.
#OpenRouter#Seedance#Veo
editor take
OpenRouter wraps Seedance, Veo, and others into one async API—switch models by changing one line.
HKR breakdown
hook knowledge resonance
open source
60
SCORE
H0·K1·R0
00:00
29d ago
AI HOT (Curated Pool)· aihot-apiZH00:00 · 08·25
How to Choose the Best AI Model Live in Your Editor — OpenRouter
OpenRouter published a practical guide arguing there is no single best model, only the best for your task, budget, and latency. It offers a six-step framework: define the task, shortlist candidates using benchmarks and live usage data, compare price and latency across providers, then test finalists on your own prompts. Judge by cost per completed task, not cost per token. The OpenRouter MCP server lets you query live rankings, pricing, and test results directly from your editor. The post notes that even within coding, there are 9 sub-tasks, each led by a different model.
#Benchmarking#Inference-opt#OpenRouter#Artificial Analysis
editor take
OpenRouter's six-step model picker: define task, filter by live rankings & price, test on your prompts — all from your editor via MCP.
HKR breakdown
hook knowledge resonance
open source
55
SCORE
H0·K1·R0
2026-08-24 · Mon

more

feeds

admin