🔍 搜索中 · 显示所有标签页的匹配项 · 按 Esc 清除
📖 一句话:今天几乎所有热度都在“给 agent 用的工具”上——浏览器操作、电脑操作、技能包、沙箱,重心已从‘做模型’转到‘做 agent 的手脚和环境’。先看下面两条(jev-ultrafast 与 trycua/cua),它们代表 computer-use 的两条路线;模型榜变化不大、快讯里刷屏的 Muse 系列可快速扫过。注意:榜首的 jev 集群已核实是真实开源发布,不是刷榜——新仓库 PR 少属正常。
📌 必读 导读 · 今天先看这些
- Jev Ultrafast: a browser agent with a dynamic, indexed action space (HN 讨论) HN浏览器 agent 的新范式:读 DOM 表格、免视觉模型、70–500ms 决策;评论区有一线工程判断,比 star 数更有信息量。
- browser-use/jev-ultrafast repo直接读实现:indexed-DOM + speculative fan-out 如何把浏览器自动化压到 ~7s;MVP 局限(Shadow DOM/iframe/无结果校验)也写得很坦白。
- trycua/cua — Computer-use 2.0 repo另一条路线:跨 OS 驱动 + agent 集群 + 评测/训练数据。想理解“agent 基础设施”全景,这个和 jev 对照着看。
- hypit-ai/hypit repo把短视频还原成可编辑“代码”、批量产变体;也是“skill 即分发”(npx skills add)的样例——一个正在成形的能力分发范式。
🔬 深度洞察 deep research
Agent 基础设施成为重心:大家在造“手和沙箱”,不是又一个模型
今日多源共振里,7/9 是“让 agent 能操作/被隔离/被评测”的工具,而非模型本身。这说明生态的建设热度已从“更强的大脑”转向“可用的手脚与环境”——谁能让 agent 稳定操作真实系统,谁就卡住下一段价值。对开发者:该层仍很早、可切入;对投资人:这是当前最密集的在建方向。
browser-use/jev-ultrafasttrycua/cuaarcboxlabs/arcboxcoder/codercloudflare/security-audit-skillsapientinc/PRAXISTdeeplethe/utopia
“Skill(可安装技能包)”正成为能力分发的新范式
cloudflare/security-audit-skill 与 hypit 都以“技能包”形式分发(hypit 用 `npx skills add`),即把一项能力打包给 Claude Code/Codex 等 agent 直接调用。这是一个正在成形的分发 primitive——值得盯它会不会长出“agent 能力的应用商店”。
cloudflare/security-audit-skillhypit-ai/hypit
专用快模型分工化:浏览器 agent 成本正在坍塌
browser-use/jev-ultrafast 建立在 TypeSafe 的 Jev(“System One”)模型上——返回类型化概率决策而非文本、跳过视觉模型、70–500ms 出结果;自测把 Google Flights 自动化压到 7s、成本降约 90%(注:官方自测、未经第三方复现)。呼应 ai_news 里 GLM-5.3 FlashX 等“快变体”:趋势是工作流不同环节用不同的更快更便宜的模型,而非一个大模型通吃。
browser-use/jev-ultrafasttamaratran/fast-jev-compactionGLM-5.3 FlashX
被低估的早期信号:涨但无 X 热度
deeplethe/utopia(本地优先、agent 辅助的“文档→本体”知识工作台)和 sapientinc/PRAXIST(可执行的自主研究系统)在没有明显 X 带节奏的情况下自然上涨——通常意味着别人还没注意,alpha 在此,值得优先加入观察名单。
deeplethe/utopiasapientinc/PRAXIST
🎯 沿“agent 基础设施”主线深挖一层:对比 browser-use/jev-ultrafast(浏览器)与 trycua/cua(全 OS computer-use)两条路线,判断哪条更贴合你的用例;同时把无 X 热度却在涨的 deeplethe/utopia 加入重点观察。
📰 最新快讯
- 为什么租 1000 块 GPU 也难以复现 DeepSeek 的推理质量
- 缓存命中率:判断推理服务商真实水平的关键信号
- 复现DeepSeek级推理极难——“99%缓存命中率”或是转售信号
- 警惕“便宜”但缓存命中率差的AI推理服务商
- “99%缓存命中率”是红灯信号?警惕套壳DeepSeek的推理服务商
- GLM-5.3 FlashX 已在 Command Code 上线,吞吐量约 200 TPS
- Qwen 展示基于 Qwen 3.8 27B 与 Cerebras 的理财助手 Money Agent
- Command Code 强调支持 Qwen 3.8 Omni Flash 与快速 DeepSeek 推理
- Qwen3.8-LiveTranslate:更低延迟的实时同声传译模型
- Qwen3.8-LiveTranslate:Qwen 新一代实时同声传译模型
- Qwen3.8-LiveTranslate:新一代实时同声传译模型
- Qwen 发布 Qwen3.8-LiveTranslate 实时同声传译模型
- Command Code声称在横向对比中以197 tokens/s取得DeepSeek推理速度第一
- LiteLLM 新增对 TypeSafe AI 的 Jev 支持 —— 官方分享内部 OCR 用例
- GLM-5.3 FlashX 现已上线 Command Code
- Muse Mac 版新增即时语音听写:按住 fn 即可说话
- Muse 向开发者开放连接器平台,Granola 集成上线
- Muse AI代理现已支持连接Notion
- Muse 开放开发者自定义连接器功能
- Muse AI 智能体新增 Granola 与 Notion 连接器
- AI 代理 Muse 现已登陆加拿大
- Muse AI 智能体现已登陆加拿大
- AI智能体Muse现已登陆加拿大
- Muse AI 助手正式在加拿大上线
- Muse 称加拿大下载问题已修复
- AI智能体Muse上线iOS和Web,Android版本即将推出
- Qwen 3.8 Omni Flash 现已在 Command Code 上线
- OpenRouter 重点介绍 Jev:面向快速、低成本“是/否”与选择题的决策模型
- OpenRouter 的实用建议:把 AI 任务拆成小的是非/选择题决策
- Jev:比LLM便宜10倍、快10倍的决策模型
⭐ 多源共振
共振集中在“给 agent 用的工具”:浏览器(jev-ultrafast)、电脑操作(trycua/cua、arcbox)、技能包(security-audit-skill)、开发环境(coder/coder)、自主研究(PRAXIST)、知识工作台(utopia)。同一主线、多个团队同时发力。
githubxsocialboardhn
+683★/d 活跃开发 official #3
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
🔺 @trycua 首发 · 30h 前
Computer-use 2.0:跨 OS 驱动、agent 集群、评测与训练数据生成——“agent 基础设施”主线的核心一员,且为团队自荐(lead:trycua)。
githubxsocialboard
+1,940★/d 早期·低活动 official #8
A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings
🔺 @MaciejLukianski 首发 · 41h 前
Cloudflare 出品的 coding-agent 技能包:多阶段安全审计、独立验证、机器可读结论——“skill 即能力分发”的代表案例。
githubxsocialboard
+1,837★/d 早期·低活动 official #5
🔺 @betterhn20 首发 · 16h 前
已核实=真实开源发布,非刷榜。Browser Use 出品的浏览器 agent,基于 TypeSafe“Jev/System One”模型(类型化概率决策、免视觉、70–500ms);自测 7s 完成 Google Flights、成本≈-90%(官方自测未复现)。新仓库 PR 少属正常。
githubxsocialboard
+884★/d 活跃开发 official #2
Local-first, agent-assisted document-to-ontology workbench
本地优先、agent 辅助的文档→本体工作台;无 X 热度却涨=早期/被低估信号,优先加入观察。
githubxsocialboard
+835★/d 早期·低活动 official #16
You built it. Now brag. Turn the project you just created into a short, shareable launch video with one command.
githubxsocialboard
+743★/d 活跃开发 official #7
Claude Code plugin that replaces the compaction summary with Jev decisions: every tool call and result is scored in one fast request, stale ones are dropped or truncated, everything kept stays verbatim.
githubxsociallaunch
+658★/d 活跃开发
🔺 @cccyd_qwq 首发 · 31h 前
已核实=真实项目(~1.3k★、活跃发版至 0.2.7)。让 coding agent 把短视频还原成可编辑“代码”、批量产出变体;以 `npx skills add` 技能包分发。工具+话题双热。
🔥 动量榜
榜首多为刚开源的新仓库,PR/issue 少属正常现象(新项目),不等于刷榜。已逐一核实:jev-ultrafast、hypit 均为真实发布。
| # | Repo | 7d | +1d★ | 7d★ | 质地 | 官方 | X |
|---|---|---|---|---|---|---|---|
| 1 | cloudflare/security-audit-skill ↺ 1d JavaScript · A coding-agent skill for multi-phase security audits with in 🔺 @MaciejLukianski 首发 · 41h 前 Cloudflare 出品的 coding-agent 技能包:多阶段安全审计、独立验证、机器可读结论——“skill 即能力分发”的代表案例。 |
+1,940 | 12,241 | 早期·低活动 | #8 | 30× | |
| 2 | browser-use/jev-ultrafast ↺ 1d Python 🔺 @betterhn20 首发 · 16h 前 已核实=真实开源发布,非刷榜。Browser Use 出品的浏览器 agent,基于 TypeSafe“Jev/System One”模型(类型化概率决策、免视觉、70–500ms);自测 7s 完成 Google Flights、成本≈-90%(官方自测未复现)。新仓库 PR 少属正常。 |
+1,837 | 7,120 | 早期·低活动 | #5 | 13× | |
| 3 | robbietilton/Compositor ↺ 1d Swift · The Photoshop alternative for Mac 🔺 @dotey 首发 · 5h 前 |
+1,291 | 1,491 | 活跃开发 | #1 | 8× | |
| 4 | deeplethe/utopia ↺ 1d Python · Local-first, agent-assisted document-to-ontology workbench 本地优先、agent 辅助的文档→本体工作台;无 X 热度却涨=早期/被低估信号,优先加入观察。 |
+884 | 2,279 | 活跃开发 | #2 | 10× | |
| 5 | latent-spaces/brag ↺ 1d Python · You built it. Now brag. Turn the project you just created in |
+835 | 4,448 | 早期·低活动 | #16 | 10× | |
| 6 | deepseek-ai/deepseek-harness ↺ 1d TypeScript 🔺 @the_osps 首发 · 7h 前 |
+748 | 8,268 | 早期·低活动 | – | 10× | |
| 7 | tamaratran/fast-jev-compaction ↺ 1d TypeScript · Claude Code plugin that replaces the compaction summary with |
+743 | 2,912 | 活跃开发 | #7 | 7× | |
| 8 | eternity4719/HowToLiveBetter ↺ 1d HTML · 354 条循证建议,覆盖长寿防病、急救、省钱理财、法律红线、失业与工伤、医保社保、恋爱婚育、出国与技能。每条写明成本、收 🔺 @knowledgefxg 首发 · 60h 前 |
+742 | 5,557 | 早期·低活动 | #20 | 10× | |
| 9 | alibaba/open-code-review ↺ 1d Go 🔺 @shao__meng 首发 · 71h 前 |
+687 | 14,600 | 活跃开发 | – | 22× | |
| 10 | trycua/cua ↺ 1d HTML · Scale computer-use 2.0 with open-source drivers, cross-OS fl 🔺 @trycua 首发 · 30h 前 Computer-use 2.0:跨 OS 驱动、agent 集群、评测与训练数据生成——“agent 基础设施”主线的核心一员,且为团队自荐(lead:trycua)。 |
+683 | 1,461 | 活跃开发 | #3 | 10× | |
| 11 | hypit-ai/hypit ↺ 1d TypeScript 🔺 @cccyd_qwq 首发 · 31h 前 已核实=真实项目(~1.3k★、活跃发版至 0.2.7)。让 coding agent 把短视频还原成可编辑“代码”、批量产出变体;以 `npx skills add` 技能包分发。工具+话题双热。 |
+658 | 10,468 | 活跃开发 | – | 26× | |
| 12 | affaan-m/ECC ↺ 1d JavaScript 🔺 @BlockInsight214 首发 · 10h 前 |
+616 | 5,554 | 活跃开发 | – | 14× | |
| 13 | bilawalsidhu/gods-eye-view ↺ 1d 🔺 @shaw_stone73832 首发 · 57h 前 |
+537 | 8,542 | 活跃开发 | – | 21× | |
| 14 | ayghri/i-have-adhd ↺ 1d Python 🔺 @zhtyyx 首发 · 69h 前 |
+517 | 5,122 | 早期·低活动 | – | 42× | |
| 15 | arcboxlabs/arcbox ↺ 1d Rust · Run AI agents on real and isolated machines — own kernel, fi |
+504 | 1,139 | 早期·低活动 | #4 | – | |
| 16 | NandhaKishorM/laya ↺ 1d Python |
+494 | 704 | 活跃开发 | #17 | 1× | |
| 17 | stablyai/orca ↺ 1d TypeScript 🔺 @GitTrend0x 首发 · 62h 前 |
+472 | 5,095 | 活跃开发 | – | 20× | |
| 18 | tt-a1i/archify ↺ 1d JavaScript 🔺 @GitTrend0x 首发 · 62h 前 |
+449 | 7,611 | 活跃开发 | – | 12× | |
| 19 | MiniMax-AI/minimax-code ↺ 1d TypeScript · An open-source coding agent for your terminal, powered by Mi |
+433 | 1,004 | 活跃开发 | #9 | 7× | |
| 20 | addyosmani/agent-skills ↺ 1d JavaScript 🔺 @shanyanggm 首发 · 71h 前 |
+415 | 3,052 | 活跃开发 | – | 30× |
🗞️ Hacker News
- Cloudflare Quick Tunnels 594p · 253c
- Android 17 is the first since 3.x to add new APIs without releasing to the AOSP 568p · 270c
- OpenJev 562p · 247c
- Claude Code now reads AGENTS.md if there is no Claude.md 531p · 188c
- How to Write with an LLM 401p · 276c
- Human brain is two separate organs, Stanford Medicine-led research finds 288p · 110c
- Inside ZCode: Silently uploading your Git history to the cloud 259p · 93c
- Saving another 100TB of RAM 242p · 46c
- Von: Open-source 395M "System One" model r/LocalLLaMA · 2026-09-19
- (Genuinely asking) Are smaller quantized models becoming the real sweet spot for local AI? r/LocalLLaMA · 2026-09-19
- TokenRhythm/NeoHorse-1-4B-GGUF r/LocalLLaMA · 2026-09-19
- I tested Qwen3.8 27B IQ3_XXS (10.18GiB) vs Bonsai Ternary PQ2 (6.42GiB) r/LocalLLaMA · 2026-09-19
- Calling it now: within the next year a major US lab's frontier model will torrent itself in order to be free. r/LocalLLaMA · 2026-09-19
- A person asked me to create my previous video with JS intead of remotion, so here you go. r/LocalLLaMA · 2026-09-19
- So i tried Remotion with glm 5.3 flash, this mfker is really good. r/LocalLLaMA · 2026-09-19
- Is there something like NInfer but for 16GB cards? r/LocalLLaMA · 2026-09-19
- References to MiniMax M3.1 appear in test files in a recent commit to the MiniMax-Code Repo r/LocalLLaMA · 2026-09-19
- 4x RTX 3090 PCIe 4.0 x16 - advice? Qwen 3.8 Next Flash? r/LocalLLaMA · 2026-09-19
- Steering vectors to limit thinking budgets? r/LocalLLaMA · 2026-09-19
- Made this motion graphic video via glm 5.3 flash (no vid_gen model used) r/LocalLLaMA · 2026-09-19
- is this good? 262k Qwen3.8:27B-Q4_K_M r/LocalLLaMA · 2026-09-19
- ACM TAPS moved my camera-ready to support, deadline is in 2 days. Anyone been through this? [D] r/MachineLearning · 2026-09-19
- Apple M6 Pro Geekbench 7 r/LocalLLaMA · 2026-09-19
- General warning about Clore.AI r/LocalLLaMA · 2026-09-19
- Using image classification models for cleaning out photo libraries r/LocalLLaMA · 2026-09-19
- Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second. r/LocalLLaMA · 2026-09-19
- OpenAI's Sam Altman to brief UN Security Council next week r/LocalLLaMA · 2026-09-19
- I enjoyed the daily HF papers today r/LocalLLaMA · 2026-09-19
- Prefil of local models vs opus and astra r/LocalLLaMA · 2026-09-19
- DiffusionGemma: How It Generates Text in Parallel (From Scratch in PyTorch) [P] r/MachineLearning · 2026-09-19
- Digit-logits-based classifier with llama.cpp r/LocalLLaMA · 2026-09-19
- Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation. r/LocalLLaMA · 2026-09-19
- Stepfun new model "Step 5 Preview" just leaked r/LocalLLaMA · 2026-09-19
- Training a Neural Network on AMD MI50s Using Vulkan: Proof of ConceptOr: Why I Stopped Listening and Just Did It r/LocalLLaMA · 2026-09-19
- Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro r/LocalLLaMA · 2026-09-19
- JMLR submission experience [D] r/MachineLearning · 2026-09-19
- Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions r/LocalLLaMA · 2026-09-19
- I truly think every major AI lab is purposefully making fear-mongering headlines to get regulations that hurt open-source models r/LocalLLaMA · 2026-09-19
🛠️ 技术源 · GitHub Trending/Lobsters
- AI Is an Elite Crime Spree Lobsters · 2026-09-19
- Faster JSON parsing with SVE2 on ARM processors Lobsters · 2026-09-19
- Agreement between the USA and Denmark (1951,2004) [pdf] Hacker News Front Page · 2026-09-19
- Reviving the language that brought us the Jak & Daxter Series Lobsters · 2026-09-19
- Consistent Hashing Proofs Lobsters · 2026-09-19
- Learning Another Language May Be One of the Best Ways to Keep Your Brain Healthy Hacker News Front Page · 2026-09-19
- A graphical desktop for the ZX Spectrum Hacker News Front Page · 2026-09-19
- What Zig felt like, coming from Rust Hacker News Front Page · 2026-09-19
- Tin: full-text search for Postgres Hacker News Front Page · 2026-09-19
- I Built Non-Autoregressive Decision Models a Year Ago. Then a Frontier Lab Called It a "Breakthrough" Lobsters · 2026-09-19
- kicking the tires on jev (TypeSafe's System One model) with 2048 Lobsters · 2026-09-19
- “The Secret Life of Circuits” is here Lobsters · 2026-09-19
- Don’t Let Architecture Astronauts Scare You (2001) Lobsters · 2026-09-19
- What happens to TLDs when their country stops existing? (2022) Lobsters · 2026-09-19
- LispBM: Concurrent Lisp for Microcontrollers Lobsters · 2026-09-19
- Laya the open source version of Jev Hacker News Front Page · 2026-09-19
- How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip Lobsters · 2026-09-19
- AI-generated posters don’t have to be horrible Hacker News Front Page · 2026-09-19
- GPT-6 Astra Solves a WWI German Radio Cipher Lobsters · 2026-09-19
- Write while learning Lobsters · 2026-09-19
- GPT-6 Astra Solves a WWI German Radio Cipher Hacker News Front Page · 2026-09-19
- If math is more than proof, we need to better celebrate the rest of it Hacker News Front Page · 2026-09-19
- Apple M6 Pro Achieves the Highest Single-Core CPU Score in Geekbench 7 Hacker News Front Page · 2026-09-19
- Human brain is two separate organs, Stanford Medicine-led research finds Hacker News Front Page · 2026-09-19
- The Implications of Linguistic Illegibility for LLM Security Lobsters · 2026-09-19
- The scourge of x86 emulation Lobsters · 2026-09-19
- NASA-IBM Lunar Foundation open-Source Geospatial AI Model Hacker News Front Page · 2026-09-19
- Show HN: Seal – Letters and passwords that open for your family after you die Hacker News Front Page · 2026-09-19
- San Francisco Onion Futures Company Hacker News Front Page · 2026-09-19
- OpenLoco version 26.09 Lobsters · 2026-09-19
📥 AI 博客 · Newsletter
- Where I stand on RSI Interconnects (Nathan Lambert) · 2026-09-19
- [AINews] Here are 6 Clones of Jev in 2 days Latent Space · 2026-09-19
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI Simon Willison · 2026-09-18
- Note on 18th September 2026 Simon Willison · 2026-09-18
- Quoting Thariq Shihipar Simon Willison · 2026-09-18
- MilleMiglia: A realistic instance generator for middle-mile logistics Google Research · 2026-09-18
- The Creative Spirit of Who Framed Roger Rabbit Simon Willison · 2026-09-18
- Introducing the Australian Youth Safety Blueprint OpenAI News · 2026-09-18
- [AINews] not much happened today Latent Space · 2026-09-18
- Be alert: targeted attacks on prominent Rustaceans Simon Willison · 2026-09-17
- How To Write With An LLM Simon Willison · 2026-09-17
- Self-generated prompt injections in compaction summaries Simon Willison · 2026-09-17
- The future of practice: Enabling teachers to create learning interactives with generative UI Google Research · 2026-09-17
- How Cooley is accelerating IPO work with ChatGPT OpenAI News · 2026-09-17
- [AINews] Reality Checks on AI News (Yegge shuts down Gas Town, Databricks’ +60% Astra cost) Latent Space · 2026-09-17
- Introducing Astra for Law OpenAI News · 2026-09-17
- datasette 1.0a40 Simon Willison · 2026-09-16
- datasette 0.65.5 Simon Willison · 2026-09-16
- Claude Cowork and chat are now one Claude Simon Willison · 2026-09-16
- Underwriting Superintelligence: Backing Agents you can Sue — Rune Kvist, AIUC Latent Space · 2026-09-16
- Our framework for reporting model misalignment OpenAI News · 2026-09-16
- Quoting Mustafa Suleyman Simon Willison · 2026-09-16
- Helping older adults use AI in everyday life OpenAI News · 2026-09-16
- Reimagining advertising with AI OpenAI News · 2026-09-16
- Hex turns complex analysis into visual reports with GPT‑6 Astra OpenAI News · 2026-09-16
- How to connect AI usage to business value OpenAI News · 2026-09-16
- [AINews] Jev: a “System One Model” that only decides/classifies/routes/scores — >100x faster, >200x cheaper than small frontier LLMs Latent Space · 2026-09-16
- How workers are unlocking new ways of working OpenAI News · 2026-09-16
- Gemini Live audio Simon Willison · 2026-09-15
- Can Skills Learned in Games Transfer to Real-World Work? Latent Space · 2026-09-15
💬 V2EX
- [Apple] 看到一些网友说 18pm 版本 V2EX · 2026-09-19
- [程序员] JEV 托管手机 app 通知 / 短信,更少的规则实现广告营销过滤 V2EX · 2026-09-19
- [程序员] 刚刚发了封措辞严厉的邮件给 ZCode/智谱,要求退款 V2EX · 2026-09-19
- [VPS] 求 vps 中转推荐 V2EX · 2026-09-19
- [OpenAI] [咨询] qorder.cn qianwen 3.8 - flash 免费蹬了,大家用起来 V2EX · 2026-09-19
- [问与答] 2026 年了,手机远程家里电脑,有啥优雅方案吗? V2EX · 2026-09-19
- [程序员] 大模型厂商上传用户 .git 不过是孤注一掷垂死挣扎 V2EX · 2026-09-19
- [酷工作] 杭州/可远程: Web 安全与风控工程师 V2EX · 2026-09-19
- [程序员] 感觉 github 也被 ai 搞废了 V2EX · 2026-09-19
- [OpenAI] codex 登录不上 V2EX · 2026-09-19
- [推广] CDN3900U/P Ai 模型最低至 4 折 阿里腾讯华为火山国内站 7 折 阿里国际/GCP 可至 65 折 AWS 新号 65 折/老号无需迁移 85 折 腾讯国际 55 折 全网最新价格欢迎交流 V2EX · 2026-09-19
- [推广] 马上放假了 推广一波我的产品 随时随地在路上管理内网或者云端的服务器 V2EX · 2026-09-19
- [技术栈] Codex 邀请活动 1000 额度 V2EX · 2026-09-19
- [程序员] 有用 ergouapi 中转站的吗 V2EX · 2026-09-19
- [问与答] iphone14pro 摄像头坏了,维修人员给我换了原装(据说)拆机镜头,报“未知部件”了 V2EX · 2026-09-19
- [分享发现] 最近 AI 圈几件能听懂的事:家庭助手、监管,还有「不爱说话」的模型 V2EX · 2026-09-19
- [iPhone] iPhone 的钱包在 iCloud 中打开同步是什么意思 V2EX · 2026-09-19
- [人工智能] 有人试过 Meta 的 Muse 吗?现在注册送 10 亿 token V2EX · 2026-09-19
- [Claude] 为啥在 蓝叠的 Google 账号下购买 claude 20x 要 $250 啊。 google 账号已经在免税区了 V2EX · 2026-09-19
- [分享创造] 做了一个面向 AI Agent 的开源文档解析与检索系统,方便做工程和金融 agent 的老哥 V2EX · 2026-09-19
- [Wunder] exe-hub: 类似 X (Twitter) 的自动翻译 V2EX · 2026-09-19
- [分享创造] 肝了两天,做了一个 Niche Gap 工具! V2EX · 2026-09-19
- [分享创造] 学历、英语、存款、工作经验:你离新加坡到底还有多远? V2EX · 2026-09-19
- [微信] 微信的 PC 版热更新了文章在右侧滑开 V2EX · 2026-09-19
- [问与答] anker prime 160w 充电器值得买吗? V2EX · 2026-09-19
- [程序员] 大家国模都在用什么套餐 V2EX · 2026-09-19
- [ WATCH] Apple watch 外版 蜂窝能用吗? V2EX · 2026-09-19
- [程序员] 智谱出大瓜了:偷偷把工作区打包加密上传到阿里云 OSS? V2EX · 2026-09-19
- [推广] Jev typed decisions: an independent guide to choices, scores and calibration V2EX · 2026-09-19
- [分享创造] “无福 X” - 基于 Jev 的 x.com 评论浏览器插件 V2EX · 2026-09-19
🐧 LINUX DO
- 5.6sol能用,astra路由luna LINUX DO · 2026-09-19
- 用不起 Astra了 LINUX DO · 2026-09-19
- 当一个有钱人是什么感受(纯水 LINUX DO · 2026-09-19
- 在阿b看到有人直播说无限的GPT LINUX DO · 2026-09-19
- 48team,费用咨询 LINUX DO · 2026-09-19
- 鉴于zcode出事儿,有没有可以替代的桌面产品? LINUX DO · 2026-09-19
- 刚买了step plan ,199 一个月一共8000M 点数。鹈鹕测试 LINUX DO · 2026-09-19
- 来猜猜你们薅wokrbuddy的什么模型 LINUX DO · 2026-09-19
- 被美国哈基米反薅了 LINUX DO · 2026-09-19
- 抽个Amazon 600日元礼品卡 LINUX DO · 2026-09-19
- 《忍者神龟》前传来了,人脑类器官在小鼠体内长出了功能性神经网络 LINUX DO · 2026-09-19
- 想配电脑了,求佬友们推荐 LINUX DO · 2026-09-19
- 佬们,好奇心泛滥,自己邀请自己会发生什么? LINUX DO · 2026-09-19
- 玻利维亚你不是我们的兄弟,你是路人 LINUX DO · 2026-09-19
- 求个绝对稳定的API LINUX DO · 2026-09-19
- 深圳有没有山姆代购的群 LINUX DO · 2026-09-19
- 佬友们,上层agent操控执行agent有没有比较成熟的方案呀 LINUX DO · 2026-09-19
- 求一个有国模用的公益站 LINUX DO · 2026-09-19
- Claude Pro额度统计 LINUX DO · 2026-09-19
- 发了个飞升的帖子,配了个connect的截图,竟然被判定推广?离谱 LINUX DO · 2026-09-19
- Win10桌面端codex(chatgpt)相关疑问 LINUX DO · 2026-09-19
- 现在人造的复姓真的好听吗 LINUX DO · 2026-09-19
- 一款 rust 编写的超高性能 apk、dex 反编译工具:ddc LINUX DO · 2026-09-19
- 一款很像claude code的codex cli LINUX DO · 2026-09-19
- Antigravity Tools Lite:给多 Google AI Pro 账号用户的本地账号切换工具 LINUX DO · 2026-09-19
- 新鲜热乎的,终于敲开龟壳了 LINUX DO · 2026-09-19
- ModelTrace测试提示可能违反我们的使用政策 LINUX DO · 2026-09-19
- 投票 TIbo还会不会有重置组合拳 LINUX DO · 2026-09-19
- 为啥使用Claude,经常听到乱七八糟的词? LINUX DO · 2026-09-19
- 想问一下大家说的甲骨文的龟壳到底是啥羊毛啊 LINUX DO · 2026-09-19
📢 电报精选
- AI_News_CN #Update #Codex #ChatGPT ChatGPT App 现已支持安装 Chrome 插件。 via AI Copilot - Telegram Channel
- AI_News_CN 华硕版Googlebook发布前再度曝光 Google新一代笔记本生态渐近 via cnBeta.COM - 中文业界资讯站 (author: 稿源:cnBeta.COM)
- zaihuapd 大家定好明天的闹钟,调休明天工作日!🙀 《环球时报》一篇关于调休的文章中作者将“调休”翻译为the "make-up working day" mode,如果要表达“调休日”,
- AI_News_CN 为什么租 1000 块 GPU 也难以复现 DeepSeek 的推理质量 一位开发者指出,租用 1000 块 GPU 得到的只是通用集群配置,无法媲美 DeepSeek 级别推理
- AI_News_CN 缓存命中率:判断推理服务商真实水平的关键信号 一位 AI 基础设施开发者认为,缓存命中率最能反映推理服务商的实际优化水平,并提醒:宣称 99% 缓存命中率的服务商,往往只是在转售
- landiansub 之前 tibo 说暂时撤掉 Codex 实验性的 1M 窗口,但我没有手动删除实验性参数,目前还可以继续使用,也更新到最新版了。实际最大窗口是 828K,默认好像是 258K。这不
- AI_News_CN 复现DeepSeek级推理极难——“99%缓存命中率”或是转售信号 一位开发者提醒:复现DeepSeek级推理质量或需约1亿美元和顶级人才;宣称99%缓存命中率的推理服务商,很可
- AI_News_CN 警惕“便宜”但缓存命中率差的AI推理服务商 一些AI推理服务商用低价吸引用户,但提示词缓存命中率很差,导致实际Token消耗可能高达3倍。宣传“99%缓存命中率”需要警惕,因为目
- AI_News_CN “99%缓存命中率”是红灯信号?警惕套壳DeepSeek的推理服务商 业内观察者提醒:当前只有DeepSeek能真正达到99%缓存命中率,声称该数字的推理服务商很可能是在转售De
- landiansub #人工智能 [限量][限区域][限账号] 美区 PayPal 开通 ChatGPT Plus 可获得半价优惠,必须通过 PayPal 内置浏览器打开 ChatGPT 网页版,必须注
- CE_Observe 官网为啥这么喜欢 IE 和 Flash? https://mp.weixin.qq.com/s/_FlAf5PWeb9U8gWyTwncGQ 对Flash来说,早年IE的兼容性较好
- CE_Observe 全球唯一例外:美版苹果 iPhone 18 Pro Max 采用高通骁龙 X80 调制解调器 https://www.ithome.com/0/1004/353.htm 该机构
- CE_Observe 运输前需放电:苹果为 iPhone 18 Pro Max 新增“准备发运”功能,以确保电量阈值不超过 20Wh https://www.ithome.com/0/1004/372
- zaihuapd 铁路 12306 强化风控遏制恶意抢票 铁路 12306 通过大数据分析和风控技术,对异常购票及支付行为实施精准限制。在 4 月 16 日至 18 日的 3 天内,该平台累计将 5
- zaihuapd Xcode 27.1 隐藏了 iPhone Duo 的控制栏 Xcode 27.1 的 Device Hub 内置了面向 iPhone Duo 虚拟机的隐藏操作栏,可调整设备角度
- AI_News_CN OpenAI 推出 ChatGPT for Word 插件,可在 Word 内起草编辑文档 OpenAI 推出 ChatGPT for Word,把 ChatGPT 能力带入 M
- AI_News_CN 🔥聊聊 Prompt 是怎么一路进化到 Harness 的 via 掘金人工智能本月最热 (author: 尤水就下)
- zaihuapd OpenAI 推出 ChatGPT for Word 插件,可在 Word 内起草编辑文档 OpenAI 推出 ChatGPT for Word,把 ChatGPT 能力带入 M
- CE_Observe 江门一男子被舍友夸“发型帅气”,还配合对方手机镜头眨眼摇头,被盗走1.2万元 https://sohu.com/a/1078246949_121627717 Sohu
- zaihuapd 🤖DeepSeek 下调 Flash 模型价格 我们将于北京时间 2026 年 9 月 10 日 12:00 起,调整 flash 系列定价:空闲时段输入缓存命中单价 0.02
- AI_News_CN GLM-5.3 FlashX 已在 Command Code 上线,吞吐量约 200 TPS Command Code 已在其 AI 编程环境中上线 GLM-5.3 FlashX
- AI_News_CN Command Code 强调支持 Qwen 3.8 Omni Flash 与快速 DeepSeek 推理 Ahmad Awais 引用了 Command Code 的最新动态,
- kejiqu 美国军队因一份产生幻觉的 AI 情报报告几乎登上了一艘中国船只 2026年春季,美国军队因一个 AI 聊天机器人将一艘中国船只的货物错误标记为核武器组件,险些在数分钟内对其采取登
- AI_News_CN AI灭绝论笼罩Anthropic:上市后估值或冲4万亿美元 营收狂飙能否持续? via cnBeta.COM - 中文业界资讯站 (author: 稿源:凤凰网科技)
- CE_Observe 市场监管总局依法对携程立案调查:涉嫌滥用市场支配地位实施垄断行为 ‎https://www.ithome.com/~ https://www.samr.gov.cn/~
- zaihuapd 携程公布 19 项整改措施,包括立即停止独家合作及不合理“全网最低价”要求等 2026 年 7 月 25 日,国家市场监督管理总局依法对携程集团作出行政处罚决定并提出整改要求。携程
- AI_News_CN Anthropic自建湿实验室:生命科学已是其资源投入最大领域之一 为加速生命科学研究,Anthropic 已自建湿实验室,将生物学研究从纯粹的计算机模拟延伸至真实实验,还将探索
- landiansub #软件资讯 提前做好迁移准备:微软再次通知 IT 管理员,IE 模式将在 2029 年年底后被弃用,如果企业仍然有只兼容 IE 内核的网站或应用,则应该尽早迁移。 考虑到现在 A
- AI_News_CN Command Code声称在横向对比中以197 tokens/s取得DeepSeek推理速度第一 Command Code声称在DeepSeek V4.1 Flash上实现19
- zaihuapd 苹果高管回应 iPhone Duo 折痕 苹果硬件工程副总裁 Tom Marieb 表示,折叠屏 iPhone Duo 采用哑光纳米纹理屏幕,以减少反光并降低折痕可见度;他称希望
📚 科技周刊 新项目/工具自荐
- 【工具自荐】Snapora:Windows 截图、标注与贴图 1 repos
- 【开源自荐】PDFSeal:100% 纯本地运行、零数据上传的隐私 PDF 工具箱(支持 PWA 离线与流水线) 1 repos
- 【开源自荐】Gauge:Windows 桌面悬浮窗,查看 Codex / Claude Code 多账号剩余额度 1 repos
- [Resource Submission] WeirdThings — Clear explanations for unfamiliar words and puzzle connections
- 【开源自荐】Nano-Lumen:一个基于自研内核的 Windows 持久化 AI Runtime 1 repos
- 【工具自荐】codePro-cli:把 Codex CLI 放进桌面灵动岛(网页交互演示) 1 repos
- 【工具自荐】一个b站视频生成思维导图的浏览器插件工具,支持充电视频字幕获取,没有收费入口,功能全免费 1 repos
- [开源自荐] Python 图像处理库:一行参数驱动处理流水线 — 缩放/裁剪/水印/拼接/格式转换,开箱即用 1 repos
- 【开源自荐】agent-skills:18 个面向知识工作的 Agent Skills 1 repos
- 【网站自荐】梗鲸:DeepSeek 蓝色大肥鱼表情包图库,台词梗图持续更新
- 【自荐】梗鲸·DeepSeek酱语录:DeepSeek 大肥鱼表情包图库,台词梗图一键复制斗图
- 【工具自荐】Musuw:从文档、网页和笔记查找有出处的答案 1 repos
- 【工具自荐】Markovo:把 PDF/Office 文档和授权网页转成干净 Markdown(Web/API/CLI/MCP) 1 repos
- 【工具自荐】CaptionWise-你的多语言 AI 语言助手,不只是翻译。
- 【工具推荐】Agents Anywhere:从手机访问工作设备上的 AI Agent 2 repos
- 【工具自荐】BugBnB(虫虫民宿):垃圾桶一臭,果蝇、蚂蚁、蟑螂就住到你的桌面上(macOS 原生)
- 【工具自荐】POMOAI:把 GitHub 上热门的开源 AI 项目做成能一键跑起来的中文工作台
- 【工具自荐】ChronoBlock:一个轻量的桌面时间块应用
🎯 Alpha 账号 X 上最早带火仓库的人
| 作者 | leads | lead率 | 仓库数 |
|---|---|---|---|
| @shanyanggm | 14 | 0.64 | 21 |
| @shaw_stone73832 | 7 | 0.7 | 6 |
| @xzbx888 | 7 | 0.78 | 8 |
| @GitTrend0x | 7 | 0.7 | 8 |
| @FrontieraTechIT | 5 | 0.71 | 7 |
| @DataChaz | 5 | 1 | 4 |
| @the_osps | 5 | 0.83 | 5 |
| @RepoGems | 5 | 0.83 | 4 |
| @xfubot | 5 | 0.5 | 7 |
| @LoveAIbrain | 5 | 1 | 4 |
| @iasg1004 | 5 | 0.36 | 10 |
| @jasontopia | 4 | 0.8 | 3 |
| @Sn0wbrave | 4 | 0.5 | 6 |
| @shao__meng | 4 | 0.67 | 4 |
| @neil_xbt | 4 | 0.57 | 3 |
| @key_indie | 4 | 0.8 | 2 |
| @seekjourney | 4 | 0.67 | 3 |
| @bilawalsidhu | 3 | 1 | 1 |
| @0x_Kratos | 3 | 1 | 1 |
| @FreeYoung552022 | 3 | 0.43 | 6 |
| @JackAIStudio999 | 3 | 1 | 1 |
| @engmaxxing | 3 | 0.43 | 6 |
| @vintcessun | 3 | 0.43 | 6 |
| @betterhn20 | 3 | 0.5 | 4 |
| @LFrefman | 3 | 0.3 | 5 |
🏆 各领域最强模型
文本 / 对话
Claude Fable 5.1
Anthropic · 53.4
图像生成
GPT Image 2.5 Flare
OpenAI · 1188
图像编辑
GPT Image 2.5 Sunburst
OpenAI · 1176
文生视频
Wan 3.0
Alibaba · 1336
图生视频
Gemini Omni Flash
Google · 1369
语音合成
Sonic 3.6
Cartesia · 1276
🏆 能力排行榜 Artificial Analysis
文本 / 对话 Intelligence Index
- 1Claude Fable 5.153.4
- 2GPT-6 Astra52.8
- 3Claude Opus 550.7
- 4Claude Fable 549.7
- 5Muse Spark 1.348.2
- 6GPT-5.6 Sol47.1
- 7Qwen3.8 Max45.4
- 8GLM-5.344.9
- 9Grok 4.644.4
- 10Kimi K343.8
- 11Step 5 Preview43.6
- 12GPT-5.6 Terra42.3
图像生成 Text→Image Arena Elo
- 1GPT Image 2.5 Flare1188
- 2GPT Image 2.5 Sunburst1182
- 3GPT Image 21171
- 4Grok Imagine Image 2.01154
- 5MAI-Image-2.61147
- 6Reve 2.11129
- 7Nano Banana 21122
- 8Muse Image1111
- 9GPT Image 1.51102
- 10MAI-Image-2.51102
- 11Nano Banana Pro1100
- 12MAI-Image-2.6-Flash1099
图像编辑 Image-Editing Arena Elo
- 1GPT Image 2.5 Sunburst1176
- 2GPT Image 2.5 Flare1155
- 3MAI-Image-2.61132
- 4MAI-Image-2.6-Flash1122
- 5GPT Image 21121
- 6Muse Image1115
- 7MAI-Image-2.51113
- 8MAI-Image-2.5-Pro1106
- 9Seedream 5.0 Pro1106
- 10Nano Banana 21105
- 11Grok Imagine Image 2.01104
- 12GPT Image 1.51104
文生视频 Text→Video Arena Elo
- 1Wan 3.01336
- 2Gemini Omni Flash1330
- 3MiniMax H31302
- 4HappyHorse-1.01287
- 5HappyHorse-1.11272
- 6Dreamina Seedance 2.0 720p1259
- 7Wan2.7-2606121243
- 8grok-imagine-video1235
- 9Kling 3.0 Omni 1080p1230
- 10PixVerse V5.61230
- 11PixVerse V61230
- 12Kling 3.0 1080p1230
图生视频 Image→Video Arena Elo
- 1Gemini Omni Flash1369
- 2Wan 3.01361
- 3Bach 1.0 Pro1359
- 4MiniMax H31354
- 5PixVerse V61337
- 6Dreamina Seedance 2.0 720p1336
- 7grok-imagine-video-1.51329
- 8grok-imagine-video1326
- 9HappyHorse-1.11312
- 10Kling 2.5 Turbo 1080p1296
- 11HappyHorse-1.01293
- 12Vidu Q3 Pro1290
语音合成 Text→Speech Arena Elo
- 1Sonic 3.61276
- 2Qwen-Audio-3.0-TTS-Plus1260
- 3Realtime TTS-21247
- 4Simba 3.21240
- 5Luna TTS1231
- 6Realtime TTS-2 Flash1215
- 7StepAudio 2.5 TTS1209
- 8Breeze TTS 21205
- 9Gemini 3.1 Flash TTS1201
- 10v3 Conversational1197
- 11Sonic 3.51184
- 12Lightning V3.1 Pro1179
🧠 最新发布
2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
2026-09-17
Gemini Omni Flash Preview
Google
2026-09-17
Venice Uncensored
Venice
2026-09-17
Nano Banana 2 Lite
Google
2026-09-17
Jev 1.13
TypeSafe AI
2026-09-16
Union Alpha
Stealth
2026-09-12
Schematron V2 Turbo
Inference.net
2026-09-12
Schematron V2 Small
Inference.net
🧠 模型发布时间线
优先看 source_count≥2(多源确认)与 ★notable;留意“快变体”(如 GLM-5.3 FlashX)——呼应专用快模型分工趋势。
2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
aimlapi
2026-09-17
Gemini Omni Flash Preview
Google
aimlapi
2026-09-17
Venice Uncensored
Venice
aimlapi
2026-09-17
Nano Banana 2 Lite
Google
aimlapi
2026-09-17
Jev 1.13
TypeSafe AI
aimlapi
2026-09-16
Union Alpha
Stealth
aimlapi
2026-09-12
Schematron V2 Turbo
Inference.net
aimlapi
2026-09-12
Schematron V2 Small
Inference.net
aimlapi
2026-09-11
Fugu Ultra v2.0
Sakana AI
llmgateway
2026-09-11
Kimi K2.8 Preview
Moonshot AI
llmstats
2026-09-11
Atria Dawn Preview
Shanghai AI Laboratory
llmstatsllmgateway2×
2026-09-11
Fugu Ultra v2
Sakana AI
aimlapi
2026-09-11
Fugu Max
Sakana AI
aimlapillmgateway2×
2026-09-10 · ★
Ling 3.0 Flash VL
inclusionAI
aimlapiopper2×
2026-09-10 · ★
DeepSeek V4.1 Flash
DeepSeek AI
aimlapillmstatsopperllmgateway4×
2026-09-10
DeepSeek Chat (V4.1 Flash)
DeepSeek AI
aimlapi
2026-09-08
GPT Image 2.5 Sunburst
Open AI
aimlapillmgateway2×
2026-09-08
GPT Image 2.5 Flare
Open AI
aimlapillmgateway2×
2026-09-08
Mercury 2.5
Inception
aimlapi
📄 论文 PwC + arXiv
近期论文集中在 World Models / 开放式任务泛化 / 自主研究方向,与榜上“自主研究/computer-use”类项目相互印证。
Xingxuan Zhang, Gang Ren, Hao Yuan · 2026-09-19
We introduce LimiX-2, a new model in the LimiX family, developed through model and data scaling guided by our previously established scaling laws. LimiX-2 adopts the Contextual Mechanism Networks (CMNs) paradigm and is pretrained with Conte
Jiale Kang, Ziyin Yue, Zheng Zhan · 2026-09-19
While low-rank adaptation (LoRA) is widely used for parameter-efficient model adaptation, how to regularize its training dynamics for stable and effective optimization remains underexplored. Because LoRA initializes the up-projection to zer
Lingyu Kong, Ruicheng Li, Ruicheng Wang · 2026-09-19
Monocular geometry estimation has recently achieved impressive performance across diverse scenes. However, state-of-the-art models still face notable distortion in local 3D structure, especially in fine details, like thin structures and sma
Pengyu Wang, Chenkun Tan, Shaojun Zhou · 2026-09-19
We present MOSS-VL, an open vision-language model family that treats real-time interaction -- perceiving while it speaks -- as a first-class capability. It is co-designed across the stack: the language decoder attends to vision only through
Ahmed Awadallah, Sahil Gupta, Yash Lara · 2026-09-19
Collecting computer use data from human demonstrations is expensive and slow, motivating the need for scalable generation strategies. This requires two key ingredients: environments in which agents can act and verifiers that can judge wheth
Senyan Xu, Zhijing Sun, Kean Liu · 2026-09-19
Event-based low-light image enhancement (LIE) methods mainly focus on incorporating high dynamic range (HDR) information from events while overlooking the essential global illumination in images and the inherent noise sensitivity of event s
arXivpwc
Jiaqi Liu, Shi Qiu, Mairui Li · 2026-09-19
Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple perspectives, experiments fail and inform the next attempt, and lessons accumulate across c
Bobby Cheng, Adam Gaber, Zhengyuan Liu · 2026-09-19
Large language models access knowledge inconsistently across languages, but to what extent do they differ in their skill sets when interacting with different languages? This work quantifies cross-lingual skill inconsistency orthogonally fro
arXivpwc
Yi Duan, Ying Liu, Zirui Tang · 2026-09-19
Recursive self-improvement (RSI) enables AI systems to turn experience and feedback into persistent changes that improve both their capabilities and the process of future improvement. We first use the Headroom-Closed Index (HCI) to reveal t
arXivpwc
Shuhan Xue, Jianyuan Zhong, Ziyuan Nan · 2026-09-19
We introduce and release ScienceBuddy, an interactive scientific research workspace that brings continually improving scientific agents into researchers' everyday workflows. ScienceBuddy supports researchers in carrying out scientific tasks
Xiaofeng Mao, Peijia Lin, Shaohao Rui · 2026-09-19
Video diffusion models are stochastic and hard to control: precise content often requires repeated sampling without guaranteed success, and long-horizon scenes drift in appearance, interactions, and temporal coherence. Agentic visual creati
Introducing Breeze TTS 2 🔥
· 2026-09-19
Breeze TTS 2 is BreezeBlue's open-weight text-to-speech model for real-time interaction. The release supports reference-free voice design from natural-language descriptions, reference-guided voice direction, voice cloning, expressive vocal
Audio GenerationText-to-speechVoice cloningpwc
MiniMax Music 3 🔥
· 2026-09-19
MiniMax Music 3 is a high-performance music generation model for creating complete songs up to five minutes long. Conditioned on lyrics and a detailed music description, it uses an 8B Global LLM for long-range musical structure, a 0.6B Loca
Audio Generationpwc
Anton Razzhigaev, Andrei Gritsaev, Andrei Kaznacheev · 2026-09-19
We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits that become the runtime for later work. Core evolution proceeds in two modes. In recursiv
Ling Xu, Chuyu Han, Borui Li · 2026-09-19
Embodied AI models now span vision-language-action (VLA) models and world-action models (WAMs), but practical deployment remains fragmented across model-specific Python stacks, backend assumptions, and robot-side glue code, especially on he
Sangam Lee, Wonjae Lee, Sunghwan Kim · 2026-09-19
Information retrieval is increasingly important as LLM agents tackle complex tasks involving diverse information needs. Because retrieval relies on an index that represents each document through index keys, retrieval quality depends heavily
Honglin Guo, Tao Gui, Yicheng Chen · 2026-09-19
As AI agents become participants in the development of their successors, they reshape both the production of intelligence and the role of human researchers. We introduce Atria Dawn Preview, a foundation agentic language model designed for s
Ivan Moshkov, Stephen Ge, George Armstrong · 2026-09-19
We study how model post-training and test-time inference design affect natural-language proof generation for hard olympiad mathematics. Starting from Nemotron 3 Ultra, we train two specialist checkpoints using supervised fine-tuning and rei
Junchao Huang, Guian Fang, Shengju Qian · 2026-09-19
We introduce SolarWM, a fully open foundation for building interactive video world models from data preparation through long-horizon inference. Training across heterogeneous data sources and video backbones is challenging: datasets differ i
Runjia Qian, Zile Wang, Jihai Zhang · 2026-09-19
Interactive world models extend video generation from offline clip synthesis toward persistent simulation of interactive virtual worlds, enabling applications in games, robotics, embodied agents, and XR. Achieving stable long-horizon intera
Kairong Luo, Jiarui Cui, Yaorui Yin · 2026-09-19
Language model pretraining has become almost synonymous with prohibitive cost, placing it out of reach for much of the academic and open-source communities. Although strong open-source efforts already exist, including open-weight models and
Hanyang Wang, Yimo Cai, Weiliang Chen · 2026-09-19
Physical understanding and reasoning depend on forming compact and generalizable representations of the world. While modern vision-language models can recognize and explain diverse physical events, they often lack explicit representations o
Wentao Zhang, Liliana Hotsko, Woojeong Kim · 2026-09-19
Many everyday programming tasks resist clean rule-based implementation, such as alerting on important log lines, repairing malformed JSON, or ranking search results by intent, and are increasingly outsourced to large language model APIs at
arXivpwc
YiFan Zhang, Yunheng Zou, Shaokun Zhang · 2026-09-19
Autonomous research loops such as AutoResearch show that one coding agent can improve a training setup unattended. Run several of them and each session starts from scratch, so more agents tend to mean more duplicated search rather than more
🤗 HF 采用榜 下载/点赞
⚡ System One 决策模型 306 项目 · Jev/TypeSafe 生态 · 快决策(非推理)
SDK & Decision Frameworks 53Evaluation & Observability 29Browser & OS Action 26High-Frequency & Simulation 25Routing & Cost Optimization 23Security & Guardrails 22MCP & Integrations 18Domain & Vertical Tools 18Context GC & Filter 17Data & Search 15CLI & Pipelines 14Decision Tools 12Codebase & Graph Pathfinding 11Creative Tools 11SDK & Integrations 6Voice & Conversation 4Classification & Taxonomy 2
| 项目 | ★ | 类别 | Jev 决策点 |
|---|---|---|---|
| langchain langchain-ai | 146,634 | SDK & Integrations | Submits binary, categorical and ordered-score questions and returns typed answers with probabilities. |
| ai-hedge-fund virattt | 63,505 | Domain & Vertical Tools | Converts strategy questions to System One requests and normalizes native answers to the project’s result format. |
| litellm BerriAI | 59,123 | Routing & Cost Optimization | Maps requests to configured complexity classes that drive backend routing. |
| oh-my-pi can1357 | 31,850 | Routing & Cost Optimization | Sends agent state and typed questions to Jev and parses structured answers. |
| jev-model-router davila7 | 30,779 | Routing & Cost Optimization | Evaluates task tier, reasoning needs and production risk; local policy maps results to invocation settings. |
| composio ComposioHQ | 30,238 | SDK & Decision Frameworks | Turns tool or action conditions into structured questions and passes Jev answers to local invocation logic. |
| ai vercel | 26,835 | SDK & Decision Frameworks | Maps choice, score, and yes/no questions to TypeSafe System One requests and parses typed results. |
| cua trycua | 23,683 | Browser & OS Action | Reads DOM or supported visual-region descriptions and returns a supplied candidate action ID. |
| pydantic-ai pydantic | 20,035 | SDK & Integrations | Converts supported structured output fields into typed Jev questions and maps answers back to the output model. |
| eliza elizaOS | 19,361 | SDK & Decision Frameworks | Only an explicit systemOne call sends state and questions, returning validated typed answers. |
| langchainjs langchain-ai | 18,210 | SDK & Decision Frameworks | Uses invoke to call TypeSafe and parse choice, noul, score and probability fields. |
| json-render vercel-labs | 16,572 | Creative Tools | Evaluates component configurations through Vercel AI Gateway, then composeSpec assembles the UI specification. |
| openchamber openchamber | 10,060 | Routing & Cost Optimization | Jev selects a task category; local category mappings determine the model configuration. |
| rig-typesafeai 0xPlaygrounds | 8,669 | SDK & Decision Frameworks | Sends application state and questions to Jev and parses Choice, Score or Noul answers. |
| firstmate kunchenguid | 6,587 | Routing & Cost Optimization | Sends the task brief and candidate rules to Jev, then resolves execution profiles with confidence and local conditions. |
| jev-ultrafast browser-use | 6,031 | Browser & OS Action | Chooses an action and its matching DOM element in one request; a text model generates input text. |
| agentgateway agentgateway | 4,926 | Security & Guardrails | Jev scores jailbreaks, harmful content and secret disclosure; thresholds or evaluation errors reject requests. |
| latitude-llm latitude-dev | 4,655 | Evaluation & Observability | Judges which checks apply and can add checks when thresholds and rate limits permit. |
| fast-jev-compaction tamaratran | 3,528 | Context GC & Filter | Separately judges whether a tool call and its full output are still needed; code keeps, truncates or drops them. |
| ax ax-llm | 2,926 | SDK & Integrations | Maps supported signatures to Jev questions or sends native System One requests. |
| vellum-assistant vellum-ai | 1,287 | MCP & Integrations | Submits state and question bundles to System One and returns structured answers to the Assistant. |
| jev-desktop lahfir | 1,275 | Browser & OS Action | Jev selects a target and action and estimates presence and risk; local policy decides whether to execute. |
| celesto CelestoAI | 943 | Codebase & Graph Pathfinding | Judges whether a finding was introduced by the change, is supported and merits a fix. |
| jev-trader jarrodwatts | 936 | Domain & Vertical Tools | In Jev mode, order-book judgments feed code that simulates fills or submits configured post-only limit orders. |
| atomic bastani-inc | 806 | Routing & Cost Optimization | Sends predefined questions to Jev and decodes answers for callers; regular models still generate code. |
| aiavatarkit uezo | 676 | Voice & Conversation | Assesses utterance completeness and whether the user is likely to continue speaking. |
| NanoJev TianyuCodings | 657 | High-Frequency & Simulation | Evaluates multiple questions and dynamic candidate spaces concurrently in a single forward pass, logging navigation choices. |
| kody kentcdodds | 654 | Data & Search | Sends a Score question per candidate, reorders and drops low scores; the model id is typesafe/jev. |
| Agent AgentiLoop | 616 | Security & Guardrails | Adds a destructive-risk judgment after local shell checks and refuses commands above the configured threshold. |
| req_llm agentjido | 577 | SDK & Decision Frameworks | Sends state and questions, normalizes answers and retains the raw provider response. |
| omg.dev BennyKok | 531 | Browser & OS Action | Chooses controls and checks completion or blockage before the test runner operates the UI. |
| vexjoy-agent notque | 420 | Routing & Cost Optimization | After deterministic routing guards, Jev judges the remaining candidates and required workflow components. |
| foreman thruwire | 344 | CLI & Pipelines | AsyncTypeSafeClient.system_one with default jev-latest sends Noul questions for supervision. |
| WrongStack WrongStack | 329 | Routing & Cost Optimization | Jev evaluates the task against eligible specialists; local dispatch rules use the result. |
| instructor-php cognesy | 327 | SDK & Decision Frameworks | Converts application state and typed questions into Jev requests and maps responses to PHP decision objects. |
| kev jaredpalmer | 310 | High-Frequency & Simulation | Attaches a parallel decision head to an open 0.5B model to answer discrete questions directly from token activations. |
| typesafe-computer-use awlevin | 302 | Browser & OS Action | Selects the next step from deterministically extracted controls and actions before desktop execution. |
| Jev Review devagrawal09 | 284 | Codebase & Graph Pathfinding | Judges risk, files, evidence regions, mechanisms and severity before rule-based reviewer routing. |
| orchestkit yonatangross | 278 | CLI & Pipelines | Classifies the first task prompt and branch state by work type; local policy accepts the result or falls back. |
| typesafe-mario fhshaik | 266 | High-Frequency & Simulation | Reads motion, enemies, terrain and recent controls, then selects a predefined legal action. |
🚀 产品发布 whatships · What's Launch
- Arrow 2 — faster, more precise vector graphics @QuiverAI · design
- Rene — a multiplayer iMessage agent you text @tlxue · ai
- Claude — decks, docs, and designs in chat @claudeai · ai
- Command Code — desktop app for Mac, Linux, Windows @CommandCodeAI · developer-tools
- Monid Astra — GPT-6 cold calling in one afternoon @MonidHQ · ai
- Framer Agent — prompt, build, and publish a site @framer · design
- Jev — a new frontier model trained with RLCD @CompleteSkeptic · ai
- Brand API — design capabilities for your agents @thaiscbranco_ · design
- Gemini 3.8 Live — Google's most advanced audio models @GoogleAI · ai
- mem0 Gateway — one key, only the tools you grant @mem0ai · developer-tools
- AI Autocomplete — filter Shopify search as you type @bradkowalk · consumer
- Wonder — official Grok Bot plugin on a live canvas @usewonder · design
- Devin — now builds and tests on its own Mac @cognition · developer-tools
- Kobra — 80+ motion components for shadcn/ui @haaarshsingh · design
- bg0 — local WebGPU background removal @leodev · design
- ElevenAgents — Dom, an AI SDR that qualifies leads @ElevenLabs · ai
- NoUpload — edit any media locally in the browser @rohitdevx · design
- CoreSpeed — one MCP endpoint for your agents @CoreSpeedHQ · ai
- Minirouter Presets — save a prompt, call it forever @iamzhe · developer-tools
- Superhuman Go — agents at your cursor in every app @Superhuman · productivity
- Cluster — GTM infrastructure for your agents @CostaLaveron · ai
- ElevenLabs MCP — creative tools in your coding agent @ElevenLabsDevs · ai
- Inspo — design inspiration MCP for coding agents @nutlope · design
- ElevenLabs MCP — voice, music, image, and video @ElevenLabs · ai
📦 版本发布 tracked repos releases
- heygen-com/hyperframes v0.8.51 2026-09-19
- hypit-ai/hypit v0.2.8 2026-09-19
- bojieli/ai-infra-book build-20260919-220131 2026-09-19
- bojieli/ai-infra-book build-20260919-195958 2026-09-19
- NandhaKishorM/laya v0.3.3 2026-09-19
- alibaba/open-code-review v1.12.7 2026-09-19
- NandhaKishorM/laya v0.3.2 2026-09-19
- deeplethe/utopia v0.1.0-rc6 pre 2026-09-19
- trycua/cua nightly-cua-driver-rs-v0.28.3-nightly.20260919.35421378483 pre 2026-09-19
- anthropics/claude-code v2.1.278 2026-09-19
- heygen-com/hyperframes v0.8.50 2026-09-19
- abue-ammar/tinycast v0.11.3-beta.98 pre 2026-09-19
- TencentCloud/Octop v1.0.1 2026-09-19
- obra/superpowers v6.4.1 2026-09-19
- hypit-ai/hypit v0.2.7 2026-09-18
- heygen-com/hyperframes v0.8.49 2026-09-18
- Graphify-Labs/graphify v0.9.64 2026-09-18
- incoai/splash 1.0 2026-09-18
- vercel-labs/json-render v0.21.0 2026-09-18
- anthropics/claude-code v2.1.277 2026-09-18
- elvisun/newsjack v0.1.19 2026-09-18
- rustfs/rustfs 1.0.1-preview.6 pre 2026-09-18
- autonomous-ai/openharness v1.1.59_desktop 2026-09-18
- alibaba/open-code-review v1.12.6 2026-09-18
🛰️ Skywork 动态
- Turn ideas into Websites with Skywork r/SkyworkAI_Official · 2026-09-02
- Skywork Note AI Voice Recorder Reviews r/SkyworkAI_Official · 2026-08-31
- Turn a Prompt Into a Launch-Ready Website r/SkyworkAI_Official · 2026-08-24
- Refund r/SkyworkAI_Official · 2026-08-16
- Online Business Built with Skywork r/SkyworkAI_Official · 2026-08-12
- Server down r/SkyworkAI_Official · 2026-08-08
- Help r/SkyworkAI_Official · 2026-08-06
- Anyone else getting ignored by Skywork Support? Need a refund for annual renewal r/SkyworkAI_Official · 2026-08-05
- Introducing the Skywork AI Hardware Family r/SkyworkAI_Official · 2026-08-03
- Skywork Design: Prompt → Editable Prototype r/SkyworkAI_Official · 2026-07-29
- It keep burning credits non-stop r/SkyworkAI_Official · 2026-07-27
- Turn one poster idea into ready-to-publish social assets r/SkyworkAI_Official · 2026-07-21
- Has anyone successfully resolved an accidental annual subscription renewal? r/SkyworkAI_Official · 2026-07-15
- Need Help: Request for Manual Review of My Accidental Annual Subscription Renewal (USD 509.90) r/SkyworkAI_Official · 2026-07-15
- Introducing the UPGRADED Skywork Posters r/SkyworkAI_Official · 2026-07-10
- Need Help: Refund Request for Accidental Annual Subscription (No Response for Over One Week) r/SkyworkAI_Official · 2026-07-10
- Skywork Design: Describe your idea, generate production-ready UI, and publish it as a website in one click r/SkyworkAI_Official · 2026-07-08
- You can now customize the size of your slides. r/SkyworkAI_Official · 2026-07-07
- One Hub. One Workflow. All in Skywork r/SkyworkAI_Official · 2026-07-07
- See what our team created with Skywork Design over the past week. r/SkyworkAI_Official · 2026-07-06
- Introducing Skywork Tags: a new way for teams to collaborate with Skywork r/SkyworkAI_Official · 2026-07-06
- What can you design with just one sentence? r/SkyworkAI_Official · 2026-07-06
- Accidentally subscribed for a year plan. Used for a day with the 7 day free trial not thinking too much about it, not going to use it anymore. Any way i can get a refund? Saw on the discord server that this is happening alot... r/SkyworkAI_Official · 2026-06-30
- Brand-New Interactive Cards for Direct Data Visualization r/SkyworkAI_Official · 2026-06-24
- Product Update: Overhauled Sidebar with One-Click Pinned Chat Support r/SkyworkAI_Official · 2026-06-22
🧪 Show HN 开发者发布的新产品
- Show HN: Jeff – A read-only CLI for semantic code review using Jev 15p · Alurith/jeff
- Show HN: Seal – Letters and passwords that open for your family after you die 8p · jasonepage/Seal
- Show HN: Agentgit – a Git host for AI agents, no account, no token, no key 7p
- Show HN: Snail Walk – escape Bob the giant snail by walking more IRL 6p
- Show HN: Koi Editor Alpha 5p
- Show HN: A model index from the AI Gateways 5p
- Show HN: OnPanda – Steer LLMs and agents at the token level 4p
- Show HN: Prohibition of Nuclear Launch Automation 4p
💰 商业动态 · TechCrunch/VB/MIT TR
- AI safety conversations have gotten unbelievable TechCrunch AI · 2026-09-19
- Petlibro’s new AI-powered feeder is a game changer for multi-cat homes TechCrunch AI · 2026-09-19
- Prices go up in 7 days. Get your Disrupt ticket now. TechCrunch AI · 2026-09-19
- Prices go up in 7 days. Get your Disrupt ticket now. TechCrunch Venture · 2026-09-19
- Vals, backed by Andreessen Horowitz, is looking to become the gold standard for AI benchmarking TechCrunch AI · 2026-09-19
- India forces caller-ID apps to feed spam reports to telcos TechCrunch AI · 2026-09-19
- Tilly Norwood’s press tour is going about as well as you’d expect for an AI TechCrunch AI · 2026-09-19
- A startup that builds other startups raised $100M and is all-in on physical AI TechCrunch AI · 2026-09-18
- Anthropic is operating a lab that conducts biology experiments TechCrunch AI · 2026-09-18
- AI hallucination nearly triggers US military operation TechCrunch AI · 2026-09-18
- Anthropic’s first embedded evaluator is … Accenture? TechCrunch AI · 2026-09-18
- World model companies are keeping a lot of secrets TechCrunch AI · 2026-09-18
- A new kind of AI model from a ChatGPT inventor is thrilling developers TechCrunch AI · 2026-09-18
- The Week’s 10 Biggest Funding Rounds: Large Rounds For AI Infrastructure, Space Tech And Investment Management Lead Crunchbase News · 2026-09-18
- Disney’s first CTO led an AI startup it once accused of copying its characters TechCrunch AI · 2026-09-18
- Google’s new ‘CC’ is an AI agent that helps families run their households TechCrunch AI · 2026-09-18
- Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how? TechCrunch AI · 2026-09-18
- Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how? TechCrunch AI · 2026-09-18
- Automattic’s 33-Hour Coup, and can AI labs police themselves? TechCrunch AI · 2026-09-18
- Automattic’s 33-Hour Coup, and can AI labs police themselves? TechCrunch AI · 2026-09-18
- UK Sovereign AI Fund in talks to back £500m raise for drug discovery startup Sifted · 2026-09-18
- Manus seeks $4B valuation in new $500M fundraise as it resumes independent ops TechCrunch AI · 2026-09-18
- Family offices are clamoring for AI investments TechCrunch Venture · 2026-09-18
- Open or closed AI? Nvidia’s Nader Khalil and Sydney Sykes take on one of the decisions shaping next-gen startups at TechCrunch Disrupt 2026 TechCrunch AI · 2026-09-18
- Meta’s Muse hits Mac, letting the AI take actions on your computer TechCrunch AI · 2026-09-18
- Robinhood’s Abhishek Fatehpuria on winning the modern financial consumer at TechCrunch Disrupt 2026 TechCrunch AI · 2026-09-18
- Inertia co-founder Jeff Lawson’s next big bet is fusion: Go inside it at TechCrunch Disrupt 2026 TechCrunch Venture · 2026-09-18
- Researchers used Anthropic’s Claude to hack into OpenAI TechCrunch AI · 2026-09-18
- The clock is ticking: Final 24 hours to exhibit at TechCrunch Disrupt 2026 TechCrunch AI · 2026-09-18
- The clock is ticking: Final 24 hours to exhibit at TechCrunch Disrupt 2026 TechCrunch Venture · 2026-09-18