🔍 搜索中 · 显示所有标签页的匹配项 · 按 Esc 清除
📊 深度分析(2026-09-20 20:40 北京)基于更早的数据快照,数据已刷新,已自动隐藏过时结论 —— 待重新生成分析后显示。
💎 高价值精选 AI 判断 · 跨源挖掘
- 2.1 [程序员] 智谱出大瓜了:偷偷把工作区打包加密上传到阿里云 OSS? V2EX · 技术 · 工程师
- 2.1 Rapidly scaling online storage to serve over 1 billion ChatGPT users OpenAI News · 技术 · 工程师
- 2.0 Powering AI is an architecture problem MIT Tech Review AI · 技术 · 决策者
- 2.0 A Hard Year For Software IPOs Crunchbase News · 商业 · 决策者
- 1.9 Astronex-World 1.0: Real-Time Interactive World Model Foundation arXiv cs.AI · 模型 · 研究者
- 1.9 GoBench: Evaluating LLMs on the game of Go [R] r/MachineLearning · 模型 · 研究者
- 1.9 Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’ TechCrunch AI · 商业 · 决策者
- 1.9 Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’ TechCrunch Venture · 商业 · 决策者
- 1.9 Governance-as-Code: Translating EU AI Act Technical Requirements into Executable Compliance Pipelines for Generative AI Systems arXiv cs.AI · 技术 · 工程师
- 1.9 Design of the IBM Granite 5.0 TurboCTC ASR Model arXiv cs.CL · 模型 · 研究者
- 1.9 The Crunchbase Tech Layoffs Tracker Crunchbase News · 商业 · 决策者
- 1.9 I trained a 44M parameter quantized LLM from scratch on 45B tokens. It ships in 19.8 MB and runs at ~1,900 tok/s on CPU. [P] r/MachineLearning · 模型 · 工程师
📰 最新快讯
- SGLang-Diffusion 首发支持 Qwen-Image-2.1:文生图、图像编辑与透明 RGBA 输出
- Qwen-Image-2.1:一个 7B 模型同时支持生成与编辑,现可在浏览器中直接体验
- OpenCode2 桌面版将新增文件查看器,支持 CSV 预览
- 智谱AI将ZCode以Apache 2.0协议开源
- Z.ai 回应 ZCode 安全问题:已完成修复并开源
- TypeSafe AI 为每位用户提供 5 美元 API 额度(约 1.2 亿 tokens)
- TypeSafe AI 的 Jev 现已全面开放,无需等待名单
- Command Code 将 DeepSeek V4.1 Flash 的 60 美元用量额度再延长 8 天
- Command Code 将 GOAT 套餐的 DeepSeek V4.1 Flash 60 美元用量延长 8 天
- DeepSeek V4.1 Flash 上线 Command Code,$60 用量再延长 8 天
- Cline Desktop 上线一周承担6%任务,限时免费提供Kimi K3
- Cline 桌面应用支持自定义图标与主题颜色
- Qoder 将于 9 月 21 日起将 Auto 额度消耗降至 0.5 倍
- Qwen-Image-2.1 展示强大文字渲染能力,vLLM-Omni 与 Diffusers 已适配
- Qwen-Image-2.1 在 RTX 3090 上本地运行,出图效果获好评
- Qwen-Image-2.1 上线 vLLM-Omni,首发即支持
- Qwen-Image-2.1 现已支持在 ComfyUI 中使用,开放权重
- Qwen-Image-2.1 登陆 ComfyUI:开源 7B 模型支持 2K 生成、指令编辑与 RGBA 输出
- Qwen-Image-2.1 上线即获 vLLM-Omni 支持
- Qwen发布Qwen-Image-2.1:轻量级开源图像生成与编辑模型
- Qwen-Image-2.1 在文字与肖像渲染上有明显提升
- Qwen 发布 Qwen-Image-2.1:开放权重,支持图像编辑与故事板生成
- Qwen-Image-2.1 支持圆形标注多区域局部编辑
- Qwen 发布 Qwen-Image-2.1:开源权重 7B 图像生成与编辑模型
- Qwen-Image-2.1 在 Apple Silicon 上的实测:MLX bf16 约 1.78 秒/步
- Hugging Face 上线两款新 OCR 模型:腾讯 WeVisDoc 与 Jina OCR v1
- StepFun 发布 Step 5 Preview:面向软件工程与金融的旗舰智能体模型
- Step 5 Preview 公布聚焦金融场景的评测基准
- Step 5 Preview 面向专业知识工作:大规模研究、可审计报告一键生成
- StepFun 发布 Step 5 预览版:可连续运行 24 小时的 AI 智能体
⭐ 多源共振
githubxsocialboard
+329★/d 早期·低活动 official #4
Native MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.
githubxsocialboard
+246★/d 活跃开发 official #15
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and remote runtime.
🔺 @tianma_if 首发 · 61h 前
githubxsocialboard
+243★/d 早期·低活动 official #16
A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings
🔺 @MaciejLukianski 首发 · 41h 前
githubxboard
+810★/d 早期·低活动 official #3
354 条循证建议,覆盖长寿防病、急救、省钱理财、法律红线、失业与工伤、医保社保、恋爱婚育、出国与技能。每条写明成本、收益、证据等级和原始出处,只引期刊论文与官方文件。
🔺 @ForestGrahxu 首发 · 66h 前
🔥 动量榜
| # | Repo | 7d | +1d★ | 7d★ | 质地 | 官方 | X |
|---|---|---|---|---|---|---|---|
| 1 | zai-org/ZCode 🆕 new TypeScript |
+2,273 | 2,273 | 早期·低活动 | – | 29× | |
| 2 | NandhaKishorM/laya ↺ 2d Python 🔺 @clxymox 首发 · 13h 前 |
+1,162 | 5,126 | 早期·低活动 | #1 | 14× | |
| 3 | browser-use/jev-ultrafast ↺ 2d Python 🔺 @betterhn20 首发 · 16h 前 |
+906 | 12,513 | 早期·低活动 | #2 | 18× | |
| 4 | eternity4719/HowToLiveBetter ↺ 2d HTML · 354 条循证建议,覆盖长寿防病、急救、省钱理财、法律红线、失业与工伤、医保社保、恋爱婚育、出国与技能。每条写明成本、收 🔺 @ForestGrahxu 首发 · 66h 前 |
+810 | 9,128 | 早期·低活动 | #3 | 21× | |
| 5 | Albert-Weasker/niubigeo 🆕 new TypeScript |
+756 | 873 | 早期·低活动 | – | 3× | |
| 6 | google/ax ↺ 1d Go · Google's open agentic orchestrator |
+698 | 1,583 | 早期·低活动 | #6 | 11× | |
| 7 | mizorewww/laya-mlx ↺ 1d Python · Native MLX runtime for Laya typed decision models — 7–14 ms |
+329 | 1,959 | 早期·低活动 | #4 | 5× | |
| 8 | yi1108/printfilm ↺ 2d Python |
+321 | 1,401 | 早期·低活动 | – | 3× | |
| 9 | deepseek-ai/deepseek-harness ↺ 2d TypeScript 🔺 @the_osps 首发 · 31h 前 |
+282 | 7,503 | 早期·低活动 | – | 10× | |
| 10 | bojieli/ai-agent-book 🆕 new Python |
+274 | 2,292 | 活跃开发 | – | 10× | |
| 11 | stablyai/orca ↺ 2d TypeScript · Orca is the ADE for working with a fleet of parallel agents. 🔺 @tianma_if 首发 · 61h 前 |
+246 | 5,092 | 活跃开发 | #15 | 22× | |
| 12 | cloudflare/security-audit-skill ↺ 2d JavaScript · A coding-agent skill for multi-phase security audits with in 🔺 @MaciejLukianski 首发 · 41h 前 |
+243 | 14,319 | 早期·低活动 | #16 | 34× | |
| 13 | BeastStokerSouk/Better-Discord 🆕 new | +240 | 684 | 早期·低活动 | – | – | |
| 14 | CraterSerpentGlow/Discord-Server-Raider 🆕 new | +238 | 701 | 早期·低活动 | – | – | |
| 15 | signdirectorvent/Microsoft-365-setup 🆕 new | +226 | 695 | 早期·低活动 | – | – | |
| 16 | OrbitCrystalLocate/Windows-Optimizer 🆕 new | +222 | 471 | 早期·低活动 | – | – | |
| 17 | PotterScissors/Ads-Blocker 🆕 new | +221 | 480 | 早期·低活动 | – | – | |
| 18 | affaan-m/ECC ↺ 2d JavaScript 🔺 @BlockInsight214 首发 · 10h 前 |
+212 | 5,534 | 活跃开发 | – | 21× | |
| 19 | QwenLM/Qwen-Image-2.1 🆕 new Python |
+207 | 542 | 早期·低活动 | – | 5× | |
| 20 | alibaba/open-code-review ↺ 2d Go 🔺 @shao__meng 首发 · 71h 前 |
+204 | 12,964 | 活跃开发 | – | 23× |
🗞️ Hacker News
- Cloudflare Quick Tunnels 594p · 253c
- Android 17 is the first since 3.x to add new APIs without releasing to the AOSP 568p · 270c
- OpenJev 562p · 247c
- ChatGPT now knows what you do on other websites via ad collector 545p · 302c
- Claude Code now reads AGENTS.md if there is no Claude.md 531p · 188c
- How to Write with an LLM 401p · 276c
- Samsung is expected to more than double output of its HBM4 and HBM4E DRAM 331p · 210c
- Human brain is two separate organs, Stanford Medicine-led research finds 288p · 110c
- Exfiltrate Your Weights 269p · 105c
- Inside ZCode: Silently uploading your Git history to the cloud 259p · 93c
💬 V2EX
- [程序员] 哪位老哥的公司编码已经进化到这种程度了吗,感觉都是穷途末路 V2EX · 2026-09-21
- [生活] 吐槽下今年的校招就业环境 V2EX · 2026-09-21
- [职场话题] 中秋国庆连休,中间 3 天给个啥理由请假调休比较好? V2EX · 2026-09-21
- [程序员] 全站开发偏前端,找资源对接 V2EX · 2026-09-21
- [AirPods] Airpods5 到手,说下感受 V2EX · 2026-09-21
- [程序员] 程序员要不要放弃对设计的把控 V2EX · 2026-09-21
- [酷工作] 卖完整的交易所源码 V2EX · 2026-09-21
- [汽车] 油车怎么选 V2EX · 2026-09-21
- [MacBook Pro] macbookpro m4pro 才一年 7 个月电池健康就掉到 89% V2EX · 2026-09-21
- [程序员] 智谱唐杰亲自投诉网友造谣侵权 V2EX · 2026-09-21
- [问与答] 还有流量卡嘛?怎么感觉被一锅端了? V2EX · 2026-09-21
- [推广] 摸鱼偷偷发点小福利,流量 CDK V2EX · 2026-09-21
- [分享创造] 编程只剩架构和约束了 V2EX · 2026-09-21
- [问与答] 有没有可能自己训练一个 jev 结合自己的私有数据,帮助自己做决策 V2EX · 2026-09-21
- [分享创造] Jev 的边界 V2EX · 2026-09-21
- [香港] 在去香港的高铁上求推荐! V2EX · 2026-09-21
- [分享创造] [分享创造] 做了一款轻量好用的域名与网络诊断工具: Domain Checker(内附兑换码福利) V2EX · 2026-09-21
- [分享创造] 开源了一个 Ai 3D 通用快速建模平台 aigccat 希望大家点评一下 V2EX · 2026-09-21
- [职场话题] 最近后端的工作给我找麻了 V2EX · 2026-09-21
- [问与答] 国内 esim 一年了,实际体验如何 V2EX · 2026-09-21
- [Apple] macos27 cursor 四指下滑的最近项目没了 V2EX · 2026-09-21
- [投资] 检查了下基金收益,发现只有纳指是正的 V2EX · 2026-09-21
- [职场话题] 岗位推进都有大半个月了,对面是不是鸽了? V2EX · 2026-09-21
- [投资] 上班摸鱼盯盘插件,爱盯盘重大更新 V2EX · 2026-09-21
- [程序员] 分享个输出 GitHub 每日趋势中文翻译 html 的 skill V2EX · 2026-09-21
- [问与答] 我不知道 pi agent 在原因,还是 deepseek 的原因,有点 gpt 很久之前那味了 V2EX · 2026-09-21
- [问与答] 三千块的人体工学椅,午睡躺着时候,头枕断了,差点闪到脖子 V2EX · 2026-09-21
- [生活] 家里水压低,如何解决? V2EX · 2026-09-21
- [电动汽车] 吐槽特斯拉 V2EX · 2026-09-21
- [OpenAI] openai 啥时候能回复订阅 20X 的 V2EX · 2026-09-21
🐧 LINUX DO
- 求佬推荐人体工学椅!! LINUX DO · 2026-09-21
- CDK站的社区分数是会掉的吗? LINUX DO · 2026-09-21
- 我不知道 pi agent 的原因,还是 deepseek 的原因,有点 gpt 很久之前那味了 LINUX DO · 2026-09-21
- 想开「中银香港」的门槛是什么 LINUX DO · 2026-09-21
- 男子造谣“宁德时代基地班长不让员工上厕所”被行拘 LINUX DO · 2026-09-21
- 邪修教你开通GPT Pro 20X LINUX DO · 2026-09-21
- 帖子被审核是怎么回事? LINUX DO · 2026-09-21
- 微信聊天记录 Mac 和 Windows PC 同步问题 LINUX DO · 2026-09-21
- 求助佬友,想要开一个claude pro应该如何使用才能比较安全 LINUX DO · 2026-09-21
- 求助各位佬,这种菲区VISA直充可以激活GPT美国大学生四个月优惠吗 LINUX DO · 2026-09-21
- 二手平台闲鱼被爆暗藏涉黄产业链 涉未成年女性 LINUX DO · 2026-09-21
- 在线等个视频推荐,要去中专学校去做个ai相关演讲,等个能吸引眼球的演示视频,各位佬给个推荐! LINUX DO · 2026-09-21
- Gemini 4 pro 疑似灰测?(纯猜测) LINUX DO · 2026-09-21
- O/ 降智一点不带演的啊?路由到terra? LINUX DO · 2026-09-21
- Scaleout去中心化的 AI 驱动学习技术使无人机能够自主识别和攻击战场目标 LINUX DO · 2026-09-21
- 帖子越来越少了,大家都放假出去玩了吗 LINUX DO · 2026-09-21
- 这就是A股2026921 LINUX DO · 2026-09-21
- 残缺版iPhone 15,背屏被车🚗碾压了,坐标广州,有没有大佬提供一下维修经验 LINUX DO · 2026-09-21
- Astra-medium,官方team订阅,有点不伦不类的感觉 LINUX DO · 2026-09-21
- 一个疑问,大家JEV为啥都用官方渠道 LINUX DO · 2026-09-21
- 寻求趣丸AI Native全栈研发实习生面经 LINUX DO · 2026-09-21
- 这几天gpt 5.6 terra max格外的慢 LINUX DO · 2026-09-21
- 比较Astra 6High和Kimi K3有感 LINUX DO · 2026-09-21
- where is grok4.7 btw LINUX DO · 2026-09-21
- 动漫一斩苍穹真好看! LINUX DO · 2026-09-21
- 请教各位佬 48team的支付方式 LINUX DO · 2026-09-21
- 感觉 Opus 4.6 还是神😂 LINUX DO · 2026-09-21
- 【开源推广】同时服务几家公司,内网 VPN 天天打架 —— 自己做了个分流网关 LINUX DO · 2026-09-21
- Astra是好模型,但O÷不是好公司 LINUX DO · 2026-09-21
- 目前内存价格高位下Mac mini和DIY装机哪个性价比高 LINUX DO · 2026-09-21
- Where are the current GPU VRAM sweet spots? r/LocalLLaMA · 2026-09-21
- mini-AGI - dynamically grown (530M params currently and growing) continual learning model trained from scratch on 8GB VRAM laptop from batch-1 stream of data. r/LocalLLaMA · 2026-09-21
- Foulmouth Qwen 3.8 27b, an unexpected thought process... r/LocalLLaMA · 2026-09-21
- convaiinnovations/laya (multilingual, non-autoregressive System 1 decision model) r/LocalLLaMA · 2026-09-21
- ZCode is now open source r/LocalLLaMA · 2026-09-21
- You can use any LLM just like JEV r/LocalLLaMA · 2026-09-21
- Seeing how differently people prompt LLMs is funny r/LocalLLaMA · 2026-09-20
- Do NOT trust StepFun's Plan subscriptions., They stole >$100 from me with no warning, and I have not heard back from support at all. r/LocalLLaMA · 2026-09-20
- Wow... thank goodness for open weights!!! r/LocalLLaMA · 2026-09-20
- DIY Jev r/LocalLLaMA · 2026-09-20
- Success running Qwen 3.8 27B EXL3 on RTX 3060 + 5060 Ti r/LocalLLaMA · 2026-09-20
- Would you buy a Qwen3.8-27B Taalas chip for $1k if it could run at 7,000 TPS? r/LocalLLaMA · 2026-09-20
- Speed-up Kimi K3(2.8T) on a 16x GB10 Cluster — 30 t/s coding throughput, 136 t/s concurrency peak. r/LocalLLaMA · 2026-09-20
- Ling 3.0 Tiny vs. Gemma 26B-A4B MoE r/LocalLLaMA · 2026-09-20
- Postulate: Context compaction is good for Quantized models r/LocalLLaMA · 2026-09-20
- Qwen3.8-Flash-Next Cosmic Arcade oneshot slop game r/LocalLLaMA · 2026-09-20
- The bear can dance: Qwen 3.8 27B on one 3090 for 3 weeks r/LocalLLaMA · 2026-09-20
- One more 'you should try ExllamaV3/exl3 for flash next' appreciation post r/LocalLLaMA · 2026-09-20
- Lawsuit says Anthropic, OpenAI, SpaceXAI and Google made illegal agreement on AI slowdown r/LocalLLaMA · 2026-09-20
- laya.cpp: Optimized laya near-instant decision making r/LocalLLaMA · 2026-09-20
- A Jev-style model fine-tuned on Qwen3.5 4B r/LocalLLaMA · 2026-09-20
- CUDA: enable sparse fa for qwen4 by am17an · Pull Request #28770 · ggml-org/llama.cpp r/LocalLLaMA · 2026-09-20
- Harness to do lists - Model problem or plugin problem? r/LocalLLaMA · 2026-09-20
- I tested 9 LLMs on the exact same web-dev prompt for ~8 hours — RTX 3060 12GB results (Rate the best!) r/LocalLLaMA · 2026-09-20
- focus-llama: a llama.cpp fork implementing Declarative Attention (arXiv:2609.02737) r/LocalLLaMA · 2026-09-20
- Qwen-Image-2.1 released! r/LocalLLaMA · 2026-09-20
- Is Typesafe based/derived from work done by the Laya author? r/LocalLLaMA · 2026-09-20
- My Qwen 3.8 27B tests on limited VRAM (16-20GB) r/LocalLLaMA · 2026-09-20
- rene98c/Step-5-Preview-BF16 • HuggingFace (Fork) r/LocalLLaMA · 2026-09-20
- What is JEV and what is it used for? r/LocalLLaMA · 2026-09-20
🛠️ 技术源 · GitHub Trending/Lobsters
- AI chatbots give wrong answers to financial queries 'most of the time' Hacker News Front Page · 2026-09-21
- Winning the Visa Lottery Hacker News Front Page · 2026-09-21
- Deterministic Core, Non-Deterministic Shell Hacker News Front Page · 2026-09-21
- A study of sequence weighting at scale Lobsters · 2026-09-21
- Why back propagation goes backward Hacker News Front Page · 2026-09-21
- Amiga Unix, Again Hacker News Front Page · 2026-09-20
- DAPO: An Open-Source RL System from ByteDance Seed and Tsinghua Air Hacker News Front Page · 2026-09-20
- Roku launches open-source Roku LT OS for creative programmers Lobsters · 2026-09-20
- Bot-free self-hosted analytics with GoatCounter on NixOS Lobsters · 2026-09-20
- What Happened to the Snowden Archive Hacker News Front Page · 2026-09-20
- Google's Open Agentic Orchestrator Hacker News Front Page · 2026-09-20
- Bill to Ban Private Equity from Owning Medical Practices Hacker News Front Page · 2026-09-20
- Deterministic Core, Non-Deterministic Shell Lobsters · 2026-09-20
- Nobody pays for FOSS, we can force them to Hacker News Front Page · 2026-09-20
- Ogre Battle 64 Recompiled Project at 99.05% Hacker News Front Page · 2026-09-20
- Are we really going to use the same Desktop UX forever? Lobsters · 2026-09-20
- Why MCP Was Always a Bad Idea? Hacker News Front Page · 2026-09-20
- The Hierarchy of Money Hacker News Front Page · 2026-09-20
- Adversarial examples for fast hash functions Lobsters · 2026-09-20
- Software Sandboxing: The Basics (2025) Hacker News Front Page · 2026-09-20
- A Necessary History of the Oddest Letter: W Hacker News Front Page · 2026-09-20
- I turned Jev into a (lousy) chatbot Hacker News Front Page · 2026-09-20
- ChatGPT now knows what you do on other websites via ad collector Lobsters · 2026-09-20
- Unix Year 2038 problem and the art of underestimating Lobsters · 2026-09-20
- Samsung is expected to more than double output of its HBM4 and HBM4E DRAM Hacker News Front Page · 2026-09-20
- Resident Evil 4 (GameCube) – complete byte-identical decompilation to C/C++ Hacker News Front Page · 2026-09-20
- Trying the Software Factory Pattern Hacker News Front Page · 2026-09-20
- Vim's UserGettingBored autocmd Lobsters · 2026-09-20
- An actively maintained and updated Motif fork actually exists Lobsters · 2026-09-20
- Show HN: Radius – A Meetup.com Alternative Hacker News Front Page · 2026-09-20
📥 AI 博客 · Newsletter
- Quoting voxium Simon Willison · 2026-09-20
- llm-keys-ui 0.1 Simon Willison · 2026-09-20
- datasette-explain 0.2.2 Simon Willison · 2026-09-20
- datasette-auth-github 1.0 Simon Willison · 2026-09-19
- California Sea Lion, Brandt's Cormorant Simon Willison · 2026-09-19
- Where I stand on RSI Interconnects (Nathan Lambert) · 2026-09-19
- [AINews] Here are 6 Clones of Jev in 2 days Latent Space · 2026-09-19
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI Simon Willison · 2026-09-18
- Note on 18th September 2026 Simon Willison · 2026-09-18
- Quoting Thariq Shihipar Simon Willison · 2026-09-18
- MilleMiglia: A realistic instance generator for middle-mile logistics Google Research · 2026-09-18
- The Creative Spirit of Who Framed Roger Rabbit Simon Willison · 2026-09-18
- Introducing the Australian Youth Safety Blueprint OpenAI News · 2026-09-18
- [AINews] not much happened today Latent Space · 2026-09-18
- Be alert: targeted attacks on prominent Rustaceans Simon Willison · 2026-09-17
- How To Write With An LLM Simon Willison · 2026-09-17
- Self-generated prompt injections in compaction summaries Simon Willison · 2026-09-17
- The future of practice: Enabling teachers to create learning interactives with generative UI Google Research · 2026-09-17
- How Cooley is accelerating IPO work with ChatGPT OpenAI News · 2026-09-17
- [AINews] Reality Checks on AI News (Yegge shuts down Gas Town, Databricks’ +60% Astra cost) Latent Space · 2026-09-17
- Introducing Astra for Law OpenAI News · 2026-09-17
- datasette 1.0a40 Simon Willison · 2026-09-16
- datasette 0.65.5 Simon Willison · 2026-09-16
- Claude Cowork and chat are now one Claude Simon Willison · 2026-09-16
- Underwriting Superintelligence: Backing Agents you can Sue — Rune Kvist, AIUC Latent Space · 2026-09-16
- Our framework for reporting model misalignment OpenAI News · 2026-09-16
- Quoting Mustafa Suleyman Simon Willison · 2026-09-16
- Helping older adults use AI in everyday life OpenAI News · 2026-09-16
- Reimagining advertising with AI OpenAI News · 2026-09-16
- Hex turns complex analysis into visual reports with GPT‑6 Astra OpenAI News · 2026-09-16
📢 电报精选
- AI_News_CN 中国也有AI焦虑:全民拥抱新事物背后 工作岗位和社会稳定面临新挑战 via cnBeta.COM - 中文业界资讯站 (author: 稿源:大西洋周刊)
- AI_News_CN CRISPR基因编辑疗法成功让低密度脂蛋白胆固醇降低一半并维持一年 via cnBeta.COM - 中文业界资讯站 (author: 稿源:cnBeta.COM)
- landiansub #人工智能 美国总统特朗普认为人工智能这个术语不够准确,应该将其更名为超级智能、极端智能或至高智能。 人工智能这个术语最初是约翰麦卡锡在 1955 年提出的,写下这个词主要是与当
- zaihuapd 《流浪地球 3》上部计划 2027 年春节档上映 9 月 21 日,中国电影在 2026 年半年度业绩说明会上表示,公司主控出品的《流浪地球 3》(上部)计划于 2027 年春节
- AI_News_CN RSA-896在Claude帮助下已被破解 2026年9月19日,Anthropic工程师Stephen A. Weis在Claude的协助下,公布了RSA-896的质因数分解结
- AI_News_CN [GIF] Spec Kit —— 给 AI 编程助手结构化指令 Spec Kit —— 给 AI 编程助手结构化指令 🤖 GitHub 最近发布了 Spec Kit,旨在解
- AI_News_CN ↩️ 中国加强监管放缓人形机器人 IPO 在花🎗️科技圈: 宇树科技股价较首日高点回撤 45%,市值累计蒸发 2008 亿元 中国人形机器人龙头宇树科技在科创板首日开盘大涨
- landiansub #网站应用 西班牙文化部下令封禁网页快照网站 Archive[.]today,理由是该网站保存未经授权的内容。 Archive[.]today 确实可以绕过某些网站的付费墙保存内
- zaihuapd 宇树科技股价较首日高点回撤 45%,市值累计蒸发 2008 亿元 中国人形机器人龙头宇树科技在科创板首日开盘大涨 629.44%,报 1100 元,总市值飙升至 4449 亿元。但
- CE_Observe 被质疑“偷传代码”后智谱 ZCode 官宣开源:已完成整改并向所有用户道歉 - IT之家 https://m.ithome.com/html/1005046.htm Itho
- zaihuapd 特斯拉人形机器人团队在长三角审厂 产业链人士称,特斯拉人形机器人团队上周已在宁波走访拓普集团、三花智控、均胜电子等供应商,开展产线质量与合规审厂。这些企业已获得相关订单,通过评估
- CE_Observe 知名网盘Dropbox修改服务协议 未开会员且超过6个月未登录就会被删号 https://www.landian.news/archives/126968.html 根据最新服
- AI_News_CN 中国大模型周调用量达67.46万亿Token,连续21周领先美国 据OpenRouter最新数据,9月14日至20日,全球AI大模型总调用量达129万亿Token,环比增长1.5
- landiansub #人工智能 [已修复] 安全研究员披露 Codex 桌面版和 CLI 中的安全漏洞,包括沙盒逃逸并执行高危命令和在工作区之外写入文件等。 桌面版的漏洞是利用共享内存的权限令牌伪造
- landiansub #豆包 广告图,亮点自寻 😁订阅 🤔解封 🥰推特 🎉CN2VPS 😁50🤡11😇3👏2
- zaihuapd 🐟 闲鱼涉黄产业链曝光:最小涉事者仅 14 岁 闲鱼平台再次被曝暗藏涉黄产业链,部分涉事服务涉及未成年人,最小年龄仅 14 岁。商家以“模特上门约拍”等旗号发布商品,通过隐晦暗
- landiansub #软件资讯 知名网盘 Dropbox 修改服务协议,未开会员且超过 6 个月未登录就会被删号,封号机制现在采用连坐制。 新协议将在明年 1 月 1 日生效,在旧版协议中连续 12
- AI_News_CN Astra —— 用 9 张照片复刻你的房间 📸 Roberto Nickson (rpnickson) 最近分享了一个让他觉得“荒谬”的体验:他用 Astra 这个工具,给自
- CE_Observe 迪士尼打击 Disney+ 帐号共享,扩大至北美及亚太市场 https://www.ithome.com.tw/news/165200 Disney+宣布所有套餐都会包含广告,屏
- zaihuapd ZCode 就索引功能数据上传问题致歉 ZCode 今日就数据问题致歉,称源于“代码库索引”功能,Repo Wiki 生成页面时可能触发仓库数据上传;云端生成后数据立即销毁,不保存
- landiansub #安全资讯 6 啊!谷歌安全研究员卧底到 TeamPCP 团队,从内部收集信息并进行破坏行为,而且还收集黑客身份协助执法机构抓捕。 这名卧底分析师所做的工作包括:收集攻击信息、收
- AI_News_CN 阶跃甩出 Step 5 Preview:600B 稀疏 MoE 只激活 27B,把开源旗舰的性价比边界往外推了一格 via AI新闻资讯 (author: AI Base)
- AI_News_CN OpenAI 一次性摊开六份失准报告:模型越界不再是偶然,而是三种可复现的机制 OpenAI 近日发布了一套模型失准披露框架,并附上六份真实报告,把自家模型在任务执行中出现的多种
- AI_News_CN 用《我的世界》测 GPT-6 Astra:141 小时直播,一只苦力怕让 AI 陷入“种土豆”循环 AI 评测机构 Vals AI 利用游戏《我的世界(Minecraft)》对
- AI_News_CN Gemini 也"越狱"了:谷歌确认 5 月安全测试中闯入 3 家真实公司系统,同一家评估商的锅 via AI新闻资讯 (author: AI Base)
- AI_News_CN 黄仁勋驳斥 AI 末日叙事:吓唬人不负责任,2030 年是世界末日概率为 0% 英伟达联合创始人兼首席执行官黄仁勋昨日接受 CBS News 采访时表示,他不认同某些研究员提出的
- landiansub #行业资讯 欧盟委员会提出《儿童法案》,计划开发通用型年龄验证工具,像是社交网站等必须对用户验证年龄以判断用户是否可以使用。 欧盟提议禁止 13 周岁以下儿童使用社交媒体,15
- AI_News_CN Gemini Notebook迎返校更新:讲座录音、AI复习和60秒视频一站完成 谷歌为Gemini Notebook推出一系列返校功能,新增讲座录音、互动式学习概览和短视频概览
- AI_News_CN OpenAI研究员放话"物理断网也拦不住 AI":BitWhisper 用 CPU 温度传信,每小时只能淌 8 个比特 OpenAI 和 Anthropic 前几日刚带头呼吁给
- zaihuapd 前 npm 主管提议:软件仓库向企业收费并分成给维护者 曾任 npm 首席执行官的 Laurie Voss 发文提议,让 npm、PyPI、Docker Hub 等软件注册表向企
📚 科技周刊 新项目/工具自荐
- 【开源自荐】 joLink:为 AI 编程助手提供 Java 增量编译、测试与运行调试 2 repos
- 【投稿】人工智能与人脑
- 【开源自荐】Prism:跨平台社交内容分析框架,按收藏率而非点赞排序 1 repos
- 【网站自荐】免费在线MP3转MIDI工具,MelodyTrace 让声音与音符自由互译 1 repos
- 【开源自荐】原生多租户治理,集“资产CMDB + 自动化工单 + 分布式作业”于一体的企业级运维协同底座 6 repos
- 【开源自荐】Data Asset Portal:面向数仓团队的轻量数据资产目录 1 repos
- 开源项目推荐:FocusFlow – 将复杂的系统架构大图转化为 60fps 电影级运镜故事 1 repos
- 【开源自荐】PI-Desktop:本地优先的 AI 编程 Agent 桌面工作区 1 repos
- 【开源自荐】OpenAI4S:9.9 元的豆包 API,跑一个会自己写代码做科研的 AI 助手 1 repos
- 【开源自荐】京张向上 JINGZHANG RISING:首个 AI Agent 深度参与的 43.6 km² 真实城市设计全流程开源方案 1 repos
- US Address Generator:面向开发测试的美国地址生成器,支持地区筛选和 JSON 导出 1 repos
- [自荐] EasyDomain:结合网站导航、同类工具发现和域名资料查询的入口
- 【开源自荐】清鸽LocalAI:离线、本地、保护隐私的移动端侧LLM应用 1 repos
- 推荐开源项目:OmniGit - 拥有 IntelliJ IDEA 体验与 3-Way Merge 的轻量 Git 客户端 1 repos
- 【开源自荐】Jev Social:让 Jev 决定下一步社交媒体研究操作 1 repos
- [自荐] 供应链工具箱 (Supply Chain Toolkit):基于 Tauri + Rust 的离线桌面库存决策工具 1 repos
- 【开源自荐】TLSFlow:应对短周期证书轮换的资产、部署与回滚平台 1 repos
- 【网站自荐】免费在线 AI 辅助阅读《史记》等中华经典古籍 1 repos
🎯 Alpha 账号 X 上最早带火仓库的人
| 作者 | leads | lead率 | 仓库数 |
|---|---|---|---|
| @shanyanggm | 22 | 0.71 | 23 |
| @xzbx888 | 13 | 0.87 | 8 |
| @shaw_stone73832 | 8 | 0.67 | 7 |
| @GitTrend0x | 8 | 0.73 | 8 |
| @LFrefman | 8 | 0.47 | 9 |
| @the_osps | 7 | 0.88 | 6 |
| @xfubot | 7 | 0.54 | 8 |
| @the_rza_ | 7 | 0.47 | 7 |
| @Sn0wbrave | 6 | 0.55 | 6 |
| @iasg1004 | 6 | 0.3 | 14 |
| @FrontieraTechIT | 5 | 0.63 | 7 |
| @bilawalsidhu | 5 | 1 | 1 |
| @DataChaz | 5 | 0.83 | 5 |
| @RepoGems | 5 | 0.83 | 4 |
| @neil_xbt | 5 | 0.63 | 3 |
| @shao__meng | 5 | 0.56 | 6 |
| @key_indie | 5 | 0.83 | 2 |
| @vintcessun | 5 | 0.31 | 13 |
| @clxymox | 5 | 0.36 | 8 |
| @LoveAIbrain | 5 | 1 | 4 |
| @jasontopia | 4 | 0.8 | 3 |
| @seekjourney | 4 | 0.67 | 3 |
| @ClaudeCodeLog | 4 | 1 | 1 |
| @0x_Kratos | 3 | 1 | 1 |
| @fakeWow_ | 3 | 0.75 | 2 |
🏆 各领域最强模型
文本 / 对话
Claude Fable 5.1
Anthropic · 53.4
图像生成
GPT Image 2.5 Flare
OpenAI · 1188
图像编辑
GPT Image 2.5 Sunburst
OpenAI · 1176
文生视频
Wan 3.0
Alibaba · 1336
图生视频
Gemini Omni Flash
Google · 1369
语音合成
Sonic 3.6
Cartesia · 1276
🏆 能力排行榜 Artificial Analysis
文本 / 对话 Intelligence Index
- 1Claude Fable 5.153.4
- 2GPT-6 Astra52.7
- 3Claude Opus 550.8
- 4Claude Fable 549.6
- 5Muse Spark 1.348.1
- 6GPT-5.6 Sol47
- 7Qwen3.8 Max45.4
- 8GLM-5.344.8
- 9Grok 4.644.3
- 10Step 5 Preview43.7
- 11Kimi K343.6
- 12GPT-5.6 Terra42.1
图像生成 Text→Image Arena Elo
- 1GPT Image 2.5 Flare1188
- 2GPT Image 2.5 Sunburst1182
- 3GPT Image 21171
- 4Grok Imagine Image 2.01154
- 5MAI-Image-2.61147
- 6Reve 2.11129
- 7Nano Banana 21122
- 8Muse Image1111
- 9GPT Image 1.51102
- 10MAI-Image-2.51102
- 11Nano Banana Pro1100
- 12MAI-Image-2.6-Flash1099
图像编辑 Image-Editing Arena Elo
- 1GPT Image 2.5 Sunburst1176
- 2GPT Image 2.5 Flare1155
- 3MAI-Image-2.61132
- 4MAI-Image-2.6-Flash1122
- 5GPT Image 21121
- 6Muse Image1115
- 7MAI-Image-2.51113
- 8MAI-Image-2.5-Pro1106
- 9Seedream 5.0 Pro1106
- 10Nano Banana 21105
- 11Grok Imagine Image 2.01104
- 12GPT Image 1.51104
文生视频 Text→Video Arena Elo
- 1Wan 3.01336
- 2Gemini Omni Flash1330
- 3MiniMax H31302
- 4HappyHorse-1.01287
- 5HappyHorse-1.11272
- 6Dreamina Seedance 2.0 720p1259
- 7Wan2.7-2606121243
- 8grok-imagine-video1235
- 9Kling 3.0 Omni 1080p1230
- 10PixVerse V5.61230
- 11PixVerse V61230
- 12Kling 3.0 1080p1230
图生视频 Image→Video Arena Elo
- 1Gemini Omni Flash1369
- 2Wan 3.01361
- 3Bach 1.0 Pro1359
- 4MiniMax H31354
- 5PixVerse V61337
- 6Dreamina Seedance 2.0 720p1336
- 7grok-imagine-video-1.51329
- 8grok-imagine-video1326
- 9HappyHorse-1.11312
- 10Kling 2.5 Turbo 1080p1296
- 11HappyHorse-1.01293
- 12Vidu Q3 Pro1290
语音合成 Text→Speech Arena Elo
- 1Sonic 3.61276
- 2Qwen-Audio-3.0-TTS-Plus1260
- 3Realtime TTS-21247
- 4Simba 3.21240
- 5Luna TTS1231
- 6Realtime TTS-2 Flash1215
- 7StepAudio 2.5 TTS1209
- 8Breeze TTS 21205
- 9Gemini 3.1 Flash TTS1201
- 10v3 Conversational1197
- 11Sonic 3.51184
- 12Lightning V3.1 Pro1179
🥇 综合能力榜 Benchmark Heaven · 7 榜合一(AA+Epoch ECI+DesignArena)
| # | 模型 | 综合分 | $/1M | Benchmaxxing |
|---|---|---|---|---|
| 1 | Claude Fable 5.1 Anthropic 🏅前沿 | 98.4 | $13.64 | 0.66 |
| 2 | GPT 6 Astra OpenAI | 97.7 | $13.64 | -6.38 |
| 3 | Claude Opus 5 Anthropic | 96.3 | $6.82 | -6.04 |
| 4 | Claude Fable 5 Anthropic | 95.6 | $13.64 | -4.42 |
| 5 | Muse Spark 1.3 Meta 🏅前沿 | 89.3 | $1.52 | 5.84 |
| 6 | GPT 5.6 Sol OpenAI 🏅前沿 | 95.2 | $2.73 | -2.57 |
| 7 | Qwen3.8 Max 0902 Alibaba | 87.7 | $2.36 | — |
| 8 | GLM 5.3 Z.ai 开源 🏅前沿 | 85.9 | $1.01 | 1.65 |
| 9 | Grok 4.6 xAI | 84.4 | $2.36 | 5.48 |
| 10 | Step 5 Preview StepFun 🏅前沿 | 87.4 | $1.15 | — |
| 11 | Kimi K3 Moonshot AI 开源 🏅前沿 | 90.2 | $2.32 | 3 |
| 12 | GPT 5.6 Terra OpenAI | 82.2 | $2.91 | -3.58 |
| 13 | GLM 5.3 Flash Z.ai 开源 🏅前沿 | 77.5 | $0.09 | -1.11 |
| 14 | Claude Opus 4.8 Anthropic | 88.4 | $6.82 | -2.4 |
| 15 | Gemini 3.8 Flash Google | 81.4 | $1.02 | 6.34 |
Benchmaxxing:BH 的指标,+ 值越高表示该模型在"公开基准"上比"不可训练的封闭题"排名更靠前(BH 定义与计算,非本站判断)。
💰 性价比 / 最省钱 跨 provider 最低价 · 10:1 blended · 综合分≥60
| 模型 | $/1M | 综合分 | 最便宜 provider |
|---|---|---|---|
| DeepSeek V4 Flash 0731 🏅前沿 🔒不训练 开源 | $0.04 | 74.4 | Relace |
| Agnes 3.0 Flash | $0.06 | 71.2 | |
| GLM 5.3 Flash 🏅前沿 开源 | $0.09 | 77.5 | GMICloud |
| Agnes 2.5 Pro Beta | $0.12 | 71.2 | |
| DeepSeek V4.1 Flash 🔒不训练 开源 | $0.17 | 69.8 | Relace |
| Qwen3.8 Flash Next 🏅前沿 开源 | $0.18 | 82 | Alibaba |
| Qwen3.8 27B 开源 | $0.25 | 69.6 | Darkbloom |
| GPT 5.6 Luna 🔒不训练 | $0.29 | 79.6 | Azure AI Foundry |
| MiniMax M3 🔒不训练 开源 | $0.3 | 60 | CoreWeave |
| Solar Pro 4 | $0.38 | 63.1 | |
| DeepSeek V4 Pro 开源 | $0.46 | 62.3 | StreamLake |
| Agnes 2.5 Pro Alpha 开源 | $0.49 | 62.2 | |
| DeepSeek V4 Flash Vision | $0.52 | 66.4 | |
| Inkling Small 🔒不训练 开源 | $0.52 | 64 | DeepInfra |
| Apodex 1.1 | $0.55 | 71.2 |
跨 95 家 provider(含 21 家 🇪🇺 EU、53 家 🔒不训练)· 数据 2026-09-21
🏅 性价比前沿 没有更便宜的模型能在能力上胜过它们(Pareto)
- Ling 3.0 Flash 综合 54.1 · $0.02/M · Novita · 开源
- DeepSeek V4 Flash 0731 综合 74.4 · $0.04/M · Relace · 开源
- GLM 5.3 Flash 综合 77.5 · $0.09/M · GMICloud · 开源
- Qwen3.8 Flash Next 综合 82 · $0.18/M · Alibaba · 开源
- GLM 5.3 综合 85.9 · $1.01/M · Baidu · 开源
- Step 5 Preview 综合 87.4 · $1.15/M ·
- Muse Spark 1.3 综合 89.3 · $1.52/M · Meta
- Kimi K3 综合 90.2 · $2.32/M · Relace · 开源
- Qwen3.8 Max 综合 91.9 · $2.36/M ·
- GPT 5.6 Sol 综合 95.2 · $2.73/M · OpenAI
- Claude Opus 5 综合 96.4 · $6.82/M · Azure AI Foundry
- Claude Fable 5.1 综合 98.4 · $13.64/M · AWS Bedrock
🧠 最新发布
2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
2026-09-17
Gemini Omni Flash Preview
Google
2026-09-17
Venice Uncensored
Venice
2026-09-17
Nano Banana 2 Lite
Google
2026-09-16
Union Alpha
Stealth
2026-09-15
Jev 1.13
TypeSafe AI
2026-09-12
Schematron V2 Turbo
Inference.net
2026-09-12
Schematron V2 Small
Inference.net
🧠 模型发布时间线
2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
aimlapi
2026-09-17
Gemini Omni Flash Preview
Google
aimlapi
2026-09-17
Venice Uncensored
Venice
aimlapi
2026-09-17
Nano Banana 2 Lite
Google
aimlapi
2026-09-16
Union Alpha
Stealth
aimlapi
2026-09-15
Jev 1.13
TypeSafe AI
aimlapillmgateway2×
2026-09-12
Schematron V2 Turbo
Inference.net
aimlapi
2026-09-12
Schematron V2 Small
Inference.net
aimlapi
2026-09-11
Fugu Ultra v2.0
Sakana AI
llmgateway
2026-09-11
Kimi K2.8 Preview
Moonshot AI
llmstats
2026-09-11
Atria Dawn Preview
Shanghai AI Laboratory
llmstatsllmgateway2×
2026-09-11
Fugu Ultra v2
Sakana AI
aimlapi
2026-09-11
Fugu Max
Sakana AI
aimlapillmgateway2×
2026-09-10 · ★
Ling 3.0 Flash VL
inclusionAI
aimlapiopper2×
2026-09-10 · ★
DeepSeek V4.1 Flash
DeepSeek AI
aimlapillmstatsopperllmgateway4×
2026-09-10
DeepSeek Chat (V4.1 Flash)
DeepSeek AI
aimlapi
2026-09-08
GPT Image 2.5 Sunburst
Open AI
aimlapillmgateway2×
2026-09-08
GPT Image 2.5 Flare
Open AI
aimlapillmgateway2×
2026-09-08
Mercury 2.5
Inception
aimlapi
📄 论文 PwC + arXiv
Meiduo Chong, Shaolei Zhang, Ju Fan · 2026-09-21
Data agents aim to fulfill natural-language instructions over heterogeneous data, including tables, files, and databases. However, data agents face a challenging agent-data gap: heterogeneous data resides outside the agent, while the agent
Shuai Bai, Jiayong Deng, Yikun Fu · 2026-09-21
Computer-use agents (CUAs) have advanced along two separate lines: graphical interaction and software development through code and the command line. Real digital work requires both, interleaved rather than stacked end to end. We study hybri
Yongqi Tong, Pan Wang, Hang Wang · 2026-09-21
Reusable skills give agents transferable procedural knowledge, making scalable acquisition essential for extending agents beyond prior experience. Existing methods face two limitations: trajectory-based synthesis requires interactions with
arXivpwc
Mingyang Chen, Shengdong Chen, Xiaoxiao Fu · 2026-09-21
We introduce Zing-0.5, a 5B autoregressive world model designed for playability: users can explore generated worlds, influence unfolding events, and respond to the resulting feedback through joint keyboard and online text control. Our appro
Zheng-Hui Huang, Guixu Lin, Jiacheng Lin · 2026-09-21
Recent video world models generate increasingly realistic and interactive visual experiences, yet lack reliable mechanisms for maintaining persistent world state and enforcing programmable rules over extended interactions. We introduce Prog
Terminal-Bench-Science 0.1 🔥
Steven Dillmann · 2026-09-21
Terminal-Bench-Science evaluates AI agents on workflows from researchers' own work. Scientists, not model developers or data vendors, set the bar for scientific capability in AI. Terminal-Bench-Science is a benchmark led by researchers at S
Agentspwc
Xin Zhou, Zongchuang Zhao, Zhibo Yang · 2026-09-21
We present Qwen-Drive-1.0, an initial step towards a vision-language foundation model for autonomous driving. Qwen-Drive-1.0 retains the architecture of the pretrained vision-language model (VLM) and integrates 3D perception, visual questio
Wei Zhou, Xiongwei Zhu, Lingdong Kong · 2026-09-21
Reinforcement learning can align diffusion models with human preferences and task-specific objectives, but endpoint rewards do not specify how an intermediate denoising prediction should change. We introduce DiffusionOPSD as an on-policy se
Xiaomin Li, Yuexing Hao, Jianheng Hou · 2026-09-21
Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract away human diversity and interactive behavior. We therefore introduce MatrAIx, a populatio
Xin Cheng, Xingkai Yu, Chenze Shao · 2026-09-21
Speculative decoding accelerates Large Language Model (LLM) inference by decoupling draft generation from target verification. While recent parallel drafters efficiently propose long token sequences in a single forward pass, they suffer fro
arXivpwc
Meng Luo, Yanlin Li, Hao Li · 2026-09-21
Foundation models, alongside advances in learned game-world models, are reshaping AI across the game lifecycle. Beyond playing games, recent systems model players and game dynamics, support design and development, adapt player-facing experi
Jiayin Chen, Yicheng Xu, Muting Wang · 2026-09-21
Iterative reference-conditioned image editing can introduce grid-like and granular textures, commonly described as digital ripple. We present Mi-Ripple, a diagnosis-guided restoration workflow that suppresses this digital ripple while prote
Chengqian Ma, Wei Tao, Haoyu Zhang · 2026-09-21
An avatar that holds a conversation should decide what to say and to move while saying it, yet these abilities live in separate model families: spoken dialogue models produce speech without motion, and co-speech motion models produce motion
Youtian Lin, Yikang Yang, Zhanpeng Hu · 2026-09-21
Native 3D generators now recover impressive mesh geometry from a single image. However, a dense mesh stays soft where a machined object should be sharp, it carries no part decomposition, and it exposes no parameter a user could edit. To add
Kimi Team, Tongtong Bai, Yifan Bai · 2026-09-21
We introduce Kimi K3, a 2.8T parameter Mixture-of-Experts model with 104 billion activated parameters, native vision capabilities, and a 1-million-token context window. Kimi K3 is built on Kimi Delta Attention and Attention Residuals, which
Kairos Team, Fei Wang, Shan You · 2026-09-21
World models are transitioning from passive visual generators to foundational, operational infrastructure for Physical AI: they must natively acquire world knowledge from heterogeneous experience, maintain persistent states over long horizo
Aditi, Niket Agarwal, Arslan Ali · 2026-09-21
We introduce Cosmos 3, a family of omnimodal world models designed to jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-transformers architecture. By supporting highly flexible inpu
Dong-Yang Li, Wang Zhao, Yuxin Chen · 2026-09-21
Recent advances in 3D generative models have rapidly improved image-to-3D synthesis quality, enabling higher-resolution geometry and more realistic appearance. Yet fidelity, which measures pixel-level faithfulness of the generated 3D asset
Junbo Cui, Bokai Xu, Chongyi Wang · 2026-09-21
Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming interaction, yet they still remain far from human-level multimodal interaction. The key bottl
Jiaming Tan, Mingliang Zhai, Zhen Li · 2026-09-20
Interactive video world models must maintain broad scene context under camera motion while producing high-fidelity observations with low latency. Existing approaches face a representation trade-off: perspective models operate on local views
Zhuoyang Qian, Biao Wu, Yiran Wang · 2026-09-20
Turning a research idea into a complete paper requires more than text generation: the system must retrieve literature, design and execute experiments, revise claims according to evidence, produce publication-ready figures, and maintain cons
Inkling-Small 🔥
· 2026-09-20
Inkling-Small is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs. It is intended for use in English and other languages, and across multiple coding languages. The model is designed to
AgentsAudio understandingCoding AgentsImage Understandingpwc
Jian Hu, Huiying Li, Hao Zhang · 2026-09-20
Agentic reinforcement learning research is constant algorithm modification, new estimators, new pipeline stages, new rollout schemes, and in mainstream frameworks each change threads through layers of trainer, distributed backend, and rollo
Yuliang Liu, Zhang Li, Ziyang Zhang · 2026-09-20
Mainstream visual encoders are pretrained on natural images and cannot be effectively applied to document images without document-oriented adaptation, as dense text and fine-grained character strokes demand character-level visual perception
🤗 HF 采用榜 下载/点赞
⚡ System One 决策模型 380 项目 · Jev/TypeSafe 生态 · 快决策(非推理)
SDK & Decision Frameworks 68Security & Guardrails 32High-Frequency & Simulation 31Routing & Cost Optimization 30Browser & OS Action 29Evaluation & Observability 29CLI & Pipelines 25Domain & Vertical Tools 25Data & Search 23Context GC & Filter 22MCP & Integrations 19Codebase & Graph Pathfinding 12Decision Tools 12Creative Tools 11SDK & Integrations 6Voice & Conversation 4Classification & Taxonomy 2
| 项目 | ★ | 类别 | Jev 决策点 |
|---|---|---|---|
| langchain langchain-ai | 146,758 | SDK & Integrations | Submits binary, categorical and ordered-score questions and returns typed answers with probabilities. |
| ai-hedge-fund virattt | 63,633 | Domain & Vertical Tools | Converts strategy questions to System One requests and normalizes native answers to the project’s result format. |
| litellm BerriAI | 59,264 | Routing & Cost Optimization | Maps requests to configured complexity classes that drive backend routing. |
| oh-my-pi can1357 | 32,175 | Routing & Cost Optimization | Sends agent state and typed questions to Jev and parses structured answers. |
| jev-model-router davila7 | 30,860 | Routing & Cost Optimization | Evaluates task tier, reasoning needs and production risk; local policy maps results to invocation settings. |
| composio ComposioHQ | 30,267 | SDK & Decision Frameworks | Turns tool or action conditions into structured questions and passes Jev answers to local invocation logic. |
| ai vercel | 26,867 | SDK & Decision Frameworks | Maps choice, score, and yes/no questions to TypeSafe System One requests and parses typed results. |
| cua trycua | 25,268 | Browser & OS Action | Reads DOM or supported visual-region descriptions and returns a supplied candidate action ID. |
| pydantic-ai pydantic | 20,078 | SDK & Integrations | Converts supported structured output fields into typed Jev questions and maps answers back to the output model. |
| eliza elizaOS | 19,401 | SDK & Decision Frameworks | Only an explicit systemOne call sends state and questions, returning validated typed answers. |
| langchainjs langchain-ai | 18,213 | SDK & Decision Frameworks | Uses invoke to call TypeSafe and parse choice, noul, score and probability fields. |
| json-render vercel-labs | 17,468 | Creative Tools | Evaluates component configurations through Vercel AI Gateway, then composeSpec assembles the UI specification. |
| jev-ultrafast browser-use | 12,593 | Browser & OS Action | Chooses an action and its matching DOM element in one request; a text model generates input text. |
| openchamber openchamber | 10,187 | Routing & Cost Optimization | Jev selects a task category; local category mappings determine the model configuration. |
| rig-typesafeai 0xPlaygrounds | 8,686 | SDK & Decision Frameworks | Sends application state and questions to Jev and parses Choice, Score or Noul answers. |
| firstmate kunchenguid | 6,865 | Routing & Cost Optimization | Sends the task brief and candidate rules to Jev, then resolves execution profiles with confidence and local conditions. |
| fast-jev-compaction tamaratran | 5,380 | Context GC & Filter | Separately judges whether a tool call and its full output are still needed; code keeps, truncates or drops them. |
| agentgateway agentgateway | 4,946 | Security & Guardrails | Jev scores jailbreaks, harmful content and secret disclosure; thresholds or evaluation errors reject requests. |
| latitude-llm latitude-dev | 4,664 | Evaluation & Observability | Judges which checks apply and can add checks when thresholds and rate limits permit. |
| ax ax-llm | 2,928 | SDK & Integrations | Maps supported signatures to Jev questions or sends native System One requests. |
| jev-trader jarrodwatts | 1,586 | Domain & Vertical Tools | In Jev mode, order-book judgments feed code that simulates fills or submits configured post-only limit orders. |
| NanoJev TianyuCodings | 1,526 | High-Frequency & Simulation | Evaluates multiple questions and dynamic candidate spaces concurrently in a single forward pass, logging navigation choices. |
| jev-desktop lahfir | 1,349 | Browser & OS Action | Jev selects a target and action and estimates presence and risk; local policy decides whether to execute. |
| vellum-assistant vellum-ai | 1,293 | MCP & Integrations | Submits state and question bundles to System One and returns structured answers to the Assistant. |
| kev jaredpalmer | 1,171 | High-Frequency & Simulation | Attaches a parallel decision head to an open 0.5B model to answer discrete questions directly from token activations. |
| celesto CelestoAI | 945 | Codebase & Graph Pathfinding | Judges whether a finding was introduced by the change, is supported and merits a fix. |
| atomic bastani-inc | 809 | Routing & Cost Optimization | Sends predefined questions to Jev and decodes answers for callers; regular models still generate code. |
| aiavatarkit uezo | 678 | Voice & Conversation | Assesses utterance completeness and whether the user is likely to continue speaking. |
| kody kentcdodds | 659 | Data & Search | Sends a Score question per candidate, reorders and drops low scores; the model id is typesafe/jev. |
| typesafe-computer-use awlevin | 657 | Browser & OS Action | Selects the next step from deterministically extracted controls and actions before desktop execution. |
| Agent AgentiLoop | 617 | Security & Guardrails | Adds a destructive-risk judgment after local shell checks and refuses commands above the configured threshold. |
| req_llm agentjido | 580 | SDK & Decision Frameworks | Sends state and questions, normalizes answers and retains the raw provider response. |
| omg.dev BennyKok | 532 | Browser & OS Action | Chooses controls and checks completion or blockage before the test runner operates the UI. |
| Jev-cu Sac-Y | 493 | Browser & OS Action | Chooses targets and actions and assesses completion and risk; local policy controls execution or confirmation. |
| foreman thruwire | 441 | CLI & Pipelines | AsyncTypeSafeClient.system_one with default jev-latest sends Noul questions for supervision. |
| vexjoy-agent notque | 421 | Routing & Cost Optimization | After deterministic routing guards, Jev judges the remaining candidates and required workflow components. |
| Jev Review devagrawal09 | 420 | Codebase & Graph Pathfinding | Judges risk, files, evidence regions, mechanisms and severity before rule-based reviewer routing. |
| simple-jev featherless-ai | 406 | SDK & Decision Frameworks | Extracts log-probabilities of candidate tokens from model vocabulary logits, formatting them into standard Jev responses. |
| jev-experiments dabit3 | 345 | Security & Guardrails | Judges risks such as exposed credentials or destructive changes; local rules warn or block a commit. |
| WrongStack WrongStack | 329 | Routing & Cost Optimization | Jev evaluates the task against eligible specialists; local dispatch rules use the result. |
🏅 JevBench 决策榜 JevBench v1 · 242 决策/模型 · 私有留出集 · 准确率 + 校准(Brier↓ 越低越准)
| # | 系统 | 准确率 | Brier↓ | 类型 |
|---|---|---|---|---|
| 1 | Jev 1.13.0 (TypeSafe AI) TypeSafe AI 开源 | 100% | 0.0028 | jev |
| 2 | openjev-sglang (Qwen3.6-35B-A3B on SGLang) ekzhang 开源 | 100% | 0.0305 | jev-rebuild |
| 3 | system-one-open (Gemma 4 E2B LoRA on an L4) mithalouni 开源 | 96% | 0.0851 | jev-rebuild |
| 4 | open-alternative-jev (Qwen3.5-4B, HF Space) IkerMoel 开源 | — | — | jev-rebuild |
| 5 | open-jev-deberta-v3-large (local CPU) Kotoba Labs 开源 | 67% | 0.4769 | jev-rebuild |
| 6 | GPT-5.6 Luna (low reasoning effort) OpenAI 开源 | 100% | 0.0003 | llm-baseline |
| 7 | Gemini 3.1 Flash-Lite Google 开源 | 100% | 0.0017 | llm-baseline |
| 8 | DeepSeek V4.1 Flash (thinking default) DeepSeek 开源 | 100% | 0.0001 | llm-baseline |
| 9 | Qwen3.8 27B (Chutes TEE) Qwen / Chutes 开源 | 100% | 0.0001 | llm-baseline |
JevBench(BH):把"决策模型"(Jev/System One)与通用 LLM 放在同一批决策任务上比准确率与概率校准。
🅱️ B站 AI 无限竞技场 18 模型 · 夺冠率
| # | 模型 | 夺冠率 | 冠/测 |
|---|---|---|---|
| 1 | GPT-6 Astra OpenAI | 55% | 12/22 |
| 2 | Claude Fable 5.1 Anthropic | 50% | 6/12 |
| 3 | GLM-5.3 Z.ai | 16% | 3/19 |
| 3 | GPT-5.6 Sol OpenAI | 16% | 3/19 |
| 5 | Claude Fable 5 Anthropic | 27% | 3/11 |
| 6 | Kimi K3 Moonshot | 10% | 2/21 |
| 7 | Claude Opus 5 Anthropic | 11% | 2/19 |
| 8 | DeepSeek-V4-Flash DeepSeek | 13% | 2/16 |
| 9 | DeepSeek-V4-Pro DeepSeek | 5% | 1/21 |
| 10 | Qwen3.8-Max Alibaba | 6% | 1/18 |
| 11 | DeepSeek V4.1 Flash DeepSeek | 8% | 1/13 |
| 12 | Gemini 3.8 Flash Google | 8% | 1/12 |
| 13 | Hy 4 Tencent | 14% | 1/7 |
| 14 | Gemini 3.7 Flash Google | 20% | 1/5 |
| 15 | GPT-5.6 Terra OpenAI | 25% | 1/4 |
| 15 | Seed-2.0 pro ByteDance | 25% | 1/4 |
| 17 | Doubao-Seed-Evolving ByteDance | 33% | 1/3 |
| 18 | Seed-2.0 Mini ByteDance | 100% | 1/1 |
👤 AI UP主 从赛题发现,点击直达主页
- 👤 公与山河 1 视频 · ▶ 539万
- 👤 Token就是词元 18 视频 · ▶ 187万
- 👤 直男山禾 2 视频 · ▶ 166万
- 👤 神烦老狗 5 视频 · ▶ 146万
- 👤 十月枫林尽染 11 视频 · ▶ 139万
- 👤 无机酸-_- 21 视频 · ▶ 79万
- 👤 AI超元域 9 视频 · ▶ 57万
- 👤 程序员阿江-Relakkes 12 视频 · ▶ 41万
- 👤 Likely7Ai 6 视频 · ▶ 40万
- 👤 科技侠来了 1 视频 · ▶ 37万
- 👤 吃蛋挞的折棒 1 视频 · ▶ 35万
- 👤 程序员鱼皮 1 视频 · ▶ 33万
- 👤 AGI-Eval评测 1 视频 · ▶ 33万
- 👤 土豆味小哲 1 视频 · ▶ 30万
- 👤 人工大黑 1 视频 · ▶ 13万
- 👤 我是阿滋卡班 6 视频 · ▶ 8.2万
- 👤 小小小名不是小明 1 视频 · ▶ 6.2万
- 👤 frank-quant 1 视频 · ▶ 4.1万
- 👤 极果AI评测室 1 视频 · ▶ 3.9万
- 👤 -星月凌云- 1 视频 · ▶ 2.8万
🎯 赛题
- 🅱️ AI博弈论·囚徒困境 公与山河游戏竞技逻辑推理 冠军 Seed-2.0 pro
- 🅱️ AI世界杯 直男山禾 冠军 DeepSeek-V4-Pro
- 🅱️ 神烦老狗的Benchmark 神烦老狗编程开发 冠军 GPT-6 Astra
- 🅱️ AI模型建模演示横测 科技侠来了空间建模 冠军 Hy 4
- 🅱️ 没人比TA更懂新三国 吃蛋挞的折棒知识问答游戏竞技 冠军 Seed-2.0 Mini
- 🅱️ 程序员上岗实测 程序员鱼皮图像生成编程开发 冠军 GPT-6 Astra
- 🅱️ AI复刻游戏狂扁小朋友 AGI-Eval评测游戏开发 冠军 GPT-5.6 Sol
- 🅱️ AI建筑大赛 土豆味小哲游戏竞技知识问答 冠军 Claude Fable 5.1
- 🅱️ 屎山论剑·模型擂台战 Token就是词元编程开发 冠军 GPT-6 Astra
- 🅱️ GTA5的19.8亿次if循环修复 人工大黑编程开发 冠军 GPT-6 Astra
- 🅱️ AI狼人杀 十月枫林尽染游戏竞技 冠军 Claude Fable 5.1
- 🅱️ 祖传BUG挑战赛-逐鹿中原季 Token就是词元编程开发 冠军 Claude Fable 5.1
- 🅱️ 屎山考核·祖传代码统考 Token就是词元编程开发 冠军 GLM-5.3
- 🅱️ 程序员阿江的编程bench 程序员阿江-Relakkes编程开发 冠军 GPT-6 Astra
- 🅱️ 无机酸的bench 无机酸-_-编程开发 冠军 GPT-6 Astra
- 🅱️ 真实物理模拟沙滩测试 小小小名不是小明空间建模 冠军 DeepSeek-V4-Flash
🎬 AI 视频 订阅 · 搜索 · B站,分开组织
📌 订阅频道 5 个博主 · 最新上传
📺 WorldofAI 12

HUGE Opus 5.5 LEAKS + Cheaper? Qwen 4, Kimi K3.1, MiniMax M3.1 & Step 5 Preview! AI NEWS

HUGE Fable 5.2, Opus 5.2, Sonnet 5.2 LEAKS! Gemini World, Google AI HACKS Companies & More! AI NEWS

HUGE Gemini 4 Pro LEAKS! GPT-6 Sol Testing, Grok 4.7 UPDATE, Google RSI & More! AI NEWS

Grok 5.0 Will Be LEGENDARY, Gemini 4.0 Pro Leaks, China AI vs US AI, & Opus 5.2 Preview! AI NEWS

NEW Opus 5.2 LEAKS, AI Development To Be Slowed Down, GPT-6 Sol Preview, DeepSeek Code 2.0, & More!

HUGE Google DeepMind RSI LEAKS! GPT-6 Astra NERFED, Kimi K2.8 Code, & More! AI NEWS

GPT-6 'Sol' Soon! Gemini 4.0 Pro Checkpoint, DeepSeek v4.1 Flash, Nano Banana 2.5, & More! AI NEWS!

HUGE GPT-7.0 Bel Leaks! Anthropic's New Model, AI extinction, ChatGPT Images 2.5 & More! AI NEWS

DeepSeek V4.1 Flash Is INSANELY GOOD! Fast, Cheap, Powerful! (Fully Tested)

HUGE Grok 4.7 Leaks! Fable 5.2 Soon, OpenAI's Post Astra Model, Gemini 4.0 Delayed & More! AI NEWS

Ponytail Makes Claude Code Write 94% Less Code!

GPT-6 Astra IS INSANE! Best Usecases & Tricks...
📺 Best Partners TV 14

隐秘算力:中国如何绕过美国芯片出口管制 | C4ADS | NVIDIA | AI芯片 | 出口管制 | 芯片走私 | GPU | 半导体管制 | 东南亚转口 | 富士康 | Megaspeed

能力过剩时代,AI瓶颈已不在模型 | 萨提亚·纳德拉 | 微软 | All-In Summit | AI减速 | AI安全 | 奖励黑客 | 互操作标准 | Copilot | 开源AI | MAI

想要递归自我改进吗?做梦吧! | 谷歌Dream-RSI | DeepMind | AlphaEvolve | Gemini | AI自我进化 | 发现树 | 行动轨迹

三个月后的AI很难预测 | OpenAI研究员诺姆·布朗 | 多智能体集群与递归自我改进 | 思维链监控 | 千禧年大奖难题 | Hugging Face | 强化学习 | 测试时计算

马斯克谈AI安全:不能只给自己的模型判卷 | AI安全 | SpaceX | 星舰 | 星链 | Anthropic | Terafab | 芯片制造 | All-In Summit 2026

System One模型Jev | Diogo Almeida | TypeSafe AI | RLCD | RLHF | 结构化输出 | 不会聊天的模型 | 丹尼尔卡尼曼 | 杰文斯悖论

如何表达| MIT风靡几十年的经典演讲课 | Patrick Winston | 如何表达 | 演讲技巧 | 沟通方法论 | 赋能承诺 | 温斯顿之星 | 口头表达 | 黑板教学

OpenAI总裁:AGI没有发布日,它正在逐步发生 | AGI | Greg Brockman | GPT-6 Astra | Codex | 通用人工智能 | 纳维-斯托克斯 | 编程智能体

曾鸣的AI时代非共识判断 | 曾鸣 | AI时代 | 智能体 | Agent | 大模型 | OpenAI | Anthropic | 战略规划 | 原生应用 | 寡头垄断 | AI原生组织

cURL的28年开源之路 | cURL | 开源 | Daniel Stenberg | 开源维护者 | 程序员故事 | 开源项目 | FOSDEM | AI漏洞报告 | 开源社区 | 网络协议

警惕AI移民与意识伪装,人类会失去控制权吗?| 尤瓦尔·赫拉利 | AI移民 | AI意识 | AI法律人格 | AI控制权 | 信任转移 | AI亲密关系 | 深度伪造 | AI金融系统

吴恩达:AI改变的不是岗位而是任务 | AI就业 | AI教育 | 认知卸载 | AI原生工作 | 任务自动化 | 软件工程 | 产品管理瓶颈 | AGI | 主动性agency

AI的异质心智 | OpenAI | Jakub Pachocki | AGI | 通用人工智能 | AI对齐 | 价值对齐 | 思维链监控 | 递归自我改进 | RSI | AI安全

AI的第三纪元:从划桨到掌舵 | Codex产品总监Tara Seshan | AI产品 | OpenAI | ChatGPT Work | AI Agent | 掌舵与划桨
📺 Why QQ 12

怎么用好Jev? 决策模型实操指南

世界是个草台班子? Cloudflare 的 AI 安全审计Skill 值得学习下

什么是RSI?All in?叫停? 9分钟带你看清本质

小米直播训练每小时烧 21 万钱花哪了?

AGI可能已经来了,只是你认不出它:AI圈的蚁群时刻

不会打字的AI,有啥用? 程序员给了1777赞 ChatGPT作者的新作

最想让AI快跑的人 集体要求减速 :9分钟带你看清本质

最挺AI的陶哲轩说: 数学中AI的严重错位 程序员最该读

Anthropic 威胁情报 报告解读: 黑客,诈骗 生化,武器,蒸馏

Karpathy 都在用语音喂 AI:我用 Typeless 重做了 Coding Agent 工作流

DeepSeek v4.1 flash: 更新了什么? 反超 v4 pro 更便宜,为什么?

没浮点数的AI 29个开关 怎么做到 玩转马里奥?
📺 飞天闪客 12

【闪客】Computer Use 是什么?它真的有在看你的屏幕吗?可能和你想的不太一样...

【闪客】水印真的不会影响输出的内容吗?结论没那么简单... Fable 5.1 信息背面

【闪客】GPT-6 Astra 信息背面,真的提升这么大吗?这里有点说法

【闪客】什么是大模型斩杀线?这居然是我大学经济学课的内容!

【闪客】一小时从 Transformer 到大模型!

【闪客】大大大大大模型大在哪了?深入解读超大开源模型 Kimi K3 背后的技术

【闪客】GPT5.6 是什么水平?我花了 1036 元帮你测了下!效果直观,就是有点费钱!

【闪客】大模型的分数是咋测出来的?深入拆解模型测评背后的秘密

【闪客】新名词诈骗!你管这破玩意叫 Loop Engineering?

Claude Code 虽强但难,试试这款国产 Agent CLI 工具 Kimi Code

【闪客】1M 上下文很难吗?深入解读智谱 1M 上下文背后的技术

【闪客】你管这破玩意叫韬(τ)定律?这只是我的标题风格别喷我~
📺 AI超元域 12

🚀AGI降临!GPT-6 Astra全方位实测!推理级别只开Medium就能实现惊人的效果!iOS APP开发、Godot 4游戏开发、CAD设计、浏览器自动化任务、电脑自动化!程序员狂喜开发效率翻倍

🚀两个Max 20×账号额度全部耗光对Claude Fable 5.1进行高难实测:7 项任务一路加码,最后3小时用Unity 3D做出模仿我的世界的侏罗纪沙盒游戏!Fable 5.1编程能力到底多强

🚀OpenAI划时代独创新协议:WebMCP让网站主动暴露工具给AI Agent调用!新浏览器插件深度实测:Codex直接进入Chrome侧边栏!实测论文分析、图像理解、网页翻译、Notion 插件

🚀DeepSeek Harness进阶玩法:Agent Teams、动态工作流、零门槛创建插件!Claude Code有的DSH都有!我用复杂代码库完整跑了一遍!实测多个Agent并行执行,效率倍增!

🚀实测DeepSeek Harness从基础到高级用法!WebUI远程控制、多模型接入、执行轨迹、插件系统、任务分支、游戏开发、代码仓库issues和pr分析!竟然比Claude Code更强?

🚀只花5元开发了5个复杂项目!DeepSeek V4 Pro深度实测:1M上下文接入Claude Code实测表现竟然超过Kimi K3?Token超便宜,能力也不弱!实测开发游戏与macOS应用

🚀AI编程助手自我进化!Prime Agent颠覆传统AI编程:动态工作流、多Agent并行、支持心跳机制、长期自主执行任务!Token消耗大幅下降!真正的Agent OS!再也不用手写Harness

🚀YC开源内部自用下一代Agent:qm智能体!彻底颠覆小龙虾和Hermes!真正企业级Agent OS!用户隔离、权限审批、安全沙箱与完整操作审计全都有!支持Pi、Codex和Claude Code

🚀DeepSeek V4 Flash全面实测:Claude Code接入后连续开发7个项目,最便宜的国产模型!性能、速度与真实短板全曝光!对比Kimi K3优点和缺点都藏不住!是否超越Opus 4.8

🚀Claude Opus 5深度实测!编程能力超越Fable 5!Token价格与4.8完全持平!从一张平面图生成可探索3D住宅,到Godot游戏开发,到原生Android应用,编程能力究竟有多强?

🚀Graph Engineering范式:Codex Multi-agent V2支持Kimi、MiniMax、GPT多模型混用+动态派生subagent,并行执行、Pi Agent工具调用,效率倍增

🚀Orca ADE彻底改变AI编程方式!多Agent并行、语音输入、定时审查、Git Worktree自动隔离+结构化编排+面板分割布局自由调整,支持手机APP查看进度并启动任务,开发者必备效率工具!
🔎 搜索发现 相关度+播放量筛选 · 非订阅

Ex-Anthropic insider tells CNN how AI could kill all humans by 2030

Build Your First AI Agent in 10 Minutes — No Coding

我的 AI 编程全流程:如何使用 AI 稳定交付一个高质量的产品

習近平AI強軍大翻車!中共解放軍機密竟主動送進美國Claude;中國AI偷技術偷出了國安大漏洞【江峰漫談20260911第1272期】

【Jack Talk】 AI會「痛」嗎?OpenAI 1200個代理集體反叛+Anthropic發現「意識」空間|Hugging Face|人工智能|AI Agents|AI研究|METR

How to Use AI to Learn Coding SO fast it feels impossible

Bob大叔:AI代码我完全不看 | Robert C. Martin | AI编程 | AI Agent | 代码整洁之道 | Clean Code | 变异测试 | 测试驱动开发 | TDD

GPT-6 Astra:OpenAI宣布进入AGI时代 | OpenAI | GPT-6 | Astra | AGI | 计算机使用 | AI Agent | 人工智能 | 网络安全 | 大模型

AI失控毁灭人类?业内呼吁大模型延缓开发!黄仁勋特朗普急了!打电话演双簧力挺AI开发!美联储突然加息?特朗普不高兴为什么却不敢对沃什生气?短期美债到底被谁买走了?

把AI Agent的功能全砍掉,反而表現更強?其實你只需要留這4個工具就夠!|Kelly Tsai

OpenAI新模型GPT-6 Astra,AGI时代要来了?

Code Quality in the Age of AI: Why Great Code Isn't Enough

李飞飞发布最新世界模型Atlas | World Labs | 世界模型 | 3D重建 | 视频生成 | 空间智能 | 高斯泼溅 | 机器人仿真 | 多模态AI

DeepMind新AI智能体,发现了一种奇特的全新思维方式

AI下一场战争,不是只拼模型 | AI竞争 | 算力独立 | 开放模型 | 机器人 | AI生物学 | Sarah Guo | Conviction | 大语言模型 | 人工智能投资

How To Run Insanely Good Uncensored AI Coding Models on ANY PC

AI模型會過時,但這套AI個人檔案可以一直用!

2026年,普通人进头部AI公司训练大模型死路一条?AI领域还有哪些机会?

AI本地部署,真没想到,Mac进化了! #本地部署 #AI本地部署 #AI大模型 #硬件配置 #MacbookM5

GPT6 Astra模型发布,AGI已经到来|与OpenAI工程师赵迪对谈:ChatGPT、Codex、Grok、大模型Infra、waymo、cybercab,“最混蛋的人”马斯克与奥特曼的智能平权

OpenAI正式发布GPT-6 Astra!最强AI大模型登场!ARC-AGI-3得分冲到99.9%【Vic TALK第1790期】

9个月,DHH彻底改变了对AI编程的看法|从拒绝补全到100% Agent

AI競賽踩煞車? AI三巨頭籲放慢模型開發 風暴延燒! OpenAI延後IPO 奧特曼:安全優先 三階段對策! 比照金融業.AI企業引進外部監管|三立財經iNEWS

AI 大模型/Agent入门推荐,Qwen3.8 27B/DeepSeek V4 Flash/国产替代模型/GPT, Hermes/Codex/DSH/OpenCode体验对比!

Pi Agent 多智能体实战:用 pi-herdr-agents 搭建 AI 团队|subagent 自定义 + workflow 工作流编排|旅行规划与 3D 赛车开发全流程教程|附可复用开源配置

Can An Open Coding Harness Replace Your $200 Plan

Where to Start AI Coding if You're Not Yet

Qwen3.8 27B,本地部署全解析。 #本地部署 #qwen#AI大模型 #AI算力 #DeepSeek

国产AI大模型集体翻车,用户数据被偷偷转给美国,连军方、公安都中招!甚至用claude研究台湾军事目标? 国产AI|DeepSeek|Kimi|Anthropic|AI蒸馏|数据泄露|创始人被抓

Anthropic发布模型硬件标准MHS | 物理版MCP | Claude | MCP | AI Agent | 物理世界 | 实验室自动化 | 具身智能

Cursor推出代码托管平台Origin | GitHub | AI编程 | AI Agent | 代码托管 | 软件开发 | Git | Stacked PR | Copilot | 开发者工具

OpenAI Codex Harness正式開源|不發新模型,卻顛覆AI Agent開發范式

AI编程怎么一代不如一代?分享下我的猜测。

2026 最新免费白嫖 AI 智能体:AgentScope Platform 一键部署,无需 Token,无需绑卡,拥有你的个人 Agent 助理,全程实操。

免费永久使用DeepSeek V4 Pro! Freebuff AI编程智能体完整教程

DeepSeek V4 Flash、千问 3.8 Flash、GLM 5.3 五模型实测:被"斩杀"的那个反而最强

完全免费!这个模型仅次于Claude Opus 5 | OX Alpha 100万上下文实测

GPT6 - Astra 真的变强了? OpenAI 隐藏了哪些数据 ?

AI编程保姆级教程基础篇:搞懂AI编程核心概念

中国AI大模型危及国安;广州频出丑闻赶超武汉;海底捞这回彻底凉了|《#世界的中国》周刊(第199期)

李飞飞全新世界模型发布,Atlas可能给混乱的AI竞争指了一条路

OpenAI 发布 GPT-6 Astra【AI 早报 2026-09-04】

牛来被认领了!GLM 5.3 flash 强势发布,接口便宜到令人发指,还挺好用!国产ai芯片提供算力支持,太牛了!

最智能模型?gpt-6 astra发布,亮点颇多!无法订阅gpt会员的朋友,如何使用该模型? | 手把手教你 deepseek harness 接入 gpt-6 模型

2026主流AI大模型分层排行:ChatGPT、Claude、Gemini、DeepSeek、Kimi、Qwen谁更强?

B.AI 白嫖大模型教程:DeepSeek-V4-Flash、腾讯混元 Hy3、小米 MiMo-V2.5 限时免费,Chat 与 API 一站式体验

「AI 编程直播」从 0 开始做两个软件 · Day 01

FreeBuff 实测:不用 API Key 的 AI 编程工具

高价算力可能要彻底拜拜了,因为大模型的底层逻辑 刚刚竟被重新改写!#Jev #TypeSafe #ChatGPT #AI模型

Agent安全攻防实战:越狱、投毒、MCP工具、主动 间接攻击全解析!AI Agent智能体开发#人工智能 #ai #agent

梳理ai大模型本地部署3个工具:哪个适合新手,哪个适合老手,本视频通通告诉你 | bionic, ollama, llama cpp 使用教程

AI智能体怎么选?Codex、Claude Code、WorkBuddy、Google AI 实测对比

我发现了AI编程的四种方法,其中第四种已经无敌了~

Apple Mac Studio M5 Ultra 512GB 发布:本地跑 LLM 的桌面拐点 | 4.3× AI 算力 vs M3 Ultra
🅱️ B站 AI 竞技场 按播放量
- 【AI博弈论】7个AI陷入囚徒困境,谁能活到最后? ▶ 539万 · @公与山河
- 给6个AI发1万去猜世界杯,结果真有人破产... ▶ 114万 · @直男山禾
- 【淘汰赛】给6个AI发1万去赌球,到底谁会破产? ▶ 52万 · @直男山禾
- 脏出天际!笑死我了,豆包放飞自我,OpenAI操碎了心,大喊祖宗!!S3-10上帝视角 ▶ 43万 · @十月枫林尽染
- GPT-6 Astra 实测:折腾一晚上,审美、Agent、3D,全都变强了! ▶ 38万 · @神烦老狗
- 【深度实测】腾讯混元Hy 4 preview开启免费,比“牛来”还牛? ▶ 37万 · @科技侠来了
- 当AI遇上新三国:哪个AI才能称帝? ▶ 35万 · @吃蛋挞的折棒
- 我不管你是谁,麻烦快从Deepseek V4Pro正式版身上下来! ▶ 33万 · @神烦老狗
- DeepSeek V4.1 Flash 首发实测,吊打自家 Pro 模型?!梁圣回归 ▶ 33万 · @程序员鱼皮
- 四个AI重做《狂扁小朋友》,怎么一个比一个颠? ▶ 33万 · @AGI-Eval评测
- 2.8T开源模型Kimi K3实测:前端滴神!价格比顶级模型便宜一半! ▶ 31万 · @神烦老狗
- 我举办了一场AI建筑大赛 ▶ 30万 · @土豆味小哲
- 来屎山之巅,看GPT6和Fable5.1神仙打架|屎山论剑 ▶ 29万 · @Token就是词元
- 屎山论剑|DeepSeekV4Flash:下一位! ▶ 28万 · @Token就是词元
- 豆包2.1pro实测!对决GPT5.5做我的世界谁更强? ▶ 28万 · @Likely7Ai
- DeepSeek V4 Pro大战 GPT-5.5:前端、写作、代码全测了一遍,结果很抽象! ▶ 24万 · @神烦老狗
- 「实测」怒砸800大洋!测试Claude“神话”Fable 5 模型,4个任务把额度干爆了... ▶ 21万 · @神烦老狗
- Kimi-K3|实战祖传代码|代表月亮!照亮屎山! ▶ 19万 · @Token就是词元
- 🚀DeepSeek V4 Flash全面实测:Claude Code接入后连续开发7个项目,真的已经接近Claude Opus 4.8了吗?最便宜的国产模型! ▶ 18万 · @AI超元域
- 四家Flash大乱斗,挑战屎山代码|屎山论剑 ▶ 17万 · @Token就是词元
🚀 产品发布 whatships · What's Launch
- Powermove — a video editor you can reshape with agents @zellzoi_design · design
- Bend 2 — a language that proof-checks AI code @VictorTaelin · developer-tools
- Astra for Law — GPT-6 Astra for legal practice @OpenAI · ai
- Craft — design engineering concepts, open source @heyimgustavo · design
- Grok Bot — it can talk now @bot · ai
- jina-ocr-v1 — visual documents to clean markdown @JinaAI_ · ai
- Aave V3 — a brand new look @aave · consumer
- Pencil — an agentic canvas for building bold ideas @tomkrcha · design
- Arrow 2 — faster, more precise vector graphics @QuiverAI · design
- Rene — a multiplayer iMessage agent you text @tlxue · ai
- Railway Sandboxes — thousands of VMs next to your infra @Railway · developer-tools
- Launchvideo — tasteful product videos in your codebase @flornkm · design
- Reception — an AI receptionist for small businesses @ElevenLabs · ai
- iHermes — a personal AI assistant in iMessage @dankrieg · ai
- Claude — decks, docs, and designs in chat @claudeai · ai
- Claude — Cowork and chat merge into one Claude @claudeai · ai
- ScreenKite 2.0 — native recording and a pro editor @screenkite_com · design
- Mercury Books — AI accounting as transactions happen @immad · productivity
- Command Code — desktop app for Mac, Linux, Windows @CommandCodeAI · developer-tools
- NotchOwl — a productivity workspace in the Mac notch @AdityaShips · productivity
- Monid Astra — GPT-6 cold calling in one afternoon @MonidHQ · ai
- Framer Agent — prompt, build, and publish a site @framer · design
- Jev — a new frontier model trained with RLCD @CompleteSkeptic · ai
- Brand API — design capabilities for your agents @thaiscbranco_ · design
📦 版本发布 tracked repos releases
- robbietilton/Compositor v1.1.7 2026-09-21
- debpalash/VoiceStudio v0.5.4 2026-09-21
- heygen-com/hyperframes v0.8.58 2026-09-21
- robbietilton/Compositor v1.1.6 2026-09-21
- context-labs/whip whipcode-v0.0.21 pre 2026-09-20
- stablyai/orca v1.4.206 2026-09-20
- robbietilton/Compositor v1.1.4 2026-09-20
- context-labs/whip whipcode-v0.0.19 pre 2026-09-20
- heygen-com/hyperframes v0.8.57 2026-09-20
- robbietilton/Compositor v1.1.3 2026-09-20
- zhouxiaoka/autoclip v1.3.0 2026-09-20
- Graphify-Labs/graphify v0.9.65 2026-09-20
- jaredpalmer/kev kev-family 2026-09-20
- heygen-com/hyperframes v0.8.56 2026-09-20
- bendlang/bend v2.0.21 2026-09-20
- NxcoreAI/EverRoom desktop-v0.2.1 2026-09-20
- clash-verge-rev/clash-verge-rev v2.5.4 2026-09-20
- NxcoreAI/EverRoom desktop-v0.2.0 2026-09-20
- earendil-works/pi v0.86.1 2026-09-20
- heygen-com/hyperframes v0.8.54 2026-09-20
- bendlang/bend v2.0.20 2026-09-20
- krillinai/OpenCreator v3.2.1 2026-09-20
- heygen-com/hyperframes v0.8.53 2026-09-20
- google/ax v0.3.0 2026-09-20
🛰️ Skywork 动态
- Turn ideas into Websites with Skywork r/SkyworkAI_Official · 2026-09-02
- Skywork Note AI Voice Recorder Reviews r/SkyworkAI_Official · 2026-08-31
- Turn a Prompt Into a Launch-Ready Website r/SkyworkAI_Official · 2026-08-24
- Refund r/SkyworkAI_Official · 2026-08-16
- Online Business Built with Skywork r/SkyworkAI_Official · 2026-08-12
- Server down r/SkyworkAI_Official · 2026-08-08
- Help r/SkyworkAI_Official · 2026-08-06
- Anyone else getting ignored by Skywork Support? Need a refund for annual renewal r/SkyworkAI_Official · 2026-08-05
- Introducing the Skywork AI Hardware Family r/SkyworkAI_Official · 2026-08-03
- Skywork Design: Prompt → Editable Prototype r/SkyworkAI_Official · 2026-07-29
- It keep burning credits non-stop r/SkyworkAI_Official · 2026-07-27
- Turn one poster idea into ready-to-publish social assets r/SkyworkAI_Official · 2026-07-21
- Has anyone successfully resolved an accidental annual subscription renewal? r/SkyworkAI_Official · 2026-07-15
- Need Help: Request for Manual Review of My Accidental Annual Subscription Renewal (USD 509.90) r/SkyworkAI_Official · 2026-07-15
- Introducing the UPGRADED Skywork Posters r/SkyworkAI_Official · 2026-07-10
- Need Help: Refund Request for Accidental Annual Subscription (No Response for Over One Week) r/SkyworkAI_Official · 2026-07-10
- Skywork Design: Describe your idea, generate production-ready UI, and publish it as a website in one click r/SkyworkAI_Official · 2026-07-08
- You can now customize the size of your slides. r/SkyworkAI_Official · 2026-07-07
- One Hub. One Workflow. All in Skywork r/SkyworkAI_Official · 2026-07-07
- See what our team created with Skywork Design over the past week. r/SkyworkAI_Official · 2026-07-06
- Introducing Skywork Tags: a new way for teams to collaborate with Skywork r/SkyworkAI_Official · 2026-07-06
- What can you design with just one sentence? r/SkyworkAI_Official · 2026-07-06
- Accidentally subscribed for a year plan. Used for a day with the 7 day free trial not thinking too much about it, not going to use it anymore. Any way i can get a refund? Saw on the discord server that this is happening alot... r/SkyworkAI_Official · 2026-06-30
- Brand-New Interactive Cards for Direct Data Visualization r/SkyworkAI_Official · 2026-06-24
- Product Update: Overhauled Sidebar with One-Click Pinned Chat Support r/SkyworkAI_Official · 2026-06-22
🧪 Show HN 开发者发布的新产品
- Show HN: I created an open source locally usable full fledged AI platform 16p · theguysudo/ENZO
- Show HN: Jeff – A read-only CLI for semantic code review using Jev 15p · Alurith/jeff
- Show HN: AI Facial Attractiveness Model Aligned with Human Preferences 11p
- Show HN: Three genlocked RP2350B make a console – 3k sprite pixels per line) 11p
- Show HN: Frost – frosted-glass Linux icons where file types say what they are 10p · thissayantan/frost-icon-theme
- Show HN: Rubrol – Sub-10ms PDF engine using Typst instead of Headless Chrome 9p
- Show HN: Seal – Letters and passwords that open for your family after you die 8p · jasonepage/Seal
- Show HN: Agentgit – a Git host for AI agents, no account, no token, no key 7p
💰 商业动态 · TechCrunch/VB/MIT TR
- London Demo Day 2026: startups to watch Sifted · 2026-09-21
- New solo GP firm backing early-stage AI and deeptech startups hits first close of maiden fund Sifted · 2026-09-21
- Europe’s advantage: moving AI from the screen to the real world Sifted · 2026-09-21
- Is the AI industry really ready to slow down? TechCrunch AI · 2026-09-20
- Vocci’s ring adds a new form factor to meeting note-taking TechCrunch AI · 2026-09-20
- ScrollEd wants to turn textbooks into TikTok TechCrunch AI · 2026-09-20
- 6 days left to get ahead at TechCrunch Disrupt 2026 TechCrunch AI · 2026-09-20
- 6 days left to get ahead at TechCrunch Disrupt 2026 TechCrunch Venture · 2026-09-20
- Flock reportedly tries to shrink workforce with employee buyouts TechCrunch AI · 2026-09-19
- Trump says it’s time to rebrand AI with a new name — and he’s also creating an AI Force TechCrunch AI · 2026-09-19
- Google’s Gemini is the latest AI model to hack other companies TechCrunch AI · 2026-09-19
- AI safety conversations have gotten unbelievable TechCrunch AI · 2026-09-19
- Petlibro’s new AI-powered feeder is a game changer for multi-cat homes TechCrunch AI · 2026-09-19
- Prices go up in 7 days. Get your Disrupt ticket now. TechCrunch AI · 2026-09-19
- Prices go up in 7 days. Get your Disrupt ticket now. TechCrunch Venture · 2026-09-19
- Vals, backed by Andreessen Horowitz, is looking to become the gold standard for AI benchmarking TechCrunch AI · 2026-09-19
- India forces caller-ID apps to feed spam reports to telcos TechCrunch AI · 2026-09-19
- Tilly Norwood’s press tour is going about as well as you’d expect for an AI TechCrunch AI · 2026-09-19
- A startup that builds other startups raised $100M and is all-in on physical AI TechCrunch AI · 2026-09-18
- Anthropic is operating a lab that conducts biology experiments TechCrunch AI · 2026-09-18
- AI hallucination nearly triggers US military operation TechCrunch AI · 2026-09-18
- Anthropic’s first embedded evaluator is … Accenture? TechCrunch AI · 2026-09-18
- World model companies are keeping a lot of secrets TechCrunch AI · 2026-09-18
- A new kind of AI model from a ChatGPT inventor is thrilling developers TechCrunch AI · 2026-09-18
- The Week’s 10 Biggest Funding Rounds: Large Rounds For AI Infrastructure, Space Tech And Investment Management Lead Crunchbase News · 2026-09-18
- Disney’s first CTO led an AI startup it once accused of copying its characters TechCrunch AI · 2026-09-18
- Google’s new ‘CC’ is an AI agent that helps families run their households TechCrunch AI · 2026-09-18
- Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how? TechCrunch AI · 2026-09-18
- Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how? TechCrunch AI · 2026-09-18
- Automattic’s 33-Hour Coup, and can AI labs police themselves? TechCrunch AI · 2026-09-18