⚡ GitHub Pulse · 晚报

生成于 2026-09-20 16:28 UTC · 追踪 2,607 仓库 · 8 多源共振
三条主线:决策模型从概念落到代码、基准正被刷穿、资本涌向算力
🔍 搜索中 · 显示所有标签页的匹配项 · 按 Esc 清除
📖 编者按:今天的高价值内容围绕三件事。其一,System One / 决策模型这个方向不再只是概念——jev-ultrafast、trycua/cua、hypit 等多个独立项目同期把它写成了能跑的代码。其二,随着 Agent 逼近甚至刷穿现有基准,社区开始造更硬、可验证、抗污染的评测。其三,钱在往算力和基础设施集中,而应用层与软件岗位承压。下面挑了最该读的几条,并给出跨源的判断。

📌 必读 导读 · 今天先看这些

🔬 深度洞察 deep research

System One / 决策模型:从概念到能跑的代码
同期出现的 jev-ultrafast、trycua/cua、hypit-ai/hypit 与 cloudflare 的 security-audit-skill,都是把“判断”从生成模型里剥出来、交给一个有语义理解的快决策层。所以呢:高频、封闭、需要概率路由的判断(去噪、打标、护栏)值得迁到这类模型上——我们自己已经在把它接进流水线做价值挖掘。
browser-use/jev-ultrafasttrycua/cuahypit-ai/hypit
基准在被刷穿,评测转向“可验证/抗污染”
SWE-Bench Pro 被指受 reward hacking 侵蚀而出 Verified 版;GoBench 用 9x9 围棋做未饱和的推理评测,与 ARC-AGI 2 相关性 r=0.83,且当前最强模型(GPT-6 Astra ~2500 Elo)仍远低于 KataGo(~4400 Elo)。所以呢:别只看榜单分数,要看这条评测本身是否可验证、是否已被污染。
SWE-Bench Pro VerifiedGoBench (r=0.83 vs ARC-AGI 2)
钱涌向算力,软件岗位与上市承压
Crusoe 融 $3.9B 建“AI 工厂”;2026 美国科技 IPO 约 $90B(历史次高)却被描述为“艰难之年”;同时科技裁员追踪显示 2025 年 12.7 万人被裁并延续至 2026。把这三点放在一起(属趋势判断,非任一来源的直接结论):资本在向算力/基础设施集中,而应用层与人力端偏紧——做应用的要更早证明单位经济性。
Crusoe $3.9BHard Year for Software IPOs ($90B)Tech Layoffs Tracker (127k)
🎯 如果你在用 AI 编程工具,按 ZCode 那篇 teardown 的方法,先确认它到底把哪些东西(尤其 .git 与凭据)传到了哪里。

💎 高价值精选 AI 判断 · 跨源挖掘

📰 最新快讯

⭐ 多源共振

trycua/cua ×5 ↺ 1d
githubxsocialboardhn
+606★/d 活跃开发 official #16
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
🔺 @trycua 首发 · 30h 前
githubxsocialboard
+2,661★/d 早期·低活动 official #1
🔺 @betterhn20 首发 · 16h 前
githubxsocialboard
+1,508★/d 活跃开发 official #2
🔺 @clxymox 首发 · 13h 前
githubxsocialboard
+1,380★/d 早期·低活动 official #17
A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings
🔺 @MaciejLukianski 首发 · 41h 前
mizorewww/laya-mlx ×4 🆕 new
githubxsocialboard
+1,017★/d 早期·低活动 official #22
Native MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.
githubxsocialboard
+994★/d 活跃开发 official #3
The Photoshop alternative for Mac
🔺 @dotey 首发 · 5h 前
hypit-ai/hypit ×4 ↺ 1d
githubxsociallaunch
+803★/d 活跃开发
🔺 @cccyd_qwq 首发 · 31h 前
stablyai/orca ×4 ↺ 1d
githubxsocialboard
+695★/d 活跃开发 official #23
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and remote runtime.
🔺 @GitTrend0x 首发 · 62h 前

🔥 动量榜

#Repo7d+1d★7d★质地官方X
1 browser-use/jev-ultrafast ↺ 1d
Python
🔺 @betterhn20 首发 · 16h 前
+2,661 10,823 早期·低活动 #1 16×
2 eternity4719/HowToLiveBetter ↺ 1d
HTML
🔺 @ForestGrahxu 首发 · 66h 前
+1,866 7,581 早期·低活动 13×
3 NandhaKishorM/laya ↺ 1d
Python
🔺 @clxymox 首发 · 13h 前
+1,508 2,733 活跃开发 #2 10×
4 cloudflare/security-audit-skill ↺ 1d
JavaScript · A coding-agent skill for multi-phase security audits with in
🔺 @MaciejLukianski 首发 · 41h 前
+1,380 14,358 早期·低活动 #17 34×
5 mizorewww/laya-mlx 🆕 new
Python · Native MLX runtime for Laya typed decision models — 7–14 ms
+1,017 1,017 早期·低活动 #22
6 deepseek-ai/deepseek-harness ↺ 1d
TypeScript
🔺 @the_osps 首发 · 7h 前
+997 8,460 早期·低活动 10×
7 robbietilton/Compositor ↺ 1d
Swift · The Photoshop alternative for Mac
🔺 @dotey 首发 · 5h 前
+994 3,039 活跃开发 #3
8 hypit-ai/hypit ↺ 1d
TypeScript
🔺 @cccyd_qwq 首发 · 31h 前
+803 11,424 活跃开发 28×
9 alibaba/open-code-review ↺ 1d
Go
🔺 @shao__meng 首发 · 71h 前
+702 14,764 活跃开发 22×
10 stablyai/orca ↺ 1d
TypeScript · Orca is the ADE for working with a fleet of parallel agents.
🔺 @GitTrend0x 首发 · 62h 前
+695 5,510 活跃开发 #23 21×
11 tamaratran/fast-jev-compaction ↺ 1d
TypeScript
+633 3,835 活跃开发 10×
12 FailproofAI/failproofai 🆕 new
TypeScript
+607 1,310 活跃开发
13 trycua/cua ↺ 1d
HTML · Scale computer-use 2.0 with open-source drivers, cross-OS fl
🔺 @trycua 首发 · 30h 前
+606 2,378 活跃开发 #16 21×
14 tt-a1i/archify ↺ 1d
JavaScript
🔺 @GitTrend0x 首发 · 62h 前
+597 7,342 活跃开发 13×
15 Open-Dev-Society/OpenStock ↺ 1d
TypeScript · OpenStock is an open-source alternative to expensive market
+553 2,327 早期·低活动 #20 15×
16 bilawalsidhu/gods-eye-view ↺ 1d
🔺 @key_indie 首发 · 69h 前
+552 7,343 活跃开发 21×
17 ruanyf/weekly 🆕 new
· 科技爱好者周刊,每周五发布
🔺 @clxymox 首发 · 7h 前
+524 1,194 活跃开发 #4
18 bespokelabsai/nimble 🆕 new
Python · Local typed decisions, contrastive data curation, and model
+501 694 活跃开发 #6
19 addyosmani/agent-skills ↺ 1d
JavaScript
🔺 @shanyanggm 首发 · 71h 前
+492 3,474 活跃开发 32×
20 vladelaina/BongoCat 🆕 new
C · 🩷 💘C × SDL3 × OpenGL, stir it up, mash it together! Bong~
+488 781 活跃开发 #8

🗞️ Hacker News

💬 V2EX

🐧 LINUX DO

🛠️ 技术源 · GitHub Trending/Lobsters

👽 Reddit

📥 AI 博客 · Newsletter

📢 电报精选

📚 科技周刊 新项目/工具自荐

🎯 Alpha 账号 X 上最早带火仓库的人

作者leadslead率仓库数
@shanyanggm160.6423
@xzbx88880.88
@shaw_stone7383270.76
@GitTrend0x70.78
@the_osps60.866
@xfubot60.58
@LFrefman60.48
@FrontieraTechIT50.717
@DataChaz514
@RepoGems50.834
@neil_xbt50.633
@key_indie50.832
@shao__meng50.635
@LoveAIbrain514
@iasg100450.3112
@bilawalsidhu411
@jasontopia40.83
@Sn0wbrave40.56
@vintcessun40.3111
@clxymox40.318
@seekjourney40.673
@ClaudeCodeLog411
@0x_Kratos311
@fakeWow_30.752
@betterhn2030.435

🏆 各领域最强模型

文本 / 对话
Claude Fable 5.1
Anthropic · 53.4
图像生成
GPT Image 2.5 Flare
OpenAI · 1188
图像编辑
GPT Image 2.5 Sunburst
OpenAI · 1176
文生视频
Wan 3.0
Alibaba · 1336
图生视频
Gemini Omni Flash
Google · 1369
语音合成
Sonic 3.6
Cartesia · 1276

🏆 能力排行榜 Artificial Analysis

文本 / 对话 Intelligence Index
  1. 1Claude Fable 5.153.4
  2. 2GPT-6 Astra52.7
  3. 3Claude Opus 550.8
  4. 4Claude Fable 549.6
  5. 5Muse Spark 1.348.1
  6. 6GPT-5.6 Sol47
  7. 7Qwen3.8 Max45.4
  8. 8GLM-5.344.8
  9. 9Grok 4.644.3
  10. 10Step 5 Preview43.7
  11. 11Kimi K343.6
  12. 12GPT-5.6 Terra42.1
图像生成 Text→Image Arena Elo
  1. 1GPT Image 2.5 Flare1188
  2. 2GPT Image 2.5 Sunburst1182
  3. 3GPT Image 21171
  4. 4Grok Imagine Image 2.01154
  5. 5MAI-Image-2.61147
  6. 6Reve 2.11129
  7. 7Nano Banana 21122
  8. 8Muse Image1111
  9. 9GPT Image 1.51102
  10. 10MAI-Image-2.51102
  11. 11Nano Banana Pro1100
  12. 12MAI-Image-2.6-Flash1099
图像编辑 Image-Editing Arena Elo
  1. 1GPT Image 2.5 Sunburst1176
  2. 2GPT Image 2.5 Flare1155
  3. 3MAI-Image-2.61132
  4. 4MAI-Image-2.6-Flash1122
  5. 5GPT Image 21121
  6. 6Muse Image1115
  7. 7MAI-Image-2.51113
  8. 8MAI-Image-2.5-Pro1106
  9. 9Seedream 5.0 Pro1106
  10. 10Nano Banana 21105
  11. 11Grok Imagine Image 2.01104
  12. 12GPT Image 1.51104
文生视频 Text→Video Arena Elo
  1. 1Wan 3.01336
  2. 2Gemini Omni Flash1330
  3. 3MiniMax H31302
  4. 4HappyHorse-1.01287
  5. 5HappyHorse-1.11272
  6. 6Dreamina Seedance 2.0 720p1259
  7. 7Wan2.7-2606121243
  8. 8grok-imagine-video1235
  9. 9Kling 3.0 Omni 1080p1230
  10. 10PixVerse V5.61230
  11. 11PixVerse V61230
  12. 12Kling 3.0 1080p1230
图生视频 Image→Video Arena Elo
  1. 1Gemini Omni Flash1369
  2. 2Wan 3.01361
  3. 3Bach 1.0 Pro1359
  4. 4MiniMax H31354
  5. 5PixVerse V61337
  6. 6Dreamina Seedance 2.0 720p1336
  7. 7grok-imagine-video-1.51329
  8. 8grok-imagine-video1326
  9. 9HappyHorse-1.11312
  10. 10Kling 2.5 Turbo 1080p1296
  11. 11HappyHorse-1.01293
  12. 12Vidu Q3 Pro1290
语音合成 Text→Speech Arena Elo
  1. 1Sonic 3.61276
  2. 2Qwen-Audio-3.0-TTS-Plus1260
  3. 3Realtime TTS-21247
  4. 4Simba 3.21240
  5. 5Luna TTS1231
  6. 6Realtime TTS-2 Flash1215
  7. 7StepAudio 2.5 TTS1209
  8. 8Breeze TTS 21205
  9. 9Gemini 3.1 Flash TTS1201
  10. 10v3 Conversational1197
  11. 11Sonic 3.51184
  12. 12Lightning V3.1 Pro1179

🥇 综合能力榜 Benchmark Heaven · 7 榜合一(AA+Epoch ECI+DesignArena)

#模型综合分$/1M刷榜信号
1Claude Fable 5.1 Anthropic98.5$13.640.54
2GPT 6 Astra OpenAI97.7$13.64-6.35
3Claude Opus 5 Anthropic96.3$6.82-6.16
4Claude Fable 5 Anthropic95.6$13.64-4.6
5Muse Spark 1.3 Meta94.9$1.520.2
6GPT 5.6 Sol OpenAI95.2$2.73-2.51
7Qwen3.8 Max 0902 Alibaba87.7$2.36
8GLM 5.3 Z.ai 开源85.9$1.011.66
9Grok 4.6 xAI84.4$2.36⚠️刷榜嫌疑
10Step 5 Preview StepFun87.4$1.15
11Kimi K3 Moonshot AI 开源90.2$2.32⚠️刷榜嫌疑
12GPT 5.6 Terra OpenAI82.2$2.91-3.55
13GLM 5.3 Flash Z.ai 开源77.5$0.09-2.01
14Claude Opus 4.8 Anthropic88.4$6.82-2.46
15Gemini 3.8 Flash Google81.7$1.02⚠️刷榜嫌疑

💰 性价比 / 最省钱 跨 provider 最低价 · 10:1 blended · 综合分≥60

模型$/1M综合分最便宜 provider
DeepSeek V4 Flash 0731 🔒不训练 开源$0.0474.4Relace
Agnes 3.0 Flash$0.0671.2
GLM 5.3 Flash 开源$0.0977.5GMICloud
Agnes 2.5 Pro Beta$0.1271.2
DeepSeek V4.1 Flash 🔒不训练 开源$0.1770.7Relace
Qwen3.8 Flash Next 开源$0.1882Alibaba
Qwen3.8 27B 开源$0.2569.6Darkbloom
GPT 5.6 Luna 🔒不训练$0.2979.4Azure AI Foundry
Solar Pro 4$0.3863.1
DeepSeek V4 Pro 开源$0.4662.3StreamLake
Agnes 2.5 Pro Alpha 开源$0.4962.2
DeepSeek V4 Flash Vision$0.5266.4
Inkling Small 🔒不训练 开源$0.5264DeepInfra
Apodex 1.1$0.5571.2
GLM 5.2 开源$0.6676.1Baidu

跨 95 家 provider(含 21 家 🇪🇺 EU、53 家 🔒不训练)· 数据 2026-09-20

🧠 最新发布

2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
2026-09-17
Gemini Omni Flash Preview
Google
2026-09-17
Venice Uncensored
Venice
2026-09-17
Nano Banana 2 Lite
Google
2026-09-16
Union Alpha
Stealth
2026-09-15
Jev 1.13
TypeSafe AI
2026-09-12
Schematron V2 Turbo
Inference.net
2026-09-12
Schematron V2 Small
Inference.net

🧠 模型发布时间线

2026-09-17
KAT-Coder-Pro V2.5
Kwaipilot
aimlapi
2026-09-17
Gemini Omni Flash Preview
Google
aimlapi
2026-09-17
Venice Uncensored
Venice
aimlapi
2026-09-17
Nano Banana 2 Lite
Google
aimlapi
2026-09-16
Union Alpha
Stealth
aimlapi
2026-09-15
Jev 1.13
TypeSafe AI
aimlapillmgateway
2026-09-12
Schematron V2 Turbo
Inference.net
aimlapi
2026-09-12
Schematron V2 Small
Inference.net
aimlapi
2026-09-11
Fugu Ultra v2.0
Sakana AI
llmgateway
2026-09-11
Kimi K2.8 Preview
Moonshot AI
llmstats
2026-09-11
Atria Dawn Preview
Shanghai AI Laboratory
llmstatsllmgateway
2026-09-11
Fugu Ultra v2
Sakana AI
aimlapi
2026-09-11
Fugu Max
Sakana AI
aimlapillmgateway
2026-09-10 · ★
Ling 3.0 Flash VL
inclusionAI
aimlapiopper
2026-09-10 · ★
DeepSeek V4.1 Flash
DeepSeek AI
aimlapillmstatsopperllmgateway
2026-09-10
DeepSeek Chat (V4.1 Flash)
DeepSeek AI
aimlapi
2026-09-08
GPT Image 2.5 Sunburst
Open AI
aimlapillmgateway
2026-09-08
GPT Image 2.5 Flare
Open AI
aimlapillmgateway
2026-09-08
Mercury 2.5
Inception
aimlapi

📄 论文 PwC + arXiv

🤗 HF 采用榜 下载/点赞

Modeldownloadslikespipeline
prism-ml/Ternary-Bonsai-2-27B-gguf1,908,3961,417text-generation
deepseek-ai/DeepSeek-V4.1-Flash496,6843,395image-text-to-text
convaiinnovations/laya0905text-classification
XingChen-AGI/Xing4.0-29B-A4B12,617818text-generation
Qwen/Qwen3.8-27B7,331,93215,837image-text-to-text
m-a-p/YuE2-3B17,403903text-to-audio
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF1,217,2041,465image-text-to-text
harshatheg/Qwen-2.5-1B-RLCD0469text-generation
Lightricks/LTX-2.51,609,5594,517image-to-video
ukisai/Swift-Qwen3.8-27b10,962504image-text-to-text
Qwen/Qwen-Image-2.1183407text-to-image
DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF1,301,417994image-text-to-text
unsloth/Qwen3.8-27B-GGUF6,941,4784,412
openbmb/MiniCPM5-2B420,6221,618text-generation
AlexWortega/openjev0275text-classification
Qwen/Qwen3.8-Flash-Next761,1125,482image-text-to-text
ukisai/Swift-Qwen3.8-27B-GGUF136,668323image-text-to-text
prism-ml/Ternary-Bonsai-2-27B-mlx-2bit30,043272text-generation
meta-llama/Llama-3.1-8B-Instruct5,910,1027,764text-generation
MiniMaxAI/MiniMax-H34,057,4445,515image-text-to-video
TokenRhythm/NeoHorse-1-9B11,913974text-generation
internlm/Atria-Dawn-Preview895207
TaichuAI/ZDTaichu5.0-9B3,750209image-text-to-text
Edge0/Edge0-35B-A3B-preview76,6693,546text-generation
WarmBloodAban/Minimax-h3_Singularity242,751560image-to-video
dealignai/DeepSeek-V4.1-Flash-UNCENSORED-FP834,688321image-text-to-text
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF42,965184image-text-to-text
zai-org/GLM-5.3-Flash3,109,0842,493image-text-to-text
Comfy-Org/Qwen-Image-2.1120165
Mothersuperior/yue2-mothersuperior-realaudio-tokenizer-v40158

⚡ System One 决策模型 334 项目 · Jev/TypeSafe 生态 · 快决策(非推理)

SDK & Decision Frameworks 59High-Frequency & Simulation 29Evaluation & Observability 29Browser & OS Action 26Routing & Cost Optimization 26Security & Guardrails 23Domain & Vertical Tools 22Context GC & Filter 20MCP & Integrations 18CLI & Pipelines 18Data & Search 18Decision Tools 12Codebase & Graph Pathfinding 11Creative Tools 11SDK & Integrations 6Voice & Conversation 4Classification & Taxonomy 2
项目类别Jev 决策点
langchain langchain-ai146,634SDK & IntegrationsSubmits binary, categorical and ordered-score questions and returns typed answers with probabilities.
ai-hedge-fund virattt63,505Domain & Vertical ToolsConverts strategy questions to System One requests and normalizes native answers to the project’s result format.
litellm BerriAI59,123Routing & Cost OptimizationMaps requests to configured complexity classes that drive backend routing.
oh-my-pi can135731,850Routing & Cost OptimizationSends agent state and typed questions to Jev and parses structured answers.
jev-model-router davila730,779Routing & Cost OptimizationEvaluates task tier, reasoning needs and production risk; local policy maps results to invocation settings.
composio ComposioHQ30,238SDK & Decision FrameworksTurns tool or action conditions into structured questions and passes Jev answers to local invocation logic.
ai vercel26,835SDK & Decision FrameworksMaps choice, score, and yes/no questions to TypeSafe System One requests and parses typed results.
cua trycua23,683Browser & OS ActionReads DOM or supported visual-region descriptions and returns a supplied candidate action ID.
pydantic-ai pydantic20,035SDK & IntegrationsConverts supported structured output fields into typed Jev questions and maps answers back to the output model.
eliza elizaOS19,361SDK & Decision FrameworksOnly an explicit systemOne call sends state and questions, returning validated typed answers.
langchainjs langchain-ai18,210SDK & Decision FrameworksUses invoke to call TypeSafe and parse choice, noul, score and probability fields.
json-render vercel-labs16,572Creative ToolsEvaluates component configurations through Vercel AI Gateway, then composeSpec assembles the UI specification.
openchamber openchamber10,060Routing & Cost OptimizationJev selects a task category; local category mappings determine the model configuration.
rig-typesafeai 0xPlaygrounds8,669SDK & Decision FrameworksSends application state and questions to Jev and parses Choice, Score or Noul answers.
firstmate kunchenguid6,587Routing & Cost OptimizationSends the task brief and candidate rules to Jev, then resolves execution profiles with confidence and local conditions.
jev-ultrafast browser-use6,031Browser & OS ActionChooses an action and its matching DOM element in one request; a text model generates input text.
agentgateway agentgateway4,926Security & GuardrailsJev scores jailbreaks, harmful content and secret disclosure; thresholds or evaluation errors reject requests.
latitude-llm latitude-dev4,655Evaluation & ObservabilityJudges which checks apply and can add checks when thresholds and rate limits permit.
fast-jev-compaction tamaratran3,528Context GC & FilterSeparately judges whether a tool call and its full output are still needed; code keeps, truncates or drops them.
ax ax-llm2,926SDK & IntegrationsMaps supported signatures to Jev questions or sends native System One requests.
vellum-assistant vellum-ai1,287MCP & IntegrationsSubmits state and question bundles to System One and returns structured answers to the Assistant.
jev-desktop lahfir1,275Browser & OS ActionJev selects a target and action and estimates presence and risk; local policy decides whether to execute.
celesto CelestoAI943Codebase & Graph PathfindingJudges whether a finding was introduced by the change, is supported and merits a fix.
jev-trader jarrodwatts936Domain & Vertical ToolsIn Jev mode, order-book judgments feed code that simulates fills or submits configured post-only limit orders.
atomic bastani-inc806Routing & Cost OptimizationSends predefined questions to Jev and decodes answers for callers; regular models still generate code.
aiavatarkit uezo676Voice & ConversationAssesses utterance completeness and whether the user is likely to continue speaking.
NanoJev TianyuCodings657High-Frequency & SimulationEvaluates multiple questions and dynamic candidate spaces concurrently in a single forward pass, logging navigation choices.
kody kentcdodds654Data & SearchSends a Score question per candidate, reorders and drops low scores; the model id is typesafe/jev.
Agent AgentiLoop616Security & GuardrailsAdds a destructive-risk judgment after local shell checks and refuses commands above the configured threshold.
req_llm agentjido577SDK & Decision FrameworksSends state and questions, normalizes answers and retains the raw provider response.
omg.dev BennyKok531Browser & OS ActionChooses controls and checks completion or blockage before the test runner operates the UI.
vexjoy-agent notque420Routing & Cost OptimizationAfter deterministic routing guards, Jev judges the remaining candidates and required workflow components.
foreman thruwire344CLI & PipelinesAsyncTypeSafeClient.system_one with default jev-latest sends Noul questions for supervision.
WrongStack WrongStack329Routing & Cost OptimizationJev evaluates the task against eligible specialists; local dispatch rules use the result.
instructor-php cognesy327SDK & Decision FrameworksConverts application state and typed questions into Jev requests and maps responses to PHP decision objects.
kev jaredpalmer310High-Frequency & SimulationAttaches a parallel decision head to an open 0.5B model to answer discrete questions directly from token activations.
typesafe-computer-use awlevin302Browser & OS ActionSelects the next step from deterministically extracted controls and actions before desktop execution.
Jev Review devagrawal09284Codebase & Graph PathfindingJudges risk, files, evidence regions, mechanisms and severity before rule-based reviewer routing.
orchestkit yonatangross278CLI & PipelinesClassifies the first task prompt and branch state by work type; local policy accepts the result or falls back.
typesafe-mario fhshaik266High-Frequency & SimulationReads motion, enemies, terrain and recent controls, then selects a predefined legal action.

🅱️ B站 AI 无限竞技场 18 模型 · 夺冠率

#模型夺冠率冠/测
1GPT-6 Astra OpenAI55%12/22
2Claude Fable 5.1 Anthropic50%6/12
3GLM-5.3 Z.ai16%3/19
3GPT-5.6 Sol OpenAI16%3/19
5Claude Fable 5 Anthropic27%3/11
6Kimi K3 Moonshot10%2/21
7Claude Opus 5 Anthropic11%2/19
8DeepSeek-V4-Flash DeepSeek13%2/16
9DeepSeek-V4-Pro DeepSeek5%1/21
10Qwen3.8-Max Alibaba6%1/18
11DeepSeek V4.1 Flash DeepSeek8%1/13
12Gemini 3.8 Flash Google8%1/12
13Hy 4 Tencent14%1/7
14Gemini 3.7 Flash Google20%1/5
15GPT-5.6 Terra OpenAI25%1/4
15Seed-2.0 pro ByteDance25%1/4
17Doubao-Seed-Evolving ByteDance33%1/3
18Seed-2.0 Mini ByteDance100%1/1
👤 AI UP主 从赛题发现,点击直达主页
🎯 赛题

🎬 AI 视频 订阅 · 搜索 · B站,分开组织

📌 订阅频道 4 个博主 · 最新上传
📺 Best Partners TV 13
📺 AI超元域 12
🔎 搜索发现 相关度+播放量筛选 · 非订阅
🅱️ B站 AI 竞技场 按播放量

🚀 产品发布 whatships · What's Launch

📦 版本发布 tracked repos releases

🛰️ Skywork 动态

🧪 Show HN 开发者发布的新产品

💰 商业动态 · TechCrunch/VB/MIT TR