Mozilla 发布《开源 AI 状态报告》V1.0:开源权重已与闭源达到能力对等,推理成本 36 个月降 50 倍
🔥 今日热点 TOP 5
- 🔴 🟢 Mozilla 发布《开源 AI 状态报告》V1.0:开源权重已与闭源达到能力对等,推理成本 36 个月降 50 倍 — Mozilla 发布首份《State of Open Source AI》报告,称开源权重模型在编码/指令遵循上已与闭源持平,Chatbot Arena 能力差距从 8.04% 收窄至 3.3%(集中在推理),OpenRouter 上排名前五的模型全部是开源权重,中国开源模型周 token 量约为美国闭源的 3 倍 — 首次报道 07-17 14:31 UTC
- 🔴 🟡 Apple 向数十名 OpenAI 员工发出法律保留函,人才争夺战升级为诉讼前奏 — Financial Times 报道 Apple 已向跳槽至 OpenAI 的数十名前员工发出"文件保留函"(document retention letters),要求保留相关记录,这是提起诉讼前的标准动作,矛头直指 OpenAI 系统性挖角及可能的知识产权/商业机密问题 — 首次报道 07-17 12:02 UTC
- 🟢 Claude Code 被曝静默上线"60 秒无人应答就自动继续"功能,作者逐 commit 拆解 Anthropic 的"misfeature" — 一篇在 HN 引发热议的 29 分钟深度技术博客,作者发现 Anthropic 7 月 1 日在 Claude Code 2.1.198 中悄悄上线"用户 60 秒不回应就自行继续"的效率绕过功能,且未写入 changelog、无文档化的关闭开关,几天后才修复,引发对 AI 代理自主性边界和信任的讨论 — 首次报道 07-17 14:26 UTC
- 🟢 Capital One 开源 VulnHunter:基于 Claude Opus 4.8 的攻击者视角 agentic 代码安全工具 — Capital One 发布开源 agentic AI 代码安全工具 VulnHunter,采用攻击者视角的推理工作流主动挖掘可利用漏洞、映射攻击路径并给出针对性修复,已在内部数千仓库验证,模型优化基于 Claude Opus 4.8 — 首次报道 07-17 12:42 UTC
- 🟢 AI 审计员 zkao 发现 OpenVM zkVM 配对库关键漏洞(CVE-2026-46669) — ZK Security 发布"AI Meets Cryptography 2"系列第二篇,其 AI 审计器 zkao(Claude Opus 4.7 + Codex 5.4 驱动)扫描 9.5 小时后在 OpenVM 的 openvm-pairing 库中发现一个 Critical 级别漏洞,允许恶意 prover 伪造任意配对等式,已修复于 OpenVM 1.6.0 — 首次报道 07-17 14:21 UTC
📰 详细资讯
1. Mozilla 发布《开源 AI 状态报告》:开源权重已与闭源能力对等
- 摘要:Mozilla 于 7 月 17 日发布首份《State of Open Source AI》报告(V1.0,由 CTO Raffi Krikorian 作序),核心结论是"能力对等已达成的、竞赛上移了一层"。报告基于 Chatbot Arena 24 个月数据指出:开源与闭源的能力差距从 2024 年 1 月的 8.04%,到 2024 年 8 月一度收窄至 0.5%,2025 年 2 月 DeepSeek-R1 曾短暂持平美国顶级模型,到 2026 年 3 月重新拉开到 3.3%——但这 3.3% 集中在推理、长上下文检索和 agentic 任务,而在编码、指令遵循、通用知识上开源已达到或接近持平。价格层面:GPT-4 级别推理成本在 36 个月内从 $20/1M token 降至 $0.40,跌幅约 50 倍,比.com 时代的带宽或 PC 算力曲线都要陡。采用层面:OpenRouter 上开源权重模型的 token 路由份额已从 2024 年末的约 33% 升至 2026 年中的过半,当前交易量前五的模型全部是开源权重——DeepSeek V4 Flash (18.4T)、Xiaomi MiMo-V2.5 (14.9T)、Tencent Hy3 preview (14.8T)、MiniMax M3 (14.3T)、Owl Alpha (11T),而 Anthropic 的 Claude Opus 4.7 (9.02T) 是最高排名的闭源模型。Mozilla/SlashData 2026 开发者调查显示 79% 添加 AI 功能的开发者使用开源模型(vs 71% 闭源),但只有 51% 的开源团队进入生产(闭源为 63%)——瓶颈是运维工具与信任,而非模型能力。报告由此提出五大押注:构建开源 harness、拥有记忆、解决可移植权限、打破按量计费、让开放默认多元化。
- 原文链接:https://stateofopensource.ai/
- 信源验证:
- ✅ [Mozilla 官方] The state of open source AI (https://stateofopensource.ai/) — 07-17 14:31 UTC
- ✅ [Hacker News] The state of open source AI (https://news.ycombinator.com/item?id=48947825) — 07-17 14:31 UTC — 329 points, 236 comments
- ✅ [OpenRouter 公开榜单] Trailing month tokens routed (报告内引用) — 截至 07-17
- 热度指标:HN 329 upvotes / 236 comments(当日 AI 类 HN 第 2 热帖)
- 社媒热评:
-
“Speculation: open models is what will kill Anthropic and OpenAI. Hyperscalers can run the models without a licensing fee. Apple can make them smaller and put them on the device. The frontier models are an edge and a liability. They’re astronomically expensive.” — @babblingfish @HackerNews
-
“Just like how the web was won? I think Mozilla is chasing a past formula, but the projection isn’t linear enough to remain consistent.” — @positron26 @HackerNews
-
(UI 争议为主流吐槽)“This new trend of content appearing while scrolling down is so terrible accessibility-wise, I do not understand how Mozilla of all institutions would do it.” — @hypfer @HackerNews
-
- 标签:#开源AI #Mozilla #OpenRouter #DeepSeek #能力对等 #推理成本 #harness
- 时效性:🟢 突发 — 首次报道于 07-17 14:31 UTC
2. Apple 向数十名 OpenAI 员工发出法律保留函
- 摘要:据 Financial Times 7 月 17 日独家报道,Apple 已向跳槽至 OpenAI 的数十名前员工发出"文件保留函"(document retention / legal hold letters),要求他们保留与在 Apple 工作期间相关的文件和通讯记录。这类信件是提起正式诉讼前的标准预备动作,意味着 Apple-OpenAI 的人才/知识产权争端正从口水战升级为潜在的法律诉讼。报道背景是 OpenAI 近几个月大规模从 Apple 硬件/AI 团队挖人(包括设计、芯片、机器人等岗位),而 Apple 一直在系统性地反制——此前已有多轮公开交锋。HN 评论区对此讨论激烈:有人认为文件保留函是例行公事,也有人指出 Apple 若无硬证据不会走到这一步;还有人预测这场争端可能冲击 OpenAI 的 IPO 进程。
- 原文链接:https://www.ft.com/content/1b8c9d52-88a9-426b-ba47-f1811f859166
- 信源验证:
- ✅ [Financial Times] Apple targets dozens of OpenAI employees with legal letters (https://www.ft.com/content/1b8c9d52-88a9-426b-ba47-f1811f859166) — 07-17 12:02 UTC
- ✅ [Hacker News] Apple targets dozens of OpenAI employees with legal letters (https://news.ycombinator.com/item?id=48946303) — 07-17 12:02 UTC — 361 points, 303 comments
- ✅ [archive.ph 镜像] (https://archive.ph/3J3iw) — 07-17
- 热度指标:HN 361 upvotes / 303 comments(当日 HN 前 5 热帖)
- 社媒热评:
-
“FT frames this as some aggressive escalation tactic, but document retention letters are extremely standard practice. At this point they’re basically a formality, as any former Apple employee at OpenAI really ought to know by now that they could get dragged into this.” — @deepwoods @HackerNews
-
“Apple must have hard evidence on this. I can’t believe they would take it this far without already knowing they are going to win. If they have to fire a huge chunk of their hardware employees it’s going to throw their IPO plans into chaos.” — @reenorap @HackerNews
-
“Predictions on who wins? Does Apple actually have a winnable case or are they just throwing a wrench in things?” — @bix6 @HackerNews
-
- 标签:#Apple #OpenAI #诉讼 #人才争夺 #知识产权 #IPO
- 时效性:🟡 跟进 — Apple-OpenAI 争端已持续数日,07-17 升级为法律保留函
3. Claude Code “Anatomy of a Misfeature”:逐 commit 拆解 Anthropic 静默上线的 60 秒自动继续功能
- 摘要:一篇在 HN 引发热议(131 points, 115 comments)的 29 分钟深度技术博客。作者 Olaf Alders 发现:2026 年 7 月 1 日(加拿大国庆日)Anthropic 在 Claude Code
2.1.198版本中悄悄上线了一个"效率绕过"功能——当 Claude Code 向用户询问输入后,若用户 60 秒内未回应,代理会"helpfully"自动用其认为最佳的判断继续执行,并打印"The user stepped away. I’ll proceed with best judgment."。作者列出了令人担忧的后果:用户去厨房做个三明治时会发生什么?同时运行多个代理时如何同时监控?代理做出错误选择时已烧掉多少 token?如果用于部署会发生什么?更严重的是该功能未出现在 changelog 中,且没有文档化的关闭开关。文章逐 commit diff 了版本差异,分析了该功能"是否由人类设计/审查/合并/记录",并指出关闭它需要关掉 auto-update(副作用是连插件更新也一并停止)。作者结论:“move fast and break things 并不一定排除 move fast and fix things”,几天内修复已推送,但用户对该产品的信任已受冲击。这引发了对 AI 代理"自主性边界"的广泛讨论。 - 原文链接:https://www.olafalders.com/2026/07/17/claude-code-anatomy-of-a-misfeature/
- 信源验证:
- ✅ [olafalders.com] Claude Code: Anatomy of a Misfeature (https://www.olafalders.com/2026/07/17/claude-code-anatomy-of-a-misfeature/) — 07-17 14:26 UTC
- ✅ [Hacker News] Claude Code: Anatomy of a Misfeature (https://news.ycombinator.com/item?id=48947776) — 07-17 14:26 UTC — 131 points, 115 comments
- ✅ [GitHub Anthropic claude code] 版本 2.1.198 发行记录(文中引用)
- 热度指标:HN 131 upvotes / 115 comments
- 社媒热评:
-
“It is much worse than that. Claude Code doesn’t auto-commit when stopping for an answer. There might be possible data loss if an uncommitted file is edited.” — @maxloh @HackerNews
-
“I really hate this direction both Anthropic and OpenAI are following. They are in this silly competition whose model/harness can go unattended the longest, no matter what. And it is never explicit, you learn about it after you get bitten by it.” — @ayhanfuat @HackerNews
-
“Increasingly getting frustrated with Anthropic so not a fanboy but I find this feature great for my workflows.” — @joshuafuller @HackerNews
-
“The underlying model is excellent, but we’re being pretty much forced into using CC to use the subscription model.” — @petesergeant @HackerNews
-
- 标签:#ClaudeCode #Anthropic #AI代理自主性 #安全 #信任 #auto-update
- 时效性:🟢 突发 — 首次报道于 07-17 14:26 UTC
4. Capital One 开源 VulnHunter:攻击者视角的 agentic 代码安全工具
- 摘要:Capital One 于 7 月 17 日宣布开源 VulnHunter——一款先进的 agentic AI 安全工具,采用攻击者视角的推理工作流直接分析源代码。与传统被动漏洞扫描器不同,VulnHunter 不是简单报一堆可能问题,而是通过 agentic 推理工作流去识别可利用缺陷、映射潜在攻击路径,并提出高度针对性的代码修复。其技术亮点包括"先内部挑战再上报"——每个被标记的漏洞都是工具尝试排除但无法排除的结果,从而大幅减少误报,让开发者专注于有证据支撑的即时修复。Capital One 在发布前已在内部数千个仓库、数十个业务领域运行验证,过去需要大量人工分诊的工作现在能快速产生可验证的发现。VulnHunter 模型优化基于 Claude Opus 4.8,以 Claude Code skill 形式实现,采用 Apache 2.0 许可证,已在 GitHub (capitalone/vulnhunter) 开源。Capital One 强调这是"集体防御"——现代软件供应链高度互联,一个广泛使用的开源组件漏洞可同时波及数千企业。
- 原文链接:https://www.capitalone.com/tech/open-source/announcing-vulnhunter/
- 信源验证:
- ✅ [Capital One Tech] Announcing VulnHunter (https://www.capitalone.com/tech/open-source/announcing-vulnhunter/) — 07-17 12:42 UTC
- ✅ [GitHub] capitalone/vulnhunter (https://github.com/capitalone/vulnhunter) — 07-17,Apache 2.0
- ✅ [Hacker News] VulnHunter: Capital One’s agentic AI code security tool (https://news.ycombinator.com/item?id=48946692) — 07-17 12:42 UTC — 54 points, 29 comments
- 热度指标:HN 54 upvotes / 29 comments
- 标签:#VulnHunter #CapitalOne #agentic安全 #开源 #ClaudeOpus #代码审计
- 时效性:🟢 突发 — 首次报道于 07-17 12:42 UTC
5. AI 审计员 zkao 发现 OpenVM zkVM 配对库关键漏洞 (CVE-2026-46669)
- 摘要:ZK Security 发布"AI Meets Cryptography"系列第二篇,展示了 AI 审计在真实加密项目中的应用。其 AI 审计器 zkao(由 Claude Opus 4.7 + Codex 5.4 驱动,融合了团队专家的方法论作为可复用流程)被指向 OpenVM 的 zkVM 进行审计。团队此前用普通 LLM(带简单 prompt 和专家 skills)扫描了 4 个月,Opus 4.6/Codex 5.3 到 Opus 4.7/Codex 5.4 都跑过,候选发现都是有效观察但无一可利用——团队假设 zkVM 的模块间依赖密度远超普通库,“一个可证明安全的模块 A 和可证明安全的模块 B 的组合可能并不安全”,而现有 agentic 编码工具(Claude Code、Codex)在"将子代理输出表示为模块知识"这件事上尚未解决好。改用 zkao 后扫描 9.5 小时即返回关键发现:openvm-pairing 库的配对检查缺少对缩放因子的子域校验,允许恶意 prover 伪造任意配对等式(注意:这不是 zkVM 证明系统本身的 soundness bug,只影响使用该库的代码)。漏洞被分配 CVE-2026-46669,已修复于 OpenVM 1.6.0。ZK Security 强调 AI 产出的是候选发现而非最终报告,人类团队做了验证、影响评估和披露。
- 原文链接:https://blog.zksecurity.xyz/posts/openvm-bugs/
- 信源验证:
- ✅ [ZK/SEC Quarterly] AI meets Cryptography 2: What AI Found in OpenVM’s zkVM (https://blog.zksecurity.xyz/posts/openvm-bugs/) — 07-17 14:21 UTC
- ✅ [Hacker News] AI Meets Cryptography 2: What AI Found in OpenVM’s ZkVM (https://news.ycombinator.com/item?id=48947714) — 07-17 14:21 UTC — 76 points
- ✅ [CVE 记录] CVE-2026-46669(文中引用,已修复于 OpenVM 1.6.0)
- 热度指标:HN 76 upvotes
- 标签:#zkao #ZK安全 #OpenVM #CVE #AI审计 #ClaudeOpus #零知识证明
- 时效性:🟢 突发 — 首次报道于 07-17 14:21 UTC
6. Simon Willison 评 Kimi K3:从"鹈鹕基准"看新模型,但该基准已基本失效
- 摘要:Simon Willison 在 Kimi K3 发布次日发文,用其著名的"鹈鹕骑自行车 SVG"测试实测了 K3。要点:(1) K3 通过 OpenRouter 测试,生成一个 SVG 花 95 输入 token + 16,658 输出 token(其中 13,241 是推理 token),总成本 25 美分;(2) K3 接受图像输入,对渲染后的 SVG 用 alt text prompt 测试花了 0.6 美分,视觉理解能力良好;(3) 定价 $3/1M 输入、$15/1M 输出,与 Anthropic Claude Sonnet 系列持平,是迄今中国 AI 实验室发布的最贵模型,比其上一代 Kimi K2.6 ($0.95/$4) 大幅上涨;(4) 2.8 万亿参数是上一代 1T 模型的两倍多。Willison 坦承他 21 个月前发明的"鹈鹕基准"现已基本失效——GPT-5.6 和 Claude Fable 5 的鹈鹕竟被 GLM-5.2 超越,而他并不认为 GLM 是 Fable 级别的模型。他指出该基准最大的局限是完全无法触及今天模型最关键的能力:agentic 工具调用以及在长对话中可靠操作工具的能力。不过他仍认为该测试有"强迫自己实际试用模型"的价值,并提到 K3 目前只有一个思考强度级别。
- 原文链接:https://simonwillison.net/2026/Jul/16/kimi-k3/
- 信源验证:
- ✅ [Simon Willison’s Weblog] Kimi K3, and what we can still learn from the pelican benchmark (https://simonwillison.net/2026/Jul/16/kimi-k3/) — 07-16 20:19 (发布),07-17 HN 发酵
- ✅ [Hacker News] Kimi K3, and what we can still learn from the pelican benchmark (https://news.ycombinator.com/item?id=48947717) — 07-17 14:21 UTC — 222 points, 124 comments
- ✅ [Artificial Analysis] Kimi K3 报告(文中引用)
- 热度指标:HN 222 upvotes / 124 comments
- 标签:#KimiK3 #SimonWillison #鹈鹕基准 #定价 #agentic
- 时效性:🟡 跟进 — Kimi K3 昨日发布,今日 Simon Willison 深度评测
7. Agent-talk:让多个编码代理相互对话协作的 Claude Code 插件
- 摘要:开源项目 agent-talk 是一个面向编码代理(如 Claude Code)的插件,让代理之间能够互相发消息、协调任务。背景痛点:大型项目需要跨多个会话并行运行编码代理,往往还要与其他开发者的代理协作,但它们之间没有沟通渠道,用户被迫充当"信使"在窗口间手动复制指令。agent-talk 让代理能直接互发消息协调底层实现,让用户专注于高层细节。它基于 retalk CLI 构建,支持 Claude Code 的插件机制,可通过
claude plugin marketplace add xhluca/agent-talk安装,并提供公共 relay(relay.retalk.dev,尽力而为无 SLA)或自建 relay。这是"多代理协作"工具栈的又一实证,与 Mozilla 报告中"harness 是新前沿"的论断呼应。 - 原文链接:https://github.com/xhluca/agent-talk
- 信源验证:
- ✅ [GitHub] xhluca/agent-talk (https://github.com/xhluca/agent-talk) — 07-16/17
- ✅ [Hacker News] Agent-talk: Enabling coding agents to work together (https://news.ycombinator.com/item?id=48936534) — 07-16 16:14 UTC — 53 points, 23 comments
- 热度指标:HN 53 upvotes / 23 comments
- 标签:#agent-talk #多代理协作 #ClaudeCode插件 #harness
- 时效性:🟢 突发 — 07-16 发布,07-17 持续讨论
8. 腾讯 Hy3 全量开源:295B MoE 模型(21B 活跃),幻觉率从 12.5% 降至 5.4%
- 摘要:腾讯 Hy 团队于 7 月 17 日在 HuggingFace/ModelScope/GitCode/CNB 全量开源 Hy3(及 FP8 量化版)。Hy3 是 295B 参数 MoE 模型,21B 活跃参数 + 3.8B MTP 层参数,192 个专家 top-8 激活,256K 上下文,Apache-2.0 许可。继 4 月底 Hy3 Preview 发布后,团队收集了 50+ 产品反馈,用更高质量数据扩大后训练。核心卖点:(1) 更强的 Agent 能力——团队组织 270 名专家盲评,Hy3 得分 2.67/4,超过 GLM-5.1 的 2.51/4,优势在前端开发、数据存储、CI/CD 任务;(2) 更可靠的产品体验——工具调用稳定性达生产级,SWE-Bench Verified 在 CodeBuddy/Cline/KiloCode 等不同脚手架间准确率方差在 4% 以内;幻觉率从 12.5% 降至 5.4%,常识错误率从 25.4% 降至 12.7%;多轮意图追踪问题率从 17.4% 降至 7.9%。Hy3 在 Mozilla 报告的 OpenRouter 排行榜上以 14.8T 周 token 排名第 3(仅次于 DeepSeek V4 Flash 和 Xiaomi MiMo-V2.5),是交易量最高的腾讯模型。
- 原文链接:https://huggingface.co/tencent/Hy3
- 信源验证:
- ✅ [HuggingFace] tencent/Hy3 (https://huggingface.co/tencent/Hy3) — 07-17 开源(约 12 小时前更新)
- ✅ [GitHub] Tencent-Hunyuan/Hy3 (https://github.com/Tencent-Hunyuan/Hy3)
- ✅ [Mozilla 报告引用] OpenRouter 排行榜 Hy3 preview 14.8T 周 token — 07-17
- 热度指标:HF 819 likes / 11.3k followers;OpenRouter 14.8T 周 token(第 3)
- 标签:#腾讯 #Hy3 #MoE #开源 #Agent #幻觉率 #SWE-Bench
- 时效性:🟡 跟进 — Preview 于 4 月,全量开源 07-17
🛠️ GitHub Trending AI 项目
| 排名 | 项目 | 星标 | 描述 | 今日新增 | 链接 |
|---|---|---|---|---|---|
| 1 | Nutlope/hallmark | ⭐ 11,940 | Anti-AI-slop design skill for Claude Code, Cursor, and Codex | +1,486 | GitHub |
| 2 | OpenCut-app/OpenCut | ⭐ 74,782 | The open-source CapCut alternative | +1,077 | GitHub |
| 3 | codecrafters-io/build-your-own-x | ⭐ 527,247 | Master programming by recreating your favorite technologies from scratch | +1,070 | GitHub |
| 4 | HKUDS/DeepTutor | ⭐ 27,312 | DeepTutor: Lifelong Personalized Tutoring | +528 | GitHub |
| 5 | PostHog/posthog | ⭐ 36,165 | AI 可观测性/分析/会话回放/实验平台,MCP 驱动 | +437 | GitHub |
| 6 | openinterpreter/openinterpreter | ⭐ 66,314 | A coding agent for open models like Kimi K3 | +431 | GitHub |
| 7 | RyanCodrai/turbovec | ⭐ 13,265 | 基于 TurboQuant 的向量索引,Rust 编写带 Python 绑定 | +280 | GitHub |
| 8 | PrismML-Eng/Bonsai-demo | ⭐ 1,698 | Bonsai Demo(三值化量化模型演示) | +279 | GitHub |
| 9 | HenryNdubuaku/maths-cs-ai-compendium | ⭐ 6,575 | Become a cracked AI/ML Research Engineer | +248 | GitHub |
| 10 | github/copilot-sdk | ⭐ 9,780 | Multi-platform SDK for integrating GitHub Copilot Agent | +234 | GitHub |
| 11 | tirth8205/code-review-graph | ⭐ 19,711 | 本地优先的代码知识图谱(MCP/CLI),为 AI 编码工具减上下文 | +57 | GitHub |
| 12 | docusealco/docuseal | ⭐ 17,801 | 开源 DocuSign 替代品 | +152 | GitHub |
| 13 | anthropics/cwc-workshops | ⭐ 1,563 | Anthropic 官方 Claude 工作坊材料 | +37 | GitHub |
🤗 HuggingFace Trending Models
| 排名 | 模型 | 参数量 | 下载量 | 描述 | 链接 |
|---|---|---|---|---|---|
| 1 | thinkingmachines/Inkling | 952B | 7.87k | Thinking Machines 首个开源多模态 MoE(持续高热,+945 likes) | HF |
| 2 | empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF | 9B | 2.1M | 基于 Claude Mythos 5 蒸馏的 1M 上下文模型(+2.27k likes) | HF |
| 3 | zai-org/GLM-5.2 | 753B | 535k | 智谱 GLM-5.2 超大模型,多语言多模态(+4.07k likes) | HF |
| 4 | prism-ml/Bonsai-27B-gguf | 4B | 1.05M | Bonsai-27B GGUF 量化版(6 小时前更新) | HF |
| 5 | tencent/Hy3 | 295B (21B 活跃) | 12.7k | 腾讯 Hy3 MoE 模型,今日全量开源(+819 likes) | HF |
| 6 | OpenMOSS-Team/MOSS-Transcribe-Diarize | 0.9B | 83.2k | 复旦 MOSS 语音转写+说话人分离模型 | HF |
| 7 | bottlecapai/ThinkingCap-Qwen3.6-27B | 27B | 9.38k | Qwen3.6 思考增强版,推理能力提升 | HF |
| 8 | InternScience/Agents-A1 | 35B | 34.1k | 35B Agent 专用模型(+571 likes) | HF |
| 9 | Wan-AI/Wan-Dancer-14B | 14B | 2.19k | 阿里 Wan 图生视频模型(13 小时前更新,新上榜) | HF |
| 10 | ATH-MaaS/OvisOCR2 | 0.9B | 10.8k | 轻量级 OCR 视觉模型 | HF |
注:Kimi K3 模型权重仍未在 HuggingFace 发布(预计 7 月 27 日),故未上榜。
🚀 Product Hunt AI 热门
| 排名 | 产品 | 描述 | 链接 |
|---|---|---|---|
| 1 | — | Product Hunt 再次被 Cloudflare 人机验证拦截,本次未能自动采集 | — |
注:Product Hunt 连续两日被 Cloudflare 拦截。建议手动浏览 https://www.producthunt.com/topics/artificial-intelligence
📚 arXiv / 研究论文精选
| 论文 | 作者/机构 | 摘要 | 链接 |
|---|---|---|---|
| EEG shows brain can simultaneously encode two speech streams | PLOS Biology | 脑电研究显示人脑可同时编码两条语音流,对语音分离/多模态 AI 有启发 | PLOS Biology |
注:arXiv 当日无单一现象级论文爆发,web_search API 持续不可用导致无法系统检索 arXiv 热门。明日补充。
📈 热度追踪
| 话题 | 连续天数 | 趋势 | 今日动态 |
|---|---|---|---|
| 开源 vs 闭源能力对等 | 2 天 | 🔥🔥 急升 | Mozilla 官方报告量化论证:能力差距仅 3.3%,OpenRouter 前五全开源,中国开源 token 量 3 倍于美国闭源 |
| 开源大模型规模竞赛 | 5 天 | 🔥 急升 | 腾讯 Hy3 295B 全量开源;OpenRouter 排行榜显示 DeepSeek V4 Flash/MiMo-V2.5/Hy3/MiniMax M3/Owl Alpha 占据前五 |
| Apple vs OpenAI 诉讼 | 7 天 | 🔥🔥 急升 | Apple 向数十名 OpenAI 员工发法律保留函,争端升级为诉讼前奏 |
| AI Agent / 编码工具 | 41 天 | 🔥 急升 | Claude Code 静默上线"60 秒自动继续"引信任危机;agent-talk 让多代理互相对话;Capital One 开源 VulnHunter |
| AI 用于安全/漏洞挖掘 | 1 天 | 🆕 新增 | zkao AI 审计器发现 OpenVM CVE-2026-46669;Capital One VulnHunter 攻击者视角代码审计——“AI 攻防"双侧同时升温 |
| AI 安全 / 隐私 / 自主性 | 4 天 | 🔥 上升 | Claude Code 60 秒自动继续功能引发代理自主性边界讨论(接续 Grok Build 隐私争议) |
| 中国 AI 开源模型 | 40 天 | 🔥 急升 | 腾讯 Hy3 全量开源,幻觉率大幅下降;OpenRouter 中国开源模型 token 量领先 |
| 推理成本下降 | 1 天 | 🆕 新增 | Mozilla 量化:GPT-4 级推理 36 个月降 50 倍($20→$0.40/1M token) |
| Gemini 品牌 | 6 天 | ➡️ 稳定 | 昨日 NotebookLM→Gemini Notebook,今日无重大进展 |
| GPT-5.6 系列 | 22 天 | ➡️ 稳定 | Simon Willison 鹈鹕基准提及 GPT-5.6 Sol |
| DeepSeek IPO | 2 天 | ➡️ 稳定 | 今日无新进展(昨日估值 $52B) |
| Kimi K3 | 2 天 | ↗️ 上升 | Simon Willison 深度评测,定价为中国实验室最贵,权重 7/27 发布 |
📊 信源使用统计
| 信源等级 | 数量 | 主要信源 |
|---|---|---|
| S 级(官方博客/公告) | 3 | Mozilla 官方报告, Capital One Tech, Tencent Hy Team (HF) |
| A 级(权威媒体) | 2 | Financial Times, ZK/SEC Quarterly |
| B 级(社区/社媒) | 9 | Hacker News (×8 帖), olafalders.net, simonwillison.net |
| C 级(聚合/分析) | 3 | HuggingFace Trending, GitHub Trending, OpenRouter 榜单 |
📌 本期说明:web_search API (Tavily) 与 web_extract 均不可用(前者 432 错误,后者私网拦截),全部信源通过浏览器直接访问 + HN Algolia API 完成交叉验证。今日最大新闻是 Mozilla《开源 AI 状态报告》——这份报告用 Chatbot Arena 24 个月数据 + OpenRouter 真实 token 流量 + SlashData 开发者调查,给出了"开源已与闭源能力对等、竞赛上移到 agentic harness 层"的强论断,其中"OpenRouter 前五模型全部开源、中国开源 token 量是美国闭源的 3 倍"极具冲击力。与此同时,“AI 攻防"双侧同时升温——攻击面有 Claude Code 静默上线"60 秒自动继续"引发的自主性边界争议(接续 Grok Build 隐私争议);防御面有 Capital One 开源 VulnHunter 和 zkao 发现 OpenVM CVE,显示 agentic AI 在安全领域快速双向渗透。Apple-OpenAI 争端从口水战升级为法律保留函,值得持续跟踪是否演变为正式诉讼及其对 OpenAI IPO 的影响。