MORNING BRIEF
HTML Archive
今日概览

晨报 · 2026-07-18

覆盖 16 个来源、78 条候选资讯,形成 7 个有效分类。飞书卡片负责提醒,HTML 负责完整阅读、归档和回看。

16信息源
78候选资讯
7有效分类
正常运行状态

生成时间 2026-07-18 06:45:53

☀️ 天气

双城速览
📍 北京·朝阳区
Smoky haze · 今日 23~31°
22°
体感 25° · 💧 100% · 🌬 4km/h
26°早晨 · 降雨27%
30°下午 · 降雨10%
26°晚间 · 降雨22%
📍 焦作·山阳区
多云 · 云量100% · 今日 21~31°
22°
体感 26° · 💧 96% · 🌬 0.4km/h

✨ 今日重点

5 条
AI HOT 精选 · 开发者热点

八天四款前沿模型发布,Kimi K3 跻身第三

过去八天内,Grok 4.5、GPT-5.6、Muse Spark 1.1 与 Kimi K3 四款前沿模型相继发布,使 Artificial Analysis Intelligence Index 得分超 50 的实验室从 6 月初的 2 家增至 6 家。(AIHOT源:X:Artificial Analysis (@ArtificialAnlys))

AI HOT 精选 · 科技前沿

苹果与 OpenAI 法律战升级:约 40 名前员工收到苹果律师函

苹果已向约40名就职于OpenAI的前员工发出律师函,要求保存相关文件。此前苹果起诉OpenAI及两名前员工,指控其通过挖角获取商业机密以加速AI硬件研发。苹果称已有超400名前员工在OpenAI工作,正寻求法院禁令阻止OpenAI使用苹果信息并要求归还机密。(AIHOT源:IT之家(RSS))

IT之家 · 国际时事

中国科学家首获门捷列夫国际基础科学奖,潘建伟量子通信成果获国际认可

IT之家 7 月 17 日消息,今日,第三届联合国教科文组织 — 俄罗斯门捷列夫国际基础科学奖颁奖典礼在联合国教科文组织巴黎总部举行。 中国科学院院士、中国科学技术大学教授潘建伟成为首位获此殊荣的中国学者。另一位获奖者为美国北卡罗来纳大学查珀尔希尔校区化学系教授谢尔盖 · 舍伊科(Sergei Sheiko),他因在基础聚合物物理学和材料科学领域的杰出贡献而获奖。 联合国教科文组织在颁奖辞中指出,潘建伟是量子光学、量子通信和量子计算领域的国际领军科学家,因在大尺度安全量子通信和可扩展量子计算方面的开创性贡献而受到表彰。他的团队研制成功“墨子号”量子科学实验卫星,实现了数千公里范...

🤖 AI/大模型

重点 1 · 共 8 条
AI HOT 精选 · AI/大模型

Sora 2 视频克隆效果惊人,真假难辨

一年后,没有任何东西能接近 Sora 的完美视频深度克隆。它捕捉到了我和 Sam 的每一块面部肌肉运动以及我们走路的方式。 如果你从这段关于我或 Sam 的视频中截取一帧,根本无法判断它是真是假。(AIHOT源:X:Gabriel (@gabriel1))

AI HOT 精选 · AI/大模型

美团LongCat发布LoHoSearch:更难搜索智能体基准

美团LongCat推出LoHoSearch,一个基于762万实体维基百科知识图谱自动生成问题的搜索智能体基准,旨在解决BrowseComp等现有基准趋于饱和的问题。在11个前沿模型测试中,最佳得分仅34.74%,远低于当前模型在BrowseComp上约90%的成绩;上下文策略仅带来+6.8个百分点的提升。该基准包含544道问题、11个领域,采用树与图结构,已开源。(AIHOT源:X:美团 LongCat (@Meituan_LongCat))

AI HOT 精选 · AI/大模型

Schema Harness 在 ARC-AGI-3 公开集上取得约 99% 成绩

Schema 框架在 ARC-AGI-3 公开集上,使用 Claude Opus 4.8 和 Fable 5 达到 99% RHAE 分数,使用 GPT-5.6 Sol 达到 95.35%。该框架不修改模型权重,而是将原始观测转化为可编辑程序,联合解决状态归因和机制发现问题。此前最强模型 GPT-5.6 Sol 在半私有集上仅得 7.78%。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))

AI HOT 精选 · AI/大模型

NVIDIA 发布 Nemotron 3 Embed 系列,8B 版本在 RTEB 基准上排名第一

NVIDIA 发布 Nemotron 3 Embed 系列,包含三个开源 checkpoint,其中 8B-BF16 版本在 RTEB 基准上以 78.46 的平均 NDCG@10 排名第一。1B-NVFP4 版本在 Blackwell 上吞吐量比 BF16 高 2 倍,精度保留 99.5%,所有模型最大序列长度 32,768 tokens。(AIHOT源:MarkTechPost(RSS))

AI HOT 精选 · AI/大模型

通义实验室发布 Wan-Streamer v0.2,端到端响应延迟仅 550ms

通义实验室发布 Wan-Streamer v0.2,这是一款将"听、看、说、演"统一进单个 Transformer 的端到端全模态模型。其端到端响应延迟仅 550ms,输出分辨率从 v0.1 的 192×336 提升至 640×368 @ 25FPS,并采用 Thinker-Performer 双通路架构在提升画质的同时维持了极低延迟。(AIHOT源:公众号:通义实验室(千问))

Follow Builders · AI/大模型

Anthropic Engineering: How we contain Claude across products

Twelve months ago, we'd have rejected out of hand the idea of granting Claude access sufficient to take down an internal Anthropic service. Today that level of access is routine, and Anthropic developers are more productive for it. The risk of these deployments has two components: how likely a failure is, and how much damage one could do. Progress on safeguards and model training has steadily driven down the first; the second—the theoretical blast radius—only grows as capabilities and access expand. Yet as agents become capable of doing work that once required a person or even a team, the cost of not deploying grows large enough that the risk-reward calculation tips heavily toward adoption, as long as products can be made safe. The engineering question becomes how to cap the blast radius. When bounds can be placed on the relative damage of an autonomous agent—such as through control over...

Follow Builders · AI/大模型

Sam Altman: we did not have our best last 12 months ever, which is mostly my fault, but w...

we did not have our best last 12 months ever, which is mostly my fault, but we are about to have our best 12 months to date. the team is doing amazing work and i think you’ll be very happy with what they’ve got cooking for you. i am happy about this for many reasons, but mostly because i care about our users winning. AI has to be about giving lots of people more freedom, agency, and wealth. we want to do the right thing, but we do not want to scare people into doing our thing.|@sama · 热度 23005 · 2026-07-16

🐙 GitHub AI趋势

重点 1 · 共 5 条
GitHub AI趋势 · GitHub AI趋势

Graphify-Labs/graphify

AI coding assistant skill (Claude Code, Codex, OpenCode, Cursor, Gemini CLI, and more). Turn any folder of code, SQL schemas, R scripts, shell scripts, docs, papers, images, or videos into a queryable knowledge graph. App code + database schema + infrastructure in one graph.|Python · ⭐90186 · Search 增补 · 近快照 +1245星 · 创建 2026-04-03

GitHub AI趋势 · GitHub AI趋势

DietrichGebert/ponytail

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.|JavaScript · ⭐85196 · Search 增补 · 近快照 +551星 · 创建 2026-06-12

GitHub AI趋势 · GitHub AI趋势

Leonxlnx/taste-skill

Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop|JavaScript · ⭐64593 · Search 增补 · 近快照 +380星 · 创建 2026-02-19

GitHub AI趋势 · GitHub AI趋势

multica-ai/andrej-karpathy-skills

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.|Unknown · ⭐193590 · Search 增补 · 近快照 +355星 · 创建 2026-01-27

GitHub AI趋势 · GitHub AI趋势

mattpocock/skills

Skills for Real Engineers. Straight from my .agents directory.|Shell · ⭐175608 · Search 增补 · 约 1064.3 星/天 · 创建 2026-02-03

🔬 科技前沿

重点 1 · 共 8 条
AI HOT 精选 · 科技前沿

苹果与 OpenAI 法律战升级:约 40 名前员工收到苹果律师函

苹果已向约40名就职于OpenAI的前员工发出律师函,要求保存相关文件。此前苹果起诉OpenAI及两名前员工,指控其通过挖角获取商业机密以加速AI硬件研发。苹果称已有超400名前员工在OpenAI工作,正寻求法院禁令阻止OpenAI使用苹果信息并要求归还机密。(AIHOT源:IT之家(RSS))

AI HOT 精选 · 科技前沿

Apple 起诉 OpenAI:诉讼背后是竞争焦虑还是时机博弈?

Apple 对 OpenAI 提起诉讼,指控其存在多项不当行为,尽管许多专家认为部分指控属于行业惯例。此举正值 Apple 发布新版软件公测版(以新 Siri AI 为核心)之际,外界猜测 Apple 究竟是担忧 OpenAI 成为潜在竞争对手,还是想利用 OpenAI 的弱势期获利。(AIHOT源:The Verge:AI(RSS))

IT之家 · 科技前沿

JDG、AL、BLG 均止步八强:《英雄联盟》2026 EWC 淘汰赛 DK 2-1 战胜 BLG 晋级半决赛,T1 2-0 战胜 HLE

IT之家 7 月 17 日消息,在今日举行的 2026 EWC 电竞世俱杯淘汰赛第一轮比赛中,来自 LCK 的 DK 战队以 2-1 的成绩战胜了 MSI 亚军队伍 BLG,晋级半决赛。 首局比赛,DK 位于蓝色方,BLG 位于红色方。开局第 3 分钟,DK 双人路持续消耗 BLG 下路吉格斯,Lucid 使用盲僧赶到下路击杀吉格斯拿下一血,但随即也被换掉,双方完成一换一。第 25 分钟,DK 拿下小龙后,Career 使用诺提勒斯开启团战,但 BLG 反打击败 DK 上野两人,随后 BLG 主动开打大龙,DK 试图阻止未果,BLG 拿下大龙并打出一换一。第 38 分钟,双方围绕...

少数派 · 科技前沿

本周看什么 | 最近值得一看的 10 部作品

📅本周新预告《蓦然回首》正式预告7月15日,电影《蓦然回首》发布了正式预告,将于9月11日在日本上映。是枝裕和执导、编剧、剪辑,出口夏希、莳田彩珠主演,菅田将晖、宫藤官九郎、野吕佳代、山田真步等出演, ... 查看全文

💻 开发者热点

重点 1 · 共 8 条
AI HOT 精选 · 开发者热点

八天四款前沿模型发布,Kimi K3 跻身第三

过去八天内,Grok 4.5、GPT-5.6、Muse Spark 1.1 与 Kimi K3 四款前沿模型相继发布,使 Artificial Analysis Intelligence Index 得分超 50 的实验室从 6 月初的 2 家增至 6 家。(AIHOT源:X:Artificial Analysis (@ArtificialAnlys))

AI HOT 精选 · 开发者热点

Cursor 评估负责人确认 Claude Fable 5 在 CursorBench 达 72.9% 新高

Cursor 的模型评估负责人 Nate Schmidt 发现,Claude Fable 5 在其内部基准 CursorBench 上以 Max effort 模式达到 72.9%,创下新高。该模型在模糊的真实编程任务中表现出全局推理能力,例如在航天模拟器中仅凭一句提示自主规划并成功登月,而此前 Claude Opus 运行 12 小时以上仍无结果。(AIHOT源:Claude:Blog(网页))

AI HOT 精选 · 开发者热点

首届"小有可为"大赛乡村教育一等奖作品"智绘科普"技术拆解

首届"小有可为"大赛乡村教育赛道一等奖作品"智绘科普"采用 Qwen3.5-397B-A17B 大语言模型与 Manim 动画引擎,通过多Agent分阶段协作与自动修复机制,将知识主题转化为可控、可编辑的教学动画。系统包含规划、草稿、实现、审查、合成五个阶段,渲染失败时可自动提取日志并修复,该工程范式可迁移至其他赛道。(AIHOT源:公众号:通义实验室(千问))

Hacker News · 开发者热点

AWS: Inaccurate Estimated Billing Data – $1.7 billion

URL already posted: https://health.aws.amazon.com/health/status I've got an estimated bill for $1.7 BILLION over this month. Normal usage is < $5. Obvs have created an urge

💹 财经投资

重点 1 · 共 6 条
新浪财经 · 财经投资

申通快递发布快递智能体平台“SClaw”

新浪科技讯 7月17日晚间消息,申通快递在2026年客户开放日上发布快递智能体平台“SClaw”,并首次明确物理AI战略应用方向。面向美妆、服装...

IT之家 · 财经投资

微软 2026 款 Surface Laptop 证明:8GB 内存难以满足 Win11 需求,日常使用多次卡顿

IT之家 7 月 17 日消息,据 Neowin 及 The Verge 今日报道,微软自家 2026 款 13 英寸 Surface Laptop 无意中证明了 Windows 11 的最低系统要求毫无意义。 这款起售价 949 美元(IT之家注:现汇率约合 6435 元人民币)的入门级产品,搭载高通骁龙 X Plus 八核处理器与 256GB 固态硬盘,而 8GB LPDDR5X 内存才是整机性能最明显的瓶颈所在。 The Verge 指出,该机型的外部设计与 2025 款保持一致,键盘手感、触控板、1080p 摄像头及续航表现(轻松达到 10 小时)均得到延续。问题在于内...

🌍 国际时事

重点 1 · 共 3 条
IT之家 · 国际时事

中国科学家首获门捷列夫国际基础科学奖,潘建伟量子通信成果获国际认可

IT之家 7 月 17 日消息,今日,第三届联合国教科文组织 — 俄罗斯门捷列夫国际基础科学奖颁奖典礼在联合国教科文组织巴黎总部举行。 中国科学院院士、中国科学技术大学教授潘建伟成为首位获此殊荣的中国学者。另一位获奖者为美国北卡罗来纳大学查珀尔希尔校区化学系教授谢尔盖 · 舍伊科(Sergei Sheiko),他因在基础聚合物物理学和材料科学领域的杰出贡献而获奖。 联合国教科文组织在颁奖辞中指出,潘建伟是量子光学、量子通信和量子计算领域的国际领军科学家,因在大尺度安全量子通信和可扩展量子计算方面的开创性贡献而受到表彰。他的团队研制成功“墨子号”量子科学实验卫星,实现了数千公里范...

🐙 开源动态

重点 1 · 共 2 条

🧾 运行审计

Source Health
源状态 / Source health
来源条目状态错误
AI HOT 精选12ok
Follow Builders12ok
微博热搜8ok
Dev.to5ok
GitHub AI趋势5ok
Hacker News5ok
Lobsters5ok
新浪财经5ok
36氪3ok
IT之家3ok
InfoQ3ok
MIT Tech Review3ok
TechCrunch3ok
少数派3ok
掘金3ok
量子位3ok
The Verge2ok
按来源展开全部条目
36氪 · 3 条
36氪

氪星晚报|阿里1688将推出AI时代B2B交易互联互通开放标准;英特尔与Google Cloud宣布深化战略合作;铁路部门试点提前60天预约购票服务

大公司: 成大生物:流感病毒裂解疫苗(高剂量)进入I期临床试验 36氪获悉,成大生物公告,公司全资子公司成大生物(本溪)研发的流感病毒裂解疫苗(高剂量)已获国家药监局临床试验批准,并完成I期临床试验筹备工作,正式进入I期临床试验阶段。该疫苗抗原含量为常规剂量四倍,适用于60岁以上老年人群及高风险人群,国内尚无同类产品获批上市,有望填补市场空白。 苏宁易购携手人工智能合作伙伴,布局家用服务机器人 36氪获悉,近日,苏宁易购旗下碧英科技入围工信部人工智能揭榜挂帅项目,将联合产业生态合作伙伴研发智能家庭陪护机器人。该项目面向老龄化背景下的居家照...

36氪

对话森博科技董事长于林义:AI应用拼的不只是技术,更是实证有效的业务闭环

7月17日,2026世界人工智能大会在上海开场。作为36氪连续第三年深入WAIC现场的重要内容窗口,「氪话未来」直播间也在大会首日同步开启现场对话。 森博科技董事长于林义 在WAIC现场接受36氪「氪话未来」特邀专访,围绕企业级AI、智能体落地、行业know-how与业务闭环等话题,分享了森博从营销服务公司转向AI驱动科技服务公司的实践路径。 本届WAIC以“智能伙伴,共创未来”为主题。相比过去几年行业对模型能力、参数规模和Demo效果的集中关注,2026年的AI产业讨论正在明显后移:企业更关心AI能否进入真实流程,能否承担具体任务,能否被业务结果验证。换句话说,AI不再只是...

36氪

2026最受投资人关注人工智能/具身智能企业50揭晓

人工智能正在进入一个新的产业周期。 过去一年,大模型能力持续演进,生成式AI、多模态交互、智能体等技术方向快速推进;而具身智能也从早期的技术探索阶段,逐渐步入产业验证的深水区,机器人开始成为人工智能与现实世界的重要载体。 市场率先给出了回应。据36氪研究院测算,中国具身智能市场规模已从2018年的2133亿元增长至2025年的9150亿元,2026年有望突破万亿关口。IT桔子数据显示,2026上半年具身智能融资总金额达到935亿元,较2025上半年提升近5倍,投融资事件达到322起,同比增长137%。 但硬币的另一面同样刺眼。智源研究院在《2026十大AI技术趋势...

AI HOT 精选 · 12 条
AI HOT 精选

Apple 起诉 OpenAI:诉讼背后是竞争焦虑还是时机博弈?

Apple 对 OpenAI 提起诉讼,指控其存在多项不当行为,尽管许多专家认为部分指控属于行业惯例。此举正值 Apple 发布新版软件公测版(以新 Siri AI 为核心)之际,外界猜测 Apple 究竟是担忧 OpenAI 成为潜在竞争对手,还是想利用 OpenAI 的弱势期获利。(AIHOT源:The Verge:AI(RSS))

AI HOT 精选

八天四款前沿模型发布,Kimi K3 跻身第三

过去八天内,Grok 4.5、GPT-5.6、Muse Spark 1.1 与 Kimi K3 四款前沿模型相继发布,使 Artificial Analysis Intelligence Index 得分超 50 的实验室从 6 月初的 2 家增至 6 家。(AIHOT源:X:Artificial Analysis (@ArtificialAnlys))

AI HOT 精选

Sora 2 视频克隆效果惊人,真假难辨

一年后,没有任何东西能接近 Sora 的完美视频深度克隆。它捕捉到了我和 Sam 的每一块面部肌肉运动以及我们走路的方式。 如果你从这段关于我或 Sam 的视频中截取一帧,根本无法判断它是真是假。(AIHOT源:X:Gabriel (@gabriel1))

AI HOT 精选

Cursor 评估负责人确认 Claude Fable 5 在 CursorBench 达 72.9% 新高

Cursor 的模型评估负责人 Nate Schmidt 发现,Claude Fable 5 在其内部基准 CursorBench 上以 Max effort 模式达到 72.9%,创下新高。该模型在模糊的真实编程任务中表现出全局推理能力,例如在航天模拟器中仅凭一句提示自主规划并成功登月,而此前 Claude Opus 运行 12 小时以上仍无结果。(AIHOT源:Claude:Blog(网页))

AI HOT 精选

美团LongCat发布LoHoSearch:更难搜索智能体基准

美团LongCat推出LoHoSearch,一个基于762万实体维基百科知识图谱自动生成问题的搜索智能体基准,旨在解决BrowseComp等现有基准趋于饱和的问题。在11个前沿模型测试中,最佳得分仅34.74%,远低于当前模型在BrowseComp上约90%的成绩;上下文策略仅带来+6.8个百分点的提升。该基准包含544道问题、11个领域,采用树与图结构,已开源。(AIHOT源:X:美团 LongCat (@Meituan_LongCat))

AI HOT 精选

苹果与 OpenAI 法律战升级:约 40 名前员工收到苹果律师函

苹果已向约40名就职于OpenAI的前员工发出律师函,要求保存相关文件。此前苹果起诉OpenAI及两名前员工,指控其通过挖角获取商业机密以加速AI硬件研发。苹果称已有超400名前员工在OpenAI工作,正寻求法院禁令阻止OpenAI使用苹果信息并要求归还机密。(AIHOT源:IT之家(RSS))

AI HOT 精选

首届"小有可为"大赛乡村教育一等奖作品"智绘科普"技术拆解

首届"小有可为"大赛乡村教育赛道一等奖作品"智绘科普"采用 Qwen3.5-397B-A17B 大语言模型与 Manim 动画引擎,通过多Agent分阶段协作与自动修复机制,将知识主题转化为可控、可编辑的教学动画。系统包含规划、草稿、实现、审查、合成五个阶段,渲染失败时可自动提取日志并修复,该工程范式可迁移至其他赛道。(AIHOT源:公众号:通义实验室(千问))

AI HOT 精选

Schema Harness 在 ARC-AGI-3 公开集上取得约 99% 成绩

Schema 框架在 ARC-AGI-3 公开集上,使用 Claude Opus 4.8 和 Fable 5 达到 99% RHAE 分数,使用 GPT-5.6 Sol 达到 95.35%。该框架不修改模型权重,而是将原始观测转化为可编辑程序,联合解决状态归因和机制发现问题。此前最强模型 GPT-5.6 Sol 在半私有集上仅得 7.78%。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))

AI HOT 精选

NVIDIA 发布 Nemotron 3 Embed 系列,8B 版本在 RTEB 基准上排名第一

NVIDIA 发布 Nemotron 3 Embed 系列,包含三个开源 checkpoint,其中 8B-BF16 版本在 RTEB 基准上以 78.46 的平均 NDCG@10 排名第一。1B-NVFP4 版本在 Blackwell 上吞吐量比 BF16 高 2 倍,精度保留 99.5%,所有模型最大序列长度 32,768 tokens。(AIHOT源:MarkTechPost(RSS))

AI HOT 精选

通义实验室发布 Wan-Streamer v0.2,端到端响应延迟仅 550ms

通义实验室发布 Wan-Streamer v0.2,这是一款将"听、看、说、演"统一进单个 Transformer 的端到端全模态模型。其端到端响应延迟仅 550ms,输出分辨率从 v0.1 的 192×336 提升至 640×368 @ 25FPS,并采用 Thinker-Performer 双通路架构在提升画质的同时维持了极低延迟。(AIHOT源:公众号:通义实验室(千问))

Dev.to · 5 条
Follow Builders · 12 条
Follow Builders

Anthropic Engineering: How we contain Claude across products

Twelve months ago, we'd have rejected out of hand the idea of granting Claude access sufficient to take down an internal Anthropic service. Today that level of access is routine, and Anthropic developers are more productive for it. The risk of these deployments has two components: how likely a failure is, and how much damage one could do. Progress on safeguards and model training has steadily driven down the first; the second—the theoretical blast radius—only grows as capabilities and access expand. Yet as agents become capable of doing work that once required a person or even a team, the cost of not deploying grows large enough that the risk-reward calculation tips heavily toward adoption, as long as products can be made safe. The engineering question becomes how to cap the blast radius. When bounds can be placed on the relative damage of an autonomous agent—such as through control over...

Follow Builders

Sam Altman: we did not have our best last 12 months ever, which is mostly my fault, but w...

we did not have our best last 12 months ever, which is mostly my fault, but we are about to have our best 12 months to date. the team is doing amazing work and i think you’ll be very happy with what they’ve got cooking for you. i am happy about this for many reasons, but mostly because i care about our users winning. AI has to be about giving lots of people more freedom, agency, and wealth. we want to do the right thing, but we do not want to scare people into doing our thing.|@sama · 热度 23005 · 2026-07-16

Follow Builders

Thibault Sottiaux: Evening! We’ve gotten lots of great feedback on the new ChatGPT desktop app (...

Evening! We’ve gotten lots of great feedback on the new ChatGPT desktop app (which we didn't get totally quite right on the first try), and as a result, we've made some changes. 1/ ChatGPT conversation history and projects are now visible in the sidebar. Also, your Chat and Work history now sync across web, mobile, and desktop. Local tasks still stay on your computer. 2/ You can now easily switch between Chat and Work modes inside ChatGPT on desktop, which is now also consistent with how it shows on web and mobile. 3/ Nothing is changing for users on Codex mode. It's still the OG and best at what it does. And overall we're continuing to fix paper cuts and improve performance, reliability, and efficiency. Keep up the feedback, hope you like the updates!|@thsottiaux · 热度 6292 · 2026-07-17

Follow Builders

Guillermo Rauch: Kimi K3 is the best performing model on https://t.co/aporqgIfIh, ahead of Fab...

Kimi K3 is the best performing model on https://t.co/aporqgIfIh, ahead of Fable, reaching a comparable success rate in less time. This is the first time that an open model is ahead of all proprietary ones for this comprehensive web engineering benchmark. Notes: ▪️ Benchmarks don’t always tell the full story, although this is important signal, adding to mounting evidence that this could be a breakthrough moment for open models ▪️ No model as of yet has reached 100% completion on this set of evals. The top performer peaks at 92% and 96% “with help”|@rauchg · 热度 3318 · 2026-07-16

Follow Builders

Guillermo Rauch: I’m excited to welcome two legends of developer tools, Pete Hunt (@floydophon...

I’m excited to welcome two legends of developer tools, Pete Hunt (@floydophone) and Nick Schrock (@schrockn), to Vercel. Pete was one of the pioneers of @reactjs at Meta. He made an early bet to power Instagram Web with ⚛️ React, evangelizing it internally and externally. He will be running Frameworks and leading @nextjs. I couldn’t imagine a better person to lead React’s most popular framework to even greater heights. Nick co-invented @graphql, solving some of the gnarliest data infrastructure and access issues at Facebook scale, with a delightful developer experience. He will be working on Agentic Developer Experience, solving the problem of enabling the next billion agents and leading the way to a future of self-improving software. It’s a dream-come-true for a founder of a startup to welcome engineering minds of this caliber who are also wonderful humans. You probably want to work with them, and they’re hiring 😁. Their DMs are open, from job applications to bug reports!|@rauchg · 热度 882 · 2026-07-16

Follow Builders

Dan Shipper: not surprising that @OpenAI is firing on all cylinders right now, and it’s an...

not surprising that @OpenAI is firing on all cylinders right now, and it’s an unbelievably interesting story - they launched GPT-5 summer of 2025 and positioned it as a pair programmer. we wrote at the time @every that they completely missed the new agentic coding that was starting to happen inside of Claude Code. they bet on agentic coding in the browser / vms and vibe coding in ChatGPT but it was too early - a small team broke off and began working on a separate Codex model line and product. didn’t have to serve the gigantic customer base of ChatGPT, and by November / December of 2025 with 5.3 it was clear they were starting to make something good and the progress was fast - Codex Desktop app launched in Feb and was just clearly superior. there’s a weird late comer advantage sometimes in AI because you get to skip to what works instead of your product having the scars of new capability improvements being bolted on every 3 months - Codex started getting popular it was clear it was a thing, and needed to get merged back in. Which they did, imo quite well—a super complicated thing that would’ve been very easy to screw up. most companies try to disrupt themselves and fail. OpenAI somehow figured out how to disrupt their main product, and then merge it back in seamlessly. incredible aura|@danshipper · 热度 761 · 2026-07-16

Follow Builders

Aaron Levie: It’s truly wild that we’re getting this level of performance from open models

It’s truly wild that we’re getting this level of performance from open models. Congrats to Kimi team on this. Every time we lower the cost of frontier intelligence, the use-cases that enterprises can take on just go up. There’s a tremendous amount of workflows that enterprises would love to deploy that are only gated by the cost of tokens. Importantly for the startup ecosystem, the combined breakthroughs from open and closed labs enable a ton of value to accrue to the layer, which can leverage a variety of models to complete full tasks for customers. This diversity of models and approaches means that the applied AI layer can tune models to their workflows and route intelligence appropriately. Huge win for all.|@levie · 热度 658 · 2026-07-16

Follow Builders

Boris Cherny: In practice that means giving Claude ways to verify its own work end to end

In practice that means giving Claude ways to verify its own work end to end. It means enabling auto mode for permissions, defaulting on automated code review and security review, and using interfaces that let you manage multiple agents at once (Agent view in CLI, Desktop app, iOS and Android apps, Tag). To get to higher levels it means /loop, /batch, dynamic workflows, and worktree isolation for subagents. It's not about a single feature, but rather using the right features with the right guardrails that enable Claude to automate entire classes of work in a way that your team can trust the output.|@bcherny · 热度 180 · 2026-07-17

Follow Builders

Aaron Levie: Here’s another awesome use-case for what we can now do with our unstructured ...

Here’s another awesome use-case for what we can now do with our unstructured data because of AI agents. Box now works with Databricks so you can take structured data from enterprise content (like contracts, financial document, supply chain data) and connect that data into Databricks. This means that I can now query large document datasets without moving or reprocessing that content. And you can connect the data with any other system, like your ERP data, CRM, or product analytics. This opens up a ton of new use-cases for enterprise content. All possible because of headless software and agents.|@levie · 热度 179 · 2026-07-16

GitHub AI趋势 · 5 条
GitHub AI趋势

Graphify-Labs/graphify

AI coding assistant skill (Claude Code, Codex, OpenCode, Cursor, Gemini CLI, and more). Turn any folder of code, SQL schemas, R scripts, shell scripts, docs, papers, images, or videos into a queryable knowledge graph. App code + database schema + infrastructure in one graph.|Python · ⭐90186 · Search 增补 · 近快照 +1245星 · 创建 2026-04-03

GitHub AI趋势

mattpocock/skills

Skills for Real Engineers. Straight from my .agents directory.|Shell · ⭐175608 · Search 增补 · 约 1064.3 星/天 · 创建 2026-02-03

GitHub AI趋势

DietrichGebert/ponytail

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.|JavaScript · ⭐85196 · Search 增补 · 近快照 +551星 · 创建 2026-06-12

GitHub AI趋势

Leonxlnx/taste-skill

Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop|JavaScript · ⭐64593 · Search 增补 · 近快照 +380星 · 创建 2026-02-19

GitHub AI趋势

multica-ai/andrej-karpathy-skills

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.|Unknown · ⭐193590 · Search 增补 · 近快照 +355星 · 创建 2026-01-27

Hacker News · 5 条
IT之家 · 3 条
IT之家

JDG、AL、BLG 均止步八强:《英雄联盟》2026 EWC 淘汰赛 DK 2-1 战胜 BLG 晋级半决赛,T1 2-0 战胜 HLE

IT之家 7 月 17 日消息,在今日举行的 2026 EWC 电竞世俱杯淘汰赛第一轮比赛中,来自 LCK 的 DK 战队以 2-1 的成绩战胜了 MSI 亚军队伍 BLG,晋级半决赛。 首局比赛,DK 位于蓝色方,BLG 位于红色方。开局第 3 分钟,DK 双人路持续消耗 BLG 下路吉格斯,Lucid 使用盲僧赶到下路击杀吉格斯拿下一血,但随即也被换掉,双方完成一换一。第 25 分钟,DK 拿下小龙后,Career 使用诺提勒斯开启团战,但 BLG 反打击败 DK 上野两人,随后 BLG 主动开打大龙,DK 试图阻止未果,BLG 拿下大龙并打出一换一。第 38 分钟,双方围绕...

IT之家

微软 2026 款 Surface Laptop 证明:8GB 内存难以满足 Win11 需求,日常使用多次卡顿

IT之家 7 月 17 日消息,据 Neowin 及 The Verge 今日报道,微软自家 2026 款 13 英寸 Surface Laptop 无意中证明了 Windows 11 的最低系统要求毫无意义。 这款起售价 949 美元(IT之家注:现汇率约合 6435 元人民币)的入门级产品,搭载高通骁龙 X Plus 八核处理器与 256GB 固态硬盘,而 8GB LPDDR5X 内存才是整机性能最明显的瓶颈所在。 The Verge 指出,该机型的外部设计与 2025 款保持一致,键盘手感、触控板、1080p 摄像头及续航表现(轻松达到 10 小时)均得到延续。问题在于内...

IT之家

中国科学家首获门捷列夫国际基础科学奖,潘建伟量子通信成果获国际认可

IT之家 7 月 17 日消息,今日,第三届联合国教科文组织 — 俄罗斯门捷列夫国际基础科学奖颁奖典礼在联合国教科文组织巴黎总部举行。 中国科学院院士、中国科学技术大学教授潘建伟成为首位获此殊荣的中国学者。另一位获奖者为美国北卡罗来纳大学查珀尔希尔校区化学系教授谢尔盖 · 舍伊科(Sergei Sheiko),他因在基础聚合物物理学和材料科学领域的杰出贡献而获奖。 联合国教科文组织在颁奖辞中指出,潘建伟是量子光学、量子通信和量子计算领域的国际领军科学家,因在大尺度安全量子通信和可扩展量子计算方面的开创性贡献而受到表彰。他的团队研制成功“墨子号”量子科学实验卫星,实现了数千公里范...

InfoQ · 3 条
Lobsters · 5 条
MIT Tech Review · 3 条
MIT Tech Review

The Download: perimenopause misinformation and China’s latest AI leap

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. There’s a lot of hype around perimenopause. Don’t buy it. Perimenopause used to be considered taboo, but not anymore. Thanks at least in part to TV doctors and...

MIT Tech Review

There’s a lot of hype around perimenopause. Don’t buy it.

Perimenopause has entered the chat. Perimenopause—and its better-known relative, menopause—used to be considered taboo. Not anymore, thanks at least in part to TV doctors and social media influencers. Perhaps it’s my age, but these days, both my algorithm and my conversations with friends increas...

MIT Tech Review

The risk of weather data sabotage is rising

Every morning, airline dispatchers, grid operators, and farmers around the world make decisions based on the same thing: a weather forecast. While these forecasts are something that most people glance at for two seconds, weather predictions influence major strategic decisions in many industries, ...

TechCrunch · 3 条
The Verge · 2 条
The Verge

TikTok is testing an AI likeness detection tool

TikTok is starting to test an opt-in tool that scans for AI likenesses and lets creators report them to the company, as spotted by social media consultant Matt Navarra. The tool is initially being tested with "some" US creators, TikTok US spokesperson Zachary Kizer tells The Verge. YouTube has be...

少数派 · 3 条
少数派

本周看什么 | 最近值得一看的 10 部作品

📅本周新预告《蓦然回首》正式预告7月15日,电影《蓦然回首》发布了正式预告,将于9月11日在日本上映。是枝裕和执导、编剧、剪辑,出口夏希、莳田彩珠主演,菅田将晖、宫藤官九郎、野吕佳代、山田真步等出演, ... 查看全文

微博热搜 · 8 条
掘金 · 3 条
掘金

正则化在机器学习中的作用

‌正则化是机器学习中防止模型过拟合的技术‌,通过限制模型复杂度来提高泛化能力 。‌‌ 核心作用‌:平衡训练误差与泛化误差,避免模型死记硬背训练数据 。 ‌常见方法‌:包括 L1 正则化(产生稀疏解)、

新浪财经 · 5 条
新浪财经

申通快递发布快递智能体平台“SClaw”

新浪科技讯 7月17日晚间消息,申通快递在2026年客户开放日上发布快递智能体平台“SClaw”,并首次明确物理AI战略应用方向。面向美妆、服装...

量子位 · 3 条