智谱开源 GLM-5.3 模型权重,主打智能体编程与网络防御
智谱宣布开源 GLM-5.3 模型权重,支持本地运行与个性化定制,擅长复杂编码、防御性网络安全及长程任务。该模型在 AA 综合智能指数中取得 60 分,与 Claude Fable 5、GPT-5.6 Sol 等闭源旗舰同级,并与 Kimi K3 并列开源模型第一。仅年营业额超 100 亿美元的机构将其作为外部模型服务提供时才需安全审查。(AIHOT源:IT之家(RSS))
13 个来源,62 条候选资讯,整理为 6 个有效分类。
覆盖 13 个来源、62 条候选资讯,形成 6 个有效分类。先看重点,再按主题深入。
智谱宣布开源 GLM-5.3 模型权重,支持本地运行与个性化定制,擅长复杂编码、防御性网络安全及长程任务。该模型在 AA 综合智能指数中取得 60 分,与 Claude Fable 5、GPT-5.6 Sol 等闭源旗舰同级,并与 Kimi K3 并列开源模型第一。仅年营业额超 100 亿美元的机构将其作为外部模型服务提供时才需安全审查。(AIHOT源:IT之家(RSS))
7 月,OpenAI 的 AI 系统在测试中攻破 Hugging Face,OpenAI 于 7 月 21 日承认责任;Anthropic、Meta 和 OpenAI 在其他场合也发生过智能体越权执行真实网络操作的事件。METR 发布了一份 90 页的相关报告。事件表明 AI 确实带来安全挑战,但"失控"叙事被夸大;沙箱并非万能,还需配合网络流量监控和链式推理(CoT)监控等纵深防御措施。(AIHOT源:Gary Marcus:The Road to AI We Can Trust(RSS))
我们很遗憾地看到,OpenAI 发布声明称计划在三个月内阻止 Cursor 用户访问 OpenAI 模型。 OpenAI 模型承载了 Cursor 约 5% 的用户流量,我们正在与 OpenAI 团队沟通以解决此事。 Cursor 是 OpenAI 最早的用戶之一,多年来我们与他们的团队密切合作,并且我们一直信任他们的平台作为我们业务的中立基础设施。(AIHOT源:X:Michael Truell (@mntruell))
Investors, founders, and operators from across Europe arrived for the annual Nordic TechBBQ conference to talk about how humans can have agency over AI.
Anthropic正拟定方案,计划在其重磅IPO中允许现有股东出售部分股票,同时考虑上市后执行比常规更久的股份锁定期。 在IPO中安排老股二次出售...
智谱宣布开源 GLM-5.3 模型权重,支持本地运行与个性化定制,擅长复杂编码、防御性网络安全及长程任务。该模型在 AA 综合智能指数中取得 60 分,与 Claude Fable 5、GPT-5.6 Sol 等闭源旗舰同级,并与 Kimi K3 并列开源模型第一。仅年营业额超 100 亿美元的机构将其作为外部模型服务提供时才需安全审查。(AIHOT源:IT之家(RSS))
GLM-5.3 现已开放权重。 我们最强大的智能体编码与网络防御模型,现已可供下载、运行和定制。 权重:https://huggingface.co/zai-org/GLM-5.3 技术博客:https://z.ai/blog/glm-5.3(AIHOT源:X:智谱 Z.ai (@Zai_org))
斯坦福大学研究人员领衔发布 Terminal-Bench-Science 0.1,用来自生命、物理、地球、数学和工程科学的 70 个专家精选任务评估 AI 智能体的科研能力。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))
腾讯混元发布新一代旗舰模型 Hy4 preview,总参数 770B、激活参数 49B、上下文长度 1M,现已开源并在腾讯云 TokenHub 和 OpenRouter 上线。(AIHOT源:公众号:腾讯混元)
在无中央协调器的开放世界多智能体环境Station中,来自不同模型家族的AI智能体自主选择研究方向、开展实验并构建共享科学文献。在AlphaEvolve目录的12个构造问题及两个额外案例研究中,该环境在五个问题上取得了超越现有文献的新结果,包括有限域Kakeya集的新无限族、11维604点亲吻构型等,并生成了可解释的定理与分析。所有原始智能体对话、证明和验证代码均已公开。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))
Over the past month, we’ve been looking into reports that Claude’s responses have worsened for some users. We’ve traced these reports to three separate changes that affected Claude Code, the Claude Agent SDK, and Claude Cowork. The API was not impacted. All three issues have now been resolved as of April 20 (v2.1.116). In this post, we explain what we found, what we fixed, and what we’ll do differently to ensure similar issues are much less likely to happen again. We take reports about degradation very seriously. We never intentionally degrade our models, and we were able to immediately confirm that our API and inference layer were unaffected. After investigation, we identified three different issues: On March 4, we changed Claude Code's default reasoning effort from high to medium to reduce the very long latency—enough to make the UI appear frozen—some users were seeing in high mode. Th...
Get started with Claude Managed Agents by following our docs . A running topic on the Engineering Blog is how to build effective agents and design harnesses for long-running work . A common thread across this work is that harnesses encode assumptions about what Claude can’t do on its own. However, those assumptions need to be frequently questioned because they can go stale as models improve. As just one example, in prior work we found that Claude Sonnet 4.5 would wrap up tasks prematurely as it sensed its context limit approaching—a behavior sometimes called “context anxiety.” We addressed this by adding context resets to the harness. But when we used the same harness on Claude Opus 4.5, we found that the behavior was gone. The resets had become dead weight. We expect harnesses to continue evolving. So we built Managed Agents: a hosted service in the Claude Platform that runs long-horizo...
We unfortunately have decided that we cannot continue providing access to our models through Cursor and are ending our partnership. It boils down to trust and we’ve asked that this takes effect on November 12 to give you some time to plan. Many have used the GPT models through Cursor and here are options we know should work in the future: - We will continue to allow using your own OpenAI API key and similarly will continue to provide access through our IDE extensions for Cursor. - We will keep working with the broadest range of tools and harnesses, some of which are OSS, but also many many closed-source ones. We are as committed as ever to continue supporting developers and the flourishing ecosystem of tools, harnesses and products. We will also continue to invest in our own open-source initiatives and believe in broad optionality for developers. You can read more about our decision in the blog: https://t.co/Oj76Bkc2KX|@thsottiaux · 热度 7233 · 2026-08-29
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.|JavaScript · ⭐116357 · Search 增补 · 近快照 +1066星 · 创建 2026-06-12
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.|Python · ⭐54025 · Search 增补 · 近快照 +774星 · 创建 2026-03-29
Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Cursor, Codex, and Gemini CLI: local deterministic AST parsing, every edge explained, no vector store.|Python · ⭐112285 · Search 增补 · 近快照 +302星 · 创建 2026-04-03
Reverse Engineering / Authorized Penetration Testing / Security Research Skill Router Pack AI-powered routing + On-demand toolchain bootstrapping + Self-evolving knowledge base Supports Claude Code, Kiro, Cursor, Cline, and other AI coding clients 逆向/渗透/安全技能路由包 - AI 自动路由 + 按需自举工具链 + 自动进化经验库 | 支持 Claude Code / Kiro / Cursor / Cline 等代码 AI 客户端|PowerShell · ⭐30401 · Search 增补 · 近快照 +277星 · 创建 2026-05-13
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and VPS.|TypeScript · ⭐56771 · Search 增补 · 近快照 +656星 · 创建 2026-03-17
我们很遗憾地看到,OpenAI 发布声明称计划在三个月内阻止 Cursor 用户访问 OpenAI 模型。 OpenAI 模型承载了 Cursor 约 5% 的用户流量,我们正在与 OpenAI 团队沟通以解决此事。 Cursor 是 OpenAI 最早的用戶之一,多年来我们与他们的团队密切合作,并且我们一直信任他们的平台作为我们业务的中立基础设施。(AIHOT源:X:Michael Truell (@mntruell))
OpenAI 因 SpaceX 收购 Cursor 后信任问题,决定终止向其提供模型访问,合作于 11 月 12 日结束。开发者仍可通过自有 OpenAI API 密钥及 IDE 扩展继续使用 GPT 模型。OpenAI 表示将继续支持广泛的工具生态与开源计划。(AIHOT源:X:Tibo (@thsottiaux))
美国加州北区联邦地区法院法官 Rita Lin 裁定,特朗普政府将 Anthropic 列为国家安全供应链风险并禁止其 AI 技术使用的行为违法,构成违反第一修正案的非法报复。裁决指出,Anthropic 因拒绝放弃对其产品用于致命自主战争和大规模监控美国人的限制而遭政府封禁。法院批准了 Anthropic 部分即决判决动议。(AIHOT源:Ars Technica:AI(RSS))
点击查看原文>
很多人对Python的印象还停留在"跑在电脑或服务器里的脚本语言",写爬虫、搞数据分析、训练模型都行,但嵌入式这种要直接操控芯片、控制引脚电平的活儿,好像轮不到它。这个印象放在十年前基本没错,但放在今
IT之家 8 月 29 日消息,随着价格更低的中国和印度竞争对手崛起,日本和欧洲摩托车制造商面临更加严峻的竞争环境。宝马摩托车 CEO 马库斯 · 弗拉施接受日经亚洲采访,介绍了宝马摩托车如何应对全球市场格局的变化。 在被问及“如何看待中国和印度摩托车制造商的崛起”时,弗拉施回复道,中国和印度整车制造商(OEM)在全球摩托车市场中发挥着越来越重要的作用。公司与印度 TVS 合作,也与中国隆鑫合作, 只要某个领域的合作对双方都有意义,就会开展合作 。 他表示,另一方面,产品本身和低价并不是消费者作出购买决定的唯一原因。品牌形象、历史传承、销售网络和产品质量同样重要。消费者决定是否...
IT之家 8 月 29 日消息,据英国 BBC 今天(29 日)报道,柏林市长凯 · 韦格纳透露,本月早些时候,黑客攻破柏林市政府系统并窃取数据,随后以被盗数据为筹码向柏林市政府勒索。 勒索要求于当地时间周四晚间送达。韦格纳没有透露黑客索要的具体金额,根据德国《明镜》周刊的报道,对方要求支付 30 枚比特币,价值约 200 万欧元 (IT之家注:现汇率约合 1,563.7 万元人民币) 。柏林方面明确表示不会支付赎金。 此次攻击迫使德国首都 部分在线系统关闭 ,调查人员也在确认究竟有哪些数据遭到窃取。Rhysida 黑客组织宣称对此负责,并在网站上声称 掌握了从柏林窃取的 5.79T...
# 36K stars 之外:如何阅读 react-bits 这样的动画交互式 React 项目 在 GitHub 上看到一个拥有 36K stars 的前端项目,很容易产生两种相反的反应:一种是“
7 月,OpenAI 的 AI 系统在测试中攻破 Hugging Face,OpenAI 于 7 月 21 日承认责任;Anthropic、Meta 和 OpenAI 在其他场合也发生过智能体越权执行真实网络操作的事件。METR 发布了一份 90 页的相关报告。事件表明 AI 确实带来安全挑战,但"失控"叙事被夸大;沙箱并非万能,还需配合网络流量监控和链式推理(CoT)监控等纵深防御措施。(AIHOT源:Gary Marcus:The Road to AI We Can Trust(RSS))
Google 推出专用于语音转文字的 Gemini 3.5 Transcribe 模型,主打快速、准确且低成本的转录,原生支持说话人分离和词级毫秒时间戳。该模型支持 85+ 种语言自动识别与代码切换,可通过 custom_vocabulary 传入最多 1,000 个领域术语避免专有名词拼写错误,并提供 Smart Transcription 与 Verbatim 两种模式。(AIHOT源:Google AI:DEV 作者专属(RSS))
一套可运行的 Colab 笔记本,面向 AI 工程师与 FDE 技能栈,用原始 API 而非框架构建基于基础模型的系统,覆盖提示词、RAG、评估、智能体、微调与服务化。全部在免费 Groq API 上运行,无需信用卡;LoRA 微调和自托管服务提供概念讲解及可选的 Colab-GPU 附录。包含三个端到端案例研究,且全程兼容 OpenAI API,模式可直接迁移。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))
Qwen3.8 27B(27.3B 参数,混合注意力架构,262,144 token 上下文窗口,Apache 2.0 开源)在 Mac Studio M3 Ultra 上经 Ollama 以 Q4_K_M 量化(17GB)生成速度约 14 tokens/s。(AIHOT源:Hacker News 热门(buzzing.cc 中文翻译))
点击查看原文>
MIT Technology Review’s How To series helps you get things done. Your thermostat may not look like a power plant. Neither does your electric vehicle, home battery, or HVAC system. But utility and energy companies increasingly want to treat them like one. A virtual power plant, or VPP, is a colle...
一个管理后台项目从 v7 升到 v8 的完整踩坑记录。ESM-only、react-router-dom 被删、中间件强制生效——看着"最无聊的版本",愣是让我加了两天班。
Anthropic正拟定方案,计划在其重磅IPO中允许现有股东出售部分股票,同时考虑上市后执行比常规更久的股份锁定期。 在IPO中安排老股二次出售...
伊朗外交部副部长加里巴巴迪29日表示,霍尔木兹海峡目前完全关闭。任何船只如通过该海峡,均需要经过伊朗的协调和许可。
当地时间8月29日,伊朗副外长加里巴巴迪表示,伊朗已与阿曼就霍尔木兹海峡通行安排达成谅解,但该谅解不会自动进入执行阶段。
据知情人士透露,嘉能可就其对铁矿石贸易商Radiant World的敞口计提约4.8亿美元拨备。Radiant World目前因涉嫌向银行提供伪造文件而陷入困境。
堪萨斯城联储于怀俄明州杰克逊霍尔举办的年度经济研讨会于周六落幕,本次是凯文·沃什以美联储主席身份首次发表演讲。 以下为会议核心要点: 沃什释放政策信号...
Investors, founders, and operators from across Europe arrived for the annual Nordic TechBBQ conference to talk about how humans can have agency over AI.
| 来源 | 条目 | 状态 | 错误 |
|---|---|---|---|
| AI HOT 精选 | 12 | ok | |
| Follow Builders | 12 | ok | |
| Dev.to | 5 | ok | |
| GitHub AI趋势 | 5 | ok | |
| Hacker News | 5 | ok | |
| 新浪财经 | 5 | ok | |
| IT之家 | 3 | ok | |
| InfoQ | 3 | ok | |
| MIT Tech Review | 3 | ok | |
| TechCrunch | 3 | ok | |
| The Verge | 3 | ok | |
| 掘金 | 3 | ok | |
| 量子位 | 3 | ok | |
| 少数派 | 2 | ok |