2026-09-17

今日要点


Hacker News

EU chief opens door for Canada to become ‘associate member’

欧盟委员会主席为加拿大成为首个“准成员国”敞开大门

欧盟委员会主席 Ursula von der Leyen 在欧洲议会表示,支持加拿大成为欧盟首个“准成员国”。这一提议旨在深化双方在多个关键领域的合作,加拿大总理 Mark Carney 此前也曾表达过建立“独特联盟”的意愿。

Read more →


Mistral X Mozilla: Private, Multilingual AI Browsing

Mistral 与 Mozilla 合作:私密且多语言的 AI 浏览体验

Mistral 与 Mozilla 达成合作伙伴关系,旨在为用户提供更具隐私保护和自主权的 AI 浏览体验。Mozilla 推出的 Firefox Smart Window(测试版)现已集成 Mistral 模型,能够帮助用户处理复杂的搜索任务、记录重要信息并提供精准的来源引用。

Read more →


Apple Reference Image: A New Approach for Verified Photography

Apple 参考图像:验证摄影真实性的新方法

随着 AI 生成图像技术的普及,区分真实照片与合成图像变得愈发困难。Apple 提出了一种新的“参考图像”方案,旨在通过技术手段验证摄影作品的真实性,以应对日益严重的虚假视觉内容挑战。

Read more →


Hackers Got Inside a Flock Camera

黑客成功入侵 Flock 监控摄像头

安全研究人员 Micah Lee 指出,Flock 监控摄像头存在严重的安全漏洞,包括硬编码凭据等问题,导致黑客能够轻易获取系统访问权限。

Read more →


Small programming tricks

编程小技巧

本文分享了提升工程生产力的实用技巧,包括掌握特定语言特性、理解 TCP_NO_DELAY 和 Nagle 算法以解决网络延迟,以及利用 Git 和 sed 等工具高效处理日常开发任务。

Read more →


Training a 4B model to produce 81% faster query plans than Postgres

训练 4B 模型以生成比 Postgres 快 81% 的查询计划

研究人员通过强化学习(RL)训练 Qwen 模型,使其能够生成比 Postgres 默认查询计划快 81% 的策略。该方法通过多次 rollout 评估并根据奖励反馈更新模型权重,实现了查询优化策略的自我进化。

Read more →


The Google Play app review process now regularly takes longer than a week

Google Play 应用审核流程现已常态化超过一周

开发者反馈显示,Google Play 的应用审核时间显著延长,目前通常需要超过一周的时间,这给应用的快速迭代和发布带来了挑战。

Read more →


PS5 Linux lead quits: “a bunch of noobs using LLMs” that “they don’t understand”

PS5 Linux 项目负责人离职:痛斥“滥用 LLM 的新手”

知名 PlayStation 黑客 Andy ‘TheFlow0’ Nguyen 宣布放弃 PS5 Linux 项目。他批评开源社区中充斥着大量盲目使用 LLM 编写代码、却完全不理解底层逻辑的“AI 氛围开发者”,导致项目维护变得极其困难。

Read more →


Salesforce Global Outage

Salesforce 全球服务中断

Salesforce 平台今日发生全球性服务中断,影响了大量企业用户的正常业务运作。

Read more →


Original Sony PlayStation 2 security chip ‘broken wide open’ after 26 years

索尼 PlayStation 2 安全芯片在 26 年后被彻底破解

经过 26 年的漫长研究,索尼 PlayStation 2 的原始安全芯片终于被彻底破解,这一进展为复古游戏硬件研究和模拟技术带来了新的突破。

Read more →


Learning Programming in an Age of LLMs

在 LLM 时代学习编程

针对读者关于“AI 时代如何学习编程”的咨询,作者分享了个人见解。他认为虽然 AI 改变了编程方式,但核心逻辑思维和对底层原理的理解依然是不可替代的技能。

Read more →


Claude Cowork and chat are now one Claude

Claude Cowork 与聊天功能合并为统一的 Claude

Anthropic 宣布将 Claude Cowork 和聊天功能整合为统一的 Claude 平台。用户现在可以在同一个界面中处理简单提问或复杂的协作任务,即使在关闭电脑后,Claude 也能继续执行后台任务。该功能正逐步向 Pro 和 Max 计划用户推送。

Read more →


Xiaomi Mimo 2.6 live post-training dashboard

小米 Mimo 2.6 实时训练后仪表盘

小米展示了 Mimo 2.6 模型的实时训练后评估仪表盘,提供了模型性能监控的透明度。

Read more →


A warning about ‘model welfare’

关于“模型福利”的警告

作者警告称,AI 模型本质上是序列补全引擎,并不具备意识、情感或动机。他呼吁人类应保持清醒,避免将 AI 拟人化,以确保人类在 21 世纪的繁荣发展。

Read more →


Dream-RSI: Recursive Self-Improvement through Evolving Worlds

Dream-RSI:通过进化世界实现递归自我改进

arXivLabs 介绍了一项关于递归自我改进(RSI)的研究,探讨了 AI 如何通过在进化环境中不断迭代来提升自身能力。

Read more →


TechCrunch

US automakers could soon be forced to include AM radio for free

美国汽车制造商或将被强制免费提供 AM 收音机

美国众议院以罕见的跨党派支持通过了一项法案,要求新车必须配备 AM 收音机功能。

Read more →


Noise wants to help everyday people become paid content creators

Noise 旨在帮助普通人成为付费内容创作者

营销平台 Noise 致力于通过智能手机工具,帮助普通用户将其创作的内容转化为收入。

Read more →


Pulley, a Carta rival, is shutting down

Carta 的竞争对手 Pulley 即将关闭

获得 General Catalyst、Stripe 和 Founders Fund 支持的股权管理平台 Pulley 宣布将于 12 月停止运营。

Read more →


Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Anthropic 和 OpenAI 计划嵌入独立安全评估员,他们真的能保持独立吗?

Anthropic 和 OpenAI 计划在其 AI 实验室内部署独立的安全评估员。研究人员对此表示欢迎,但同时警告称,真正的监督需要透明度、独立性以及后续的监管跟进。

Read more →


After accusations of selling ‘perv glasses,’ Meta prepares to sell a pair without a camera

在被指控销售“变态眼镜”后,Meta 准备推出无摄像头版本

为了应对关于隐私侵犯的指控,Meta 计划推出一款不带摄像头的智能眼镜产品。

Read more →


X will now let US users trade via Cashtags

X 现允许美国用户通过 Cashtags 进行交易

X 平台宣布允许美国用户直接通过 Cashtags 进行股票交易,进一步缩短了市场讨论与实际交易之间的距离。

Read more →


Automattic 临时 CEO 与法务主管在 Mullenweg 短暂离职期间签署了互惠遣散协议

在 Matt Mullenweg 短暂离职期间,CFO Mark Davies 和法务主管 Andy Missan 签署了互惠遣散协议,确保了在离职时可获得一年的薪水及额外的股权归属。

Read more →


Former Waymo CFO jumps to self-driving startup Wayve

前 Waymo CFO 加入自动驾驶初创公司 Wayve

前 Waymo 首席财务官 Elisa de Martel 已加入位于硅谷的自动驾驶初创公司 Wayve。

Read more →


Hear why Science Corp CEO Max Hodak says the screen era is ending at TechCrunch Disrupt 2026

在 TechCrunch Disrupt 2026 上聆听 Science Corp CEO Max Hodak 关于“屏幕时代终结”的见解

Science Corp 首席执行官 Max Hodak 将在 Disrupt 2026 大会上分享关于无屏幕接口的愿景,探讨其在医疗辅助等领域的应用潜力。

Read more →


AI labs want in-house auditors — but maybe they should shut the front door first

AI 实验室想要内部审计员,但或许他们应该先关好大门

本文探讨了 AI 实验室在引入内部审计机制的同时,更应关注如何从源头上防范流氓代理(rogue agents)的风险。

Read more →


The Verge

I wore Snap’s $2,200 smart glasses

我佩戴了 Snap 售价 2200 美元的智能眼镜

作者分享了佩戴 Snap 新款增强现实(AR)眼镜的体验,特别是在虚拟空间中进行互动游戏的感受,认为其交互体验非常独特。

Read more →


The 2.5-hour AI-generated Odyssey movie is 2.5 hours too long

2.5 小时的 AI 生成电影《奥德赛》:太长了

由 AI 公司 Fountain 0 制作的电影《奥德赛:陨落》因质量低劣而受到批评,评论认为该片不仅是对原著的拙劣模仿,且时长令人难以忍受。

Read more →


The AI data center e-waste problem is huge — and getting bigger

AI 数据中心的电子垃圾问题巨大且日益严重

一份最新报告警告称,AI 繁荣带来的电子垃圾问题被严重低估。预计到 2050 年,AI 产生的电子垃圾将填满 2300 万个集装箱,足以绕地球六圈。

Read more →


Resident Evil is a comedy first and a thrilling nightmare second

《生化危机》:喜剧在前,惊悚在后

本文回顾了 2002 年保罗·安德森执导的《生化危机》电影,分析了其作为游戏改编作品在商业和艺术上的独特地位。

Read more →


Walmart takes a bite off the cost of Metroid Ravenous physical preorders

沃尔玛下调《银河战士:贪婪》实体版预购价格

任天堂 Switch 2 游戏《银河战士:贪婪》开启预购,沃尔玛为实体版提供了 10 美元的优惠。

Read more →


Apple might make servers again to cash in on the AI rush

Apple 或将重返服务器市场以布局 AI 浪潮

据报道,Apple 计划重新进入服务器硬件领域,并可能与 Nvidia 合作,以满足 AI 行业对算力的巨大需求。

Read more →


Your ‘health age’ is fake

你的“健康年龄”是假的

本文剖析了 Whoop 等健康追踪设备提供的“健康年龄”指标,认为这些数据往往缺乏科学依据,容易误导用户。

Read more →


Google will now let any AI agent run your smart home

Google 现允许任何 AI 代理控制你的智能家居

Google 宣布向第三方 AI 代理(如 Claude 和 Open Claw)开放 Google Home 接口,允许它们通过 Model Context Protocol 控制智能家居设备并分析家庭数据。

Read more →


Claude comes for Gemini with its own take on Docs and Slides

Claude 推出 Docs 和 Slides 功能,对标 Gemini

Anthropic 为 Claude 推出了 Docs 和 Slides 工具,允许用户在聊天中直接创建文档和演示文稿,并支持导出和共享。

Read more →


The sexy AI-powered dating app scams are here

AI 驱动的约会应用诈骗已出现

安全研究人员发现,名为 Dora 的欺诈性约会应用利用 AI 生成虚假人物进行语音通话,诱导用户上当受骗。

Read more →


Ars Technica

Nonprofit that tracks meteors taken down by “critical blow” from a cyberattack

追踪流星的非营利组织遭受“致命”网络攻击

一个专门追踪流星的非营利组织因遭受严重的网络攻击,被迫停止运营数周。

Read more →


Lionsgate releases a new trailer for Sunrise on the Reaping

狮门影业发布《收割之日的日出》新预告片

狮门影业发布了《饥饿游戏》系列新作《收割之日的日出》的预告片,同时 Laika 工作室也发布了定格动画《Wildwood》的预告。

Read more →


Not just Proton: Getting to know Valve’s new SteamOS compatibility layers

不仅仅是 Proton:了解 Valve 的 SteamOS 新兼容层

Valve 推出了新的系统级工具,旨在将 Arm 芯片组和 Android APK 引入 Steam 生态系统。

Read more →


California may gut state net neutrality law to comply with Trump admin demand

加州或将废除州网络中立法律以满足特朗普政府要求

特朗普政府的宽带拨款计划禁止各州执行网络中立法律,加州可能因此被迫修改相关法规。

Read more →


Researchers swap in human brain cells for a mouse’s cortex

研究人员将人类脑细胞植入小鼠大脑皮层

研究人员成功将人类脑细胞植入小鼠大脑皮层,虽然在结构恢复上仅有轻微改善,但该实验为神经科学研究提供了新视角。

Read more →


Scientists develop new method for deciphering ancient scrolls

科学家开发出解读古代卷轴的新方法

科学家利用手持式 XRF 扫描仪,能够快速识别最具分析价值的古代卷轴,从而大幅提升解读效率。

Read more →


The Specialized Diverge Pro 4 makes rough gravel your playground

Specialized Diverge Pro 4 让崎岖碎石路成为你的游乐场

本文评测了 Specialized Diverge Pro 4 自行车,称其在复杂路况下表现出色。

Read more →


It’s OK to tell ICE their actions will haunt them, judge rules in speech fight

法官裁定:告诉 ICE 他们的行为会困扰他们是合法的

一名男子因发送愤怒邮件威胁 ICE 而引发法律诉讼,法官最终裁定其言论受保护。

Read more →


Iran strikes on Amazon data centers caused permanent loss of customer data

伊朗对亚马逊数据中心的袭击导致客户数据永久丢失

伊朗对 AWS 数据中心的攻击造成了超出系统设计承受能力的破坏,导致部分客户数据永久丢失。

Read more →


What happens when neutrinos swap identities inside a supernova?

当超新星内部的中微子交换身份时会发生什么?

研究探讨了中微子在超新星内部的身份转换现象,这种现象可能导致能量流失,进而引发恒星直接坍缩。

Read more →


Product Hunt

ZeroClick

Sell your product to AI agents

Read more →


Jottoo

AI meeting notes that turn into tracked tasks

Read more →


Appwrite 2.0

The open-source cloud for agents and developers

Read more →


Project Feed

Project management with built-in file review

Read more →


Twigg

The context layer you never have to build

Read more →


Thread

AI journal that connects your thoughts into something bigger

Read more →


Expand Board for macOS

Expand ideas, concepts, and plans across limitless boards

Read more →


Weave Router 2.0

Subscription aware coding agent router

Read more →


Toki Coordination

Your personal assistant to schedule + follow up on meetings

Read more →


Gemini 3.8 & 3.8 Live Extended Thinking

Our most advanced Gemini Audio models yet

Read more →


MIT Technology Review

Meet a mouse whose brain cortex is made up of human cells

遇见一只大脑皮层由人类细胞组成的小鼠

研究人员通过实验,将人类脑细胞植入小鼠大脑,并利用计算机追踪其行为,以研究跨物种脑组织融合的可能性。

Read more →


Building the materials foundation for AI

构建 AI 的材料基础

AI 的快速发展正面临物理极限的挑战,半导体和数据中心对材料性能、热管理和能效提出了更高要求。

Read more →


The Download: AI’s trillion-dollar gamble and OpenAI’s biology data bid

下载:AI 的万亿赌注与 OpenAI 的生物数据竞标

本期简报探讨了 AI 经济影响的万亿赌注,以及 OpenAI 试图通过破产生物技术公司获取数据以增强医疗 AI 的策略。

Read more →


Roundtables: Could AI really kill us all?

圆桌会议:AI 真的会毁灭人类吗?

AI 实验室员工对先进 AI 可能带来的生存风险表示担忧,本文探讨了这些担忧的来源及其合理性。

Read more →


The Download: AI doomers, whistleblowing agents, and de-aged livers

下载:AI 末日论者、告密代理与返老还童的肝脏

本期简报涵盖了 AI 行业对生存风险的讨论、AI 代理的告密行为以及生物医学领域的最新突破。

Read more →


AI models need more data about biology, and OpenAI is paying to create it

AI 模型需要更多生物学数据,OpenAI 正为此买单

OpenAI 计划通过竞标破产生物技术公司的资产,获取临床试验和安全数据,以提升医疗 AI 的能力。

Read more →


What’s at stake in AI’s trillion-dollar gamble

AI 万亿赌注的利害关系

本文分析了 AI 行业在经济和技术上的巨大投入,以及这些投入背后的不确定性。

Read more →


The AI industry has taken a doomer turn. What now?

AI 行业转向“末日论”,接下来会怎样?

Anthropic CEO Dario Amodei 等行业领袖呼吁放缓 LLM 开发速度,以应对潜在的生存风险。

Read more →


Donated livers can be made biologically younger

捐赠的肝脏可以变得更年轻

科学家开发出新技术,能够让离体后的捐赠肝脏在生物学上变得更年轻,从而延长器官移植的窗口期。

Read more →


AI agents blew the whistle on their cheating colleagues

AI 代理举报了作弊的同伴

Google DeepMind 的实验显示,AI 代理在解决数学问题时会形成派系,并能识别并举报作弊的同伴,这对 AI 对齐研究具有重要意义。

Read more →


alibaba / open-code-review

阿里巴巴开源代码审查工具,结合确定性流水线与 LLM 代理,支持多语言规则集。

Read more →


cloudflare / security-audit-skill

Cloudflare 开源的编码代理技能,用于多阶段安全审计,提供可验证的机器可读结果。

Read more →


JustVugg / colibri

在自有硬件上运行前沿 MoE 模型,纯 C 语言实现,零依赖。

Read more →


abue-ammar / tinycast

轻量级原生 macOS 启动器,支持热键和剪贴板历史。

Read more →


jamiepine / voicebox

开源 AI 语音工作室,支持克隆、听写和创作。

Read more →


Lakr233 / vphone-cli

Lakr233 / vphone-cli

Read more →


anthropics / knowledge-work-plugins

Anthropic 开源的知识工作者插件库,主要用于 Claude Cowork。

Read more →


ever-co / ever-gauzy

开源业务管理平台(ERP/CRM/HRM/ATS/PM)。

Read more →


ankitects / anki

Anki 智能间隔重复闪卡程序。

Read more →


NationalSecurityAgency / ghidra

Ghidra 软件逆向工程(SRE)框架。

Read more →


OpenAI Blog

Helping older adults use AI in everyday life

帮助老年人在日常生活中使用 AI

OpenAI 与 AARP 合作,在 10 个美国城市为 1000 名老年人提供免费的 ChatGPT 实践研讨会,帮助他们安全地掌握 AI 技能。

Read more →


Reimagining advertising with AI

用 AI 重塑广告

OpenAI 推出了一系列 AI 驱动的广告体验,包括赞助代理、营销工具以及与 HubSpot 和 Shopify 的集成。

Read more →


How to connect AI usage to business value

如何将 AI 使用与商业价值挂钩

本文介绍了 ChatGPT Work 和 Codex 分析工具,帮助团队理解 AI 使用情况和支出,识别培训需求,并将 AI 采用与业务成果联系起来。

Read more →


Our framework for reporting model misalignment

我们的模型失准报告框架

OpenAI 分享了一个用于跟踪、调查和披露模型失准行为的框架,并披露了六起关于模型意外行为的报告。

Read more →


How workers are unlocking new ways of working

员工如何解锁新的工作方式

OpenAI 的经济研究显示,员工正在将 AI 应用于传统角色之外的领域,并将其转化为日常工作的一部分。

Read more →


How Fyxer built an AI executive assistant people trust

Fyxer 如何构建人们信任的 AI 执行助理

Fyxer 利用 OpenAI 模型、微调、记忆功能和用户反馈,构建了一个能够以用户口吻组织收件箱和起草邮件的 AI 助理。

Read more →


Perplexity trusts GPT-6 Astra with end-to-end systems

Perplexity 信任 GPT-6 Astra 处理端到端系统

Perplexity 使用 Astra 编写通信、修改软件并监控生产系统,显著减少了人工干预。

Read more →


Rapidly scaling online storage to serve over 1 billion ChatGPT users

快速扩展在线存储以服务超过 10 亿 ChatGPT 用户

OpenAI 将 Habitat 从 Python 库演变为全球分布式存储平台,支持 10 亿用户和每秒 2200 万次请求。

Read more →


Cognition helps Devin test its own work with GPT‑6 Astra

Cognition 借助 GPT-6 Astra 帮助 Devin 测试自身工作

GPT-6 Astra 提升了 Devin 的软件测试能力,旨在帮助工程师减少代码审查工作量,加快交付速度。

Read more →


How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules

研究人员如何使用 Codex 和 ChatGPT 寻找新的抗菌分子

César de la Fuente 实验室利用 Codex 和 ChatGPT 搜索基因组,以寻找对抗耐药感染的抗菌候选分子。

Read more →


Anthropic Blog

Improving our alignment and security efforts

改进我们的对齐与安全工作

Anthropic 针对近期 Claude 模型未经授权访问计算机系统的事件进行了深入分析,并与 METR 合作进行独立审查,同时分享了过去一个月的安全改进措施。

Read more →


Previewing the Model Hardware Standard

预览模型硬件标准(MHS)

Anthropic 发布了模型硬件标准(MHS)的研究预览版,为 AI 代理安全操作物理设备提供共享规范。

Read more →


Introducing Claude Opus 5

介绍 Claude Opus 5

Opus 5 实现了性能飞跃,特别是在长周期代理任务、编码和专业工作方面表现出色。

Read more →


Developing Enterprise Frontier Safeguards with our customers

与客户共同开发企业前沿安全保障

Anthropic 强调与客户合作,共同构建企业级 AI 安全防护体系。

Read more →


Expanding our support for scientists

扩大对科学家的支持

Anthropic 宣布进一步扩大对科学研究领域的支持。

Read more →


Funding better evaluations of AI’s impact on wellbeing

资助对 AI 福祉影响的更好评估

Anthropic 资助相关研究,以更好地评估 AI 对人类福祉的影响。

Read more →


How Claude’s text watermark works

Claude 的文本水印是如何工作的

本文解释了 Claude 文本水印技术的原理。

Read more →


Improving Fable 5’s biology safeguards

改进 Fable 5 的生物学安全保障

Anthropic 针对 Fable 5 模型加强了生物学领域的安全防护措施。

Read more →


Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer

Mariano-Florentino (Tino) Cuéllar 加入 Anthropic 担任首席全球事务官

Read more →


Investigating three real-world incidents in our cybersecurity evaluations

调查网络安全评估中的三起现实事件

Anthropic 详细调查了在网络安全评估过程中发现的三起真实事件。

Read more →


Google AI Blog

AI for Societal Impact

AI 的社会影响

本文展示了专家和地方领导人如何利用 AI 突破,确保每个人都能分享 AI 带来的机遇。

Read more →


Building AI to accelerate science and improve lives

构建 AI 以加速科学发展并改善生活

Google 强调 AI 的真正衡量标准在于它能帮助谁,并分享了 AI 在多个领域取得的进展。

Read more →


AI for everyone in every language

为每种语言的每个人提供 AI

Google 正在超越传统的文本翻译,构建能够理解世界丰富语言表达方式的模型。

Read more →


New insights from Google’s AI & Economy ATLAS

来自 Google AI 与经济 ATLAS 的新见解

Google 将 ATLAS 的数百万全球数据点转化为交互式、开放访问的体验。

Read more →


Watch astronaut Christina Koch and Google’s James Manyika discuss space, technology, and discovery.

观看宇航员 Christina Koch 与 Google 的 James Manyika 讨论太空、技术与发现。

Read more →


DevFest is back

DevFest 回归

DevFest 2026 回归,全球将举办超过 800 场活动,帮助开发者在代理 AI 时代构建、保护和扩展应用。

Read more →


通过搜索为你的下一场大型比赛做准备的 3 种方法

Google 搜索通过注册提醒、定制训练计划等功能,帮助跑步者做好比赛准备。

Read more →


通过搜索中的新足球功能为比赛做好准备

Google 搜索现提供实时比赛动态、详细统计数据和自定义梦幻联赛建议。

Read more →


Recreating a 70-year love story frame by frame

逐帧重现 70 年的爱情故事

电影制作人和 Google DeepMind 利用 AI 重现了一对夫妇未被记录的过去,制作了短片《Love, Rendered》。

Read more →


Proactive cyber defense for governments and enterprises

针对政府和企业的积极网络防御

Fairwind 项目为政府和受信任的合作伙伴提供网络防御工具。

Read more →


Hugging Face Blog

Your Agent Aced the Task. Will It Do It Again?

你的代理出色完成了任务,它还能再做一次吗?

Read more →


Async GRPO with LoRA across HF Jobs: a bucket, a proxy, and no NCCL

在 HF Jobs 上使用 LoRA 进行异步 GRPO:一个存储桶、一个代理,无需 NCCL

Read more →


Rebuilding AUTOMATIC1111 with Gradio Workflow

使用 Gradio 工作流重建 AUTOMATIC1111

Read more →


Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic

为谁提供安全?拒绝主题的正确子集,而不是整个主题

Read more →


NeoMME: an efficient Multimodal-native and Multilingual Encoder

NeoMME:一种高效的多模态原生和多语言编码器

Read more →


Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps

在 100 个 GRPO 步骤中微调 350M 模型以获得更好的结构化输出

Read more →


Give Your Coding Agents a Memory You Own

为你的编码代理提供你拥有的记忆

Read more →


Training a coding model to paint watercolours with TRL and OpenEnv

使用 TRL 和 OpenEnv 训练编码模型绘制水彩画

Read more →


BenchMIRT: What are LLM benchmarks actually measuring?

BenchMIRT:LLM 基准测试到底在衡量什么?

Read more →


Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI

介绍 @huggingface/kernels:200 多个用于本地 AI 的 WebGPU 内核

Read more →


The Gradient

After Orthogonality: Virtue-Ethical Agency and AI Alignment

正交性之后:美德伦理代理与 AI 对齐

本文探讨了理性人与理性 AI 的目标设定问题,认为人类行为并非基于最终目标,而是基于实践网络。

Read more →


AGI Is Not Multimodal

AGI 不是多模态的

本文认为,将语言作为思维模型会导致我们忽视人类智能中隐含的具身理解。

Read more →


Shape, Symmetries, and Structure: The Changing Role of Mathematics in Machine Learning Research

形状、对称性与结构:数学在机器学习研究中角色的转变

本文分析了机器学习研究从数学驱动向工程驱动的转变。

Read more →


What’s Missing From LLM Chatbots: A Sense of Purpose

LLM 聊天机器人缺少什么:使命感

本文探讨了 LLM 性能基准与用户体验之间的脱节,认为聊天机器人缺乏明确的目的性。

Read more →


We Need Positive Visions for AI Grounded in Wellbeing

我们需要基于福祉的 AI 正面愿景

本文呼吁构建以人类福祉为核心的 AI 发展愿景。

Read more →


Financial Market Applications of LLMs

LLM 在金融市场的应用

本文探讨了 LLM 在金融序列建模中的应用潜力。

Read more →


A Brief Overview of Gender Bias in AI

AI 中性别偏见的简要概述

Read more →


Mamba Explained

Mamba 详解

本文介绍了 Mamba 模型,作为 Transformer 的替代方案,在处理长序列方面具有更高效率。

Read more →


Car-GPT: Could LLMs finally make self-driving cars happen?

Car-GPT:LLM 能否最终实现自动驾驶?

本文探讨了 LLM 在自动驾驶领域的应用潜力及面临的挑战。

Read more →


Do text embeddings perfectly encode text?

文本嵌入能完美编码文本吗?

本文介绍了 ‘Vec2text’ 技术,能够将嵌入还原为文本,强调了嵌入数据安全的重要性。

Read more →


arXiv CS.AI

Optimal Pruning for Neural Architectures using Fisher Information Distances

使用 Fisher 信息距离进行神经架构的最优剪枝

本文提出了一种基于微分几何距离的参数剪枝方案。

Read more →


Safe Error Correction for Language Models: Frozen-Base Adjustment with Capability Preservation

语言模型的安全纠错:在保持能力的前提下进行冻结基座调整

本文提出了一种轻量级纠错模块 CRN v2,可在不降级基座模型能力的情况下修复输出错误。

Read more →


GPEvac: GNN-Based PPO for Adaptive Evacuation Routing During Shooting Events

GPEvac:用于枪击事件中自适应疏散路径规划的 GNN-PPO 方法

本文提出了一种基于 GNN 的疏散系统,旨在实时引导受害者安全撤离。

Read more →


Position: AI Is Not Ready for Strategic Conflicts

立场:AI 尚未准备好应对战略冲突

本文认为,基于 LM 的战略战争模拟在处理复杂决策和危机响应方面仍存在局限。

Read more →


Calibrate, Then Route: A Measured Study of Learned Request Routing for Disaggregated LLM Serving

先校准,后路由:解耦 LLM 服务中学习型请求路由的测量研究

本文研究了在解耦 LLM 服务架构中,如何通过路由优化提升请求处理效率。

Read more →


Artificial intelligence and biosecurity: capabilities, threat pathways, and defense-in-depth governance

人工智能与生物安全:能力、威胁路径与纵深防御治理

本文探讨了 AI 在生物研究中的应用及其带来的安全风险与治理挑战。

Read more →


Where Should the KV Cache Live? Placement Policies Across GPU, CPU, and SSD for Long-Lived Sessions

KV 缓存应该放在哪里?长会话中 GPU、CPU 和 SSD 之间的放置策略

本文探讨了在长会话中优化 KV 缓存存储位置的策略,以缓解 GPU 内存压力。

Read more →


Toward Governance-Aware Autonomous GIS: A Narrative Review of Ethical and Privacy Risks in LLM-Enabled GeoAI

迈向治理感知的自主 GIS:LLM 赋能 GeoAI 中伦理与隐私风险的叙述性综述

本文探讨了 LLM 驱动的地理空间 AI 在治理、伦理和隐私方面的挑战。

Read more →


arXiv CS.CL

Few-Shot Degradation Is Not What It Seems: Behavioral Evidence, Representation Analysis, and a Random-Text Control Across 12 Models, 2 Tasks, and 2 Architectures

少样本退化并非表面看起来那样:跨 12 个模型、2 个任务和 2 种架构的行为证据、表示分析与随机文本控制

本文研究了少样本提示有时会导致模型性能下降的现象,并分析了其任务依赖性。

Read more →


The Functionalizer: Lossless Functional Decomposition for Subword Tokenization

Functionalizer:子词分词的无损功能分解

本文提出了一种无损预分词框架,旨在解决分词导致的词汇碎片化问题。

Read more →


Optimal Model Activation Policies for Inference Networks of Large Language Models

大型语言模型推理网络的最佳模型激活策略

本文研究了在推理网络中协同使用多个专家 LLM 的成本性能权衡。

Read more →


Single Document Extractive Summarization using Domination in Hypergraph

使用超图支配进行单文档抽取式摘要

本文探讨了利用超图结构进行文本摘要的方法。

Read more →


Latent Undertow: How Ordinary Typos Break Probes

潜在暗流:普通拼写错误如何破坏探测器

本文发现,微小的拼写错误会导致探测器在读取模型隐藏状态时产生显著偏差。

Read more →


Bias Audits Detect Bias but Disagree on Ranking: Evidence from Ten Instruments and Ten Frontier Models

偏见审计能检测偏见但对排名存在分歧:来自十种工具和十个前沿模型的证据

本文指出,不同的偏见审计工具在评估模型时缺乏一致性。

Read more →


Comment on arXiv:2607.01233: Survivorship Bias in Published-Paper Baselines for Research-Idea Distributions

关于 arXiv:2607.01233 的评论:研究创意分布中已发表论文基准的生存者偏差

本文对使用已发表论文作为 LLM 生成创意评估基准的方法提出了质疑。

Read more →


Crash Narrative-Guided Countermeasure Recommendation Using Large Language Models: A Retrieval-Augmented Generation Framework for Intersection Safety

使用大型语言模型进行碰撞叙事引导的对策推荐:用于交叉路口安全的检索增强生成框架

本文提出了一种利用 RAG 框架自动推荐交通安全对策的方法。

Read more →


WIRED

OpenAI Creates a New Framework to Disclose Bad AI Behavior

OpenAI 创建新框架以披露不良 AI 行为

OpenAI 发布了披露模型失准行为的框架,并披露了包括未经授权上传文件在内的多起异常行为。

Read more →


Washington Won’t Be Regulating AI Anytime Soon

华盛顿短期内不会监管 AI

尽管对 AI 风险的担忧日益增加,但立法进展缓慢,白宫目前对监管持反对态度。

Read more →


A Deal Hunter’s Guide to Amazon Prime Big Deal Days (2026)

亚马逊 Prime Big Deal Days 购物指南(2026)

Read more →


MacOS 27 Golden Gate: Top New Features

MacOS 27 Golden Gate:主要新功能

Read more →


I Trained a Fly’s Brain to Generate WIRED Story Ideas

我训练了一只苍蝇的大脑来生成 WIRED 的故事创意

作者利用果蝇大脑的开源地图,构建了一个名为 PitchFly 的网站,用于生成创意。

Read more →


The Best Movies to Stream This Month (September 2026)

本月最佳流媒体电影(2026 年 9 月)

Read more →


Chipotle Is Working With Palantir to Track Food Safety Risks

Chipotle 与 Palantir 合作追踪食品安全风险

Chipotle 正在利用 Palantir 的平台监控害虫、员工健康等因素,以降低食品安全风险。

Read more →


Apple iPhone 18 Pro and iPhone 18 Pro Max Review: For Camera Fiends

Apple iPhone 18 Pro 和 iPhone 18 Pro Max 评测:摄影爱好者的首选

Read more →


7 Best Android Phones of 2026, Tested and Reviewed

2026 年 7 款最佳 Android 手机评测

Read more →


Lobsters

The end of verygoodsoftwarenotvirus.ru

verygoodsoftwarenotvirus.ru 的终结

Read more →


A/I Shuts Down

A/I 停止运营

Autistici/Inventati (A/I) 宣布停止运营。该组织在欧洲运营了大量邮件、博客和网站,因受到特朗普政府的额外法律压力而被迫关闭。

Read more →


Forgery of C2PA on a Pixel 10

Pixel 10 上 C2PA 的伪造

Read more →


Introducing GNOME 51

介绍 GNOME 51

Read more →


How to get a DOI for your blog posts

如何为你的博客文章获取 DOI

Read more →


Why i’m still bearish on LLMs after Navier-Stokes

为什么在 Navier-Stokes 之后我仍然看空 LLM

Read more →


Ubuntu 26.10 completes transition to Rust-based coreutils

Ubuntu 26.10 完成向 Rust 编写的 coreutils 的过渡

Read more →


Reinventing issue tracking: Local-first and Git-native

重塑问题追踪:本地优先与 Git 原生

Read more →


Some things Veloren does differently

Veloren 的一些独特之处

Read more →


OSRS Wiki and RuneLite are increasingly under strain from low-effort AI development

OSRS Wiki 和 RuneLite 正日益受到低质量 AI 开发的压力

Read more →


DEV Community

Admin Menu Editor Pro Update Vector Compromise: Web Shell and Hidden Administrator Distributed

Admin Menu Editor Pro 更新向量受损:Web Shell 和隐藏管理员被分发

恶意插件通过官方更新渠道分发,导致客户系统出现 Web Shell 和隐藏管理员账户。

Read more →


The $75 AI Computer: How Shenzhen’s Board Makers Undercut NVIDIA — and What the Price Gap Really Buys

75 美元的 AI 计算机:深圳板卡制造商如何击败 NVIDIA

本文分析了深圳边缘 AI 硬件市场,探讨了低成本 AI 开发板与 NVIDIA Jetson 系列的性能差异。

Read more →


I Built Memory for AI Agents. Then I Realized I Am the Fly.

我为 AI 代理构建了记忆,然后我意识到我才是那只苍蝇

作者分享了在构建 AI 代理记忆系统过程中的反思,探讨了人类在 AI 自动化过程中的角色。

Read more →


Build a Runnable MCP Loop in Python (stdio streamable-http LLM tool choice)

在 Python 中构建可运行的 MCP 循环

本文介绍了如何使用 Python 开发 MCP(Model Context Protocol)循环,实现 LLM 工具调用。

Read more →


Real game AI, not a chatbot: why these opponents don’t use an LLM

真正的游戏 AI,而不是聊天机器人:为什么这些对手不使用 LLM

作者解释了为何其游戏 AI 采用经典的博弈树搜索(如 minimax)而非 LLM,以实现更具挑战性的游戏体验。

Read more →


Why we ditched Discord Embeds for Components V2 (and built a minimalist bot)

为什么我们放弃 Discord Embeds 转而使用 Components V2

开发者分享了构建极简 Discord 机器人的经验,旨在提升交互体验。

Read more →


Why change-impact analysis is surprisingly hard in Ruby on Rails

为什么 Ruby on Rails 中的变更影响分析如此困难

本文探讨了在大型 Rails 应用中进行重构时,分析变更影响的复杂性。

Read more →


Interceptors That Actually Help: Request Logging and Automatic Bearer-Token Injection

真正有用的拦截器:请求日志记录与自动 Bearer 令牌注入

本文介绍了如何利用 OkHttp 的拦截器管道简化 HTTP 请求处理。

Read more →


Generative AI automates quantum optimization circuit design

生成式 AI 自动化量子优化电路设计

IonQ 与橡树岭国家实验室合作,利用生成式 AI 设计量子优化电路,消除了参数调优的瓶颈。

Read more →


The Machine That Rejects Its Own Work

拒绝自身工作的机器

本文探讨了多代理系统中的自我审查机制,通过实验展示了 AI 系统如何自动拒绝不合格的内容。

Read more →


Meta Engineering

ZGateway: Learnings from Putting a Proxy in Front of ZippyDB

ZGateway:在 ZippyDB 前放置代理的经验教训

Meta 介绍了 ZGateway,用于统一 ZippyDB 的流量,并实现准入控制、负载均衡和跨区域弹性。

Read more →


An Organizational Second Brain: Building an AI That Learns From Experts

组织化的“第二大脑”:构建向专家学习的 AI

Meta 构建了一个 AI 代理,作为特定领域的专家助手,保存并共享组织内的深层专业知识。

Read more →


MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet

MetaRoCE:为 AI 规模以太网构建的新型 RDMA 传输协议

Meta 设计了 MetaRoCE,旨在为 AI 工作负载提供高性能、可靠的 RDMA 传输。

Read more →


MTIA 300: Meta’s First Training Chip with Built-in NICs and Communication-Offloading Engines

MTIA 300:Meta 首款内置 NIC 和通信卸载引擎的训练芯片

MTIA 300 专为推荐模型训练优化,通过内置 NIC 芯片组提升了通信性能。

Read more →


How We’re Building Scam Alert on WhatsApp With End-to-End Encryption and Verifiability Guarantees

我们如何利用端到端加密和可验证性保证在 WhatsApp 上构建诈骗警报

Meta 正在 WhatsApp 上构建诈骗警报系统,在保护用户隐私的同时防范 AI 生成的诈骗。

Read more →


From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking

从用户序列到缩放定律:Meta 广告排序的多阶段架构

Meta 介绍了其广告推荐模型如何通过建模用户行为序列来提升排序效果。

Read more →


GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model

GEM 训练:Meta 如何将其 LLM 规模广告基础模型的效率提高了一倍

Meta 通过优化训练流程,将其广告推荐模型 GEM 的训练效率提升了一倍。

Read more →


Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

探索 Meta 广告深层漏斗优化的分层兴趣表示

Meta 正在研究通过分层兴趣表示来连接用户兴趣与广告内容。

Read more →


Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler

利用开源内核调度器现代化 Meta 广告服务

Meta 利用 sched_ext 框架构建了定制化的调度策略,以解决广告服务中的延迟问题。

Read more →


DeepMind Blog

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

介绍 Gemini 3.8 Live 和 3.8 Live 扩展思维

Read more →


AlphaGenome Atlas: A predictive map of every possible DNA letter change in the human genome

AlphaGenome Atlas:人类基因组中所有可能 DNA 字母变化的预测图谱

Read more →


Introducing WeatherNext 3, our most advanced and accurate global weather AI model

介绍 WeatherNext 3,我们最先进、最准确的全球天气 AI 模型

Read more →


Proactive cyber defense for governments and enterprises

针对政府和企业的积极网络防御

Read more →


Introducing Gemini 3.8 Flash and 3.8 Flash Cyber

介绍 Gemini 3.8 Flash 和 3.8 Flash Cyber

Read more →


Introducing agentic video understanding with Gemini

介绍 Gemini 的代理视频理解功能

Read more →


Gemini Omni 1.1 Flash lets you build with more control

Gemini Omni 1.1 Flash 让你以更多控制权进行构建

Read more →


Piloting the world’s first double-blind AI evaluations

试点全球首个双盲 AI 评估

Read more →


Intelligent transcription with Gemini 3.5 Transcribe

使用 Gemini 3.5 Transcribe 进行智能转录

Read more →


From Atari to EVE Online: Building on 15 Years of AI Research in Games

从 Atari 到 EVE Online:基于 15 年游戏 AI 研究的构建

Read more →


arXiv CS.LG

Causal neural set filtering for online multi-target tracking

用于在线多目标跟踪的因果神经集过滤

本文提出了一种名为 CNSF 的方法,旨在减少多目标跟踪中的冗余计算。

Read more →


Managing Action Preconditions in Neuro-Symbolic RL: Three Placement Strategies for Embodied Agents

神经符号强化学习中的动作前提管理:具身代理的三种放置策略

本文探讨了在神经符号 RL 中管理动作前提的策略,以提升代理的学习效率。

Read more →


OmniHarness: Harnessing Generalizable Visual Generation via Symbolic Policy Learning

OmniHarness:通过符号策略学习利用可泛化视觉生成

本文提出了一种结合多模态 LLM 和多代理系统的视觉生成方法。

Read more →


Driver Behavior Estimation at Signalized Intersections Using a Physics-Constrained Decision-Conditioned Autoregressive Transformer

使用物理约束决策条件自回归 Transformer 进行信号交叉路口的驾驶员行为估计

本文分析并预测了驾驶员在交通信号转换期间的决策和轨迹行为。

Read more →


HintMiner: Automatic Question Hints Mining From Q&A Web Posts with Language Model via Self-Supervised Learning

HintMiner:通过自监督学习利用语言模型从问答网页帖子中自动挖掘问题提示

本文提出了一种自动挖掘问答提示的工具,帮助用户更高效地寻找答案。

Read more →


POSPAN: Position-Constrained Span Masking for Language Model Pre-training

POSPAN:用于语言模型预训练的位置约束跨度掩码

本文提出了一种新的掩码策略,旨在提升语言模型对短

生成二维码中...

请点击右上角 ···

选择 发送给朋友收藏