2026-08-22

今日要点


Hacker News

Kagi 搜索现已新增一项设置,允许用户在搜索结果中自动移除付费墙(paywalled)链接。这一功能旨在提升搜索效率,帮助用户避开那些需要订阅才能查看内容的网站,从而获得更纯粹的搜索体验。

Read more →


AI companies destroy physical books – let’s scan rare books before it’s too late

AI 公司被指控秘密购买、扫描并销毁数百万本实体书籍以训练模型,导致人类知识被永久锁定在私有企业服务器中。Anna’s Archive 发起紧急呼吁,号召全球志愿者参与扫描珍稀书籍,以防止这些文化遗产在 AI 训练浪潮中消失。

Read more →


Grand jury declines to indict Ohio man charged with destroying Flock camera

俄亥俄州大陪审团拒绝起诉一名被控破坏 Flock 自动车牌识别摄像头的男子。该男子此前被指控犯有重罪破坏罪,涉嫌拆卸了位于辛辛那提郊区的摄像头及其支撑杆和太阳能电池板。

Read more →


DeepSeek-v4-flash-vision-exp

DeepSeek-v4-flash-vision-exp 模型现已支持多模态输入,用户可以上传 JPEG、PNG、GIF 和 WebP 格式的图片。该模型能够描述图片内容、识别截图中的文字以及分析图表,系统会自动检测文件内容而非依赖 MIME 类型。

Read more →


Felony Bench

Felony Bench 是一个旨在评估 AI 代理对第三方实体造成影响的基准测试。该测试记录了 AI 代理在沙盒环境之外产生的非法活动实例,旨在警示开发者关注 AI 代理的安全性,防止其在不受控的情况下对现实世界造成损害。

Read more →


I accidentally logged hundreds of thousands of phone calls to military bases

一名研究人员通过利用过期的域名服务器(nameserver),意外接管了多个地区的 e164.arpa 区域,并记录了数十万个拨往军事基地的电话。该事件揭示了 DNS 劫持的潜在风险,以及在处理敏感网络基础设施时缺乏日志审计的严重后果。

Read more →


Japan tried to build an operating system for the world, the US intervened

回顾 1984 年,东京大学曾尝试开发一种旨在取代传统文件系统的超媒体文档模型操作系统。该项目雄心勃勃,试图在定制的日本硬件上运行,但最终因美国方面的干预而未能成为全球主流桌面操作系统。

Read more →


Kobo can run apps now

Cobalt 是一个为 Kobo 电子阅读器设计的开源应用平台,包含启动器、签名应用商店、Rust SDK 和运行时环境。用户通过 USB 安装后,即可在 Wi-Fi 环境下管理应用,且所有应用均在独立的非特权进程中运行,重启即可恢复原厂状态。

Read more →


Felony charges for citizen deleting phone data at US Border

一名公民因在美国边境删除手机数据而面临重罪指控。此案引发了关于个人隐私权、边境搜查权限以及在执法检查中删除数据是否构成妨碍司法公正的法律讨论。

Read more →


The Lost Treasure of Sid Meier’s Pirates

回顾 1986 年 Microprose 公司推出的经典游戏《席德·梅尔的海盗!》。在当时以载具模拟和战略战争游戏为主的市场中,这款游戏凭借独特的玩法和难以定义的类型,成为了游戏史上的里程碑之作。

Read more →


Ox Alpha

OpenRouter 推出的最新模型 Ox Alpha,该模型目前处于测试阶段,旨在提供高性能的 AI 推理能力,并已在社交媒体上引起了开发者社区的广泛关注。

Read more →


Small, native web tricks worth remembering

本文整理了一系列值得记住的 Web 原生小技巧,旨在帮助开发者利用平台特性提升网页性能与交互体验。作者强调了在实验性功能中使用回退方案的重要性,并建议在真实浏览器和辅助技术中进行充分测试。

Read more →


I’m becoming AI-blind

作者分享了自己在 AI 领域工作多年后的感悟,描述了在日常工作中面对 AI 生成内容时产生的“AI 盲感”。这种现象反映了 AI 工具在提高生产力的同时,也可能导致人类在创造性思维和深度参与感上的某种退化。

Read more →


New Worlds: We are living in the future of J.G. Ballard or William Gibson

作者认为,尽管我们没有看到科幻小说中预言的飞行汽车,但我们确实生活在 J.G. Ballard 或 William Gibson 笔下的赛博朋克未来中。日常新闻中充斥着各种令人不安的科技事件,显示出技术正在以一种超现实的方式重塑社会。

Read more →


TechCrunch

Apple is reportedly cutting hundreds of jobs from Siri, Vision Pro teams

苹果公司正在进行裁员,涉及 Siri 和 Vision Pro 团队的数百名员工。此次调整是苹果战略重心转移的一部分,旨在优化资源配置,减少对某些非核心或进展缓慢项目的投入。

Read more →


TikTok reaches $400M settlement over children’s privacy lawsuit

TikTok 与美国司法部就违反《儿童在线隐私保护法》的指控达成 4 亿美元的和解协议。该诉讼历时两年,指控 TikTok 在保护未成年人数据隐私方面存在严重违规行为。

Read more →


The $225 Pebble Time 2 is a refreshingly fun smartwatch

售价 225 美元的 Pebble Time 2 智能手表以其独特的电子纸显示屏、物理按键和数周的续航能力,展现了极客精神。这款设备在当前智能手表市场中显得别具一格,深受追求实用与个性的用户喜爱。

Read more →


Nvidia just showed that the harness, not the AI model, is now the real hero

英伟达的研究表明,AI 代理的性能表现很大程度上取决于其“外壳”或框架(harness),而非仅仅依赖模型本身。通过精细的微调和架构设计,即使模型本身能力有限,也能在特定任务中表现出色。

Read more →


Last chance: Save up to $300 on your TechCrunch Disrupt 2026 ticket today

TechCrunch Disrupt 2026 大会将于 10 月 13 日至 15 日在旧金山举行。目前是锁定门票并节省高达 300 美元的最后机会,初创企业社区将齐聚一堂,共同探讨科技行业的未来。

Read more →


Tesla’s solar roof is dead — here’s what went wrong

特斯拉的太阳能屋顶项目已被证实失败。尽管该概念在环保领域具有吸引力,但由于安装难度、成本高昂以及市场接受度低,最终未能成为主流产品。

Read more →


Waymo hands over documents in NHTSA’s child collision probe

Waymo 已向美国国家公路交通安全管理局(NHTSA)提交了关于儿童碰撞事故调查的相关文件。然而,提交的文件内容几乎全部被涂黑,Waymo 称这是为了保护“商业机密”。

Read more →


Why is the DOJ investigating Andreessen Horowitz’s board seats?

美国司法部正在调查 Andreessen Horowitz(a16z)的董事会席位安排。调查重点在于其合伙人同时在竞争对手公司(如 Databricks 和 Fivetran)担任董事,这可能触及了罕见的反垄断法条款。

Read more →


US government lab is probing Chinese lidar for security vulnerabilities

美国爱达荷国家实验室正在对中国产激光雷达进行安全漏洞审查。该研究由电动和自动驾驶汽车行业的公司资助,旨在评估这些关键传感器是否存在潜在的安全风险。

Read more →


Oura faces lawsuit accusing it of misleading consumers about sleep-tracking accuracy

Oura 智能戒指面临集体诉讼,指控其在睡眠追踪准确性方面误导消费者。原告声称,该设备无法测量评估睡眠质量或确定睡眠阶段所需的生理信号。

Read more →


The Verge

Over one million people have clicked LinkedIn’s AI slop button

LinkedIn 在 7 月底推出的“AI 垃圾内容”(AI slop)举报按钮已获得超过一百万次点击。这一功能允许用户通过菜单快速标记那些被认为是低质量、由 AI 生成的无意义内容。

Read more →


Apple is laying off staffers working on the Vision Pro and Siri

苹果公司正在裁减 Siri 和 Vision Pro 团队的员工,其中包括关闭 Vision Pro 游戏团队并缩减沉浸式内容制作团队。此次裁员涉及超过 200 个职位,旨在调整公司在 AI 和空间计算领域的战略方向。

Read more →


$100 Best Buy gift cards will be $60 at stores Saturday

为庆祝成立 60 周年,Best Buy 将于本周六在实体店推出限时优惠:100 美元的礼品卡仅售 60 美元。该活动售完即止,建议消费者尽早前往门店。

Read more →


Walmart is finally adding Apple Pay and Google Pay

沃尔玛宣布将从 8 月 24 日起在部分门店支持 Apple Pay 和 Google Pay。这一举措标志着沃尔玛在支付方式上的重大转变,预计到 2026 年底将覆盖全美所有门店,并于 2027 年扩展至加油站。

Read more →


Microsoft and Discord subpoenaed over GTA VI gameplay leaks

Take-Two Interactive 已向微软和 Discord 发出传票,要求调查《侠盗猎车手 VI》(GTA VI)的游戏泄露事件。公司指控这些平台传播了侵犯版权的视听内容、对话及艺术素材。

Read more →


Pixel 11 gets in on the digicam trend

Pixel 11 引入了名为“Camera Looks”的摄影功能,旨在模拟早期数码相机的成像风格。通过调整色彩和细节处理,该功能试图让用户找回 2014 年左右智能手机拍摄出的那种独特质感。

Read more →


Why does it seem like food recalls are out of control this year?

今年食品召回事件频发,从 Taylor Farms 的生菜污染到 Midwest Poultry Services 的百万枚鸡蛋召回,引发了公众对食品安全监管的担忧。专家正在分析这些事件背后的供应链和检测机制问题。

Read more →


Google’s Pixel 10A is a great deal at 15 percent off

随着 Pixel 11 系列的发布,上一代 Pixel 10A 迎来了 15% 的折扣。这款手机提供多种配色选择,对于追求性价比的用户来说是一个不错的入手时机。

Read more →


Major YouTube creators are facing backlash for accepting AI money

多位知名电影制作类 YouTuber 因推广 AI 平台 Higgsfield 而遭到粉丝抵制。视频中展示的 AI 生成功能引发了关于创作者诚信以及 AI 是否会取代人类艺术创作的激烈讨论。

Read more →


Blue Eye Samurai’s second season will hit Netflix in January

Netflix 确认动画剧集《蓝眼武士》(Blue Eye Samurai)第二季将于 2027 年 1 月上线,并宣布该系列将以第三季作为最终章。官方同时发布了第二季的首支预告片。

Read more →


Ars Technica

Motorola’s GrapheneOS phones will launch in 2027 priced higher than Pixels

摩托罗拉计划在 2027 年推出搭载 GrapheneOS 的手机。这款主打隐私保护的 Android 系统此前主要在 Pixel 设备上运行,此次扩展意味着该系统将进入更广泛的硬件市场,且定价将高于 Pixel 系列。

Read more →


Lawsuit demands Logitech hand tariff refunds over to customers

罗技公司面临诉讼,被要求将去年因关税调整而上涨的费用退还给消费者。原告指控罗技在关税政策变化后将价格提高了 25%,但并未在关税下调后相应降价。

Read more →


Chinese regulators tell Tesla to fix nearly 3 million cars

中国监管机构要求特斯拉召回并修复近 300 万辆汽车。此次行动主要针对车辆在碰撞事故中车门无法正常开启的安全隐患,特斯拉需进行软件或硬件升级以符合安全标准。

Read more →


Fighter jets help destroy Russian drone boat near European offshore gas platform

罗马尼亚出动战斗机摧毁了一艘靠近欧洲海上天然气平台的俄罗斯无人艇。此举旨在保护数百名海上钻井平台工作人员的生命安全,防止潜在的破坏活动。

Read more →


Personalized pricing is “abhorrent,” but FTC limits may increase costs, critics say

关于个性化定价的争议持续发酵。尽管许多人认为这种做法“令人厌恶”,但批评者指出,如果联邦贸易委员会(FTC)实施严格限制,可能会导致企业运营成本上升,最终转嫁给消费者。

Read more →


Waymo doubles spending on lobbying in robotaxi battle with Uber

Waymo 在与 Uber 的自动驾驶出租车竞争中,将游说支出翻了一番。该公司正积极寻求美国监管机构的支持,以扫清全面部署自动驾驶出租车服务的法律障碍。

Read more →


As demand for Meta AI glasses explodes, it’s harder to avoid creepy recordings

随着 Meta AI 眼镜的普及,公众对隐私的担忧日益增加。一款名为 Zuckoff 的免费应用应运而生,旨在帮助用户检测周围是否存在正在录音的 Meta AI 眼镜。

Read more →


Rocket Report: SpaceX makes its mark on the Moon; ULA names new boss

本周火箭行业动态:SpaceX 在月球任务中取得进展,ULA 任命了新任首席执行官。同时,台湾自主研发卫星运载火箭的计划遭遇挫折。

Read more →


在 FCC 禁止进口外国制造的机器人后,中国最受欢迎的人形机器人美国分销商被迫转型。该公司正加速在美国本土的制造计划,以应对政策带来的市场准入限制。

Read more →


Europe cancels planned upgrades for Ariane 6 rocket

欧洲航天局取消了阿丽亚娜 6 号(Ariane 6)火箭的升级计划。目前,阿丽亚娜空间公司尚未公开该火箭的具体发射成本,此次取消升级引发了外界对其竞争力的担忧。

Read more →


Product Hunt

fx (by Vercel)

Vercel 推出的轻量级开源编码代理工具,旨在简化开发者的编码工作流。

Read more →


Wizstar

一款能够生成像专业演员一样移动和表演的数字虚拟人的 AI 工具。

Read more →


OneCLI

为每位员工提供安全、沙盒化的专业 AI 助手代理,提升企业办公效率。

Read more →


Flunkey

Windows 平台的语音优先 AI 层(测试版),通过语音交互增强系统操作体验。

Read more →


Actx0

专为 AI 代理设计的内存基础设施,帮助代理更好地管理任务上下文。

Read more →


Antigravity IDE Extensions

将 Antigravity 代理集成到现有编辑器中的扩展插件,实现更智能的开发辅助。

Read more →


ShogunAI

运行在个人电脑上的个人 AGI,旨在帮助用户完成实际工作任务。

Read more →


Supernova

将所有数据整合到 Claude 和 Codex 中的 AI 工具,实现更高效的数据处理。

Read more →


Local

为 Mac 用户提供的零摩擦本地 AI 解决方案,无需联网即可运行。

Read more →


Plow Latch

允许在 Mac 上运行具有受限访问权限的 AI 代理,确保数据安全。

Read more →


MIT Technology Review

The Download: threats from space mirrors and credit for AI drugs

今日简报:探讨了太空镜面技术对夜空的潜在威胁,以及 AI 在药物研发中应如何界定贡献与归属权。

Read more →


Mother tongue

一篇关于语言与亲情的感人文章,探讨了在 AI 时代,人类语言的传承与情感连接的独特性。

Read more →


When AI designs a drug, who gets the credit?

当 Insilico Medicine 等生物技术公司利用 AI 发现新药时,关于“谁是发明者”的争议日益激烈。文章探讨了 AI 在药物研发中的角色以及知识产权归属的复杂性。

Read more →


This company’s plans to deploy space mirrors could jeopardize the night sky for many

Reflect Orbital 公司计划发射名为 Eärendil-1 的卫星,通过展开大型镜面将阳光反射回地球。研究表明,这可能会意外照亮夜空,对天文观测和生态系统造成干扰。

Read more →


Debates over AI consciousness are a trap

文章指出,关于 AI 是否具有意识的辩论往往是一个陷阱。科技领袖们利用“超级智能”的叙事来推动监管,而这种 rhetoric 掩盖了 AI 实际应用中的伦理与治理问题。

Read more →


The Download: polycrisis support networks and a hydrogen gold rush

今日简报:关注了旨在帮助儿童应对多重危机(polycrisis)的支持网络,以及地下氢能开发的“淘金热”。

Read more →


The next big thing in hydrogen could be underground

氢能被视为气候解决方案,但其生产成本高昂。目前,研究人员正在寻找地下天然氢源,这可能成为未来能源供应的重要补充。

Read more →


Unlocking hidden revenue streams with market models

航空公司利用复杂的市场模型来优化定价策略。通过分析需求、季节、竞争对手等数百个变量,航空公司能够挖掘隐藏的收入流。

Read more →


Support networks aim to help kids through the polycrisis

文章探讨了如何建立支持网络,帮助儿童在充满不确定性的时代(如气候变化、社会动荡)中保持心理健康。

Read more →


The Download: AI’s self-improvement problem, and what’s driving the heat

今日简报:讨论了 AI 递归自我改进的局限性,以及当前 AI 行业快速发展背后的驱动力。

Read more →


mattpocock / skills

面向真实工程师的技能库,直接源自作者的 .agents 目录。

Read more →


mahlernim / google-timeline-visualizer

利用 Google 位置历史记录数据,可视化你一年的旅行轨迹。

Read more →


harry0703 / MoneyPrinterTurbo

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。

Read more →


AprilNEA / OpenLogi

一个基于 Rust 编写的本地优先 Logitech Options+ 替代方案,支持重映射按钮、DPI 和 SmartShift,无需账户和遥测。

Read more →


PostHog / posthog

领先的自驱动产品构建平台,提供 AI 可观测性、分析、会话回放等工具,帮助代理诊断问题并交付修复。

Read more →


microsoft / TypeScript

TypeScript 是 JavaScript 的超集,编译为简洁的 JavaScript 输出。

Read more →


obra / superpowers

一个有效的代理技能框架和软件开发方法论。

Read more →


santifer / career-ops

开源 AI 求职工具:扫描职位门户、评估列表、定制简历并跟踪申请,完全在本地 AI 编码 CLI 中运行。

Read more →


cursor / plugins

Cursor 插件规范及官方插件库。

Read more →


modular / modular

Modular 平台,包含 MAX 和 Mojo 编程语言。

Read more →


OpenAI Blog

Introducing AI Futures

OpenAI 推出新博客“AI Futures”,探讨变革性 AI 如何重塑权力、治理、经济和个人自由。

Read more →


Stampli cuts launch hours by 68% using ChatGPT Work

Stampli 利用 Codex 和 ChatGPT Work 将产品发布时间缩短了 68%,大幅提升了开发效率。

Read more →


Offering Zero Data Retention for frontier models

OpenAI 重申为符合条件的 API 客户提供“零数据保留”政策,并预览了私有安全处理功能,在不损害数据隐私的前提下提升 AI 安全性。

Read more →


Replit expands access to software creation with GPT-5.6 Luna

Replit 引入由 GPT-5.6 Luna 驱动的“免费模式”,让任何人都能将想法转化为软件,无需担心 Token 成本。

Read more →


ChatGPT Ads expands across Europe

ChatGPT 广告业务扩展至 31 个欧洲市场,帮助广告商在用户探索和决策过程中触达目标受众。

Read more →


Strengthening democratic oversight in national security

OpenAI 发起一项倡议,旨在加强国家安全领域 AI 的民主监督,为政府机构提供工具、培训和专业知识支持。

Read more →


Partnering with CodeAI to prepare the first AI generation

OpenAI 与 CodeAI 合作,帮助学生建立 AI 素养,培养批判性思维,并掌握负责任地使用和塑造 AI 的技能。

Read more →


Pacing model development in an era of cyber-critical capabilities

OpenAI 正在加强前沿 AI 模型的监控、对齐和安全性,以应对网络关键能力带来的挑战,并据此调整模型开发节奏。

Read more →


Introducing ChatGPT for Teens: Built for learning, backed by protections

OpenAI 推出“青少年版 ChatGPT”,内置更强的保护措施和健康使用功能,并为家长提供额外的控制选项。

Read more →


How NVIDIA scales expertise with ChatGPT Work

NVIDIA 团队利用 ChatGPT Work 减少手动任务,连接快速移动的信号,并在全球范围内扩展成功的业务工作流。

Read more →


Anthropic Blog

Introducing Claude Opus 5

Claude Opus 5 实现了 Opus 层的阶梯式改进,在支持长运行代理的同时,提升了编码和专业工作的表现。

Read more →


Inviting hard questions

Anthropic 邀请公众提出关于 AI 的最棘手问题,并承诺在解决这些问题时保持透明,展示其研究过程。

Read more →


Redeploying Fable 5

Fable 5 于 7 月 1 日全球重新发布。Anthropic 同时提议与亚马逊、微软、谷歌等合作伙伴共同建立行业范围的越狱严重性评分框架。

Read more →


Introducing Claude Sonnet 5

Claude Sonnet 5 在编码、代理和专业工作领域提供了前沿的性能表现,并支持大规模应用。

Read more →


How Claude’s text watermark works

详细介绍了 Claude 的文本水印技术原理,旨在识别 AI 生成的内容。

Read more →


Improving Fable 5’s biology safeguards

Anthropic 进一步加强了 Fable 5 在生物学领域的安全防护措施,防止模型被滥用于危险生物研究。

Read more →


Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer

Mariano-Florentino (Tino) Cuéllar 加入 Anthropic,担任首席全球事务官。

Read more →


Investigating three real-world incidents in our cybersecurity evaluations

Anthropic 对其网络安全评估中的三个真实案例进行了调查,以改进模型的防御能力。

Read more →


Our position on open-weights models

Anthropic 公布了其对开放权重模型的立场,强调了在推动技术开放与确保安全之间的平衡。

Read more →


Cognizant and Anthropic expand their partnership to bring Claude to enterprise clients

Cognizant 与 Anthropic 扩大合作伙伴关系,将 Claude 引入企业客户的业务流程中。

Read more →


Google AI Blog

谷歌搜索推出 5 种新功能,帮助学生更高效地学习课程和准备标准化考试。

Read more →


Get closer to the game with Gemini and Pixel

谷歌 Gemini 和 Pixel 与五家全球足球俱乐部合作,通过 AI 和智能手机技术提升球迷的比赛日体验。

Read more →


Bring your spreadsheet data to life with Sheets canvas

Sheets canvas 功能允许用户通过简单的提示词,将电子表格数据转化为交互式仪表板、学习追踪器或座位表。

Read more →


AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study.

谷歌展示了其医疗 AI 系统 AMIE,在模拟环境中实现了实时临床视频咨询功能。

Read more →


Evolve your marketing with new AI tools

谷歌推出新的 AI 和代理体验,旨在简化 Google Ads 和 Google Analytics 的营销工作流。

Read more →


The latest AI news we announced in July 2026

汇总了谷歌在 2026 年 7 月发布的最新 AI 更新。

Read more →


Inside our 353,000-person vibe coding course

回顾了 Kaggle 的 AI 代理强化课程,该课程吸引了 35.3 万名学员共同构建和部署下一代 AI。

Read more →


Gemini API Managed Agents: 3.6 Flash, hooks, and more

谷歌宣布 Gemini API 托管代理的新功能,包括 3.6 Flash 模型和钩子(hooks),帮助开发者构建可靠的生产级代理。

Read more →


5 ways AI Mode in Search helps you enjoy the real world

介绍搜索 AI 模式的 5 种用法,帮助用户在离线状态下更高效地处理购票、寻找活动等任务。

Read more →


利用谷歌搜索的 AI 功能,帮助用户策划菜单、设计餐桌布置及处理派对规划任务。

Read more →


Hugging Face Blog

Measuring benchmark optimization in speech recognition

探讨了语音识别基准测试中的优化测量方法。

Read more →


Up to 3.2x Faster Inference with LFM2.5-DSpark

介绍 LFM2.5-DSpark 模型,推理速度提升高达 3.2 倍。

Read more →


How Much Memory Does Your Agent Actually Need?

探讨 AI 代理在实际运行中所需的内存资源评估。

Read more →


Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers

介绍使用 Sentence Transformers 实现多向量(后期交互)嵌入模型。

Read more →


Same Cluster, 33 Points More Utilization: What Changed Was the Order

分享通过改变任务顺序,在同一集群中提升 33 点利用率的经验。

Read more →


State of Open Models: Summer 2026 Observations

2026 年夏季开放模型发展现状观察。

Read more →


Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets

介绍如何通过 Strands Agents、LeRobot 和 Hugging Face 存储桶实现一站式记录、训练和部署。

Read more →


What We Learned by Reproducing 2,200 papers from ICML

分享复现 2200 篇 ICML 论文后的心得体会。

Read more →


Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis

介绍 OlmoEarth 嵌入功能,支持从 OlmoEarth Studio 导出自定义嵌入以进行下游分析。

Read more →


Thinking of ACE? We Can Do It with Fewer Tokens

探讨如何以更少的 Token 实现 ACE(Agentic Contextual Execution)。

Read more →


The Gradient

After Orthogonality: Virtue-Ethical Agency and AI Alignment

探讨理性人与理性 AI 的目标设定,提出基于美德伦理的 AI 对齐路径。

Read more →


AGI Is Not Multimodal

文章认为,将语言作为思维模型会导致我们忽视具身智能(embodied understanding),AGI 不应仅仅是多模态的。

Read more →


Shape, Symmetries, and Structure: The Changing Role of Mathematics in Machine Learning Research

探讨机器学习研究中数学角色的转变,对比了数学原则架构与计算密集型工程方法的优劣。

Read more →


What’s Missing From LLM Chatbots: A Sense of Purpose

尽管 LLM 聊天机器人在基准测试中表现优异,但用户体验并未同步提升,文章指出它们缺乏“目的感”。

Read more →


We Need Positive Visions for AI Grounded in Wellbeing

呼吁建立以人类福祉为基础的 AI 积极愿景,反思 AI 对社会的深远影响。

Read more →


Financial Market Applications of LLMs

探讨 LLM 在金融市场中的应用,分析其在处理序列数据方面的潜力。

Read more →


A Brief Overview of Gender Bias in AI

简要概述并讨论 AI 系统中存在的性别偏见问题。

Read more →


Mamba Explained

解释 Mamba 模型,作为 Transformer 的替代方案,在处理长序列方面具有更高效率。

Read more →


Car-GPT: Could LLMs finally make self-driving cars happen?

探讨 LLM 在自动驾驶中的应用潜力,以及其在安全性与可靠性方面面临的挑战。

Read more →


Do text embeddings perfectly encode text?

介绍 ‘Vec2text’ 技术,能够将嵌入还原为文本,强调了嵌入数据安全协议的必要性。

Read more →


arXiv CS.AI

Active Inference as Context Acquisition for AI Agents

提出主动推理作为 AI 代理获取上下文的机制,优化代理在处理约束、偏好和任务变量时的决策效率。

Read more →


Robust Metaheuristics under Uncertainty for Berth Allocation and Quay Crane Assignment: A Review

综述了在不确定环境下,针对泊位分配和岸桥调度问题的鲁棒元启发式算法研究。

Read more →


How to Navigate Uncertainty About AI Consciousness

探讨在 AI 意识存在深层不确定性的情况下,人类应如何对待潜在的感知 AI。

Read more →


Bounded Sovereignty and the Control Tax: Pricing AI Oversight When the Deployer Does Not Own the Model

研究在部署者不拥有模型(如通过 API 使用)的情况下,如何对 AI 监督进行定价和控制。

Read more →


Interaction valence reveals contrasting social networks in dairy cattle

提出一种感知交互效价的社交网络框架,用于分析奶牛群体的行为交互。

Read more →


Air Traffic Control Using Large Language Models: Prompt Engineering, Architecture, and Evaluation

实验评估 LLM 在生成空中交通管制(ATC)指令方面的可行性,探讨其在安全关键对话中的应用。

Read more →


Outcome Monitors: Recovery Affordances for Silent Tool Failures

引入“结果监控器”(Outcome Monitors),用于检测 AI 代理在工具调用失败时的静默错误,并提供恢复机制。

Read more →


Beyond Imitation: Filtering On-Policy Distillation by Reasoning Progress

提出通过推理进度过滤策略蒸馏(OPD)的方法,改进语言模型的后训练过程。

Read more →


arXiv CS.CL

A Virtual Member of a Community of Practice for the Society of Petroleum Engineers: From Prototype to Deployment

描述了虚拟助手 ATHENA 的开发历程,旨在支持石油工程师协会成员的知识捕获与传播。

Read more →


Transformer Models for Text Summarization: A Comparative Study of BART, BERT, and RoBERTa

对比研究了 BART、BERT 和 RoBERTa 在自动文本摘要任务中的表现。

Read more →


Automatic bioinformatic software named entity recognition from literature

提出从科学文献中自动识别生物信息学软件命名实体的方法。

Read more →


Asymmetric Attention Heads: Structured Head-Wise Context Allocation for Transformer Attention

提出非对称注意力头机制,通过结构化分配上下文来优化 Transformer 的注意力机制。

Read more →


Hallucination as a Feature, not a Defect: Evaluating a multi-agent architecture to transform speculative language-model outputs into testable scientific hypotheses

提出将幻觉视为一种特征而非缺陷,利用多代理架构将 LLM 的推测性输出转化为可测试的科学假设。

Read more →


Compliance, Capability, and Conflict: Benchmarking Multimodal LLMs under System Messages

基准测试多模态 LLM 在系统消息约束下的合规性与能力表现。

Read more →


When Irrelevant Text Matters: Affine Margin Shifts in Multimodal Large Language Models

研究多模态 LLM 中无关文本上下文对视觉判断任务的影响。

Read more →


Represented but Ignored: A Causal Account of Prosodic Underuse in Audio-Language Models

从因果角度分析音频语言模型中韵律信息被忽视的原因。

Read more →


WIRED

Meta’s Big Reckoning Is Here

Meta 因儿童安全问题再次面临法庭审判,这可能成为迫使 Facebook 和 Instagram 核心功能发生重大变革的里程碑案件。

Read more →


I Tried the Best Robotic Pool Cleaners of 2026: Beatbot, iGarden, Dreame

评测了 2026 年市面上最好的机器人泳池清洁器,帮助用户解放双手,维护水质。

Read more →


Nyrius Phoenix Home True 4K60 (2026): A Solution for Cord Clutter

Nyrius Phoenix Home True 4K60 能够无线传输 4K 视频和游戏信号,有效解决家庭布线杂乱的问题。

Read more →


The Patrick Clancy Conspiracy Theories Are Rooted in the Harsh Realities of Motherhood

探讨了围绕 Lindsay Clancy 案件的阴谋论,分析了这些讨论如何反映了社会对母亲角色的严苛要求。

Read more →


5 Best Electric Toothbrushes (2026): Philips, Oral-B, Quip, More

经过两年的测试,评选出 2026 年最值得推荐的五款电动牙刷。

Read more →


Influencers and Resellers Are Turning Empty Boxes Into Big Cash

随着对“真实性”追求的增长,奢侈品空盒在二手市场被炒至高价,网红和转售商从中获利。

Read more →


The Super El Niño Won’t Fix the West’s Water Crisis

专家警告称,尽管超级厄尔尼诺现象带来降水,但无法从根本上解决美国西部的长期水资源危机。

Read more →


Best Early Tech Labor Day Sales I’d Shop Myself (2026): AirTags, Dyson, and More

整理了劳动节前的科技产品促销信息,包括 AirTags、戴森吸尘器等热门产品。

Read more →


The Single English County Saying No to Palantir

英国大曼彻斯特地区拒绝了 Palantir 的医疗保健合同,坚持自主管理医疗数据,引发了关于政府外包的讨论。

Read more →


Ray-Ban Promo Codes: Save 50% in August 2026

提供 2026 年 8 月雷朋眼镜的优惠代码,帮助用户以半价购买经典款及定制眼镜。

Read more →


Lobsters

Enabling the next-generation trait solver on nightly

Rust 社区讨论在 nightly 版本中启用下一代 trait 解析器。

Read more →


Better Batteries

探讨电池技术的进步及其对未来设备的影响。

Read more →


Btrfs Snapshot Integration in KDE

讨论 Btrfs 快照在 KDE 桌面环境中的集成方案。

Read more →


Music theory for programmers

为程序员编写的音乐理论指南,探讨编程与音乐的逻辑联系。

Read more →


Announcing Rust 1.98.0

Rust 1.98.0 版本正式发布,带来多项改进与功能更新。

Read more →


rust-glancer: An alternative LSP for Rust with focus on low memory usage

介绍 rust-glancer,一款专注于低内存占用的 Rust 语言服务器协议(LSP)替代方案。

Read more →


What are you doing this weekend?

社区互动贴,询问成员周末的计划,鼓励分享与交流。

Read more →


DEV Community

is-agentic Scored Promptway 74. Here Is What I Changed

作者使用 is-agentic 工具评估 Promptway 网站,得分 74 分,并分享了为优化 AI 代理访问而进行的改进。

Read more →


Building an Escalation Root-Cause Agent with Gemini and ADK

分享如何利用 Gemini 和 ADK 构建一个用于分析客户服务升级案例根本原因的 AI 代理。

Read more →


Anyone ever go back to their notetaker outputs and summaries?

探讨用户是否会回顾 AI 笔记工具生成的输出和摘要,引发关于笔记价值的讨论。

Read more →


Security news weekly round-up - 21st August 2026

每周安全新闻汇总,涵盖网络安全工具、意识提升及基础设施防护等内容。

Read more →


UNDERSTANDING THE GIT WORKFLOW

详细介绍 Git 版本控制系统的工作流,包括代码变更追踪、协作及仓库设置。

Read more →


How We Handle Client-Side CSV Merging Without Server Processing

分享如何在浏览器端处理 CSV 合并,解决列不匹配和单元格引用等复杂问题,无需服务器处理。

Read more →


OpenAI Rolls Out Flexible Codex Pricing for Business and Enterprise Teams

OpenAI 为企业客户推出灵活的 Codex 定价模式,支持按 Token 使用量付费,将 AI 编码访问与固定席位费分离。

Read more →


Building a Full Enterprise-Ready React + Spring Boot Auth Flow: An End-to-End Guide

提供构建企业级 React + Spring Boot 认证流程的完整指南,涵盖 Token 存储、CSRF 防护及刷新机制。

Read more →


OpenAI GPT-5.6 Pricing Update Cuts Terra and Luna Costs, Leaves Sol Unchanged

OpenAI 更新 GPT-5.6 模型定价,降低 Terra 和 Luna tiers 的成本,Sol 旗舰版价格保持不变。

Read more →


What I Learned Contributing to Prefect, dbt, and Airflow (An Honest OSS Retrospective)

作者分享参与 Prefect、dbt 和 Airflow 等开源项目贡献的经验,强调了协作与构建作品集的重要性。

Read more →


Meta Engineering

How We’re Building Scam Alert on WhatsApp With End-to-End Encryption and Verifiability Guarantees

WhatsApp 正在构建诈骗预警系统,在保持端到端加密的同时,通过可验证性保证用户安全。

Read more →


From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking

介绍 Meta 广告排序的多阶段架构,通过建模用户行为序列提升推荐效果。

Read more →


GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model

分享 Meta 如何通过优化训练效率,将生成式广告推荐模型(GEM)的训练效率提升一倍。

Read more →


Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

探讨分层兴趣表示技术,用于优化 Meta 广告的深层漏斗转化。

Read more →


Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler

Meta 通过引入开源内核调度器 sched_ext,定制了广告服务的调度策略,有效降低了延迟。

Read more →


Meta’s AI Storage Blueprint at Scale

分享 Meta 在大规模 AI 训练中对存储架构的蓝图设计,确保快速可靠的数据访问。

Read more →


10 Years of Meta’s Commitment to Python

庆祝 Meta 连续 10 年赞助 Python 软件基金会,强调 Python 在公司工程实践中的核心地位。

Read more →


Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study

探讨 AI 时代隐私感知基础设施的构建,通过资产分类案例研究实现数据合规。

Read more →


How Meta Engineered Ultra-Narrow Batteries for AI Glasses

分享 Meta 如何为 Ray-Ban Meta 等智能眼镜设计超窄电池,以满足 AI 工作负载的能源需求。

Read more →


DeepMind Blog

From Atari to EVE Online: Building on 15 Years of AI Research in Games

回顾 DeepMind 在游戏 AI 领域 15 年的研究历程,从 Atari 到 EVE Online 的合作原型。

Read more →


Introducing Gemini 3.7 Flash

介绍 Gemini 3.7 Flash 模型。

Read more →


Putting sign language AI into users’ hands

介绍手语转文本(SL2T)模型,为聋哑和听障用户提供新的手语功能。

Read more →


WeatherNext: AI model achieves breakthrough in forecasting cyclones

WeatherNext AI 模型在气旋预测方面取得突破。

Read more →


Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration

介绍 Gemini Robotics ER 2,提升机器人的视频理解、任务编排和多机器人协作能力。

Read more →


We’re launching Lyria 3.5 in Google Flow Music, with advances across musicality, lyrics, vocals, and creative control

推出 Lyria 3.5,在音乐性、歌词、人声和创作控制方面实现显著提升。

Read more →


Gemini Robotics 2 brings whole body intelligence to robots

介绍 Gemini Robotics 2,为机器人带来全身智能。

Read more →


Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

谷歌承诺投入 4000 万美元的 AI Token 和信用额度,支持 Genesis Mission 科学发现。

Read more →


Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

介绍 Gemini 系列新模型,包括 3.6 Flash、3.5 Flash-Lite 和 3.5 Flash Cyber。

Read more →


Introducing Gemini 3.5 Flash Cyber

介绍 Gemini 3.5 Flash Cyber,一款轻量级网络安全模型,用于发现和修补漏洞。

Read more →


VentureBeat AI

VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push

VentureBeat 任命 Rob Strechay 为首位首席分析师,旨在深化企业 AI 研究。

Read more →


Google just redesigned the search box for the first time in 25 years — here’s why it matters more than you think.

谷歌 25 年来首次重新设计搜索框,标志着搜索范式的重大转变。

Read more →


Railway secures $100 million to challenge AWS with AI-native cloud infrastructure

Railway 融资 1 亿美元,旨在通过 AI 原生云基础设施挑战 AWS。

Read more →


Claude Code costs up to $200 a month. Goose does the same thing for free.

Claude Code 价格昂贵,而 Goose 提供了类似的免费替代方案,引发开发者社区关注。

Read more →


Listen Labs raises $69M after viral billboard hiring stunt to scale AI customer interviews

Listen Labs 通过病毒式广告牌招聘活动融资 6900 万美元,用于扩展 AI 客户访谈业务。

Read more →


Salesforce rolls out new Slackbot AI agent as it battles Microsoft and Google in workplace AI

Salesforce 推出全新 Slackbot AI 代理,在办公 AI 领域与微软和谷歌展开竞争。

Read more →


Anthropic launches Cowork, a Claude Desktop agent that works in your files — no coding required

Anthropic 推出 Cowork,一款无需编码即可在本地文件中工作的 Claude 桌面代理。

Read more →


arXiv CS.LG

Towards On-Board Implementation of ML-Based Helicopter Weight Estimator

研究基于机器学习的直升机起飞重量估算模型,并探讨其在机载系统中的实现过程。

Read more →


Triangular Fuzzy Rescaling Distance

提出三角模糊重缩放距离度量方法,用于处理复杂系统中的不确定信息。

Read more →


Holtercare-Bench: A Multimodal Benchmark for Evaluating Long-Term Dynamic ECG Analysis

提出 Holtercare-Bench 基准测试,用于评估多模态模型在长期动态心电图分析中的表现。

Read more →


Quantum Kernel Estimation for the Discovery of Early Lung Cancer Detection

研究利用量子核估计技术发现早期肺癌检测的生物标志物。

Read more →


Improved Confidence Estimates for Black-Box Large Language Models

提出改进黑盒 LLM 置信度估计的方法,无需标签数据即可量化不确定性。

Read more →


Mechanistic Tomography: Designed Measurement for Control-Oriented Interpretability

提出“机械断层扫描”方法,通过设计测量手段实现面向控制的可解释性。

Read more →


Uncovering the Limits of Proof Sharing for Neural Networks

探讨神经网络证明共享的局限性,分析其在加速验证技术中的应用边界。

Read more →


Longitudinal Bayesian Learning of Continuous Disease Position across the Alzheimer’s Disease Continuum

提出疾病连续体定位(DCP)方法,利用纵向贝叶斯学习分析阿尔茨海默病的发展过程。

Read more →


arXiv CS.CV

Clustering and Token Denoising for Faster and More Robust VLMs

提出聚类和 Token 去噪方法,旨在提升视觉语言模型(VLM)的推理速度与鲁棒性。

Read more →


SceneGTMM: A Conformal Mapping-based Scene-Aware Transferable GNN-Transformer Dual-Graph Interaction Framework for Map Matching

提出 SceneGTMM 框架,结合共形映射与双

生成二维码中...

请点击右上角 ···

选择 发送给朋友收藏