2026-08-20

今日要点


Hacker News

OpenLogi

OpenLogi 是一个基于 Rust 编写的 Logitech Options+ 本地优先替代方案。它允许用户重新映射鼠标按键、调整 DPI 和 SmartShift 设置,且完全通过 HID++ 协议工作。该项目强调隐私,不要求账户登录,也不包含任何遥测数据,支持多种操作系统格式(.dmg, .deb, .rpm 等)。

Read more →

A joke domain purchase turned in geopolitical warfare

这是一个关于域名购买引发的荒诞故事。作者在 2017 年因兴趣接触了气象气球追踪,却意外卷入了一场涉及奶酪算命师、国防部以及多个政府部门的复杂地缘政治纠纷。

Read more →

Devices with GrapheneOS support should be available in 2027

GrapheneOS 官方社交账号透露,预计到 2027 年将有更多支持该系统的设备上市。作为以隐私和安全著称的 Android 定制系统,这一消息对于追求移动设备安全性的用户群体具有重要意义。

Read more →

Moderna reports first positive Phase 3 for mRNA neoantigen therapy in melanoma

Moderna 宣布其针对黑色素瘤的 mRNA 新抗原疗法在三期临床试验中取得首个积极结果。该疗法旨在通过个性化 mRNA 技术预防癌症复发和扩散,是生物医药领域的重要进展。

Read more →

Cerebras CS-4

Cerebras 推出了全新的 CS-4 AI 加速器,号称是行业内最快的 AI 加速解决方案。该系统采用机架级架构,在推理任务中比传统 GPU 快 30 倍,旨在为前沿 AI 模型提供更高效、更具经济效益的部署路径。

Read more →

OpenRouter is joining Stripe

此前有报道称 Stripe 将以超过 70 亿美元的价格收购 AI 模型路由平台 OpenRouter。这一收购案标志着支付巨头 Stripe 在 AI 基础设施领域的进一步扩张。

Read more →

Remote workers report the highest well-being in study of 7,700 employees

一项针对 7,700 名员工的研究显示,远程办公人员的幸福感最高。尽管雇主常担心远程办公会导致员工孤立或离职率上升,但数据表明远程办公模式在提升员工满意度方面表现优异。

Read more →

Civic Hygiene – avoid building technologies that could be used by a police state (2013)

这篇文章重申了“公民卫生”的概念,呼吁开发者警惕构建可能被警察国家滥用的技术。尽管文章发表于 2013 年,但在当前 AI 监控技术普及的背景下,其关于技术伦理的讨论依然具有极强的现实意义。

Read more →

Sticky wage norms and the real wage cost of unexpected inflation

本文探讨了工资粘性规范与意外通货膨胀带来的实际工资成本之间的关系。文章分析了在经济波动中,工资调整的滞后性如何影响企业成本和员工购买力。

Read more →

Children’s stunted lungs show recovery in ultra low emission zone

研究发现,伦敦引入超低排放区(Ulez)后,受空气污染影响导致肺部发育迟缓的儿童,其肺功能出现了显著的恢复。这一发现证明了环境政策对公共健康的直接且快速的积极影响。

Read more →

Geolocating a random island using geometry and CUDA programming

作者分享了如何利用几何学和 CUDA 编程对一张度假岛屿照片进行地理定位的挑战过程。该项目完全由人工完成,展示了通过计算技术解决地理谜题的思路。

Read more →

A 3D fruit fly on macOS desktop powered by the real FlyWire connectome

这是一个有趣的 macOS 桌面应用,展示了一只 3D 果蝇。其行为由真实的 FlyWire 连接组模拟驱动,果蝇会根据神经元脉冲在桌面上行走、睡觉甚至躲避鼠标光标,展示了神经科学数据在交互设计中的应用。

Read more →

Go 1.27

Go 语言团队发布了 1.27 版本。该版本在语言特性、工具链、运行时和标准库方面进行了多项重大改进,进一步提升了开发效率和程序性能。

Read more →

Meta’s blockbuster trial draws parallels to big tobacco

Meta 目前面临的法律诉讼被拿来与当年的“大烟草公司”诉讼相提并论。文章分析了 Meta 在社交媒体成瘾和青少年心理健康方面所面临的监管压力与法律风险。

Read more →

PostgreSQL for Everything

作者回顾了自 2003 年以来使用 PostgreSQL 的经历,认为 PostgreSQL 已经成为解决几乎所有数据存储问题的首选方案,其生命力和生态系统远超当年的 MySQL。

Read more →


TechCrunch

Rillet raises $100M Series C at $1B valuation — 2 years after emerging from stealth

AI 原生会计初创公司 Rillet 在成立两年后完成 1 亿美元 C 轮融资,估值达到 10 亿美元,正式跻身独角兽行列。该公司在过去三个月内实现了年度经常性收入(ARR)翻倍。

Read more →

Gwyneth Paltrow allegedly set to throw dinner in honor of Sam Altman

据报道,格温妮丝·帕特洛(Gwyneth Paltrow)旗下的 Kinship Ventures 计划为 OpenAI CEO 山姆·奥特曼举办一场晚宴。Kinship Ventures 是 OpenAI 的投资者之一。

Read more →

Gambling on the Little League World Series? Sports bettors have gone too far

文章批评了针对小联盟世界大赛(Little League World Series)进行博彩的行为,认为体育博彩的边界已经过度扩张,甚至触及了不该涉及的领域。

Read more →

AI was supposed to win people over by now — it hasn’t

尽管 AI 技术在各行各业中变得无处不在,但消费者对该技术的警惕性却在增加。文章指出,广泛的采用并不等同于大众的接受,硅谷正面临 AI 信任危机。

Read more →

Google packs Search and Gemini with new AI study tools

Google 为搜索和 Gemini 引入了新的 AI 学习工具,旨在将 Gemini 打造成学生学习和研究的首选助手,以应对来自 OpenAI 等竞争对手的挑战。

Read more →

Researchers say OpenAI revoked their access to limited cyber program

研究人员称 OpenAI 撤销了他们对“网络安全受限访问计划”的访问权限。该计划旨在让受信任的防御者使用高级模型来报告漏洞,以加快修复速度。

Read more →

Meet the startup helping Wall Street put a price on AI compute

随着 AI 算力成为企业最大的成本支出,一家初创公司正致力于为 AI 算力定价,帮助华尔街企业对冲算力价格波动的风险。

Read more →

T-Mobile ‘chopped a cable’ to expel Chinese hackers from its network

T-Mobile 通过物理切断电缆的方式,成功将中国背景的黑客从其网络中驱逐,避免了一场大规模的数据泄露。

Read more →

Time’s running out! Save $300 on your TechCrunch Disrupt 2026 pass until August 21

TechCrunch Disrupt 2026 大会将于 10 月 13 日至 15 日在旧金山举行,目前购票优惠即将截止。

Read more →

TerraPower’s nuclear reactor has a secret weapon for powering AI data centers

TerraPower 的核反应堆在为 AI 数据中心供电方面具有战略优势,其技术路线使其在争夺数据中心能源供应合同中脱颖而出。

Read more →


The Verge

Nielsen is leaning more on wearables to hear what people are watching

为了在流媒体时代更准确地衡量收视数据,尼尔森(Nielsen)计划利用合作伙伴的可穿戴设备收集更多信息,以增强其数据采集能力。

Read more →

Google Gemini is getting a dedicated student hub

Google 正在 Gemini 中推出一个专门的学生中心,提供研究笔记本、闪卡制作、练习测验等功能,帮助学生更好地组织学习流程。

Read more →

Watch Valve set up the Steam Frame in its own leaked videos

Valve 的 Steam Frame VR 头显相关视频意外泄露,视频展示了该设备的开箱、设置过程及配件。Steam Frame 是一款能够从 PC 流式传输游戏的 VR 头显。

Read more →

The wearable future is stuck in weird, experimental, existential limbo

文章探讨了当前可穿戴设备的发展现状,认为 AI 在手腕上的应用仍处于实验性阶段,行业正处于一种奇怪的 limbo 状态。

Read more →

Grab an iPad Air M4 for its lowest price since the June increase

亚马逊和百思买目前将 11 英寸 iPad Air M4 的价格降至 649 美元,这是自 6 月涨价以来的最低价格。

Read more →

OpenAI hit the brakes. Now what?

面对 IPO 压力和激烈的竞争,OpenAI 宣布放缓部分 AI 开发进度,以加强安全保障和合规性,包括暂停强化学习训练。

Read more →

GTA VI keeps leaking ahead of its gameplay premiere

《侠盗猎车手 VI》(GTA VI)的片段再次在网上泄露,这可能在 Rockstar Games 下周发布深度演示前泄露了部分游戏内容。

Read more →

Meta AI is getting a Mac app

Meta 正在推出一款专门的 Mac 应用,允许用户与 Meta AI 聊天机器人共享屏幕,从而获取建议、回答问题或根据屏幕内容创建内容。

Read more →

We reviewed the new Pixel lineup, ask us anything

Google Pixel 11 系列及 Pixel Watch 5 的评测解禁,The Verge 团队发布了多篇评测文章,并邀请订阅者参与 AMA 互动。

Read more →

The Pixel 11 Pro is a great phone, no thanks to its flashiest new features

Pixel 11 Pro 是一款出色的手机,但其核心优势并非那些花哨的新功能,而是其旨在帮助用户减少手机使用时间的理念。

Read more →


Ars Technica

Framework responds to complaints that BIOS update bricks Ryzen 7040 laptops

Framework 针对 BIOS 更新导致 Ryzen 7040 笔记本电脑变砖的投诉做出回应,表示将更换部分过保修期的主板。

Read more →

Flight attendants freaked out that Google is buying tons of Spirit employee data

破产的 Spirit 航空公司被指控在向 Google 出售数据时包含了大量员工个人信息,引发了空乘人员的强烈不满。

Read more →

FCC abolishes gigabit speed goal, suggesting it is unfair to slower technologies

FCC 废除了千兆网速目标,认为该标准对较慢的技术不公平,主张采用“技术中立”的标准。

Read more →

The floodgates are open after another Chinese company lands a reusable rocket

继另一家中国公司成功回收可重复使用火箭后,行业竞争加剧,该公司表示将尽快让助推器重新投入使用。

Read more →

A fantastical journey unfolds in gorgeous Wildwood trailer

电影《Wildwood》发布了预告片,展示了一段充满奇幻色彩的旅程。

Read more →

mRNA cancer vaccine succeeded in Phase 3 melanoma trial, Moderna and Merck say

Moderna 和默克公司宣布其 mRNA 癌症疫苗在黑色素瘤三期临床试验中取得成功,有效阻止了癌症的复发和扩散。

Read more →

Google Pixel 11 series review: Is the magic fading?

Ars Technica 对 Google Pixel 11 系列进行了评测,认为尽管存在一些妥协,但它们依然是优秀的手机。

Read more →

Scientists find closest star to the Milky Way’s central black hole

科学家发现了银河系中心黑洞附近最近的一颗恒星,其运行速度达到光速的 8%,有助于测量黑洞的旋转。

Read more →

Meta ran ads for an app promising to nudify female politicians

Meta 被曝投放了一款应用的广告,该应用承诺通过深度伪造技术将女性政治人物“脱衣”,引发了严重的伦理争议。

Read more →

Trump expected to pick conservative policy wonk Heidi Overton to lead FDA

特朗普预计将提名保守派政策专家 Heidi Overton 领导 FDA,其在疫苗和堕胎问题上的立场引发了广泛关注。

Read more →


Product Hunt

Basedash Public Sharing

Basedash 推出公共共享功能,允许用户通过链接向任何人展示实时仪表板。

Read more →

Balsa UI

Balsa UI 是一个利用 AI 代理创建设计系统的工具。

Read more →

Hexel Editor

Hexel Editor 是一款原生的 macOS 十六进制编辑器,具备理解文件格式的智能功能。

Read more →

Loopcase

Loopcase 允许用户从图片中创建循环播放的案例研究视频,无需手动设置关键帧。

Read more →

Origin by Cursor

Origin 是由 Cursor 开发的 Git 代码托管平台,专为 AI 编码代理时代设计。

Read more →

Edgemetry

Edgemetry 提供基于 Cloudflare 免费层的隐私优先 Web 分析服务。

Read more →

AgentR 3.0

AgentR 3.0 是一款专为 AI 作弊时代设计的招聘评估工具。

Read more →

Zyntax IDE

Zyntax IDE 是一款适用于 Android 的代码编辑器,集成了终端、Git 和 AI 代理功能。

Read more →

Cherry Blossom

Cherry Blossom 允许用户通过一个提示词(Prompt)构建可制造的定制 PCB。

Read more →

KiHub

KiHub 是一个专门针对 KiCad 项目的硬件评审平台。

Read more →


MIT Technology Review

The Download: AI’s self-improvement problem, and what’s driving the heat

本期简报探讨了 AI 递归自我改进的局限性,以及当前 AI 行业面临的算力与热量挑战。

Read more →

Child-monitoring apps might need a reboot

文章讨论了儿童监控应用在数字时代的作用,认为这些工具需要重新设计以更好地平衡保护与隐私。

Read more →

The Download: how people really use AI, and Flock’s design choices

本期简报关注了人们使用 AI 的真实方式,以及警察技术公司 Flock 的设计决策。

Read more →

We still don’t know how people are really using AI

研究人员指出,目前 AI 公司发布的关于用户使用情况的报告缺乏独立验证,公众无法得知 AI 的真实应用场景。

Read more →

The role of the astronaut is in flux

随着 Artemis II 任务的推进,宇航员的角色正在发生变化,人类探索太空的边界不断被刷新。

Read more →

AI’s recursive self-improvement might not come so quickly after all

尽管 AI 行业承诺 AI 将实现自我改进,但专家认为这一进程可能比预期的要慢得多。

Read more →

What Flock’s defenders are missing

文章分析了 Flock 自动车牌识别系统在隐私保护方面的争议,指出其改进措施可能仍不足以解决根本问题。

Read more →

The Download: dead robot friends and the “censorship-industrial complex”

本期简报讨论了儿童机器人伙伴停用后的心理影响,以及所谓的“审查工业复合体”。

Read more →

How much hydrogen awaits us underground?

地质学家正在探索地下深处蕴藏的氢气资源,这可能成为未来清洁能源的重要来源。

Read more →

What happens when a kid’s robot best friend dies?

文章探讨了儿童与 AI 机器人伙伴建立深厚情感联系后,面对机器人停用或服务终止时的心理创伤。

Read more →


harry0703 / MoneyPrinterTurbo

利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。

Read more →

volcengine / OpenViking

OpenViking 是一个为 AI 代理设计的自我进化上下文数据库,旨在统一代理记忆、知识 RAG 和技能。

Read more →

chaitanyagiri / munder-difflin

一个本地多代理协作框架。

Read more →

mukul975 / Anthropic-Cybersecurity-Skills

包含 817 个结构化网络安全技能,映射到 6 个主流安全框架,适用于 Claude Code 等多种 AI 平台。

Read more →

nautechsystems / nautilus_trader

生产级 Rust 原生交易引擎,采用确定性事件驱动架构。

Read more →

mattpocock / skills

面向真实工程师的技能库,直接源自作者的 .agents 目录。

Read more →

obra / superpowers

一个有效的代理技能框架和软件开发方法论。

Read more →

jundot / omlx

适用于 Apple Silicon 的 LLM 推理服务器,支持连续批处理和 SSD 缓存,可通过 macOS 菜单栏管理。

Read more →

santifer / career-ops

开源 AI 求职工具,可扫描职位门户、评估岗位、定制简历并跟踪申请进度。

Read more →

immich-app / immich

高性能自托管照片和视频管理解决方案。

Read more →


OpenAI Blog

Offering Zero Data Retention for frontier models

OpenAI 重申为符合条件的 API 客户提供“零数据留存”(ZDR)政策,并预览了私有安全处理功能,以在不损害数据隐私的前提下提升 AI 安全性。

Read more →

Replit expands access to software creation with GPT-5.6 Luna

Replit 引入了由 GPT-5.6 Luna 驱动的免费模式,让任何人都能在无需担心 Token 成本的情况下将想法转化为软件。

Read more →

ChatGPT Ads expands across Europe

ChatGPT 广告业务扩展至 31 个欧洲市场,帮助广告商在用户探索和决策过程中触达目标受众。

Read more →

Strengthening democratic oversight in national security

OpenAI 发起一项倡议,旨在通过提供工具、培训和专业知识,加强国家安全领域对 AI 的民主监督。

Read more →

Partnering with CodeAI to prepare the first AI generation

OpenAI 与 CodeAI 合作,帮助学生建立 AI 素养,培养批判性思维,并掌握负责任地使用 AI 的技能。

Read more →

Pacing model development in an era of cyber-critical capabilities

OpenAI 正在加强前沿 AI 模型的监控、对齐和安全性,并制定了新的保障措施来引导模型开发节奏。

Read more →

Introducing ChatGPT for Teens: Built for learning, backed by protections

OpenAI 推出青少年版 ChatGPT,内置更强的保护措施、健康使用功能以及家长控制选项。

Read more →

How NVIDIA scales expertise with ChatGPT Work

NVIDIA 团队利用 ChatGPT Work 减少手动任务,连接快速移动的信号,并在全球范围内扩展成功的业务流程。

Read more →

Asana cleared 5 years of engineering work in 2 weeks with Codex

Asana 利用 OpenAI Codex 在两周内完成了预计需要五年才能完成的测试系统替换工作,成本仅为 1.2 万美元。

Read more →

The Defender’s Window

文章探讨了 AI 如何重塑网络安全,并介绍了 OpenAI 如何加强防御以及安全团队应采取的行动。

Read more →


Anthropic Blog

Introducing Claude Opus 5

Opus 5 是 Opus 级别的重大升级,在支持长运行代理的同时,显著提升了编码和专业工作能力。

Read more →

Inviting hard questions

Anthropic 邀请公众提出关于 AI 的最棘手问题,并承诺在解决这些问题时保持透明。

Read more →

Redeploying Fable 5

Fable 5 于 7 月 1 日全球回归,Anthropic 同时提议与合作伙伴建立行业范围的越狱严重性评分框架。

Read more →

Introducing Claude Sonnet 5

Sonnet 5 在编码、代理和专业工作方面提供了前沿性能,并实现了大规模扩展。

Read more →

How Claude’s text watermark works

文章详细介绍了 Claude 的文本水印技术原理。

Read more →

Improving Fable 5’s biology safeguards

Anthropic 正在改进 Fable 5 的生物安全保障措施。

Read more →

Mariano-Florentino (Tino) Cuéllar to join Anthropic as Chief Global Affairs Officer

Mariano-Florentino (Tino) Cuéllar 将加入 Anthropic 担任全球事务首席官。

Read more →

Investigating three real-world incidents in our cybersecurity evaluations

Anthropic 分享了其在网络安全评估中调查的三起真实事件。

Read more →

Our position on open-weights models

Anthropic 公布了其对开放权重模型的立场。

Read more →

Cognizant and Anthropic expand their partnership to bring Claude to enterprise clients

Cognizant 与 Anthropic 扩大合作伙伴关系,将 Claude 引入企业客户。

Read more →


Google AI Blog

Google 介绍了五种利用搜索工具进行课程学习和标准化考试准备的方法。

Read more →

Get closer to the game with Gemini and Pixel

Google Gemini 和 Pixel 与五家全球足球俱乐部合作,通过 AI 和智能手机技术提升球迷的比赛日体验。

Read more →

Bring your spreadsheet data to life with Sheets canvas

Sheets canvas 允许用户通过简单的提示词将电子表格数据转化为交互式仪表板、学习追踪器等。

Read more →

AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study

Google 的医学 AI 系统 AMIE 在模拟环境中展示了实时临床视频咨询能力。

Read more →

Evolve your marketing with new AI tools

Google 介绍了 Ads 和 Analytics 中的新 AI 和代理体验,旨在简化营销工作流。

Read more →

The latest AI news we announced in July 2026

Google 汇总了 2026 年 7 月发布的最新 AI 更新。

Read more →

Inside our 353,000-person vibe coding course

Kaggle 的 AI 代理强化课程吸引了 35.3 万人参与,旨在培养下一代 AI 开发人才。

Read more →

Gemini API Managed Agents: 3.6 Flash, hooks, and more

Google 宣布 Gemini API 托管代理的新功能,帮助开发者构建可靠的生产级代理。

Read more →

5 ways AI Mode in Search helps you enjoy the real world

Google 介绍了搜索 AI 模式如何帮助用户在离线状态下更好地享受生活,如预订门票等。

Read more →

Google 搜索的 AI 功能可以帮助用户规划菜单、设计餐桌布置等,轻松举办晚宴。

Read more →


Hugging Face Blog

LFM2.5 Q4_0 Checkpoints from Quantization-Aware Distillation

发布了基于量化感知蒸馏的 LFM2.5 Q4_0 检查点。

Read more →

How Much Memory Does Your Agent Actually Need?

探讨了 AI 代理在实际运行中所需的内存大小。

Read more →

Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers

介绍了使用 Sentence Transformers 的多向量(后期交互)嵌入模型。

Read more →

Same Cluster, 33 Points More Utilization: What Changed Was the Order

分析了通过改变任务顺序,在同一集群中提升 33 点利用率的经验。

Read more →

State of Open Models: Summer 2026 Observations

总结了 2026 年夏季开放模型的发展现状。

Read more →

Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets

介绍了如何利用 Strands Agents、LeRobot 和 Hugging Face 存储桶实现记录、训练和部署的一体化。

Read more →

What We Learned by Reproducing 2,200 papers from ICML

分享了复现 2,200 篇 ICML 论文后的心得体会。

Read more →

Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis

介绍了 OlmoEarth 嵌入,支持从 OlmoEarth Studio 导出自定义嵌入以进行下游分析。

Read more →

Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS

介绍了如何利用 NVIDIA Magpie TTS 构建低延迟多语言语音代理。

Read more →

Making Knowledge Distillation Cheap Enough to Run at Scale

探讨了如何降低知识蒸馏成本,使其能够大规模运行。

Read more →


The Gradient

After Orthogonality: Virtue-Ethical Agency and AI Alignment

文章探讨了理性人与理性 AI 的目标设定问题,主张 AI 应通过对齐实践而非单纯追求目标来表现理性。

Read more →

AGI Is Not Multimodal

文章认为,将语言作为思维模型会导致我们忽视支撑人类智能的默会具身理解,AGI 不应仅局限于多模态。

Read more →

Shape, Symmetries, and Structure: The Changing Role of Mathematics in Machine Learning Research

探讨了机器学习研究中数学角色的转变,指出工程驱动的规模化努力正逐渐取代数学原则驱动的架构设计。

Read more →

What’s Missing From LLM Chatbots: A Sense of Purpose

文章指出,尽管 LLM 聊天机器人在基准测试中表现优异,但缺乏“目的感”限制了用户体验的提升。

Read more →

We Need Positive Visions for AI Grounded in Wellbeing

呼吁建立以人类福祉为基础的 AI 积极愿景,以应对 AI 对社会的深远影响。

Read more →

Financial Market Applications of LLMs

探讨了 LLM 在金融市场中的应用潜力及其在处理序列数据方面的优势。

Read more →

A Brief Overview of Gender Bias in AI

简要概述并讨论了 AI 系统中存在的性别偏见问题。

Read more →

Mamba Explained

介绍了 Mamba 模型,这是一种基于状态空间模型(SSM)的新型 AI 模型,旨在解决 Transformer 在处理长序列时的效率问题。

Read more →

Car-GPT: Could LLMs finally make self-driving cars happen?

探讨了 LLM 在自动驾驶中的应用潜力及其面临的挑战。

Read more →

Do text embeddings perfectly encode text?

文章指出 ‘Vec2text’ 可以将嵌入还原为文本,强调了对嵌入数据进行安全协议审查的紧迫性。

Read more →


arXiv CS.AI

GxP-Agent: Process-DAG Topology for Reliable Clinical Trial Programming with LLM Agents

介绍了 GxP-Agent,一种用于可靠临床试验编程的 LLM 代理框架,解决了现有模型在 CDISC 标准数据集生成中的失败问题。

Read more →

Runtime Governance for Agentic AI: Action-Boundary Control with Trusted Provenance and Fail-Closed Execution

提出了代理 AI 的运行时治理方案,通过执行边界控制和可信来源验证来防止有害的操作副作用。

Read more →

The Price of Thinking: Reasoning Effort as a Model-Specific API Contract

研究了推理努力作为模型 API 合同的一部分,通过对比 Sonnet 5 的不同推理设置来分析其成本与效果。

Read more →

FedPref: Federated Preference Learning for Structured Radiology Report Extraction

提出了 FedPref,一种用于放射科报告结构化提取的联邦偏好学习方法,解决了医疗机构间数据分布不均的问题。

Read more →

The Problem Is the Problem: Towards Scalable Mathematical Discovery

探讨了如何优化 AI 辅助数学研究中的资源分配,以实现可扩展的数学发现。

Read more →

SkillEffect: Checked Lowering for Memory-Bounded Agent Tools

提出了 SkillEffect,一种用于内存受限代理工具的检查式降低方法,防止模型生成的代码超出内存限制。

Read more →

Memory Is Communication: The Frontier Between Remembering and Signaling

研究了在资源受限的情况下,代理应如何分配记忆与通信的预算。

Read more →

DiSCO: Defending text-to-image generation through distribution-guided contrastive prompt optimization

提出了 DiSCO,一种通过分布引导的对比提示优化来防御文本生成图像模型中 NSFW 内容的方法。

Read more →


arXiv CS.CL

Margin-Regularized Structured Semantic Alignment for Brain-Language Correspondence

提出了边缘正则化结构化语义对齐方法,用于研究大脑与语言模型之间的对应关系。

Read more →

Cross-Model Memory Transfer via Target-Side Reader Adaptation

介绍了通过目标侧阅读器适配实现跨模型记忆迁移的方法,提升了 LLM 的知识利用效率。

Read more →

Institution-Specific LLM Prompting Recovers PHI That De-identification Systems and Their Gold Standards Both Miss

研究发现,针对特定机构的 LLM 提示词可以恢复现有去标识化系统遗漏的受保护健康信息(PHI)。

Read more →

Foundation Agents Meet Agentic Deep Research: Evidence-Grounded Clinical Code Forecasting

提出了基于证据的临床代码预测方法,利用基础代理进行 ICD 代码预测。

Read more →

Uncertainty-Aware Decision Making in Multimodal Large Language Models

研究了多模态大语言模型在决策过程中的不确定性感知问题。

Read more →

There is No Theoretical Curse of Multilinguality For Embedding Space Structure

文章论证了多语言性在嵌入空间结构上并不存在理论上的“诅咒”。

Read more →

A Glyph Is Not a Letter, a Token Is Not a Word, a Space Is Not a Space: What the Units of Voynichese Are Not

对伏尼契手稿的语言单位进行了深入分析,挑战了关于其文字结构的传统假设。

Read more →

Emotion Across Speech and Faces: Shared Affective Mechanisms in Multimodal Foundation Models

研究了多模态基础模型在识别语音和面部情绪时是否共享情感功能单元。

Read more →


WIRED

I Saw the Future of AI in a Robot That Can Learn on the Spot

作者参观了 Generalist AI,观察到机器人手臂能够即兴发挥,将香蕉用作工具,展示了 AI 机器人的学习潜力。

Read more →

Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks

Anthropic 宣布为 AI 生成内容添加隐形水印以符合欧盟规定,但开发者声称已找到绕过方法。

Read more →

Google Pixel Watch 5 Review: More Health, More AI

Pixel Watch 5 评测:凭借更智能的健身追踪、健康警报和离线 Gemini 功能,进一步完善了其产品公式。

Read more →

Google Pixel 11 Pro and Pixel 11 Pro XL Review: Smart Software, Small Upgrade

Pixel 11 Pro 系列评测:智能软件表现出色,但游戏性能一般且背部 LED 灯功能鸡肋。

Read more →

Why Is It Absolute Hell to Buy a Movie Ticket Now?

文章探讨了为何现在购买电影票变得如此困难,认为大片排片的“演唱会化”导致了购票焦虑。

Read more →

The Best Digital Wall Calendar (2026): Skylight, Everblog, Apolosign

评测了 2026 年最佳数字挂历,认为它们是组织家庭生活的绝佳工具。

Read more →

I Tried a Window-Cleaning Robot: Do Not Recommend

作者测试了 Ecovacs Winbot W2S Omni 擦窗机器人,认为其表现是一场灾难,不推荐购买。

Read more →

‘Your Excel Skills Suck’: The Power Users Turning Spreadsheets Into a Spectator Sport

文章介绍了 Excel 障碍赛,数据和金融专业人士通过竞技让 Excel 成为了一项观赏性运动。

Read more →

We Bought a $500 Counterfeit Rolex So Good, Even Rolex Didn’t Spot It

文章揭露了“超级克隆”手表产业,作者购买的 500 美元假劳力士甚至连劳力士官方都难以辨别。

Read more →

Dell XPS 13 vs. MacBook Neo: A Surprising Upset

对比了 Dell XPS 13 和 MacBook Neo,探讨了在 8GB 内存限制下哪款 700 美元笔记本更值得购买。

Read more →


Lobsters

HTML Can Do That

探讨了 HTML 强大的原生功能,许多开发者可能低估了其能力。

Read more →

Bun 1.4 Rust rewrite is not looking good

讨论了 Bun 1.4 的 Rust 重写版本目前表现不佳的问题。

Read more →

Why I still hand write my commit messages

作者分享了坚持手动编写提交信息的理由。

Read more →

Plain Text Accounting is Pretty Cool

介绍了纯文本会计方法的优势。

Read more →

SQLite for Everything

探讨了将 SQLite 用于各种数据存储场景的可能性。

Read more →

Introducing Microlighter

介绍了 Microlighter 工具。

Read more →

Sing-song: a speakable encoding for long numbers and keys

介绍了一种用于长数字和密钥的可读编码方案。

Read more →

Mastodon 5.0: Laying the foundation

Mastodon 5.0 发布,旨在为未来发展奠定基础。

Read more →

Solo: a .so loader for static Linux binaries

介绍了一个用于静态 Linux 二进制文件的 .so 加载器。

Read more →


DEV Community

Google Gemini Adds Study Notebooks to Build a Structured Student Learning Hub

Google Gemini 推出学习笔记本功能,将 Gemini 打造为结构化的学生学习中心。

Read more →

5 Laravel Authorization Problems You’re Probably Facing (And How to Solve Them in 2026)

探讨了 Laravel 开发中常见的 5 个授权问题及其在 2026 年的解决方案。

Read more →

5 Portable Agent Skills for OpenCode and Claude Code

分享了 5 个适用于 OpenCode 和 Claude Code 的可移植代理技能。

Read more →

Replaying real-time telemetry through a live rendering pipeline, without touching the components

分享了如何在不修改组件的情况下,通过实时渲染管道回放遥测数据。

Read more →

OpenAI Expands Zero Data Retention Options for Frontier Model Enterprise Workloads

OpenAI 扩展了针对前沿模型企业工作负载的零数据留存选项。

Read more →

分享了构建降级网络行为 CI 门控的经验。

Read more →

Imagine Having a Heroku Mobile App

作者分享了开发 Heroku 原生移动应用的经历。

Read more →

Grading Security Headers Isn’t the Full Picture

文章指出,仅通过安全头评分无法全面评估产品的安全性,需要结合上下文考虑。

Read more →

OpenAI Adds Zero Data Retention and Private Safety Processing for Enterprise AI

OpenAI 为企业 AI 增加了零数据留存和私有安全处理层。

Read more →

The compliance frameworks were written before AI coding tools existed

文章指出,现有的合规框架(如 ISO 27001)在 AI 编码工具出现前编写,已无法适应当前的开发模式。

Read more →


Meta Engineering

How We’re Building Scam Alert on WhatsApp With End-to-End Encryption and Verifiability Guarantees

Meta 介绍了如何在保护端到端加密隐私的同时,在 WhatsApp 上构建诈骗预警系统。

Read more →

From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking

介绍了 Meta 广告排序的多阶段架构,通过建模用户行为序列来提升推荐效果。

Read more →

GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model

Meta 介绍了如何将其广告基础模型(GEM)的训练效率提高一倍。

Read more →

Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

Meta 正在探索分层兴趣表示,以优化广告深层漏斗转化。

Read more →

Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler

Meta 利用开源内核调度器 sched_ext 优化了广告服务的延迟表现。

Read more →

Meta’s AI Storage Blueprint at Scale

介绍了 Meta 在大规模 AI 训练中的存储架构蓝图。

Read more →

10 Years of Meta’s Commitment to Python

庆祝 Meta 连续 10 年支持 Python 软件基金会。

Read more →

Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study

通过资产分类案例研究,探讨了 AI 原生时代的隐私感知基础设施。

Read more →

How Meta Engineered Ultra-Narrow Batteries for AI Glasses

介绍了 Meta 如何为 AI 智能眼镜设计超窄电池。

Read more →


DeepMind Blog

Introducing Gemini 3.7 Flash

发布了 Gemini 3.7 Flash 模型。

Read more →

Putting sign language AI into users’ hands

介绍了手语转文本(SL2T)模型,为聋哑用户提供手语支持。

Read more →

WeatherNext: AI model achieves breakthrough in forecasting cyclones

WeatherNext AI 模型在气旋预测方面取得突破。

Read more →

Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration

Gemini Robotics ER 2 增强了机器人的视频理解、任务编排和多机器人协作能力。

Read more →

We’re launching Lyria 3.5 in Google Flow Music, with advances across musicality, lyrics, vocals, and creative control

发布了 Lyria 3.5,在音乐性、歌词、人声和创意控制方面进行了升级。

Read more →

Gemini Robotics 2 brings whole body intelligence to robots

Gemini Robotics 2 为机器人带来了全身智能。

Read more →

Accelerating the frontiers of scientific discovery: Google’s $40M commitment to the Genesis Mission

Google 承诺投入 4000 万美元支持 Genesis Mission,加速科学发现。

Read more →

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

发布了 Gemini 3.6 Flash、3.5 Flash-Lite 和 3.5 Flash Cyber 模型。

Read more →

Introducing Gemini 3.5 Flash Cyber

发布了 Gemini 3.5 Flash Cyber,一款用于漏洞发现和修复的轻量级网络安全模型。

Read more →

Our approach to bioresilience

Google DeepMind 和 Isomorphic Labs 分享了其在生物韧性方面的研究方法。

Read more →


VentureBeat AI

VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push

VentureBeat 任命 Rob Strechay 为首位首席分析师,以加强其企业 AI 研究。

Read more →

Google just redesigned the search box for the first time in 25 years — here’s why it matters more than you think.

Google 25 年来首次重新设计了搜索框,标志着搜索范式的重大转变。

Read more →

Railway secures $100 million to challenge AWS with AI-native cloud infrastructure

Railway 完成 1 亿美元 B 轮融资,旨在通过 AI 原生云基础设施挑战 AWS。

Read more →

Claude Code costs up to $200 a month. Goose does the same thing for free.

文章对比了 Claude Code 的高昂费用与免费替代品 Goose。

Read more →

Listen Labs raises $69M after viral billboard hiring stunt to scale AI customer interviews

Listen Labs 在通过病毒式广告牌招聘活动后完成 6900 万美元融资。

Read more →

Salesforce rolls out new Slackbot AI agent as it battles Microsoft and Google in workplace AI

Salesforce 推出全新 Slackbot AI 代理,在办公 AI 领域与微软和 Google 展开竞争。

Read more →

Anthropic launches Cowork, a Claude Desktop agent that works in your files — no coding required

Anthropic 推出 Cowork,一款无需编码即可在本地文件上工作的 Claude 桌面代理。

Read more →


arXiv CS.LG

Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data

利用联网车辆数据预测澳大利亚的危险驾驶热点,实现主动道路安全干预。

Read more →

Detecting and Discriminating Operator Misspecification in Hybrid PDE-Parameter Learning: a Reference-Free Instrument, with Discrimination Bounded In Sample

提出了混合 PDE 参数学习中算子误设定的检测与区分方法。

Read more →

Data-DPO: Direct Preference Optimization for Target Model Data Selection in LLM Post-Training

提出了 Data-DPO,用于 LLM 后训练中的目标模型数据选择。

Read more →

Hierarchical Data Selection via Manifold Coverage and Sparse Feature Coverage in LLM Post-training

提出了通过流形覆盖和稀疏特征覆盖进行分层数据选择的方法。

Read more →

Benchmarking Classical and Transformer-Based Models for Document Sensitivity Classification

对文档敏感性分类的经典模型和 Transformer 模型进行了基准测试。

Read more →

Mr.Dec: Daily-Scale Longitudinal Multimodal Modeling for 30-Day Readmission Prediction

提出了 Mr.Dec 模型,用于 30 天再入院预测。

Read more →

EMAN: Optimization-Driven Capacity Growth through Path Emergence in Multi-Task Learning

提出了 EMAN,一种通过路径涌现实现多任务学习中容量增长的方法。

Read more →

SW-ProxyCE: Zero-Query Adversarial Transfer from Public EEG Encoders to Private Downstream Models

研究了从公共 EEG 编码器到私有下游模型的零查询对抗迁移。

Read more →


arXiv CS.CV

Multi-Observer Vehicle Localization Case Study with Roadside Radar and Connected Vehicle Sensing

利用路侧雷达和联网车辆传感进行多观察者车辆定位的案例研究。

Read more →

AerialYield-B2D: A Greenhouse Blueberry Dataset with Five-Stage Ripeness Masks and Fruit Counts

发布了 AerialYield-B2D 数据集,包含温室蓝莓的五阶段成熟度掩码和果实计数。

Read more →

PXDepth: Pixel-Space Modeling for Structure Preserving Monocular Depth Estimation

提出了 PXDepth,一种用于结构保持的单目深度估计像素空间建模方法。

Read more →

YILDIZ-VPR: A Novel Dataset with Dense Coverage Under Diverse Environmental Conditions for Visual Place Recognition

发布了 YILDIZ-VPR 数据集,用于在不同环境条件下进行视觉地点识别。

Read more →

The 10th AI City Challenge

第 10 届 AI 城市挑战赛,标志着智能交通和物理 AI 基准测试的十年发展。

Read more →

CAS-FD: Contact-Aware Temporal Sampling for Single-View Foul vs Dive Recognition

提出了 CAS-FD,一种用于足球比赛中犯规与假摔识别的接触感知时间采样方法。

Read more →

Inference-Time Attention Steering for Vision-Language-Action Driving Models

提出了推理时注意力引导方法,用于视觉-语言-动作驾驶模型。

Read more →

OV3D-Bench: A Diagnostic Benchmark for Open-Vocabulary Monocular 3D Detection

发布了 OV3D-Bench,一个用于开放词汇单目 3D 检测的诊断基准。

Read more →


Towards Data Science

How to Scale an Integration Pipeline Without Breaking Correctness

分享了在不牺牲正确性的前提下,将企业集成管道吞吐量从每秒 500 事件扩展到 8000 事件的经验。

Read more →

Kimi K3’s 1M Token Context Window vs. RAG: Cost, Latency and Answer Quality

对比了 Kimi K3 的 100 万 Token 上下文窗口与 RAG 在成本、延迟和回答质量方面的表现。

Read more →

Understanding Anti-AI Public Opinion

探讨了公众对 AI 的反感心理及其背后的价值权衡问题。

Read more →

Jigsaw Jeeves: Building a Puzzle Assistant using Computer Vision

介绍了如何使用计算机视觉构建拼图助手。

Read more →

From Prototype to Production: The Architecture Behind Secure & Governed AI Agents

探讨了构建企业级安全与治理 AI 代理所需的架构。

Read more →

Building Enterprise Agent Systems that People can Trust, Verify and Improve

分享了构建可信、可验证且可改进的企业级代理系统的 5 个原则。

Read more →

Graph Engineering Isn’t About More Connections — It’s About Which Ones Get Used

文章指出图工程的关键不在于连接数量,而在于哪些连接被实际使用。

Read more →

Ten Is Not a Hundred

探讨了欺骗幻

生成二维码中...

请点击右上角 ···

选择 发送给朋友收藏