今日焦点 · 研究前沿

深度研究|递归自我改进(RSI)全景 2026:AI 正在加速 AI,但「验证瓶颈」决定它能走多远

梳理 RSI 从 Good 1965 到 2026 的思想史、技术图谱与一手实证:Anthropic 承认其代码库 >80% 合并代码由 Claude 撰写、METR 测得 AI 可完成任务时长约每 4 个月翻倍、AlphaEvolve 优化了支撑自身的计算栈。核心判断——有界自我精炼已工程化,开放式 RSI 尚未发生;而进步与安全共享同一个「验证瓶颈」。
来源:本站原创2026-09-14
阅读全文 →

研究前沿

共 68 条
CARDEA: Auditable Reasoning Grounded in Spatial Evidence for End-to-End Coronary Angiography Interpretation
Invasive coronary angiography (CAG) is the gold standard for diagnosing coronary artery disease, but interpretation vari…
研究前沿 HuggingFace Daily Papers 9-7
A*-Thought-V2: Efficient Latent Reasoning via Geometric Dynamics of LLM
Chain-of-Thought (CoT) improves the reasoning ability of Large Language Models (LLMs) but incurs substantial computation…
研究前沿 HuggingFace Daily Papers 9-7
LG 智能电视会在待机状态下扫描家庭网络和记录麦克风音频
根据 YouTube 主播 Gamers Nexus、Level1Techs 以及独立安全研究员合作展开的调查,测试了包括 G5 在内的零售 LG OLED 电视机,发现 LG 智能电视会在屏幕关闭但没有断电的待机状态下扫描家庭网络,寻找手…
研究前沿 Solidot 9-7
2026 年 Ig Nobel 宣布
从美国波士顿迁往瑞士苏黎世的 Ig Nobel 奖颁奖典礼宣布了 2026 年的获奖者。明年的颁奖典礼将在德国 Flanders 的 Antwerp 举行,2028 年重返瑞士,以后的偶数年颁奖典礼都在苏黎世举行。获奖名单包括: 生物力学奖…
研究前沿 Solidot 9-7
中国游戏市场规模在 2025 年首次突破 500 亿美元
根据 Niko Partners 的报告,中国游戏市场规模在 2025 年首次突破 500 亿美元达到 518 亿美元,2026 年预计将增长 4% 达到 539 亿美元,2030 年将达到 598 亿美元,到 2030 年中国游戏玩家将达…
研究前沿 Solidot 9-7
GPT-6不只Astra!Sol内测结果曝光,速度快6倍
OpenAI研究院人均带3个AI实习生
研究前沿 量子位 9-7
DF26: We Cannot Tell Fake From Real Anymore
We introduce DF26, a novel benchmark for detecting AI-generated videos containing fully synthetic clips produced by rece…
研究前沿 HuggingFace Daily Papers 9-7
ReactVAU: A Slow-Fast Decoupled Framework for Streaming Video Anomaly Understanding
In this paper, we propose ReactVAU, a Slow-Fast Decoupled Framework for real-time streaming Video Anomaly Understanding …
研究前沿 HuggingFace Daily Papers 9-7
Revisiting Complete Reasoning Traces for Post-Training
Large language models (LLMs) are often post-trained on pre-collected reasoning trajectories to improve their reasoning c…
研究前沿 HuggingFace Daily Papers 9-7
Reason Through the Latent! Making Latent Visual Reasoning Necessary
Latent visual reasoning aims to perform multimodal reasoning through hidden-state computation rather than explicit textu…
研究前沿 HuggingFace Daily Papers 9-6
VDiff-Bench: A Challenging Benchmark for Fine-Grained Image Difference Identification
Multimodal Large Language Models (MLLMs) perform strongly on general visual understanding tasks such as visual question …
研究前沿 HuggingFace Daily Papers 9-5
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data
Recent advances in wearable sensing enable continuous monitoring of physiological and behavioral signals, yet existing b…
研究前沿 HuggingFace Daily Papers 9-4
← 上一页 第 5 / 6 页 · 共 68 条 下一页 →