今日焦点 · 本站原创 · 研究前沿

深度研究|递归自我改进(RSI)全景 2026:AI 正在加速 AI,但「验证瓶颈」决定它能走多远

梳理 RSI 从 Good 1965 到 2026 的思想史、技术图谱与一手实证:Anthropic 承认其代码库 >80% 合并代码由 Claude 撰写、METR 测得 AI 可完成任务时长约每 4 个月翻倍、AlphaEvolve 优化了支撑自身的计算栈。核心判断——有界自我精炼已工程化,开放式 RSI 尚未发生;而进步与安全共享同一个「验证瓶颈」。
来源:本站原创2026-09-14
阅读全文 →

最新文章

LATEST 共 514 条
SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents
SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering agents on challenging repository-l…
智能体 HuggingFace Daily Papers 9-8
GPT-6 让 48 个网页验证码失效了,最聪明的 AI 和最笨的人类相遇了
人类的验证码已经拦不住 AI #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
大模型 爱范儿 9-8
Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout
Autoregressive (AR) video diffusion models have shown great potential in real-time video generation. Recent methods dist…
行业动态 HuggingFace Daily Papers 9-8
The Price of Sparsity: Sufficient Conditions for Sparse Recovery using Sparse and Sparsified Measurements
We consider the problem of support recovery for sparse binary signals from noisy linear measurements. For sparse Gaussia…
行业动态 HuggingFace Daily Papers 9-8
TANGO: Humanoid Navigation in Cluttered Environments with a Whole-Body Vision-Language-Action Model
We study the problem of navigating cluttered indoor environments with a humanoid robot. Unlike conventional methods that…
智能体 HuggingFace Daily Papers 9-8
深度智控获宁德时代、沙特阿美战投等重磅加码,加速打造物理AI时代算力与能源底座
近日,物理AI企业深度智控(DeepCtrls)完成新一轮B+轮数亿元融资。
行业动态 量子位 9-8
现场围观金融AI决赛,大厂挑人的逻辑我悟了
百万奖金、大厂直通、VC跟投
行业动态 量子位 9-8
SynthGait-19K: A Physically Grounded Synthetic Video Dataset for Gait Parameter Estimation
Accurate estimation of clinically meaningful gait parameters from monocular video is important for scalable mobility ass…
行业动态 HuggingFace Daily Papers 9-8
深入马来西亚AI现场!WAIC CONNECT MALAYSIA首日亮点全速递
从看市场,到见场景;从认识伙伴,到寻找合作。
行业动态 量子位 9-8
SAEScientist-Bench: Can AI Agents Conduct Autonomous SAE Interpretability Research?
While research on recursive self-improvement (RSI) has predominantly automated model training pipelines, reliable autono…
智能体 HuggingFace Daily Papers 9-8
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation
Weak-to-strong generalization asks whether stronger models can learn from weaker supervisors and surpass them. This ques…
行业动态 HuggingFace Daily Papers 9-8
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models
Training capable cyber agents is often treated primarily as a problem of model scale, yet open-weight post-training is c…
智能体 HuggingFace Daily Papers 9-8
← 上一页 第 30 / 43 页 · 共 514 条 下一页 →