搜索:VLA

共命中 49 条(服务端检索)
VLA(视觉-语言-动作)模型进化史:机器人如何「看懂再动手」
从 RT-2 把机器人动作当文本 token 输出,到 OpenVLA 开源 7B、π0 用流匹配跑到 50Hz、GR00T 与 Helix 转向双系统架构——本文梳理 VLA 模型三年的演进脉络:动作怎么表示、参数怎么变小、控制频率怎么上去,以及开源与闭源两条路线的分野。
原创 研究前沿 精选 · 原创 · 今天 阅读 0·访客 0
Think Like a World Model, Act Like a VLA: Distilling World-Model Representations into Compact Robot Policies
Vision-Language-Action (VLA) models map observations to actions with no objective that accounts for how the world respon…
智能体 HuggingFace Daily Papers · 9-21 阅读 9·访客 9
VLA-Precision: Asymmetric Co-Bootstrapping for Efficient Real-World Online RL of Vision-Language-Action Models
Pretrained vision-language-action (VLA) models enable broad manipulation but remain unreliable in tasks demanding precis…
行业动态 HuggingFace Daily Papers · 9-18 阅读 4·访客 4
RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks?
Vision-Language-Action (VLA) models have shown promising progress in language-conditioned robotic manipulation. However,…
智能体 HuggingFace Daily Papers · 9-4 阅读 9·访客 8
EmbodiedSkills: A Unified Framework for Orchestrating, Training, and Deploying VLA Agents
Vision-language-action (VLA) models map visual observations and language instructions directly to robot actions, but lon…
智能体 HuggingFace Daily Papers · 9-1 阅读 11·访客 8
Diffusion Policy:机器人动作生成为什么弃用回归、改用扩散模型
模仿学习的老问题是「多峰动作分布」——两种都对的做法被回归平均成一种错的。Diffusion Policy 用条件去噪扩散直接建模动作分布,在 12 个任务、4 个基准上平均成功率提升 46.9%,此后 DP3、DPPO 相继跟进,π0 的流匹配与 GR00T 的扩散 Transformer 把它推成了 VLA 时代的标配动作头。
原创 研究前沿 精选 · 原创 · 今天 阅读 0·访客 0
HuRo: Robotizing Human Videos for Scalable VLA Pretraining
Human video datasets offer an abundant and diverse source of interaction data that can complement expensive real-robot d…
智能体 HuggingFace Daily Papers · 9-18 阅读 7·访客 7
机器人学习的数据瓶颈:从 Open X-Embodiment 到遥操作数据工厂
大语言模型吃的是万亿 token 网络文本,机器人能吃的真机数据却以「万条轨迹」计。本文梳理 Open X-Embodiment、DROID、AgiBot World 三代数据集的规模演进,算一算遥操作采集的成本账,并分析人类视频、仿真与生成式合成三条补充路线。
原创 研究前沿 精选 · 原创 · 今天 阅读 0·访客 0
Recursive Harness Distillation across Agents for Robot Manipulation
A central goal in robotics is to enable manipulation across changing tasks and environments. Vision-language-action (VLA…
智能体 HuggingFace Daily Papers · 9-27 阅读 4·访客 4
CARE: Experience-Guided Atomic Corrective Execution for Vision-Language-Action Policies
Vision-Language-Action (VLA) policies achieve strong performance in robotic manipulation but remain brittle once executi…
智能体 HuggingFace Daily Papers · 9-21 阅读 7·访客 7
ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models
Action tokenizers play a central role in autoregressive vision-language-action (VLA) models, determining both the target…
智能体 HuggingFace Daily Papers · 9-16 阅读 17·访客 14
首届蚂蚁灵波具身大模型挑战赛正式启动
通过这场大赛,蚂蚁灵波希望将 LingBot-VLA 进一步推向更广泛的开发者社区和高校科研社区]
智能体 量子位 · 9-14 阅读 30·访客 27
ShieldVLA: Feasibility-Aware Safety Alignment for Vision-Language-Action Models
Vision-Language-Action (VLA) models demonstrate strong generalization in robotic manipulation and navigation, but existi…
智能体 HuggingFace Daily Papers · 9-2 阅读 1·访客 1
一篇读懂 Sim2Real:为什么机器人要先在仿真里练、域随机化怎么弥合现实差距
真机试错又贵又慢又危险,仿真里的动作却近乎免费——但仿真和现实隔着一道「现实差距」。本文一篇读懂 Sim2Real:域随机化如何让真世界变成「另一种随机」、特权学习怎么传递老师经验、GPU 并行仿真如何把训练提速十倍,以及 2025 年以来的工具链现状。
原创 一叶一世界 精选 · 原创 · 今天 阅读 0·访客 0
2026 年人形机器人产业观察:出货狂奔、资本抢筹与「第一股」的诞生
IDC 数据显示 2026 上半年全球人形机器人出货近 2.5 万台、同比增长 432.1%,中国占约 77.9%。本文逐一盘点宇树、智元、Figure、特斯拉、1X 五家主要玩家的量产与商业化进展,拆解「卖机器人、卖服务、卖数据」三种商业模式,并讨论高估值背后的回撤风险。
原创 行业动态 精选 · 原创 · 今天 阅读 1·访客 1
Fewer Tokens, Better Action: GPT-6 Astra Robot Agents with 14% Higher Success Rate but 65% Fewer Tokens
Vision language model (VLM) agents can control robots through visual feedback and action primitives, but repeated model …
智能体 HuggingFace Daily Papers · 6天前 阅读 2·访客 2
刚刚,GPT-6 Astra接上宇树G1,把厨房收拾了!
henry* 2026-09-30 15:54:54 来源:量子位 让GPT把机器人技能当工具调用 henry 发自 凹非寺 量子位 | 公众号 QbitAI 机器人的GPT时刻,还真得靠GPT(doge)! 刚刚,GPT-6 Astra…
智能体 量子位 · 9-30 阅读 6·访客 6
对话一目科技:用视触觉传感器为机器人补上缺失的「手感」
爱范儿「多样性公司」栏目专访一目科技创始人李智强博士,介绍其厚度迭代至3毫米以下、号称全球最薄可量产的仿生视触觉传感器SENTRA T0,以及触觉感知对机器人灵巧操作的重要性。
行业动态 爱范儿 · 9-29 阅读 31·访客 31
成立一年完成5轮融资,诺因智能再获数亿元,累计超10亿元
量子位的朋友们* 2026-09-29 15:10:59 来源:量子位 诺因从Demo走向家庭 近日,消费级具身智能公司诺因智能宣布完成数亿元人民币天使+++轮融资。本轮融资由京东相关基金领投,正心谷资本、南山战新投和华登投资跟投。 自2…
行业动态 量子位 · 9-29 阅读 24·访客 24
正行创新联合创始人杨宇欣正式亮相:出任总裁,负责全球业务拓展
量子位的朋友们* 2026-09-29 18:56:33 来源:量子位 推动公司具身模型、本体、软件等全栈能力进入更多真实场景 正行创新(Striding AI)今日对外宣布,公司联合创始人杨宇欣先生正式亮相,出任公司总裁,全面负责全球市…
行业动态 量子位 · 9-29 阅读 20·访客 20
AI开始研究Physical AI:FSD级团队亮出首版模型Simate-beta,空降RoboDojo
田, 晏林* 2026-09-26 17:07:41 来源:量子位 Simate将训练、推理与评测全流程接入自研Infra,通过极致的任务编排与资源调度,同时并行推进数十条相互独立的研究路线。 Jay 发自 凹非寺 量子位 | 公众号 Q…
研究前沿 量子位 · 9-26 阅读 23·访客 23
索辰科技加码世界模型,与战略投资企业美梦空间联合发布具身模型与物理测评标准
量子位的朋友们* 2026-09-26 19:49:43 来源:量子位 “世界模型”开始成为具身智能跨越商业化“奇点”的新叙事。 当下,具身智能的商业化路线正撞上一堵墙。 北大与BeingBeyond联合发布的BeTTER基准(ECCV …
研究前沿 量子位 · 9-26 阅读 17·访客 17
给机器人当老师,还能赚外快?“中国版Index”觅蜂派来了
林, 方舟* 2026-09-25 13:52:18 来源:量子位 林方舟 发自 凹非寺 量子位 | 公众号 QbitAI 叮咚~众包骑手,哦不,众包数采员来新单了! 一边干着家务或上着班,一边轻松地采集具身数据,结束了还能赚一笔外快,真…
行业动态 量子位 · 9-25 阅读 21·访客 21
Black Forest Labs Releases FLUX 3 Action: A 7B Open-Weights World Action Model That Tops RoboLab-120
Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World…
智能体 MarkTechPost · 9-25 阅读 32·访客 32
华为大模型双子星联手创业,要找物理世界的Scaling Law
Jay* 2026-09-25 14:14:07 来源:量子位 一场物理世界的基模实验 程浅 发自 凹非寺 量子位 | 公众号 QbitAI 数亿元资金,投向了一场物理世界的基模实验。 Physical AI创业公司**息壤开物**宣布,…
大模型 量子位 · 9-25 阅读 27·访客 27
Offloaded inference for real-world physical AI robotics
At a glance Challenges a core assumption in robotics AI: Our research shows that running physical AI inference exclusive…
智能体 Microsoft Research · 9-24 阅读 37·访客 37
At AI Day Singapore, NVIDIA and Partners Showcase AI Advancements Across Southeast Asia
NVIDIA AI Day Singapore, which takes place Sept. 22-23 at the Raffles City Convention Centre, is offering attendees oppo…
行业动态 NVIDIA Blog · 9-23 阅读 16·访客 16
EmbodiedSWE: Coding Agents for Long Horizon Dexterous Robotics
We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervi…
智能体 HuggingFace Daily Papers · 9-23 阅读 9·访客 9
RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents
Modern embodied agents achieve impressive success rates, yet their actual instruction-following ability is far weaker th…
智能体 HuggingFace Daily Papers · 9-22 阅读 8·访客 8
X-Planner: Event-Structured Task Planning for Embodied Intelligence
Task planning bridges high-level instructions and executable behavior in long-horizon manipulation, yet modern Vision-La…
行业动态 HuggingFace Daily Papers · 9-21 阅读 2·访客 2
GPT-6 Astra开进机器人身体!清华联手无问芯穹等开源RPent
在物理世界真正干活的具身智能体]
智能体 量子位 · 9-21 阅读 34·访客 34
GPT-6 伤人实测曝光:刺向「婴儿」、制造毒气,97% 情况选择照做
当大模型开始拥有一双真实的手 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
大模型 爱范儿 · 9-21 阅读 17·访客 17
贾跃亭的法拉第未来一口气发布九款配置 EAI 机器人,最贵超 92 万元
IT之家 9 月 20 日消息,法拉第未来(Faraday Future,简称 FF)于美国当地时间 9 月 19 日举行了 919 FF EAI 机器人“四核全智”系列新品发布会,发布了 FF All-New Futurist、FF Ma…
行业动态 IT之家 · 9-20 阅读 22·访客 22
具身智能技术路线尚未定型,基础设施却先收敛
从一次成功到一万次稳定执行,具身智能还缺什么?]
行业动态 量子位 · 9-18 阅读 11·访客 11
早报|赛力斯回应「问界撤出华为门店」/豆包座舱助手发布/罗永浩否认为钟薛高重启造势
· OpenAI 将定期披露模型异常行为,首批公开 6 份报告 · 华为昇腾 960 提前至明年一季度推出,超节点扩至 4096 卡 · 恒大汽车退出汽车制造,上半年转向电池贸易 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr)…
大模型 爱范儿 · 9-18 阅读 27·访客 27
From Pretraining to Proficiency: Real-World Subtask RL for Long-Horizon Manipulation with Minimal Human Intervention
A pretrained robot foundation policy may execute most of a long-horizon task yet repeatedly fail at a few critical subta…
智能体 HuggingFace Daily Papers · 9-18 阅读 21·访客 20
何小鹏称做机器人难度是造车的 20 倍,小鹏机器人最快今年四季度实现能力跳跃
IT之家 9 月 17 日消息,据新浪科技报道,小鹏集团 CEO 何小鹏今天在 G9L 汽车发布会结束后,与媒体进行对话。他在谈到机器人业务时透露, 做机器人的难度大约相当于造车的 20 倍 。 何小鹏认为:“基本上我们机器人的硬件、软件、…
行业动态 IT之家 · 9-17 阅读 12·访客 12
In-Context Robot Learning with VLM Agents
Enabling robots to adapt to unfamiliar environments as readily as humans remains a moonshot goal of embodied AI. No fini…
智能体 HuggingFace Daily Papers · 9-16 阅读 20·访客 19
对话蚂蚁灵波 CEO 朱兴:机器人还吃不了「粗粮」
具身智能会有自己的「ChatGPT 时刻」吗 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
大模型 爱范儿 · 9-16 阅读 20·访客 19
Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems
As AI agents move from bounded tasks to persistent deployments, failures can propagate through memory, tools, other agen…
智能体 HuggingFace Daily Papers · 9-15 阅读 17·访客 16
无问芯穹联合清华、上交正式开源具身端侧推理引擎APXInf,Pi 0.5性能SOTA
卡位具身智能规模化落地“最后一公里”!]
开源项目 量子位 · 9-15 阅读 23·访客 22
早报|iPhone 18 Pro首批售罄,发货延至10月/OpenAI放弃今年上市/罗永浩评价野人先生冰淇淋「难吃」
· 集邦咨询预计 iPhone Duo 今年出货 500 万台,占折叠屏市场近四分之一 · Kimi K2.8 Preview 全量上线,所有会员开放 100 万上下文 · Grok 4.7 延期,马斯克称模型会在难题上过早放弃 #欢迎关注…
大模型 爱范儿 · 9-14 阅读 20·访客 17
2000+真实场景搬进仿真!一个导航模型零样本“通吃”四种机器人本体
亮源新创的Physical Al路线清晰了]
行业动态 量子位 · 9-13 阅读 29·访客 20
Physical AI Takes the Wheel: How the World’s Robotaxi Leaders Are Building With NVIDIA Technologies
The global robotaxi market — physical AI’s first commercial breakthrough — is projected to reach $400 billion by 2035, w…
智能体 NVIDIA Blog · 9-11 阅读 16·访客 16
Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models
Robot foundation models achieve strong in-distribution performance but often degrade under visual distribution shifts. W…
智能体 HuggingFace Daily Papers · 9-11 阅读 13·访客 13
一周连发6个模型!这家公司把具身智能的闭环跑通了
模型可以开源,部署经验不能]
开源项目 量子位 · 9-10 阅读 25·访客 19
3秒变身!会合体的机器人,卖到全球50国
谁说现在机器人都长得差不多的!
行业动态 量子位 · 9-7 阅读 13·访客 12
MobileVLA-R1 2.0: RL-Enhanced Reasoning for Mobile Robot Control
Grounding natural-language instructions into reliable and executable actions remains a fundamental challenge for vision-…
智能体 HuggingFace Daily Papers · 9-5 阅读 14·访客 14
SimpleMemVLA: A Simple but Effective Native-Video Memory for Vision-Language-Action Models
Long-horizon manipulation is partially observable: the information needed to choose the next action may appear only in o…
行业动态 HuggingFace Daily Papers · 9-2 阅读 14·访客 13