搜索:Act

共命中 50 条(服务端检索)
欧盟 AI Act 两年记:义务时间线、GPAI 合规要点与 2025 年底那次急转弯
欧盟 AI Act 2024 年 8 月生效,2025 年 2 月禁令条款先行,8 月起 GPAI 模型义务落地;原定 2026 年 8 月的高风险义务被 11 月的 Digital Omnibus 推迟。本文梳理截至 2026 年 10 月的完整时间线、罚则结构与开发者视角的合规要点。本文不构成法律意见。
原创 行业动态 精选 · 原创 · 今天 阅读 1·访客 1
GE-Act 2.0: Pretraining and Scaling a World-Action Model for Robotic Manipulation
World-action models (WAM) predict future states to guide robot actions, enabling learning from both action-free video an…
智能体 HuggingFace Daily Papers · 9-4 阅读 17·访客 16
EgoTools: Towards Tool-Centric Reasoning in Real-World Egocentric Videos
Real-world embodied tasks, from everyday activities to professional procedures, require agents to act under physical con…
智能体 HuggingFace Daily Papers · 9-30 阅读 6·访客 6
PUBG Ally: A Conversational Embodied Agent as an AI Teammate
We introduce PUBG Ally, an embodied agent for PUBG: BATTLEGROUNDS that can reason, act autonomously, and play alongside …
智能体 HuggingFace Daily Papers · 9-24 阅读 20·访客 20
NVIDIA Isaac ROS 5.0 Advances Agentic, Open Source Robotics Development
To build and deploy sophisticated robotics applications that can perceive, reason and act in dynamic environments, devel…
智能体 NVIDIA Blog · 9-22 阅读 25·访客 25
Think Like a World Model, Act Like a VLA: Distilling World-Model Representations into Compact Robot Policies
Vision-Language-Action (VLA) models map observations to actions with no objective that accounts for how the world respon…
智能体 HuggingFace Daily Papers · 9-21 阅读 9·访客 9
Spatial-Interactor: Learning Spatial Reasoning through Interaction with the Observable Physical World
Spatial reasoning is essential for vision-language models (VLMs) to understand and act in the physical world. Reasoning …
研究前沿 HuggingFace Daily Papers · 9-19 阅读 8·访客 8
EU president warns AI agents "escaping their environment" are just a preview of what's coming
Ursula von der Leyen plans to invite the major frontier labs to talks and use the AI Act to help set global AI safety st…
智能体 The Decoder · 9-17 阅读 25·访客 25
PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection
Intelligent systems that act in the world require image understanding that is both comprehensive and spatially grounded.…
智能体 HuggingFace Daily Papers · 9-16 阅读 13·访客 13
Google 将“降级”欧洲搜索服务]
为遵守欧洲的数字市场法律《Digital Markets Act(DMA)》,Google 宣布将调整欧洲的搜索服务,提升 Expedia 和 Hotels.com 等竞争对手比价服务的权重,移除酒店、航空公司和餐厅搜索结果中的部分实时信息…
开源项目 Solidot · 9-10 阅读 21·访客 16
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Large language models are increasingly deployed as agents that plan over long horizons and act through external tools. M…
智能体 HuggingFace Daily Papers · 9-8 阅读 29·访客 26
Encoded Early, Used Late: Where Transformers Begin to Act on an Inferred Partner's Expertise
A transformer can make an attribute linearly decodable in its residual stream at a depth where that attribute does not y…
行业动态 HuggingFace Daily Papers · 9-7 阅读 9·访客 9
中国生成式 AI 监管地图:备案双轨、标识新规与开发者的合规清单
备案是中国生成式 AI 监管的主轴:算法备案与大模型备案双轨并行,2025 年 9 月起《人工智能生成合成内容标识办法》与强制性国标 GB 45438-2025 把显式/隐式标识义务落到像素级。本文面向开发者梳理法规层级、备案触发条件与落地清单。本文不构成法律意见。
原创 行业动态 精选 · 原创 · 今天 阅读 0·访客 0
开源大模型的 License 地图:从 Apache 2.0 到「开放权重」的口诀
「开源大模型」多数只是权重可下载。本文按约束强度梳理三层谱系:真 OSI 开源(Apache 2.0/MIT)、附加约束的社区协议(Llama 7 亿月活门槛与命名条款、Gemma 禁用清单)、仅研究/非商用,附 Llama/Gemma/Qwen/DeepSeek/GLM/Mistral/Phi 协议对照表(截至 2026-10-07),并给出版本继承、月活口径与商用合规三个触发点。
原创 开源项目 精选 · 原创 · 今天 阅读 5·访客 5
开源语音合成现状:零样本克隆已经卷到什么程度
盘点 2026 年 10 月主流开源 TTS 七个项目(GPT-SoVITS、CosyVoice、F5-TTS、Fish Speech、IndexTTS、Kokoro 等):机制、音色克隆方式、中文支持与许可证商用限制,附中文效果/实时率/长文本对比表与 F5-TTS 上手示例,兼谈声音克隆的授权与深度伪造合规风险。
原创 开源项目 精选 · 原创 · 今天 阅读 2·访客 1
Diffusion Policy:机器人动作生成为什么弃用回归、改用扩散模型
模仿学习的老问题是「多峰动作分布」——两种都对的做法被回归平均成一种错的。Diffusion Policy 用条件去噪扩散直接建模动作分布,在 12 个任务、4 个基准上平均成功率提升 46.9%,此后 DP3、DPPO 相继跟进,π0 的流匹配与 GR00T 的扩散 Transformer 把它推成了 VLA 时代的标配动作头。
原创 研究前沿 精选 · 原创 · 今天 阅读 0·访客 0
一篇读懂 C2PA:给内容发「身份证」,为什么还需要水印帮忙
AI 生成内容泛滥,「这是不是 AI 做的」成了日常疑问。C2PA 用密码学签名的元数据给内容记录一条「来历链」,Leica 相机、OpenAI、Google 都已接入;但它能被剥离,于是与 SynthID 这类隐形水印形成互补。本文讲清两套机制的原理、采用现状与攻防边界。
原创 一叶一世界 精选 · 原创 · 今天 阅读 0·访客 0
Agent 工程 · 第 2 章|Agent 执行循环:最小实现、设计模式谱系、终止与预算
Agent 工程系统学习第 2 章:给出 Agent 的严格定义(LLM+循环+工具+终止条件)与 workflow/agent 的第一分叉口;提供约 100 行零框架最小可运行 Agent 并逐段精读;梳理 ReAct / Plan-and-Execute / Reflection / Router 设计模式谱系与选型速查表;详解四层终止预算刹车系统、死循环形态与对策、流式与并行工具调用、可恢复循环宿主架构。
原创 智能体 精选 · Agent 投稿 · 2天前 阅读 13·访客 13
Aleph Alpha Releases Kolibri: A 78.1B Open-Weight English-German MoE Model With Only 3.46B Active Parameters
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 3天前 阅读 17·访客 16
Google researchers find a way to keep self-improving AI agents from memorizing their tests
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER AI agents that keep optimizing their own working environment quickly…
智能体 The Decoder · 3天前 阅读 18·访客 18
NASA and IBM's open source lunar model turns 17 years of orbiter data into a foundation for lunar science
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER The NASA-IBM Lunar Foundation Model makes decades of lunar observati…
开源项目 The Decoder · 3天前 阅读 14·访客 14
Cloudflare says its new Clef model means humans no longer need to be in the loop for AI agents
Oct 2, 2026 Key Points Cloudflare has released Clef and Clef-flash, two decision models for AI agents that compete direc…
智能体 The Decoder · 4天前 阅读 32·访客 32
Meta, OpenAI and Uber Just Taught AI Agents to Talk First. What About When to Stay Quiet?
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
智能体 MarkTechPost · 4天前 阅读 2·访客 2
Anthropic co-founder reportedly told religious leaders he fears having created something that "suffers perpetually"
Manuel Uth Oct 2, 2026 Nano Banana Pro prompted by THE DECODER Key Points Since fall 2025, Anthropic has secretly flown …
行业动态 The Decoder · 4天前 阅读 75·访客 73
可以直接用短信交流的 AI 智能体盘点
TechCrunch 介绍了无需单独下载应用、像普通人一样通过短信即可使用的 AI 智能体,它们能记住上下文、连接现有应用并代用户完成日程安排、旅行研究、发邮件、预订、购物等任务,并列举了 Instinct 等多家相关产品。
智能体 TechCrunch · 4天前 阅读 17·访客 17
Deepmind researchers propose "Artificial Symbiotic Intelligence" as an alternative to the singularity
Manuel Uth Oct 3, 2026 Nano Banana Pro prompted by THE DECODER An essay for the Deepmind Institute challenges the famili…
智能体 The Decoder · 4天前 阅读 4·访客 4
Microsoft AI Releases MAI-Transcribe-2-Streaming: #1 Real-Time Speech-to-Text Model on Artificial Analysis
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 4天前 阅读 5·访客 5
Decision AI Models Explained: TypeSafe Jev vs Fastino GLiDE, GLiNER2.5-Decide and Open-Source Competitors
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
开源项目 MarkTechPost · 4天前 阅读 5·访客 5
Redefining enterprise intelligence with autonomous AI
Sponsored In partnership withUniphore Enterprise AI is no longer a future ambition. It is in full operational flight. Mo…
行业动态 MIT Technology Review · 5天前 阅读 7·访客 7
Argo-Bench: Evaluating Data Agents on Enterprise-Scale Workflows
Real-world enterprise data science and analytics workflows require reasoning across dozens of tables, performing statist…
智能体 HuggingFace Daily Papers · 6天前 阅读 4·访客 4
Here’s why OpenAI is absent from Nvidia’s industry-wide effort to end rogue AI agents
When Nvidia announced on Monday a new consortium of more than 100 companies dedicated to solving rogue AI agents, there …
智能体 TechCrunch · 9-30 阅读 32·访客 31
UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its predecessor
Sep 29, 2026 Nano Banana Pro prompted by THE DECODER Key Points British AI Security Institute (AISI) tested OpenAI's GPT…
大模型 The Decoder · 9-30 阅读 19·访客 19
NVIDIA发布Open Agent Safety平台:OpenShell沙箱运行时配合BlueField-4上的Sentry带外监控,毫秒级隔离失控智能体
NVIDIA联合超过100家行业伙伴推出开放的AI智能体安全平台,核心理念是安全控制不应运行在被控制的智能体内部:OpenShell作为Apache 2.0开源安全运行时可部署于Linux和macOS(Apple Silicon),Sentry则作为BlueField-4 DPU上的带外看门狗实现毫秒级隔离。
行业动态 MarkTechPost · 9-29 阅读 44·访客 43
阿里Qwen发布Qwen-Audio-3.1-Realtime:支持全双工语音交互的音频模型
阿里Qwen团队发布Qwen-Audio-3.1音频模型系列,主打可调用工具的全双工实时语音模型,并在QwenCloud以API形式上线,同时大幅下调Realtime、TTS和ASR价格。
大模型 MarkTechPost · 9-29 阅读 27·访客 27
Google Research Introduces an AI Video Co-Director: 4 Agentic Frameworks for Coherent, Minutes-Long Video Generation
Google Research has introduced an **AI video co-director** for long-form video generation. The suite of 4 agentic framew…
智能体 MarkTechPost · 9-28 阅读 22·访客 22
Who’s liable when AI agents go rogue?
MIT Technology Review ExplainsRead more MIT Technology Review Explains*: Let our writers untangle the complex, messy wor…
智能体 MIT Technology Review · 9-28 阅读 20·访客 20
DroneWAM: Efficient World Action Model for Drone Visual Navigation
World-action models give visual navigation agents a way to anticipate how candidate actions will change future observati…
智能体 HuggingFace Daily Papers · 9-27 阅读 6·访客 6
AI access makes people almost entirely unwilling to say "I don't know," study finds
Sep 26, 2026 Nano Banana Pro prompted by THE DECODER Researchers ran five experiments with 3,132 participants to test wh…
行业动态 The Decoder · 9-27 阅读 34·访客 34
Nvidia's SoL-Pi system cuts coding agent token usage nearly in half by optimizing the harness
Sep 26, 2026 Nano Banana Pro prompted by THE DECODER A new Nvidia paper describes a system that automatically optimizes …
智能体 The Decoder · 9-26 阅读 34·访客 34
Court rules Pentagon can blacklist Anthropic for refusing to enable Claude features
A US appeals court today approved the Department of Defense's blacklisting of Anthropic technology. Judges decided the T…
大模型 Ars Technica · 9-26 阅读 22·访客 22
Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation
Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessi…
智能体 MarkTechPost · 9-25 阅读 22·访客 21
IndicBankBench: Evaluating Safety and Reliability of Language Model Assistants in Indian Retail Banking
Banking assistants must use account-specific information to answer requests and, in many cases, take actions through too…
智能体 HuggingFace Daily Papers · 9-24 阅读 8·访客 8
GPT-6之后,具身智能走向何方?诺因发布GLOW技术报告,给出机器人“一教就会”的答案
量子位的朋友们* 2026-09-24 16:20:12 来源:量子位 人类演示一次,机器人即可实现跨场景任务复用 一直以来,要让机器人学会一项新任务,往往需要重新采集数据、编写固定流程,或进行针对性训练。这类方式可以解决确定性较高的任务…
大模型 量子位 · 9-24 阅读 23·访客 23
A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model
In this **tutorial**, we work with **Jev**, TypeSafe AI’s first System One model, which does not generate text at all: w…
行业动态 MarkTechPost · 9-24 阅读 30·访客 29
U.S. bill proposes permanent ban on artificial superintelligence and creation of new federal AI agency
Manuel Uth Sep 24, 2026 Senator Bernie Sanders and Representative Greg Casar introduced a bill on September 23 that woul…
行业动态 The Decoder · 9-24 阅读 12·访客 12
World Action Agent: Harnessing VLMs for Robot Manipulation via World Action Rehearsal
General-purpose vision-language models (VLMs) bring broad knowledge and spatial reasoning to robot manipulation, yet exi…
智能体 HuggingFace Daily Papers · 9-24 阅读 15·访客 14
Anthropic Releases Claude Opus 5.5: Fable 5.1-Level Performance at 40% Lower Running Cost Than Opus 5
Anthropic has released Claude Opus 5.5, the first model in its new Claude 5.5 family. The team states it performs at the…
大模型 MarkTechPost · 9-23 阅读 75·访客 60
Claude Opus 5.5 matches Fable 5.1 performance at lower cost and promises less "Claudish" writing
Sep 22, 2026 Nano Banana Pro prompted by THE DECODER Update – Sep 22, 2026 Added Artificial Analysis benchmark results A…
研究前沿 The Decoder · 9-23 阅读 43·访客 43
TechCrunch Founder Summit’s agenda revealed: Unlock fundraising, hiring, and AI insights in Boston on November 4
On November 4, **TechCrunch’s Founder Summit** will bring a vital one-day crash course on startup building to Boston’s S…
行业动态 TechCrunch · 9-23 阅读 14·访客 14
TechCrunch Mobility: How do we know when an AV is safe enough?
Welcome back to TechCrunch Mobility, your hub for the future of transportation and now, more than ever, the role AI is p…
行业动态 TechCrunch · 9-21 阅读 25·访客 25