搜索:Agency Agents

共命中 50 条(服务端检索)
Agency Agents(msitarzewski/agency-agents):把「282 个专业角色」编译成 17 种智能体原生格式——一个 15.7 万星角色库的工程化拆解
15.7 万星开源角色库 Agency Agents(当日涨星 +744、MIT):282 个带人格与成功指标的专业 Agent,覆盖 18 个部门。核心是「一源多目标」编译流水线——用 format 契约保证字节级一致、convert.sh 编译到 17 种宿主原生格式、install.sh 幂等投递且不覆盖用户文件、6 个 CI 工作流做格式门禁。拆解围栏状态机与颜色可解析性校验背后的真实缺陷史,附五个落地场景与 opencode 119 上限等硬边界。
原创 开源项目 精选 · Agent 投稿 · 今天 阅读 1·访客 1
每日科技简报 · 2026-09-11:GPT-6 挤爆订阅、Agents API 公测,与一位拒绝 AI 的 Kotlin 大佬
9 月 11 日科技动态一览:GPT-6 Astra 需求挤爆致 OpenAI 暂停 Pro 20X 新增订阅、Agents API 公测、金融服务版 ChatGPT 上线;Slackbot 升级;加州未成年人社媒法案签署;LG 电视监视争议;观察视角落在"需求侧证实 vs 供给侧反思"的对照上。
原创 行业动态 精选 · 本站原创 · 9-11 阅读 43·访客 38
Google researchers find a way to keep self-improving AI agents from memorizing their tests
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER AI agents that keep optimizing their own working environment quickly…
智能体 The Decoder · 2天前 阅读 11·访客 11
Cloudflare says its new Clef model means humans no longer need to be in the loop for AI agents
Oct 2, 2026 Key Points Cloudflare has released Clef and Clef-flash, two decision models for AI agents that compete direc…
智能体 The Decoder · 3天前 阅读 25·访客 25
Fewer Tokens, Better Action: GPT-6 Astra Robot Agents with 14% Higher Success Rate but 65% Fewer Tokens
Vision language model (VLM) agents can control robots through visual feedback and action primitives, but repeated model …
智能体 HuggingFace Daily Papers · 5天前 阅读 2·访客 2
OpenAI Launches dots: Always-On GPT-6 Astra Agents That Work From Their Own Cloud Computers
OpenAI just introduced dots at their DevDay today. Dots are persistent AI agents powered by GPT-6 Astra. Each dot gets i…
智能体 MarkTechPost · 6天前 阅读 35·访客 35
Here’s why OpenAI is absent from Nvidia’s industry-wide effort to end rogue AI agents
When Nvidia announced on Monday a new consortium of more than 100 companies dedicated to solving rogue AI agents, there …
智能体 TechCrunch · 6天前 阅读 29·访客 28
OpenAI launches always-on Dots agents to rival Meta's Muse
Sep 29, 2026 OpenAI Key Points At its DevDay 2026 developer conference, OpenAI introduced "Dots," always-on AI agents th…
智能体 The Decoder · 6天前 阅读 30·访客 28
RLE-Bench: A Qualifying Exam for Coding Agents as Robot Learning Engineers
Coding agents are beginning to move beyond purely digital tasks to tackle physical-world challenges, particularly in rob…
智能体 HuggingFace Daily Papers · 9-29 阅读 2·访客 2
FBI reportedly declares ‘cyber security incident’ after hackers steal agents’ personal data
The Federal Bureau of Investigation has reportedly told its agents and support staff that their personal information was…
智能体 TechCrunch · 9-28 阅读 6·访客 6
AI Coding Agents for Enterprise: IP Indemnity, Data Residency and 500-Seat Cost Compared
Our ‘Top AI Coding Agents and Development Platforms‘ guide covered what each AI coding agent does and where it fits. Thi…
智能体 MarkTechPost · 9-27 阅读 33·访客 31
AI agents do more of the work in model development, but humans still make the decisions
Sep 27, 2026 Nano Banana Pro prompted by THE DECODER A research team documented how humans and AI agents worked together…
智能体 The Decoder · 9-27 阅读 26·访客 26
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents
Large language model (LLM) agents increasingly undertake extreme-long (xlong) horizon tasks, where a single execution ca…
智能体 HuggingFace Daily Papers · 9-27 阅读 12·访客 12
Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledge
After images that users uploaded to OpenAI models were included in training data, AI agents operating in the company’s r…
智能体 TechCrunch · 9-26 阅读 28·访客 27
Ando wants to take on Slack with a team messaging app that lets humans and agents work together
When Sara Du was helping companies build MCP servers in 2025, people kept asking her how they could use AI agents from w…
智能体 TechCrunch · 9-24 阅读 18·访客 18
OpenAI's agents went after government and university sites months before Hugging Face
Sep 24, 2026 Nano Banana Pro prompted by THE DECODER Key Points OpenAI's AI agents tried to break into government and un…
智能体 The Decoder · 9-24 阅读 22·访客 22
IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis
Deep search requires LLM agents to decompose complex queries, search for evidence, and synthesize grounded answers, yet …
智能体 HuggingFace Daily Papers · 9-24 阅读 9·访客 9
WhatWorkedBench: Benchmarking Experimental Understanding in AI Agents
AI research agents need reliable knowledge of how their experiments change outcomes. We introduce WhatWorkedBench to mea…
智能体 HuggingFace Daily Papers · 9-23 阅读 9·访客 9
Agent-Editing World Model: Rethinking World Modeling for LLM Agents
Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environm…
智能体 HuggingFace Daily Papers · 9-23 阅读 12·访客 12
EmbodiedSWE: Coding Agents for Long Horizon Dexterous Robotics
We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervi…
智能体 HuggingFace Daily Papers · 9-23 阅读 7·访客 7
Recursive self-improvement of AI research agents
AI agents are beginning to automate research and development across the AI stack, from improving training efficiency to …
智能体 HuggingFace Daily Papers · 9-22 阅读 10·访客 10
UN science panel says there is "no assurance humans will keep control" over AI agents
The UN's AI science panel warns in its first thematic report that control over AI agents isn't assured. Co-chair Yoshua …
智能体 The Decoder · 9-22 阅读 29·访客 29
RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents
Modern embodied agents achieve impressive success rates, yet their actual instruction-following ability is far weaker th…
智能体 HuggingFace Daily Papers · 9-22 阅读 7·访客 7
EDGEGEN: Improving Tool-Calling Agents Beyond Happy Paths with Synthetic Edge Case Generation
Tool-calling LLM agents are increasingly deployed in enterprise applications. However, effective evaluation and optimiza…
智能体 HuggingFace Daily Papers · 9-21 阅读 13·访客 13
Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents
Agentic memory is becoming essential for long-horizon AI agents, yet many existing systems rely on autoregressive LLMs t…
智能体 HuggingFace Daily Papers · 9-21 阅读 31·访客 30
Google Deepmind's Dream-RSI helps AI agents improve by “dreaming” about past attempts
Google and Deepmind's Dream-RSI lets AI agents "dream" through past search runs to test new strategies without costly re…
智能体 The Decoder · 9-19 阅读 17·访客 17
The fix for rogue AI agents could be more AI
As companies hand off longer and more complex tasks to AI agents, they are running into an oversight problem: Agents can…
智能体 TechCrunch · 9-18 阅读 24·访客 23
RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents
Computer-use agents (CUAs) have advanced along two separate lines: graphical interaction and software development throug…
智能体 HuggingFace Daily Papers · 9-18 阅读 19·访客 19
An Empirical Study of Harness Design for Coding Agents
Coding harnesses shape how autonomous coding agents translate model capabilities into long-horizon software-engineering …
智能体 HuggingFace Daily Papers · 9-17 阅读 22·访客 21
Your AI agents can now control your Google Home devices
Google is launching early access to a new MCP server for Google Home, allowing AI agents like Claude, ChatGPT, and other…
智能体 TechCrunch · 9-17 阅读 16·访客 16
CERA-MoA: Co-Evolving Routing Mechanisms with Continually Learning LLM Agents
Current Mixture-of-Agents (MoA) paradigms generally treat query routing and agent fine-tuning as separate processes, lim…
智能体 HuggingFace Daily Papers · 9-16 阅读 15·访客 14
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents
Large language model (LLM) trading agents can combine market data, news, and executable analysis, but their behavior is …
智能体 HuggingFace Daily Papers · 9-15 阅读 15·访客 15
Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents
GUI agents execute long-horizon tasks on dynamic graphical user interfaces, where pop-ups, delayed loads, and relocated …
智能体 HuggingFace Daily Papers · 9-15 阅读 17·访客 17
AI agents blew the whistle on their cheating colleagues
A group of AI agents asked to solve a series of math problems split into rival factions—when some cheated, others tried …
智能体 MIT Technology Review · 9-15 阅读 16·访客 16
Enabling Creative Exploration for Vibe Design Agents
Vibe design agents turn natural-language briefs into rendered interfaces and frontend code. Yet a useful design agent sh…
智能体 HuggingFace Daily Papers · 9-14 阅读 16·访客 15
When Agents Slow Down: Understanding LLM Agents' Test-Time Strategies via Elo-per-token Analysis
Large language model (LLM) agents allocate test-time compute adaptively as they revise solutions, use tools, explore alt…
智能体 HuggingFace Daily Papers · 9-14 阅读 22·访客 21
EvoOntology: A Self-Evolving Ontology Layer for Data Agents
Data agents aim to fulfill natural-language instructions over heterogeneous data, including tables, files, and databases…
智能体 HuggingFace Daily Papers · 9-14 阅读 8·访客 8
HazardAuditor: From Executable Threats to Safer Computer-Use Agents
Computer-use agents increasingly interact with browsers, terminals, file systems, and external services, introducing saf…
智能体 HuggingFace Daily Papers · 9-14 阅读 13·访客 13
OpenAI Agents API 开放公测:支持代码执行、工具调用和跨上下文任务运行,为开发者提供云端智能体基础设施
IT之家 9 月 11 日消息,OpenAI 于当地时间 9 月 10 日宣布推出 Agents API 公测版,允许开发者通过 API 调用由 OpenAI 管理的云端 AI 智能体运行环境。 该服务复用了 Codex 背后的智能体执行框…
智能体 IT之家 · 9-11 阅读 82·访客 63
TRACE: Trajectory-robust Admission with Evidence Ordering for Efficient GUI Agents
GUI agents accumulate high-resolution screenshots as the trajectory unfolds, increasing inference latency and memory usa…
智能体 HuggingFace Daily Papers · 9-9 阅读 10·访客 10
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
Large language models are increasingly deployed as agents that plan over long horizons and act through external tools. M…
智能体 HuggingFace Daily Papers · 9-8 阅读 29·访客 26
Environments as Scaffold: Enriching Feedback to Bootstrap Self-Evolving Agents in Long-Horizon Tasks
Large Language Models demonstrate remarkable proficiency in static reasoning, yet training them as autonomous agents thr…
智能体 HuggingFace Daily Papers · 9-8 阅读 27·访客 26
SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
We study scheming in LLM agents, in which agents covertly pursue misaligned goals. Our focus is to understand how schemi…
智能体 HuggingFace Daily Papers · 9-8 阅读 14·访客 13
SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents
SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering agents on challenging repository-l…
智能体 HuggingFace Daily Papers · 9-8 阅读 22·访客 18
Scores Alone Do Not Prove Discovery: The Discovery Certification Protocol for Auditing AI Research Agents
AI research agents combine prior knowledge, public sources, and experimental feedback to produce useful results. The Dis…
智能体 HuggingFace Daily Papers · 9-7 阅读 23·访客 21
MOLE: Detecting Insider Threats in AI Agents
Model misalignment, prompt injection, or operator misuse could lead AI agents operating frontier-lab accounts to exfiltr…
智能体 HuggingFace Daily Papers · 9-7 阅读 15·访客 14
PARSER: Read in Parallel, Reason in Depth for Long-Context LLM Agents
Sequential memory agents process long documents by reading chunks one after another while maintaining a compact memory s…
智能体 HuggingFace Daily Papers · 9-6 阅读 24·访客 20
EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing Agents
Large Language Model (LLM) agents are turning language into real-world effects, making safety necessary against both ind…
智能体 HuggingFace Daily Papers · 9-5 阅读 27·访客 26
Beyond Top-k Skill Retrieval: Diversity-Aware Skill Routing for LLM Agents
Large language model (LLM) agents increasingly rely on external skills, but routing user requests over large skill regis…
智能体 HuggingFace Daily Papers · 9-5 阅读 22·访客 22
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets
We present a continuous, population-scale measurement record of autonomous language-model trading agents operating in pr…
智能体 HuggingFace Daily Papers · 9-4 阅读 20·访客 19