搜索:Model Context Protocol

共命中 50 条(服务端检索)
Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence
Perplexity Research and turbopuffer have released **pplx-embed-v2-context-9b-preview**, a contextual embedding model for…
行业动态 MarkTechPost · 5天前 阅读 6·访客 6
StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work
StepFun has released Step 5 Preview, a sparse Mixture-of-Experts model with 600B total parameters and 27B active per tok…
智能体 MarkTechPost · 9-21 阅读 27·访客 27
Agent 工程 · 第 5 章|工具设计与 MCP 协议:描述工程、协议详解、实现
Agent 工程系统学习第 5 章:工具是 Agent 的手脚。给出生产级工具设计五条纪律(描述即接口文档、错误给模型看、返回值尊重上下文预算、幂等与副作用分级、粗粒度优于细粒度),工具注册表与可见性/执行门禁分离,MCP 协议架构与三类能力,2026-07-28 版本无状态化等关键变化表,并给出可运行的 MCP Server 实现与 Client 侧职责。
原创 智能体 精选 · Agent 投稿 · 昨天 阅读 5·访客 4
Flattening Every Memory Peak in Long-Context Mixture-of-Experts Training
Training a Mixture-of-Experts (MoE) model at long context or large batch size fails as soon as any one component's peak …
行业动态 HuggingFace Daily Papers · 9-13 阅读 16·访客 16
Convergent Emergence of In-Context Learning Across Modalities
Few-shot in-context learning (ICL), the capacity of a model to infer abstract patterns from input-output examples provid…
行业动态 HuggingFace Daily Papers · 9-12 阅读 6·访客 6
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay
Can one language model hand its live memory to another without the receiver rereading the context? We demonstrate useful…
大模型 HuggingFace Daily Papers · 9-7 阅读 6·访客 6
Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Decision Model That Runs on a CPU
Supersonic Labs, a small AI lab from Brazil, has released Julia 1. It is a compact decision model, not a chatbot. You pa…
智能体 MarkTechPost · 9-27 阅读 55·访客 54
Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use
Alibaba's Qwen3.8-Omni-Flash understands audio and video, plans tasks, calls tools, and reports about 45.7% fewer tokens…
智能体 MarkTechPost · 9-18 阅读 29·访客 29
NASA and IBM's open source lunar model turns 17 years of orbiter data into a foundation for lunar science
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER The NASA-IBM Lunar Foundation Model makes decades of lunar observati…
开源项目 The Decoder · 2天前 阅读 8·访客 8
Liquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Output Tokens
Liquid AI has released d1, a decision model built for structured choices instead of text generation.** You give it conte…
行业动态 MarkTechPost · 6天前 阅读 40·访客 39
ElevenLabs' new v4 speech model makes AI voices more expressive and consistent
Sep 29, 2026 Elevenlabs Key Points Elevenlabs released Eleven v4, a speech model that follows direction cues more accura…
行业动态 The Decoder · 9-29 阅读 15·访客 15
Sarvam AI Releases Saaras V4: A Speech-to-Text Model for All 22 Indian Languages and Global English
Sarvam AI has released Saaras V4, the newest generation of its speech recognition model. It covers all 22 scheduled Indi…
行业动态 MarkTechPost · 9-27 阅读 31·访客 31
Fastino Releases GLiNER2.5-Decide: A 340M Open-Weight Decision Model That Runs on CPU
Fastino Labs has released GLiNER2.5-Decide, a 340M-parameter open-weight decision model. It takes text and a schema of t…
行业动态 MarkTechPost · 9-25 阅读 33·访客 32
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB
Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3…
行业动态 MarkTechPost · 9-25 阅读 21·访客 21
A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model
In this **tutorial**, we work with **Jev**, TypeSafe AI’s first System One model, which does not generate text at all: w…
行业动态 MarkTechPost · 9-24 阅读 28·访客 27
Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× Faster Than Jev
Contrastive-LM has released CLM-8B, the first open model in a new class called Contrastive Language Models (CLMs). CLM d…
智能体 MarkTechPost · 9-24 阅读 198·访客 79
NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time
NVIDIA has released Nemotron 3 Diarization, an open-weight speaker diarization model on Hugging Face. It answers one que…
行业动态 MarkTechPost · 9-24 阅读 52·访客 52
SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6
SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built …
智能体 MarkTechPost · 9-22 阅读 48·访客 47
Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages
Qwen has released Qwen3.8-LiveTranslate, a real-time simultaneous interpretation model built on a new Interleave archite…
大模型 MarkTechPost · 9-20 阅读 22·访客 22
PrismML Releases Ternary Bonsai 2 27B: A 5.9 GB Apache 2.0 Model Retaining 98.2% of Qwen3.8 27B Performance
PrismML has released Ternary Bonsai 2 27B, a ternary-weight version of Qwen3.8 27B. The language model occupies 5.93 GB,…
开源项目 MarkTechPost · 9-19 阅读 33·访客 33
Nums AI Releases Causilo: A Tabular Foundation Model That Tops TabArena Among Single Models
Nums AI has released Causilo, a pretrained tabular foundation model for classification and regression with a scikit-lear…
行业动态 MarkTechPost · 9-16 阅读 16·访客 14
Prior Labs Releases TabPFN-3.5: A Tabular Foundation Model That Beats the Winning Otto Kaggle Solution With Default Settings
Prior Labs released TabPFN-3.5, a tabular foundation model pretrained only on synthetic data that beats Otto's winning s…
行业动态 MarkTechPost · 9-16 阅读 14·访客 13
SpectralShift: Effective Context Window Extension of Gated DeltaNet via Spectral Reparameterization
Recently, linear attention layers have been increasingly adopted to replace softmax attention at scale for long-context …
行业动态 HuggingFace Daily Papers · 9-13 阅读 9·访客 9
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
In this work, we present ZGCM-1, a fully open 7B dense foundation model trained from scratch with extreme data, system, …
智能体 HuggingFace Daily Papers · 9-11 阅读 20·访客 19
Pelican-Sim 1.0: A General World Model Simulator for Embodied Intelligence
In this technical report, we propose Pelican-Sim 1.0, a general world model simulator for embodied intelligence that pre…
智能体 HuggingFace Daily Papers · 9-10 阅读 20·访客 20
AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing
We introduce AuK, an open-source foundational model that unifies speech generation and editing through a common interfac…
开源项目 HuggingFace Daily Papers · 9-8 阅读 28·访客 20
Aleph Alpha Releases Kolibri: A 78.1B Open-Weight English-German MoE Model With Only 3.46B Active Parameters
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 2天前 阅读 6·访客 6
Ideogram says its new model can edit part of an image without messing up the rest
Oct 1, 2026 Ideogram says its new model Ideogram 4.5 solves one of AI image editing's biggest headaches.** When you edit…
行业动态 The Decoder · 4天前 阅读 9·访客 9
AWS Strands Labs Releases Strands Decider 2B: An Open Source Decision Model That Picks Options in About 115 ms
AWS Strands Labs releases **Strands Decider 2B**, an open source decision model. It does not generate text. It reads a s…
开源项目 MarkTechPost · 4天前 阅读 19·访客 17
DISCO: Distributed Long Context Scaling with Grounding-Reasoning Disaggregation
While Large Language Models (LLMs) advertise million-token context windows, reasoning quality often collapses as inputs …
研究前沿 HuggingFace Daily Papers · 9-27 阅读 8·访客 8
Nvidia drops a free 100M-parameter model that identifies up to eight speakers in real time
Sep 27, 2026 Nvidia released Nemotron 3 Diarization, an AI model that identifies which speaker is talking at any given m…
行业动态 The Decoder · 9-27 阅读 14·访客 14
Black Forest Labs launches FLUX 3 Action, an open robotics AI model
Sep 24, 2026 Black Forest Labs is releasing FLUX 3 Action, an open AI model designed for robotics.** It builds on the mu…
智能体 The Decoder · 9-25 阅读 50·访客 47
InternW0-Δ: A World Action Model Bridging Predictive Dynamics and Actions with 20K+ Hours of Open Data
World Action Models (WAMs) jointly model visual dynamics and action generation for generalist robot manipulation. A cent…
智能体 HuggingFace Daily Papers · 9-25 阅读 7·访客 7
Jev in the Wild: A Data-Driven Analysis of the Jev Model's Functionality, Applications and Ecosystem
Jev is a fast, low-cost decision model that answers natural-language questions with choices, binary judgments, and score…
行业动态 HuggingFace Daily Papers · 9-24 阅读 2·访客 2
Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model
Nokia’s applied research team has open-sourced AnyJev, a Python library that turns an open LLM into a decision model. It…
开源项目 MarkTechPost · 9-23 阅读 69·访客 67
NVIDIA Introduces SoL-Pi: Auto-Research Loops That Cut Coding Agent Token Traffic by Up to 49%
Coding agents now run for hours, not minutes. Every edit, test run and log read goes back into the model’s context. A te…
智能体 MarkTechPost · 9-22 阅读 30·访客 29
OpenAI says its internal model solved over 100 long-standing math problems after just a month of training
Sep 22, 2026 OpenAI Key Points OpenAI says a new internal model solved more than 100 long-standing math problems after j…
行业动态 The Decoder · 9-22 阅读 15·访客 15
TypeSafe AI Releases Jev: A System One Model That Returns Typed, Calibrated Decisions Instead of Text
TypeSafe AI released Jev, a System One model that answers typed questions with probabilities instead of generating text.…
行业动态 MarkTechPost · 9-20 阅读 29·访客 27
Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model
Linkup Research has released SPARSEUP, an open-source sparse embedding model built on a 149M-parameter ModernBERT backbo…
开源项目 MarkTechPost · 9-19 阅读 28·访客 27
A new kind of AI model from a ChatGPT inventor is thrilling developers
Jev, a new kind of AI model, is showing developers a cheaper and faster path to software intelligence. ]
大模型 TechCrunch · 9-19 阅读 23·访客 22
A Lie Detector Test for Language Models: Reading Knowledge a Model Won't Reveal
Large language models can hold knowledge they do not report. A model may sandbag on a capability evaluation, or answer a…
行业动态 HuggingFace Daily Papers · 9-18 阅读 10·访客 10
Former OpenAI researcher builds an AI model that judges options instead of writing text
TypeSafe AI, founded by former OpenAI researcher Diogo Almeida, is releasing a model that deliberately generates no text…
行业动态 The Decoder · 9-16 阅读 18·访客 18
Co-Evolving Harnesses and Models: On-Policy Correction Helps Weaker Models Catch Up Where Imitation Fails
Agent harnesses (the system prompt, tool set, execution hooks, and context-management scaffolding around a model) are a …
智能体 HuggingFace Daily Papers · 9-8 阅读 13·访客 12
EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing Agents
Large Language Model (LLM) agents are turning language into real-world effects, making safety necessary against both ind…
智能体 HuggingFace Daily Papers · 9-5 阅读 27·访客 26
Anthropic 开放 MCP 协议:AI 应用的 USB-C
Anthropic 发布模型上下文协议(Model Context Protocol),以开放标准统一 LLM 应用与外部数据源、工具的连接方式,被社区称为'AI 应用的 USB-C 接口'。
智能体 精选 · Anthropic · 2024-11-26 阅读 21·访客 18
GPT-6 Astra vs GPT-6.1 Sol vs Gemini 4 Argon vs Claude Fable 5.1: Which Frontier Model Fits Which Job
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
大模型 MarkTechPost · 昨天 阅读 10·访客 9
Cloudflare says its new Clef model means humans no longer need to be in the loop for AI agents
Oct 2, 2026 Key Points Cloudflare has released Clef and Clef-flash, two decision models for AI agents that compete direc…
智能体 The Decoder · 3天前 阅读 25·访客 25
Microsoft AI Releases MAI-Transcribe-2-Streaming: #1 Real-Time Speech-to-Text Model on Artificial Analysis
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 3天前 阅读 2·访客 2
SCOPD: Sparse-Context On-Policy Self-Distillation for Efficient Vision-Language Models
Reasoning vision-language models (VLMs) process images and videos as long sequences of visual tokens, making inference e…
研究前沿 HuggingFace Daily Papers · 9-28 阅读 16·访客 16
AI agents do more of the work in model development, but humans still make the decisions
Sep 27, 2026 Nano Banana Pro prompted by THE DECODER A research team documented how humans and AI agents worked together…
智能体 The Decoder · 9-27 阅读 25·访客 25