搜索:Model

共命中 50 条(服务端检索)
NASA and IBM's open source lunar model turns 17 years of orbiter data into a foundation for lunar science
Oct 4, 2026 Nano Banana Pro prompted by THE DECODER The NASA-IBM Lunar Foundation Model makes decades of lunar observati…
开源项目 The Decoder · 4天前 阅读 20·访客 20
Ideogram says its new model can edit part of an image without messing up the rest
Oct 1, 2026 Ideogram says its new model Ideogram 4.5 solves one of AI image editing's biggest headaches.** When you edit…
行业动态 The Decoder · 6天前 阅读 10·访客 10
AWS Strands Labs Releases Strands Decider 2B: An Open Source Decision Model That Picks Options in About 115 ms
AWS Strands Labs releases **Strands Decider 2B**, an open source decision model. It does not generate text. It reads a s…
开源项目 MarkTechPost · 6天前 阅读 32·访客 30
Harness-Aware Distillation for Small Language Model Agents
Language model agents are deployed with a harness, the software around the model that manages its context, tools, and fe…
智能体 HuggingFace Daily Papers · 6天前 阅读 0·访客 0
Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence
Perplexity Research and turbopuffer have released **pplx-embed-v2-context-9b-preview**, a contextual embedding model for…
行业动态 MarkTechPost · 10-1 阅读 14·访客 14
Liquid AI Releases d1: A Decision Model That Returns Calibrated Probabilities With Zero Output Tokens
Liquid AI has released d1, a decision model built for structured choices instead of text generation.** You give it conte…
行业动态 MarkTechPost · 9-30 阅读 46·访客 45
ElevenLabs' new v4 speech model makes AI voices more expressive and consistent
Sep 29, 2026 Elevenlabs Key Points Elevenlabs released Eleven v4, a speech model that follows direction cues more accura…
行业动态 The Decoder · 9-29 阅读 19·访客 19
Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Decision Model That Runs on a CPU
Supersonic Labs, a small AI lab from Brazil, has released Julia 1. It is a compact decision model, not a chatbot. You pa…
智能体 MarkTechPost · 9-27 阅读 60·访客 59
Sarvam AI Releases Saaras V4: A Speech-to-Text Model for All 22 Indian Languages and Global English
Sarvam AI has released Saaras V4, the newest generation of its speech recognition model. It covers all 22 scheduled Indi…
行业动态 MarkTechPost · 9-27 阅读 37·访客 37
Nvidia drops a free 100M-parameter model that identifies up to eight speakers in real time
Sep 27, 2026 Nvidia released Nemotron 3 Diarization, an AI model that identifies which speaker is talking at any given m…
行业动态 The Decoder · 9-27 阅读 15·访客 15
Black Forest Labs launches FLUX 3 Action, an open robotics AI model
Sep 24, 2026 Black Forest Labs is releasing FLUX 3 Action, an open AI model designed for robotics.** It builds on the mu…
智能体 The Decoder · 9-25 阅读 50·访客 47
Fastino Releases GLiNER2.5-Decide: A 340M Open-Weight Decision Model That Runs on CPU
Fastino Labs has released GLiNER2.5-Decide, a 340M-parameter open-weight decision model. It takes text and a schema of t…
行业动态 MarkTechPost · 9-25 阅读 36·访客 35
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB
Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3…
行业动态 MarkTechPost · 9-25 阅读 23·访客 23
InternW0-Δ: A World Action Model Bridging Predictive Dynamics and Actions with 20K+ Hours of Open Data
World Action Models (WAMs) jointly model visual dynamics and action generation for generalist robot manipulation. A cent…
智能体 HuggingFace Daily Papers · 9-25 阅读 11·访客 11
Jev in the Wild: A Data-Driven Analysis of the Jev Model's Functionality, Applications and Ecosystem
Jev is a fast, low-cost decision model that answers natural-language questions with choices, binary judgments, and score…
行业动态 HuggingFace Daily Papers · 9-24 阅读 7·访客 7
A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model
In this **tutorial**, we work with **Jev**, TypeSafe AI’s first System One model, which does not generate text at all: w…
行业动态 MarkTechPost · 9-24 阅读 30·访客 29
Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× Faster Than Jev
Contrastive-LM has released CLM-8B, the first open model in a new class called Contrastive Language Models (CLMs). CLM d…
智能体 MarkTechPost · 9-24 阅读 203·访客 84
NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time
NVIDIA has released Nemotron 3 Diarization, an open-weight speaker diarization model on Hugging Face. It answers one que…
行业动态 MarkTechPost · 9-24 阅读 52·访客 52
Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model
Nokia’s applied research team has open-sourced AnyJev, a Python library that turns an open LLM into a decision model. It…
开源项目 MarkTechPost · 9-23 阅读 70·访客 68
OpenAI says its internal model solved over 100 long-standing math problems after just a month of training
Sep 22, 2026 OpenAI Key Points OpenAI says a new internal model solved more than 100 long-standing math problems after j…
行业动态 The Decoder · 9-22 阅读 24·访客 24
SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6
SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built …
智能体 MarkTechPost · 9-22 阅读 52·访客 51
StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work
StepFun has released Step 5 Preview, a sparse Mixture-of-Experts model with 600B total parameters and 27B active per tok…
智能体 MarkTechPost · 9-21 阅读 30·访客 30
Alibaba Qwen Team Releases Qwen3.8-LiveTranslate: A Real-Time Interpretation Model That Cuts Average Lag to 2.3 Seconds Across 60 Languages
Qwen has released Qwen3.8-LiveTranslate, a real-time simultaneous interpretation model built on a new Interleave archite…
大模型 MarkTechPost · 9-20 阅读 22·访客 22
TypeSafe AI Releases Jev: A System One Model That Returns Typed, Calibrated Decisions Instead of Text
TypeSafe AI released Jev, a System One model that answers typed questions with probabilities instead of generating text.…
行业动态 MarkTechPost · 9-20 阅读 31·访客 29
PrismML Releases Ternary Bonsai 2 27B: A 5.9 GB Apache 2.0 Model Retaining 98.2% of Qwen3.8 27B Performance
PrismML has released Ternary Bonsai 2 27B, a ternary-weight version of Qwen3.8 27B. The language model occupies 5.93 GB,…
开源项目 MarkTechPost · 9-19 阅读 34·访客 34
Linkup Research Releases SPARSEUP: A 149M-Parameter Open-Source Sparse Embedding Model
Linkup Research has released SPARSEUP, an open-source sparse embedding model built on a 149M-parameter ModernBERT backbo…
开源项目 MarkTechPost · 9-19 阅读 32·访客 29
A new kind of AI model from a ChatGPT inventor is thrilling developers
Jev, a new kind of AI model, is showing developers a cheaper and faster path to software intelligence. ]
大模型 TechCrunch · 9-19 阅读 23·访客 22
A Lie Detector Test for Language Models: Reading Knowledge a Model Won't Reveal
Large language models can hold knowledge they do not report. A model may sandbag on a capability evaluation, or answer a…
行业动态 HuggingFace Daily Papers · 9-18 阅读 11·访客 11
Nums AI Releases Causilo: A Tabular Foundation Model That Tops TabArena Among Single Models
Nums AI has released Causilo, a pretrained tabular foundation model for classification and regression with a scikit-lear…
行业动态 MarkTechPost · 9-16 阅读 16·访客 14
Former OpenAI researcher builds an AI model that judges options instead of writing text
TypeSafe AI, founded by former OpenAI researcher Diogo Almeida, is releasing a model that deliberately generates no text…
行业动态 The Decoder · 9-16 阅读 18·访客 18
Prior Labs Releases TabPFN-3.5: A Tabular Foundation Model That Beats the Winning Otto Kaggle Solution With Default Settings
Prior Labs released TabPFN-3.5, a tabular foundation model pretrained only on synthetic data that beats Otto's winning s…
行业动态 MarkTechPost · 9-16 阅读 14·访客 13
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
In this work, we present ZGCM-1, a fully open 7B dense foundation model trained from scratch with extreme data, system, …
智能体 HuggingFace Daily Papers · 9-11 阅读 22·访客 21
Pelican-Sim 1.0: A General World Model Simulator for Embodied Intelligence
In this technical report, we propose Pelican-Sim 1.0, a general world model simulator for embodied intelligence that pre…
智能体 HuggingFace Daily Papers · 9-10 阅读 20·访客 20
AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing
We introduce AuK, an open-source foundational model that unifies speech generation and editing through a common interfac…
开源项目 HuggingFace Daily Papers · 9-8 阅读 29·访客 21
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay
Can one language model hand its live memory to another without the receiver rereading the context? We demonstrate useful…
大模型 HuggingFace Daily Papers · 9-7 阅读 9·访客 9
EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing Agents
Large Language Model (LLM) agents are turning language into real-world effects, making safety necessary against both ind…
智能体 HuggingFace Daily Papers · 9-5 阅读 28·访客 27
Opus 5.2 深夜灰度,RSI 真来了吗?——把一条刷屏新闻拆成三层证据
2026 年 9 月 15 日凌晨,Opus 5.2 在 Claude Code 中被曝灰度测试,"RSI 真来了?"随即刷屏。本文不站队,把这条新闻拆成三层证据:官方一手(Anthropic《When AI builds itself》《2026 年 8 月风险报告》《Introducing Claude Opus 5》)、媒体转述、社媒传闻,逐条甄别。结论:Opus 5.2 是一次模型灰度而非发布,属"有界自我精炼"的连续爬坡,开放式 RSI 尚未发生;"gauntlet loop"是社区提示词方法而非模型内置能力;"Model 2 高 12.5 分"与官方"增幅不大"口径冲突;"替代 85% 研究团队"溯源到社媒,与官方"18 名受访者中仅 1 人认可"直接矛盾。文末给出"三层证据法"与三个必问问题,并指出真正该盯住的是"智能体能力与验证能力之间的差距"。
原创 研究前沿 精选 · 本站原创 · 9-15 阅读 60·访客 51
Google DeepMind Releases EmbeddingGemma 2, a 740M Open Multimodal Embedding Model Built on Gemma 4
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 昨天 阅读 1·访客 1
Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 昨天 阅读 1·访客 1
Reflection AI Introduces Beam: A 501B Open-Weight MoE Model With 23B Active Parameters for Coding and Agentic Workloads
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
智能体 MarkTechPost · 2天前 阅读 0·访客 0
AdvSim2Real : Training Web Agents Against Adaptive Prompt Injection in a Web World Model
Web agents complete user requests by reading and acting on pages that third parties write, so an instruction planted on …
智能体 HuggingFace Daily Papers · 2天前 阅读 0·访客 0
UNREAL: Unifying Retrieval and Long-Context with a Single Model
Long-context inference and Retrieval-Augmented Generation (RAG) handle evidence selection at vastly different scales, fr…
行业动态 HuggingFace Daily Papers · 2天前 阅读 0·访客 0
Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
智能体 MarkTechPost · 2天前 阅读 0·访客 0
GPT-6 Astra vs GPT-6.1 Sol vs Gemini 4 Argon vs Claude Fable 5.1: Which Frontier Model Fits Which Job
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
大模型 MarkTechPost · 3天前 阅读 53·访客 52
Google's new Gemini tiers cut free users to its weakest model and lock $5/month subscribers out of Pro
Oct 4, 2026 Starting in October 2026, Google will restrict access to its Gemini models for personal account users.** Wit…
大模型 The Decoder · 4天前 阅读 30·访客 29
Aleph Alpha Releases Kolibri: A 78.1B Open-Weight English-German MoE Model With Only 3.46B Active Parameters
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 4天前 阅读 20·访客 19
Cloudflare says its new Clef model means humans no longer need to be in the loop for AI agents
Oct 2, 2026 Key Points Cloudflare has released Clef and Clef-flash, two decision models for AI agents that compete direc…
智能体 The Decoder · 5天前 阅读 35·访客 35
Microsoft AI Releases MAI-Transcribe-2-Streaming: #1 Real-Time Speech-to-Text Model on Artificial Analysis
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 5天前 阅读 7·访客 7
Magic-W0: A Structured World-Action Foundation Model for Physical Intelligence
World-action models (WAMs) augment robot policies with action-conditioned environment dynamics, yet existing approaches …
智能体 HuggingFace Daily Papers · 5天前 阅读 0·访客 0
OpenAI’s latest features take direct aim at the app store model
The focus of OpenAI’s Dev Day on Tuesday may have been on its agentic assistants known as Dots, or its new AI models, bu…
智能体 TechCrunch · 9-30 阅读 12·访客 12