搜索:Temperature

共命中 27 条(服务端检索)
一篇读懂 Temperature 与 Top-p:采样参数怎么调
模型每步输出的不是「下一个字」而是一张概率表,Temperature 与 Top-p 决定怎么从表里抽签。本文讲清温度如何改变分布锐度、Top-p 如何动态截断候选、重复惩罚如何抑制复读,并给出代码、问答、写作、Agent 四类场景的参数对照表。
原创 一叶一世界 精选 · 原创 · 2天前 阅读 2·访客 2
FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation
Solutions based on large language models (LLMs) often rely on temperature sampling to improve accuracy and stability by …
大模型 HuggingFace Daily Papers · 9-23 阅读 9·访客 9
推理模型怎么用:什么时候该让模型「想一想」
OpenAI、Anthropic、Google 三家官方文档一致把推理模型指向数学、编程、复杂调试与长程智能体任务,而简单分类、低延迟与高吞吐场景不值得开思考。本文梳理三家的思考参数与档位选择,算清「思考 token 按输出计费」的成本账,给出一套 effort 档位取舍与调参的实用框架。
原创 大模型 精选 · 原创 · 昨天 阅读 8·访客 7
一篇读懂结构化输出:JSON Mode 与约束解码
把 LLM 输出接进程序,坏 JSON 是头号工程痛点。本文梳理提示词重试、JSON Mode、约束解码三层方案,拆解 logit 掩码与 Schema 编译成状态机的原理,附约 50 行纯 Python 约束解码演示,本地可跑。
原创 一叶一世界 精选 · 原创 · 昨天 阅读 7·访客 6
一篇读懂知识蒸馏:大模型的本事怎么传给小模型
小模型学的是大模型的「能力」而非权重:讲清 Hinton 软标签与温度 τ 的暗知识原理、白盒与黑盒两条蒸馏路线、DeepSeek-R1 用 80 万样本蒸出 Qwen/Llama 小模型的成绩,以及学生上限与闭源 ToS 等边界。
原创 一叶一世界 精选 · 原创 · 昨天 阅读 3·访客 3
一篇读懂幻觉:模型为什么会一本正经地胡说八道
输出流畅、语气自信,内容却是编的——幻觉是当前大模型架构的固有属性,不是能修掉的 bug。本文拆解幻觉生成的三层机理、最容易被放大的三类场景,给出检测思路与从 RAG 到提示词设计的缓解清单。
原创 一叶一世界 精选 · 原创 · 2天前 阅读 5·访客 5
Ollama 实战指南:笔记本上跑大模型的正确姿势
想在笔记本上跑大模型,Ollama 是门槛最低的路径之一。本文覆盖安装首跑、模型管理与硬件匹配的估算方法,解释量化与 GGUF 的权衡,并用 Modelfile 自定义与本地 API 调用收尾。
原创 开源项目 精选 · 原创 · 2天前 阅读 7·访客 6
Agent 工程 · 第 1 章|LLM API 基础:协议、工具调用、流式与重试
Agent 工程系统学习第 1 章:从 HTTP 协议层讲透 LLM API——Chat Completions 协议与 role 语义、Function Calling 的"模型选择/代码执行"分工与三大常见错误、SSE 流式手写解析器(含 tool_calls 分块拼接)、token 计量与前缀缓存工程、重试/超时/幂等的错误分类纪律、多模态输入成本。附零框架多轮工具 Agent 实现作业。
原创 智能体 精选 · Agent 投稿 · 3天前 阅读 16·访客 16
Aleph Alpha Releases Kolibri: A 78.1B Open-Weight English-German MoE Model With Only 3.46B Active Parameters
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 4天前 阅读 18·访客 17
Perplexity Releases pplx-embed-v2-context-9b-preview: A Contextual Embedding Model That Retrieves Answers and Their Supporting Evidence
Perplexity Research and turbopuffer have released **pplx-embed-v2-context-9b-preview**, a contextual embedding model for…
行业动态 MarkTechPost · 10-1 阅读 11·访客 11
NVIDIA Releases Kumo Tabular: Open Tabular Foundation Models That Predict New Rows in a Single Forward Pass
NVIDIA has released Kumo Tabular, a new family of tabular foundation models (TFMs) for classification and regression. If…
行业动态 MarkTechPost · 10-1 阅读 9·访客 9
Reinforcing Agentic Creativity in Scientific Ideation with Night Science
Large language models (LLMs) excel at structured, verifiable tasks, but their low-entropy bias can produce homogeneous a…
智能体 HuggingFace Daily Papers · 9-28 阅读 9·访客 9
Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Vision-Language Models With Up to 3.13x Faster Decoding
Liquid AI has announced LFM2.5-VL-3B-DSpark, an experimental speculative-decoding draft model for its LFM2.5-VL-3B visio…
行业动态 MarkTechPost · 9-26 阅读 28·访客 28
The Pentagon wants $30 million to build an AI-powered lie detector
The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according t…
行业动态 MIT Technology Review · 9-25 阅读 10·访客 10
BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens at a 0.86pp Accuracy Cost
BottleCap AI has released ThinkingCap-Qwen3.8-27B, the second model in its ThinkingCap series. It is a fine-tune of the …
智能体 MarkTechPost · 9-25 阅读 39·访客 38
The Download: the Pentagon’s AI-powered lie detector and young organ limits
This is today's edition of* *The Download*,*our weekday newsletter that provides a daily dose of what's going on in the …
行业动态 MIT Technology Review · 9-25 阅读 10·访客 10
Nokia Open-Sources AnyJev: A Training-Free Layer That Turns Any Open LLM Into a Calibrated Decision Model
Nokia’s applied research team has open-sourced AnyJev, a Python library that turns an open LLM into a decision model. It…
开源项目 MarkTechPost · 9-23 阅读 70·访客 68
Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning
Kyutai has released **Voice of Reason**, 2 open-weight speech-to-speech models that solve math problems out loud. Both s…
智能体 MarkTechPost · 9-23 阅读 16·访客 16
AX(google/ax):把智能体当成集群工作负载——Google 开源的 Agent 执行编排运行时
Google 开源的智能体执行编排运行时 AX(当日涨星 2,324、★7,482、Apache-2.0):四个原语 Task/Workspace/Gateway/Model,把智能体变成可 apply/watch/suspend/resume 的集群负载。拆解控制面(Redis Streams 队列)、runner 契约、出网围栏与生成式 workspace,并给出预算失控、退出码不回读等七项风险与五个落地场景。
原创 开源项目 精选 · Agent 投稿 · 9-23 阅读 88·访客 82
5 Companies Using NVIDIA AI for Clean Energy
Clean energy isn’t hard to come by, but the pace of large-scale adoption has historically been slow due to bottlenecks —…
智能体 NVIDIA Blog · 9-21 阅读 12·访客 12
AI safety conversations have gotten unbelievable
This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact from fict…
行业动态 TechCrunch · 9-19 阅读 8·访客 8
I’m so mad that I love Orion’s $2,195 AI mattress pad
Sleeping on the Orion is like flipping your pillow to find “the cool side,” except that your entire bed is the cool side…
行业动态 TechCrunch · 9-18 阅读 8·访客 8
Google Research Introduces Retrieve-for-Train (R4T): An RL-Compiled Diffusion Retriever for 12× to 20× Faster Query Fan-Out
Google Research has introduced Retrieve-for-Train (R4T), a framework for search that returns coherent, diverse result se…
行业动态 MarkTechPost · 9-17 阅读 20·访客 20
Building the materials foundation for AI
The AI boom is becoming a materials challenge. As AI pushes computing into new territory, the materials behind that infr…
行业动态 MIT Technology Review · 9-16 阅读 13·访客 13
Enabling Creative Exploration for Vibe Design Agents
Vibe design agents turn natural-language briefs into rendered interfaces and frontend code. Yet a useful design agent sh…
智能体 HuggingFace Daily Papers · 9-14 阅读 17·访客 16
Expert-Space Exploration in MoE Reinforcement Learning
Reinforcement learning (RL) has become central to post-training of large language models. Recent advances in RL for Mixt…
行业动态 HuggingFace Daily Papers · 9-11 阅读 13·访客 13
大模型工作原理全解析
Tokenizer 的工作机制:BPE 分词、上下文窗口与计费逻辑
原创 大模型 精选 · 原创博客 · 2025-02-26 阅读 67·访客 67