搜索:train

共命中 8 条(服务端检索)
Train Smarter, Not Harder: Switching Signal-Guided Training in Active Learning
Training strategy, namely whether to retrain from scratch or fine-tune from the previous checkpoint, is an overlooked de…
行业动态 HuggingFace Daily Papers 9-6 阅读 0 · 访客 0
大模型工作原理全解析
Tokenizer 的工作机制:BPE 分词、上下文窗口与计费逻辑
原创 大模型 原创博客 精选 · 2025-02-26 阅读 2 · 访客 2
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics
We study how model post-training and test-time inference design affect natural-language proof generation for hard olympi…
行业动态 HuggingFace Daily Papers 9-9 阅读 2 · 访客 0
Difficulty-Adaptive Tree-Structured Policy Optimization for Expanding Reasoning Coverage in RLVR
Reinforcement Learning with Verifiable Rewards (RLVR) has been central to the recent success of Large Reasoning Models. …
研究前沿 HuggingFace Daily Papers 9-8 阅读 1 · 访客 0
Diffs vs. Whole Files: An Empirical Comparison of Iterative Edit-Based and Direct Generation for Flutter/Dart Code Models
Large language models used for code editing can be trained and deployed in at least two output regimes: direct generatio…
行业动态 HuggingFace Daily Papers 9-5 阅读 3 · 访客 0
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions
Many recurring text functions are easy to describe but difficult to implement with rules, while calling a large remote m…
行业动态 HuggingFace Daily Papers 9-3 阅读 0 · 访客 0
LLM周边一览
大语言模型的基本原理、发展脉络与能力边界总览
原创 大模型 原创博客 精选 · 2024-05-25 阅读 3 · 访客 3
GPT-2模型微调
从零微调 GPT-2 的完整流程:数据准备、训练与效果评估
原创 大模型 原创博客 精选 · 2024-04-26 阅读 2 · 访客 2