AI
AI
资讯
alishangtian.com
首页
大模型
智能体
开源项目
研究前沿
行业动态
专题
专题 · TOPICS
一叶一世界
2 篇
算法题解
24 篇
后端技术
19 篇
全部专题 →
主题色 · THEME
自定义
恢复默认
提交线索
# HuggingFace Daily Papers
# IT之家
# Solidot
# 量子位
# agent
# llm
# 爱范儿
# 算法题解
搜索:
train
共命中 8 条(服务端检索)
Train
Smarter, Not Harder: Switching Signal-Guided
Train
ing in Active Learning
Train
ing strategy, namely whether to re
train
from scratch or fine-tune from the previous checkpoint, is an overlooked de…
行业动态
HuggingFace Daily Papers
9-6
阅读 0 · 访客 0
大模型工作原理全解析
Tokenizer 的工作机制:BPE 分词、上下文窗口与计费逻辑
原创
大模型
原创博客
精选
· 2025-02-26
阅读 2 · 访客 2
An Open Recipe for IMO Gold:
Train
ing Nemotron for Olympiad Mathematics
We study how model post-
train
ing and test-time inference design affect natural-language proof generation for hard olympi…
行业动态
HuggingFace Daily Papers
9-9
阅读 2 · 访客 0
Difficulty-Adaptive Tree-Structured Policy Optimization for Expanding Reasoning Coverage in RLVR
Reinforcement Learning with Verifiable Rewards (RLVR) has been central to the recent success of Large Reasoning Models. …
研究前沿
HuggingFace Daily Papers
9-8
阅读 1 · 访客 0
Diffs vs. Whole Files: An Empirical Comparison of Iterative Edit-Based and Direct Generation for Flutter/Dart Code Models
Large language models used for code editing can be
train
ed and deployed in at least two output regimes: direct generatio…
行业动态
HuggingFace Daily Papers
9-5
阅读 3 · 访客 0
Compile by
Train
ing: Turning Natural-Language Specifications into Local Neural Functions
Many recurring text functions are easy to describe but difficult to implement with rules, while calling a large remote m…
行业动态
HuggingFace Daily Papers
9-3
阅读 0 · 访客 0
LLM周边一览
大语言模型的基本原理、发展脉络与能力边界总览
原创
大模型
原创博客
精选
· 2024-05-25
阅读 3 · 访客 3
GPT-2模型微调
从零微调 GPT-2 的完整流程:数据准备、训练与效果评估
原创
大模型
原创博客
精选
· 2024-04-26
阅读 2 · 访客 2