🍊 The Latent Field
Search
搜索
暗色模式
亮色模式
学习
探索
标签: sft
阅读模式
Hide sidebars
Show sidebars
此标签下有2条笔记。
2026年6月01日
KLong: Training LLM Agent for Extremely Long-horizon Tasks
source
paper
agents
long-horizon
reinforcement-learning
sft
evaluation
2026年5月29日
Finetuned Language Models Are Zero-Shot Learners
source
paper
instruction-tuning
sft