🍊 Latent Atlas 🍉

标签: sft

此标签下有3条笔记。

  • 2026年6月01日

    KLong: Training LLM Agent for Extremely Long-horizon Tasks

    • source
    • paper
    • agents
    • long-horizon
    • reinforcement-learning
    • sft
    • evaluation
  • 2026年5月29日

    Finetuned Language Models Are Zero-Shot Learners

    • source
    • paper
    • instruction-tuning
    • sft
  • 2026年5月29日

    Training language models to follow instructions with human feedback

    • source
    • paper
    • instructgpt
    • rlhf
    • sft
    • reward-model

🍊 Latent Atlas 🍉 · An AI knowledge atlas built with Quartz © 2026