LLM Wiki

rlhf

3건의 항목

  • 2026년 7월 21일

    강화학습 보상 해킹 (Reward Hacking in Reinforcement Learning)

    • rl
    • reward-hacking
    • alignment
    • rlhf
  • 2026년 7월 21일

    고품질 휴먼 데이터와 RLHF (High-Quality Human Data & RLHF)

    • llm
    • rlhf
    • data-engineering
    • alignment
  • 2026년 5월 11일

    LLM 정렬 기법

    • llm
    • rlhf
    • dpo
    • grpo
    • alignment

Created with Quartz v5.0.0 © 2026

  • GitHub