도형준 지식 볼트

태그: RLHF

6건의 항목

  • 2026년 7월 21일

    Claude (1~3.5 시리즈): Anthropic의 Constitutional AI 기반 대형 언어 모델

    • Alignment
    • Anthropic
    • Claude
    • Constitutional-AI
    • Harmless-AI
    • Multimodal
    • RLAIF
    • RLHF
    • Safety
  • 2026년 7월 21일

    GPT-5 업데이트: 지속적 개선과 에이전트 강화

    • Agentic-AI
    • Continual-Learning
    • GPT-5
    • GPT-5.2
    • Model-Updates
    • OpenAI
    • RLAIF
    • RLHF
    • Safety-Alignment
  • 2026년 7월 21일

    Claude Opus 4: Anthropic의 에이전틱 AI 플래그십 모델

    • Agentic-AI
    • Anthropic
    • ASL-3
    • Claude
    • Constitutional-AI
    • Opus-4
    • RLHF
    • Safety
    • Safety-Evaluation
    • SWE-bench
  • 2026년 7월 21일

    GPT-4: 멀티모달 대형 언어 모델

    • 128K-Context
    • Function-Calling
    • GPT-4
    • Mixture-of-Experts
    • MMLU
    • Multimodal
    • OpenAI
    • RLHF
  • 2026년 7월 21일

    Claude Opus 4.5: SWE-bench 80.9%를 달성한 Anthropic의 최강 에이전트 모델

    • Agentic-AI
    • Anthropic
    • Claude
    • Coding-Agent
    • Constitutional-AI
    • Dense-Transformer
    • Multimodal
    • Opus-4.5
    • RLHF
    • Safety
    • SWE-bench
  • 2026년 7월 21일

    LLaMA 3: 오픈소스 대규모 언어 모델

    • GQA
    • Grouped-Query-Attention
    • Instruction-Tuning
    • LLaMA-3
    • Meta
    • Open-Source-LLM
    • RLHF

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community