도형준 지식 볼트
Search
검색
다크 모드
라이트 모드
탐색기
태그: RLHF
6건의 항목
2026년 7월 21일
Claude (1~3.5 시리즈): Anthropic의 Constitutional AI 기반 대형 언어 모델
Alignment
Anthropic
Claude
Constitutional-AI
Harmless-AI
Multimodal
RLAIF
RLHF
Safety
2026년 7월 21일
GPT-5 업데이트: 지속적 개선과 에이전트 강화
Agentic-AI
Continual-Learning
GPT-5
GPT-5.2
Model-Updates
OpenAI
RLAIF
RLHF
Safety-Alignment
2026년 7월 21일
Claude Opus 4: Anthropic의 에이전틱 AI 플래그십 모델
Agentic-AI
Anthropic
ASL-3
Claude
Constitutional-AI
Opus-4
RLHF
Safety
Safety-Evaluation
SWE-bench
2026년 7월 21일
GPT-4: 멀티모달 대형 언어 모델
128K-Context
Function-Calling
GPT-4
Mixture-of-Experts
MMLU
Multimodal
OpenAI
RLHF
2026년 7월 21일
Claude Opus 4.5: SWE-bench 80.9%를 달성한 Anthropic의 최강 에이전트 모델
Agentic-AI
Anthropic
Claude
Coding-Agent
Constitutional-AI
Dense-Transformer
Multimodal
Opus-4.5
RLHF
Safety
SWE-bench
2026년 7월 21일
LLaMA 3: 오픈소스 대규모 언어 모델
GQA
Grouped-Query-Attention
Instruction-Tuning
LLaMA-3
Meta
Open-Source-LLM
RLHF