코난쌤 블로그
Search
검색
다크 모드
라이트 모드
탐색기
홈
전체 글
카테고리
소개
연락처
개인정보처리방침
태그: LLM-agent
5건의 항목
2026년 7월 24일
JANUS: 에이전트가 위험해지기 전에 미리 본다 — 긴 호라이즌 에이전트 안전을 위한 예측형 가드레일
agent-safety
guardrails
reinforcement-learning
long-horizon
LLM-agent
2026년 7월 19일
GRASP: 강화학습으로 에이전트 RAG의 검색 도구를 자유자재로 다루게 만드는 방법
agentic-rag
reinforcement-learning
retrieval-augmented-generation
LLM-agent
tool-use
multi-hop-reasoning
GRPO
search-policy
2026년 7월 17일
SEED: 에이전트 RL이 완료된 궤적에서 스스로 배우는 자가진화 증류
agentic-rl
on-policy-distillation
self-evolving
hindsight-learning
LLM-agent
2026년 7월 13일
리더보드 너머: LLM 에이전트의 6대 실패 클러스터 — 도구·계획·추론 실패 종합 분석
LLM-agent
failure-taxonomy
tool-use
planning
multi-agent
benchmark
survey
2026년 7월 12일
행동 상태 붕괴를 막아라 — Meta의 사기억 에이전트(Proactive Memory Agent) 완전 해부
LLM-agent
memory
long-horizon
behavioral-state-decay
Meta-AI
intervention