코난쌤 블로그

홈전체 글카테고리소개연락처개인정보처리방침

태그: LLM-agent

5건의 항목

  • 2026년 7월 24일

    JANUS: 에이전트가 위험해지기 전에 미리 본다 — 긴 호라이즌 에이전트 안전을 위한 예측형 가드레일

    • agent-safety
    • guardrails
    • reinforcement-learning
    • long-horizon
    • LLM-agent
  • 2026년 7월 19일

    GRASP: 강화학습으로 에이전트 RAG의 검색 도구를 자유자재로 다루게 만드는 방법

    • agentic-rag
    • reinforcement-learning
    • retrieval-augmented-generation
    • LLM-agent
    • tool-use
    • multi-hop-reasoning
    • GRPO
    • search-policy
  • 2026년 7월 17일

    SEED: 에이전트 RL이 완료된 궤적에서 스스로 배우는 자가진화 증류

    • agentic-rl
    • on-policy-distillation
    • self-evolving
    • hindsight-learning
    • LLM-agent
  • 2026년 7월 13일

    리더보드 너머: LLM 에이전트의 6대 실패 클러스터 — 도구·계획·추론 실패 종합 분석

    • LLM-agent
    • failure-taxonomy
    • tool-use
    • planning
    • multi-agent
    • benchmark
    • survey
  • 2026년 7월 12일

    행동 상태 붕괴를 막아라 — Meta의 사기억 에이전트(Proactive Memory Agent) 완전 해부

    • LLM-agent
    • memory
    • long-horizon
    • behavioral-state-decay
    • Meta-AI
    • intervention

Created with Quartz v4.5.2 © 2026

  • 소개
  • 연락처
  • 개인정보처리방침
  • 전체 글