코난쌤 블로그

홈전체 글카테고리소개연락처개인정보처리방침

태그: prompt-injection

6건의 항목

  • 2026년 9월 03일

    SafeEvolve 정리 — 하네스와 정책을 같이 진화시키는 에이전트 안전 정렬

    • agent
    • safety
    • harness
    • RL
    • LLM
    • prompt-injection
  • 2026년 8월 30일

    프레이밍 갭 — 같은 공격인데 문구만 바꾸면 침투율 0%에서 100%로

    • security
    • prompt-injection
    • llm-agents
    • tool-use
    • exfiltration
  • 2026년 8월 28일

    SARA: 도구 증강 LLM 에이전트의 행동 유도와 실행 인가의 분리

    • agent
    • security
    • tool-use
    • llm
    • prompt-injection
    • runtime-authorization
  • 2026년 8월 15일

    ToolHazard: 에이전트 보안 테스트 환경을 자동 합성으로 키우는 프레임워크

    • agent
    • security
    • prompt-injection
    • RL
    • alignment
    • benchmark
    • tool-use
    • LLM
    • evaluation
  • 2026년 7월 28일

    코딩 에이전트에게 악성 이슈를 던지면 어떻게 되는가 — IssueTrojanBench가 폭로한 66.5% 뚫림의 현실

    • coding-agent
    • security
    • prompt-injection
    • LLM
    • agent
    • benchmark
    • tool-use
    • harness
    • automation
  • 2026년 7월 15일

    LLM 에이전트 안전성의 새로운 패러다임: 5개 격리 경계로 보는 시스템 안전

    • LLM-agents
    • agent-safety
    • isolation
    • prompt-injection
    • multi-agent
    • tool-use
    • security

Created with Quartz v4.5.2 © 2026

  • 소개
  • 연락처
  • 개인정보처리방침
  • 전체 글