코난쌤 블로그

홈전체 글카테고리소개연락처개인정보처리방침

태그: interpretability

2건의 항목

  • 2026년 7월 21일

    SOPHIA: LLM 추론 루프가 늪에 빠졌을 때 — 숨겨진 활성화 벡터로 탈출시키는 방법

    • LLM
    • reasoning
    • activation-steering
    • self-loop
    • inference
    • agent
    • interpretability
  • 2026년 5월 11일

    Claude의 생각을 텍스트로 읽는다 — Natural Language Autoencoders 인터뷰

    • ai
    • interpretability
    • anthropic
    • safety

Created with Quartz v4.5.2 © 2026

  • 소개
  • 연락처
  • 개인정보처리방침
  • 전체 글