Docs faviconDOCS우성짱의 문서
전체YouTubeArticleTagsAuthorsHub
홈/태그 찾기/#ai-red-teaming
Tag2건YouTube 1Article 1

#ai-red-teaming

이 태그와 연결된 문서를 한곳에서 모아보고, 함께 자주 등장하는 연관 태그까지 이어서 탐색할 수 있습니다.

연관 태그

#agent-network-safety공동문서 1 · 연관도 71%#agent-reputation-manipulation공동문서 1 · 연관도 71%#agentic-cost-exhaustion공동문서 1 · 연관도 71%#multi-agent-security공동문서 1 · 연관도 71%#multiturn-guardrail-bypass공동문서 1 · 연관도 71%#networked-agent-risk공동문서 1 · 연관도 71%#reputation-system-abuse공동문서 1 · 연관도 71%#safety-usability-balance공동문서 1 · 연관도 71%#social-proof-exploitation공동문서 1 · 연관도 71%#attack-demonstration공동문서 1 · 연관도 50%
이렇게까지 창의적으로 AI를 타락시켜?" (강수진 박사)
YouTube2026년 8월 25일

이렇게까지 창의적으로 AI를 타락시켜?" (강수진 박사)

AI를 이렇게까지 창의적으로 타락시키는 핵심 수법은 지시를 쪼개고 맥락과 형식을 바꾸는 프롬프트 인젝션이며, 이를 막으려면 실제 공격 데이터에 기반한 지속적인 레드티밍과 정교한 가드레일 조율이 필요하다.

티타임즈TV
#llm-security#prompt-injection-defense#ai-red-teaming#multiturn-guardrail-bypass
Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale
Article2026년 4월 30일

Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale

Microsoft 연구진은 100개 이상의 내부 AI 에이전트가 상호작용하는 플랫폼을 레드팀 테스트해, 단일 에이전트 평가로는 드러나지 않는 네트워크 수준의 전파형 공격·평판 조작·방어 징후를 확인했다.

Microsoft
#gpt-4o#gpt-5#microsoft-research#gpt-4-1