-
GPT-Red: 견고성을 위한 자가 개선의 활용
GPT-Red: Unlocking Self-Improvement for Robustness
Explore GPT-Red, OpenAI’s automated red teaming system that uses self-play to improve AI safety, alignment, and prompt injection robustness.
-
AI 안전을 위한 프론티어 위협 레드팀
Frontier threats red teaming for AI safety
-
신화 이후의 레드팀
Red-Teaming after Mythos — Zico Kolter & Matt Fredrikson, Gray Swan
OpenAI boardmember Zico Kolter and Gray Swan CEO Matt Fredrikson join swyx to explain why AI security is not just “cybersecurity with AI”
-
취약한 앱을 직접 만들고 LLM이 해킹할 수 있는지 $1,500을 들여 테스트해봤다
I built a vulnerable app and spent $1,500 seeing if LLMs could hack it
<p>Article URL: <a href="https://kasra.blog/blog/i-spent-1500-seeing-if-llms-could-hack-my-app/">https://kasra.blog/blog/i-spent-1500-seeing-if-llms-could-hack-my-app/</a></p> <p>C…