-
2,000명이 내 AI 어시스턴트를 해킹하려 했을 때 일어난 일
What happened after 2,000 people tried to hack my AI assistant — Fernando Irarrázaval
<p>Article URL: <a href="https://www.fernandoi.cl/posts/hackmyclaw/">https://www.fernandoi.cl/posts/hackmyclaw/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?…
-
악성코드 개발자들이 스파이웨어에 핵·생물무기 문구를 추가함
악성코드 개발자들이 스파이웨어에 핵·생물무기 문구를 추가함 | GeekNews
<ul> <li>AI 보안 스캐너의 분석을 막기 위해 스파이웨어에 <strong>LLM 안전 거부</strong>를 유발하는 핵·생물무기 문구가 삽입됨</li> <li><strong>1차 안전 정렬</strong>에 과도하게 의존하면 실제 보안 분석에서 공격자가 악용할 수 있는 맹점이 생김</li> <li>폐쇄형 모델과 …
-
임포트 AI 457: AI 스턱스넷; 저주받은 뮤온 최적화기; 긍정적 정렬
Import AI 457: AI stuxnet; cursed Muon optimizer; and positive alignment
Welcome to Import AI, a newsletter about AI research.
-
Import AI 441: 내 에이전트는 작동 중이야. 너의는?
Import AI 441: My agents are working. Are yours?
Plus: Corrupting AI systems with a poison fountain