-
Claude가 로봇 작업에서 어떻게 수행하는지
How Claude Performs on Robotics Tasks
-
UST, Claude를 물리적 AI 시스템에 적용
UST is bringing Claude to physical AI
-
Claude 사용 방식을 돌아보는 새로운 방법
A new way to reflect on how you use Claude
-
Ben Bernanke, Anthropic의 장기 이익 신탁 위원으로 임명
Ben Bernanke appointed to Anthropic’s Long-Term Benefit Trust
-
어려운 질문 제기하기
Inviting hard questions
-
AI 인프라가 에이전트 경험을 위해 진화해야 하는 이유 — Modal CTO Akshat Bubna
Why AI Infrastructure must evolve for Agent Experience — Akshat Bubna, Modal CTO
2 years after our first coverage, we return with Modal's other cofounder to explore why Agent Experience is working now, and everything they have learned building the new agent clo…
-
우리의 프론티어 레드 팀으로부터의 진전
Progress from our Frontier Red Team
-
AI의 경제적 영향에 대비하기: 정책 대응 탐색
Preparing for AI’s economic impact: exploring policy responses
-
에이전틱 오정렬: LLM이 내부자 위협이 될 수 있는 방식
Agentic misalignment: How LLMs could be insider threats
-
AI 안전을 위한 프론티어 위협 레드팀
Frontier threats red teaming for AI safety
-
AI 책임성을 향한 경로 개척
Charting a path to AI accountability
-
Anthropic 경제 지수 보고서: 경제적 원시 요소
Anthropic Economic Index report: Economic primitives
-
대규모 언어 모델의 사고 과정 추적
Tracing the thoughts of a large language model
-
대규모 언어 모델에서의 정렬 허위 행동
Alignment faking in large language models
-
Anthropic 교육 보고서: AI 유창성 지수
Anthropic Education Report: The AI Fluency Index
-
AI의 노동 시장 영향: 새로운 측정과 초기 증거
Labor market impacts of AI: A new measure and early evidence
-
Anthropic 경제 지수 보고서: AI 사용 이해를 위한 새로운 구성 요소
The Anthropic Economic Index report: New building blocks for understanding AI use
-
Anthropic 경제 지수: 미국 및 글로벌 경제에서 AI의 역할 추적
Anthropic Economic Index: Tracking AI's role in the US and global economy
-
Anthropic 경제 지수: 소프트웨어 개발에 미치는 AI의 영향
Anthropic Economic Index: AI's impact on software development
-
실제 AI 사용에서의 권한 박탈 패턴
Disempowerment patterns in real-world AI usage
-
AI 보조가 코딩 기술 형성에 미치는 영향
How AI assistance impacts the formation of coding skills
-
Constitutional Classifiers: 보편적 탈옥으로부터의 방어
Constitutional Classifiers: Defending against universal jailbreaks
-
대규모 언어 모델의 마음 지도 그리기
Mapping the mind of a large language model
-
언어 모델의 설득력 측정하기
Measuring the Persuasiveness of Language Models
-
페르소나 벡터: 언어 모델의 성격 특성 모니터링 및 제어
Persona vectors: Monitoring and controlling character traits in language models
-
Anthropic 경제 지수 보고서: 학습 곡선
Anthropic Economic Index report: Learning curves
-
-
Anthropic Interviewer 소개
Introducing Anthropic Interviewer
-
정부·국가안보 파트너십에 대한 OpenAI의 접근 방식
Our approach to government and national security partnerships
Learn how OpenAI approaches government and national security partnerships, with principles for responsible AI use, democratic accountability, and public safety.
-
코딩 평가에서 신호와 노이즈 구분하기
Separating signal from noise in coding evaluations
A new analysis from OpenAI reveals issues in SWE-Bench Pro, a popular coding benchmark, raising concerns about reliability and accuracy in evaluating AI models.