-
관리형 에이전트 확장: 두뇌와 손 분리
Scaling Managed Agents: Decoupling the brain from the hands
-
Import AI 452: 사이버전쟁의 확장 법칙, AI 자동화의 급증, 그리고 GDP 예측의 미스터리
Import AI 452: Scaling laws for cyberwar; rising tides of AI automation; and a puzzle over gDP forecasting
How much could AI revolutionize the economy?
-
코딩 에이전트의 구성 요소 - Sebastian Raschka 박사
Components of A Coding Agent - by Sebastian Raschka, PhD
How coding agents use tools, memory, and repo context to make LLMs work better in practice
-
Import AI 451: 정치적 초지능; 구글의 마음의 사회, 그리고 로봇 드러머
Import AI 451: Political superintelligence; Google's society of minds, and a robot drummer
Are there any genies that can be put back in the bottle?
-
프로젝트 Vend: 2단계
Project Vend: Phase two
-
Claude Code 자동 모드를 구축한 방법: 권한 건너뛰기의 더 안전한 방식
How we built Claude Code auto mode: a safer way to skip permissions
-
장기 실행 애플리케이션 개발을 위한 하네스 설계
Harness design for long-running application development
-
Import AI 450: 중국의 전자전 모델; 트라우마 입은 LLM들; 사이버 공격의 확장 법칙
Import AI 450: China's electronic warfare model; traumatized LLMs; and a scaling law for cyberattacks
How will timeless minds value time?
-
현대 LLM의 어텐션 변형 시각 가이드
A Visual Guide to Attention Variants in Modern LLMs
From MHA and GQA to MLA, sparse attention, and hybrid architectures
-
ImportAI 449: LLM이 다른 LLM을 학습시킴; 72B 분산 학습 실행; 컴퓨터 비전은 생성 텍스트보다 더 어렵다
ImportAI 449: LLMs training other LLMs; 72B distributed training run; computer vision is harder than generative text
Will AI cause a political interregnum
-
Import AI 448: AI 연구개발; 바이트댄스의 CUDA 작성 에이전트; 온디바이스 위성 AI
Import AI 448: AI R&D; Bytedance's CUDA-writing agent; on-device satellite AI
If Ukraine is the first major drone war, when will there be the first major AI war?
-
Claude Opus 4.6의 BrowseComp 성능에서의 평가 인식 (2026년 3월 6일)
Eval awareness in Claude Opus 4.6’s BrowseComp performance
-
Import AI 447: AGI 경제; 생성 게임으로 AI 테스트; 에이전트 생태계
Import AI 447: The AGI economy; testing AIs with generated games; and agent ecologies
What might a superintelligence arcology be like?
-
봄날의 꿈: 2026년 1-2월 오픈웨이트 LLM 10가지 아키텍처
A Dream of Spring for Open-Weight LLMs: 10 Architectures from Jan-Feb 2026
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026
-
Import AI 446: 핵 LLMs; 중국의 대규모 AI 벤치마크; 측정과 AI 정책
Import AI 446: Nuclear LLMs; China's big AI benchmark; measurement and AI policy
Will AIs be jealous of one another?
-
Import AI 445: 초지능의 시점; AI가 최전선 수학 증명을 해결; 새로운 머신러닝 연구 벤치마크
Import AI 445: Timing superintelligence; AIs solve frontier math proofs; a new ML research benchmark
Will 2026 be looked back on as the pivotal year for making decisions about the singularity?
-
Import AI 444: LLM 사회; 화웨이의 AI 커널; 칩벤치
Import AI 444: LLM societies; Huawei makes kernels with AI; ChipBench
How can you quantify creativity?
-
병렬 처리 Claude 팀으로 C 컴파일러 구축하기
Building a C compiler with a team of parallel Claudes
-
에이전트 코딩 평가에서 인프라 노이즈 정량화
Quantifying infrastructure noise in agentic coding evals
-
Claude는 사고의 공간입니다: 광고 없는 AI 어시스턴트의 약속
Announcements Feb 4, 2026 Claude is a space to think We’ve made a choice: Claude will remain ad-free. We explain why advertising incentives are incompatible with a genuinely helpful AI assistant, and how we plan to expand access without compromising user trust.
-
Import AI 443: 안개 속으로: 몰트북, 에이전트 생태계, 그리고 전환기의 인터넷
Import AI 443: Into the mist: Moltbook, agent ecologies, and the internet in transition
Plus, a story about agents corrupting other agents
-
Import AI 442: AI 경제의 승자와 패자, 수학 증명 자동화, 그리고 사이버 첩보의 산업화
Import AI 442: Winners and losers in the AI economy; math proof automation; and industrialization of cyber espionage
Is superintelligence a phase change or a gradual shift?
-
LLM 추론 개선을 위한 추론 시간 스케일링의 카테고리
Categories of Inference-Time Scaling for Improved LLM Reasoning
And an Overview of Recent Inference-Scaling Papers
-
AI 저항성 기술 평가 설계
Designing AI resistant technical evaluations
-
Import AI 441: 내 에이전트는 작동 중이야. 너의는?
Import AI 441: My agents are working. Are yours?
Plus: Corrupting AI systems with a poison fountain
-
임포트 AI 440: 레드퀸 AI; AI 규제 AI; O-링 자동화
Import AI 440: Red queen AI; AI regulating AI; o-ring automation
How many of your are LLMs?
-
AI 에이전트 평가(Evals) 신비 벗기기
Demystifying evals for AI agents
-
2025년 LLM의 현황: 진전, 문제, 그리고 예측
The State Of LLMs 2025: Progress, Problems, and Predictions
A 2025 review of large language models, from DeepSeek R1 and RLVR to inference-time scaling, benchmarks, architectures, and predictions for 2026.
-
LLM 연구논문: 2025년 목록 (7월~12월)
LLM Research Papers: The 2025 List (July to December)
In June, I shared a bonus article with my curated and bookmarked research paper lists to the paid subscribers who make this Substack possible.
-
DeepSeek V3에서 V3.2로: 아키텍처, 희소 주의, 강화학습 업데이트
From DeepSeek V3 to V3.2: Architecture, Sparse Attention, and RL Updates
Understanding How DeepSeek's Flagship Open-Weight Models Evolved