-
Claude Opus 5: Opus 가격에 Fable 수준의 성능
[AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable)
ain't nobody beats Anthropic at distilling Fable!
-
Claude 5 모델을 위한 새로운 컨텍스트 엔지니어링 규칙
Claude 5 모델을 위한 새로운 컨텍스트 엔지니어링 규칙 | GeekNews
<ul> <li>Claude Opus 5와 Claude Fable 5에서는 Claude Code의 <strong>시스템 프롬프트를 80% 이상 줄이고도</strong> 코딩 평가에서 측정 가능한 성능 저하가 없었음</li> <li>과거 모델의 최악 상황을 막던 세부 규칙은 시스템 프롬프트·Skills·CLAUDE.md·사…
-
GitHub - marcelroed/gigatoken: GB/s 속도의 언어 모델 토큰화
GitHub - marcelroed/gigatoken: Language model tokenization at GB/s
<p>Article URL: <a href="https://github.com/marcelroed/gigatoken/">https://github.com/marcelroed/gigatoken/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=4…
-
Claude Code, Rust로 재작성된 Bun 사용
Claude Code, Rust로 재작성된 Bun 사용 | GeekNews
<ul> <li>Claude Code <strong>v2.1.181</strong>부터 Rust로 포팅된 Bun을 내장해 Linux 시작 속도가 10% 빨라졌지만, 대부분의 사용자는 변화를 거의 알아차리지 못함</li> <li>실행 파일의 문자열을 조사하면 정식 태그가 아직 없는 <strong>Bun v1.4.0</str…
-
Claude Code, Rust로 작성된 Bun 도입
Claude Code uses Bun written in Rust now
<p>Article URL: <a href="https://simonwillison.net/2026/Jul/19/claude-code-in-bun-in-rust/">https://simonwillison.net/2026/Jul/19/claude-code-in-bun-in-rust/</a></p> <p>Comments UR…
-
Kimi K3의 순간
Kimi K3의 순간 | GeekNews
<ul> <li>일상적인 코딩 작업에서 <strong>Kimi K3와 Claude</strong>를 병행한 결과, 출력 품질과 답변에 필요한 토큰 수에서 실질적인 차이를 구별하기 어려웠음</li> <li>Kimi K3 API는 100만 토큰당 입력 <strong>$3</strong>, 출력 $15로, 각각 $10와 $50…
-
코드를 거의 보지 않고 만든 70배 빠른 SQL 파서
코드를 거의 보지 않고 만든 70배 빠른 SQL 파서 | GeekNews
<ul> <li>PostHog는 ANTLR 기반 C++ SQL 파서를 여러 Claude Code 세션으로 재작성해 <strong>16K줄의 Rust 파서</strong>와 5K줄의 도구, 수천 줄의 테스트를 만들었으며 노트북 기준 약 70배의 속도 향상을 얻음</li> <li>새 구현은 <strong>예측형 재귀 하강 파…
-
프로덕션 AI 에이전트를 GPT-5.6으로 전환해 2.2배 빠르고 27% 저렴해진 과정
프로덕션 AI 에이전트를 GPT-5.6으로 전환해 2.2배 빠르고 27% 저렴해진 과정 | GeekNews
<ul> <li>Ploy는 프로덕션 마케팅 웹사이트를 계획·구축·검증하는 에이전트를 Claude Opus 4.8에서 <strong>GPT-5.6 Sol</strong>로 전환하고 모든 워크스페이스의 기본 모델로 지정함</li> <li>평가 하네스의 모델별 가정을 바로잡은 뒤 홈페이지 재구축 작업에서 평균 실행 시간이 <s…
-
CursorBench 3.1 모델 평가 결과
CursorBench 3.1 모델 평가 결과 | GeekNews
<ul> <li>Cursor의 코딩 모델 평가표에서 <strong>Fable 5 Max</strong>가 72.9%로 1위를 기록해, 상위권 경쟁의 기준점이 됨</li> <li><strong>Fable 5 계열</strong>은 Max, Extra High, High, Medium이 1~4위를 모두 차지하며 다른 모델군과…
-
DSpark: Speculative decoding을 활용한 LLM 추론 가속화
DSpark: Speculative decoding을 활용한 LLM 추론 가속화 [pdf] | GeekNews
<ul> <li>DSpark: 준자기회귀(semi-autoregressive) 생성과 신뢰도 스케줄링을 결합한 추측 디코딩(speculative decoding) 프레임워크</li> <li><strong>병렬 드래프터(parallel drafter)</strong> 가 한 번의 순전파로 긴 토큰 블록을 제안하지만 토큰 간…
-
OpenAI와 Broadcom, LLM 최적화 추론 칩 공개
OpenAI and Broadcom unveil LLM-optimized inference chip
OpenAI and Broadcom introduce Jalapeño, a custom AI chip built for LLM inference to improve performance, efficiency, and scale across AI systems.
-
Codex SQLite 피드백 로그의 연간 640TB 기록과 SSD 내구성 문제
Codex SQLite feedback logs can write ~640 TB/year and rapidly consume SSD endurance · Issue #28224 · openai/codex
<p>Article URL: <a href="https://github.com/openai/codex/issues/28224">https://github.com/openai/codex/issues/28224</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/i…
-
DuckDB 내부 구조: DuckDB는 왜 빠른가? (1부)
DuckDB Internals: Why Is DuckDB Fast? (Part 1)
<p>Article URL: <a href="https://www.greybeam.ai/blog/duckdb-internals-part-1">https://www.greybeam.ai/blog/duckdb-internals-part-1</a></p> <p>Comments URL: <a href="https://news.y…
-
Claude Fable 5: 신화급 과대광고, 벤치마크 조작, 그리고 명예의 전당 항목들
Claude Fable 5: Mythos-grade hype, record cheating, and a few hall-of-fame entries | Blog | Endor Labs
<p>Article URL: <a href="https://www.endorlabs.com/learn/claude-fable-5-mythos-grade-hype">https://www.endorlabs.com/learn/claude-fable-5-mythos-grade-hype</a></p> <p>Comments URL:…
-
MiMo-V2.5-Pro-UltraSpeed: 초당 1000토큰을 생성하는 1T 모델
MiMo-V2.5-Pro-UltraSpeed: 초당 1000토큰을 생성하는 1T 모델 | GeekNews
<ul> <li><strong>1조(1T) 파라미터 모델</strong>에서 디코딩 속도 <strong>1000 tokens/s</strong>를 처음으로 돌파한 모델</li> <li>전용 하드웨어가 아닌 <strong>commodity GPU</strong>만으로 속도를 달성했으며, 단일 표준 <strong>8-GPU …