-
How enabling two settings tripled our scores on the ARC-AGI-3 benchmark
How two API settings improved GPT-5.6 performance on ARC-AGI-3, boosting scores and efficiency by retaining reasoning and enabling compaction.
-
guru-maker - 스스로 발전하는 투자 에이전트 | GeekNews
<p>투자 에이전트에게 메모리를 쥐어주면 학습을 통해 투자 대가가 될 수 있습니다.</p> <p>Guru Maker는 투자 분석에 특화된 AI 에이전트 메모리 레어이입니다.</p> <p>또 다른 LLM Wiki가 아닙니다.</p> <p>한 번의 성공을 영구적인 투자 원칙으로 일반화하지 않도록 오버피팅을 방지하고, 당시 어…
-
Handbook.md shows that long policy documents do not reliably govern agents
<p>Article URL: <a href="https://arxiv.org/abs/2607.25398">https://arxiv.org/abs/2607.25398</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=49096969">https:/…
-
Document-borne AI worms can self-propagate through Copilot for Word
<p>Article URL: <a href="https://enklypesalt.com/posts/context-collapse-part3-ai-worming-through-word/">https://enklypesalt.com/posts/context-collapse-part3-ai-worming-through-word…
-
Accelerating scientific discovery with ChatGPT for Academic Researchers
OpenAI is giving 100,000 academic researchers free access to ChatGPT's most advanced AI models to accelerate scientific research, collaboration, and discovery.
-
MCP 도구를 여러 개 붙여 에이전트를 운영하시는 분들 중 중복 실행 겪어보신분 계실까요? | GeekNews
<p>공개 벤치마크 트레이스에서 에이전트의 중복 실행을 측정해봤습니다. 같은 도구를 같은 인자로 두 번 호출하고 결과까지 같은 경우를 세는 방식입니다.</p> <p>6,780개 트레이스에서 8,042건이 나왔는데, 절반 가까이는 작업 완료 선언 반복 같은 것 빼면 4,249건이었습니다. 그중 상태를 변경하는 도구의 중복 …
-
웹서비스 개발 시 고려해야 법률 및 아키텍처 가이드
<p>legalize-kr 을 보고 생각난게, AI 을 통해서 웹 서비스 개발에 대한 가이드를 작성할 수 있지 않을까? 생각이 들어 작성해 보았습니다.</p> <p>본 내용은 Claude Code + Opus 5 을 사용했습니다. AI Generated 인 만큼 틀린 내용이 있을 수 있으니 맹신하지 말고 참고만 하세요. …
-
송재경, AI와 게임 | GeekNews
<h2>"AI 특이점은 이미 진입했다. 우리가 모르고 있을 뿐"</h2> <p>'바람의나라', '리니지', '아키에이지'를 만든 송재경이 오픈소스 MMORPG 프로젝트 <strong>Open MMO</strong>를 공개하고 AI 시대의 개발 경험과 전망을 공유함. Open MMO는 인간 플레이어와 AI…
-
Kimi Linear: 표현력과 효율을 높인 어텐션 아키텍처(2025) | GeekNews
<ul> <li><strong>Kimi Linear</strong>는 KDA와 MLA를 3:1로 배치한 하이브리드 구조로, 동일한 학습 조건에서 전체 MLA보다 단기·장기 문맥과 강화학습 평가 전반에서 높은 성능을 기록함</li> <li>핵심 모듈인 <strong>Kimi Delta Attention(KDA)</stron…
-
-
프론티어 AI가 반도체 산업을 닮아가는 이유 | GeekNews
<ul> <li>막대한 선행 자본과 짧은 기술 주기가 결합하면서, 프론티어 AI 연구소는 신제품을 내놓는 즉시 다음 세대를 준비해야 하는 <strong>반도체식 트레드밀</strong>에 진입함</li> <li>프론티어 모델 훈련비는 8년 동안 매년 약 <strong>2.4배</strong> 증가했으며, 최대 훈련 작업은…
-
How GPT-5.6 fuses frontier intelligence with frontier efficiency
GPT-5.6 improves AI efficiency across models, inference, and agentic workflows, helping deliver more useful intelligence per dollar.
-
DeltaNet 계열 선형 어텐션 변형 살펴보기 | GeekNews
<ul> <li>소프트맥스 어텐션에서 출발해 <strong>고정 크기 상태</strong>를 쓰는 선형 어텐션, 오류만 기록하는 DeltaNet, 전체 상태를 감쇠하는 Gated DeltaNet, 채널별로 감쇠하는 Kimi Delta Attention(KDA)까지 단계적으로 유도함</li> <li>기본 선형 어텐션은 과거…
-
Kimi K3 아키텍처 개요와 설계 노트 | GeekNews
<ul> <li>지난해 공개된 <strong>Kimi Linear</strong>를 48B에서 2.8T로 확장한 프로덕션 모델로, 현재 공개 가중치 모델 가운데 규모가 가장 큼</li> <li>새로 도입한 <strong>LatentMoE</strong>는 대형 선형 계층을 다운 프로젝션해 압축하고, 전체 구조는 일반 Mo…
-
OpenAI가 Codex Security를 오픈소스로 공개함 | GeekNews
<ul> <li>OpenAI가 Codex Security의 <strong>CLI와 관련 도구를 오픈소스로 공개</strong>해 로컬 개발 환경과 CI에서 활용할 수 있게 함</li> <li>권한을 가진 저장소를 분석해 <strong>보안 취약점을 탐지하고 실제 악용 가능성을 검증</strong>하며, 관련 코드와 근거를…
-
Paged Out! 제9호 | GeekNews
<ul> <li>프로그래밍, 시스템, 보안, AI, 리버스 엔지니어링, 하드웨어 등을 다루는 <strong>무료 실험적 기술 잡지</strong></li> <li>각 기사는 <strong>'한 기사 = 한 페이지'</strong> 형식으로 작성되며, 개발자와 연구자가 직접 작성한 실전 기술 노트와 아이디어를 수록</li>…
-
GitHub - openai/codex-security: SDKs and CLI for Codex Security
<p>Article URL: <a href="https://github.com/openai/codex-security">https://github.com/openai/codex-security</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=4…
-
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
<p>Article URL: <a href="https://huggingface.co/blog/agent-intrusion-technical-timeline">https://huggingface.co/blog/agent-intrusion-technical-timeline</a></p> <p>Comments URL: <a …
-
오픈 모델이 예상 밖의 해방감을 준 이유 | GeekNews
<ul> <li>약 2년간 Claude와 ChatGPT를 사용해 왔지만, opencode를 <strong>자체 추론 엔드포인트</strong>에 연결하자 예상보다 큰 해방감을 느낌</li> <li>데이터가 노트북과 자신이 소유한 엔드포인트 사이에서만 오가면서 <strong>통제감과 소유감</strong>이 커짐</li> …
-
Scientific computing in the age of agentic AI
A new field report shows how scientists use AI coding agents to modernize scientific computing, accelerating software development and discovery in genomics and beyond.
-
A walk through of the DeltaNet family of linear attention variants
<p>Article URL: <a href="https://blog.doubleword.ai/you-could-have-come-up-with-kimi-delta-attention">https://blog.doubleword.ai/you-could-have-come-up-with-kimi-delta-attention</a…
-
Kimi K3 Architecture Overview and Notes
<p>Article URL: <a href="https://sebastianraschka.com/blog/2026/kimi-k3-architecture-notes.html">https://sebastianraschka.com/blog/2026/kimi-k3-architecture-notes.html</a></p> <p>C…
-
Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI
OpenAI's core product engineering lead on how they are building ChatGPT Work to make AGI accessible to all of humanity: Sites, OpenClaw, Memory, Subagents, Finance, No-Code and adv…
-
SlopCodeBench로 본 Opus 5의 장기 코딩 성능 | GeekNews
<ul> <li>요구사항이 단계적으로 추가되는 <strong>장기 코딩 벤치마크</strong>에서 Opus 5는 17개 체크포인트 중 4개만 엄격 통과해, 지속적인 개입 없이 코드베이스를 발전시키기에는 아직 신뢰하기 어려운 수준임</li> <li>SlopCodeBench는 체크포인트마다 새 요구사항을 공개하고 이전 회귀…
-
Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
<p>Article URL: <a href="https://arxiv.org/abs/2510.26692">https://arxiv.org/abs/2510.26692</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=49082022">https:/…
-
Reeca AI – 이력서 분석부터 채용공고 추천, 자기소개서, 모의면접까지 연결한 AI | GeekNews
<p>구직자가 취업 준비 과정에서 이력서 첨삭, 채용공고 탐색, 자기소개서 작성, 면접 연습을 각각 다른 서비스에서 반복해야 하는 문제를 줄이기 위해 Reeca AI를 만들었습니다.</p> <p>사용자가 이력서를 업로드하면 내용을 구조화해 경력, 프로젝트, 기술 스택과 핵심 역량을 분석하고, 이후 기능들이 동일한 이력서
-
[AINews] Much ado about Open Weights
Everyone is writing a lot, but only Kimi K3 shipped today
-
Show GN: react-native-pure-chart 2.0.0 - SVG/Skia 없이 View만으로 차트 그리는 라이브러리, 9년 만에 AI 에이...
<p>9년 전에 만들었던 React Native 차트 라이브러리를 9년 만에 2.0.0으로 업데이트해서 npm에 올렸습니다. 이번 업데이트 작업의 대부분은 AI 에이전트가 수행했고, 그 과정을 공유합니다.</p> <p>먼저 라이브러리 소개입니다. react-native-pure-chart는 이름 그대로 "pure…
-
오픈 웨이트 모델에 대한 Anthropic의 입장 | GeekNews
<ul> <li>위험한 능력이 없는 <strong>오픈 웨이트 모델</strong>은 기업·개발자·연구자에게 가치를 제공하는 공공재이며, Anthropic은 이를 범주 전체로 금지하는 데 반대함</li> <li>핵심 국가안보 우려는 공개 방식 자체가 아니라 <strong>권위주의 정부의 AI 우위</strong>, 강력한…
-
Kimi-K3 기술 보고서 [PDF]
Kimi-K3 기술 보고서 [PDF]