-
Echo – 오픈웨이트 모델로 Fable 수준의 결과를 1/3 비용으로 달성
Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models | Hacker News
<p>I’ve been building Echo (<a href="https://echo.tracerml.ai/" rel="nofollow">https://echo.tracerml.ai/</a>), an experiment in making one AI system out of a pool of open-weight mo…
-
코딩 에이전트를 위한 고급 컨텍스트 엔지니어링
advanced-context-engineering-for-coding-agents/wsff.md at main · humanlayer/advanced-context-engineering-for-coding-agents
<p>Article URL: <a href="https://github.com/humanlayer/advanced-context-engineering-for-coding-agents/blob/main/wsff.md">https://github.com/humanlayer/advanced-context-engineering-…
-
GitHub - marcelroed/gigatoken: GB/s 속도의 언어 모델 토큰화
GitHub - marcelroed/gigatoken: Language model tokenization at GB/s
<p>Article URL: <a href="https://github.com/marcelroed/gigatoken/">https://github.com/marcelroed/gigatoken/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=4…
-
Laguna S 2.1 공개
Laguna S 2.1 공개 | GeekNews
<ul> <li>Poolside가 장기 작업과 추론 능력을 강화한 <strong>Laguna S 2.1</strong>을 공개함. 총 118B MoE 중 토큰당 8B 매개변수를 활성화하며, thinking·no-thinking 모드 모두 최대 <strong>1M 토큰 컨텍스트</strong>를 지원함</li> <li>학습…
-
태스크 이코노미 - 데이터가 만드는 다음 1조 달러 시장
태스크 이코노미 - 데이터가 만드는 다음 1조 달러 시장 | GeekNews
<ul> <li>AI 사용량을 나타내는 <strong>토큰</strong>처럼, 모델 능력을 개선하는 데이터 작업 단위인 <strong>태스크(task)</strong> 가 AI 투자를 측정하는 대규모 시장을 형성할 전망임</li> <li>인터넷 학습 데이터와 단순·범용 능력이 포화되면서, 실제 환경에서 전문가의 암묵지를…
-
AI 코드 검토가 타당한 반론이 될 수 없는 이유
AI 코드 검토가 타당한 반론이 될 수 없는 이유 | GeekNews
<ul> <li>LLM 코딩 도구의 잦은 오류를 사람이 모두 검토하면 된다는 해법은 <strong>코드 리뷰의 처리 한계</strong> 때문에 품질과 생산성을 함께 보장하기 어려움</li> <li>경험적 연구에 따르면 효과적인 리뷰는 한 번에 <strong>1시간·400 LOC</strong> 정도가 상한이며, 이를 넘…
-
안녕히, 그동안의 모든 Bikeshed에 감사드립니다
안녕히, 그동안의 모든 Bikeshed에 감사드립니다 | GeekNews
<ul> <li>FOSS의 미래를 바꿀 두 요인으로 <strong>LLM 기반 코드 리뷰</strong>와 <strong>연령 확인</strong>을 꼽으며, 특히 연령 확인이 현재 형태의 FOSS를 끝내는 계기가 될 것으로 내다봄</li> <li>LLM 코드 리뷰는 사람이 살피기 어려운 탐색 공간을 더 넓고 깊게 조사하…
-
LLM은 사랑하지만 과대광고는 싫다
LLM은 사랑하지만 과대광고는 싫다 | GeekNews
<ul> <li>LLM, 자율주행차, 영상 생성 모델, 코딩 에이전트는 실제로 유용하고 흥미롭지만, 이를 둘러싼 <strong>공포와 종말론적 과대광고</strong>에는 동의하지 않음</li> <li>기회의 창이 닫히면 영구적 하층민이 된다는 담론은 사실이라기보다 사람을 불안하게 만들어 <strong>San Franci…
-
Mesh LLM: iroh에서의 분산 AI 컴퓨팅
Mesh LLM: distributed AI computing on iroh
<p>Article URL: <a href="https://www.iroh.computer/blog/mesh-llm">https://www.iroh.computer/blog/mesh-llm</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=488…
-
오픈 웨이트 LLM과 폐쇄형 LLM의 격차
오픈 웨이트 LLM과 폐쇄형 LLM의 격차 | GeekNews
<ul> <li>Artificial Analysis Intelligence Index에서는 <strong>오픈 웨이트 LLM</strong>이 폐쇄형 LLM의 과거 성능을 따라잡는 시간이 2024년 여름부터 꾸준히 줄어드는 흐름을 보임</li> <li>이 단일 지표에 추세선을 그으면 격차가 <strong>2026년 12월…
-
Show GN: 『내 문장이 그렇게 이상한가요?』에서 영감을 받아 grep으로 대충 훑지 않고 모든 문장을 LLM이 읽고 ...
<p>Claude code나 codex에서 글을 교정하라고 시키면 grep으로 읽어 단어를 빠뜨리곤 합니다.<br /> 그래서 모든 문장을 구분자 단위로 강제로 LLM한테 읽혀서 윤문하는 스킬을 만들어봤어요.</p> <p>im-not-ai 스킬과 『내 문장이 그렇게 이상한가요?』에서 영감을 받아 적의를 가진 것들 등 사람…
-
Apertus – 주권 AI를 위한 오픈 파운데이션 모델
Apertus – Open Foundation Model for Sovereign AI
<p>Article URL: <a href="https://apertvs.ai/">https://apertvs.ai/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48622778">https://news.ycombinator.com/item…
-
LLM이 작성한 인시던트 보고서의 미래가 두렵다
LLM이 작성한 인시던트 보고서의 미래가 두렵다 | GeekNews
<ul> <li>인시던트 보고서에서 <strong>LLM은 자료 수집과 정리</strong>를 돕는 데 유용하지만, 보고서 본문까지 맡기면 검증 과정이 약해짐</li> <li>직접 쓰는 과정은 <strong>증거와 설명의 일관성을 확인</strong>하게 만들며, <strong>글쓰기 자체가 이해 부족을 드러내는 장치</…
-
당신은 가중치 속에 있나요?
Show HN: Are You in the Weights?
<p>With more traffic moving off-web and into LLMs, I got curious about what traces we leave "in the weights". My design partner and I built a site in the past few weeks that checks…
-
LLM이 발견한 0-day 취약점
LLM-discovered 0 days
-
LLM을 위한 사이버 도구 모음
Cyber toolkits for LLMs
-
LLMs과 생물 위험
LLMs and biorisk
-
질문: Claude/GPT를 로컬 모델로 대체했나요? (일일 코딩용)
Ask HN: Has anyone replaced Claude/GPT with a local model for daily coding? | Hacker News
<p>Has anyone here fully swapped Claude/GPT for a local model as their main coding tool, not just for side experiments? If so, please share your setup and performance (e.g tok/s)</…
-
리우데자네이루의 "자체 개발" LLM이 기존 모델의 병합으로 보임
리우데자네이루의 “자체 개발” LLM이 기존 모델의 병합으로 보임 | GeekNews
<ul> <li>GitHub 상태는 <strong>Open</strong>이며, <a href="https://huggingface.co/prefeitura-rio/Rio-3.5-Open-397B/commit/a778c1ec4e21180ee55c3ea016a348e549e75f09">a778c1ec4e21180ee55c3…
-
애플 파운데이션 모델
Apple Foundation Models
<p>Article URL: <a href="https://platform.claude.com/docs/en/cli-sdks-libraries/libraries/apple-foundation-models">https://platform.claude.com/docs/en/cli-sdks-libraries/libraries/…
-
repo-slopscore: 커밋 기록 분석으로 Git 저장소의 AI/LLM 기여 감지
repo-slopscore: 커밋 기록 분석으로 Git 저장소의 AI/LLM 기여 감지 | GeekNews
<ul> <li><strong>repo-slopscore</strong>는 Git 저장소의 커밋 기록을 분석해 AI/LLM 기여를 감지하는 도구로 소개됨</li> <li>서비스는 홈, <strong>저장소 스캔</strong>, 소스 코드 링크를 제공하며 소스 코드는 codeberg.org/polyphony/repo-sl…
-
우리 직장의 LLM 집단 망상
우리 직장의 LLM 집단 망상 | GeekNews
<ul> <li><strong>자금난</strong>에 시달리는 직장에서 핵심 업무 예산은 깎이는데도 <strong>AI 도입</strong>에는 돈이 흘러가는 모순적 상황을 직접 겪은 경험 기록</li> <li>수년간 보너스가 취소되고 인력·라이선스·데이터베이스가 삭감된 와중에 <strong>컨설턴트, LLM 워크숍, …
-
DiffusionGemma: 4배 빠른 텍스트 생성
DiffusionGemma: 4배 빠른 텍스트 생성 | GeekNews
<ul> <li><strong>DiffusionGemma</strong>는 텍스트 확산 방식으로 전체 텍스트 블록을 동시에 생성하는 Apache 2.0 라이선스의 26B MoE 실험용 공개 모델임</li> <li>일반적인 자기회귀 LLM의 순차적 토큰 생성 대신 <strong>256토큰 병렬 생성</strong>을 사용해…
-
Mythos와 일하는 느낌은 이렇습니다
Mythos와 일하는 느낌은 이렇습니다 | GeekNews
<ul> <li>일반 공개된 첫 <strong>Mythos급 모델 Claude 5 Fable</strong>은 다단계 사양서를 받아 최대 십수 시간 동안 스스로 작업을 수행하며, 이전에 사용해 본 모든 모델을 상당한 격차로 능가</li> <li>단일 프롬프트와 한 차례 피드백만으로 <strong>정교한 사회과학 논문</s…
-
Tokenomics: 에이전트형 소프트웨어 엔지니어링에서 토큰이 어디에 사용되는지 정량화
Tokenomics: 에이전트형 소프트웨어 엔지니어링에서 토큰이 어디에 사용되는지 정량화 | GeekNews
<ul> <li>LLM 기반 다중 에이전트 소프트웨어 개발 시스템의 실행 추적을 SDLC 단계에 매핑해, 토큰 소비가 초기 생성보다 <strong>코드 리뷰</strong>와 검증에 집중되는 구조를 측정한 연구</li> <li>ChatDev가 수행한 30개 소프트웨어 개발 태스크에서 코드 리뷰 단계가 평균 <strong>…
-
Odysseus - 셀프 호스팅 AI 워크스페이스
Odysseus - 셀프 호스팅 AI 워크스페이스 | GeekNews
<ul> <li>ChatGPT/Claude의 UI 경험을 <strong>자체 하드웨어</strong>에서 직접 운영하는 <strong>로컬 퍼스트</strong> 통합 AI 워크스페이스</li> <li>구독자 1.1억명 유튜버인 PewDiePie가 12개월간 개발해 출시, 일주일만에 GitHub 스타 5만 개 돌파</li…
-
코드는 더 싸졌다
코드는 더 싸졌다 | GeekNews
<ul> <li>AI 코딩 도구 확산으로 코드 작성 비용은 급격히 낮아졌지만, 정작 생성된 코드를 <strong>이해하는 비용</strong>은 더 커진 현실</li> <li>LLM은 비결정적이고 원본 소스를 보존하지 않으며 출력 범위가 일반 소프트웨어 전체로 넓어, <strong>컴파일러 출력과 동일시할 수 없음</st…
-
어시스턴트 축: 대규모 언어 모델의 특성 규명 및 안정화
The assistant axis: situating and stabilizing the character of large language models
-
대규모 언어 모델의 창발적 내성 인식
Emergent introspective awareness in large language models
-
Ask HN: GenAI로 느낀 "아, 큰일 났다" 순간은?
Ask HN: What was your "oh shit" moment with GenAI? | Hacker News
<p>Most of us were amused when DALL-E and its peers went mainstream, and we were quick to point out the obvious flaws.<p>Then ChatGPT hit the scene and again, many of us dismissed …