-
모델 팩토리 안을 들어다보다 — Eiso Kant, Poolside AI
Inside the Model Factory — Eiso Kant, Poolside AI
Poolside's co-CEO on how his small team of top researchers built a model factory capable of training Laguna S - a 118B MOE beating Thinky's ~1T open weights model... and this is ju…
-
Ornith-1.0: 에이전트 코딩을 위한 자가 개선 오픈소스 모델
Ornith-1.0: self-improving open-source models for agentic coding
<p>Article URL: <a href="https://github.com/deepreinforce-ai/Ornith-1">https://github.com/deepreinforce-ai/Ornith-1</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/i…
-
품질이 낮은 강화학습 환경 배포를 중단하는 방법 (예제 포함)
How to Stop Shipping Low-Quality RL Environments (with Examples)
Your broken harness is actively making the model worse. Here's what I keep seeing after years of eyeballing trajectories, and what you need to fix.
-
CS336: 처음부터 만드는 언어 모델링 | GeekNews
<ul> <li><strong>언어 모델</strong>은 현대 NLP 애플리케이션의 기반이며, 하나의 범용 시스템으로 다양한 하위 작업을 다루는 새 패러다임을 연다</li> <li>이 과정은 사전학습용 <strong>데이터 수집·정제</strong>, Transformer 구축, 학습, 배포 전 평가까지 언어 모델 개발…
-
ImportAI 449: LLM이 다른 LLM을 학습시킴; 72B 분산 학습 실행; 컴퓨터 비전은 생성 텍스트보다 더 어렵다
ImportAI 449: LLMs training other LLMs; 72B distributed training run; computer vision is harder than generative text
Will AI cause a political interregnum