-
Laguna S 2.1 공개
Laguna S 2.1 공개 | GeekNews
<ul> <li>Poolside가 장기 작업과 추론 능력을 강화한 <strong>Laguna S 2.1</strong>을 공개함. 총 118B MoE 중 토큰당 8B 매개변수를 활성화하며, thinking·no-thinking 모드 모두 최대 <strong>1M 토큰 컨텍스트</strong>를 지원함</li> <li>학습…
-
Qwen 3.8 출시
Qwen 3.8 출시 | GeekNews
<ul> <li>Alibaba Qwen이 <strong>2.4조 파라미터</strong> 규모의 Qwen3.8을 출시하며, 조만간 오픈 웨이트도 공개할 예정</li> <li>Qwen3.8은 계속 진화 중인 모델로, <strong>선도적인 프런티어 AI 모델과 호환</strong>된다고 밝힘</li> <li>Qwen 측은 …
-
Qwen 토큰 요금제
Qwen (@Alibaba_Qwen) on X
<p><a href="https://www.qwencloud.com/pricing/token-plan" rel="nofollow">https://www.qwencloud.com/pricing/token-plan</a></p> <hr /> <p>Comments URL: <a href="https://news.ycombina…
-
Kimi K3의 순간
The Kimi K3 Moment
<p>Article URL: <a href="https://stephen.bochinski.dev/blog/2026/07/18/the-kimi-k3-moment/">https://stephen.bochinski.dev/blog/2026/07/18/the-kimi-k3-moment/</a></p> <p>Comments UR…
-
Kimi K3 공개 - 개방형 프론티어 인텔리전스
Kimi K3 공개 - 개방형 프론티어 인텔리전스 | GeekNews
<ul> <li>Kimi K3는 <strong>2.8조 파라미터</strong>, 네이티브 비전, 100만 토큰 컨텍스트를 갖추고 장시간 코딩/지식 작업/추론을 겨냥한 세계 최초의 공개 3T급 모델</li> <li><strong>Kimi Delta Attention/Attention Residuals</strong>와 8…
-
Kimi K3와 펠리칸 벤치마크로부터 배울 수 있는 것
Kimi K3, and what we can still learn from the pelican benchmark
<p>Chinese AI lab Moonshot AI <a href="https://www.kimi.com/blog/kimi-k3">announced Kimi K3</a> this morning, describing it as their "most capable model to date, with 2.8 trillion …
-
에이전틱 코딩 및 지식 작업을 위해 설계된 Kimi AI K3
Kimi AI with K3 | Built for Agentic Coding & Knowledge Work
<p>Article URL: <a href="https://www.kimi.com/en">https://www.kimi.com/en</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48935342">https://news.ycombinator.…
-
Inkling: 우리의 오픈 웨이트 모델
Inkling: Our Open-Weights Model
<p>Article URL: <a href="https://thinkingmachines.ai/news/introducing-inkling/">https://thinkingmachines.ai/news/introducing-inkling/</a></p> <p>Comments URL: <a href="https://news…
-
Leanstral 1.5
<p>Article URL: <a href="https://docs.mistral.ai/models/model-cards/leanstral-1-5-26-06">https://docs.mistral.ai/models/model-cards/leanstral-1-5-26-06</a></p> <p>Comments URL: <a …
-
Apple Foundation Models에 Claude 탑재하기
Apple Foundation Models에 Claude 탑재하기 | GeekNews
<ul> <li>Apple의 <strong>Foundation Models 프레임워크</strong>에 Claude를 서버 사이드 모델로 연결하는 Swift 패키지로, 개발자가 <strong>Apple 온디바이스 모델과 똑같은 코드 경로</strong>로 Claude를 호출할 수 있게 됨</li> <li>WWDC 2026…
-
Kimi K2.7-Code: 더 나은 토큰 효율성을 갖춘 오픈소스 코딩 모델
Kimi K2.7-Code: open-source coding model with better token efficiency
<p>Article URL: <a href="https://huggingface.co/moonshotai/Kimi-K2.7-Code">https://huggingface.co/moonshotai/Kimi-K2.7-Code</a></p> <p>Comments URL: <a href="https://news.ycombinat…
-
Claude Fable 5 첫인상
Initial impressions of Claude Fable 5
<p>I didn't have early access to today's <a href="https://www.anthropic.com/news/claude-fable-5-mythos-5">Claude Fable 5</a> release, but I've spent the past ~5.5 hours putting it …
-
GPT-2: 공개하기에 너무 위험했던 모델 (2019)
GPT-2: Too Dangerous To Release (2019) – Naoki Shibuya
<p>Article URL: <a href="https://naokishibuya.github.io/blog/2022-12-30-gpt-2-2019/">https://naokishibuya.github.io/blog/2022-12-30-gpt-2-2019/</a></p> <p>Comments URL: <a href="ht…
-
MAI-Code-1-Flash 소개 | Microsoft AI
Introducing MAI-Code-1-Flash | Microsoft AI
<p><a href="https://microsoft.ai/models/mai-code-1-flash/" rel="nofollow">https://microsoft.ai/models/mai-code-1-flash/</a><p><a href="https://microsoft.ai/pdf/MAI-Code-1-Flash-Mod…
-
CS336: 처음부터 만드는 언어 모델링 | GeekNews
<ul> <li><strong>언어 모델</strong>은 현대 NLP 애플리케이션의 기반이며, 하나의 범용 시스템으로 다양한 하위 작업을 다루는 새 패러다임을 연다</li> <li>이 과정은 사전학습용 <strong>데이터 수집·정제</strong>, Transformer 구축, 학습, 배포 전 평가까지 언어 모델 개발…
-
GPT-5.5를 위한 펠리칸: 반공식적 Codex 백도어 API
A pelican for GPT-5.5 via the semi-official Codex backdoor API
<p><a href="https://openai.com/index/introducing-gpt-5-5/">GPT-5.5 is out</a>. It's available in OpenAI Codex and is rolling out to paid ChatGPT subscribers. I've had some preview …