-
GitHub - marcelroed/gigatoken: GB/s 속도의 언어 모델 토큰화
GitHub - marcelroed/gigatoken: Language model tokenization at GB/s
<p>Article URL: <a href="https://github.com/marcelroed/gigatoken/">https://github.com/marcelroed/gigatoken/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=4…
-
Nativ - 프런티어급 오픈 모델을 Mac에서 로컬 실행
Nativ - 프런티어급 오픈 모델을 Mac에서 로컬 실행 | GeekNews
<ul> <li><strong>Nativ</strong>은 Apple Silicon Mac에서 오픈 AI 모델을 내려받아 계정·구독·클라우드 없이 실행하는 MIT 라이선스 오픈소스 앱임</li> <li>Google, Cohere, Liquid AI 등의 모델을 제공하며, Mac 하드웨어에 적합한 모델을 추천하고 모든 응답…
-
Claude가 load-bearing을 반복하는 것을 멈추는 방법
How to stop Claude from saying load-bearing | jola.dev
<p>Article URL: <a href="https://jola.dev/posts/how-to-stop-claude-from-saying-load-bearing">https://jola.dev/posts/how-to-stop-claude-from-saying-load-bearing</a></p> <p>Comments …
-
새로운 GPT-5.6 패밀리: Luna, Terra, Sol
The new GPT-5.6 family: Luna, Terra, Sol
<p>OpenAI's latest flagship model <a href="https://openai.com/index/gpt-5-6/">hit general availability this morning</a>, and comes in three sizes: Luna, Terra, and Sol (from smalle…
-
OpenAI GPT 5.6이 출시되었습니다.
OpenAI GPT 5.6이 출시되었습니다. | GeekNews
<p>OpenAI가 2026년 7월 9일, GPT-5.6 패밀리를 정식 출시(GA)했습니다. 6월 26일 프리뷰 이후의 정식 공개이며, <strong>Sol</strong>(플래그십), <strong>Terra</strong>(균형형), <strong>Luna</strong>(최저비용) 세 개의 티어로 구성됩니다.</p>…
-
GPT-5.6: 야망에 맞춰 확장되는 프론티어 지능
GPT-5.6: Frontier intelligence that scales with your ambition
More intelligence from every token, stronger performance per dollar, and more capability on demand for your hardest work.
-
언어 모델의 글로벌 워크스페이스
A global workspace in language models
-
Leanstral 1.5: 모두를 위한 증명
Leanstral 1.5: Proof abundance for all
<p>Article URL: <a href="https://mistral.ai/news/leanstral-1-5/">https://mistral.ai/news/leanstral-1-5/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48780…
-
로컬 LLM 실행에 관한 모든 것
GitHub - jamesob/local-llm: Everything I know about running LLMs locally
<p>Article URL: <a href="https://github.com/jamesob/local-llm">https://github.com/jamesob/local-llm</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48775921"…
-
Qwen 3.6 27B는 로컬 개발의 최적 지점
Qwen 3.6 27B는 로컬 개발의 최적 지점 | GeekNews
<ul> <li><strong>Qwen 3.6 27B</strong>는 로컬 모델에 회의적이던 사용자에게도 범용 작업에서 의미 있는 선택지로 보이며, 35B A3B보다 느리지만 더 강력한 dense 모델로 추천됨</li> <li>창작·코딩 테스트에서는 <strong>제약 조건 준수</strong>가 강점으로 드러났고, O…
-
DSpark: 추측적 디코딩을 통한 LLM 추론 가속
DSpark: Speculative decoding accelerates LLM inference [pdf]
<p>Article URL: <a href="https://github.com/deepseek-ai/DeepSpec/blob/main/DSpark_paper.pdf">https://github.com/deepseek-ai/DeepSpec/blob/main/DSpark_paper.pdf</a></p> <p>Comments …
-
GPT-5.5가 MIT 라이선스 GLM-5.2보다 3배 더 환각함
GPT-5.5 hallucinates 3x more than MIT-licensed GLM-5.2
<p>Article URL: <a href="https://arrowtsx.dev/bigger-models/">https://arrowtsx.dev/bigger-models/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48600167">h…
-
Claude의 생물정보학 연구 능력을 BioMysteryBench로 평가하기
Evaluating Claude’s bioinformatics research capabilities with BioMysteryBench
-
GLM-5.2: 세계 최고의 프론트엔드 코딩 모델, 추측 디코딩을 위한 IndexShare
[AINews] GLM-5.2: the top Frontend Coding model in the world, IndexShare for Speculative Decoding
We have a new top open model in the world!
-
GPT-NL: 네덜란드의 주권 언어 모델
GPT‑NL: a sovereign language model for the Netherlands
<p>Article URL: <a href="https://www.tno.nl/en/digital/artificial-intelligence/gpt-nl/">https://www.tno.nl/en/digital/artificial-intelligence/gpt-nl/</a></p> <p>Comments URL: <a hr…
-
LLM 연구 논문: 2026년 목록 (1월~5월)
LLM Research Papers: The 2026 List (January to May)
A curated roundup of notable LLM research papers that came out this year
-
자동화된 정렬 연구자들: 대규모 언어 모델을 활용한 확장 가능한 감시
Automated Alignment Researchers: Using large language models to scale scalable oversight
-
그들은 가중치로 만들어졌다
그들은 가중치로 만들어졌다 | GeekNews
<ul> <li><strong>가중치</strong>는 AI 모델 안에 사전, 문법 규칙, 작은 사람, 언어 모듈, 추론 장치, 데이터베이스가 따로 없고 80개 층의 부동소수점 숫자 곱셈이 문장과 추론을 만든다는 풍자적 전제임</li> <li>성과 평가의 어조 완화와 추도문 작성도 기술적으로 <strong>다음 토큰 예측…
-
Gemma 4 12B: 통합형 인코더 없는 멀티모달 모델
Gemma 4 12B: 통합형 인코더 없는 멀티모달 모델 | GeekNews
<ul> <li><strong>Gemma 4 12B</strong>는 노트북에서 에이전트형 멀티모달 지능을 실행하도록 설계된 중간 크기 모델이며, edge 친화적인 E4B와 더 고급인 26B MoE 사이의 간극을 메움</li> <li><strong>인코더 없는 통합 아키텍처</strong>로 이미지와 오디오 입력을 별도 …
-
자연어 오토인코더: Claude의 생각을 텍스트로 변환하기
Natural Language Autoencoders
-
봄날의 꿈: 2026년 1-2월 오픈웨이트 LLM 10가지 아키텍처
A Dream of Spring for Open-Weight LLMs: 10 Architectures from Jan-Feb 2026
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026
-
LLM 추론 개선을 위한 추론 시간 스케일링의 카테고리
Categories of Inference-Time Scaling for Improved LLM Reasoning
And an Overview of Recent Inference-Scaling Papers
-
Qwen3을 처음부터 이해하고 구현하기
Understanding and Implementing Qwen3 From Scratch
A Detailed Look at One of the Leading Open-Source LLMs
-
주요 LLM 아키텍처 비교
The Big LLM Architecture Comparison
From DeepSeek-V3 to Kimi K2: A Look At Modern LLM Architecture Design
-
LLM 연구 논문: 2025년 목록 (1월~6월)
LLM Research Papers: The 2025 List (January to June)
A topic-organized collection of 200+ LLM research papers from 2025
-
LLM 추론을 위한 강화학습의 현황
The State of Reinforcement Learning for LLM Reasoning
Understanding GRPO and New Insights from Reasoning Model Papers