-
Laguna S 2.1 출시: Deepseek v4 Flash보다 저렴하고 V4 Pro보다 우수
[AINews] "Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro"
a quiet day lets us highlight a new neolab win.
-
DSpark: 추측적 디코딩을 통한 LLM 추론 가속
DSpark: Speculative decoding accelerates LLM inference [pdf]
<p>Article URL: <a href="https://github.com/deepseek-ai/DeepSpec/blob/main/DSpark_paper.pdf">https://github.com/deepseek-ai/DeepSpec/blob/main/DSpark_paper.pdf</a></p> <p>Comments …
-
Unlimited OCR — Baidu의 원샷 장문 파싱 모델
Unlimited OCR — Baidu의 원샷 장문 파싱 모델 | GeekNews
<ul> <li><strong>DeepSeek OCR</strong>를 기반으로 디코더의 모든 어텐션을 교체해, 수십 페이지 문서를 <strong>한 번의 순전파(forward pass)</strong> 로 전사하는 E2E OCR 모델</li> <li>핵심은 <strong>참조 슬라이딩 윈도우 어텐션(R-SWA)</str…
-
DeepSeek Vision 발표
DeepSeek Introduces Vision
<p>Article URL: <a href="https://chat.deepseek.com/">https://chat.deepseek.com/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48581458">https://news.ycombi…
-
미국, DeepSeek 및 보안 위험으로 분류된 100여 개 기업의 블랙리스트 등재 모두
미국, DeepSeek 및 보안 위험으로 분류된 100여 개 기업의 블랙리스트 등재 모두 | GeekNews
<ul> <li>미국 정부가 국가안보 위험으로 분류된 중국 <strong>DeepSeek</strong>, 메모리 칩 제조사 <strong>CXMT</strong> 등 100개 이상 기업의 <strong>무역 블랙리스트(Entity List) 등재를 보류</strong> 중이며, 베이징과의 긴장 고조를 피하려는 트럼프 행…
-
미국, DeepSeek 제재 보류, 100개 이상 기업 보안 위험으로 지정
US holds off blacklisting DeepSeek, more than 100 firms deemed security risks
<p><a href="https://archive.ph/MlU1U" rel="nofollow">https://archive.ph/MlU1U</a></p> <hr /> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48565498">https://news.y…
-
DeepSeek V4 Pro, 정밀도에서 GPT-5.5 Pro를 앞서다
DeepSeek V4 Pro, 정밀도에서 GPT-5.5 Pro를 앞서다 | GeekNews
<ul> <li>사전 준비가 불가능하도록 즉석 생성된 <strong>4개 텍스트 과제</strong> 1:1 비교에서 DeepSeek V4 Pro가 <strong>38.0점</strong>, GPT-5.5 Pro가 33.0점 기록</li> <li>두 모델 모두 강력했으나, DeepSeek는 더 엄격하고 직역적이며 <str…
-
딥씨크 V4 - 거의 최고 수준의 성능, 저렴한 가격
DeepSeek V4 - almost on the frontier, a fraction of the price
<p>Chinese AI lab DeepSeek's last model release was V3.2 (and V3.2 Speciale) <a href="https://simonwillison.net/2025/Dec/1/deepseek-v32/">last December</a>. They just dropped the f…
-
2025년 LLM의 현황: 진전, 문제, 그리고 예측
The State Of LLMs 2025: Progress, Problems, and Predictions
A 2025 review of large language models, from DeepSeek R1 and RLVR to inference-time scaling, benchmarks, architectures, and predictions for 2026.
-
주요 LLM 아키텍처 비교
The Big LLM Architecture Comparison
From DeepSeek-V3 to Kimi K2: A Look At Modern LLM Architecture Design