-
Kimi K3는 Fable과 경쟁하며, 두 모델의 조합은 최고 수준의 성능을 달성함
Kimi K3는 Fable과 경쟁하며, 두 모델의 조합은 최고 수준의 성능을 달성함 | GeekNews
<ul> <li>약 <strong>1,030개 에이전트 작업</strong>에서 Kimi K3와 Fable 5를 비교한 결과, 작업별 라우팅은 <strong>93% 정확도</strong>로 개별 모델보다 높은 품질을 달성함</li> <li>SWE·터미널·알고리듬·다중 언어·법률 작업에서 전체 성능은 비슷했지만, 두 모델이…
-
Kimi K3는 Fable과 경쟁 중, Kimi K3와 Fable은 최첨단 기술
Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
<p>Also: Kimi K3: second only to Fable 5 on AA-Briefcase <a href="https://artificialanalysis.ai/articles/kimi-k3-agentic-knowledge-benchmark" rel="nofollow">https://artificialanal…
-
Fable 5 vs. GPT-5.6 Sol: NP-난제에서 /goal이 도움이 될까?
Fable 5 vs. GPT-5.6 Sol on an NP-Hard Problem: Does /goal Help? - Charles AZAM
<p>Article URL: <a href="https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/">https://charlesazam.com/blog/fable-5-gpt-5-6-sol-goal/</a></p> <p>Comments URL: <a href="https://ne…
-
Apple SpeechAnalyzer API, Whisper·이전 API와 비교 벤치마크
Apple SpeechAnalyzer API, Whisper·이전 API와 비교 벤치마크 | GeekNews
<ul> <li>Apple M2 Pro에서 5,559개 LibriSpeech 음성을 동일한 프로덕션 코드로 처리한 결과, <strong>SpeechAnalyzer</strong>가 깨끗한 음성 2.12%, 잡음이 많은 음성 4.56%의 단어 오류율(WER)로 테스트한 모든 엔진보다 정확했음</li> <li>기존 <stro…
-
Claude Code는 프롬프트를 읽기 전 3.3만 토큰, OpenCode는 7천 토큰을
Claude Code는 프롬프트를 읽기 전 3.3만 토큰, OpenCode는 7천 토큰을 | GeekNews
<ul> <li>동일한 모델·머신·작업에서 API 경계를 측정한 결과, Sonnet 4.5 첫 요청의 고정 오버헤드는 <strong>Claude Code 약 32,800토큰</strong>, OpenCode 약 6,900토큰으로 4.7배 차이 났으며 Fable 5에서는 약 3.3배로 줄어듦</li> <li>격차의 대부분은…
-
Claude Code가 프롬프트를 읽기 전에 OpenCode보다 4.7배 더 많은 토큰을 전송
Claude Code Sends 4.7x More Tokens Than OpenCode Before Reading Your Prompt
<p>This started based off of a hunch. We usually use OpenCode, but were 'forced' to use Claude Code for a while due to issues with Meridian. In that time, we saw the usage meter ri…