-
arXiv 논문 12,750편에서 AI 문체를 측정한 방법과 한계
arXiv 논문 12,750편에서 AI 문체를 측정한 방법과 한계 | GeekNews
<ul> <li>arXiv 논문 <strong>12,750편의 본문 전체</strong>를 분석한 결과, 최신 논문의 약 3분의 1이 기계 작성처럼 읽혔으며 최근 완결 분기 판정률은 약 32%였음</li> <li>2021~2022년 논문을 인간 작성 대조군으로 삼아 <strong>오탐률 0.4%</strong> 에 맞춰 …
-
arXiv에서 AI 글쓰기를 측정한 방법과 그 한계
How we measured AI writing across arXiv, and where the measurement breaks
<p>Article URL: <a href="https://unslop.run/blog/measuring-ai-writing-on-arxiv">https://unslop.run/blog/measuring-ai-writing-on-arxiv</a></p> <p>Comments URL: <a href="https://news…
-
Code as Agent Harness — 코드를 에이전트의 실행 기반으로 보는 102페이
Code as Agent Harness — 코드를 에이전트의 실행 기반으로 보는 102페이 | GeekNews
<p><strong>UIUC × Meta × Stanford</strong> 합작. 5월 arXiv에 올라온 서베이 논문인데, 관점이 꽤 재밌다.</p> <h3>핵심 주장</h3> <p><strong>"코드는 더 이상 LLM이 생성하는 결과물이 아니다. 에이전트가 추론하고, 행동하고, 상태를 저장하고, 피드백을 …