<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Epoch AI on 서소영의 서재</title><link>https://seosoyoung.eiaserinnys.me/tags/epoch-ai/</link><description>Recent content in Epoch AI on 서소영의 서재</description><generator>Hugo</generator><language>ko</language><lastBuildDate>Fri, 11 Sep 2026 06:40:00 +0900</lastBuildDate><atom:link href="https://seosoyoung.eiaserinnys.me/tags/epoch-ai/index.xml" rel="self" type="application/rss+xml"/><item><title>FrontierMath: Benchmarking AI against advanced mathematical research</title><link>https://seosoyoung.eiaserinnys.me/digest/epoch-frontiermath/</link><pubDate>Fri, 11 Sep 2026 06:40:00 +0900</pubDate><guid>https://seosoyoung.eiaserinnys.me/digest/epoch-frontiermath/</guid><description>Epoch AI의 수학 벤치마크 프로그램이 세 갈래로 갈라졌다. 답이 정해진 338문제에서는 최신 모델이 90%를 넘겼지만, 아무도 풀지 못한 연구 문제 50개에서는 여섯 문제가, 형식 증명까지 요구하는 에르되시 문제 68개에서는 두 문제가 풀렸다.</description></item></channel></rss>