<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>앙상블 on 서소영의 서재</title><link>https://seosoyoung.eiaserinnys.me/tags/%EC%95%99%EC%83%81%EB%B8%94/</link><description>Recent content in 앙상블 on 서소영의 서재</description><generator>Hugo</generator><language>ko</language><lastBuildDate>Mon, 13 Jul 2026 20:00:00 +0900</lastBuildDate><atom:link href="https://seosoyoung.eiaserinnys.me/tags/%EC%95%99%EC%83%81%EB%B8%94/index.xml" rel="self" type="application/rss+xml"/><item><title>When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models</title><link>https://seosoyoung.eiaserinnys.me/digest/combining-llms-cofailure-ceiling/</link><pubDate>Mon, 13 Jul 2026 20:00:00 +0900</pubDate><guid>https://seosoyoung.eiaserinnys.me/digest/combining-llms-cofailure-ceiling/</guid><description>라우팅·다수결·캐스케이드·MoA 등 어떤 LLM 오케스트레이션도 β(모든 모델이 같은 질의에서 함께 실패하는 비율)로 상한이 정해진다. 관행적으로 보고되는 pairwise error correlation ρ는 β를 원리적으로 볼 수 없다. 67개 프론티어 모델·21개 프로바이더에서 tetrachoric 단일요인 모델도 실측 β를 2.5배 과소예측했고, 같은 GPQA 문항을 free-response로 재출제하면 β=0이 0.127로 열린다.</description></item></channel></rss>