ended5월 14일· 1 sources
The Real Performance History of AI Models: Exposing Post-Launch Changes
AI 모델의 진짜 성능 기록, 출시 후 저하 현상을 데이터로 증명하다
Why it matters
AI model performance doesn't remain constant after launch—labs often introduce performance degradation through quantization, aggressive content filtering, or behavioral changes to reduce compute costs. LM Arena's objective ELO leaderboard, built from thousands of crowdsourced human evaluations, provides transparent tracking of these hidden trends that marketing announcements typically obscure. This data reveals critical differences between API performance and consumer web interfaces, empowering users to make informed choices based on real-world model capabilities.
1
Sources
+0
24h
—
Growth
8d
Active
LM ArenaELO RatingModel DegradationQuantizationLeaderboard