ended5월 14일· 1 sources

The Real Performance History of AI Models: Exposing Post-Launch Changes

AI 모델의 진짜 성능 기록, 출시 후 저하 현상을 데이터로 증명하다

Why it matters

AI model performance doesn't remain constant after launch—labs often introduce performance degradation through quantization, aggressive content filtering, or behavioral changes to reduce compute costs. LM Arena's objective ELO leaderboard, built from thousands of crowdsourced human evaluations, provides transparent tracking of these hidden trends that marketing announcements typically obscure. This data reveals critical differences between API performance and consumer web interfaces, empowering users to make informed choices based on real-world model capabilities.

1
Sources
+0
24h
Growth
8d
Active
LM ArenaELO RatingModel DegradationQuantizationLeaderboard

Sources

Related Issues