ended3월 30일· 1 sources
From Rendering to Reasoning: Measuring Cognitive Intelligence in World Models
AI가 정말 '생각'할까?: 인지 지능을 측정하는 WM Bench
Why it matters
Current world models excel at generating visually convincing content but lack metrics to measure actual cognitive understanding. WM Bench introduces the first comprehensive benchmark evaluating perception, cognition, and embodiment across 100 scenarios, shifting focus from output quality to real reasoning capability. This is critical for developing AI systems that make intelligent decisions in dynamic environments rather than simply produce photorealistic visuals.
1
Sources
+0
24h
—
Growth
175d
Active
WM Benchworld modelscognitive intelligencereasoningembodiment