World Embedding Benchmark
2 Oct 2026
Physical fidelity has received increasing attention in world models and video generation, yet how video representations encode physical information remains less understood.
Trending research and the full catalog - each paper linked to the benchmarks, methods, and models it introduces.
Filtering here covers the 2,000 most recent papers, as much as one page can hold in memory. See the full index of 22,347 papers.
2 Oct 2026
Physical fidelity has received increasing attention in world models and video generation, yet how video representations encode physical information remains less understood.
2 Oct 2026
Omni-modal large language models (OmniLLMs) unify text, images, audio, and video, yet hallucinate when generation relies on the wrong evidence. Existing inference-time methods can reduce hallucinations, but rarely reveal which evidence sustains a generated commitment.