ECCV 2026 Workshop

3D in the Era of World Models

Exploring the role of explicit 3D structure, spatial intelligence, video generation, and physical reasoning in scalable world models.

Full day September 9, 2026 Malmo, Sweden

Introduction

A workshop for the 3D vision, graphics, generative modeling, and world-modeling communities.

Research in 3D vision and graphics has built explicit 3D-aware pipelines that provide geometric consistency, interpretability, and controllability. At the same time, recent progress in generative video models suggests that large models may implicitly acquire spatial, geometric, and physical knowledge from scale.

This workshop asks how the role of 3D vision research should evolve in the era of large-scale foundation models for spatially aware world models, and how explicit 3D representations can work together with video-based implicit 3D priors.

AudienceResearchers and practitioners in 3D vision, graphics, video, robotics, and embodied AI
Posters20 to 30 requested poster boards
Scope3D/4D vision, world models, video generation, and spatial intelligence

Topics

We welcome work and discussion across the foundations, evaluation, and applications of 3D-aware world models.

Representation, Structure, and Synergy

When are explicit 3D representations necessary, and when can implicit world models suffice?

Learning Paradigms and Scalability

Can large-scale video pretraining yield spatially grounded understanding, and what is the role of 3D inductive bias?

4D and Long Context Modeling

How should we model dynamic scenes, long-term temporal consistency, and physically plausible interaction?

Evaluation and Benchmarks

What metrics and protocols measure spatial understanding, physical grounding, and controllability beyond pixel fidelity?

Applications

How can 3D reasoning interface with language, planning, embodied agents, and next-generation simulation?

Important Dates

All dates are tentative and will be updated with final submission portal information.

Submission DeadlineAugust 15, 2026 EOD AoEAugust 16, 2026, 4:59 AM PT
NotificationAugust 24, 2026
Final VersionAugust 31, 2026
WorkshopSeptember 9, 2026

Schedule

Full-day program with keynotes, spotlight talks, posters, and a panel discussion.

Morning Session

Opening Remarks
Keynote Talk #1
Keynote Talk #2
Spotlight Talk #1
Spotlight Talk #2
Coffee Break
Keynote Talk #3
Keynote Talk #4
Lunch Break

Afternoon Session

Keynote Talk #5
Keynote Talk #6
Poster Session
Spotlight Talk #3
Spotlight Talk #4
Keynote Talk #7
Keynote Talk #8
Panel Discussion
Closing Remarks

Call for Papers

Accepted papers are non-archival and may be presented as posters or spotlights.

Submission Format

We accept non-archival long papers up to 8 pages and short papers up to 4 pages. Please submit through the OpenReview submission portal.

Relevant Areas

Relevant submissions include 3D/4D vision, world models, video generation, neural rendering, spatial intelligence, benchmarks, physical reasoning, and embodied applications.

Organizers

A cross-institutional team spanning academia and industry.

Contact

Please contact the workshop chairs for questions about submissions and program updates.

Qixing Huang
huangqx@cs.utexas.edu
Hanwen Jiang
hwjiang1510@gmail.com