WACV 2027 Workshop 1st Edition

Latent Visual Reasoning: Perception, Imagination, and Multimodal Thought

Date: January 4 or 5, 2027 (TBD)  ·  Location: Disney Springs, Buena Vista, Florida

Submission countdown

· 23:59:59 AoE

--Days
--Hours
--Minutes
--Seconds

Anywhere on Earth (UTC−12)

About

What should a model think with when words are not enough? Should it keep an internal scene, imagine how objects move, revise a spatial hypothesis, or switch among visual, linguistic, symbolic, and tool-based computation?

Current multimodal large language models typically encode an image as context and carry out intermediate reasoning in language. Language is well suited to abstraction and communication, but it often struggles with precise pose, depth, correspondence, occlusion, and temporal or geometric structure. People, by contrast, often simulate how a scene might unfold rather than narrating every detail. When an apple falls from a table, we can imagine its motion, the collision, and where it lands without putting the whole process into words. That kind of spatial and dynamic reasoning may be essential for visual intelligence in video, 3D vision, robotics, medical imaging, and scientific vision.

The Latent Visual Reasoning (LVR) workshop focuses on visual reasoning that does not rely exclusively on text as the intermediate representation. We welcome work on visual tokens, latent scenes, imagined visual states, and other structured representations that keep task-relevant visual information, as well as hybrid systems that combine these with language, symbols, explicit visual operations, or tools. In every case, the intermediate representation should play a testable computational role, not merely coincide with an ordinary hidden activation.

Invited Speakers

Min Xu

Min Xu

Associate Professor

Carnegie Mellon University

Call for Papers

We invite submissions on latent visual reasoning: work whose intermediate computation is not limited to text, especially when visual structure is kept and updated across reasoning steps. Topics include, but are not limited to:

Tracks and publication

We welcome submissions in the following three categories:

All submissions will be handled through OpenReview. Submission portal: Submit via OpenReview. Formatting: please use the official WACV 2027 LaTeX Author Kit and follow the WACV 2027 Author Guidelines.

Timeline

Event Date
Workshop website live by August 25, 2026
Submission deadline October 10, 2026 (AoE)
Author notification deadline (archival papers) October 30, 2026
Metadata of accepted papers due to IEEE (archival papers) November 2, 2026
Camera-ready deadline (archival papers) November 20, 2026
Workshop date January 4 or 5, 2027 (TBD)

Tentative Schedule

LVR will be a half-day, in-person workshop at WACV 2027. Times and room: TO BE ADDED.

Session Details
Opening and research agenda TO BE ADDED
Invited talks TO BE ADDED
Contributed orals and spotlights TO BE ADDED
Posters and demonstrations TO BE ADDED
Panel: “How Do We Know a Model Is Reasoning Visually?” TO BE ADDED
Awards and closing remarks TO BE ADDED

Organizing Committee

Xi Xiao

Xi Xiao

University of Alabama at Birmingham

Contact: xxiao@uab.edu · chen.liu.cl2482@yale.edu · Qizhen.Lan@uth.tmc.edu

Program Committee

Venue

WACV 2027 · Florida

Disney Springs

Buena Vista, Florida

LVR will be held in person in conjunction with WACV 2027 on January 4 or 5, 2027. The exact workshop date and room will be announced.

Open in Google Maps