Grounded and Faithful Vision-Language Models for Real-World Deployment.

2nd Workshop on VLM4RWD | NeurIPS 2026 | Dec, 2026 | Sydney, Australia

Workshop Overview

Vision-language(-action) models are rapidly evolving from systems that describe the world to agents that must perceive, reason, and act within it. This workshop focuses on the principles, methods, and evaluation needed to build grounded, faithful, and reliable multimodal intelligence for real-world deployment.

Grounded Understanding & Faithful Reasoning

AI systems must reliably connect language, perception, and actions with relevant entities and observations in the environment, ensuring that predictions, reasoning, and decisions remain supported by underlying visual and physical evidence.

Visual grounding Faithful Reasoning Evidence alignment Hallucination mitigation

Reliable Interaction & Decision-Making

Real-world agents must integrate perception, reasoning, planning, and action in a manner that remains robust under dynamic and uncertain conditions, with decisions grounded in the causal structure of the environment.

Embodied AI Autonomous Systems Causal Decision-Making

Evaluation & Deployment Reliability

Systems deployed in robotics and autonomous environments require principled evaluation, calibrated uncertainty, predictable failure modes, and rigorous assessment across distribution shifts and counterfactual scenarios.

Benchmarks Uncertainty Robust Evaluation

Call for Papers

Overview

We invite high-quality submissions that advance grounded and faithful vision-language and vision-language-action models for real-world deployment.

We welcome research addressing key challenges in visual grounding, faithful reasoning, hallucination mitigation, robustness, embodied intelligence, robotics, autonomous systems, and world models.

Accepted papers will be presented during the workshop poster sessions, and selected submissions will be invited for contributed spotlight talks.

Submission Guidelines

  • Formatting: Workshop papers may be up to 8 pages, excluding references and appendices, and should follow the NeurIPS 2026 conference format.
  • Review process: Submissions will undergo double-blind review and must be fully anonymized.
  • Submission portal: Papers will be submitted through OpenReview.
  • Submission types: We welcome full workshop papers, demo papers, extended abstracts, position papers, datasets, benchmarks, and emerging research ideas.
  • Previously published work: Relevant published work may be submitted for presentation but will not be eligible for workshop awards.

Important Dates

Paper Submission

Aug 31th, 2026

Notification

Sep 29th, 2026

Camera Ready

Oct , 2026

Workshop Date

Dec , 2026

Schedule

To be announced! Full-day workshop at NeurIPS 2026, Dec, Sydney, Australia.

Speakers

We are excited to welcome leading researchers from academia and industry. To be completed.

Organizing Committee

Contact Us

Questions or feedback? Feel free to reach out. We would love to hear from you.

Email

For general inquiries:

mnasraza@uwaterloo.ca

yimu.wang@uwaterloo.ca

Paper Submissions

Paper submissions will be managed through:

OpenReview Submission System

Reviewer Self-Nomination

Please complete our Reviewer & Area Chair Self-Nomination form.

Workshop Location

NeurIPS 2026

Sydney, Australia

Exact venue details will be announced closer to the event

Frequently Asked Questions

Is the workshop in-person or virtual?

The workshop will be held in-person at NeurIPS 2026 in Sydney, Australia.

Will the workshop proceedings be archival?

No, the workshop proceedings will be non-archival. Authors of accepted papers retain the full copyright of their work and are free to submit extended versions to conferences or journals.

Can I join the program committee?

Yes! If you are interested in serving as a reviewer or Area Chair for VLM4RWD, please complete our Reviewer & Area Chair Self-Nomination form. We welcome researchers with relevant expertise in vision-language models, multimodal AI, robotics, autonomous systems, and related areas.