Article
Physic Grounded Vision Foundation Models for Human Computer Interaction in Embodied Environments
2025-10-10
Abstract excerpt
Physic-ground vision foundation models for human-computer interaction represent a transformative paradigm in artificial intelligence, as they extend beyond conventional data-driven perception to incorporate explicit reasoning about the physical laws and causal structures that govern the real world. Unlike earlier generations of vision models that excelled at pattern recognition but faltered when faced with tasks d...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 78c20611-0a00-5525-944f-fe7810fc7c6e
- DOI
- 10.20944/preprints202510.0649.v1
