ChronoGraph represents a changing environment as a sequence of scene graphs.
Shared scene graphs connect interaction understanding with grounded robot planning
ChronoGraphVLM represents observed and predicted scene changes as action-linked graphs, enabling a single model to interpret interactions and plan affordance-level robot actions.
Big Tech
Chenyangguang Zhang · Malgorzata Gwiazda · Guanlong Jiao · Yuanchen Ju · Federico Tombari · Koushil Sreenath · +2 more
ETH Zurich · Technical University of Munich · University of British Columbia · University of California, Berkeley · Google
Research Digest··2 min read
Zhang and colleagues introduce ChronoGraph, a representation that links actions on functional object parts, such as handles, to subsequent semantic and geometric changes in a scene.
Why this paper
From Google and 5 others
In one line
ChronoGraph links actions on affordance parts to semantic and geometric state changes, enabling vision-language models to understand and plan 4D interactions.
What we could check
- ·No code link found
- ·No weights link found
- ·No dataset link found
- ·No compute details found
- ·No stated limitations found
- ·No benchmark numbers found
Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.
§