multimodal
fact
bullish
Using LRMs significantly simplifies the 3D human-object interaction reconstruction procedure by reframing it as interpreting the LRM mesh rather than fitting models to 2D images
This significantly simplifies the reconstruction procedure, reframing the problem as interpreting the LRM mesh: we segment it into human and object components, fit a parametric body model to the human part, and optionally align an object template to the object part
Computer Vision30 Aug 2026