Boxer Lifts 2D Detections into Fused 3D Boxes

Loading video
Loading videoMeta Reality Labs' Boxer lifts open-vocabulary 2D detections into globally fused 3D oriented boxes from posed images and point clouds, with offline fusion and online tracking.
3D Object DetectionOpen-Vocabulary PerceptionOriented Bounding BoxesEgocentric VisionMeta Reality LabsBoxer
Category: perception
Author: @ddetone
Date: 2026-09-12T00:00:00
Duration: 16.66s
Reference: https://facebookresearch.github.io/boxer/





