Ground3D-LMM โ€” CA-VQA (3D+2D)

Fine-tuned Ground3D-LMM checkpoint for CA-VQA (paper Table 9 โ€” generality beyond ScanNet), 3D+2D (point cloud + RGB views) setting. Reproduces the paper's Table-9 average (~48).

Usage

CKPT=$(huggingface-cli download amolharsh/Ground3D-LMM-CAVQA-4B-Joint pytorch_model.pth)
# evaluate with the CA-VQA harness: infer_cavqa.py -> merge_preds.py -> evaluate_cavqa.py
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support