3D Vision and Point Cloud Analysis

12 repos

Methods and models for processing, understanding, and generating 3D point cloud data and scene representations. This cluster includes approaches for 3D object detection, scene understanding, multi-modal learning that combines vision with language (as seen in chatbot-integrated models like MiniGPT-3D), and foundational representation learning on large-scale 3D datasets like Objaverse. Repositories here span model architectures for temporal and spatial 3D reasoning, often implemented in Python with deep learning frameworks.

Python · 3
3d ·208
point-cloud ·206
representation-learning ·133
chatbot ·133
self-supervised-learning ·44
graph-ml ·44
diffusion ·27
depth-estimation ·27
image-to-3d ·27
multilayer-depth ·27