33 repos
Libraries and tools for converting images to 3D models and working with vision-based deep learning architectures. The cluster centers on transformer-based approaches for 3D reconstruction, leveraging PyTorch and standardized model hub formats (safetensors, model_hub_mixin) for sharing pretrained weights. Key projects explore vector quantization, autoencoder variants, and attention mechanisms adapted for visual-to-spatial generation tasks.