facebook/map-anything-v1

Model

26

stars

6

commits

4

repos using this model

7

linked in READMEs

Dec 19, 2025

updated

3d-reconstruction
camera-pose
computer-vision
covisibility
depth-estimation
image-to-3d
mapanything
model_hub_mixin
multi-view-stereo
pytorch_model_hub_mixin
safetensors

README

Overview

MapAnything is a simple, end-to-end trained transformer model that directly regresses the factored metric 3D geometry of a scene given various types of modalities as inputs. A single feed-forward model supports over 12 different 3D reconstruction tasks including multi-image sfm, multi-view stereo, monocular metric depth estimation, registration, depth completion and more.

This is the CC-BY-NC-4.0 variant of the model released in September 2025, i.e., the V1 Version.

Latest CC-BY-NC-4.0 model here: https://huggingface.co/facebook/map-anything

Quick Start

Please refer to our Github Repo: https://github.com/facebookresearch/map-anything

Citation

If you find our repository useful, please consider giving it a star ⭐ and citing our paper in your work:

@inproceedings{keetha2026mapanything,
  title={{MapAnything}: Universal Feed-Forward Metric 3D Reconstruction},
  author={Keetha, Nikhil and M{\"u}ller, Norman and Sch{\"o}nberger, Johannes and Porzi, Lorenzo and Zhang, Yuchen and Fischer, Tobias and Knapitsch, Arno and Zauss, Duncan and Weber, Ethan and Antunes, Nelson and others},
  booktitle={International Conference on 3D Vision (3DV)},
  year={2026},
  organization={IEEE}
}

Contributors

NikV09

5 commits

meta-bot

1 commits

facebook/map-anything-v1

Model

26

stars

6

commits

4

repos using this model

7

linked in READMEs

Dec 19, 2025

updated

3d-reconstruction
camera-pose
computer-vision
covisibility
depth-estimation
image-to-3d
mapanything
model_hub_mixin
multi-view-stereo
pytorch_model_hub_mixin
safetensors

README

Overview

MapAnything is a simple, end-to-end trained transformer model that directly regresses the factored metric 3D geometry of a scene given various types of modalities as inputs. A single feed-forward model supports over 12 different 3D reconstruction tasks including multi-image sfm, multi-view stereo, monocular metric depth estimation, registration, depth completion and more.

This is the CC-BY-NC-4.0 variant of the model released in September 2025, i.e., the V1 Version.

Latest CC-BY-NC-4.0 model here: https://huggingface.co/facebook/map-anything

Quick Start

Please refer to our Github Repo: https://github.com/facebookresearch/map-anything

Citation

If you find our repository useful, please consider giving it a star ⭐ and citing our paper in your work:

@inproceedings{keetha2026mapanything,
  title={{MapAnything}: Universal Feed-Forward Metric 3D Reconstruction},
  author={Keetha, Nikhil and M{\"u}ller, Norman and Sch{\"o}nberger, Johannes and Porzi, Lorenzo and Zhang, Yuchen and Fischer, Tobias and Knapitsch, Arno and Zauss, Duncan and Weber, Ethan and Antunes, Nelson and others},
  booktitle={International Conference on 3D Vision (3DV)},
  year={2026},
  organization={IEEE}
}

Contributors

NikV09

5 commits

meta-bot

1 commits