ooonesevennn/CVPR_2026_Oral_Papers

This repository is a curated collection of CVPR 2026 oral papers.

38

6 commits

updated Apr 23, 2026

See the code

README

CVPR 2026 Oral Papers Collection

A community-maintained index of 141 CVPR 2026 Oral papers, organized by research topic, with direct links to the virtual conference page, arXiv preprints and code repositories where available.

中文版本 / Chinese README

Status

  • Total orals: 141
  • arXiv links: 102 (72%)
  • GitHub links: 29 (20%)

Papers not yet on arXiv may be uploaded closer to the conference — PRs welcome.

Per-category counts

#CategoryPapersarXivCode
13D Reconstruction, Geometry, Gaussian / Radiance Field, SfM, Registration45337
2Generative Models, Diffusion, Image / Video / 3D Generation & Editing20173
3Medical Imaging: VLM, Segmentation, Reconstruction, Registration15104
4Segmentation, Detection, Recognition, Tracking, Correspondence302110
5VLM / MLLM / Multimodal Reasoning / Safety / Benchmarks12100
6Robotics, Embodied AI, Navigation, Manipulation, Autonomous Driving, World Models862
7Model Compression, Training Efficiency, Optimization, Federated Learning, Task Arithmetic1394
8Privacy, Security, Watermarking, Copyright, Membership Inference530
9Datasets & Benchmarks11102
—Total (de-dup)14110229

Sub-categories may cross-list a paper (e.g. 3D editing appears in both §1 and §2); the table de-duplicates within each top-level category, so the column sum exceeds 141.

Files

PathPurpose
cvpr2026_orals.jsonFull structured records (title, authors, abstract, arxiv, github, related)
cvpr2026_orals.csvFlat spreadsheet of the same data

Contributing

If you spot a missing arXiv or GitHub link, feel free to open an issue or submit a pull request.


1. 3D Reconstruction, Geometry, Gaussian / Radiance Field, SfM, Registration

1.1 3D Reconstruction, Gaussian Splatting, NeRF, Dynamic Scenes

1.2 Reflection, Transparency, Materials, Lighting, Inverse Rendering

1.3 Multi-view Geometry, SfM, Cameras, Pose Graph

1.4 3D Representations, Shape Decomposition, Shape Modeling

1.5 Occupancy / Scene Completion / Active Mapping

1.6 Multi-view / Point Cloud Registration

2. Generative Models, Diffusion, Image / Video / 3D Generation & Editing

2.1 Image Generation & Editing

2.3 3D Generation & Editing

2.4 Multimodal Generation / Tokenizer / Interleaved

3. Medical Imaging: VLM, Segmentation, Reconstruction, Registration

3.1 Medical Segmentation

3.2 Medical Reconstruction / Inverse Problems / Imaging

3.3 Medical VLMs / Uncertainty / Reliability

3.4 Medical Registration

4. Segmentation, Detection, Recognition, Tracking, Correspondence

4.1 Image / Video Segmentation

4.2 Detection / OOD / Long-tail / Category Discovery

4.3 Correspondence, Keypoints, Optical Flow, Tracking, Representations

4.4 Specialized Recognition / Human Motion / Dance / Sign Language

5. VLM / MLLM / Multimodal Reasoning / Safety / Benchmarks

5.1 Multimodal Reasoning & Capability Analysis

5.2 MLLM / VLM Safety & Attack / Defense

5.3 Video / Vision-Language Model Capabilities

6. Robotics, Embodied AI, Navigation, Manipulation, Autonomous Driving, World Models

6.1 Embodied / Navigation / Manipulation

6.2 Autonomous Driving / World Models / Scenario Mining

6.3 Game Agents

7. Model Compression, Training Efficiency, Optimization, Federated Learning, Task Arithmetic

7.1 Compression / Acceleration / Edge

7.2 Inference & Architecture Efficiency

7.3 Optimization / Federated / Clustering / Theory

7.4 Dataset Distillation / Data Quality

9. Datasets & Benchmarks

ooonesevennn/CVPR_2026_Oral_Papers

This repository is a curated collection of CVPR 2026 oral papers.

38

6 commits

updated Apr 23, 2026

See the code

README

CVPR 2026 Oral Papers Collection

A community-maintained index of 141 CVPR 2026 Oral papers, organized by research topic, with direct links to the virtual conference page, arXiv preprints and code repositories where available.

中文版本 / Chinese README

Status

  • Total orals: 141
  • arXiv links: 102 (72%)
  • GitHub links: 29 (20%)

Papers not yet on arXiv may be uploaded closer to the conference — PRs welcome.

Per-category counts

#CategoryPapersarXivCode
13D Reconstruction, Geometry, Gaussian / Radiance Field, SfM, Registration45337
2Generative Models, Diffusion, Image / Video / 3D Generation & Editing20173
3Medical Imaging: VLM, Segmentation, Reconstruction, Registration15104
4Segmentation, Detection, Recognition, Tracking, Correspondence302110
5VLM / MLLM / Multimodal Reasoning / Safety / Benchmarks12100
6Robotics, Embodied AI, Navigation, Manipulation, Autonomous Driving, World Models862
7Model Compression, Training Efficiency, Optimization, Federated Learning, Task Arithmetic1394
8Privacy, Security, Watermarking, Copyright, Membership Inference530
9Datasets & Benchmarks11102
—Total (de-dup)14110229

Sub-categories may cross-list a paper (e.g. 3D editing appears in both §1 and §2); the table de-duplicates within each top-level category, so the column sum exceeds 141.

Files

PathPurpose
cvpr2026_orals.jsonFull structured records (title, authors, abstract, arxiv, github, related)
cvpr2026_orals.csvFlat spreadsheet of the same data

Contributing

If you spot a missing arXiv or GitHub link, feel free to open an issue or submit a pull request.


1. 3D Reconstruction, Geometry, Gaussian / Radiance Field, SfM, Registration

1.1 3D Reconstruction, Gaussian Splatting, NeRF, Dynamic Scenes

1.2 Reflection, Transparency, Materials, Lighting, Inverse Rendering

1.3 Multi-view Geometry, SfM, Cameras, Pose Graph

1.4 3D Representations, Shape Decomposition, Shape Modeling

1.5 Occupancy / Scene Completion / Active Mapping

1.6 Multi-view / Point Cloud Registration

2. Generative Models, Diffusion, Image / Video / 3D Generation & Editing

2.1 Image Generation & Editing

2.3 3D Generation & Editing

2.4 Multimodal Generation / Tokenizer / Interleaved

3. Medical Imaging: VLM, Segmentation, Reconstruction, Registration

3.1 Medical Segmentation

3.2 Medical Reconstruction / Inverse Problems / Imaging

3.3 Medical VLMs / Uncertainty / Reliability

3.4 Medical Registration

4. Segmentation, Detection, Recognition, Tracking, Correspondence

4.1 Image / Video Segmentation

4.2 Detection / OOD / Long-tail / Category Discovery

4.3 Correspondence, Keypoints, Optical Flow, Tracking, Representations

4.4 Specialized Recognition / Human Motion / Dance / Sign Language

5. VLM / MLLM / Multimodal Reasoning / Safety / Benchmarks

5.1 Multimodal Reasoning & Capability Analysis

5.2 MLLM / VLM Safety & Attack / Defense

5.3 Video / Vision-Language Model Capabilities

6. Robotics, Embodied AI, Navigation, Manipulation, Autonomous Driving, World Models

6.1 Embodied / Navigation / Manipulation

6.2 Autonomous Driving / World Models / Scenario Mining

6.3 Game Agents

7. Model Compression, Training Efficiency, Optimization, Federated Learning, Task Arithmetic

7.1 Compression / Acceleration / Edge

7.2 Inference & Architecture Efficiency

7.3 Optimization / Federated / Clustering / Theory

7.4 Dataset Distillation / Data Quality

9. Datasets & Benchmarks