dronefreak/rdd2022-rfdetr-medium

Model

RF-DETR Medium Finetuned on RDD2022 Road Damage

0

3 commits

3 linked in READMEs

updated Oct 1, 2026

See the code

README

RF-DETR Medium Finetuned on RDD2022 Road Damage

Fine-tuned RF-DETR Medium object detector on the RDD2022 Road Damage benchmark dataset, trained and evaluated as part of DetectionBench -- a framework for reproducibly benchmarking modern object detectors with identical training recipes and evaluation metrics across multiple real-world datasets.

RDD2022 Road Damage Detection Demo


Task Framework Base Model
mAP@50 mAP@50:95 Params
License Source

Usage

Install Dependencies

pip install rfdetr huggingface_hub

Load Model from Hugging Face

from huggingface_hub import hf_hub_download
import rfdetr

weights = hf_hub_download(
    repo_id="dronefreak/rdd2022-rfdetr-medium",
    filename="checkpoint_best_total.pth"
)

model = rfdetr.RFDETRMedium(pretrain_weights=weights)

Run Inference

detections = model.predict("image.jpg", threshold=0.25)

Performance

Evaluated on the RDD2022 Road Damage test split, using DetectionBench's standard evaluation pipeline (detectionbench-evaluate).

MetricScore (%)
mAP@5065.08
mAP@50-9536.02
Precision71.18
Recall55.98
F1 Score62.67
Parameters33.7M
FLOPsN/A (not published upstream)

RDD2022 Road Damage Model Zoo

Every model DetectionBench has trained and evaluated on RDD2022 Road Damage so far, for full transparency -- see DetectionBench for the smaller, curated comparison set used on the project README.

ModelmAP@50mAP@50-95PrecisionRecall
RF-DETR Medium65.0836.0271.1855.98
RF-DETR Small64.7135.7365.6959.41
YOLOv8m62.0334.0865.6157.42
YOLOv8s61.4533.5364.5557.18
YOLO26s61.2733.364.4257.13
YOLO26m61.2433.3863.757.27
RF-DETR Nano60.8533.2265.4954.3
YOLOv8n58.832.0562.0356.08
YOLO11x51.0526.5456.6449.54

Per-Class Performance

ClassmAP@50mAP@50-95
longitudinal_crack59.5731.99
transverse_crack59.8630.22
alligator_crack66.2235.37
pothole74.6846.48

This model was evaluated with Supervision's detection metrics, which report mAP/Precision/Recall directly but don't produce a confusion-matrix plot the way Ultralytics' validator does.


Dataset

This model was trained on RDD2022 Road Damage. For the full dataset description, provenance, license, and citation, see the dataset card:

https://huggingface.co/datasets/dronefreak/RDD2022

Classes

  • longitudinal_crack
  • transverse_crack
  • alligator_crack
  • pothole

Training Configuration

SettingValue
DatasetRDD2022 Road Damage
FrameworkRF-DETR
Training ToolkitDetectionBench
Epochs (configured max)50
Epochs (actually trained)24
Early Stopping Patience10
Batch Size5
Resolution576
Optimizeradamw
Learning Rate0.0001
Seed42

Repository Contents

checkpoint_best_total.pth
metrics.csv
config.json
rdd2022_rfdetr-medium_showcase.jpg
README.md


Training Framework

This model was trained using DetectionBench, an open-source framework for benchmarking object detectors across multiple real-world datasets with a common pipeline.

Features include:

  • A dataset-adapter registry for converting real-world datasets into a canonical format
  • Identical training/evaluation recipes across model families (Ultralytics YOLO/RT-DETR, RF-DETR)
  • Hardware profiling (latency, FPS, VRAM, parameters, FLOPs)
  • One-command reproducibility via versioned Hydra configs

If you find this model useful, please consider starring the repository.


Known Limitations

  • Not comparable to the official CRDDC2022 leaderboard: the challenge test set has no public labels, so the test split here is a held-out 15% slice (70/15/15 split) of the publicly-labelled images, merged across countries -- scores are only comparable between the models listed in this card's Model Zoo.
  • Class imbalance: longitudinal_crack (44.0%) is the most common class, while pothole (18.1%) and alligator_crack (17.9%) are the rarest of the four -- per-class accuracy differs noticeably between them.
  • Four-class taxonomy only: the source data's 5th "other" bucket (block cracks, road repairs and country-specific codes, ~6.5k boxes) was dropped to match the four damage types the CRDDC2022 challenge scores, so those damage types are not detected.
  • Sparse, thin targets: about a third of images contain no in-taxonomy damage (clean-road frames), with 1.5 boxes per image on average, and cracks are thin structures that are easily lost when large frames (some over 4000 pixels wide) are downscaled to the model's input size.
  • Uneven country and imaging-setup mix: the images come from six countries and several capture setups (smartphone, dashboard camera, drone) in very different proportions, so performance can vary substantially by country and generalization to unseen regions or damage conventions is untested.
  • Share-alike data: the RDD2022 images are CC BY-SA 4.0 -- see the Dataset section above for attribution and the dataset card for the full terms.

Citation

If you use this model in your research, please consider citing the dataset and the model architecture:

@article{arya2022rdd2022,
  title = {RDD2022: A multi-national image dataset for automatic Road Damage Detection},
  author = {Arya, Deeksha and Maeda, Hiroya and Ghosh, Sanjay Kumar and Toshniwal, Durga and Sekimoto, Yoshihide},
  journal = {arXiv preprint arXiv:2209.08538},
  year = {2022}
}
@inproceedings{robinson2026rfdetr,
  title     = {RF-DETR: Real-Time Detection Transformer},
  author    = {Robinson, Isaac and Robicheaux, Peter and Popov, Matvei and Ramanan, Deva and Peri, Neehar},
  booktitle = {International Conference on Learning Representations (ICLR)},
  year      = {2026},
  url       = {https://arxiv.org/abs/2511.09554}
}

@article{oquab2023dinov2,
  title={DINOv2: Learning Robust Visual Features without Supervision},
  author={Oquab, Maxime and Darcet, Timoth{\'e}e and Moutakanni, Theo and Vo, Huy and Szafraniec, Marc and Khalidov, Vasil and Fernandez, Pierre and Haziza, Daniel and Massa, Francisco and El-Nouby, Alaaeldin and others},
  journal={arXiv preprint arXiv:2304.07193},
  year={2023}
}
autonomous-driving
computer-vision
detectionbench
infrastructure
model-index
object-detection
pavement-distress
pothole-detection
pytorch
rfdetr
road-damage

dronefreak/rdd2022-rfdetr-medium

Model

RF-DETR Medium Finetuned on RDD2022 Road Damage

0

3 commits

3 linked in READMEs

updated Oct 1, 2026

See the code

README

RF-DETR Medium Finetuned on RDD2022 Road Damage

Fine-tuned RF-DETR Medium object detector on the RDD2022 Road Damage benchmark dataset, trained and evaluated as part of DetectionBench -- a framework for reproducibly benchmarking modern object detectors with identical training recipes and evaluation metrics across multiple real-world datasets.

RDD2022 Road Damage Detection Demo


Task Framework Base Model
mAP@50 mAP@50:95 Params
License Source

Usage

Install Dependencies

pip install rfdetr huggingface_hub

Load Model from Hugging Face

from huggingface_hub import hf_hub_download
import rfdetr

weights = hf_hub_download(
    repo_id="dronefreak/rdd2022-rfdetr-medium",
    filename="checkpoint_best_total.pth"
)

model = rfdetr.RFDETRMedium(pretrain_weights=weights)

Run Inference

detections = model.predict("image.jpg", threshold=0.25)

Performance

Evaluated on the RDD2022 Road Damage test split, using DetectionBench's standard evaluation pipeline (detectionbench-evaluate).

MetricScore (%)
mAP@5065.08
mAP@50-9536.02
Precision71.18
Recall55.98
F1 Score62.67
Parameters33.7M
FLOPsN/A (not published upstream)

RDD2022 Road Damage Model Zoo

Every model DetectionBench has trained and evaluated on RDD2022 Road Damage so far, for full transparency -- see DetectionBench for the smaller, curated comparison set used on the project README.

ModelmAP@50mAP@50-95PrecisionRecall
RF-DETR Medium65.0836.0271.1855.98
RF-DETR Small64.7135.7365.6959.41
YOLOv8m62.0334.0865.6157.42
YOLOv8s61.4533.5364.5557.18
YOLO26s61.2733.364.4257.13
YOLO26m61.2433.3863.757.27
RF-DETR Nano60.8533.2265.4954.3
YOLOv8n58.832.0562.0356.08
YOLO11x51.0526.5456.6449.54

Per-Class Performance

ClassmAP@50mAP@50-95
longitudinal_crack59.5731.99
transverse_crack59.8630.22
alligator_crack66.2235.37
pothole74.6846.48

This model was evaluated with Supervision's detection metrics, which report mAP/Precision/Recall directly but don't produce a confusion-matrix plot the way Ultralytics' validator does.


Dataset

This model was trained on RDD2022 Road Damage. For the full dataset description, provenance, license, and citation, see the dataset card:

https://huggingface.co/datasets/dronefreak/RDD2022

Classes

  • longitudinal_crack
  • transverse_crack
  • alligator_crack
  • pothole

Training Configuration

SettingValue
DatasetRDD2022 Road Damage
FrameworkRF-DETR
Training ToolkitDetectionBench
Epochs (configured max)50
Epochs (actually trained)24
Early Stopping Patience10
Batch Size5
Resolution576
Optimizeradamw
Learning Rate0.0001
Seed42

Repository Contents

checkpoint_best_total.pth
metrics.csv
config.json
rdd2022_rfdetr-medium_showcase.jpg
README.md


Training Framework

This model was trained using DetectionBench, an open-source framework for benchmarking object detectors across multiple real-world datasets with a common pipeline.

Features include:

  • A dataset-adapter registry for converting real-world datasets into a canonical format
  • Identical training/evaluation recipes across model families (Ultralytics YOLO/RT-DETR, RF-DETR)
  • Hardware profiling (latency, FPS, VRAM, parameters, FLOPs)
  • One-command reproducibility via versioned Hydra configs

If you find this model useful, please consider starring the repository.


Known Limitations

  • Not comparable to the official CRDDC2022 leaderboard: the challenge test set has no public labels, so the test split here is a held-out 15% slice (70/15/15 split) of the publicly-labelled images, merged across countries -- scores are only comparable between the models listed in this card's Model Zoo.
  • Class imbalance: longitudinal_crack (44.0%) is the most common class, while pothole (18.1%) and alligator_crack (17.9%) are the rarest of the four -- per-class accuracy differs noticeably between them.
  • Four-class taxonomy only: the source data's 5th "other" bucket (block cracks, road repairs and country-specific codes, ~6.5k boxes) was dropped to match the four damage types the CRDDC2022 challenge scores, so those damage types are not detected.
  • Sparse, thin targets: about a third of images contain no in-taxonomy damage (clean-road frames), with 1.5 boxes per image on average, and cracks are thin structures that are easily lost when large frames (some over 4000 pixels wide) are downscaled to the model's input size.
  • Uneven country and imaging-setup mix: the images come from six countries and several capture setups (smartphone, dashboard camera, drone) in very different proportions, so performance can vary substantially by country and generalization to unseen regions or damage conventions is untested.
  • Share-alike data: the RDD2022 images are CC BY-SA 4.0 -- see the Dataset section above for attribution and the dataset card for the full terms.

Citation

If you use this model in your research, please consider citing the dataset and the model architecture:

@article{arya2022rdd2022,
  title = {RDD2022: A multi-national image dataset for automatic Road Damage Detection},
  author = {Arya, Deeksha and Maeda, Hiroya and Ghosh, Sanjay Kumar and Toshniwal, Durga and Sekimoto, Yoshihide},
  journal = {arXiv preprint arXiv:2209.08538},
  year = {2022}
}
@inproceedings{robinson2026rfdetr,
  title     = {RF-DETR: Real-Time Detection Transformer},
  author    = {Robinson, Isaac and Robicheaux, Peter and Popov, Matvei and Ramanan, Deva and Peri, Neehar},
  booktitle = {International Conference on Learning Representations (ICLR)},
  year      = {2026},
  url       = {https://arxiv.org/abs/2511.09554}
}

@article{oquab2023dinov2,
  title={DINOv2: Learning Robust Visual Features without Supervision},
  author={Oquab, Maxime and Darcet, Timoth{\'e}e and Moutakanni, Theo and Vo, Huy and Szafraniec, Marc and Khalidov, Vasil and Fernandez, Pierre and Haziza, Daniel and Massa, Francisco and El-Nouby, Alaaeldin and others},
  journal={arXiv preprint arXiv:2304.07193},
  year={2023}
}
autonomous-driving
computer-vision
detectionbench
infrastructure
model-index
object-detection
pavement-distress
pothole-detection
pytorch
rfdetr
road-damage