arXiv · 2511.04394
DORAEMON: A Unified Library for Visual Object Modeling and Representation Learning at Scale
Abstract
DORAEMON is an open-source PyTorch library that unifies visual object modeling and representation learning across diverse scales. A single YAML-driven workflow covers classification, retrieval and metric learning; more than 1000 pretrained backbones are exposed through a timm-compatible interface, together with modular losses, augmentations and distributed-training utilities. Reproducible recipes match or exceed reference results on ImageNet-1K, MS-Celeb-1M and Stanford online products, while one-command export to ONNX or HuggingFace bridges research and deployment. By consolidating datasets, models, and training techniques into one platform, DORAEMON offers a scalable foundation for rapid experimentation in visual recognition and representation learning, enabling efficient transfer of research advances to real-world applications. The repository is available at https://github.com/wuji3/DORAEMON.
Explore related subjects
Keep this discovery
Ke Du, Yimin Peng, Chao Gao, Fan Zhou, Siqiao Xue. 2025-11-06. DORAEMON: A Unified Library for Visual Object Modeling and Representation Learning at Scale. https://arxiv.org/abs/2511.04394
Cite the original work for its findings. Save a collection to share your selection of sources.