PMLB v1.0: An open source dataset collection for benchmarking machine learning methods
Abstract
Domain fit: AI-adjacent · Paper appears method- or tooling-adjacent to AI workflows with partial ecosystem coverage.
Motivation: Novel machine learning and statistical modeling studies rely on standardized comparisons to existing methods using well-studied benchmark datasets. Few tools exist that provide rapid access to many of these datasets through a standardized, user-friendly interface that integrates well with popular data science workflows. Results: This release of PMLB provides the largest collection of diverse, public benchmark datasets for evaluating new machine learning and data science methods aggregated in one location. v1.0 introduces a number of critical improvements developed following discussions with the open-source community. Availability: PMLB is available at https://github.com/EpistasisLab/pmlb. Python and R interfaces for PMLB can be installed through the Python Package Index and Comprehensive R Archive Network, respectively.
Results and benchmarks
Motivation: Novel machine learning and statistical modeling studies rely on standardized comparisons to existing methods using well-studied benchmark datasets.
Benchmark evidence is limited
Evidence graph: 4 refs, 4 links.
Utility signals: depth 60/100, grounding 85/100, status medium.
Implementation
Best maintained implementation now
PMLB: A large, curated repository of benchmark datasets for evaluating supervised machine learning algorithms.
873 stars · 143 forks · Last push Feb 25, 2025 · MIT license
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Matched via arXiv identifier search
EpistasisLab/pmlb is the strongest maintained implementation based on ranking signals. CI workflows are present. License is declared (MIT).
Open EpistasisLab/pmlb- Dependency manifest is missing
- Selected EpistasisLab/pmlb as the strongest maintained implementation for new work.
- Includes CI workflow signals.
- Repository activity is within the last 24 months.
- Official repository is preserved separately as historical context.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 873
- Last push
- Feb 25, 2025 (547d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No Docker setup
- Dependency manifest missing
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 11
- Last push
- Mar 11, 2025 (533d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No Docker setup
- Dependency manifest missing
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 0
- Last push
- Oct 14, 2020 (2142d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No tagged releases
- No Docker setup
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- EpistasisLab/pmlb has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
- Last push was 547 days ago.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Validation caveat
Repositories and ecosystem
Official
- EpistasisLab/pmlb-manuscriptConfidence: High
Manuscript for PMLB v1.0
0 stars · 5 forks · Last push Oct 14, 2020 · NOASSERTION license
Community
No additional community repositories detected yet.
These repositories had low-confidence matching signals and are hidden by default.
- nyuvis/SliceLens
Confidence: Low · 20 stars
- cran/pmlbr
Confidence: Low · 0 stars
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- Granddyser/flux-klein-9b-Biglove-Collection
67,799 downloads · 45 likes
- strangerzonehf/Flux-Ultimate-LoRA-Collection
10,339 downloads · 123 likes
- BigDannyPt/Illustrious-GGUF-Collection
10,082 downloads · 37 likes
Broaden model search
Datasets
- Ujjwal-Tyagi/ai-ml-foundations-book-collection
2,557 downloads · 78 likes · Updated Apr 24, 2026
- dbarbedillo/SMS_Spam_Multilingual_Collection_Dataset
359 downloads · 14 likes · Updated Jan 13, 2023
Broaden dataset search
Spaces
Broaden space search
Research context
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXData includes links from Papers with Code ( CC-BY-SA-4.0 ).