End-to-End Training of Neural Retrievers for Open-Domain Question Answering
Results and benchmarks
End-to-End Training of Neural Retrievers for Open-Domain Question Answering is the primary contribution described in this paper.
Benchmark evidence is limited
Evidence graph: 4 refs, 4 links.
Utility signals: depth 55/100, grounding 85/100, status medium.
Implementation
Best maintained implementation now
Ongoing research training transformer models at scale
17,507 stars · 4,382 forks · Last push Aug 21, 2026 · NOASSERTION license
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Community adoption signal (17507 stars)
NVIDIA/Megatron-LM is the strongest maintained implementation based on ranking signals. CI workflows are present. License is declared (NOASSERTION).
Open NVIDIA/Megatron-LM- No repository-level red flags were detected, but paper-specific preprocessing and hyperparameter details may still be under-specified.
- Selected NVIDIA/Megatron-LM as the strongest maintained implementation for new work.
- Includes CI workflow signals.
- Includes dependency/environment manifest signals.
- Repository activity is within the last 24 months.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Active
- Confidence
- High
- Reproducibility
- Strong
- Stars
- 17,507
- Last push
- Aug 21, 2026 (4d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No Docker setup
- Maintenance
- Stale
- Confidence
- Low
- Reproducibility
- Limited
- Stars
- 101
- Last push
- Nov 27, 2022 (1367d)
Partial overlap with paper title keywords · Community adoption signal (101 stars)
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Stale
- Confidence
- Low
- Reproducibility
- Limited
- Stars
- 8
- Last push
- Jun 6, 2023 (1177d)
Matched via arXiv identifier search · Repository appears stale (>24 months since last push)
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
Reproduction readiness
Ready to run
Ready to reproduce
- Clone NVIDIA/Megatron-LM and install dependencies from pyproject.toml.
- CI pipeline detected, so automated tests are in place.
- Last updated 4 days ago.
Quick start
git clone https://github.com/NVIDIA/Megatron-LM.git
pip install -e . Validation caveat
Repositories and ecosystem
No additional verified repositories beyond the primary recommendation.
These repositories had low-confidence matching signals and are hidden by default.
- DevSinghSachan/unsupervised-passage-reranking
Confidence: Low · 101 stars
- ShangQingTu/OpenChatLog
Confidence: Low · 8 stars
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- MaRiOrOsSi/t5-base-finetuned-question-answering
541 downloads · 32 likes
- iarfmoose/t5-base-question-generator
216,879 downloads · 63 likes
- shahrukhx01/question-vs-statement-classifier
29,824 downloads · 47 likes
Broaden model search
Datasets
- google-research-datasets/natural_questions
22,017 downloads · 127 likes · Updated Mar 11, 2024
- stanfordnlp/web_questions
11,716 downloads · 40 likes · Updated Jan 4, 2024
Broaden dataset search
Spaces
- keras-io/question_answering
11 likes
- basakbuluz/turkish-question-answering
7 likes
Broaden space search
Research context
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXData includes links from Papers with Code ( CC-BY-SA-4.0 ).