Results and benchmarks
A Question-Entailment Approach to Question Answering is the primary contribution described in this paper.
| Task | Dataset | Metric | Value | Source |
|---|---|---|---|---|
| Question answering | SNLI | Accuracy | 79.50 | paper-derived |
| Question answering | MultiNLI | Accuracy | 73.71 | paper-derived |
Audit each benchmark finding before selecting an implementation path. Evidence refs map to the disclosure below.
Evidence graph: 4 refs, 4 links.
Utility signals: depth 100/100, grounding 95/100, status high.
Implementation
Best maintained implementation now
Medical Question-Answering datasets prepared for the TREC 2017 LiveQA challenge (Medical Task)
56 stars · 18 forks · Last push May 13, 2026
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Strong overlap with paper title keywords
abachaa/LiveQA_MedicalTask_TREC2017 is the strongest maintained implementation based on ranking signals.
Open abachaa/LiveQA_MedicalTask_TREC2017- License metadata missing
- No CI workflows detected
- Dependency manifest is missing
- Selected abachaa/LiveQA_MedicalTask_TREC2017 as the strongest maintained implementation for new work.
- Repository activity is within the last 24 months.
- Official repository is preserved separately as historical context.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Limited
- Stars
- 462
- Last push
- Oct 17, 2023 (1043d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No CI pipeline detected
- No tagged releases
- Maintenance
- Recently updated
- Confidence
- High
- Reproducibility
- Limited
- Stars
- 56
- Last push
- May 13, 2026 (104d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No CI pipeline detected
- No tagged releases
- No Docker setup
Reproduction readiness
Major work
No dependency manifest, manual reconstruction required
- abachaa/LiveQA_MedicalTask_TREC2017 has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
- You will need to reverse-engineer dependencies from import statements in the source code.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- MaRiOrOsSi/t5-base-finetuned-question-answering
542 downloads · 32 likes
- vasudevgupta/bigbird-roberta-natural-questions
1,065 downloads · 10 likes
- iarfmoose/t5-base-question-generator
214,571 downloads · 63 likes
Broaden model search
Datasets
- google-research-datasets/natural_questions
22,424 downloads · 127 likes · Updated Mar 11, 2024
- stanfordnlp/web_questions
11,509 downloads · 40 likes · Updated Jan 4, 2024
Broaden dataset search
Spaces
- keras-io/question_answering
11 likes
- basakbuluz/turkish-question-answering
7 likes
Broaden space search
Research context
Tasks
Question answering
Methods
None detected
Domains
None detected
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).