Results and benchmarks
Scaling Sparse and Dense Retrieval in Decoder-Only LLMs focuses on retrieval / indexing.
| Task | Dataset | Metric | Value | Source |
|---|---|---|---|---|
| Retrieval | ArguAna | Lion-DS-8B | 0.494 | paper-derived |
Audit each benchmark finding before selecting an implementation path. Evidence refs map to the disclosure below.
Evidence graph: 2 refs, 1 links.
Utility signals: depth 95/100, grounding 68/100, status medium.
Implementation
Historical official implementation (not recommended for new builds)
Only a historical official implementation is available
Use with caution for new projects; verify against current tooling and maintained community alternatives.
hansizeng/scaling-retriever · 22 stars · Last push Mar 31, 2025
Only historical official repository was found (hansizeng/scaling-retriever).
Open hansizeng/scaling-retriever- Only historical official implementation is available
- No direct maintained implementation is currently verified.
- Only historical official repository was found: hansizeng/scaling-retriever.
- No maintained paper-verified implementation met reliability thresholds.
Reproduction readiness
Setup required
Dependencies pinned, manual setup needed
- hansizeng/scaling-retriever has requirements.txt but requires manual environment setup.
- Last push was 512 days ago, so expect possible dependency version conflicts.
- No Dockerfile, so you will set up the environment manually.
- No CI pipeline, so test coverage is unknown.
Hardware requirements
- Expect multi-day setup/compute for meaningful reproduction based on current guidance.
Quick start
git clone https://github.com/hansizeng/scaling-retriever.git
pip install -r requirements.txt Hugging Face artifacts
No trustworthy direct or curated related Hugging Face artifacts were found yet. Use targeted searches to quickly locate candidate models, datasets, and demos.
Models
Tip: start with models, then check datasets and spaces if you need evaluation data or demos.
Research context
Tasks
Retrieval / indexing
Methods
Retrieval-augmented generation
Domains
Information Retrieval
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).