Skip to content
OpenTrain AIFor AI Companies

Scaling Sparse and Dense Retrieval in Decoder-Only LLMs

Published Feb 1, 2025
arXiv PDF
Researcher verdict
Starting point
Use as implementation starting point
Benchmark evidence
Thin evidence
Verify before relying
Time to first repro
A few days
Plan setup time
Risk flags
1
Review before use

Results and benchmarks

Freshness tier: cold
Scaling Sparse and Dense Retrieval in Decoder-Only LLMs focuses on retrieval / indexing.
Task Dataset Metric Value Source
Retrieval ArguAna Lion-DS-8B 0.494 paper-derived

Audit each benchmark finding before selecting an implementation path. Evidence refs map to the disclosure below.

Implementation

Historical official implementation (not recommended for new builds)

Why this implementation
Confidence: low

Only historical official repository was found (hansizeng/scaling-retriever).

Open hansizeng/scaling-retriever
Reproduction risks
  • Only historical official implementation is available
  • No direct maintained implementation is currently verified.
  • Only historical official repository was found: hansizeng/scaling-retriever.
  • No maintained paper-verified implementation met reliability thresholds.

Reproduction readiness

Time to first repro: days
Last checked: Aug 25, 2026

Setup required

Dependencies pinned, manual setup needed

  • hansizeng/scaling-retriever has requirements.txt but requires manual environment setup.
  • Last push was 512 days ago, so expect possible dependency version conflicts.
  • No Dockerfile, so you will set up the environment manually.
  • No CI pipeline, so test coverage is unknown.
Open hansizeng/scaling-retriever

Hardware requirements

  • Expect multi-day setup/compute for meaningful reproduction based on current guidance.

Quick start

git clone https://github.com/hansizeng/scaling-retriever.git
pip install -r requirements.txt

Hugging Face artifacts

No trustworthy direct or curated related Hugging Face artifacts were found yet. Use targeted searches to quickly locate candidate models, datasets, and demos.

Tip: start with models, then check datasets and spaces if you need evaluation data or demos.

Research context

Tasks

Retrieval / indexing

Methods

Retrieval-augmented generation

Domains

Information Retrieval

Evaluation and human feedback data

Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.

Open in HFEPX
Explore similar papers

Jump to Paper2Code search queries derived from this paper's research context.

Data includes links from Papers with Code ( CC-BY-SA-4.0 ).