Skip to content
OpenTrain AIFor AI Companies
implementation starting point
Benchmarks: thin evidence
Time to repro: a few hours

Results & Benchmarks

Freshness tier: cold
Direct + Inferred Evidence

Some benchmark signal exists in the extracted evidence, but it is not structured strongly enough yet for a confident benchmark decision.

Generating Long Sequences with Sparse Transformers presents a transformer method.

Use This Implementation Because…

Confidence: medium

kyegomez/SparseAttention is the best available implementation candidate based on ranking signals, but recommendation confidence is not yet high. CI workflows are present. License is declared (MIT).

Open kyegomez/SparseAttention

Reproduction Risks

  • No repository-level red flags were detected, but paper-specific preprocessing and hyperparameter details may still be under-specified.
Evidence disclosure

Evidence graph: 3 refs, 3 links.

Utility signals: depth 90/100, grounding 85/100, status high.

Implementation Comparison

Top 3 paths

Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.

openai/sparse_attention
historical official
Maintenance: Archived
Confidence: High
Reproducibility: Limited

Official implementation from Papers with Code · Strong overlap with paper title keywords

Stars
1,614
Last push
Aug 12, 2020 (2181d ago)

Risk flags

  • Repository archived
  • No push in 12+ months
  • No CI pipeline detected
Maintenance: Stale
Confidence: Low
Reproducibility: Moderate

Community adoption signal (1079 stars)

Stars
1,079
Last push
Sep 18, 2024 (683d ago)
Dependencies

Risk flags

  • No push in 12+ months
  • No CI pipeline detected
  • No tagged releases
Maintenance: Archived
Confidence: Low
Reproducibility: Limited

Community adoption signal (10837 stars) · Repository is archived

Stars
10,837
Last push
Jun 16, 2026 (47d ago)
Releases Dependencies

Risk flags

  • Repository archived
  • No CI pipeline detected
  • No Docker setup

Best implementation now

kyegomez/SparseAttention
Confidence: Medium
Reproducibility: Strong

Pytorch Implementation of the sparse attention from the paper: "Generating Long Sequences with Sparse Transformers"

Stars: 97
Forks: 6
Last push: Jul 27, 2026
License: MIT
Matched via arXiv identifier search
Strong overlap with paper title keywords
Community adoption signal (97 stars)
License ✓
CI ✓
Deps ✓
Docker –
  • Selected kyegomez/SparseAttention as the strongest maintained implementation for new work.
  • Includes CI workflow signals.
  • Includes dependency/environment manifest signals.
  • Repository activity is within the last 24 months.

Historical official implementation

Preserved for provenance. Not recommended as the default path for new builds.

openai/sparse_attention
Stars: 1,614
Last push: Aug 12, 2020
Archived

Reproduction readiness

Ready to Run
Time to first repro: hours
Last checked: Aug 2, 2026

Ready to reproduce

  • · Clone kyegomez/SparseAttention and install dependencies from requirements.txt.
  • · CI pipeline detected — automated tests are in place.
  • · Last updated 6 days ago.
Open kyegomez/SparseAttention

Quick start

git clone https://github.com/kyegomez/SparseAttention.git
pip install -r requirements.txt

Additional implementations

No additional verified repositories beyond the primary recommendation.

These repositories had low-confidence matching signals and are hidden by default.

Hugging Face artifacts

No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.

Research context

Tasks

Transformer

Methods

Transformer

Domains

None detected

Evaluation & Human Feedback Data

Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.

Open in HFEPX

Explore Similar Papers

Jump to Paper2Code search queries derived from this paper's research context.