Skip to content
OpenTrain AIFor AI Companies

Datasets and Benchmarks for Offline Safe Reinforcement Learning

Zuxin Liu, Zijian Guo, Haohong Lin, Yihang Yao, Jiacheng Zhu +6 morePublished Jun 15, 2023
arXiv PDF
Researcher verdict
Starting point
Use as implementation starting point
Benchmark evidence
Missing
Not verified yet
Time to first repro
A few days
Plan setup time
Risk flags
1
Review before use

Abstract

Domain fit: AI-core · Core AI workload signals detected from paper context and implementation/artifact evidence.

This paper presents a comprehensive benchmarking suite tailored to offline safe reinforcement learning (RL) challenges, aiming to foster progress in the development and evaluation of safe learning algorithms in both the training and deployment phases. Our benchmark suite contains three packages: 1) expertly crafted safe policies, 2) D4RL-styled datasets along with environment wrappers, and 3) high-quality offline safe RL baseline implementations. We feature a methodical data collection pipeline powered by advanced safe RL algorithms, which facilitates the generation of diverse datasets across 38 popular safe RL tasks, from robot control to autonomous driving. We further introduce an array of data post-processing filters, capable of modifying each dataset's diversity, thereby simulating various data collection conditions. Additionally, we provide elegant and extensible implementations of prevalent offline safe RL algorithms to accelerate research in this area. Through extensive experiments with over 50000 CPU and 800 GPU hours of computations, we evaluate and compare the performance of these baseline algorithms on the collected datasets, offering insights into their strengths, limitations, and potential areas of improvement. Our benchmarking framework serves as a valuable resource for researchers and practitioners, facilitating the development of more robust and reliable offline safe RL solutions in safety-critical applications. The benchmark website is available at \url{www.offline-saferl.org}.

Results and benchmarks

Freshness tier: cold
This paper presents a comprehensive benchmarking suite tailored to offline safe reinforcement learning (RL) challenges, aiming to foster progress in the development and evaluation of safe learning algorithms in both the training and deployment phases.

Implementation

Best maintained implementation now

Recommended
Confidence: High
Reproducibility: Moderate

🚀 A fast safe reinforcement learning library in PyTorch

254 stars · 35 forks · Last push Sep 30, 2024 · MIT license

  • License
  • CI
  • Dependencies
  • Docker

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Strong overlap with paper title keywords

Why this implementation
Confidence: high

liuzuxin/fsrl is the strongest maintained implementation based on ranking signals. CI workflows are present. License is declared (MIT).

Open liuzuxin/fsrl
Reproduction risks
  • Dependency manifest is missing
  • Selected liuzuxin/fsrl as the strongest maintained implementation for new work.
  • Includes CI workflow signals.
  • Repository activity is within the last 24 months.
  • Official repository is preserved separately as historical context.

Compare implementation paths

Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.

liuzuxin/fsrl
best maintained
Maintenance
Stale
Confidence
High
Reproducibility
Moderate
Stars
254
Last push
Sep 30, 2024 (694d)

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

  • No push in 12+ months
  • No tagged releases
  • No Docker setup
liuzuxin/osrl
historical official
Maintenance
Stale
Confidence
High
Reproducibility
Limited
Stars
249
Last push
Sep 13, 2024 (711d)

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

  • No push in 12+ months
  • No CI pipeline detected
  • No Docker setup
liuzuxin/dsrl
alternative
Maintenance
Stale risk
Confidence
High
Reproducibility
Limited
Stars
136
Last push
Nov 12, 2025 (286d)

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

  • No CI pipeline detected
  • No Docker setup
  • Dependency manifest missing

Reproduction readiness

Time to first repro: days
Last checked: Aug 23, 2026

Major work

No dependency manifest, manual reconstruction required

  • liuzuxin/fsrl has no requirements.txt, environment.yml, pyproject.toml, or Dockerfile.
  • You will need to reverse-engineer dependencies from import statements in the source code.
  • Last push was 694 days ago.
Open liuzuxin/fsrl

Hardware requirements

  • Through extensive experiments with over 50000 CPU and 800 GPU hours of computations, we evaluate and compare the performance of these baseline algorithms on the collected datasets,

Repositories and ecosystem

Official

  • liuzuxin/dsrl
    Confidence: High

    🔥 Datasets and env wrappers for offline safe reinforcement learning

    136 stars · 7 forks · Last push Nov 12, 2025 · Apache-2.0 license

Community

No additional community repositories detected yet.

Hugging Face artifacts

No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.

Research context

Tasks

Autonomous driving

Methods

Reinforcement learning

Domains

Autonomous Driving

Evaluation and human feedback data

Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.

Open in HFEPX
Explore similar papers

Jump to Paper2Code search queries derived from this paper's research context.

Data includes links from Papers with Code ( CC-BY-SA-4.0 ).