Deep Reinforcement Learning for Multi-Agent Interaction
Abstract
Domain fit: AI-core · Core AI workload signals detected from paper context and implementation/artifact evidence.
The development of autonomous agents which can interact with other agents to accomplish a given task is a core area of research in artificial intelligence and machine learning. Towards this goal, the Autonomous Agents Research Group develops novel machine learning algorithms for autonomous systems control, with a specific focus on deep reinforcement learning and multi-agent reinforcement learning. Research problems include scalable learning of coordinated agent policies and inter-agent communication; reasoning about the behaviours, goals, and composition of other agents from limited observations; and sample-efficient learning based on intrinsic motivation, curriculum learning, causal inference, and representation learning. This article provides a broad overview of the ongoing research portfolio of the group and discusses open problems for future directions.
Results and benchmarks
The development of autonomous agents which can interact with other agents to accomplish a given task is a core area of research in artificial intelligence and machine learning.
Benchmark evidence is limited
Evidence graph: 4 refs, 4 links.
Utility signals: depth 55/100, grounding 85/100, status medium.
Implementation
Best maintained implementation now
An extension of the PyMARL codebase that includes additional algorithms and environment support
732 stars · 191 forks · Last push Sep 24, 2024 · Apache-2.0 license
- License
- CI
- Dependencies
- Docker
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata · Community adoption signal (732 stars)
uoe-agents/epymarl is the strongest maintained implementation based on ranking signals. License is declared (Apache-2.0). Dependency/environment manifests are present.
Open uoe-agents/epymarl- No CI workflows detected
- Selected uoe-agents/epymarl as the strongest maintained implementation for new work.
- Includes dependency/environment manifest signals.
- Repository activity is within the last 24 months.
- Official repository is preserved separately as historical context.
Compare implementation paths
Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 732
- Last push
- Sep 24, 2024 (700d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No CI pipeline detected
- No Docker setup
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 74
- Last push
- Sep 15, 2024 (709d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No tagged releases
- No Docker setup
- Maintenance
- Stale
- Confidence
- High
- Reproducibility
- Moderate
- Stars
- 60
- Last push
- Sep 15, 2024 (709d)
Official implementation from Papers with Code · Repository link is mentioned in the paper metadata
- No push in 12+ months
- No tagged releases
- No Docker setup
Reproduction readiness
Setup required
Dependencies pinned, manual setup needed
- uoe-agents/epymarl has requirements.txt but requires manual environment setup.
- Last push was 700 days ago, so expect possible dependency version conflicts.
- No Dockerfile, so you will set up the environment manually.
- No CI pipeline, so test coverage is unknown.
Quick start
git clone https://github.com/uoe-agents/epymarl.git
pip install -r requirements.txt Validation caveat
Repositories and ecosystem
Official
- uoe-agents/lb-foragingConfidence: High
Level-Based Foraging (LBF): A multi-agent reinforcement learning environment
60 stars · 16 forks · Last push Sep 15, 2024 · MIT license
Community
No additional community repositories detected yet.
Hugging Face artifacts
No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.
Models
- jdopensource/JoyAI-VL-Interaction-Preview
1,548 downloads · 79 likes
- jdopensource/JoyAI-VL-Interaction
2,689 downloads · 16 likes
- BabyLM-community/babylm-interaction-baseline-simpo
2,352 downloads · 2 likes
Broaden model search
Datasets
No trustworthy datasets matches right now.
Search datasets on Hugging FaceSpaces
Research context
Tasks
Agentic tool use
Methods
Reinforcement learning
Domains
AI Agents
Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.
Open in HFEPXJump to Paper2Code search queries derived from this paper's research context.
Data includes links from Papers with Code ( CC-BY-SA-4.0 ).