Latest Research in Multi-Agent Reinforcement Learning

886 research papers · 0.0 average citations · 2026 median publication year

Top Research Topics in Multi-Agent Reinforcement Learning

Highest-Cited Papers

  1. AI Teacher Latent Basis Reorientation
  2. Context-Aware visual scene analysis and adaptation for intelligent english tutoring systems
  3. Attention Is All You Need: A Technical Review of the Transformer Architecture and Its Impact on Modern Artificial Intelligence
  4. Attention Is All You Need: A Technical Review of the Transformer Architecture and Its Impact on Modern Artificial Intelligence
  5. Accelerating Q-learning through Efficient Value-Sharing across Actions
  6. Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces
  7. Reward Shaping based on Trajectory Quality for offline and hybrid reinforcement learning
  8. Source-Free Domain Adaptation with Vision-Language Prior
  9. CFR variant and iterate reporting in small imperfect-information games
  10. Neuromorphic-Inspired Language Identification for Low-Resource Code-Switched Texts Using Spiking Neural Networks
  11. Opponent-modeling-enhanced multi-agent reinforcement learning for dynamic games under incomplete information
  12. State-based Selection of Expert Policies in Disaster Environments and Reinforcement Learning of a Unified Student Policy via Distillation
  13. Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL
  14. Not All Layers Need Tuning: Diagnosing and Directing Adaptation in Vision-Language-Action Models
  15. Region-Level Policy Optimization for Fine-grained MLLM Perception
  16. Precision autotuning for linear solvers via contextual bandit-based RL
  17. CARE-VI: Conservative Adaptive Reliability Estimation for Value Improvement in Off-Policy Actor-Critic Learning
  18. Absence is Presence: Understanding Visual Scene Negative Events Under Safety Cognitive Constraint
  19. L2R: Low-Rank and Lipschitz-Controlled Routing for Mixture-of-Experts
  20. EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

Sub-Regions

Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
L2 Region - - 2026 Sep Q3

Multi-Agent Reinforcement Learning

886 papers

Top Topics (10)

Machine Learning268
Computer Vision and Pattern Recognition259
Artificial Intelligence123
Computation and Language62
Robotics31
Machine Learning16
Multimodal Machine Learning Applications13
Reinforcement Learning in Robotics12
Multiagent Systems10
Systems and Control9

Top Publications (20)

1.AI Teacher Latent Basis Reorientation2.Context-Aware visual scene analysis and adaptation for intelligent english tutoring systems3.Attention Is All You Need: A Technical Review of the Transformer Architecture and Its Impact on Modern Artificial Intelligence4.Attention Is All You Need: A Technical Review of the Transformer Architecture and Its Impact on Modern Artificial Intelligence5.Accelerating Q-learning through Efficient Value-Sharing across Actions6.Scalable Policy Optimization for Networked Multi-Agent Reinforcement Learning with Continuous State-Action Spaces7.Reward Shaping based on Trajectory Quality for offline and hybrid reinforcement learning8.Source-Free Domain Adaptation with Vision-Language Prior9.CFR variant and iterate reporting in small imperfect-information games10.Neuromorphic-Inspired Language Identification for Low-Resource Code-Switched Texts Using Spiking Neural Networks11.Opponent-modeling-enhanced multi-agent reinforcement learning for dynamic games under incomplete information12.State-based Selection of Expert Policies in Disaster Environments and Reinforcement Learning of a Unified Student Policy via Distillation13.Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL14.Not All Layers Need Tuning: Diagnosing and Directing Adaptation in Vision-Language-Action Models15.Region-Level Policy Optimization for Fine-grained MLLM Perception16.Precision autotuning for linear solvers via contextual bandit-based RL17.CARE-VI: Conservative Adaptive Reliability Estimation for Value Improvement in Off-Policy Actor-Critic Learning18.Absence is Presence: Understanding Visual Scene Negative Events Under Safety Cognitive Constraint19.L2R: Low-Rank and Lipschitz-Controlled Routing for Mixture-of-Experts20.EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

Sub-Regions (6)

AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.