Skip to content

About

CASIA Intelligent Flight Technology Team — profile page

Resources

Stars

3 stars

Watchers

0 watching

Forks

Latest commit

 

History

18 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 

Repository files navigation

CASIA-Collect-AI

Unmanned Collective Intelligence Team Institute of Automation, Chinese Academy of Sciences (CASIA)

Curating high-quality open-source research code in Multi-Agent RL, LLM, and Robotics

English | 中文

GitHub Lab Contact


👨‍🏫 About the Team

We are the Unmanned Collective Intelligence Team (无人集群智能团队), led by Prof. Zhiqiang Pu (蒲志强) at the Institute of Automation, Chinese Academy of Sciences (CASIA).

CASIA-Collect-AI is our open-source code collection platform, curating and maintaining high-quality research code from our team and affiliated researchers in MARL, LLM, and robotics.

Our group is affiliated with the National Key Laboratory of Cognition and Decision Intelligence for Complex Systems (认知与决策智能全国重点实验室), where Prof. Pu serves as deputy director. We are also part of the School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS).

Our research operates at the intersection of:

  • 🤖 Multi-Agent Reinforcement Learning (MARL) — cooperative, heterogeneous, large-scale
  • 🧠 LLM × MARL — RL-based LLM fine-tuning, alignment, and parameter-efficient methods
  • ⚽ Football / Sports AI — data-driven tactic generation, VLM-guided policy alignment
  • 🚁 Autonomous Unmanned Systems — UAV formation control, nonlinear flight control

🔬 Research Directions

Direction Description Representative Work
MARL Foundations Heterogeneity, parameter sharing, policy distance MADPS (paper) (AAMAS'24 Oral), HetDPS (paper) (AAMAS'26 Oral)
Sparse Reward MARL Lazy agents, responsibility diffusion, curriculum learning LazyAgents (paper) (ICML'23)
LLM × MARL Fine-tuning LLMs with cooperative MARL, MoE PEFT CORY (paper) (NeurIPS'24)
MARL Platforms General simulation environments for MARL research Unreal-MAP (paper) (AAAI'26 Oral)
Football AI VLM-based reward shaping, generative tactic discovery V-GEPF (paper) (AAAI'25), TacEleven (paper)
Hierarchical Agents Nested LM agents for complex long-horizon tasks agent-matrix

📚 Some of the Representative Publications

2026

Paper Venue Repo
Unreal-MAP: Unreal-Engine-Based General Platform for Multi-Agent RL AAAI 2026 (Oral) →
HetDPS: Heterogeneity in Multi-Agent Reinforcement Learning AAMAS 2026 (Oral) →

2025

Paper Venue Repo
TacEleven: Generative Tactic Discovery for Football Open Play arXiv:2511.13326 →
V-GEPF: Vision-Based Generic Potential Function for Policy Alignment in MARL AAAI 2025 →
CoMoE: Contrastive Representation for MoE in Parameter-Efficient Fine-tuning EMNLP 2025 —
Cognition-Oriented Multiagent Reinforcement Learning IEEE TNNLS —
A Policy Resonance Approach to Solve Responsibility Diffusion in MARL IEEE TNNLS —
Self-Clustering Hierarchical MARL With Extensible Cooperation Graph IEEE TETCI —
Hybrid Actor-Critic for Physically Heterogeneous MARL IEEE TCDS —
Efficient Multitask RL via Task-Specific Action Correction IEEE TCDS —

2024

Paper Venue Repo
CORY: Coevolving with the Other You — Fine-Tuning LLM with Sequential Cooperative MARL NeurIPS 2024 →
MAPD: Measuring Policy Distance for Multi-Agent Reinforcement Learning AAMAS 2024 (Oral) →
Orientation and Decision-Making for Soccer Based on Sports Analytics and AI IEEE/CAA JAS —
Fuzzy Feedback MARL for Adversarial Dynamic Multiteam Competitions IEEE TFS —
QFuture: Learning Future Expectation Cognition in MARL IEEE TCDS —
Long-Term and Short-Term Opponent Intention Inference for Football MARL IEEE TCDS —
Multiexperience-Assisted Efficient Multiagent Reinforcement Learning IEEE TNNLS —

2023

Paper Venue Repo
Lazy Agents: A New Perspective on Solving Sparse Reward in MARL ICML 2023 →
Attention Enhanced Reinforcement Learning for Multi-Agent Cooperation IEEE TNNLS —
Deep RL for Multiagent Formation Control With Collision Avoidance IEEE TSMC:S —
Deep-RL-Based Multitarget Coverage With Connectivity Guaranteed IEEE TII —
Automatic Curriculum Learning for Large-Scale Cooperative MARL IEEE TETCI —
Cognition-Driven Multiagent Policy Learning for Promoting Cooperation IEEE TG —
Learning to Play Football From Sports Domain Perspective IEEE TG —

2022

Paper Venue
ConcNet: Concentration Network for RL of Large-Scale Multi-Agent Systems AAAI 2022
Multi-Target Encirclement with Collision Avoidance via Deep RL ICRA 2022
Relative Distributed Formation and Obstacle Avoidance with Multi-Agent RL ICRA 2022
Fixed-Time Adaptive Fuzzy Control for Uncertain Nonstrict-Feedback Systems (Highly Cited, WoS) IEEE TFS

🗂️ Repository Index

Repository Topic Paper
LLM-MARL-CORY LLM fine-tuning via cooperative MARL NeurIPS 2024
LLM-Football-TacEleven LLM-based generative football tactic discovery arXiv 2025
MARL-Diversity-HetDPS Heterogeneity & dynamic parameter sharing AAMAS 2026 (Oral)
MARL-Diversity-MADPS Policy distance metric for MARL AAMAS 2024 (Oral)
MARL-Environment-UnrealMAP Unreal Engine MARL simulation platform AAAI 2026 (Oral)
MARL-Football-VGEPF VLM-based reward shaping for football MARL AAAI 2025
MARL-Reward-LazyAgents Sparse reward & lazy agent problem in MARL ICML 2023
agent-matrix Hierarchical nested LM agents (new project) —

📢 Recruitment | 招聘

🔥 ZKDX Intelligence (中科序智能) is hiring!

Embodied AI · Multimodal LLMs · Robotics

👉 View details: recruitment.md


📬 Contact


CASIA-Collect-AI curates and maintains open-source research code from the Unmanned Collective Intelligence Team at CASIA. All repos include bilingual (English/Chinese) documentation.

About

CASIA Intelligent Flight Technology Team — profile page

Resources

Stars

3 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors