北京通用人工智能研究院BIGAI

科研成果

Towards Human-Level Bimanual Dexterous Manipulation with Reinforcement Learning

MATE: Benchmarking Multi-Agent Reinforcement Learning in Distributed Target Coverage Control

TarGF: Learning Target Gradient Field for Object Rearrangement

LIGS:Learning Intrinstic-reward Generation Selection for Multi-Agent Learning

ToM2C: Target-oriented Multi-agent Communication and Cooperation with Theory of Mind

Multi-agent Modeling of Crowd Dynamics under Mass Shooting Cases