北京通用人工智能研究院BIGAI

科研成果

JurisBench: A Deep Benchmark for Assessing Large Language Models in Professional Legal Practice

Simple Role Assignment is Extraordinarily Effective for Safety Alignment

An LLM-based Agent Simulation Approach to Study Moral Evolution

Action-Perturbation Backdoor Attacks on Partially Observable Multiagent Systems

Modeling Decision-Making with Will for Cooperation in Social Dilemmas

Linking Process to Outcome: Conditonal Reward Modeling for LLM Reasoning