Publications
* equal contribution · † corresponding author. A complete and continuously updated record is available on Google Scholar.
Preprints & Workshop Papers
- Agent Cybernetics arXiv preprint, 2026
- AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations arXiv preprint, 2026
- MatSeek: An Automated Knowledge-Driven Framework for Materials Research AI4Mat Workshop at the International Conference on Learning Representations (ICLR), 2026
- MATAI: A Generalist Machine Learning Framework for Property Prediction and Inverse Design of Advanced Alloys arXiv preprint, 2025
- AutoMAT: A Hierarchical Framework for Autonomous Alloy Discovery arXiv preprint, 2025
2026
- StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley The 19th European Conference on Computer Vision (ECCV), 2026
- FinMaster: A Holistic Benchmark for Full-Pipeline Financial Management with Large Language Models The 64th Annual Meeting of the Association for Computational Linguistics (ACL), Findings, 2026
- LLM-Based World Models Can Make Decisions Solely, But Rigorous Evaluations Are Needed Transactions on Machine Learning Research (TMLR), 2026
- Nondeterministic Polynomial-time Problem Challenge: An Ever-Scaling Reasoning Benchmark for LLMs Transactions on Machine Learning Research (TMLR), 2026
- FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures Trading Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD), 2026
- FinWorld: An All-in-One Open-Source Platform for End-to-End Financial AI Research and Deployment Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD), 2026
- The Avengers: A Routing Recipe for Collective Intelligence in Language Models Proceedings of the 40th AAAI Conference on Artificial Intelligence (AAAI), 2026
- GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization Proceedings of the 40th AAAI Conference on Artificial Intelligence (AAAI), 2026
2025
- Efficient Integration of External Knowledge to LLM-Based World Models via Retrieval-Augmented Generation and Reinforcement Learning The 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP), Findings, 2025
- FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation The 63rd Annual Meeting of the Association for Computational Linguistics (ACL), 2025 SAC Highlights Award
- Double Oracle Neural Architecture Search for Game Theoretic Deep Learning Models IEEE Transactions on Image Processing (TIP), 2025
- Cradle: Empowering Foundation Agents Towards General Computer Control Proceedings of the 42nd International Conference on Machine Learning (ICML), 2025
- AgentStudio: A Toolkit for Building General Virtual Agents Proceedings of the 2025 International Conference on Learning Representations (ICLR), 2025
2024
- A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and Generalist Proceedings of the 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD), 2024
- Configurable Mirror Descent: Towards a Unification of Decision Making Proceedings of the 41st International Conference on Machine Learning (ICML), 2024
- Reinforcement Nash Equilibrium Solver Proceedings of the 33rd International Joint Conference on Artificial Intelligence (IJCAI), 2024 A previous version appeared at the 23rd International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS), Extended Abstract, 2024 Note: an extended abstract is a bad option — remember to opt out when submitting to AAMAS.
- Self-Adaptive PSRO: Towards an Automatic Population-Based Game Solver Proceedings of the 33rd International Joint Conference on Artificial Intelligence (IJCAI), 2024
- Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control Proceedings of the 2024 International Conference on Learning Representations (ICLR), 2024
- True Knowledge Comes from Practice: Aligning Large Language Models with Embodied Environments via Reinforcement Learning Proceedings of the 2024 International Conference on Learning Representations (ICLR), 2024
- Greedy Sequential Execution: Solving Homogeneous and Heterogeneous Cooperative Tasks with a Unified Framework Proceedings of the 2024 International Conference on Learning Representations (ICLR), 2024
- Grasper: A Generalist Pursuer for Pursuit-Evasion Problems Proceedings of the 23rd International Conference on Autonomous Agents and Multi-Agent Systems (AAMAS), 2024
- Transition-Informed Reinforcement Learning for Large-Scale Stackelberg Mean-Field Games Proceedings of the 38th AAAI Conference on Artificial Intelligence (AAAI), 2024
- EarnHFT: Efficient Hierarchical Reinforcement Learning for High Frequency Trading Proceedings of the 38th AAAI Conference on Artificial Intelligence (AAAI), 2024
- Market-GAN: Adding Control to Financial Market Data Generation with Semantic Context Proceedings of the 38th AAAI Conference on Artificial Intelligence (AAAI), 2024
2023
- Offline RL with Discrete Proxy Representations for Generalizability in POMDPs Proceedings of the 37th Conference on Neural Information Processing Systems (NeurIPS), 2023
- TradeMaster: A Holistic Quantitative Trading Platform Empowered by Reinforcement Learning Proceedings of the 37th Conference on Neural Information Processing Systems (NeurIPS), Datasets and Benchmarks Track, 2023
- Mastering Stock Markets with Efficient Mixture of Diversified Trading Experts Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD), 2023
- Controlling Type Confounding in Ad Hoc Teamwork with Instance-Wise Teammate Feedback Rectification Proceedings of the 40th International Conference on Machine Learning (ICML), 2023
- PRUDEX-Compass: Towards Systematic Evaluation of Reinforcement Learning in Financial Markets Transactions on Machine Learning Research (TMLR), 2023
- Enhancing Meta Learning via Multi-Objective Soft Improvement Functions The 11th International Conference on Learning Representations (ICLR), 2023
- Population-Size-Aware Policy Optimization for Mean-Field Games The 11th International Conference on Learning Representations (ICLR), 2023
- Solving Large-Scale Pursuit-Evasion Games Using Pre-Trained Strategies Proceedings of the 37th AAAI Conference on Artificial Intelligence (AAAI), 2023
2022
- DO-GAN: A Double Oracle Framework for Generative Adversarial Networks Proceedings of the 2022 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2022
2021
- RMIX: Learning Risk-Sensitive Policies for Cooperative Reinforcement Learning Agents Proceedings of the 35th Conference on Neural Information Processing Systems (NeurIPS), 2021
- Neural Regret Matching for Distributed Constraint Optimization Problems Proceedings of the 30th International Joint Conference on Artificial Intelligence (IJCAI), 2021
- CFR-MIX: Solving Imperfect Information Extensive-Form Games with Combinatorial Action Space Proceedings of the 30th International Joint Conference on Artificial Intelligence (IJCAI), 2021
2020
- We Mind Your Well-Being: Preventing Depression in Uncertain Social Networks by Sequential Interventions Proceedings of the 30th International Conference on Automated Planning and Scheduling (ICAPS), 2020
- Learning Expensive Coordination: An Event-Based Deep RL Approach Proceedings of the 2020 International Conference on Learning Representations (ICLR), 2020
2019
- When Players Affect Target Values: Modeling and Solving Dynamic Partially Observable Security Games Proceedings of the 10th Conference on Decision and Game Theory for Security (GameSec), 2019 Outstanding Student Paper Award
- Who Should Pay the Cost: A Game-Theoretic Model for Government Subsidized Investments to Improve National Cybersecurity Proceedings of the 28th International Joint Conference on Artificial Intelligence (IJCAI), 2019
2018
- Catching Captain Jack: Efficient Time and Space Dependent Patrols to Combat Oil-Siphoning in International Waters Proceedings of the 32nd AAAI Conference on Artificial Intelligence (AAAI), 2018
2017
- Stop Nuclear Smuggling Through Efficient Container Inspection Proceedings of the 16th International Joint Conference on Autonomous Agents and Multi-Agent Systems (AAMAS), 2017