Research
Search
Conference & Journals
FUSCO: High-Performance Distributed Data Shuffling via Transformation-Communication Fusion
24th USENIX Symposium on Networked Systems Design and Implementation (NSDI 2027)
·
2027
ICPL: Few-shot In-Context Preference Learning via LLMs
Reinforcement Learning Conference (RLC 2026)
·
2026
Human-Guided Online Reward Adaptation for Real-Robot Arm Manipulation
IEEE Robotics and Automation Letters (RA-L 2026)
·
2026
DynaRL: Flexible and Dynamic Scheduling of Large-Scale Reinforcement Learning Training
20th USENIX Symposium on Operating Systems Design and Implementation (OSDI 2026)
·
2026
LAMP: Latent Motion Prior-Guided Real-World Learning for Dexterous Hand Manipulation
arXiv preprint arXiv:2607.06323 (2026)
·
2026
Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents
arXiv preprint arXiv:2607.08448 (2026)
·
2026
LaWAM: Latent World Action Models for Efficient Dynamics-Aware Robot Policies
Conference on Robot Learning (CoRL 2026)
·
2026
STEAM: Self-Supervised Temporal Ensemble Advantage Modeling for Real-World Robot Learning
arXiv preprint arXiv:2606.29834 (2026)
·
2026
Verifiable Process Rewards for Agentic Reasoning
arXiv preprint arXiv:2605.10325 (2026)
·
2026
StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation
arXiv preprint arXiv:2603.28565 (2026)
·
2026
Translate Policy to Language: Flow Matching Generated Rewards for LLM Explanations
International Conference on Learning Representations (ICLR 2026)
·
2026
AED: Automatic Discovery of Effective and Diverse Vulnerabilities for Autonomous Driving Policy with Large Language Models
IEEE International Conference on Automation Science and Engineering (CASE 2026)
·
2026
VS-Bench: Evaluating VLMs for Strategic Reasoning and Decision-Making in Multi-Agent Environments
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2026), Oral
·
2026
Exploring the Secondary Risks of Large Language Models
Machine Intelligence Research (MIR 2026)
·
2026
D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning
IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026)
·
2026
RLinf: Flexible and Efficient Large-scale Reinforcement Learning via Macro-to-Micro Flow Transformation
20th USENIX Symposium on Operating Systems Design and Implementation (OSDI 2026)
·
2026
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
IEEE Robotics and Automation Letters (RA-L 2026)
·
2026
JuggleRL: Mastering Ball Juggling with a Quadrotor via Deep Reinforcement Learning
IEEE International Conference on Robotics and Automation (ICRA 2026)
·
2026
RE-PO: Robust Enhanced Policy Optimization as a General Framework for LLM Alignment
International Conference on Learning Representations (ICLR 2026)
·
2026
SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling
International Conference on Learning Representations (ICLR 2026)
·
2026
RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models
Robotics: Science and Systems (RSS 2026)
·
2026
MARSHAL: Incentivizing Multi-Agent Reasoning via Self-Play with Strategic LLMs
International Conference on Learning Representations (ICLR 2026)
·
2026
RoboScape-R: Unified Reward-Observation World Models for Generalizable Robotics Training via RL
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2026), Findings
·
2026
WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning
arXiv preprint (2026)
·
2026
USER: A Unified and Extensible System for Real-World Online Policy Learning in Embodied AI
Robotics: Science and Systems (RSS 2026)
·
2026
Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models
Conference on Robot Learning (CoRL 2026)
·
2026
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
Conference on Robot Learning (CoRL 2026)
·
2026
Tex3D: Objects as Attack Surfaces via Adversarial 3D Textures for Vision-Language-Action Models
ACM International Conference on Multimedia (ACM MM 2026)
·
2026
What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study
IEEE Robotics and Automation Letters (RA-L 2025)
·
2025
Few-shot In-context Preference Learning using Large Language Models
International Conference on Learning Representations (ICLR 2025)
·
2025
CityLight: A Universal Model Towards Real-world City-scale Traffic Signal Control Coordination
ACM International Conference on Information and Knowledge Management (CIKM 2025)
·
2025
FlightBench: A Comprehensive Benchmark of Spatial Planning Methods for Quadrotors
IEEE Robotics and Automation Letters (RA-L 2025)
·
2025
Human-Robot Cooperative Distribution Coupling for Hamiltonian-Constrained Social Navigation
IEEE International Conference on Robotics and Automation (ICRA 2025)
·
2025
Multi-UAV Behavior-based Formation with Static and Dynamic Obstacles Avoidance via Reinforcement Learning
IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025)
·
2025
Neural Internal Model Control: Learning a Robust Control Policy via Predictive Error Feedback
IEEE Robotics and Automation Letters (RA-L 2025)
·
2025
Learning Global Nash Equilibrium in Team Competitive Games with Generalized Fictitious Cross-Play
Journal of Machine Learning Research (JMLR 2025), Volume 26
·
2025
Learning from Suboptimal Data in Continuous Control via Auto-Regressive Soft Q-Network
International Conference on Machine Learning (ICML 2025)
·
2025
VolleyBots: A Testbed for Multi-Drone Volleyball Game Combining Motion Control and Strategic Play
Annual Conference on Neural Information Processing Systems (NeurIPS 2025)
·
2025
Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization
International Conference on Machine Learning (ICML 2025)
·
2025
Multi-Robot System for Cooperative Exploration in Unknown Environments: A Survey
arXiv preprint arXiv:2503.07278 (2025)
·
2025
Hysteresis-Aware Neural Network Modeling and Whole-Body Reinforcement Learning Control of Soft Robots
IEEE Robotics and Automation Letters (RA-L 2025)
·
2025
Mastering Multi-Drone Volleyball through Hierarchical Co-Self-Play Reinforcement Learning
Conference on Robot Learning (CoRL 2025)
·
2025
Fine-tuning Diffusion Policies with Backpropagation Through Diffusion Timesteps
arXiv preprint arXiv:2505.10482 (2025)
·
2025
Toward Real-World Cooperative and Competitive Soccer with Quadrupedal Robot Teams
Conference on Robot Learning (CoRL 2025)
·
2025
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
Annual Conference on Neural Information Processing Systems (NeurIPS 2025)
·
2025
What Can RL Bring to VLA Generalization? An Empirical Study
Annual Conference on Neural Information Processing Systems (NeurIPS 2025)
·
2025
Online Planning for Multi-UAV Pursuit-Evasion in Unknown Environments Using Deep Reinforcement Learning
IEEE Robotics and Automation Letters (RA-L 2025)
·
2025
Spec-VLA: Speculative Decoding for Vision-Language-Action Models with Relaxed Acceptance
Conference on Empirical Methods in Natural Language Processing (EMNLP 2025), Main Conference
·
2025
Long-horizon Locomotion and Manipulation on a Quadrupedal Robot with Large Language Models
IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025)
·
2025
πRL: Online RL Fine-tuning for Flow-based Vision-Language-Action Models
arXiv preprint arXiv:2510.25889 (2025)
·
2025
Red Teaming Large Reasoning Models
Annual Meeting of the Association for Computational Linguistics (ACL 2025)
·
2025
Multi-Agent Vulnerability Discovery for Autonomous Driving Policy by Finding AV-Responsible Scenarios
IEEE International Conference on Automation Science and Engineering (CASE 2024)
·
2024
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
International Conference on Machine Learning (ICML 2024)
·
2024
MASP: Scalable Graph-based Planning towards Multi-Agent Navigation
IEEE Robotics and Automation Letters (RA-L 2024)
·
2024
LLM-Powered Hierarchical Language Agent for Real-time Human-AI Coordination
International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2024)
·
2024
Sharing Minds during MARL Training for Enhanced Cooperative LLM Agents
Annual Conference on Neural Information Processing Systems (NeurIPS 2024)
·
2024
OmniDrones: An Efficient and Flexible Platform for Reinforcement Learning in Drone Control
IEEE Robotics and Automation Letters (RA-L 2024)
·
2024
Accelerate Multi-Agent Reinforcement Learning in Zero-Sum Games with Subgame Curriculum Learning
AAAI Conference on Artificial Intelligence (AAAI 2024)
·
2024
Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study
International Conference on Machine Learning (ICML 2024), Oral
·
2024
LAGOON: Language-Guided Motion Control
IEEE International Conference on Robotics and Automation (ICRA 2024)
·
2024
A Survey on Self-play Methods in Reinforcement Learning
arXiv preprint arXiv:2408.01072 (2024)
·
2024
Reward-Robust RLHF in LLMs
arXiv preprint arXiv:2409.15360 (2024)
·
2024
SleepNetZero: Zero-Burden Zero-Shot Reliable Sleep Staging With Neural Networks Based on Ballistocardiograms
ACM International Conference on Ubiquitous Computing (UbiComp 2024)
·
2024
Learning Graph-Enhanced Commander-Executor for Multi-Agent Navigation
International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2023)
·
2023
Asynchronous Multi-Agent Reinforcement Learning for Efficient Real-Time Multi-Robot Cooperative Exploration
International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2023)
·
2023
Learning Zero-Shot Cooperation with Humans, Assuming Humans Are Biased
International Conference on Learning Representations (ICLR 2023)
·
2023
Automatic Truss Design with Reinforcement Learning
International Joint Conference on Artificial Intelligence (IJCAI 2023)
·
2023
Fictitious Cross-Play: Learning Global Nash Equilibrium in Mixed Cooperative-Competitive Games
International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2023)
·
2023
Active Neural Topological Mapping for Multi-Agent Exploration
IEEE Robotics and Automation Letters (RA-L 2023)
·
2023
Revisiting Some Common Practices in Cooperative Multi-Agent Reinforcement Learning
International Conference on Machine Learning (ICML 2022)
·
2022
VMAPD: Generate Diverse Solutions for Multi-Agent Games with Recurrent Trajectory Discriminators
IEEE Conference on Games (CoG 2022)
·
2022
SAVE: Spatial-Attention Visual Exploration
IEEE International Conference on Image Processing (ICIP 2022)
·
2022
Learning Efficient Multi-Agent Cooperative Visual Exploration
European Conference on Computer Vision (ECCV 2022)
·
2022
A Benchmark of Planning-based Exploration Methods in Photo-Realistic 3D Simulator
IEEE International Conference on Robotics and Biomimetics (ROBIO 2022)
·
2022
The Surprising Effectiveness of PPO in Cooperative Multi-Agent Games
Conference on Neural Information Processing Systems (NeurIPS 2022), Datasets and Benchmarks Track
·
2022
INCAME: Interruptible CNN Accelerator for Multirobot Exploration
IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems (TCAD 2022), 41(4):964-978
·
2022
Discovering Diverse Multi-agent Strategic Behavior Via Reward Randomization
International Conference on Learning Representations (ICLR 2021)
·
2021
Unlocking the Potential of MAPPO with Asynchronous Optimization
CAAI International Conference on Artificial Intelligence (CICAI 2021)
·
2021
CNN-based Feature-point Extraction for Real-time Visual SLAM on Embedded FPGA
IEEE Symposium on Field-Programmable Custom Computing Machines (FCCM 2020)
·
2020
CNN-based Monocular Decentralized SLAM on embedded FPGA
IEEE Reconfigurable Architectures Workshop (RAW 2020)
·
2020
INCA: INterruptible CNN Accelerator for Multi-tasking in Embedded Robots
ACM/IEEE Design Automation Conference (DAC 2020)
·
2020
Benchmarking Multi-agent Deep Reinforcement Learning Algorithms
Conference on Neural Information Processing Systems (NeurIPS 2020), Workshop
·
2020
Learning Safety-Aware Policy with Imitation Learning for Context-Adaptive Navigation
Workshop Paper / Technical Report (2019)
·
2019
A DenseNet feature-based loop closure method for visual SLAM system
IEEE International Conference on Robotics and Biomimetics (ROBIO 2019)
·
2019
Long-Sighted Imitation Learning for Partially Observable Control
International Conference on Control and Robot Technology (ICCRT 2019)
·
2019
DS-SLAM: A Semantic Visual SLAM towards Dynamic Environments
IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2018)
·
2018
Multi-robot coordination for high-speed pick-and-place tasks
IEEE International Conference on Robotics and Biomimetics (ROBIO 2017)
·
2017