Publications

Many of these publications are copyrighted by their respective publishers. Downloadable versions are not necessarily identical to the published versions. They are made available here for personal use only.

2026

  1. eventvl.png
    EventVL: Understand Event Streams via Multimodal Large Language Model
    Pengteng Li, Yunfan Lu, Pinhao Song, Wuyang Li, Huizai Yao, and Hui Xiong
    IEEE Transactions on Multimedia, 2026
  2. fu2025a-0.jpg
    Tasc: Task-aware shared control for teleoperated manipulation
    Ze Fu, Pinhao Song, Yutong Hu, and Renaud Detry
    In IEEE/RSJ International Conference on Intelligent Robots and Systems, 2026
  3. rtv2.gif
    Robot Trajectron V2: A Probabilistic Shared Control Framework for Navigation
    Pinhao Song, Yurui Du, Ophelie Saussus, Sofie De Schrijver, Irene Caprara, Peter Janssen, and Renaud Detry
    International Journal of Robotics Research, 2026
  4. raglro.png
    RAGLRO: Retrieval-Augmented Generation With Large Language Models for Robotic Operations
    Wenrui Wang, Penghong Wang, Yang Chen, Xianqi Zhang, Pinhao Song, Oleg Cherkasov, and Xiaopeng Fan
    CAAI Transactions on Intelligence Technology, 2026
  5. du2026a-0.jpg
    ELVIS: Ensemble-Calibrated Latent Imagination for Long-Horizon Visual MPC
    Yurui Du, Pinhao Song, Yutong Hu, and Renaud Detry
    In Robotics: Science and Systems (RSS), 2026
  6. deblursplat.png
    DeblurSplat: Traditional SfM-free 3D Gaussian Splatting with Event Camera for Robust Deblurring
    Pengteng Li, Pinhao Song, Weiyu Guo, Huizai Yao, Yunfan Lu, F. Richard Yu, and Hui Xiong
    IEEE Transactions on Multimedia, 2026
  7. rtv3.gif
    Robot Trajectron V3: A Probabilistic Shared Control Framework for SE(3) Manipulation
    Pinhao Song, Zhongxi Li, Ze Fu, Federico Ulloa Rios, and Renaud Detry
    arXiv preprint arXiv:2607.09315, 2026

2025

  1. minidi.gif
    Mini Diffuser: Fast Multi-Task Diffusion Policy Training Using Two-Level Mini-Batches
    Yutong Hu, Pinhao Song, Kehan Wen, and Renaud Detry
    IEEE Robotics and Automation Letters, 2025
  2. evg.gif
    Equivariant Volumetric Grasping
    Pinhao Song, Yutong Hu, Pengteng Li, and Renaud Detry
    arXiv preprint arXiv:2507.18847, 2025
  3. see-trek.png
    See&Trek: Training-Free Spatial Prompting for Multimodal Large Language Model
    Pengteng Li, Pinhao Song, Wuyang Li, Huizai Yao, Weiyu Guo, Yijie Xu, Dugang Liu, and Hui Xiong
    In Annual Conference on Neural Information Processing Systems (NeurIPS), 2025

2024

  1. mpode.jpg
    Neural ordinary differential equation for irregular human motion prediction
    Yang Chen, Hong Liu, Pinhao Song, and Wenhao Li
    Pattern Recognition Letters, 2024
  2. edge.png
    Edge-guided representation learning for underwater object detection
    Linhui Dai, Hong Liu, Pinhao Song, Hao Tang, Runwei Ding, and Shengquan Li
    CAAI Transactions on Intelligence Technology, 2024
  3. Underwater image clarifying based on human visual colour constancy using double-opponency
    Bin Kong, Jing Qian, Pinhao Song, Jing Yang, and Amir Hussain
    CAAI Transactions on Intelligence Technology, 2024
  4. GCCNet.png
    A gated cross-domain collaborative network for underwater object detection
    Linhui Dai, Hong Liu, Pinhao Song, and Mengyuan Liu
    Pattern Recognition, 2024
  5. rt.gif
    Robot Trajectron: Trajectory Prediction-based Shared Control for Robot Manipulation
    Pinhao Song, Pengteng Li, Erwin Aertbeliën, and Renaud Detry
    In 2024 IEEE International Conference on Robotics and Automation (ICRA), 2024
  6. igd.gif
    Implicit grasp diffusion: Bridging the gap between dense prediction and sampling-based grasping
    Pinhao Song, Pengteng Li, and Renaud Detry
    In 8th Annual Conference on Robot Learning (CoRL), 2024
  7. otocc.png
    OTOcc: Optimal Transport for Occupancy Prediction.
    Pengteng Li, Ying He, F Richard Yu, Pinhao Song, Xingchen Zhou, Guang Zhou, and K Larson
    In IJCAI, 2024

2023

  1. boosting-rcnn.gif
    Boosting R-CNN: Reweighting R-CNN samples by RPN’s error for underwater object detection
    Pinhao Song, Pengteng Li, Linhui Dai, Tao Wang, and Zhan Chen
    Neurocomputing, 2023
  2. DMCL.png
    Achieving domain generalization for underwater object detection by domain mixup and contrastive learning
    Yang Chen, Pinhao Song, Hong Liu, Linhui Dai, Xiaochuan Zhang, Runwei Ding, and Shengquan Li
    Neurocomputing, 2023
  3. baggingrcnn.png
    Bagging R-CNN: Ensemble for object detection in complex traffic scenes
    Pengteng Li, Ying He, Dongfu Yin, F Richard Yu, and Pinhao Song
    In 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2023
  4. IGG.png
    IGG: Improved graph generation for domain adaptive object detection
    Pengteng Li, Ying He, F Richard Yu, Pinhao Song, Dongfu Yin, and Guang Zhou
    In Proceedings of the 31st ACM international conference on multimedia (ACMMM), 2023
  5. Domain Similarity-Perceived Label Assignment for Domain Generalized Underwater Object Detection
    Xisheng Li, Wei Li, Pinhao Song, Mingjun Zhang, and Jie Zhou
    arXiv preprint arXiv:2401.05401, 2023

2022

  1. pfd.png
    Pose-guided feature disentangling for occluded person re-identification based on transformer
    Tao Wang, Hong Liu, Pinhao Song, Tianyu Guo, and Wei Shi
    In Proceedings of the AAAI conference on artificial intelligence, 2022
  2. aoodetr.png
    AO2-DETR: Arbitrary-oriented object detection transformer
    Linhui Dai, Hong Liu, Hao Tang, Zhiwei Wu, and Pinhao Song
    IEEE transactions on circuits and systems for video technology, 2022
  3. EA.png
    Excavating roi attention for underwater object detection
    Xutao Liang and Pinhao Song
    In 2022 IEEE international conference on image processing (ICIP), 2022

2020

  1. DG-YOLO.png
    Towards domain generalization in underwater object detection
    Hong Liu, Pinhao Song, and Runwei Ding
    In 2020 IEEE international conference on image processing (ICIP), 2020