Selected Publications (*Corresponding Author)
TourMLLM: A Retrieval-Augmented Multimodal Large Language Model for Multitask Learning in the Tourism Domain
,
H. Yamanishi, L. Xiao* (corresponding author), and T. Yamasaki,
ACM ICMR , pp. 1654–1663, 2025, Best paper award!
News page!
Fig. 1. Overview of TourMLLM: retrieval-augmented pipeline for tourism tasks.
A Multihead Continual Learning Framework for Fine-Grained Fashion Image Retrieval with Contrastive Learning and Exponential Moving Average Distillation
,
L. Xiao and T. Yamasaki,
IEEE Transactions on Multimedia , pp. 1-10, 2026.
Fig. 1. Detailed pipeline of the proposed MCL-FIR model.
Multi-level Knowledge Distillation for Fine-grained Fashion Image Retrieval
,
L. Xiao and T. Yamasaki,
Knowledge-Based Systems , vol. 310, p. 112955, 2025.
Fig. 1. Details of the proposed MKD.
MAction-SocialNav: Multi-Action Socially Compliant Navigation via Reasoning-enhanced Prompt Tuning
,
Z. Wang, X. Zhang, Z. Liu, T. Kawabata, D. Song, X. Xiao, and L. Xiao* ,
IEEE Robotics and Automation Letters (RA-L) , vol. 11, no. 8, pp. 9747-9754, 2026.
Fig. 1. Detailed pipeline of the proposed MAction-SocialNav model.
E-SocialNav: Efficient Socially Compliant Navigation with Language Models
,
L. Xiao , D. Song, X. Xiao, and T. Yamasaki,
2026 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , pp. 20077-20081, 2026.
Fig. 1. The detailed structure of E-SocialNav.
Fig. 2. Visualizations: E-SocialNav accurately captures social-compliance cues from the image.