Surgical-lvlm: Learning to adapt large vision-language model for grounded visual question answering in robotic surgery
Published in ICLR 2024, 2024
Recommended citation: G Wang, L Bai, WJ Nah, J Wang, Z Zhang, Z Chen, J Wu, M Islam, H Liu, et al. (2024). "Surgical-lvlm: Learning to adapt large vision-language model for grounded visual question answering in robotic surgery." ICLR 2024.
Download Paper
