Meissa: Multi-modal Medical Agentic Intelligence
Multi-modal large language models (MM-LLMs) have shown strong performance in medical image understanding and clinical reasoning. Recent medical agent systems extend them with tool use and multi-agent collaboration, enabling complex decision-making. However, these systems rely almost entirely on frontier models (e.g., GPT), whose API-based deployment incurs high cost, high latency, and privacy risks that conflict with on-premise clinical requirement
By Yixiong Chen, Xinyi Bai, Yue Pan, Zongwei Zhou, Alan Yuille