MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models
Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain challenging. Existing approaches either verbalize regions as coordinate strings or rely on external modules that decouple perception from understanding, creating representation gaps for region-language alignment. We present MedUP, a Med-VLM that natively unifies perception and understanding within a shared token space. At its core lies UniMedTok, a regi
Record details
Published: 11 August 2026
Source: arXiv
Category: Research
Topics: Safety & alignment · Healthcare
Retrieved: 12 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
UniMod: Enhancing Multi-Modal Medical Diagnosis through Cross-Modality and Within-Modality Alignment
arXiv cs.LG · 10 August 2026
Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation
arXiv · 11 August 2026
Mr3D-VL: A generalist vision language foundation model for Multiparametric 3D Magnetic Resonance Imaging
arXiv · 13 August 2026
Charting Public Health: A Taxonomic Study of Visualization Practices in the Public Health Field
arXiv cs.HC · 9 August 2026
Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance
arXiv cs.CY · 14 August 2026
Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation
arXiv · 7 August 2026
How to cite this record
ethics.ai (11 August 2026), “MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models,” evidence record 18421, https://ethics.ai/record/18421 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.