Skip to content
arXiv cs.CV · Papers

Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation

arXiv:2608.11335v1 Announce Type: new Abstract: Clinical text can narrow down what to segment, but recent text-guided designs emphasize spatial alignment while overlooking frequency content that governs texture and boundaries. We propose Dual-Domain Cross-Modal Decoding (DD-CMD) for clinical text-guided pulmonary infec