From Alignment to Synthesis: Contrastive Volumetric Grounding for Text-to-CT Generation
arXiv:2506.00633v4 Announce Type: replace-cross Abstract: Generating semantically controllable 3D CT volumes from radiology reports requires more than a rich text encoder, it requires vision-language alignment grounded in volumetric space. Existing Text-to-CT
arXiv cs.AI··Updated just now·38 sightings