OnCo
technologiesTechnologyEmerging

Merlin (Stanford abdominal CT vision-language model)

Merlin is a model trained on 15,000 CT scans with their reports that can find and describe hundreds of findings.

Merlin (2024) is a 3D vision-language model trained on 6M images and 6M EHR codes from abdominal CTs; supports zero-shot finding classification and report generation.

Generic schematic · not to scale · placeholder for the ai computation front
Flagged finding · Neural network

How it works

3D image encoder aligned with report text and structured codes.

Strengths
  • Report-supervised 3D learning
Limitations
  • Single institution
  • Abdomen only
Since
2024

Latest papers

top
Latest papers · live from Europe PMC
Open in Europe PMC

Query for this technology: (TITLE:"Merlin" OR ABSTRACT:"Merlin" OR TITLE:"Stanford abdominal CT vision-language model" OR ABSTRACT:"Stanford abdominal CT vision-language model") AND (cancer OR tumor OR tumour OR oncology OR carcinoma OR lymphoma OR leukemia OR leukaemia OR myeloma OR sarcoma OR melanoma OR glioma). Results are unfiltered search hits about Merlin (Stanford abdominal CT vision-language model), not a curated reading list.

Connected

5top

Pages like this

not linked directly; found by shared links