Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation
REVEAL (Representation-driven Endoscopic Visual Embedding Alignment), the largest generative foundation model for endoscopy to date, trained on GastroNet-5M (GN-5M), a multicenter dataset of 5 million endoscopic frames, delivers performance that is competitive with, and in several cases exceeds, endoscopic foundation m...