Real-time foundation model for endoscopy supports task-specific fine-tuning
Researchers present woma, a foundation model trained without labels on roughly one million gastrointestinal endoscopy frames. Task-specific models are then fine-tuned from this base, and the authors describe a systematic design intended for production deployment, including requirements and performance targets. The work targets real-time use in clinical endoscopy workflows.