Knowledge-Guided Visual Analytics for Scientific Image Annotation

A deep learning pipeline for automatic annotation and retrieval of scientific figures. The system builds on CLIP and BiomedCLIP — including fine-tuned variants — to embed biomedical images and text into a shared semantic space, and integrates Wikidata-based knowledge graphs to enrich that space with structured domain knowledge. The result is stronger semantic alignment between figures and their descriptions, along with interpretable visualisations that make the annotation and retrieval behaviour of the models easier to inspect.