Zürich, Switzerland
Curriculum Vitae
The short version. The full PDF is linked below, and this page is always current because it is built from the same content as the rest of the site.
Experience
Aug 2026 — Present
AI Engineer · Studio Jadu
Building multimodal text–image–video models for Studio Jadu's animation pipeline: tooling that extends what a small team can produce, with artists and storytellers directing the work at every step.
Sept 2025 — Feb 2026

Visiting researcher — text-to-image generative models · University of Zurich — Digital Society Initiative
Interpretability of text-to-image generative models, under the supervision of Dr. Eva Cetinić.
- Training data attribution — working out what a generated image owes to the specific examples it was trained on.
- Sparse autoencoder feature analysis on the vision encoders that condition these models, CLIP and DINO.
Oct 2023 — Present

PhD candidate, Computer Science & Mathematics · University of Bari Aldo Moro
Vision–language models and multimodal AI at CILab, supervised by Prof. Giovanna Castellano and Prof. Gennaro Vessio, funded by Italy's NRRP Mission 4. Two threads run through it: getting these models to hold up in a domain as demanding as art, and working out how they arrive at what they say. Thesis submitted September 2026; defence expected December 2026.
- ArtSeek — a multimodal RAG system pairing late-interaction retrieval (ColQwen2) with agentic reasoning. Its classification head re-uses the retriever's own vision encoder through multi-task learning, reaching state of the art on artwork attribute classification.
- I Dream My Painting (WACV 2025) — the first work on multi-mask inpainting: several regions filled at once, each from its own prompt, in a single generation pass rather than one region at a time.
- Label Anything (ECAI 2025) — few-shot semantic segmentation driven by visual prompts, state of the art on COCO-20ⁱ among multi-class methods.
- WikiFragments — a Wikipedia-scale multimodal dataset, over 5M image–text pairs across 42M examples.
- Large-scale training on the CINECA LEONARDO supercomputer through an EuroHPC ISCRA allocation.
May 2022 — Sept 2022
Software engineer — research collaboration · National Research Council of Italy (CNR) — Institute for Educational Technologies
Natural language processing for education, carrying on from the text complexity work I had started in my BSc thesis.
- Built NLP tooling for dyslexia detection in Python and spaCy.
- Contributed to COURAGE, an EU-funded social media literacy platform, on Django and Vue.js.
2021
Student, Samsung Innovation Campus · Samsung Italy & University of Bari Aldo Moro
Samsung Italy's training programme, run with the University of Bari. Selected among 25 students and placed in the top five on the final project, winning a scholarship.
- Built SmartGym, an Android app that counts exercise repetitions from the camera with an on-device computer vision model.
Education
Sep 2021 – Jul 2023

MSc in Computer Science — Artificial Intelligence110/110 cum laude
The AI curriculum, and the two years where the research direction actually formed: the thesis grew straight out of the coursework and became my first paper.
CourseworkMachine Learning · Deep Learning · Natural Language Processing · Computer Vision · Fundamentals of AI · Information Theory
Leveraging VLP and Transformer Models to Describe Artworks with ChatGPT Data
Nov 2018 – Jul 2021

BSc in Computer Science110/110 cum laude
A broad computer science foundation — mathematics, algorithms, software engineering, databases — with the first machine learning courses that pointed where I ended up.
CourseworkMathematics — discrete and analysis · Algorithms & Data Structures · Object-Oriented Programming · Software Engineering · Machine Learning · Databases
Creazione di un Tool per la Valutazione della Complessità del Testo
Awards
2025
Best Paper Honorable Mention — IEEE CIS Italy Chapter, IJCNN 2025
I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
Publications12
2026
- . Understanding How MLLMs Describe Artworks Using Token Activation Maps. PRESTIGE Workshop at the International Conference on Pattern Recognition (ICPR).
- . Art2Mus: Artwork-to-Music Generation via Visual Conditioning and Large-Scale Cross-Modal Alignment. arXiv preprint.
2025
- . ArtSeek: Deep Artwork Understanding via Multimodal In-Context Reasoning and Late Interaction Retrieval. arXiv preprint.
- . Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts. 28th European Conference on Artificial Intelligence (ECAI).
- . I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting. IEEE/CVF Winter Conference on Applications of Computer Vision (WACV).
- . Unveiling Visual Features in Artwork Classification: Towards Explainable Vision Transformers in the Arts. International Conference on Image Analysis and Processing (ICIAP).
- . Multimodal Artwork Topic Modeling via Fine-Tuned CLIP and Knowledge-Driven Prompts. IEEE 35th International Workshop on Machine Learning for Signal Processing (MLSP).
- . Generative AI Across Modalities: Insights from Our Research on Domain-Aware Content Generation. Ital-IA.
2024
- . Art2Mus: Bridging Visual Arts and Music through Cross-Modal Generation. AI4VA Workshop at the European Conference on Computer Vision (ECCV).
- . Converso: Improving LLM Chatbot Interfaces and Task Execution via Conversational Forms. LUHME Workshop at the European Conference on Artificial Intelligence (ECAI).
2023
- . Exploring the Synergy Between Vision-Language Pretraining and ChatGPT for Artwork Captioning: A Preliminary Study. FAPER Workshop at the International Conference on Image Analysis and Processing (ICIAP).
- . Exploring New Frontiers at the Intersection of AI and Art. Lecture Notes in Computer Science.
Service & training
Reviewing
- NeurIPS · CVPR · ECAI · IJCNN · ICAART · ICPR
- Journal of Intelligent Systems · Journal of Cultural Heritage · Information Systems Frontiers · Information Processing and Management · IEEE Access · Knowledge-Based Systems
Summer schools
- DeepLearn — Porto, Portugal
- ICVSS — Sicily, Italy
Details
Interests
Vision–language models · Multimodal deep learning · Interpretability · Video generation · Diffusion models · MLLMs · AI for art and cultural heritage
Contact