Zürich, Switzerland

Curriculum Vitae

The short version. The full PDF is linked below, and this page is always current because it is built from the same content as the rest of the site.

Full CV (PDF)Email me

Experience

Aug 2026 — Present

AI Engineer · Studio Jadu

Zürich, Switzerland — hybrid

Building multimodal text–image–video models for Studio Jadu's animation pipeline: tooling that extends what a small team can produce, with artists and storytellers directing the work at every step.

Sept 2025 — Feb 2026

Visiting researcher — text-to-image generative models · University of Zurich — Digital Society Initiative

Zürich, Switzerland

Interpretability of text-to-image generative models, under the supervision of Dr. Eva Cetinić.

  • Training data attribution — working out what a generated image owes to the specific examples it was trained on.
  • Sparse autoencoder feature analysis on the vision encoders that condition these models, CLIP and DINO.

Oct 2023 — Present

PhD candidate, Computer Science & Mathematics · University of Bari Aldo Moro

Bari, Italy

Vision–language models and multimodal AI at CILab, supervised by Prof. Giovanna Castellano and Prof. Gennaro Vessio, funded by Italy's NRRP Mission 4. Two threads run through it: getting these models to hold up in a domain as demanding as art, and working out how they arrive at what they say. Thesis submitted September 2026; defence expected December 2026.

  • ArtSeek — a multimodal RAG system pairing late-interaction retrieval (ColQwen2) with agentic reasoning. Its classification head re-uses the retriever's own vision encoder through multi-task learning, reaching state of the art on artwork attribute classification.
  • I Dream My Painting (WACV 2025) — the first work on multi-mask inpainting: several regions filled at once, each from its own prompt, in a single generation pass rather than one region at a time.
  • Label Anything (ECAI 2025) — few-shot semantic segmentation driven by visual prompts, state of the art on COCO-20ⁱ among multi-class methods.
  • WikiFragments — a Wikipedia-scale multimodal dataset, over 5M image–text pairs across 42M examples.
  • Large-scale training on the CINECA LEONARDO supercomputer through an EuroHPC ISCRA allocation.

May 2022 — Sept 2022

Software engineer — research collaboration · National Research Council of Italy (CNR) — Institute for Educational Technologies

Palermo, Italy — remote

Natural language processing for education, carrying on from the text complexity work I had started in my BSc thesis.

  • Built NLP tooling for dyslexia detection in Python and spaCy.
  • Contributed to COURAGE, an EU-funded social media literacy platform, on Django and Vue.js.

2021

Student, Samsung Innovation Campus · Samsung Italy & University of Bari Aldo Moro

Bari, Italy

Samsung Italy's training programme, run with the University of Bari. Selected among 25 students and placed in the top five on the final project, winning a scholarship.

  • Built SmartGym, an Android app that counts exercise repetitions from the camera with an on-device computer vision model.

Education

Sep 2021 – Jul 2023

MSc in Computer Science — Artificial Intelligence110/110 cum laude

University of Bari Aldo Moro · Bari, Italy

The AI curriculum, and the two years where the research direction actually formed: the thesis grew straight out of the coursework and became my first paper.

CourseworkMachine Learning · Deep Learning · Natural Language Processing · Computer Vision · Fundamentals of AI · Information Theory

Leveraging VLP and Transformer Models to Describe Artworks with ChatGPT Data

Nov 2018 – Jul 2021

BSc in Computer Science110/110 cum laude

University of Bari Aldo Moro · Bari, Italy

A broad computer science foundation — mathematics, algorithms, software engineering, databases — with the first machine learning courses that pointed where I ended up.

CourseworkMathematics — discrete and analysis · Algorithms & Data Structures · Object-Oriented Programming · Software Engineering · Machine Learning · Databases

Creazione di un Tool per la Valutazione della Complessità del Testo

Awards

2025

Best Paper Honorable Mention — IEEE CIS Italy Chapter, IJCNN 2025

I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting

Publications12

2026

  1. Nicola Fanelli, Pasquale De Marinis, Raffaele Scaringi, Eva Cetinić, Gennaro Vessio, Giovanna Castellano. Understanding How MLLMs Describe Artworks Using Token Activation Maps. PRESTIGE Workshop at the International Conference on Pattern Recognition (ICPR).
  2. Ivan Rinaldi, Matteo Mendula, Nicola Fanelli, Florence Levé, Matteo Testi, Giovanna Castellano, Gennaro Vessio. Art2Mus: Artwork-to-Music Generation via Visual Conditioning and Large-Scale Cross-Modal Alignment. arXiv preprint.

2025

  1. Nicola Fanelli, Gennaro Vessio, Giovanna Castellano. ArtSeek: Deep Artwork Understanding via Multimodal In-Context Reasoning and Late Interaction Retrieval. arXiv preprint.
  2. Pasquale De Marinis, Nicola Fanelli, Raffaele Scaringi, Emanuele Colonna, Giuseppe Fiameni, Gennaro Vessio, Giovanna Castellano. Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts. 28th European Conference on Artificial Intelligence (ECAI).
  3. Nicola Fanelli, Gennaro Vessio, Giovanna Castellano. I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting. IEEE/CVF Winter Conference on Applications of Computer Vision (WACV).
  4. Raffaele Scaringi, Nicola Fanelli, Gennaro Vessio, Giovanna Castellano. Unveiling Visual Features in Artwork Classification: Towards Explainable Vision Transformers in the Arts. International Conference on Image Analysis and Processing (ICIAP).
  5. Raffaele Scaringi, Giacomo Stea, Nicola Fanelli, Gennaro Vessio, Giovanna Castellano. Multimodal Artwork Topic Modeling via Fine-Tuned CLIP and Knowledge-Driven Prompts. IEEE 35th International Workshop on Machine Learning for Signal Processing (MLSP).
  6. Giovanna Castellano, Emanuele Colonna, Nicola Fanelli, Luca Laraspata, Ivan Rinaldi, Andrea Gabriele Valerio, Gennaro Vessio. Generative AI Across Modalities: Insights from Our Research on Domain-Aware Content Generation. Ital-IA.

2024

  1. Ivan Rinaldi, Nicola Fanelli, Giovanna Castellano, Gennaro Vessio. Art2Mus: Bridging Visual Arts and Music through Cross-Modal Generation. AI4VA Workshop at the European Conference on Computer Vision (ECCV).
  2. Gianfranco Demarco, Nicola Fanelli, Gennaro Vessio, Giovanna Castellano. Converso: Improving LLM Chatbot Interfaces and Task Execution via Conversational Forms. LUHME Workshop at the European Conference on Artificial Intelligence (ECAI).

2023

  1. Giovanna Castellano, Nicola Fanelli, Raffaele Scaringi, Gennaro Vessio. Exploring the Synergy Between Vision-Language Pretraining and ChatGPT for Artwork Captioning: A Preliminary Study. FAPER Workshop at the International Conference on Image Analysis and Processing (ICIAP).
  2. Giovanna Castellano, Nicola Fanelli, Raffaele Scaringi, Gennaro Vessio. Exploring New Frontiers at the Intersection of AI and Art. Lecture Notes in Computer Science.

Service & training

Reviewing

  • NeurIPS · CVPR · ECAI · IJCNN · ICAART · ICPR
  • Journal of Intelligent Systems · Journal of Cultural Heritage · Information Systems Frontiers · Information Processing and Management · IEEE Access · Knowledge-Based Systems

Summer schools

  • DeepLearn — Porto, Portugal
  • ICVSS — Sicily, Italy

Details

Interests

Vision–language models · Multimodal deep learning · Interpretability · Video generation · Diffusion models · MLLMs · AI for art and cultural heritage