Program

This is our tentative program. More details coming soon.

Tutorial

  • 09h00-09h10: Welcome
  • 09h10-10h30: Tutorial on Mechanistic Interpretability – Part I
  • 10h30-11h00: Coffee Break
  • 11h00-12h40: Tutorial on Mechanistic Interpretability – Part II
  • 12h40-12h55: Discussion and Concluding Remarks.
  • 12h00-14h00:  Lunch Break

Workshop

  • 14h00-15h00: Keynote Talk
  • 15h00-15h15 (Paper presentation): Concepts Worth Having: Refining VLM-Guided Concept Bottleneck Models with Minimal Annotations
  • 15h15-15h30 (Paper presentation): COCOLogic-V2: Identifying Logical Inconsistencies via Truly Hard-Negatives
  • 15h30-15h45 (Paper presentation): Evaluating Meta-Feature Fidelity: Detecting Information Leakage in Interpretable Latent Spaces
  • 15h45-16h00 (Paper presentation): How smoothing the affinity matrix affects neighborhood preservation in t-SNE
  • 16h00-16h30: Coffee Break
  • 16h30-16h45 (Paper presentation): How smoothing the affinity matrix affects neighborhood preservation in t-SNE
  • 16h45-17h00 (Paper presentation): Hallucination Neurons and Where to Find Them: An Investigation into the existence of Hallucination Neurons
  • 17h00-17h15 (Paper presentation): When Models Disagree: Contrastive Explainability for Clinical Mortality Prediction
  • 17h15-18h00: Poster Session