Elisabetta Rocchetti
ActiveResearch focus
I am focusing on interpretability, exploring different approaches for modelling neural networks, and understanding how models organise concepts in their latent space.
Open to collaboration
Yes
Looking for
volunteers and independent contributors, or collaborators from other research labs
Join the Damaqu community to see this team's full profile.
Papers on Damaqu (11)
- Processing and Representation of Linguistic Properties in Large Language Modelsother
- Augustine of Hippo at SemEval-2023 Task 4: An Explainable Knowledge Extraction Method to Identify Human Values in Arguments with SuperASKEother
- A Discussion on Open Issues Regarding Human Value Detection in Argumentsother
- Causal Mediation Analysis for Interpreting Large Language Modelsother
- The Cow of Rembrandt Analyzing Artistic Prompt Interpretation in Text-to-Image Modelsother
- Refusal Beyond a Single Direction: A Preliminary Comparison of Diff-in-Means and INLParxiv· 11 Jun 2026
- Unveiling Transformer Perception by Exploring Input Manifoldsother· 23 Apr 2026
- How LLMs Follow Instructions: Skillful Coordination, Not a Universal Mechanismarxiv· 7 Apr 2026
- Modeling Transformers as complex networks to analyze learning dynamicsarxiv· 18 Sept 2025
- How Instruction-Tuning Imparts Length Control: A Cross-Lingual Mechanistic Analysisarxiv· 2 Sept 2025
- What's Taboo for You? - An Empirical Evaluation of LLMs Behavior Toward Sensitive Contentarxiv· 31 Jul 2025