Artificial Scientific Intelligence Lab - Science of AI Team
ActiveResearch focus
We consider AI systems as scientific objects and try to develop theories and algorithms to enhance our understanding of AI systems, especially the representation of multimodal models.
Open to collaboration
Yes
Looking for
We are looking for collaborators in general mechanistic interpretability.
Contact
Join the Damaqu community to see this team's full profile.
Papers on Damaqu (6)
- When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Modelsarxiv· 7 May 2026
- Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Controlarxiv· 6 Jan 2026
- A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minimaarxiv· 5 Dec 2025
- How does My Model Fail? Automatic Identification and Interpretation of Physical Plausibility Failure Modes with Matryoshka Transcodersarxiv· 13 Nov 2025
- Rep2Text: Decoding Full Text from a Single LLM Token Representationarxiv· 9 Nov 2025
- Human-like Content Analysis for Generative AI with Language-Grounded Sparse Encodersarxiv· 20 Aug 2025