Oliver Eberle

Research on the Foundations of AI Interpretability

oeberle_website.png

Senior Researcher
Technische Universität Berlin
The Berlin Institute for the Foundations of Learning (BIFOLD)

Explainable AI and Deep Learning

  • Natural Language Processing
  • Computer Vision
  • Digital Humanities
  • Human Cognition

news

Jul 08, 2026 Our paper on discovering algorithmic primitives was presented at ICML 2026. AlgoTrace introduces a framework for tracing and steering algorithmic primitives in LLM latent space, showing primitives can be extracted, injected, and composed to steer multi-step reasoning. Sources: ICML poster page · OpenReview
Jul 08, 2026 Our new preprint, coauthored by Maximilian S. Ernst, Lorenz Linhardt, and Aaron Peikert, introduces Distributed Sparse Interventions (DSI), a method for identifying neuron sets whose targeted interventions elicit task-relevant behaviour in instruction‑tuned LMs. Sources: arXiv abstract · PDF
Oct 24, 2025 Oliver Eberle has been awarded the 2025 Heinz‑Billing Prize for the Advancement of Scientific Computing. The award was announced 24 Oct 2025 (ceremony 23–24 Oct 2025) in recognition of contributions to digital humanities and scientific computing. Sources: MPIWG announcement · BIFOLD announcement

selected publications

  1. MedIA
    Beyond attention heatmaps: How to get better explanations for multiple instance learning models in histopathology
    Mina Jamshidi Idaji, Julius Hense, Tom Neuhäuser, and 12 more authors
    Medical Image Analysis, Sep 2026
  2. Preprint
    Distributed Sparse Interventions in Language Models
    Maximilian S. Ernst, Lorenz Linhardt, Aaron Peikert, and 1 more author
    Jul 2026
    arXiv:2607.07128 [cs.LG]
  3. ICML
    AlgoTrace: Algorithmic Primitives and Compositional Geometry of Reasoning in Language Models
    Samuel Lippl, Thomas Austin McGee, Kimberly Lopez, and 5 more authors
    In Forty-third International Conference on Machine Learning, 2026
  4. ICML
    Position: We Need An Algorithmic Understanding of Generative AI
    Oliver Eberle, Thomas Austin Mcgee, Hamza Giaffar, and 2 more authors
    In Proceedings of the 42nd International Conference on Machine Learning, Oct 2025
  5. Preprint
    RelP: Faithful and Efficient Circuit Discovery in Language Models via Relevance Patching
    Farnoush Rezaei Jafari, Oliver Eberle, Ashkan Khakzar, and 1 more author
    Oct 2025
    arXiv:2508.21258 [cs.LG]
  6. NeurIPS
    Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
    Laura Kopf, Nils Feldhus, Kirill Bykov, and 4 more authors
    In Advances in Neural Information Processing Systems, 2025
  7. Sci Adv
    Historical insights at scale: A corpus-wide machine learning analysis of early modern astronomic tables
    Oliver Eberle, Jochen Büttner, Hassan el-Hajj, and 3 more authors
    Science Advances, Oct 2024
  8. NeurIPS
    xMIL: Insightful Explanations for Multiple Instance Learning in Histopathology
    Julius Hense, Mina Jamshidi Idaji, Oliver Eberle, and 7 more authors
    In Advances in Neural Information Processing Systems, 2024
  9. NeurIPS
    MambaLRP: Explaining Selective State Space Sequence Models
    Farnoush Rezaei Jafari, Grégoire Montavon, Klaus-Robert Müller, and 1 more author
    In Advances in Neural Information Processing Systems 37, 2024
  10. ACL
    Explaining Text Similarity in Transformer Models
    Alexandros Vasileiou and Oliver Eberle
    In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), 2024
  11. IJDH
    Explainability and transparency in the realm of digital humanities: toward a historian XAI
    Hassan El-Hajj, Oliver Eberle, Anika Merklein, and 7 more authors
    International Journal of Digital Humanities, Nov 2023
  12. EMNLP
    Rather a Nurse than a Physician - Contrastive Explanations under Investigation
    Oliver Eberle, Ilias Chalkidis, Laura Cabello, and 1 more author
    In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, 2023
  13. ICML
    XAI for Transformers: Better Explanations through Conservative Propagation
    Ameen Ali, Thomas Schnake, Oliver Eberle, and 3 more authors
    In Proceedings of the 39th International Conference on Machine Learning, Jun 2022
  14. Springer
    An Ever-Expanding Humanities Knowledge Graph: The Sphaera Corpus at the Intersection of Humanities, Data Management, and Machine Learning
    Hassan El-Hajj, Maryam Zamani, Jochen Büttner, and 8 more authors
    Datenbank-Spektrum, Jul 2022
  15. ACL
    Do Transformer Models Show Similar Attention Patterns to Task-Specific Human Gaze?
    Oliver Eberle, Stephanie Brandl, Jonas Pilot, and 1 more author
    In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 2022
  16. PAMI
    Higher-Order Explanations of Graph Neural Networks via Relevant Walks
    Thomas Schnake, Oliver Eberle, Jonas Lederer, and 4 more authors
    IEEE Transactions on Pattern Analysis and Machine Intelligence, Nov 2022
  17. PAMI
    Building and Interpreting Deep Similarity Models
    Oliver Eberle, Jochen Buttner, Florian Krautli, and 3 more authors
    IEEE Transactions on Pattern Analysis and Machine Intelligence, Mar 2022