publications

publications by categories in reversed chronological order. generated by jekyll-scholar.

2026

  1. A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification
    Stephanie Brandl and Oliver Eberle
    In Proceedings of the 6th Workshop on Trustworthy NLP (TrustNLP 2026), Jul 2026
  2. MedIA
    Beyond attention heatmaps: How to get better explanations for multiple instance learning models in histopathology
    Mina Jamshidi Idaji, Julius Hense, Tom Neuhäuser, and 12 more authors
    Medical Image Analysis, Sep 2026
  3. Preprint
    Distributed Sparse Interventions in Language Models
    Maximilian S. Ernst, Lorenz Linhardt, Aaron Peikert, and 1 more author
    Jul 2026
    arXiv:2607.07128 [cs.LG]
  4. ICML
    AlgoTrace: Algorithmic Primitives and Compositional Geometry of Reasoning in Language Models
    Samuel Lippl, Thomas Austin McGee, Kimberly Lopez, and 5 more authors
    In Forty-third International Conference on Machine Learning, 2026

2025

  1. ICML
    Position: We Need An Algorithmic Understanding of Generative AI
    Oliver Eberle, Thomas Austin Mcgee, Hamza Giaffar, and 2 more authors
    In Proceedings of the 42nd International Conference on Machine Learning, Oct 2025
  2. Preprint
    RelP: Faithful and Efficient Circuit Discovery in Language Models via Relevance Patching
    Farnoush Rezaei Jafari, Oliver Eberle, Ashkan Khakzar, and 1 more author
    Oct 2025
    arXiv:2508.21258 [cs.LG]
  3. NeurIPS
    Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
    Laura Kopf, Nils Feldhus, Kirill Bykov, and 4 more authors
    In Advances in Neural Information Processing Systems, 2025
  4. ACL
    Trick or Neat: Adversarial Ambiguity and Language Model Evaluation
    Antonia Karamolegkou, Oliver Eberle, Phillip Rust, and 2 more authors
    In Findings of the Association for Computational Linguistics: ACL 2025, 2025

2024

  1. Sci Adv
    Historical insights at scale: A corpus-wide machine learning analysis of early modern astronomic tables
    Oliver Eberle, Jochen Büttner, Hassan el-Hajj, and 3 more authors
    Science Advances, Oct 2024
  2. NeurIPS
    xMIL: Insightful Explanations for Multiple Instance Learning in Histopathology
    Julius Hense, Mina Jamshidi Idaji, Oliver Eberle, and 7 more authors
    In Advances in Neural Information Processing Systems, 2024
  3. NeurIPS
    MambaLRP: Explaining Selective State Space Sequence Models
    Farnoush Rezaei Jafari, Grégoire Montavon, Klaus-Robert Müller, and 1 more author
    In Advances in Neural Information Processing Systems 37, 2024
  4. ACL
    Explaining Text Similarity in Transformer Models
    Alexandros Vasileiou and Oliver Eberle
    In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), 2024
  5. LREC-COLING
    Evaluating Webcam-based Gaze Data as an Alternative for Human Rationale Annotations
    Stephanie Brandl, Oliver Eberle, Tiago Ribeiro, and 2 more authors
    In Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), May 2024

2023

  1. IJDH
    Explainability and transparency in the realm of digital humanities: toward a historian XAI
    Hassan El-Hajj, Oliver Eberle, Anika Merklein, and 7 more authors
    International Journal of Digital Humanities, Nov 2023
  2. EMNLP
    Rather a Nurse than a Physician - Contrastive Explanations under Investigation
    Oliver Eberle, Ilias Chalkidis, Laura Cabello, and 1 more author
    In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, 2023

2022

  1. ICML
    XAI for Transformers: Better Explanations through Conservative Propagation
    Ameen Ali, Thomas Schnake, Oliver Eberle, and 3 more authors
    In Proceedings of the 39th International Conference on Machine Learning, Jun 2022
  2. Springer
    An Ever-Expanding Humanities Knowledge Graph: The Sphaera Corpus at the Intersection of Humanities, Data Management, and Machine Learning
    Hassan El-Hajj, Maryam Zamani, Jochen Büttner, and 8 more authors
    Datenbank-Spektrum, Jul 2022
  3. ACL
    Do Transformer Models Show Similar Attention Patterns to Task-Specific Human Gaze?
    Oliver Eberle, Stephanie Brandl, Jonas Pilot, and 1 more author
    In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 2022
  4. PAMI
    Higher-Order Explanations of Graph Neural Networks via Relevant Walks
    Thomas Schnake, Oliver Eberle, Jonas Lederer, and 4 more authors
    IEEE Transactions on Pattern Analysis and Machine Intelligence, Nov 2022
  5. PAMI
    Building and Interpreting Deep Similarity Models
    Oliver Eberle, Jochen Buttner, Florian Krautli, and 3 more authors
    IEEE Transactions on Pattern Analysis and Machine Intelligence, Mar 2022