08-15-2026, 07:24 PM
Understanding Large Language Models
BY Jenny Kunz
Summary
This pdf file contains Jenny Kunz's 2024 doctoral dissertation from Linköping University, titled "Understanding Large Language Models: Towards Rigorous and Targeted Interpretability Using Probing Classifiers and Self-Rationalisation".
The research explores methods for evaluating and understanding the inner workings and outputs of large language models (LLMs), focusing on two primary areas: model interpretability via probing classifiers to analyze intermediate layer representations, and model explainability by evaluating the quality, properties, and alignment of natural language explanations generated by self-rationalizing models.
BOOK [PDF]
BY Jenny Kunz
Summary
This pdf file contains Jenny Kunz's 2024 doctoral dissertation from Linköping University, titled "Understanding Large Language Models: Towards Rigorous and Targeted Interpretability Using Probing Classifiers and Self-Rationalisation".
The research explores methods for evaluating and understanding the inner workings and outputs of large language models (LLMs), focusing on two primary areas: model interpretability via probing classifiers to analyze intermediate layer representations, and model explainability by evaluating the quality, properties, and alignment of natural language explanations generated by self-rationalizing models.
BOOK [PDF]
┌────────────────────────────────┐
│ KONSTANTINOS MICHAILIDIS │
└────────────────────────────────┘
│ KONSTANTINOS MICHAILIDIS │
└────────────────────────────────┘

