↓Skip to main content

Mechanistic Interpretability

Why Large Language Models Reason and Behave Like Humans
·2142 words·11 mins
Large Language Models Artificial Intelligence Transformer Mechanistic Interpretability Sparse Autoencoders Machine Learning Reinforcement-Learning Reasoning Models Embodied AI