FeynmanWiki
ExploreLibraryRoadmapsCreateMy blogsPricing

Library

Search the public knowledge base.

Filters

Reset all

Categories

Machine Learning22Reinforcement Learning11Transformers5LLMs4Attention3Diffusion2AI Interpretability1BPE1causal interventions1context paradigm1Cordis1Databases1Distributed Machine Learning Systems1Distributed Training1DTensor1Dynamic software composition1energy-based models1FSDP1GPU-Parallel Robot Learning1importance sampling1

Sort by

LatestMost readRecently updated

Library

Search the public knowledge base.

K
Machine LearningReinforcement LearningTransformersLLMsAttentionDiffusionAI InterpretabilityBPEcausal interventionscontext paradigm

1 results

Sort by
LatestMost readRecently updated

Mechanistic Interpretability: Reverse-Engineering the Algorithms Inside Neural Networks

Take a simple completion: “Alice gave Bob the book because wanted it.” Suppose the model assigns a high target logit $y$ to the corr...

AI Interpretabilitymechanistic interpretabilitytransformer circuits
Sep 8, 2026 19 min read 106