Select a category to explore research frontiers
Loading categories...
Research focusing on reverse-engineering the internal computational mechanisms and decision-making processes within transformer-based language models through circuit analysis and activation patching.