#LLM

Research

Defining and extracting generalizable interaction primitives from DNNs
Lu Chen, 
Siyu Lou, 
Benhao Huang, 
Quanshi Zhang
Sep 13th 2024
ICLR-2024
#LLM
#AI Interpretability

Given different DNNs trained for the same task, developed a new method to extract interactions that are shared by these DNNs. Experiments show that the extracted interactions can better reflect common ...

paper
GitHub
Last Updated on Aug 24th 2025 Powered by greatest-gatsby-academic-template.