This paper helps us understand how text-to-image diffusion transformers work by analyzing the role of "template tokens" in generating images from text prompts. Practitioners might care because it shows how to improve the efficiency of these models without sacrificing their performance.
Firehose
Filtered to Papers, tagged “causal interpretability” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News