This paper improves the performance and efficiency of diffusion transformers, a type of AI model used for video generation, by reducing the computational cost of attention mechanisms. Practitioners caring about accelerating AI models on hardware can benefit from this research.
Firehose
Filtered to Papers, tagged “tensor cores” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives