Skip to content

attention-mechanism.

Listed 12Updated Sep 7, 2026

  1. Rank 1. vit-pytorchImplementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in PytorchGitHub stars: 25.5k
  2. Rank 2. x-transformersA concise but complete full-attention transformer with a set of promising experimental features from various papersGitHub stars: 5.9k
  3. Rank 3. soundstorm-pytorchImplementation of SoundStorm, Efficient Parallel Audio Generation from Google Deepmind, in PytorchGitHub stars: 1.5k
  4. Rank 4. perceiver-pytorchImplementation of Perceiver, General Perception with Iterative Attention, in PytorchGitHub stars: 1.2k
  5. Rank 5. OpenSTLOpenSTL: A Comprehensive Benchmark of Spatio-Temporal Predictive LearningGitHub stars: 1.1k
  6. Rank 6. tab-transformer-pytorchImplementation of TabTransformer, attention network for tabular data, in PytorchGitHub stars: 1.1k
  7. Rank 7. enformer-pytorchImplementation of Enformer, Deepmind's attention network for predicting gene expression, in PytorchGitHub stars: 573
  8. Rank 8. slot-attentionImplementation of Slot Attention from GoogleAIGitHub stars: 498
  9. Rank 9. MultiModalMambaA novel implementation of fusing ViT with Mamba into a fast, agile, and high performance Multi-Modal Model. Powered by Zeta, the simplest AI framework ever.GitHub stars: 474
  10. Rank 10. deformable-attentionImplementation of Deformable Attention in Pytorch from the paper "Vision Transformer with Deformable Attention"GitHub stars: 368
  11. Rank 11. se3-transformer-pytorchImplementation of SE3-Transformers for Equivariant Self-Attention, in Pytorch. This specific repository is geared towards integration with eventual Alphafold2 replication.GitHub stars: 335
  12. Rank 12. bidirectional-cross-attentionA simple cross attention that updates both the source and target in one stepGitHub stars: 199