feifeibear / long-context-attention

USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference
357Updated this week

Related projects

Alternatives and complementary repositories for long-context-attention