๐ Efficient implementations for emerging model architectures
Star History
Momentum
+145
STARS ยท LAST 30 DAYS
5
PER DAY
#159
MOST-STARRED Python
| Window | 7 days | 30 days | 90 days |
|---|---|---|---|
| Stars gained | +35 | +145 | +450 |
| Per day | 5 | 5 | 5 |
| Forks gained | +4 | +15 | +36 |
flash-linear-attention gained 145 stars in the last 30 days, about 5 a day, and now has 5.8K. It is about 3 years old and has averaged roughly 1.9K stars a year. It ranks #159 among Python repositories and #1,007 across all languages on GitHubRepo.
Trending Record
3
DAYS ON TRENDING
#1252
BEST RANK
Sep 26, 2026
FIRST APPEARANCE
Active
STATUS TODAY
flash-linear-attention has maintained a continuous presence across global trending indexes, peaking at #1252. Below is the 30-day activity profile:
๐ก Overview
flash-linear-attention is an open-source project written in Python: ๐ Efficient implementations for emerging model architectures.
Engineered for speed, consistency, and developer ease, it solves common hurdles in large-language-models, machine-learning-systems, natural-language-processing. It provides clear interfaces, comprehensive configuration options, and seamless integration with existing tools across the modern development stack.
โก Key Features
Optimized execution pipeline written in Python for predictable speed.
Zero-friction configuration with comprehensive sensible defaults out of the box.
Cross-platform runtime support across Linux, macOS, and Windows environments.
Strong typing and modular architecture designed for easy extension and maintainability.
Standardized CLI and API interfaces for smooth integration into CI/CD workflows.
Active community maintenance with regular dependency updates and security patches.
๐ฅ Installation
$ pip install flash-linear-attention
โ System Requirements
Platforms
- โข macOS
- โข Linux
- โข Windows
Runtime & Dependencies
Python >= 3.9, pip, virtualenv
Architecture
x86_64, ARM64 (Apple Silicon & Graviton)
๐ง How It Works
flash-linear-attention coordinates its core functionality through a modular Python pipeline. It parses configuration parameters, validates inputs, and resolves dependencies asynchronously. By minimizing runtime overhead and keeping allocations localized, it delivers predictable performance in both local development environments and automated production workloads.
๐ฏ Production Use Cases
Production System Integration
Embed flash-linear-attention into Python backend services to handle core application logic.
CI/CD Automated Pipelines
Run automated validation, builds, and integration suites during deployments.
Developer Tooling & Workflows
Accelerate developer onboarding with pre-configured project utilities.
Open Source Extension
Fork and customize internal modules under the repository's open MIT license.
๐ Getting Started
Install flash-linear-attention using your package manager: `pip install flash-linear-attention`
Initialize your project workspace or configuration file for flash-linear-attention.
Import flash-linear-attention into your codebase or invoke it directly from your terminal.
Execute your test suite or run `flash-linear-attention --help` to verify successful setup.
๐ Strengths
โ ๏ธ Considerations
โ Alternatives & Direct Competitors
๐ฅ Who Should Use This
Developers and engineering teams building with Python, seeking reliable, tested, and actively maintained tooling for production workloads.
๐ Nearby in the Rankings
fla-org/flash-linear-attention is currently ranked #1,007 by stars across every repository tracked on GitHubRepo. These are adjacent projects:
| Rank | Repository | Language | Stars | Action |
|---|---|---|---|---|
| #1,002 | openai-php/client | PHP | โ 5.8K | Compare โ |
| #1,003 | GeyserMC/Geyser | Java | โ 5.8K | Compare โ |
| #1,004 | agentscope-ai/agentscope-java | Java | โ 5.8K | Compare โ |
| #1,005 | scenee/FloatingPanel | Swift | โ 5.8K | Compare โ |
| #1,006 | paradigmxyz/reth | Rust | โ 5.8K | Compare โ |
| #1,007 | fla-org/flash-linear-attention This Project | Python | โ 5.8K | |
| #1,008 | ayn2op/discordo | Go | โ 5.8K | Compare โ |
| #1,008 | lemonade-sdk/lemonade | C++ | โ 5.8K | Compare โ |
| #1,010 | alphaXiv/OpenResearch | Rust | โ 5.8K | Compare โ |
| #1,011 | Eventual-Inc/Daft | Rust | โ 5.8K | Compare โ |
| #1,012 | Tianyu199509/DeskBox | C# | โ 5.8K | Compare โ |
Frequently Asked Questions
What does flash-linear-attention do? +
๐ Efficient implementations for emerging model architectures
What language is flash-linear-attention written in? +
The primary language is Python. Topics include: large-language-models, machine-learning-systems, natural-language-processing, sequence-modeling.
Is flash-linear-attention actively maintained? +
Yes, the last recorded push was on Sep 26, 2026 with 109 open issues being tracked.
How many stars does flash-linear-attention have? +
flash-linear-attention has 5,793 stars and 726 forks on GitHub.
How does flash-linear-attention rank among GitHub repositories? +
With 5,793 stars, fla-org/flash-linear-attention is ranked #1,007 globally across all repositories tracked on GitHubRepo and #159 among Python projects.
What license is flash-linear-attention distributed under? +
The repository reports a MIT license. Always verify the repository LICENSE file for legal terms.