RWKV-LM

RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.

14.7k
Stars
+576
Gained
4.1%
Growth
Python
Language

🎭 Best For

🏷️ Topics & Ecosystem

attention-mechanism chatgpt deep-learning gpt gpt-2 gpt-3 language-model linear-attention lstm pytorch rnn rwkv transformer transformers

📊 Activity

Latest commit: 2026-09-02. Over the past 294 days, this repository gained 576 stars (+4.1% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.