Trending

Content tagged with "deep-learning"

deep-learning

Hacker News

Top stories from the Hacker News community• Updated 8 minutes ago

Reddit

Top posts from tech subreddits• Updated 2 minutes ago

Hugging Face Trending

Popular models from Hugging Face• Updated 20 minutes ago

GitHub Trending

Popular repositories from GitHub• Updated 34 minutes ago

RWKV-LM

RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.

Python
13,338
900