BlinkDL

BlinkDL

AI & ML interests

RWKV is all you need

Recent Activity

Organizations

RWKV's profile picture FreedomAI's profile picture rwkv-x-dev's profile picture Social Post Explorers's profile picture

BlinkDL's activity

posted an update 2 months ago
view post
Post
3872
RWKV-7 "Goose" 0.4B trained w/ ctx4k automatically extrapolates to ctx32k+, and perfectly solves NIAH ctx16k ๐Ÿคฏ 100% RNN and attention-free. Only trained on the Pile. No finetuning. Replicable training runs. tested by our community: https://github.com/Jellyfish042/LongMamba
posted an update 3 months ago
view post
Post
6336
RWKV-6-world-v3 (+3.1T tokens) is our best multilingual 7B model as of now: BlinkDL/rwkv-6-world

It's 100% RNN and attention-free. MMLU 54.2% (previous world-v2.1 = 47.9%. note: without eval-boosting tricks such as annealing).

RWKV-7-world-v4 soon :)
  • 1 reply
ยท
posted an update 5 months ago
view post
Post
5561
RWKV-7 "Goose" preview rc2 => Peak RNN architecture?๐Ÿ˜ƒWill try to squeeze more performance for the final release. Preview code & model: https://github.com/BlinkDL/RWKV-LM/tree/main/RWKV-v7
  • 2 replies
ยท