Sparse MoE models compressed using REAP (Router-weighted Expert Activation Pruning) method
-
cerebras/Qwen3-Coder-REAP-363B-A35B-FP8
Text Generation • 363B • Updated • 66 • 15 -
cerebras/Qwen3-Coder-REAP-246B-A35B-FP8
Text Generation • 246B • Updated • 118 • 20 -
cerebras/Qwen3-Coder-REAP-363B-A35B
Text Generation • 363B • Updated • 40 • 4 -
cerebras/Qwen3-Coder-REAP-246B-A35B
Text Generation • 246B • Updated • 32 • 7