Ouch

#1
by Propercode - opened

I'm sorry to report that I feel I wasted time and resources testing this on DGX Spark, llama.cpp and Hermes Agent: full version did not fit with meaningful context size, 8bit quantization behaves like drugged Qwen 27B or worse. Why release something like this?? Any bozo can lobotomize random Chinese open-source model distilled from OpenAI and Clumsy Claude with LoRa - why waste other's time with this, you're not going to get funded releasing crap like that.

Sign up or log in to comment

MiniMax H3 Video Generator 20 free credits · Text & image to video Try Free →