ThoxMicro-1bit-16M

License Params GGUF Format

Your AI. Your Data. Your Rules.

From-scratch BitNet b1.58 ternary Llama (16M) trained on TinyStories — deeper sibling of the 9M edge research model.

What this is

  • Trained from scratch — no upstream base model.
  • 86.6% of weights are ternary (BitNet b1.58).
  • Same pending-review TinyStories licensing caveat as the 9M.
  • Neither the 9M nor 16M supersedes the other.

Architecture (from config)

Field Value
Architecture BitNet b1.58 ternary Llama decoder
Layers 16
Hidden size 256
Attention heads 8
KV heads 8
FFN / intermediate 768
Vocab 8,192 (own byte-level BPE)
Max context 512
Tied embeddings yes

Intended use

On-device / edge text generation within the THOX stack. Not a safety-aligned public assistant unless deployed behind THOX guardrails.

Usage

llama.cpp

huggingface-cli download Thox-ai/ThoxMicro-1bit-16M --include '*.gguf' --local-dir ./ThoxMicro-1bit-16M
llama-cli -m ./ThoxMicro-1bit-16M/model-TQ2_0.gguf -p "Hello"

Links

  • Ollama: ollama.com/thox-ai/<slug>verify with the Ollama lane (task 80017303)
  • Docs: https://docs.thox.ai

THOX.ai LLC — Your AI. Your Data. Your Rules. · On-device and private by design.

Downloads last month
74
GGUF
Model size
15.7M params
Architecture
llama
Hardware compatibility
Log In to add your hardware

2-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Thox-ai/ThoxMicro-1bit-16M

Quantizations
1 model

Space using Thox-ai/ThoxMicro-1bit-16M 1

MiniMax H3 Video Generator 20 free credits · Text & image to video Try Free →