Free
$0 forever

No credit card. No expiry.

  • Compaxis Hub for Windows 10 / 11
  • Qwen 3 1.7B SQ4 GGUF (the free model)
  • Runs on any 4 GB+ GPU
  • Use in llama.cpp / Ollama / LM Studio
  • Community support
  • Larger SQ4 models (Pro)
  • Priority email support (Pro)
Download Hub
Pro · Monthly
$9 / month

Cancel anytime. Keep working SQ4 files.

  • Everything in Free
  • All 7 SQ4 production models
  • Qwen 3 8B / 14B / 27B SQ4
  • Gemma 4 E4B + 26B MoE SQ4
  • Devstral Small 24B SQ4 (coding)
  • New SQ4 models as we add them
  • Priority email support
Start Monthly →
Lifetime
$149 once

One-time purchase. Pay once, keep using.

  • Everything in Pro Annual
  • All SQ4 models released in V1.x
  • No subscription renewal
  • Priority email support
  • 14-day money-back guarantee
  • Future major version may upgrade-price
Buy Lifetime →

14-day money-back guarantee on all paid plans

If Compaxis SQ4 doesn't measurably outperform Q4_K_M for the models you actually use, email within 14 days for a full refund — no questions asked. Full terms on the refund policy page.

Frequently asked questions

If your question isn't here, email .

What's the difference between SQ4 and standard Q4_K_M?

Compaxis SQ4 uses a more efficient compression algorithm than the standard Q4_K_M. The result is a 4-bit GGUF that's measurably closer to FP16 — 21% to 78% closer across our 7-model production line — at roughly the same file size, same speed, and same loader. Drop it into llama.cpp / Ollama / LM Studio just like any other GGUF.

Is the SQ4 file a real GGUF, or do I need a new runtime?

Real GGUF. Compaxis SQ4 files load in unmodified llama.cpp, Ollama, LM Studio, KoboldCpp, and anything else that already reads GGUF. No new runtime to install, no API to integrate. The Hub is a convenience download manager — you can also drop the .gguf into your existing tools by hand.

What hardware do I need?

Same as any 4-bit GGUF: roughly the model file size + a couple of gigabytes of context cache. Qwen 3 1.7B SQ4 runs comfortably on a 4 GB GPU. Qwen 3 8B SQ4 fits in 8 GB. Qwen 3 14B SQ4 wants 12 GB. The 27B-class models want 20–24 GB. Compaxis SQ4 files are about 15% larger than Q4_K_M on average — the bits we add buy the quality.

Can I use Compaxis SQ4 commercially?

Yes. Your Pro or Lifetime license covers commercial use of the SQ4 files we publish. Note that each underlying model has its own upstream licence (Apache 2.0 for Qwen and Gemma 4, MIT for the DeepSeek distill) — Compaxis re-quantizes them but doesn't change those upstream terms. We list each model's licence on its download page.

What happens to my files if I cancel?

Any Compaxis SQ4 GGUF you downloaded while your subscription was active is yours to keep and use. The Hub stops pulling new updates and stops downloading additional Pro models, but every file already on disk keeps working in llama.cpp.

How is Pro Lifetime different from "lifetime updates"?

Lifetime covers every SQ4 model we release during the V1.x series (V1.0, V1.1, V1.2 …). If we ship a V2.0 with a fundamentally different product — for example moving beyond GGUF, or new quantization formats — there may be a discounted upgrade price. The V1.x line itself is included indefinitely.

Do you offer team / volume pricing?

Not yet for V1. If you need 5+ seats or a site licence, email and we'll work something out.

Ready to try it?

Download the Hub free, run Qwen 3 1.7B SQ4 against the Q4_K_M build of the same model, and see the difference on your hardware.

Hub V1.0 ships on Windows 10 / 11 (x64). Linux support coming in a later release. This link downloads a short README; the full signed installer ships shortly and will appear at the same location.