Gata 0.01 12B Web Game Dev (GGUF)

This repository contains the quantized GGUF versions of Gata 0.01 12B, a specialized model fine-tuned for browser game development and Three.js/TypeScript engineering.

The GGUF files were compiled using the --no-mtp flag for maximum compatibility with vanilla llama.cpp builds.

Available Quantizations

  • Q4_K_M (q4 km): Excellent balance between size (6.46 GB) and quality. Highly recommended for 24GB or smaller GPUs.
  • Q8_0 (q8 xl): High-precision 8-bit quantization (11.01 GB) with minimal perplexity degradation.

Quick Start (llama.cpp)

Run the model in server mode:

llama-server --model Gata0.01-12b-web-game-dev-trunk-Q4_K_M.gguf -c 4096 --port 8080

License & Copyright Notices

This model is based on the Qwen model family developed by Alibaba Cloud. It inherits the Qwen License Agreement. Please refer to the Qwen License for usage constraints and commercial application rules.

Downloads last month
116
GGUF
Model size
11B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for quimmedes/Gata0.01-12b-web-game-dev-GGUF

Finetuned
Qwen/Qwen3.5-9B
Quantized
(567)
this model