# atomicmilkshake/godzilla-llama.cpp
*public · C++ · ★ 6 · 0 trails · 0 tours*

🦖 Godzilla: Next-gen LLM inference engine featuring 512K context scaling, Six Flags over Texas stack, TriAttention, TurboQuant, DFlash draft sidecars, and --kv-vram-only allocation.

**Topics:** cuda, kv-cache, llama-cpp, llm-inference, speculative-decoding, triattention, turboquant, vram-optimization

## Trails
_No trails published for this repo yet._

## Tours
_No tours published for this repo yet._

---
Repository: https://github.com/atomicmilkshake/godzilla-llama.cpp
Interactive view: https://app.principal-ade.com/atomicmilkshake/godzilla-llama.cpp
JSON: https://app.principal-ade.com/api/repos/atomicmilkshake/godzilla-llama.cpp
To open a trail or tour locally in the interactive viewer, see https://app.principal-ade.com for the Principal CLI quickstart.
