LL4nc33 / llama-tq Star 6 Code Issues Pull requests llama.cpp fork tuned for running modern models (Gemma-4, Qwen3.x) at full context on 12 GB Turing GPUs (RTX 2060/2070/2080, T4). TurboQuant KV cache (KTQ+VTQ, 2.78 bpw f16-quality), SWA-aware KV, MTP+n-gram speculation. turing fine-tuning kv-cache llama-cpp speculative-decoding low-vram turboquant llama-tq Updated Aug 17, 2026 C++
LL4nc33 / accountant Star 0 Code Issues Pull requests CRM/Selbstbuch für AT-EPUs mit lokalem KI-Assistenten. Paragraf 11 UStG, UVA, Mahnwesen, Projekte, Paperless-ngx. Angular + Express + SQLite. AGPL-3.0. accounting invoicing self-hosted austria xrechnung ai-assistant kleinunternehmer paperless-ngx llama-cpp local-llm ollama llama-tq Updated Aug 17, 2026 TypeScript