kaptaan45/QaptaanLM-0.75B-Instruct
Text Generation • 0.8B • Updated • 489
Official suite of QaptaanLM-0.75B hybrid linear-attention models, quantization variants (GGUF, BitsAndBytes, ONNX WebGPU), and training datasets.
Note Supervised Fine-Tuned (SFT) Master Model (752M params)
Note Continued Pre-Trained (CPT) Base Foundation Model (752M params)
Note All GGUF quantizations for llama.cpp & Ollama (SFT)
Note All GGUF quantizations for llama.cpp & Ollama (Base)
Note Unified 4-Bit NF4 & 8-Bit Int8 BitsAndBytes suite (SFT)
Note Unified 4-Bit NF4 & 8-Bit Int8 BitsAndBytes suite (Base)
Note ONNX & WebGPU client-side / in-browser export (SFT)
Note ONNX & WebGPU client-side / in-browser export (Base)
Note 1 Billion Token CPT Programming & Technical Dataset
Note Multi-turn Instruction Tuning Dataset