AI & ML interests

High-performance local inference, NVIDIA RTX 5080 optimisation, CUDA, quantisation, long-context inference, multimodal models, speculative decoding, and reproducible model builds.

Recent Activity

tmballin  updated a model 7 days ago
ninfer-5080/Qwen3.8-27B-RTX5080
tmballin  updated a Space 16 days ago
ninfer-5080/README
tmballin  published a Space 16 days ago
ninfer-5080/README
View all activity