DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation Paper • 2607.05147 • Published Jul 6 • 48
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated 21 days ago • 106
Navigating Text-To-Image Customization:From LyCORIS Fine-Tuning to Model Evaluation Paper • 2309.14859 • Published Sep 26, 2023 • 6
view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • Jul 27 • 483
view article Article Beyond LoRA: Can you beat the most popular fine-tuning technique? +2 BenjaminB, sayakpaul, hubnemo, kashif • Jun 18 • 100
FedPara: Low-Rank Hadamard Product for Communication-Efficient Federated Learning Paper • 2108.06098 • Published Aug 13, 2021 • 4
Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential Paper • 2507.11851 • Published Jul 16, 2025 • 1
Bone: Block Affine Transformation as Parameter Efficient Fine-tuning Methods for Large Language Models Paper • 2409.15371 • Published Sep 19, 2024 • 3
view article Article Introducing Modular Diffusers - Composable Building Blocks for Diffusion Pipelines +2 YiYiXu, OzzyGT, dn6, sayakpaul • Mar 5 • 55
view article Article GGML and llama.cpp join HF to ensure the long-term progress of Local AI +4 ggerganov, ngxson, allozaur, lysandre, victor, julien-c • Feb 20 • 510
PVeRA: Probabilistic Vector-Based Random Matrix Adaptation Paper • 2512.07703 • Published Dec 8, 2025 • 1
Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning Paper • 2205.05638 • Published May 11, 2022 • 6
view article Article Transformers v5: Simple model definitions powering the AI ecosystem +2 lysandre, ArthurZ, cyrilvallez, reach-vb • Dec 1, 2025 • 313
llama.vim Collection Recommended models for the llama.vim and llama.vscode plugins • 10 items • Updated Jul 30 • 82