
USAF: Fine-tune MoE models on consumer-grade GPUs
USAF is a new sparse fine-tuning method for Mixture-of-Experts (MoE) models that enables fine-tuning on hardware typically limited to inference. By training sparse expert weights and the router instead of traditional adapters, it significantly lowers the barrier for local model customization.






