JIT LoRA: Real-Time Conversational Knowledge Injection on Apple Silicon via MLX
Background LoRA training updates a running language model. On M4 Max, the paper reports 61/105 pooled recall, 60/60 general-knowledge preservation, and 69.6 seconds for 180 steps.