Pre-warm LLM into CPU RAM at startup (avoids first-call GPU timeout) ce5cab3 verified sush0401 commited on Jun 15