{ "model_type": "unknown", "falconsai.synthesized": true, "falconsai.tool": "FALCONS.AI Model Surgeon V7.26", "falconsai.attn_note": "attention kernel is chosen at load time (e.g. attn_implementation='flash_attention_2' on CUDA/ROCm hosts that have it); nothing in this file selects it" }