JEV / code /scripts /vllm_humaneval.py

Commit History

vLLM inference: adapter_vllm/ (decision head as lm_head LoRA), model card section "Inference with vLLM", OpenAI client, measured speed
b16c3a6
verified

cloudyu commited on