Kwame Mensah
kwamemen
ยท
AI & ML interests
Efficient LLM inference, KV cache optimization, quantization, speculative decoding, model pruning
Recent Activity
upvoted a paper about 11 hours ago
Gricea: An Open Science Platform for Conversational AI Research liked a dataset 1 day ago
open-llm-leaderboard-old/details_xformAI__opt-125m-gqa-ub-6-best-for-KV-cache liked a dataset 1 day ago
ergt2025/kvcache_offloadOrganizations
None yet