AI & ML interests

MoE architectures, Chimera models, Assembly of Experts

Recent Activity

GevatterGaul  updated a model 2 days ago
tngtech/Qwen3.6-27B-NVFP4-GGUF
GevatterGaul  published a model 7 days ago
tngtech/Qwen3.6-27B-NVFP4-GGUF
SR-TNG  updated a model 7 days ago
tngtech/Qwen3.6-27B-NVFP4-GGUF
View all activity

Articles

BM-TNG 
published an article about 1 year ago
view article
Article

How Long Prompts Block Other Requests - Optimizing LLM Performance

tngtech
•
• 14
SR-TNG 
published an article over 1 year ago
view article
Article

Finetuning olmOCR to be a faithful OCR-Engine

tngtech
•
• 19
BM-TNG 
published an article over 1 year ago
view article
Article

Prefill and Decode for Concurrent Requests - Optimizing LLM Performance

tngtech
•
• 86
BM-TNG 
published an article over 1 year ago
view article
Article

Efficient Request Queueing – Optimizing LLM Performance

tngtech
•
• 27