Kshitij Thakkar PRO
AI & ML interests
Organizations
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 31 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 37 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 24 • 1 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
-
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs4-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx2048
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs1-ctx2048
1B • Updated
- Build error2
Loggenix Moe 0.3B A0.1B Demo
🏢2Demo Space for my model loggenix-moe-0.3B-A0.1B
-
kshitijthakkar/loggenix-synthetic-ai-tasks-eval-with-outputs
Viewer • Updated • 28 • 17 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v6-with-outputs
Viewer • Updated • 170 • 18 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v5-with-outputs-v7-sft-v1
Viewer • Updated • 170 • 9
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore experiment logs and sync findings with a coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Explore and edit experiment logbooks with AI agent help
-
kshitijthakkar/deepseek-v4-mini-300M-init
Text Generation • 0.3B • Updated • 57 -
kshitijthakkar/deepseek-v4-mini-1B-init
Text Generation • 1B • Updated • 62 • 1 -
kshitijthakkar/deepseek-v4-mini-3B-init
Text Generation • 3B • Updated • 152 • 3 -
kshitijthakkar/deepseek-v4-mini-6B-init
Text Generation • 8B • Updated • 114 • 5
-
kshitijthakkar/qwen3.5-moe-0.87B-d0.8B
Image-Text-to-Text • 1B • Updated • 65 • 1 -
kshitijthakkar/qwen3.5-moe-2.3B-d2B
Image-Text-to-Text • 3B • Updated • 27 -
kshitijthakkar/qwen3.5-moe-4.7B-d4B
Image-Text-to-Text • 5B • Updated • 37 -
kshitijthakkar/qwen3.5-tiny-test
Image-Text-to-Text • 0.1B • Updated • 17
- Sleeping10
TraceMind MCP Server
🤖10MCP server for agent evaluation with Gemini 2.5 Flash
- SleepingAgents22
TraceMind AI
🧠22AI agent evaluation with MCP-powered intelligence
-
MCP-1st-Birthday/smoltrace-recruitment-tasks
Viewer • Updated • 101 • 44 -
MCP-1st-Birthday/smoltrace-smart-home-tasks
Viewer • Updated • 100 • 19
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore experiment logs and sync findings with a coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Explore and edit experiment logbooks with AI agent help
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 31 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 37 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 24 • 1 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
-
kshitijthakkar/deepseek-v4-mini-300M-init
Text Generation • 0.3B • Updated • 57 -
kshitijthakkar/deepseek-v4-mini-1B-init
Text Generation • 1B • Updated • 62 • 1 -
kshitijthakkar/deepseek-v4-mini-3B-init
Text Generation • 3B • Updated • 152 • 3 -
kshitijthakkar/deepseek-v4-mini-6B-init
Text Generation • 8B • Updated • 114 • 5
-
kshitijthakkar/qwen3.5-moe-0.87B-d0.8B
Image-Text-to-Text • 1B • Updated • 65 • 1 -
kshitijthakkar/qwen3.5-moe-2.3B-d2B
Image-Text-to-Text • 3B • Updated • 27 -
kshitijthakkar/qwen3.5-moe-4.7B-d4B
Image-Text-to-Text • 5B • Updated • 37 -
kshitijthakkar/qwen3.5-tiny-test
Image-Text-to-Text • 0.1B • Updated • 17
-
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs4-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx2048
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs1-ctx2048
1B • Updated
- Sleeping10
TraceMind MCP Server
🤖10MCP server for agent evaluation with Gemini 2.5 Flash
- SleepingAgents22
TraceMind AI
🧠22AI agent evaluation with MCP-powered intelligence
-
MCP-1st-Birthday/smoltrace-recruitment-tasks
Viewer • Updated • 101 • 44 -
MCP-1st-Birthday/smoltrace-smart-home-tasks
Viewer • Updated • 100 • 19
- Build error2
Loggenix Moe 0.3B A0.1B Demo
🏢2Demo Space for my model loggenix-moe-0.3B-A0.1B
-
kshitijthakkar/loggenix-synthetic-ai-tasks-eval-with-outputs
Viewer • Updated • 28 • 17 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v6-with-outputs
Viewer • Updated • 170 • 18 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v5-with-outputs-v7-sft-v1
Viewer • Updated • 170 • 9