Kshitij Thakkar PRO
AI & ML interests
Recent Activity
Organizations
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 31 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 35 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 23 • 1 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
-
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs4-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx2048
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs1-ctx2048
1B • Updated
- Build error2
Loggenix Moe 0.3B A0.1B Demo
🏢2Demo Space for my model loggenix-moe-0.3B-A0.1B
-
kshitijthakkar/loggenix-synthetic-ai-tasks-eval-with-outputs
Viewer • Updated • 28 • 15 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v6-with-outputs
Viewer • Updated • 170 • 16 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v5-with-outputs-v7-sft-v1
Viewer • Updated • 170 • 9
-
kshitijthakkar/nirnaya-0.4b-0.2a-decision
Text Generation • 0.4B • Updated • 1.12k -
kshitijthakkar/nirnaya-jevstyle-tasksource-v2
Viewer • Updated • 456k • 116 -
kshitijthakkar/nirnaya-probability-distill-v1
Viewer • Updated • 1.53M • 97 -
kshitijthakkar/nirnaya-probability-distill-v2
Viewer • Updated • 1.95M • 100
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore experiment logs and sync findings with a coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Explore and edit experiment logbooks with AI agent help
-
kshitijthakkar/deepseek-v4-mini-300M-init
Text Generation • 0.3B • Updated • 56 -
kshitijthakkar/deepseek-v4-mini-1B-init
Text Generation • 1B • Updated • 59 • 1 -
kshitijthakkar/deepseek-v4-mini-3B-init
Text Generation • 3B • Updated • 177 • 3 -
kshitijthakkar/deepseek-v4-mini-6B-init
Text Generation • 8B • Updated • 134 • 6
-
kshitijthakkar/qwen3.5-moe-0.87B-d0.8B
Image-Text-to-Text • 1B • Updated • 68 • 1 -
kshitijthakkar/qwen3.5-moe-2.3B-d2B
Image-Text-to-Text • 3B • Updated • 30 -
kshitijthakkar/qwen3.5-moe-4.7B-d4B
Image-Text-to-Text • 5B • Updated • 38 -
kshitijthakkar/qwen3.5-tiny-test
Image-Text-to-Text • 0.1B • Updated • 13
- Sleeping10
TraceMind MCP Server
🤖10MCP server for agent evaluation with Gemini 2.5 Flash
- SleepingAgents22
TraceMind AI
🧠22AI agent evaluation with MCP-powered intelligence
-
MCP-1st-Birthday/smoltrace-recruitment-tasks
Viewer • Updated • 101 • 42 -
MCP-1st-Birthday/smoltrace-smart-home-tasks
Viewer • Updated • 100 • 24
-
kshitijthakkar/trace-path-lab
Viewer • Updated • 8 • 53 - Running on CPU Upgrade5
OpenEnv Arena
🏟5Submit an OpenEnv environment and train a model on it
-
kshitijthakkar/arena-eight-domain-rl-tasks
Viewer • Updated • 400 • 88 -
kshitijthakkar/OpenEnv-Arena-MultiDomain-RL-Mix-v1
Viewer • Updated • 1.63k • 28
-
kshitijthakkar/nirnaya-0.4b-0.2a-decision
Text Generation • 0.4B • Updated • 1.12k -
kshitijthakkar/nirnaya-jevstyle-tasksource-v2
Viewer • Updated • 456k • 116 -
kshitijthakkar/nirnaya-probability-distill-v1
Viewer • Updated • 1.53M • 97 -
kshitijthakkar/nirnaya-probability-distill-v2
Viewer • Updated • 1.95M • 100
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore experiment logs and sync findings with a coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Explore and edit experiment logbooks with AI agent help
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 31 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 35 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 23 • 1 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
-
kshitijthakkar/deepseek-v4-mini-300M-init
Text Generation • 0.3B • Updated • 56 -
kshitijthakkar/deepseek-v4-mini-1B-init
Text Generation • 1B • Updated • 59 • 1 -
kshitijthakkar/deepseek-v4-mini-3B-init
Text Generation • 3B • Updated • 177 • 3 -
kshitijthakkar/deepseek-v4-mini-6B-init
Text Generation • 8B • Updated • 134 • 6
-
kshitijthakkar/qwen3.5-moe-0.87B-d0.8B
Image-Text-to-Text • 1B • Updated • 68 • 1 -
kshitijthakkar/qwen3.5-moe-2.3B-d2B
Image-Text-to-Text • 3B • Updated • 30 -
kshitijthakkar/qwen3.5-moe-4.7B-d4B
Image-Text-to-Text • 5B • Updated • 38 -
kshitijthakkar/qwen3.5-tiny-test
Image-Text-to-Text • 0.1B • Updated • 13
-
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs4-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx2048
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs2-ctx1024
1B • Updated -
kshitijthakkar/moe-1083m-781m-16x8-8L-large-moe-1.3b-bs1-ctx2048
1B • Updated
- Sleeping10
TraceMind MCP Server
🤖10MCP server for agent evaluation with Gemini 2.5 Flash
- SleepingAgents22
TraceMind AI
🧠22AI agent evaluation with MCP-powered intelligence
-
MCP-1st-Birthday/smoltrace-recruitment-tasks
Viewer • Updated • 101 • 42 -
MCP-1st-Birthday/smoltrace-smart-home-tasks
Viewer • Updated • 100 • 24
- Build error2
Loggenix Moe 0.3B A0.1B Demo
🏢2Demo Space for my model loggenix-moe-0.3B-A0.1B
-
kshitijthakkar/loggenix-synthetic-ai-tasks-eval-with-outputs
Viewer • Updated • 28 • 15 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v6-with-outputs
Viewer • Updated • 170 • 16 -
kshitijthakkar/loggenix-synthetic-ai-tasks-eval_v5-with-outputs-v7-sft-v1
Viewer • Updated • 170 • 9