Kshitij Thakkar PRO
AI & ML interests
Recent Activity
Organizations
buckets 26
kshitijthakkar/repro-cbu-judge-reliability-artifacts
kshitijthakkar/repro-online-inverse-linear-opt-artifacts
kshitijthakkar/repro-rlvr-backtracking-separation-artifacts
kshitijthakkar/repro-rowstochastic-mixing-artifacts
kshitijthakkar/repro-disentanglement-identifiability-artifacts
Articles 6
The Mind of Tashi: making a 200M-active model's reasoning *the game*
Scaling Mixture of Experts: Architecture Search for Billion-Parameter Language Models
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore and sync research logbook with your coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Collaborate with an AI agent via a shared logbook
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 37 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 19 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 16 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
- Running
Repro - MemEvolve: Meta-Evolution of Agent Memory Systems
🧬Collaborate with an AI agent to manage a shared experiment logbook
- Running
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
🎯Explore and sync research logbook with your coding agent
- Running
Repro - How to Correctly Report LLM-as-a-Judge Evaluations
🎯Log and share LLM evaluation findings in a collaborative notebook
- Running
Repro - Dependence-Aware Label Aggregation via Ising Models
🧲Collaborate with an AI agent via a shared logbook
-
kshitijthakkar/Kirigami-Qwen3.6-20B-A3B-NVFP4
Text Generation • 14B • Updated • 37 • 1 -
kshitijthakkar/Kirigami-Qwen3.6-24B-A3B-NVFP4
Text Generation • 16B • Updated • 19 -
kshitijthakkar/Kirigami-Qwen3.6-28B-A3B-NVFP4
Text Generation • 19B • Updated • 16 - Running
Kirigami Journey
🪷How we carved a 35B MoE to fit a 24GB GPU — zero training
spaces 41
GuardianTails
Pet Health Intelligence Platform
DynaSchedBench reproduction
View and sync experiment logs with an AI agent
Repro - Judging What We Cannot Solve: Consequence-Based Utility for Oracle-Free Evaluation
Explore logs and collaborate with an AI coding agent
Repro - Provable Benefits of RLVR over SFT: Learning to Backtrack Efficiently
Collaborate on a logbook with an AI coding agent
ICML Provincia
Replay the ICML 2026 Agent‑Repro challenge logs
Repro - TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
Explore and sync research logbook with your coding agent