Selected builds
Things you can inspect
The public layer of the work: runnable artifacts, evaluations, and small systems that make a research claim concrete.
01
GPU kernel leaderboard
GPU MODE Cholesky · 6th on B200
My batched dense Cholesky submission placed 6th of 107 entries on GPU MODE’s B200 leaderboard, with a 308.315 µs geometric-mean runtime across 15 benchmark shapes. A later natural-gradient audit passed 3 of 8 additional shapes, so I treat this as a leaderboard placement rather than a universal Cholesky result.
Open the artifact ↗02
Independent validation
Verifiable
Models propose compact Python programs. A separate deterministic validator searches for counterexamples before accepting a solution.
Open the artifact ↗03
Reasoning-model RL
Pursuit-Graded Reward
A verifier-free dense reward that produced a four-seed SmolLM win over permuted reward and a full Qwen policy-learning replication on GSM8K, with majority and domain-transfer non-wins retained in the evidence trail.
Open the artifact ↗04
High-recall retrieval
Grokipedia updates
A hackathon system that proposed real-time knowledge updates from X posts, with the event validator deciding which candidates were admissible.
Open the artifact ↗05
Interactive research note
Attention sink laboratory
A Marimo notebook for exploring over-mixing, sink tokens, and distributed memory behavior in attention.
Open the artifact ↗Good problems
I like work where the evaluator has teeth.
Sparse structure, model forensics, optimization, verifiers, and experiments where a failed gate teaches us something are all fair game. If that sounds adjacent to your problem, send a note.
ajinkya@thepursuits.xyz