Shreshth Rajan
I study Pure Math and Philosophy at Harvard. I took leave to be the first person on the team at Forge and founding engineer at Pillar. I'm an investor or advisor to Arda (seed), Etched (seed), Applied Compute (seed), Sable (seed), Transfyr AI (advisor), Advanced Machine Intelligence (seed), and Operating Intelligence (seed).
I study Pure Math and Philosophy at Harvard. I took leave to be the first person on the team at Forge and founding engineer at Pillar. Research Envaudit. Auditing reward hackability in code RL training environments. 28.5% hackability on SWE-bench Verified; +14.14pp Pass@1 inflation across 134 model submissions. [arxiv] Arithmetic Intensity Aware Quantization. Post-training quantization optimizing FLOPs/byte. 1.66x speedup. [arxiv] Uncertainty Gated Retrieval. Test-time retrieval for semantic segmentation. +11% accuracy, 87.5% cost reduction. [arxiv] Multiver. Multi-agent…
saved by
related reading
- Composer2.pdfcursor.com
- As Rocks May Think | Eric Jangevjang.com
- Asymmetry of verification and verifier’s rule - Jason Weijasonwei.net
- Trending Papers - Hugging Facepaperswithcode.com
- Explore | alphaXivalphaxiv.org
- rsrch spacersrch.space
- The Verification Horizon: No Silver Bullet for Coding Agent Rewardsarxiv.org
- When AI Writes the World's Software, Who Verifies It? — Leonardo de Mouraleodemoura.github.io
- rsrch spacersrch.space
- AI Will Write All the Code. Mathematics Will Prove It Works.menlovc.com
- LLM-as-a-Verifier: A General-Purpose Verification Framework | alphaXivalphaxiv.org
- Our Problems · Proximalproximal.ai