Kartik Garg
Source-listed: CS undergraduate student · Birla Institute of Technology and Science, Pilani (BITS Pilani), K. K. Birla Goa Campus
- Language models
- Interpretability
- Evaluation
Selected work
4- Alignment Faking - the Train -> Deploy Asymmetry: Through a Game-Theoretic Lens with Bayesian-Stackelberg EquilibriaFirst author · 2025
- LLM Anchoring Mechanistic InterpretabilityContributor · 2026
- Forge — LLM-Powered Coding Agent with Integrated Benchmark SuiteContributor · 2026
- AI Meeting Intelligence AgentContributor · 2026