← People

Kartik Garg

Source-listed: CS undergraduate student · Birla Institute of Technology and Science, Pilani (BITS Pilani), K. K. Birla Goa Campus

  • Language models
  • Interpretability
  • Evaluation

Selected work

4
  1. Alignment Faking - the Train -> Deploy Asymmetry: Through a Game-Theoretic Lens with Bayesian-Stackelberg EquilibriaFirst author · 2025
  2. LLM Anchoring Mechanistic InterpretabilityContributor · 2026
  3. Forge — LLM-Powered Coding Agent with Integrated Benchmark SuiteContributor · 2026
  4. AI Meeting Intelligence AgentContributor · 2026