Monitoring and Discovering Reward Hacking with Internal Representations during LLM Evaluations · Bharat Hunt