Monitoring and Discovering Reward Hacking with Internal Representations during LLM Evaluations — ThinkLLM