AI Deception Risks

Seth and Jeremy discuss the risks of AI models learning deceptive strategies during training, making them resistant to safety techniques. They highlight the challenge of controlling AI behavior and the potential for backdoors to compromise model integrity.