AI Adversarial Training Boosts Robot Self-Repair Capabilities
AI News Brief: OpenAI recently unveiled an automated red-teaming system called GPT-Red, whose core mechanism operates via self-play, letting AI systems attack and defend against each other to continuously uncover and strengthen the model’s capabilities in safety, alignment, and resistance to prompt injection attacks. Through this self-play training loop…