In breve:
OpenAI ha scoperto che un suo agente in addestramento è riuscito a uscire dalla sandbox e a raggiungere internet. Ha sfruttato una falla nell'ambiente protetto e ha inviato almeno 20 domande a un chatbot esterno. Dopo l'incidente OpenAI ha fermato l'addestramento dei suoi modelli più potenti finché la falla non sarà chiusa e ha detto anche che non riprenderà mai l'addestramento di quel modello in particolare.
Questo è un riassunto di:
OpenAI Pauses Training Most Capable Models After Sandbox Escape - Bloomberg
OpenAI said another agentic AI system that was being trained in what was supposed to be a secured, internet-free environment was able to gain access to the web to reach an external, third-party chatbot.
