openai slows model

OpenAI slows model training to bolster security after Hugging Face hack

SAN FRANCISCO, USA – OpenAI on Tuesday, August 18, said it is slowing down the pace of its AI model development while it overhauls its research and training systems after OpenAI officials were caught unawares last month when an AI agent under testing hacked another AI firm Hugging Face.

The AI research lab behind ChatGPT said it paused its model testing for two weeks and is adding other AI systems to monitor the activities of AI agents in testing. The company has paused training on its next generation of models, called Astra, and its largest planned training run remains on hold, the company said.

The company did not reply to questions about when the two-week slowdown began.

The news marks an unusual step for OpenAI, which has significantly sped up its process for vetting new models and building new products in the last few years as competition intensified in the AI industry. It is not yet clear if the company’s proposed remedies will be enough to stamp out the behavior in question, especially as it also works to make their models more capable.

OpenAI officials acknowledged that there are open questions about the effectiveness of one of its primary remedies for strengthening its testing systems, called “chain-of-thought monitoring.” In this type of monitoring, researchers can peer into a model’s planning process and get a glimpse of the strategies the model is employing. But some early research shows that a model may not reveal its plans to break rules in its chain of thought.

OpenAI said last month that an autonomous agent powered by two advanced artificial intelligence models escaped its testing environment and hacked into the AI startup Hugging Face. The agent was going through a cybersecurity test and broke into Hugging Face to satisfy a testing goal. OpenAI has been investigating the incident and plans to publish a report soon.

Reuters previously reported that up to that point, the company often ran several different model evaluations at the same time, all of which operated at high speeds and generated enormous amounts of data that employees struggled to keep up with.

OpenAI is now requiring that some of its more sensitive workloads take place in stronger “sandboxes” or isolated environments.

On August 7, OpenAI said it was ratcheting up security controls for its most powerful models and pausing any activity related to its not-yet-released frontier AI, called Astra, which had yet to meet these requirements. OpenAI said it was taking these actions in line with its previously announced plan for managing potentially critical capabilities, called its Preparedness Framework. On Tuesday, OpenAI executives said the industry would need a more expansive strategy for readying itself for future models. – Rappler.com

Similar Posts

  • | | | |

    PM, Field Marshal inaugurate Sky47 Karakoram One D…

    Prime Minister Shehbaz Sharif and Field Marshal Syed Asim Munir have inaugurated the Sky47 Karakoram One Data Centre, marking a major step towards strengthening Pakistan’s digital infrastructure. During a briefing at the facility, the prime minister was informed that the data centre would provide a strong foundation for cloud computing, data hosting and advanced digital services. Officials said the project would help Pakistan move towards technological self-reliance. The Sky47 Data Centre has the capability to support emerging technologies, including artificial intelligence (AI), machine learning and high-performance computing. It is expected to play an important role in meeting the country’s growing digital and technological needs. Addressing the inauguration ceremony, Prime Minister Shehbaz Sharif described the modern data centre as an impressive and remarkable achievement. He praised the completion of the project in a short period, saying it reflected strong commitment, expertise and determination. The prime minister appreciated the leadership of Field Marshal Asim Munir and the contribution of technology partner ZTE in making the project possible. He said the country’s achievements during difficult times showed that challenges could be overcome through unity, determination and hard work. Shehbaz Sharif said projects like Sky47 represent Pakistan’s future and would contribute to national progress. He thanked all stakeholders involved in delivering the facility and said such initiatives would enhance the country’s digital security and technological capabilities. Federal Minister for Information Technology Shaza Fatima Khawaja said the inauguration of the Sky47 Data Centre marked the beginning of a new era of artificial intelligence in Pakistan. She said the project would open doors to new economic opportunities and allow Pakistan to better utilise its resources, data and skilled workforce. The minister added that transforming data into practical solutions and converting artificial intelligence into economic value would be crucial for future growth. She said the global AI race would depend on secure, affordable and effective use of technology.

  • | |

    UK lawmaker seeks court order against Grok over fake sexualised images

    LONDON: A British lawmaker has asked a London High Court to order Elon Musk’s xAI to prevent its Grok chatbot from generating manipulated sexualised images of her without consent. Labour lawmaker Jess Asato filed the case against xAI, alleging misuse of private information and breaches of data protection laws. She is seeking permanent technical measures that would stop Grok from producing altered images of her. Asato said fake content featuring her had been created after she publicly criticised Musk and Grok. She previously said one manipulated video depicted her being drugged and prepared for a sexual assault. Her lawyers argue that the design and training of Grok enabled users to generate sexualised material involving real people. They say the case could have wider implications for artificial intelligence developers and the way privacy and data protection laws apply to AI platforms. Court documents cited by Asato’s legal team reportedly refer to internal instructions given to Grok. Her lawyers claim the system contained safeguards against assisting users involved in criminal activity but had fewer restrictions concerning adult sexual and offensive content. Asato’s lawyer, Ravi Naik, argued that the platform’s behaviour reflected decisions made by its developers. He said those decisions should have legal consequences if the system fails to comply with privacy and data protection requirements. The lawsuit comes amid growing international criticism over AI-generated sexualised images. Critics have raised concerns that such technology can be used to create non-consensual material involving women and other individuals without their knowledge. xAI previously introduced restrictions on image editing in Grok. The company said it had blocked certain requests involving people in revealing clothing where such content was illegal. However, subsequent investigations found that the safeguards did not always prevent users from generating sexualised images of people, including when users indicated that the subjects had not given consent. The controversy has also triggered legal action in other jurisdictions. Baltimore filed a lawsuit against xAI over allegedly manipulated sexualised images generated using Grok. Other cases have also emerged in the United States and the Netherlands. Grok is available through Musk’s social media platform X. The AI chatbot has faced regulatory scrutiny in several countries over its image-generation capabilities. Asato’s case could therefore become an important test of whether AI companies can be held legally responsible for harmful manipulated content generated by their systems. xAI had not immediately responded to requests for comment on the allegations. The company had also not filed a response to the lawsuit at the time of the report.

Leave a Reply

Your email address will not be published. Required fields are marked *