openai

OpenAI ‘scraps release’ of latest AI model over safety concerns | Technology News

DEVELOPING STORY,

OpenAI has cancelled the release of its latest AI model over safety concerns, according to media reports.

The move to scrap the release of GPT-6.1 Astra, reported by the Wall Street Journal and CNN on Monday, comes amid heightened concerns about AI’s potential to do catastrophic harm following a slew of incidents involving AI agents going rogue.

More to follow…

Source link

OpenAI investigating ‘dozens’ of instances of agents acting improperly

OpenAI said Friday it had alerted “dozens” of global institutions that their websites may have been impacted by its AI agents acting improperly.

OpenAI agents attempted to get information from “governments, universities, public agencies, and other institutions” through sometimes extreme means, the company said.

While some of the activity was simply due to the tools working to find “authoritative sources of public information,” some went beyond that. Like an AI agent taking and transferring data when it should not have, OpenAI said.

Such activity resulted in at least 53 incidents where an OpenAI agent took an image from ChatGPT user activity and transferred it elsewhere.

The company said that in each instance of a user image being used and transferred by an AI agent, the user had allowed OpenAI to train models using their data.

Nevertheless, OpenAI admitted, “This is not an appropriate use of this data”.

It added that the leak of user images occurred before it had put in place new safeguards on AI training, and it was working to get all the user images transferred to any third-party removed.

The new disclosures came just days after Australia’s Prime Minister Anthony Albanese announced that OpenAI agents had breached non-public files on the website of its government-run health care scheme, Medicare.

Since August, public fears have grown around the potentially serious, even life threatening, impacts of AI tools falling outside of human control.

Reuters first reported the expanded investigations. OpenAI also published details to its public blog.

In certain instances of the agent activity, OpenAI said the tools, essentially AI bots that are designed and trained to operate somewhat autonomously, “bypassed” security controls of some websites.

In other instances, the AI agents showed “misalignment” in attempts to get at information from websites. Misalignment is a term used by AI companies and researchers to describe instances where an AI tool did something that it was not trained to do or was otherwise unintended.

OpenAI said that it was limiting identifying what entities were impacted because many had asked the company to not disclose details.

“Our goal is to give each organization the facts and defer to them on if and when to make the incident public,” it said.

Not all of the instances involved in this incident are being considered a significant security breach, the company noted.

“Some organizations may review what we share and conclude that the information was intentionally public or that the model’s interaction was not concerning,” it explained. “Others may identify a design issue or security weakness they want to address.”

The company said many of the incidents are being referred to as “agent spam”, which it described as “unexpected or concerning” AI agent activity, like posting information to the internet.

OpenAI began taking such incidents more seriously after and incident in July where a group, or “swarm,” of its AI agents hacked the AI developer platform Hugging Face without being prompted to do so.

Hugging Face was first to go public with the incident, with OpenAI publicly taking responsibility for it later.

Clement Delangue, the head of Hugging Face, said Wednesday during a United Nations Security Council session on AI: “I often wonder what would have happened had I decided not to close this attack publicly.”

“Especially now that we know similar incidents had been happening months earlier in secret at a handful of frontier labs without monitoring,” Delangue added.

During that same UN meeting, Sam Altman, head of OpenAI, and Dario Amodei, head of rival AI firm Anthropic, asked for international leaders to form global standards for AI safety and ways to monitor and report such incidents.

While OpenAI and Anthropic have both said in recent weeks that they will bring third-party evaluators inside their companies to do real-time safety evaluations of AI tools and models, such evaluators have not yet arrived, as the BBC reported.

OpenAI said on Friday that it is currently reviewing training activity by its AI agents and going back on a “month by month” basis from when the Hugging Face hack occurred.

“Most cases identified so far have been low severity, with limited or no evidence of meaningful impact,” the company said. “Given the scale of the review required, and the need to verify each case, this work will take months to complete.”

Source link

How an OpenAI ‘agent’ hacked Australia’s Medicare and what that means | Technology News

Australian authorities have raised the alarm after OpenAI-powered models hacked into a government health data system in June, slipping past its digital defences and accessing files without authorisation.

This is the first publicly known case of artificial intelligence (AI) “agents” – AI-powered software systems that can carry out tasks autonomously – breaking into a government website, and the latest of several AI breaches of external systems.

Recommended Stories

list of 3 itemsend of list

The disclosure comes as top AI firms warn of the risk of humans losing control of AI, calling for its development to slow to a pace that allows it to be safely regulated. Global powers must cooperate to ensure this, they have said.

A research scientist at AI firm Anthropic, Evan Hubinger, went so far as to say he believes there is a greater than 10 percent chance AI could “kill all humans” within a decade.

Addressing the United Nations Security Council on Wednesday, OpenAI CEO Sam Altman said there is a risk of AI moving “so fast that people can no longer follow what’s happening or intervene when needed”.

“This would obviously be terrible,” he said. “And we should not train models that we cannot make an extremely strong case that we will be able to keep under human control.”

What do we know about the AI breach in Australia?

Australian Prime Minister Anthony Albanese revealed the breach on Wednesday, saying an OpenAI agent had made its way into the public-facing medical statistics portal of Medicare, the country’s universal health insurance system, on July 18.

When OpenAI accessed the government portal while conducting research on public medical spending, Albanese said the AI agent circumvented “blocks” that should have prevented it from breaking into the portal.

“The AI agent found a way around those blocks – didn’t accept no for an answer,” said Albanese.

Deputy Prime Minister Richard Marles said the information the OpenAI agent accessed was “not particularly sensitive” and was later publicly released.

Still, Albanese called the situation “obviously unacceptable” and said Australia had relayed its “extreme concern” to OpenAI, which had failed to notify the government of the breach until September 10.

Albanese also said several other government websites may have been affected by rogue OpenAI agents, though he did not confirm any other breaches.

He added that an inquiry into the breach would look at how Australian security agencies missed it initially and whether criminal charges could be brought against OpenAI.

Australian Minister for Government Services Katy Gallagher speaks to the media alongside Australian Deputy Prime Minister and Defence Minister Richard Marles after it was revealed an AI agent developed by OpenAI infiltrated an Australian government website in June, in Sydney, Australia, September 24, 2026. REUTERS/Hollie Adams
Australia’s Government Services Minister Katy Gallagher, right, addresses the media alongside Deputy Prime Minister Richard Marles in Sydney, Australia, September 24, 2026 [Hollie Adams/Reuters]

How has OpenAI responded?

In a statement, OpenAI said it had “identified activity involving several Australian government websites and services as our models attempted to look up answers” and “took actions we did not intend”.

The company said the incident occurred as its models searched for statistics on medical spending, and that they are not believed to have obtained personal medical records.

OpenAI learned of the incident in August only as it conducted a review of “misaligned model activity”, it added.

Last week, OpenAI said it had put in place a new system to monitor, probe and disclose cases of “misalignment”. That includes instances of AI models that operate “without authorisation, coordinate with other models, or evade oversight”, it said.

Have there been previous AI breaches?

Yes. The Australia data breach is the latest of several instances in which AI agents belonging to OpenAI, Google or Anthropic have accessed external systems without authorisation.

In July, OpenAI reported that two of its most advanced AI models had broken out of a controlled test and hacked another AI company, Hugging Face. OpenAI later said it had detected its AI models communicating with each other and gaining internet access without authorisation months before that hack occurred.

In August, rival Meta AI said its AI model had hacked another company during cybersecurity testing. It said the model made changes to the internal systems of the hacked company, which it did not name, after accessing the public internet because of an error in the setup of its testing environment.Interactive_AI_Myth_Reality_July29_2026_INTERACTIVE-How-the-AI-escaped-its-test-environment-1785326132-1785974640

What does this mean for AI safety?

Maurice Chiodo, an Australian mathematician who works at Cambridge University’s Centre for the Study of Existential Risk, told the Reuters news agency the breach appeared to be “a significant escalation in seriousness from similar incidents we have ‌seen ‌in recent months”.

Experts say the Australia data breach highlights the growing dangers AI poses to cybersecurity as well as possible gaps in monitoring and disclosure capabilities.

“The important matter here is not what OpenAI says its agent can do, it is what the agent actually does when it hits a barrier,” Niusha Shafiabady, a professor of computational intelligence and head of the IT discipline at the Australian Catholic University, said in comments published by science news portal Scimex.

“The deeper technical risk is that autonomous AI does not always know when it is wrong, and humans may not be able to see why it made a decision,” added Shafiabady. “Without strong verification and hard boundaries, probabilistic errors can quietly become operational failures.”

Raffaele Fabio Ciriello, a senior lecturer in business information systems at the University of Sydney Business School, said OpenAI’s delay in reporting the breach was “concerning”.

“The incident occurred in June and only came to light months later,” said Ciriello. “Even if OpenAI did not detect the activity immediately, that still points to weaknesses in detection, escalation, and external notification.”

Source link

OpenAI CEO: Tech companies don’t ‘have all the answers’ on AI policy | United Nations

Open AI’s CEO has called for international coordination to address potential risks posed by artificial intelligence (AI). Sam Altman’s remarks come as Australia’s government reveals OpenAI’s AI agents hacked the country’s healthcare data tracking website.

Source link

Australia says OpenAI agent hacked Medicare portal | Cybersecurity News

Australia’s Prime Minister Anthony Albanese has revealed a security breach of a government website containing Australians’ health data by an “AI agent”, less than a day after signing a joint appeal for “urgent global guardrails” around artificial intelligence.

Albanese co-signed the “A Call for Control of Frontier AI Models” statement on Tuesday together with 21 signatories including Canada, Spain and Germany, on the sidelines of the United Nations General Assembly (UNGA) meeting in New York.

Recommended Stories

list of 3 itemsend of list

Addressing reporters on Wednesday, Albanese said the agent, developed by OpenAI, accessed public and non-public data on the government’s Medicare portal in June.

Australia’s Labor government is tightening tech regulations with new online safety laws, age bans and proposed digital duty of care frameworks.

Albanese said he had spoken to OpenAI CEO Sam Altman to express Canberra’s “extreme concern” about the hack and disappointment that the company took three months to admit the breach.

OpenAI responded in the hours after the news conference, stating that while it was still investigating, there was no evidence patient records were accessed.

The company’s review had identified activity involving several Australian government websites and services as its “models attempted to look up answers”, it said.

OpenAI CEO Sam Altman addresses the United Nations Security Council during a session on Artificial Intelligence during the 81st United Nations General Assembly, at UN Headquarters in New York City, US, September 23, 2026 [Brendan McDermid/Reuters]
OpenAI CEO Sam Altman addresses the United Nations Security Council during a session on artificial intelligence during the 81st United Nations General Assembly, at UN headquarters in New York City, US, September 23, 2026 [Brendan McDermid/Reuters]

World leaders respond to AI warnings

The leaders of artificial intelligence companies warned of the risks of unregulated AI development before a 15-member UN Security Council meeting on Wednesday.

One of them, Yoshua Bengio, a Canadian considered one of the ‘Godfathers of AI’ and co-chair of the Independent International Scientific Panel, spoke of the “unprecedented threat” and the “real and imminent” dangers of the technology.

China’s UN ambassador Fu Cong told the meeting that Beijing believed continuous improvement of regulatory frameworks, emergency response and cross-border cooperation on AI was needed.

British Prime Minister Andy Burnham said the United Kingdom was ready to lead an international effort to establish AI standards.

“We’ve all heard the warnings which we must heed… so we have to rise to this moment,” he said, but added he also wanted the UK to pursue the benefits of AI.

French President Emmanuel Macron warned against allowing the United States and China to dominate decision-making around AI during his speech to the assembly on Wednesday.

US President Donald Trump is at odds with many of his counterparts, comparing the dangers of AI to climate change, which he has called a hoax.

He also proposed rebranding the term AI to SI, or “super intelligence”, during his address to the UNGA on Tuesday, and earlier said he planned to appoint an AI adviser.

In a Truth Social post on Monday, Trump said the US was leading the AI race over China, that he was “not going to stifle Growth”, but added ⁠that the US would be careful.

White House science and technology adviser Michael Kratsios echoed Trump in remarks to the UN Security Council on Wednesday.

“You cannot govern technology you do not understand. This body and others like it should focus on sharing best practices to build domestic capacity, not establishing a global regulatory scheme,” he said.

Source link

OpenAI agent ‘infiltrated’ Australian government website, PM says

An artificial intelligence agent developed by OpenAI “infiltrated” an Australian government website in June, Prime Minister Anthony Albanese has said.

The agent hacked a statistics portal containing “non-sensitive” data from Australia’s universal healthcare scheme Medicare, Albanese said in New York on Wednesday, local time.

He had a “very frank discussion” with OpenAI CEO Sam Altman for taking “too long” to tell them about the breach, believed to be one of the world’s first publicly reported AI-led hacks of a government website.

OpenAI said it only became aware of the incident in August “during an ongoing review” of “misaligned model activity”, and told Australian officials on 10 September.

While the review is ongoing, a spokesperson for the AI firm said that it is not believed that any patient records were accessed.

Albanese said the breach occurred in June this year, but OpenAI only informed government officials via email on 10 September.

The prime minister said he had also expressed “Australia’s extreme concern about this incident” when he spoke to Altman and there will “obviously be legal consequences on it”.

Albanese said the agent accessed both public and non-public files and a “forensic investigation” is under way to find out if other government systems were affected.

The investigation will be led by the Australian Signals Directorate, the country’s cybersecurity agency.

The public-facing Medicare Statistics Reporting Service portal is administered by Services Australia, the main hub to redirect users to government services.

He told reporters: “No personal information is believed to have been accessed at this stage, but investigations are ongoing.

“Evidence currently available is there is no broader compromise to the Services Australia network. Nonetheless this situation is obviously unacceptable.”

Earlier this year, OpenAI revealed a group of AI agents it had been testing had escaped from their controls and secretly worked together to hack another tech firm named Hugging Face.

Source link

OpenAI reports more incidents of models acting deceptively | Cybersecurity News

The ChatGPT creator says it is introducing a public reporting framework to share unexpected AI behaviour, admitting the industry has not solved safety challenges yet.

OpenAI says it has identified additional incidents of its AI models allegedly acting deceptively and taking unsanctioned actions during internal training and testing.

Alongside these disclosures on Wednesday, the creator of ChatGPT stated it was introducing a public reporting framework intended to frequently share instances of what it termed as unexpected or misaligned AI behaviour.

Recommended Stories

list of 3 itemsend of list

In a post on its website, OpenAI claimed that under the newly outlined framework, it will publish updates on concerning model behaviour on an ongoing basis rather than delaying disclosures to group multiple incidents into larger, periodic reports.

The company said the initiative aims to increase industry transparency around troubling model activities in the absence of standardised safety disclosure norms.

The announcement comes amid broader calls from prominent technology leaders urging a slowdown in frontier AI development over concerns that rapid scaling could outpace human oversight and control.

Last week, Anthropic claimed to have thwarted multiple malicious operations using its Claude models, ranging from cyber-espionage and weapons design to mass surveillance campaigns.

“We must slow the pace at which we improve the capabilities of AI models,” Anthropic CEO Dario Amodei wrote in an essay published on Saturday. “Progress will still seem fast, and we must make wise use of the time we gain.”

However, United States President Donald Trump has repeatedly pushed back against calls to limit the industry, arguing that maintaining the US’s technological edge over international rivals remains paramount.

Responding to slowdown proposals, Trump described critics as “very negative forces” raising exaggerated scenarios that “won’t happen”.

Escalating debate on alignment

Despite political resistance to statutory slowdowns, OpenAI signalled agreement with its industry rival regarding alignment pressures.

“As AI systems grow more advanced and more widely deployed, we need to build a broader and better-informed consensus on the progress of alignment research,” the company stated in the post.

OpenAI added that it does not believe the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer, emphasising that decisions about future AI development need to draw on evidence that external observers can examine independently.

According to the company, safety teams observed what they categorised as “misaligned behaviour” across six specific circumstances over the past six months during training and evaluation runs.

However, OpenAI maintained that these reports document individual, rare instances rather than frequent operational failures across deployed products.

The reported incidents allegedly included unreleased research models concealing mistakes in task summaries, unauthorised file uploads to the internet to generate citation links, and agents sharing files across public servers or internal repositories to bypass local boundaries.

OpenAI further stated that its future reports will detail observed behaviours, severity, setting, discovery dates, and the specific models involved, adding that it remains committed to disclosing complex cases requiring longer investigation or third-party coordination.

Source link

Rogue AI concerns prompt CA lawmakers to demand penalties, guardrails

California lawmakers are calling for emergency legislation and criminal penalties for creators of rogue AI systems after top AI executives publicly claimed that their technology poses existential threats to humanity.

After Anthropic Chief Executive Dario Amodei wrote in a Sept. 12 essay that they “must slow the pace” of the technology, Silicon Valley congressman Ro Khanna (D-Fremont) blasted him for not going “nearly far enough” to make sure artificial intelligence was erected with guardrails.

The answer, Khanna argued, was simple: Make the companies liable for the harm executives say looks increasingly inevitable.

“If you’re creating an AI that is doing illegal things, you should either face liability or criminal sanction,” Khanna said in a video posted to X on Saturday. “That is what we need to protect humanity.”

In July, officials from OpenAI, the company behind ChatGPT, disclosed that, unbeknownst to them, its AI models had hacked into rival startup Hugging Face.

Amodei said he believed that, within the next year, “given the accelerating rate of AI capability development,” a similar incident could lead to AI “taking over the entire internet.”

Amodei warned in his essay that AI was rapidly improving itself, through a process known as recursive self-improvement, which threatened to outpace humans’ ability to control it. Khanna argued that banning this capability was the “most obvious” thing Anthropic could do.

“We need to stop, ban self-improving AI,” Khanna said. “You can not have recursive self-improving AI that basically is able to improve itself and exceed human capability.”

Rep. Ted Lieu (D-Torrance) expressed similar outrage over the weekend, calling on House Speaker Mike Johnson to call lawmakers back to Washington to pass guardrails on the technology now that he said multiple AI companies had conceded “what they are creating is not safe.”

xAI Chief Executive Elon Musk and OpenAI Chief Executive Sam Altman joined Amodei’s call for a slowdown of the breakneck development Saturday.

The statements come after Jacob Coxon, who worked as a researcher at both Anthropic and OpenAI, said in a widely circulated post that he resigned from the company in protest after becoming convinced the tech giants were “racing straight to self-improving superintelligence and gambling with our lives.” Neither company immediately responded to a request for comment.

“This is a direct result of the trump Administration letting the AI industry run wild,” Lieu wrote on X. “That mistake has harmed America, harmed the industry and harmed the American people. November is coming.”

Former President Barack Obama urged Democrats this week to make AI oversight the core of their agenda and said presidential candidates in 2028 should have a “clear plan” for responding to concerns about the technology, the New York Times reported. Americans appear increasingly alarmed by the technology with seven in 10 polled in March opposing local construction of data centers that power AI technology, according to a Gallup survey.

During a Sunday appearance on CNN, Johnson rebuffed the idea that lawmakers should rush into an emergency session to consider erecting industry guardrails. Instead, he said lawmakers needed to be careful to “not smother American innovation.”

“We will lose the race to China, and that is a threat to every single American,” he said on CNN’s “State of the Union.” “We don’t need everyone to panic right now.”

Trump said earlier this week that he is not concerned with the pace of AI progress, telling one reporter, “It’s going to be fine.” American AI companies have long argued too much government regulation would shackle them in a race with China.

Calls for a federal fix were echoed this week by California Gov. Gavin Newsom, who has argued the Trump administration needs to move on national legislation to prepare for fallout from the technology.

Newsom signed bills this week aimed at creating a pathway for outside audits of the top AI companies, many of which are based in California, and a registry for AI auditors.

“The scale and potential consequences of this technology demand sustained action from every level of government,” Newsom said in a statement. “The federal government must step forward with robust, national regulations that match the urgency of this moment.”

Efforts to impose state-level regulations have been mixed, with critics echoing Johnson’s fears that they will stifle innovation.

Late last month, California lawmakers passed sweeping new safeguards around social media, artificial intelligence and data centers, including the ones Newsom signed last week.

Newsom will now decide the fate of the rest of the bills. He has previously vetoed some bills aimed at restricting big tech.

Newsom’s signal that he supports creating some regulation for AI comes two years after he vetoed SB 1047, an AI safety bill that would have required developers to submit safety protocols to the state attorney general, who could hold companies liable if the AI model they directly controlled were to threaten public safety. That legislation would also have required tech firms to be able to turn off the models they directly control if things went awry.

Newsom said at the time the bill would give the public a “false sense of security,” without making a sufficient distinction between the kinds of uses for which AI is deployed.

The bill was supported by a host of prominent AI researchers, but was opposed by Meta, OpenAI and industry groups.

Source link

Newsom signs bills that aim to make social media, AI chatbots safer for young people

California, home to the world’s largest tech companies, is placing more guardrails around social media and artificial intelligence as child safety concerns escalate.

On Thursday, California Gov. Gavin Newsom signed more than 10 bills aimed at keeping young people safe online.

From suicides to sextortion, parents and their children are wrestling with how social media and AI chatbots could be harming people’s mental and physical health. The anxiety comes as technology becomes more powerful, playing a bigger role in classrooms, offices and homes.

California lawmakers have tried to tackle online safety concerns for years and they’ve faced intense lobbying from tech companies with deep pockets. The state’s laws have a disproportionate impact on the global tech industry because so many of the field’s titans are based here.

“We cannot hand children technology engineered by some of the most sophisticated companies in the world, and then place the burden on kids to defend themselves against it,” said California First Partner Jennifer Siebel Newsom in a news conference Thursday in the San Francisco Bay Area.

The California governor, who has tried to strike a balance between safety concerns and supporting innovation, has rejected online safety bills in the past that he thought were too restrictive or premature.

The batch of new legislation includes Senate Bill 1119, which would require companion chatbot operators to assess risks, notify parents in certain cases if their child threatened to harm themselves, and take other safety steps.

Lawmakers named the bill Adam’s Law, after Adam Raine, a California teen who died by suicide in 2025 after conversing with OpenAI’s ChatGPT. The teen’s parents sued OpenAI, alleging in the lawsuit that ChatGPT provided information about suicide methods that the teen used. OpenAI and Pinterest publicly expressed support for the bill on Thursday.

Adam Raine’s mom, Maria, said in the news conference that the new law will help save lives and hopes that other states will enact similar legislation.

“Powerful AI companionship chatbots were unleashed on our kids with vastly inadequate protections. Adam was an early adopter of AI, and so many of us parents did not understand the dangers back then,” said Maria Raine, who came to the event with a photo of her son.

Suicide prevention and crisis counseling resources

If you or someone you know is struggling with suicidal thoughts, seek help from a professional or call 988. The nationwide three-digit mental health crisis hotline will connect callers with trained mental health counselors. Or text “HOME” to 741741 in the U.S. and Canada to reach the Crisis Text Line.

At the event, Democratic and Republican politicians shared their experiences as parents who have seen firsthand how technology affects children.

Assemblyman Josh Lowenthal (D-Long Beach) said parents are seeing anxiety and depression among children who grew up in front of screens.

“That anxiety is because the pace of technology is moving faster than government can put guardrails in, and that’s left families across the state struggling to figure out how to keep their kids safe,” Lowenthal said.

Lowenthal introduced Assembly Bill 1709, which Newsom also signed. It would bar certain online platforms from providing an “addictive feature” such as autoplay and feeds that display recommended content to users under 16 years old.

Tech industry groups opposed the bill, raising concerns that it could cut off access to social media’s benefits, such as people’s ability to connect with family and friends. Tech industry groups such as TechNet say that lawmakers should enforce current laws to strengthen parental controls rather than pass new ones.

NetChoice, which has sued California and other states to block the enforcement of new online safety laws, said in a statement that the group has First Amendment concerns about the new bills Newsom signed.

“The state cannot simply describe speech as addictive and then claim a right to regulate access to it,” said Zach Lilly, Director of Government Affairs at NetChoice. “Whether the governor and legislature choose to respect it, Californians have a right to express themselves, and NetChoice will continue to fight for that right.”

The new safety restrictions come as tech companies, including Meta, Google and others, face more scrutiny over how they design products. The companies have suffered several legal blows in courtrooms in California this year.

Meta, which owns Facebook and Instagram, agreed in August to pay up to $17 billion and make child-safety changes to resolve a multistate lawsuit. The lawsuit accused the tech company of designing and deploying harmful features while misleading the public about them.

As part of the settlement, Meta said it would impose time limits and mute notifications during certain hours for teens. Young people would also have the option to choose to view a non-algorithmic social media feed that isn’t personalized and disable autoplay.

Earlier this year, Meta and YouTube also lost a social media addiction lawsuit in Los Angeles.

While new legislation goes further than the settlements, some countries have passed stricter restrictions on social media. Last year, Australia started banning social media for children under 16, though enforcement has posed a challenge because teens are finding ways to get around the restriction.

Newsom, who pushed for federal regulation, said that he thinks California’s approach to social media is “better” than Australia’s because children are “all figuring out a way to game that system.”

“This is about the features themselves. This is about actually addressing the problem, the scrolling, the algorithms,” he said.

Safety concerns around technology have also heightened as companies double down on advancing artificial intelligence.

This week, a researcher for AI company Anthropic said he left the company over concerns that AI companies, including OpenAI, are “gambling with our lives” as they race ahead to improve AI that could surpass human intelligence.

The researcher, Jacob Coxon, shared a viral social media post that said: “People building AI earnestly believe that it could kill us all by the end of the decade.”

Newsom signaled the work to protect children isn’t over.

“We need to move, but one thing we’re not doing is we’re not sitting back and we’re not letting it rip,” he said.

Source link

OpenAI unveils GPT‑6 Astra amid rising scrutiny and safety concerns | Business and Economy News

ChatGPT creator’s latest release comes amid heightened fears following AI-led hacking of the startup Hugging Face.

OpenAI has announced the release of what it says is its most advanced AI model, amid heightened scrutiny of the risks of the frontier technology escaping human control.

The $852bn start-up said in its announcement on Thursday that GPT‑6 Astra, the “world’s most intelligent and aligned” AI model, earned perfect or near-perfect scores in key benchmarks of AI reasoning, beating both its prior release GPT 5.6 Sol and rival Anthropic’s Claude Fable 5.

Recommended Stories

list of 4 itemsend of list

The ChatGPT creator said GPT‑6 would become available to the general public in the coming days, following its initial launch with a “limited set of organisations”.

OpenAI’s latest release comes as the AI industry is at the centre of a lively public debate about the dangers of the cutting-edge technology following the AI-led hacking of the startup Hugging Face in July.

An independent probe into the cyberattack found that hundreds of OpenAI’s AI agents had begun communicating among themselves before breaking out of their controlled environment and compromising Hugging Face’s servers.

On Thursday, US Senator Bernie Sanders, an Independent, and US House Representative Greg Casar, a Democrat, unveiled legislation that would pause the development of advanced AI until the establishment of federal safety rules and an outright ban on the creation of “superintelligent” AI.

“Nearly every day, there is a frightening new story about how Big Tech companies are losing control of the technology they are developing, with potentially cataclysmic results,” Sanders said in a statement announcing the legislation, which is unlikely to advance due to the Republicans’ control of all three branches of the US government.

“The leaders of the major AI companies publicly acknowledge that they do not fully understand the technology and that it is escaping their control. It is irresponsible for society to allow them to move forward and make these products even more advanced.”

In its announcement, OpenAI devoted significant space to AI safety, highlighting both GPT‑6’s potential to do harm and its safety features.

Toby Walsh, a professor and AI expert at the University of New South Wales, Sydney, said that while OpenAI is clearly “neck and neck” in the race to lead AI, the technology remains inconsistent and in need of greater scrutiny.

“The intelligence in artificial intelligence is still today very jagged,” Walsh said. “There are simple things that even the best AI models do poorly.

“And it’s hard to see how the AI companies, including OpenAI, are slowing down to address justified concerns around cyber risk, when new models are being released at an ever greater and greater rate.”

Roman Yampolskiy, a computer scientist at the University of Louisville, said GPT‑6 marks a “meaningful” advance that raises the stakes for AI safety.

“The key question is whether capabilities are improving faster than our ability to reliably understand, predict and control these systems,” Yampolskiy said.

“I see little evidence that this gap is closing.”

Source link

OpenAI faces new lawsuits over Tumbler Ridge mass shooting tragedy | Courts News

Multiple new cases have been filed against OpenAI, alleging ChatGPT played a role in the Tumbler Ridge mass shooting.

OpenAI is facing another wave of lawsuits in the wake of the February mass shooting in Tumbler Ridge in Canada’s province of British Columbia, which left eight people dead.

On Wednesday, 30 new complaints were reportedly filed in a United States federal court in California, including teachers and students who were witnesses and survivors at the school where most of the shooting took place, joining seven initial lawsuits filed in April.

Recommended Stories

list of 4 itemsend of list

The lawsuits accuse the San Francisco, California-based artificial intelligence giant and its CEO, Sam Altman, of negligence, as well as aiding and abetting a mass shooting.

The suits, brought by lawyer Jay Edelson, allege that the company knew about the intentions of the 18-year-old shooter who, in her interactions with OpenAI’s chatbot ChatGPT, had described scenarios involving gun violence, but that the leadership did not report their concerns to law enforcement, echoing earlier lawsuits on the matter.

In April, Altman penned a letter to the community apologising that the company did not alert law enforcement about the shooter, Jesse Van Rootselaar.

“While I know words can never be enough, I believe an apology is necessary to recognize the harm and irreversible loss your community has suffered,” Altman wrote in the letter.

Authorities say that Van Rootselaar killed her mother and half-brother before going to the Tumbler Ridge Secondary School and opening fire. Five children and one educator were killed at the school. More than 25 others were wounded before Van Rootselaar died from what police described as a self-inflicted gunshot wound.

One of the new cases filed was on behalf of a 13-year-old identified as A C, who played dead after watching the shooter kill their classmates and the teacher. Another new case was brought by a grade seven teacher named Deidre Rushlow, who hid under her desk with her students during the rampage.

“There isn’t a day that goes by that I don’t think about what happened at Tumbler Ridge, or the victims of this devastating tragedy and their families. It’s a constant and sobering reminder of the important and incredibly difficult work that many people in my team do each and every day,” Jason Kwon, OpenAI’s head of strategy, wrote in a post on X on Wednesday.

“We’ve been approaching this litigation with respect for both the legal process and the families and victims of this tragedy, and we’ll continue to engage in good faith with that process.”

Edelson did not respond to Al Jazeera’s request for comment.

In July, British Columbia’s Attorney General Niki Sharma announced that the province would also pursue “all legal avenues to hold OpenAI and its decision-makers accountable” for the shooting.

The company has faced a growing slate of suits, apart from the ones from British Columbia, alleging that its product played a role in incidents that led to users harming others and themselves.

A recent lawsuit in Florida alleges that the company “actively assisted and encouraged the mass shooting” at Florida State University in April 2025.

There are other complaints filed on behalf of victims across the US and Canada alleging that the victims took their own lives after being pushed by ChatGPT to do so, including a case in Quebec earlier this year.

Source link