GPT 6 Is Out In OpenAI’s Sandbox
The new GPT 6 model comes out amid a growing concern about releasing these new models – how they might be used to hack systems, and overrun network defenses. So the model isn’t coming out all at once.
In the software development world of the late twentieth century, the word “beta” was used quite liberally. A company would release a piece of software to a specific community before fully releasing it to the world. Those advance users would help look for bugs.
Now, a lot of the focus of this sandboxed user phase is to limit a model’s use to a small trusted circle – because otherwise, its power might harm business as usual.
Here’s how Katelyn Chedraoui explains it at CNET:
“OpenAI said Astra has been extensively tested for cybersecurity-related vulnerabilities, and it will stagger its rollout. GPT-6-Astra will be available to approved cybersecurity defenders in OpenAI’s Daybreak program starting on Thursday, with plans to bring the new model to paying ChatGPT subscribers and its API over the next several days.”
If the Daybreak folks only have it for a few days, is that enough time to really vet the system?
Chedraoui cites a statement by OpenAI Chief Scientist Jakub Pachocki:
“As AI takes on more of its own development, we need to keep people meaningfully involved,” Pachocki said. “People must remain able to decide the direction of further progress and the future that it creates.”
There’s also OpenAI’s inclusion in a short list of parties that the U.S. White House is comfortable with, according to recent policy. The Trump administration has reportedly put together a “voluntary AI framework” for regulating models like Mythos and GPT 6. But the details are a secret, which is raising quite a few questions.
“Four federal agencies have been sued amid calls to release information about the secret framework that the Trump administration uses to conduct safety reviews of frontier AI models prior to release,” writes Ashley Belanger at Ars Tecnica . “In a Wednesday press release announcing the lawsuit, a nonpartisan nonprofit called Protect Democracy alleged that ‘almost no details’ have been released to the public or Congress. To everyone except a few vague ‘trusted partners,’ it remains unclear what the government’s review process looks like, which companies are involved in constructing the framework, or what legal authority Trump officials have to conduct the reviews.”
Presumably, as one of those trusted partners, top people at OpenAI do now what the criteria are. But they’re not telling the rest of us. The idea, expressed in this piece and elsewhere, is that corruption and favoritism can thrive in this type of setup.
In general, many feel like the U.S. has been behind in regulating AI.
“The industry is begging governments to take action,” writes Hadas Gold at CNN . “Leaders of the major AI companies have joined calls for a slowdown and worldwide cooperation on standards and safety.”
Gold notes that a British agency is considered one of the top regulators of AI, with a lot more power than its U.S. counterpart, even though most of the top AI companies come out of the United States. The report also brings us these quotes from two top staffers, one at Meta, and the other at OpenAI:
“This feels like early COVID … There’s an emergency vibe that’s appropriate here.” – Joshua Saxe, former Meta senior technical expert
“There’s a 20-30% chance that existing methods for alignment and control break down before we reach broadly superhuman AI.” - Paul Cristiano, a former top staffer at OpenAI
Simply speaking, there’s a good deal of nervousness about a process that’s located in the White House, with undescribed protocols.
The Apprentice Generation
This last interesting piece comes from the CNET coverage above.
“Astra was also OpenAI’s first model that involved previous AI models very heavily in the training process, with the older generations helping train the new one,” Chedraoui writes.
That line brought to mind, for me, a funny sort of visualization, of patient GPTs mentoring one another.
So I asked GPT about this.
“Since I am GPT-5.6 Sol,” the model responded, “models from my generation could have been used in Astra’s development. That doesn’t necessarily mean this particular running instance of me participated. Rather, GPT-5.6-class models could be deployed internally to do things such as generate or curate training data, evaluate outputs, provide training signals/supervision, identify failures, help with coding and debugging of the training infrastructure, and monitor or troubleshoot training runs.”
So there you have it. GPT 6 is standing on the shoulders of giants, as it were.
As for the White House framework, like so much else, we’re waiting and watching to see how the whole thing shakes out. Stay tuned.
Loading article...