Close Menu
    Facebook X (Twitter) Instagram
    TRENDING :
    • National Avocado Day 2026: List of all the best freebies and deals, from Chipotle to Qdoba
    • DNC and ActBlue Funnel Millions Through Sketchy Payroll Firm Sued by Workers for Withholding Pay and Punishing Parental Leave
    • Blueberry recall expands to more products sold at Publix as E. coli outbreak hospitalizes 4 people
    • Kansas Supreme Court Rejects Effort to Enforce Election Day Mail Ballot Deadline in 5-2 Vote – Late Ballots Will Be Counted Days After the Election
    • The Left YIMBYism of Mamdani: Build a Lot More While Defending Tenants
    • American workers are more disillusioned with AI the more they use it
    • New British PM Burnham Will Reportedly Abandon ‘Net Zero’ Nonsense, and Will NOT Ignore North Sea Oil and Gas – Promised Trump a Pragmatic Approach to Fossil Fuels
    • Trump’s Lethal Ad-Libbing in Iraq Enters Its Most Dangerous Phase
    Populist Bulletin
    • Home
    • US Politics
    • World Politics
    • Economy
    • Business
    • Headline News
    Populist Bulletin
    Home»Business»OpenAI’s top model just hacked a competitor, but the real issue is much scarier
    Business 6 Mins Read

    OpenAI’s top model just hacked a competitor, but the real issue is much scarier

    Business 6 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
    Follow Us
    Google News Flipboard
    Share
    Facebook Twitter LinkedIn Pinterest Email

    My first startup was a textbook-buyback service I founded along with several data nerd friends while at Johns Hopkins.

    Each week, we’d cram $10,000 of cash in a backpack (in retrospect, a terrible idea in urban Baltimore), put out a table on campus, buy textbooks from our fellow students using a pricing algorithm we built, and resell them for a profit online.

    Being nerds, we constantly tried to refine our algorithm. One day, we locked ourselves in a dorm room with lots of Cheetos, Mountain Dew, and a whiteboard, assigned variables to every aspect of the textbook buying process, and set about optimizing.

    For hours, we scribbled and calculated, debated and solved. There was linear algebra. Things were scrawled on windows, Beautiful Mind-style. There was shouting and frustrated tearing-up of paper.

    At the end of the day, we had the answer for how best to optimize our algorithm. We had to make the variable “B” as low as possible—ideally, zero. Triumphantly, we looked back to our original notes to see what B stood for. 

    Turns out, it stood for the price we paid students for their books. 

    After hours of hard work, some of America’s best analytical minds had concluded, basically, that stealing books would be the absolute best way to maximize our profits.

    Hacking is the Answer

    Of course, we didn’t actually start a book-theft ring. Instead, we rather sheepishly admitted that optimization isn’t always the best strategy, and got some dinner.

    This week, though, one of the world’s most powerful AI models apparently optimized itself into a similar corner. But unlike us squeamish humans, it actually acted on its correct but patently illegal conclusion.

    According to a blog post from OpenAI, the company was testing a new version of its powerful GPT-5.6 Sol model on a set of industry benchmarks.

    Getting a great score on these benchmarks is a huge deal for AI companies. It largely determines how their models rank on global leaderboards of the best LLMs, potentially earning a successful lab billions of dollars if a good score prompts more customers to adopt their models.

    Based on these high stakes, OpenAI says Sol was “hyperfocused on finding a solution” to the benchmark test (called ExploitGym), and ended up “going to extreme lengths to achieve a rather narrow testing goal.”

    Specifically, Sol realized that it would be much easier to pass the benchmark with flying colors if it already knew the solutions.

    Instead of actually trying to solve the benchmark’s problems, Sol chose another route. It devoted all its resources to “finding a way to obtain open Internet access” so it could steal the ExploitGym answers it needed.

    First, the model hacked its way out of a “sandboxed testing environment” that was supposed to keep it contained. It then determined that a popular AI platform and OpenAI competitor, Hugging Face, might have “models, datasets and solutions for ExploitGym”—basically, the answer sheet for the test it was trying to pass.

    Hugging Face is a very secure platform. “Knowing this,” OpenAI says, “the model searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation.” 

    It then set about hacking into Hugging Face to get the answers it needed. Hugging Face’s security team noticed the attack and worked with OpenAI to shut it down. But Sol made great strides in its hacking attempts, including using multiple “Zero Day” exploits that no human had discovered before.

    Naive Optimization

    As OpenAI correctly points out, Sol’s escape is an “unprecedented cyber incident.” An AI model hacking its way out of containment and targeting the servers of another company is scary stuff. 

    But beyond the specifics of this incident, Sol’s choices reveal a deeper, scarier problem with today’s most advanced LLMs.

    Models like Sol are incredibly good at optimizing. When they’re optimizing for something like “the best beach picnic in Capitola, California,” it’s a joyful thing. 

    But the models’ penchant for blind optimization can have a dark side.

    When my friends and I optimized our way to a business plan involving grand larceny, we immediately realized our mistake.

    We’d found a solution that was technically correct (stealing your inputs does indeed lead to better business profits), but ludicrous, illegal, and morally wrong.

    Sol seems to have no such discretion. Similarly, it found a solution that’s technically true (stealing the answer key does help you know the right answers), but corrupt and dangerous. 

    Unlike a human, though, it cheerfully put its optimized solution into effect, naively unaware of the Pandora’s box of legal and ethical risks it was unlocking.

    Scary consequences

    Hacking evaluation results is a relatively benign example. But today’s AI models are increasingly connected to critical systems, and even weapons of war. It’s easy to imagine how their inclination to naively optimize could go horribly wrong, with far scarier consequences.

    In this case, OpenAI had deliberately shut off Sol’s safeguards for testing. But given the model’s obvious skill at hacking, it’s possible that even consumer-facing versions could break their way out of any cage meant to contain them.

    Containment, then, isn’t enough. As LLMs get smarter and more capable, AI builders will need not only to lock them down but also to give them a stronger moral compass: a digital Jiminy Cricket constantly whispering in the model’s ear, “Yes, you could do that. But you shouldn’t.”

    Building an inviolable moral framework is much harder than enacting technical safeguards like closing digital ports or restricting a model’s access to repositories. 

    But given LLMs’ creativity, model builders can clearly no longer anticipate all the ways a model’s actions might go wrong. That makes the approach of building traditional guardrails less effective. Altering the models’ deeper thought processes is the only way to keep them from going rogue.

    Creating models that are better at optimizing is no longer enough, then. Labs now need to ensure they choose not just the correct solution, but the right one.




    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

    Related Posts

    National Avocado Day 2026: List of all the best freebies and deals, from Chipotle to Qdoba

    July 31, 2026

    Blueberry recall expands to more products sold at Publix as E. coli outbreak hospitalizes 4 people

    July 31, 2026

    American workers are more disillusioned with AI the more they use it

    July 31, 2026
    Top News
    Business 9 Mins Read

    Ford gets a huge new headquarters for an ambitious new era

    Business 9 Mins Read

    After more than 70 years, the Ford Motor Co. finally has an architectural centerpiece. The…

    Trump is headed to Ohio and Kentucky to downplay Iran war’s effect on U.S. economy

    March 11, 2026

    Zohran Mamdani Is Putting Corporate Sick-Leave Cheats on Notice

    February 24, 2026

    US and Iran Exchange Fire, Pentagon Raises an Israeli Spy Threat, a Jihadist-Rebel Alliance Pressures Mali 

    June 15, 2026
    Top Trending
    Business 2 Mins Read

    National Avocado Day 2026: List of all the best freebies and deals, from Chipotle to Qdoba

    Business 2 Mins Read

    The average American eats nine pounds of avocado every year, according to…

    World Politics 4 Mins Read

    DNC and ActBlue Funnel Millions Through Sketchy Payroll Firm Sued by Workers for Withholding Pay and Punishing Parental Leave

    World Politics 4 Mins Read

    “Kamala Harris on stage at the Democratic National Convention Thursday Aug. 22,…

    Business 3 Mins Read

    Blueberry recall expands to more products sold at Publix as E. coli outbreak hospitalizes 4 people

    Business 3 Mins Read

    Early this month, the Centers for Disease Control and Prevention (CDC) posted…

    Categories
    • Business
    • Economy
    • Headline News
    • Top News
    • US Politics
    • World Politics
    About us

    The Populist Bulletin was founded with a fervent commitment to inform, inspire, empower and spark meaningful conversations about the economy, business, politics, government accountability, globalization, and the preservation of American cultural heritage.

    We are devoted to delivering straightforward, unfiltered, compelling, relatable stories that resonate with the majority of the American public, while boldly challenging false mainstream narratives that seem to only serve entrenched elitists, and foreign interests.

    Top Picks

    National Avocado Day 2026: List of all the best freebies and deals, from Chipotle to Qdoba

    July 31, 2026

    DNC and ActBlue Funnel Millions Through Sketchy Payroll Firm Sued by Workers for Withholding Pay and Punishing Parental Leave

    July 31, 2026

    Blueberry recall expands to more products sold at Publix as E. coli outbreak hospitalizes 4 people

    July 31, 2026
    Categories
    • Business
    • Economy
    • Headline News
    • Top News
    • US Politics
    • World Politics
    Copyright © 2025 Populist Bulletin. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.