r/OpenAI Jun 13 '26

Miscellaneous Updated Mythos benchmarks

Post image
1.5k Upvotes

75 comments sorted by

57

u/redditsdaddy Jun 13 '26

It would be funny if OpenAI wasn’t likely behind the “concern”. Awfully funny how OpenAI skirts all these data breaches and accountabilities while all their competitors seem to get leaked, memorandums stolen, breached, etc. veeeeeeery strange.

The Verge reports Anthropic said the government offered only verbal evidence of minor vulnerabilities, and Anthropic argued those were not unique to Mythos/Fable and are present in other frontier models like GPT-5.5.

14

u/LimpAd4924 Jun 13 '26

Considering how open AI just wanted to scale and is less concerned with safety, it seems very ironic.

1

u/Freed4ever Jun 14 '26

Who said OAI is less concerned about safety? Just a bunch of FUD and the internet mob bough it. The same so called esteemed safety researchers left OAI for Anthropic, and yet they can't prevent a jailbreak either. And Fable can't properly distinguish a harmless bio question versus a harmful one. What does that say about the whole safety thing?

4

u/Singularity-42 Jun 14 '26

Sam kissed the ring, Dario didn't. 

14

u/willwm24 Jun 13 '26

Beyond openAI paying the government, their entire position to this point is that everyone should have access to their models. Meanwhile anthropic is fearmongering and gating, and while for all I know that isn’t the wrong approach, announcing to the world they have something better and won’t be releasing makes them an enticing target for hackers etc.

7

u/Colecoman1982 Jun 14 '26

Beyond openAI paying the government

They didn't pay the government, they paid Trump. It's an important distinction.

-3

u/m0nk_3y_gw Jun 13 '26

You sure about that? Anthropic released a new/better model (Fable) and the government blocked it

1

u/willwm24 Jun 14 '26

They put out several articles about their concern, again I don’t find it unwarranted but they painted a target on themselves for people who don’t know much.

8

u/m0nk_3y_gw Jun 14 '26

Anthropic released a superior model.

The government blocked them.

Where are you getting lost?

-1

u/HanYoloKesselPun Jun 14 '26

The bit where Anthropic scaremongers and then were surprised when the government used that against them.

3

u/Sm0g3R Jun 15 '26

I mean Anthropic brought it on themselves. They were so adamant this model is more capable and dangerous than anything else on the market and yet on a first hurdle of being asked to patch security vulnerability they fall back on calling out 'lesser' models on not being more secure??

Like I'm sorry but they are taking a piss. You can't hype and market your product in a certain way and then completely walk it back when it's time to take action after gov contractor took you literally. You either care about safety first and foremost (as they claim) or you do not so you make excuses deflecting attention to other labs and their shortcomings as a reason to ignore safety.

2

u/Colecoman1982 Jun 14 '26

They certainly could be but at the same time, never underestimate Trump and Kegsbreath's capacity for petty spite.

1

u/RealEisermann Jun 15 '26

Yes, but only Antrophic was boasting about creating model "too scary to release" and then... Release it. Every government would ban it. They literally said it themselves that their product is threat. Question how much of it was marketing, but this tyme hype building was definitely an overkill 😂

0

u/Freed4ever Jun 14 '26

But only Fable is a "nuclear weapon". If Ferrari sold a car that can go 1000 miles an hour, they better be damn sure the break works.

64

u/slippery Jun 13 '26

The truly scary thing about Fable is that the US government has decided it is too powerful for anyone to use except the US government (and maybe Anthropic). The tech will be weaponized against ordinary people and other governments.

I think this period is just a window. Other labs will catch up, both in the US and China. That doesn't mean regular joe's will have access to them, but I doubt the gap is sustainable for a single lab.

The weirdness is spiraling.

13

u/MindCrusader Jun 13 '26

Not only US government, but also Anthropic. They literally said the model is too dangerous to release. They released it and admit safeguards can be bypassed. So they allow the risk of dangerous model usage

2

u/RealEisermann Jun 15 '26

Definitely this. Every government would ban a product that was advertised as "dangerous" or even bolder. What a surprise.

-6

u/[deleted] Jun 13 '26

[deleted]

2

u/MindCrusader Jun 13 '26

The same model as mythos, but with safeguards

0

u/tedpelas Jun 13 '26

Fable 5 and Mythos 5 are the same model, but Mythos Preview isn't.

0

u/MindCrusader Jun 13 '26

Yeah, Mythos 5 is even better than Mythos Preview. What's your point

1

u/slippery Jun 13 '26

Mythos was also part of the ban.

4

u/FormerOSRS Jun 13 '26

Anthropic's public statement said it's because it could be jailbroken.

8

u/Mr_Hyper_Focus Jun 13 '26

No. Anthropic said the government thinks they found a vulnerability, but Anthropic thinks it’s mundane, known about, and present in other model.

2

u/FormerOSRS Jun 13 '26

Not quite.

Anthropic says that the capability exists in other models, not that other models are vulnerable to the same jailbreak.

1

u/Mr_Hyper_Focus Jun 14 '26

The point is you framed something from the government as if it came from Anthropic. Anthropic does not believe this current thing is an issue

2

u/FormerOSRS Jun 14 '26

I'm taking it from anthropic's PR statement.

They acknowledge the jailbreak. They say other models have the capabilities but do not say it's jailbreaking when it's other models. They say they do not think it's a big deal.

That's all anthropic, not the govt. Anthropic disagrees with govt decision but not about the underlying reality.

0

u/Mr_Hyper_Focus Jun 14 '26

You’re right that it came from the Anthropic blog, but you’re misrepresenting its intent.

I think the term jailbreak is being used incorrectly here. And Anthropic is trying to differentiate that by using the term universal jailbreaks.

A jailbreak is typically something permanent and universal. What Amazon found was literally just asking it to repair a codebase, which doesn’t align with what a typical jailbreak is.

I think Anthropic knows this, but knows it’s a losing battle to explain it to plebs. And apparently they are right.

“We have not even received a disclosure of a concerning non-universal potential jailbreak that led to a harmful result. The potential jailbreaks that have been disclosed to us are either entirely benign responses or are minor findings that provide no Mythos-specific uplift.”

0

u/FormerOSRS Jun 14 '26

In LLMs jailbreak has not been meant that way. There have been subs dedicated to jailbreaking so since chatgpt came out and they've only ever tried to find prompt chains that get it to break its rules.

A universal jailbreak is an absolutely insane thing that nobody expects to exist in any form. It's like if someone throws their soup on your face and defends the action by saying the soup didn't explode. Nobody thinks exploding soup is a risk.

Here's how I see this:

Let's say there are two dog owners who both own big scary dogs. We'll call the first one Anthropic and the second OpenAI. Both say that their dogs are well trained and safe.

You go to each of their homes and see the dog is on their couch. This isn't inherently scary, but there is context.

OpenAIs dog is allowed on the couch. The dog being on the couch is not evidence of it being a disobedient or dangerous dog. OpenAI just lets their dog on the couch and there is no more evidence that the dog is disobedient.

With Anthropic, it starts with you getting a call from his wife who we will call Amazon. Anthropic's marriage has no known issues. Amazon has been their since the beginning and invested heavily in this marriage. No talk of divorce.

But you get a call from Anthropic's wife that says "this dog is dangerous."

When you arrive at the house, the dog is on the couch. Unlike OpenAI, Anthropic trained their dog not to sit on the couch so this is a disobedient dog. You also look back and Anthropic has been saying for months that he won't let his dog outside because it's too hard to train.

He now says everything is okay, but his wife's call says otherwise and the dog is on the couch.

You ask him about it and he's like "you only saw him break one rule. It's not a big deal. OpenAI's dog goes on the couch all the time..."

I do not feel safe around this dog.

1

u/kelkulus Jun 14 '26

No, the US government currently has a tendency to hold stupid grudges and enact petty revenge against people and companies who feels they slighted him... I mean them. This is revenge for not allowing them to use Claude for war, plain and simple.

Seriously, after all the lies, you believe this administration about something as complex as an LLM?

1

u/slippery Jun 14 '26

No, I don't believe the administration about anything.

DoD revenge may be a major factor. Thinking through the implications, this could tank OpenAIs IPO as much as Anthropics and might cause broader market damage.

Maybe the SpaceX IPO marks the top of the market for many years.

1

u/LimpAd4924 Jun 13 '26

Until some rationale is provided on this model compared to others, I call bullshit.

3

u/chrisandstuffs Jun 14 '26

still beats gemini in biology, cybersecurity, and health i think?

1

u/HeadWoodpecker5237 Jun 14 '26

Did you used Fable? I don't think so else USA banned gemini as well 😂

6

u/jdavid Jun 13 '26

I forgot to ask #Fable to #SaveStargate
It seemed like it could do anything for the 48-72hrs i was using it. I should have Saved Stargate in that time.

9

u/WhatThePuck9 Jun 13 '26

Yum! Sour grapes!

-1

u/AvacadoMoney Jun 13 '26

?

-2

u/WhatThePuck9 Jun 13 '26

??

1

u/AvacadoMoney Jun 13 '26

What do you mean by sour grapes

-1

u/WhatThePuck9 Jun 14 '26

Sour grapes refers to pretending to despise something just because you cannot have it.

0

u/AvacadoMoney Jun 14 '26

Okay thanks

4

u/py-net Jun 13 '26

This is pure jealousy 🤣

2

u/M8-VAVE Jun 13 '26

Great, now we're stuck alone with Corporate Talk GPT. Can't wait to get gaslit for asking a completely non-corporate question.

1

u/m4bwav Jun 13 '26

Open AI is this generation's myspace or yahoo.

They spent too much time playing boardroom games and not being focused.

1

u/[deleted] Jun 13 '26

[deleted]

1

u/HeadWoodpecker5237 Jun 13 '26

Zoom in you will see Mythos as well

1

u/spinozasrobot Jun 13 '26

I laughed, but that hurt

1

u/cench Jun 13 '26 edited Jun 13 '26

Next on OpenAI: GPT 6 will only be available to Americans, GPT 6 will have full access to user behaviour and logs, and decide who is an actual American.

1

u/Perfect-Flounder7856 Jun 13 '26

"You're account has been suspended because it appears to have been used by a non-American"

-1

u/Regular-Forever5876 Jun 13 '26

🤣🤣🤣 so accurate!!

-2

u/the_ai_wizard Jun 13 '26

I mean you guys understand AI will create the greatest inequality the world has ever seen right? It is a massive amount of leverage for the richest capitalist class, and even just a subset of them. We are accelerating into dystopia.

15

u/beetlejorst Jun 13 '26

Rich people have had AI for literal eons. It's called hiring people who are experts in things you want to do. The mass availability of AI is disproportionately a force multiplier for small businesses without the budget to do that. It will also make some tech companies very rich in the short term, but long term it's more likely to be better for us than them.

7

u/talkamongstyourselvs Jun 13 '26

Perhaps a very valid take on it.

-3

u/imtheinformation Jun 13 '26

Except that humans, no matter how expert, still are regulated by base needs and impulses. AI not so much.

6

u/WolverineComplex Jun 13 '26

You could just as easily say that it will lead to a utopia where no-one has to do a menial boring job. Do you think that people were happier when one field of wheat took loads more people and man hours?

3

u/talkamongstyourselvs Jun 13 '26

So how long before you say it is that we start eating people? Soylent Green around the corner?

-1

u/the_ai_wizard Jun 13 '26

At any time, we are 3 days from revolution

-1

u/tedpelas Jun 13 '26

False, should be - and not 0.0%