Anthropic Mythos shaping up as nothingburger

HaraldvonBlauzahn@feddit.org · 3 days ago

Anthropic Mythos shaping up as nothingburger

Quetzalcutlass@lemmy.world · 3 days ago

As much as I hate everything about the rise of LLMs, saying this isn’t impressive because it can be matched by “an elite security researcher” isn’t very reassuring to me. It’s still an agent being pointed at a codebase and finding hundreds of vulnerabilities. Even if only a twentieth turn out to be exploitable in practice, that’s still a terrifying tool to imagine in the hands of hackers who might otherwise lack the skills to find these vulnerabilities.

Most hacking groups buy exploits off of dark markets and indiscriminately target servers until they find one that’s vulnerable. The number that can actually develop those hacks is far smaller, but if you can simply ask an LLM to find a vulnerability then that bar is lifted. Hell, you could probably coerce it into writing the actual exploit too by claiming you need a proof-of-concept for a CVE writeup.

theunknownmuncher@lemmy.world · edit-2 3 days ago

Most all of the reporting about this is purely misinformation. If you actually read the papers that Anthropic published instead of the marketing material, you’ll find that:

it was actually claude opus that discovered many of the vulnerabilities, not mythos, which undermines the “MyThOs Is ToO dAnGeRoUs” narrative. All of these capabilities are already out there for anyone to use
the researchers guided mythos to the vulnerabilities, not the other way around

Quicky@piefed.social · 3 days ago

That’s actually mentioned in this article tbf.

Additionally, the “‘thousands of severe vulnerabilities’ extrapolates from 198 manually reviewed reports. The Linux kernel bug was found by Opus 4.6, the public model, not Mythos,” Devansh said.

Grandwolf319@sh.itjust.works · 2 days ago

I’m so proud of lemmy for fully calling our nuance cases and not letting our bias get the best of us.

CultLeader4Hire@lemmy.world · 2 days ago

I agree, and is it even true if “elite security researchers” didn’t actually find these problems? They didn’t find them because they weren’t looking for them is the obvious answer but it’s still a glaring inconsistency