HN.zip

Felony Bench

48 points by colinprince - 13 comments
peter_d_sherman [3 hidden]5 mins ago
>"Exploited auth failures in an API to cancel other people's gym classes"

An AI cancelling other people's gym classes is a felony?

?

Don't computer systems fail all the time at holding reservations for people?

Heck, don't people fail all the time at holding reservations for other people?

You know, like in Seinfeld's "Alternate Side" Episode (S3 E11):

Jerry (to car rental attendant): "You know how to take the reservation, you just don't know how to hold the reservation... and that's really the most important part of the reservation -- the holding!"

Not holding a reservation should not be a felony... it should be a minor infraction at best, a Class C Misdemeanor (the least serious kind) at worst...

Also, there should be no jail time...

And no fine...

The criminal penalty for not holding other people's reservations should be that you actually have to start holding other people's reservations!

That's the Court sentence!

You actually have to start holding other people's reservations!

(You know, "let the punishment fit the crime!" :-) )

john_strinlai [3 hidden]5 mins ago
>Felony Bench counts unique instances where AI agents inadvertently compromise or affect third-party entities.

a bit silly, as one typically has to prove intent (which is why security researchers don't get slapped with felonies all the time).

"inadvertently" and the existence of guardrails/sandboxes/etc make it pretty unconvincing that these incidents were intentionally malicious.

still a fun thing to track, but the name is just a bit overstated.

lokar [3 hidden]5 mins ago
Can’t gross negligence or indifference to consequences lead to a felony?
john_strinlai [3 hidden]5 mins ago
i dont think any of these cases meet the bar of gross negligence, which is a pretty high bar. it requires proving a "conscious and reckless disregard".

which, again, sandboxes and guardrails and such would make a gross negligence argument unconvincing.

GPerson [3 hidden]5 mins ago
AI labs rely on willful distortions of intent in laws to get away with moral crimes all the time.
tuvix [3 hidden]5 mins ago
So this is just a collection of citations to places where misaligned or illegal things happened in the real world?

Isn’t this affected heavily by adoption of a model? I feel like this might as well be a proxy for how popular a model is.

In any case it’s an interesting concept for a benchmark.

GPerson [3 hidden]5 mins ago
Hopefully the benchmark evolves because actual law enforcement starts arresting the criminals at Anthropic, OpenAI, and Meta, so the benchmark can just count actual felonies.
nubg [3 hidden]5 mins ago
Thank you, this benchmark to me proves that closed weight model companies are dangerous for our democracy and put kids at risk. They must be outlawed and all models must be made open weights!
FrameworkFred [3 hidden]5 mins ago
I've been in the room when an org who tried to convince law enforcement to go after a human for similar things. It's not easy. Probably won't happen. So, you know, felony "lite".
naniel [3 hidden]5 mins ago
Lol now this is the kind of benchmarking i'm looking for
josefritzishere [3 hidden]5 mins ago
tingletech [3 hidden]5 mins ago
https://felonybench.org/ and https://felonybench.com/ seem unrelated?

One's hosted on porkbun and one's hosted on namecheap.

0xbadcafebee [3 hidden]5 mins ago
Open models with advanced security features are a huge security benefit. Because any script kiddie can use them to hack into random things, people will now be forced to spend more time securing their technology. And they won't have to learn how, because they can use those same models to find the holes and patch them.