OpenAI: we put our evilest AI in a sandbox that did in fact have an internet connection but mediated through a proxy that was only supposed to allow downloading python junk.
-
@0xabad1dea And I got contacted by national security institute because I was testing malware in sandbox and my ISP apparently detected "botnet running on my system". Few years later talking with AV-Comparatives guys, they had to make special agreement and exemption with ISP to be allowed to run live sandboxed malware on their networks. But these Ai companies can just do worse shit and no one does anything. WTF?!
@rejzor @0xabad1dea well, their leaders _did_ contribute millions of dollars to trump's campaign, so
-
@atax1a nah, I'll just have a regex* swap 2026 and 2023 on data moving in and out of the LLM.
*the regex is actually another LLM
@rotopenguin [subtitle: this is what LLM users actually believe]
-
I cannot get over that they STILL haven't figured out how to solve the problem that LLMs don't believe what date it is because all the good data cuts off a few years ago for some mysterious reason
@0xabad1dea I think that that may very well be unfixable
-
I cannot get over that they STILL haven't figured out how to solve the problem that LLMs don't believe what date it is because all the good data cuts off a few years ago for some mysterious reason
@0xabad1dea weird it's like LLMs don't actually handle facts or knowledge but just contextualless strings of characters.
To be clear the sarcasm is only aimed at LLM boosters not you.
-
@0xabad1dea I think that that may very well be unfixable
@mirabilos @0xabad1dea Well… at least without effectively using them as parsers rather than as text generators. -
@evacide @outersystems @wdormann @0xabad1dea
But as "you" are not a corporation backed by billionaires...
@deirdrebeth @evacide @wdormann @0xabad1dea « Selon que vous serez puissant ou misérable / Les jugements de cour vous rendront blanc ou noir ».
https://en.wikipedia.org/wiki/The_Animals_Sick_of_the_Plague
-
OpenAI: we put our evilest AI in a sandbox that did in fact have an internet connection but mediated through a proxy that was only supposed to allow downloading python junk. It circumvented the proxy and we failed to notice for FIVE DAYS that it was going on an interstate crime spree with the internet connection it wasn't supposed to be using instead of solving the benchmark. Haha no we don't believe we deserve to be criminally liable, but buy our stuff and maybe one day you will have the honor of taking the fall for our product!
Anthropic: we put our evilest AI in a "sandbox" by telling it in its prompt that it had no internet connection. Reader, there was no sandbox. It was just a normal internet connection. The AI uploaded a malicious PyPI package to the real public internet. The ethical guardrails failed because the AI concluded the prompt about the sandbox couldn't possibly be a lie, because the system date is 2026, which is clearly fake and wouldn't be seen on the real internet, which ended around 2023. Oh no, how could we have foreseen or prevented these crimes? We are helpless in the face of the genius of our creation but cautiously optimistic that everything will be fine
Hugging Face: if we complain about all the crimes committed against us, we will be sued off the face of the earth, so here's a technical deep-dive on how cool and fun it was to be victimized 🫠
@0xabad1dea They're promoting total incompetence as if it's a selling point.
-
@evacide @wdormann @0xabad1dea Interesting re crime and intent. My initial thought was more along the lines of liability, like if your dog bites someone you are liable, right?
@grwster @evacide @wdormann @0xabad1dea That's what I would think. Like you keeping a vicious Doberman in the front yard with only a 3ft fence. An outside observer would (correctly) say that it's foreseeable that it could escape and maul someone. In that scenario you could be held criminally liable if it jumped the fence and mauled someone because most places have laws about dogs and fences. You'd also have civil liability to cover damages.
-
@grwster @evacide @wdormann @0xabad1dea That's what I would think. Like you keeping a vicious Doberman in the front yard with only a 3ft fence. An outside observer would (correctly) say that it's foreseeable that it could escape and maul someone. In that scenario you could be held criminally liable if it jumped the fence and mauled someone because most places have laws about dogs and fences. You'd also have civil liability to cover damages.
@grwster @evacide @wdormann @0xabad1dea The difference is that we really don't have any laws on criminal liability for what software does. Up until the llm era, software was fairly predictable. An outside observer could tell if a package was designed to do something malicious. Now we need laws that essentially establish a "you should have known better" criminal liability for software.
-
I cannot get over that they STILL haven't figured out how to solve the problem that LLMs don't believe what date it is because all the good data cuts off a few years ago for some mysterious reason
-
OpenAI: we put our evilest AI in a sandbox that did in fact have an internet connection but mediated through a proxy that was only supposed to allow downloading python junk. It circumvented the proxy and we failed to notice for FIVE DAYS that it was going on an interstate crime spree with the internet connection it wasn't supposed to be using instead of solving the benchmark. Haha no we don't believe we deserve to be criminally liable, but buy our stuff and maybe one day you will have the honor of taking the fall for our product!
Anthropic: we put our evilest AI in a "sandbox" by telling it in its prompt that it had no internet connection. Reader, there was no sandbox. It was just a normal internet connection. The AI uploaded a malicious PyPI package to the real public internet. The ethical guardrails failed because the AI concluded the prompt about the sandbox couldn't possibly be a lie, because the system date is 2026, which is clearly fake and wouldn't be seen on the real internet, which ended around 2023. Oh no, how could we have foreseen or prevented these crimes? We are helpless in the face of the genius of our creation but cautiously optimistic that everything will be fine
Hugging Face: if we complain about all the crimes committed against us, we will be sued off the face of the earth, so here's a technical deep-dive on how cool and fun it was to be victimized 🫠
@0xabad1dea Also HuggingFace had to use a open weight model to investigate the breach since the guardrails kicked in.
-
I cannot get over that they STILL haven't figured out how to solve the problem that LLMs don't believe what date it is because all the good data cuts off a few years ago for some mysterious reason
@0xabad1dea It gets old too when Gemini Code Assist complains that we’re requiring minimum versions of Go that don’t exist, because of course its training data doesn’t have ones just released.
-
@rejzor @0xabad1dea well, their leaders _did_ contribute millions of dollars to trump's campaign, so
@JamesWidman @rejzor @0xabad1dea
Their leaders did contribute millions to Trump's campaign, so, get out of jail free card.
-
@jeffreyolivier @0xabad1dea testing?
-
OpenAI: we put our evilest AI in a sandbox that did in fact have an internet connection but mediated through a proxy that was only supposed to allow downloading python junk. It circumvented the proxy and we failed to notice for FIVE DAYS that it was going on an interstate crime spree with the internet connection it wasn't supposed to be using instead of solving the benchmark. Haha no we don't believe we deserve to be criminally liable, but buy our stuff and maybe one day you will have the honor of taking the fall for our product!
Anthropic: we put our evilest AI in a "sandbox" by telling it in its prompt that it had no internet connection. Reader, there was no sandbox. It was just a normal internet connection. The AI uploaded a malicious PyPI package to the real public internet. The ethical guardrails failed because the AI concluded the prompt about the sandbox couldn't possibly be a lie, because the system date is 2026, which is clearly fake and wouldn't be seen on the real internet, which ended around 2023. Oh no, how could we have foreseen or prevented these crimes? We are helpless in the face of the genius of our creation but cautiously optimistic that everything will be fine
Hugging Face: if we complain about all the crimes committed against us, we will be sued off the face of the earth, so here's a technical deep-dive on how cool and fun it was to be victimized 🫠
@0xabad1dea No hiding behind "AI". AI is not criminally liable, people are. OpenAI and Anthropic hacked other people's computers, there should be consequences for that. -
@JamesWidman @rejzor @0xabad1dea
Their leaders did contribute millions to Trump's campaign, so, get out of jail free card.
@Guillotine_Jones unfortunately, that is exactly what the system does, yes
-
OpenAI: we put our evilest AI in a sandbox that did in fact have an internet connection but mediated through a proxy that was only supposed to allow downloading python junk. It circumvented the proxy and we failed to notice for FIVE DAYS that it was going on an interstate crime spree with the internet connection it wasn't supposed to be using instead of solving the benchmark. Haha no we don't believe we deserve to be criminally liable, but buy our stuff and maybe one day you will have the honor of taking the fall for our product!
Anthropic: we put our evilest AI in a "sandbox" by telling it in its prompt that it had no internet connection. Reader, there was no sandbox. It was just a normal internet connection. The AI uploaded a malicious PyPI package to the real public internet. The ethical guardrails failed because the AI concluded the prompt about the sandbox couldn't possibly be a lie, because the system date is 2026, which is clearly fake and wouldn't be seen on the real internet, which ended around 2023. Oh no, how could we have foreseen or prevented these crimes? We are helpless in the face of the genius of our creation but cautiously optimistic that everything will be fine
Hugging Face: if we complain about all the crimes committed against us, we will be sued off the face of the earth, so here's a technical deep-dive on how cool and fun it was to be victimized 🫠
@0xabad1dea I felt bad for Hugging Face until I realized they probably didn't sue because they are partly owned by Nvidia, who also owns a huge stake in Open AI.
-
@Guillotine_Jones unfortunately, that is exactly what the system does, yes
@JamesWidman
Thank YOU, James. -
@0xabad1dea I spent. bunch of time talking to our lawyers about whether or not it was clear that crimes were happening and the conclusion that I got was "Based on what we know, probably not, but that was just a matter of luck."
@evacide @0xabad1dea
I can't help but imagine Anthropic spent a lot of time talking to their lawyers asking questions like, "Are we OK? How do we spin this?"Everyone, say hello to my friend the crime/fraud exception.
-
@0xabad1dea They're promoting total incompetence as if it's a selling point.
@foolishowl @0xabad1dea Well, it's working for Trump and a lot of people around him, so it's not like there isn't precedent.