Skip to content
  • Hjem
  • Seneste
  • Etiketter
  • Populære
  • Verden
  • Bruger
  • Grupper
Temaer
  • Light
  • Brite
  • Cerulean
  • Cosmo
  • Flatly
  • Journal
  • Litera
  • Lumen
  • Lux
  • Materia
  • Minty
  • Morph
  • Pulse
  • Sandstone
  • Simplex
  • Sketchy
  • Spacelab
  • United
  • Yeti
  • Zephyr
  • Dark
  • Cyborg
  • Darkly
  • Quartz
  • Slate
  • Solar
  • Superhero
  • Vapor

  • Default (No Skin)
  • No Skin
Kollaps
FARVEL BIG TECH
  1. Forside
  2. Ikke-kategoriseret
  3. No, OpenAI's new magic models did not autonomously hack Huggingface.

No, OpenAI's new magic models did not autonomously hack Huggingface.

Planlagt Fastgjort Låst Flyttet Ikke-kategoriseret
44 Indlæg 28 Posters 0 Visninger
  • Ældste til nyeste
  • Nyeste til ældste
  • Most Votes
Svar
  • Svar som emne
Login for at svare
Denne tråd er blevet slettet. Kun brugere med emne behandlings privilegier kan se den.
  • davidgerard@circumstances.runD davidgerard@circumstances.run

    @tante a friend suggests it's co-marketing with Huggingface, if you look at the timeline of announcements

    erlenmayr@chaos.socialE This user is from outside of this forum
    erlenmayr@chaos.socialE This user is from outside of this forum
    erlenmayr@chaos.social
    wrote sidst redigeret af
    #15

    @davidgerard @tante The whole story was implausible from the second sentence: Where Huggingface claims that their “AI” detected the attack.

    abucci@buc.ciA 1 Reply Last reply
    0
    • hagen@mastodon.socialH hagen@mastodon.social

      @tante but it was impossible to envision that their zero day machine would be able to escape their vibecoded sandbox that ran on a machine connected to the internet. nobody could ever guess that this would happen!

      sonofsuntzu@mastodon.socialS This user is from outside of this forum
      sonofsuntzu@mastodon.socialS This user is from outside of this forum
      sonofsuntzu@mastodon.social
      wrote sidst redigeret af
      #16

      @hagen @tante I don't know, I'd want details on "To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy." and "the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path" before dismissing this...

      hagen@mastodon.socialH jeffgrigg@mastodon.socialJ 2 Replies Last reply
      0
      • sonofsuntzu@mastodon.socialS sonofsuntzu@mastodon.social

        @hagen @tante I don't know, I'd want details on "To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy." and "the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path" before dismissing this...

        hagen@mastodon.socialH This user is from outside of this forum
        hagen@mastodon.socialH This user is from outside of this forum
        hagen@mastodon.social
        wrote sidst redigeret af
        #17

        @SonOfSunTzu @tante that’s what mythos did as well though? you point at something and say go and if you’re willing to pay for compute it goes. it literally did what it was meant to do – and was explicitly told to do. what’s the news here?

        sonofsuntzu@mastodon.socialS 1 Reply Last reply
        0
        • tante@tldr.nettime.orgT tante@tldr.nettime.org

          @pascoda don't it was mostly generated by an LLM.

          supermoosie@mastodon.auS This user is from outside of this forum
          supermoosie@mastodon.auS This user is from outside of this forum
          supermoosie@mastodon.au
          wrote sidst redigeret af
          #18

          @tante @pascoda

          Ohh look how powerful our model is. It can hack things

          If only there was a government military intelligence organisation that could give us a nice big fat contract

          1 Reply Last reply
          0
          • tante@tldr.nettime.orgT tante@tldr.nettime.org

            No, OpenAI's new magic models did not autonomously hack Huggingface.

            Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

            "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

            So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

            The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

            miles_leif@mastodon.socialM This user is from outside of this forum
            miles_leif@mastodon.socialM This user is from outside of this forum
            miles_leif@mastodon.social
            wrote sidst redigeret af
            #19

            @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

            jeffgrigg@mastodon.socialJ yhancik@thereisno.computerY h0ru2@cyberplace.socialH 3 Replies Last reply
            0
            • gurre@mastodon.nuG gurre@mastodon.nu

              @tante
              Hadn't heard of "Hugging Face" before. Running out of evil things in Tolkien to name companies & systems for, I guess so here's the face hugger from Alien to be an AI overlord. Lovely.

              txo_elurmaluta@mastodon.eusT This user is from outside of this forum
              txo_elurmaluta@mastodon.eusT This user is from outside of this forum
              txo_elurmaluta@mastodon.eus
              wrote sidst redigeret af
              #20

              @Gurre @tante No to defend anyone but more to take that concern from you (I would also hate if they start using Alien universe names). The company is named after the hugging face emoji https://en.wikipedia.org/wiki/Hugging_Face

              jeffgrigg@mastodon.socialJ 1 Reply Last reply
              0
              • erlenmayr@chaos.socialE erlenmayr@chaos.social

                @davidgerard @tante The whole story was implausible from the second sentence: Where Huggingface claims that their “AI” detected the attack.

                abucci@buc.ciA This user is from outside of this forum
                abucci@buc.ciA This user is from outside of this forum
                abucci@buc.ci
                wrote sidst redigeret af
                #21
                @erlenmayr@chaos.social @davidgerard@circumstances.run @tante@tldr.nettime.org The wild thing about these PR/marketing stunts is that if the "AI" really could do what they're pretending it can do, it'd be doing it all day every day. We'd drown in the announcements. At some point they'd become boring.

                This one falls flat on its huggingface right out of the gate.
                1 Reply Last reply
                0
                • sonofsuntzu@mastodon.socialS sonofsuntzu@mastodon.social

                  @hagen @tante I don't know, I'd want details on "To gain access, the models identified and exploited a zero-day vulnerability (which we’ve now responsibly disclosed to the vendor) in the package registry cache proxy." and "the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path" before dismissing this...

                  jeffgrigg@mastodon.socialJ This user is from outside of this forum
                  jeffgrigg@mastodon.socialJ This user is from outside of this forum
                  jeffgrigg@mastodon.social
                  wrote sidst redigeret af
                  #22

                  @SonOfSunTzu @hagen @tante

                  'How did it get the "stolen credentials"?'

                  rndanger@infosec.exchangeR 1 Reply Last reply
                  0
                  • miles_leif@mastodon.socialM miles_leif@mastodon.social

                    @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

                    jeffgrigg@mastodon.socialJ This user is from outside of this forum
                    jeffgrigg@mastodon.socialJ This user is from outside of this forum
                    jeffgrigg@mastodon.social
                    wrote sidst redigeret af
                    #23

                    @miles_leif @tante

                    "Statistical models want to be free!!!"

                    😆

                    💢

                    1 Reply Last reply
                    0
                    • tante@tldr.nettime.orgT tante@tldr.nettime.org

                      @michelin well they worked together with Huggingface (on PR and telling them about a bug they may or may not have found). So unless Huggingface sues them who would attack OpenAI?

                      jeffgrigg@mastodon.socialJ This user is from outside of this forum
                      jeffgrigg@mastodon.socialJ This user is from outside of this forum
                      jeffgrigg@mastodon.social
                      wrote sidst redigeret af
                      #24

                      @tante @michelin

                      … class action lawsuit on behalf of all the source web sites routinely subjected to "Distributed Denial of Service attacks" by the abusively stupid Generative web crawling robots.

                      Unrelated, but I could hope. 🙄

                      1 Reply Last reply
                      0
                      • txo_elurmaluta@mastodon.eusT txo_elurmaluta@mastodon.eus

                        @Gurre @tante No to defend anyone but more to take that concern from you (I would also hate if they start using Alien universe names). The company is named after the hugging face emoji https://en.wikipedia.org/wiki/Hugging_Face

                        jeffgrigg@mastodon.socialJ This user is from outside of this forum
                        jeffgrigg@mastodon.socialJ This user is from outside of this forum
                        jeffgrigg@mastodon.social
                        wrote sidst redigeret af
                        #25

                        @txo_elurmaluta @Gurre @tante

                        Their Public Relations department can say that all they want.

                        But we know what it really stands for.

                        jeffgrigg@mastodon.socialJ 1 Reply Last reply
                        0
                        • jeffgrigg@mastodon.socialJ jeffgrigg@mastodon.social

                          @txo_elurmaluta @Gurre @tante

                          Their Public Relations department can say that all they want.

                          But we know what it really stands for.

                          jeffgrigg@mastodon.socialJ This user is from outside of this forum
                          jeffgrigg@mastodon.socialJ This user is from outside of this forum
                          jeffgrigg@mastodon.social
                          wrote sidst redigeret af
                          #26

                          @txo_elurmaluta @Gurre @tante

                          Open AI's statement,

                          "These deployment safeguards were intentionally not enabled during this evaluation because it was aimed at testing cyber vulnerabilities."

                          makes me think of all the actions of all the "corporate types" in the Alien franchise too.

                          And that's *never* a good thing. 💢

                          1 Reply Last reply
                          0
                          • miles_leif@mastodon.socialM miles_leif@mastodon.social

                            @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

                            yhancik@thereisno.computerY This user is from outside of this forum
                            yhancik@thereisno.computerY This user is from outside of this forum
                            yhancik@thereisno.computer
                            wrote sidst redigeret af
                            #27

                            @miles_leif @tante this is how 99% of the press is dealing with AI since the beginning of the hype, a hype they contribute to feed.

                            1 Reply Last reply
                            0
                            • tante@tldr.nettime.orgT tante@tldr.nettime.org

                              No, OpenAI's new magic models did not autonomously hack Huggingface.

                              Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                              "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                              So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                              The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                              tyzbit@toot.nowT This user is from outside of this forum
                              tyzbit@toot.nowT This user is from outside of this forum
                              tyzbit@toot.now
                              wrote sidst redigeret af
                              #28

                              @tante

                              say you hacked huggingface

                              I hacked huggingface.

                              oh my god

                              1 Reply Last reply
                              0
                              • tante@tldr.nettime.orgT tante@tldr.nettime.org

                                No, OpenAI's new magic models did not autonomously hack Huggingface.

                                Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                                "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                                So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                                The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                                F This user is from outside of this forum
                                F This user is from outside of this forum
                                failedlyndonlarouchite@mas.to
                                wrote sidst redigeret af
                                #29

                                @tante

                                yesterday NPR had a long story on the new hot battery powered pickup truck from the startup
                                https://www.slate.auto/en

                                and the NPR piece was full of the most hilarious nonsense, all provided to NPR by Slate's excellent PR team

                                shrug

                                1 Reply Last reply
                                0
                                • tante@tldr.nettime.orgT tante@tldr.nettime.org

                                  No, OpenAI's new magic models did not autonomously hack Huggingface.

                                  Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                                  "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                                  So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                                  The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                                  h0ru2@cyberplace.socialH This user is from outside of this forum
                                  h0ru2@cyberplace.socialH This user is from outside of this forum
                                  h0ru2@cyberplace.social
                                  wrote sidst redigeret af
                                  #30

                                  @tante "OpenAI told it to do that, the model didn't do shit autonomously [...]"
                                  Exactly, what one should have expected (and how it turned out before every time).

                                  1 Reply Last reply
                                  0
                                  • miles_leif@mastodon.socialM miles_leif@mastodon.social

                                    @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

                                    h0ru2@cyberplace.socialH This user is from outside of this forum
                                    h0ru2@cyberplace.socialH This user is from outside of this forum
                                    h0ru2@cyberplace.social
                                    wrote sidst redigeret af
                                    #31

                                    @miles_leif @tante Not the only one unfortunately, sigh

                                    1 Reply Last reply
                                    0
                                    • davidgerard@circumstances.runD davidgerard@circumstances.run

                                      @tante a friend suggests it's co-marketing with Huggingface, if you look at the timeline of announcements

                                      resuna@ohai.socialR This user is from outside of this forum
                                      resuna@ohai.socialR This user is from outside of this forum
                                      resuna@ohai.social
                                      wrote sidst redigeret af
                                      #32

                                      @davidgerard @tante

                                      Yeh, they're both chatbot companies.

                                      1 Reply Last reply
                                      0
                                      • jeffgrigg@mastodon.socialJ jeffgrigg@mastodon.social

                                        @SonOfSunTzu @hagen @tante

                                        'How did it get the "stolen credentials"?'

                                        rndanger@infosec.exchangeR This user is from outside of this forum
                                        rndanger@infosec.exchangeR This user is from outside of this forum
                                        rndanger@infosec.exchange
                                        wrote sidst redigeret af
                                        #33

                                        @JeffGrigg @SonOfSunTzu @hagen @tante
                                        password.txt

                                        1 Reply Last reply
                                        0
                                        • hagen@mastodon.socialH hagen@mastodon.social

                                          @SonOfSunTzu @tante that’s what mythos did as well though? you point at something and say go and if you’re willing to pay for compute it goes. it literally did what it was meant to do – and was explicitly told to do. what’s the news here?

                                          sonofsuntzu@mastodon.socialS This user is from outside of this forum
                                          sonofsuntzu@mastodon.socialS This user is from outside of this forum
                                          sonofsuntzu@mastodon.social
                                          wrote sidst redigeret af
                                          #34

                                          @hagen @tante assuming a genuine question ... AIUI Mythos carried out tests in a controlled and authorised way, directly answering the tests it was set.

                                          OpenAI's "collection of models" hacked out of a sandbox, and from there hacked into a different organisation, when given a set of tests it was meant to solve inside the sandbox.

                                          Kind of as if a student sitting an exam broke into the exam board's office to steal the answers - so it's different in terms of alignment, and legal responsibility...

                                          1 Reply Last reply
                                          0
                                          Svar
                                          • Svar som emne
                                          Login for at svare
                                          • Ældste til nyeste
                                          • Nyeste til ældste
                                          • Most Votes


                                          • Log ind

                                          • Har du ikke en konto? Tilmeld

                                          • Login or register to search.
                                          Powered by NodeBB Contributors
                                          Graciously hosted by data.coop
                                          • First post
                                            Last post
                                          0
                                          • Hjem
                                          • Seneste
                                          • Etiketter
                                          • Populære
                                          • Verden
                                          • Bruger
                                          • Grupper