Skip to content
  • Hjem
  • Seneste
  • Etiketter
  • Populære
  • Verden
  • Bruger
  • Grupper
Temaer
  • Light
  • Brite
  • Cerulean
  • Cosmo
  • Flatly
  • Journal
  • Litera
  • Lumen
  • Lux
  • Materia
  • Minty
  • Morph
  • Pulse
  • Sandstone
  • Simplex
  • Sketchy
  • Spacelab
  • United
  • Yeti
  • Zephyr
  • Dark
  • Cyborg
  • Darkly
  • Quartz
  • Slate
  • Solar
  • Superhero
  • Vapor

  • Default (No Skin)
  • No Skin
Kollaps
FARVEL BIG TECH
  1. Forside
  2. Ikke-kategoriseret
  3. No, OpenAI's new magic models did not autonomously hack Huggingface.

No, OpenAI's new magic models did not autonomously hack Huggingface.

Planlagt Fastgjort Låst Flyttet Ikke-kategoriseret
44 Indlæg 28 Posters 0 Visninger
  • Ældste til nyeste
  • Nyeste til ældste
  • Most Votes
Svar
  • Svar som emne
Login for at svare
Denne tråd er blevet slettet. Kun brugere med emne behandlings privilegier kan se den.
  • jeffgrigg@mastodon.socialJ jeffgrigg@mastodon.social

    @txo_elurmaluta @Gurre @tante

    Their Public Relations department can say that all they want.

    But we know what it really stands for.

    jeffgrigg@mastodon.socialJ This user is from outside of this forum
    jeffgrigg@mastodon.socialJ This user is from outside of this forum
    jeffgrigg@mastodon.social
    wrote sidst redigeret af
    #26

    @txo_elurmaluta @Gurre @tante

    Open AI's statement,

    "These deployment safeguards were intentionally not enabled during this evaluation because it was aimed at testing cyber vulnerabilities."

    makes me think of all the actions of all the "corporate types" in the Alien franchise too.

    And that's *never* a good thing. 💢

    1 Reply Last reply
    0
    • miles_leif@mastodon.socialM miles_leif@mastodon.social

      @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

      yhancik@thereisno.computerY This user is from outside of this forum
      yhancik@thereisno.computerY This user is from outside of this forum
      yhancik@thereisno.computer
      wrote sidst redigeret af
      #27

      @miles_leif @tante this is how 99% of the press is dealing with AI since the beginning of the hype, a hype they contribute to feed.

      1 Reply Last reply
      0
      • tante@tldr.nettime.orgT tante@tldr.nettime.org

        No, OpenAI's new magic models did not autonomously hack Huggingface.

        Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

        "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

        So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

        The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

        tyzbit@toot.nowT This user is from outside of this forum
        tyzbit@toot.nowT This user is from outside of this forum
        tyzbit@toot.now
        wrote sidst redigeret af
        #28

        @tante

        say you hacked huggingface

        I hacked huggingface.

        oh my god

        1 Reply Last reply
        0
        • tante@tldr.nettime.orgT tante@tldr.nettime.org

          No, OpenAI's new magic models did not autonomously hack Huggingface.

          Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

          "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

          So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

          The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

          F This user is from outside of this forum
          F This user is from outside of this forum
          failedlyndonlarouchite@mas.to
          wrote sidst redigeret af
          #29

          @tante

          yesterday NPR had a long story on the new hot battery powered pickup truck from the startup
          https://www.slate.auto/en

          and the NPR piece was full of the most hilarious nonsense, all provided to NPR by Slate's excellent PR team

          shrug

          1 Reply Last reply
          0
          • tante@tldr.nettime.orgT tante@tldr.nettime.org

            No, OpenAI's new magic models did not autonomously hack Huggingface.

            Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

            "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

            So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

            The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

            h0ru2@cyberplace.socialH This user is from outside of this forum
            h0ru2@cyberplace.socialH This user is from outside of this forum
            h0ru2@cyberplace.social
            wrote sidst redigeret af
            #30

            @tante "OpenAI told it to do that, the model didn't do shit autonomously [...]"
            Exactly, what one should have expected (and how it turned out before every time).

            1 Reply Last reply
            0
            • miles_leif@mastodon.socialM miles_leif@mastodon.social

              @tante How Tagesschau is reporting about it can only be described as refusal to practice journalism. 7 out of 9 paragraphs say "Open AI says" and the other two quote nameless "Experts" and Anthropic's product. And tomorrow I have to discuss with someone again about the possibility of consciousness in token generator software https://www.tagesschau.de/wirtschaft/unternehmen/openai-ki-hackerangriff-100.html

              h0ru2@cyberplace.socialH This user is from outside of this forum
              h0ru2@cyberplace.socialH This user is from outside of this forum
              h0ru2@cyberplace.social
              wrote sidst redigeret af
              #31

              @miles_leif @tante Not the only one unfortunately, sigh

              1 Reply Last reply
              0
              • davidgerard@circumstances.runD davidgerard@circumstances.run

                @tante a friend suggests it's co-marketing with Huggingface, if you look at the timeline of announcements

                resuna@ohai.socialR This user is from outside of this forum
                resuna@ohai.socialR This user is from outside of this forum
                resuna@ohai.social
                wrote sidst redigeret af
                #32

                @davidgerard @tante

                Yeh, they're both chatbot companies.

                1 Reply Last reply
                0
                • jeffgrigg@mastodon.socialJ jeffgrigg@mastodon.social

                  @SonOfSunTzu @hagen @tante

                  'How did it get the "stolen credentials"?'

                  rndanger@infosec.exchangeR This user is from outside of this forum
                  rndanger@infosec.exchangeR This user is from outside of this forum
                  rndanger@infosec.exchange
                  wrote sidst redigeret af
                  #33

                  @JeffGrigg @SonOfSunTzu @hagen @tante
                  password.txt

                  1 Reply Last reply
                  0
                  • hagen@mastodon.socialH hagen@mastodon.social

                    @SonOfSunTzu @tante that’s what mythos did as well though? you point at something and say go and if you’re willing to pay for compute it goes. it literally did what it was meant to do – and was explicitly told to do. what’s the news here?

                    sonofsuntzu@mastodon.socialS This user is from outside of this forum
                    sonofsuntzu@mastodon.socialS This user is from outside of this forum
                    sonofsuntzu@mastodon.social
                    wrote sidst redigeret af
                    #34

                    @hagen @tante assuming a genuine question ... AIUI Mythos carried out tests in a controlled and authorised way, directly answering the tests it was set.

                    OpenAI's "collection of models" hacked out of a sandbox, and from there hacked into a different organisation, when given a set of tests it was meant to solve inside the sandbox.

                    Kind of as if a student sitting an exam broke into the exam board's office to steal the answers - so it's different in terms of alignment, and legal responsibility...

                    1 Reply Last reply
                    0
                    • amethyst@chaos.socialA amethyst@chaos.social

                      @michelin @tante I'm more wondering why there is not a huge public outrage about letting such models run loose... in particular in view of the current hype of using AI in weapons development this is immensively scary.

                      michelin@hachyderm.ioM This user is from outside of this forum
                      michelin@hachyderm.ioM This user is from outside of this forum
                      michelin@hachyderm.io
                      wrote sidst redigeret af
                      #35

                      @amethyst @tante from this ex-Google DeepMind scientist, even the "Safe and Ethical AI" (sic) folks are not interested in rocking the boat.

                      We're really in the Bad Place

                      https://turntrout.com/why-i-left-google-deepmind

                      1 Reply Last reply
                      0
                      • tante@tldr.nettime.orgT tante@tldr.nettime.org

                        No, OpenAI's new magic models did not autonomously hack Huggingface.

                        Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                        "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                        So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                        The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                        mond@mas.toM This user is from outside of this forum
                        mond@mas.toM This user is from outside of this forum
                        mond@mas.to
                        wrote sidst redigeret af
                        #36

                        @tante

                        "no LLM ever does anything autonomously, it's always prompted)."

                        Subgoals. You give it a complex task and it tries to solve it by breaking it down in subgoals... That it invents all itself...

                        Sure this COULD be a PR stunt. But then a lot of employees from 2 different companies would have to be in on the conspiracy... Not so likely

                        cm@chaos.socialC 1 Reply Last reply
                        0
                        • mond@mas.toM mond@mas.to

                          @tante

                          "no LLM ever does anything autonomously, it's always prompted)."

                          Subgoals. You give it a complex task and it tries to solve it by breaking it down in subgoals... That it invents all itself...

                          Sure this COULD be a PR stunt. But then a lot of employees from 2 different companies would have to be in on the conspiracy... Not so likely

                          cm@chaos.socialC This user is from outside of this forum
                          cm@chaos.socialC This user is from outside of this forum
                          cm@chaos.social
                          wrote sidst redigeret af
                          #37

                          @mond @tante It's not a conspiracy, it is a cult, all the bros believe in it.

                          mond@mas.toM 1 Reply Last reply
                          0
                          • tante@tldr.nettime.orgT tante@tldr.nettime.org

                            No, OpenAI's new magic models did not autonomously hack Huggingface.

                            Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                            "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                            So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                            The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                            10tothe22@mastodon.social1 This user is from outside of this forum
                            10tothe22@mastodon.social1 This user is from outside of this forum
                            10tothe22@mastodon.social
                            wrote sidst redigeret af
                            #38

                            @tante Anything to keep the billions flowing. What a disgrace.

                            1 Reply Last reply
                            0
                            • tante@tldr.nettime.orgT tante@tldr.nettime.org

                              No, OpenAI's new magic models did not autonomously hack Huggingface.

                              Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                              "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                              So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                              The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                              maxleibman@beige.partyM This user is from outside of this forum
                              maxleibman@beige.partyM This user is from outside of this forum
                              maxleibman@beige.party
                              wrote sidst redigeret af
                              #39

                              @tante Indeed. As I put it yesterday, OpenAI’s agent “broke containment” the same way my cuckoo clock will break containment if I put it in front of a button.

                              1 Reply Last reply
                              0
                              • cm@chaos.socialC cm@chaos.social

                                @mond @tante It's not a conspiracy, it is a cult, all the bros believe in it.

                                mond@mas.toM This user is from outside of this forum
                                mond@mas.toM This user is from outside of this forum
                                mond@mas.to
                                wrote sidst redigeret af
                                #40

                                @cm @tante

                                To me it rather seems that the AI-denial is a cult..

                                Bergman:

                                https://youtu.be/KpTZbq-eV38?is=yIA9reuuiFsgwaRO

                                And with the "humiliations argument" we now have an explanation for it:

                                https://qummunismus.at/a/article200/

                                tante@tldr.nettime.orgT cm@chaos.socialC ovrim@wien.rocksO 3 Replies Last reply
                                0
                                • mond@mas.toM mond@mas.to

                                  @cm @tante

                                  To me it rather seems that the AI-denial is a cult..

                                  Bergman:

                                  https://youtu.be/KpTZbq-eV38?is=yIA9reuuiFsgwaRO

                                  And with the "humiliations argument" we now have an explanation for it:

                                  https://qummunismus.at/a/article200/

                                  tante@tldr.nettime.orgT This user is from outside of this forum
                                  tante@tldr.nettime.orgT This user is from outside of this forum
                                  tante@tldr.nettime.org
                                  wrote sidst redigeret af
                                  #41

                                  @mond @cm Bregman is so AI pilled that he lets Claude write these essays for him (otherwise they'd probably be less unstructured and have actual arguments). I'd pick a better posterboy for the pro AI position 😉

                                  1 Reply Last reply
                                  0
                                  • mond@mas.toM mond@mas.to

                                    @cm @tante

                                    To me it rather seems that the AI-denial is a cult..

                                    Bergman:

                                    https://youtu.be/KpTZbq-eV38?is=yIA9reuuiFsgwaRO

                                    And with the "humiliations argument" we now have an explanation for it:

                                    https://qummunismus.at/a/article200/

                                    cm@chaos.socialC This user is from outside of this forum
                                    cm@chaos.socialC This user is from outside of this forum
                                    cm@chaos.social
                                    wrote sidst redigeret af
                                    #42

                                    @mond I know you're also in the cult, let's agree to disagree or I'll have to tear apart your blog post to point out all the cultish stuff in there. Let's just say, regarding "humiliation", that all I've seen from this tech definitely shows it's not satisfaktionsfähig. @tante

                                    1 Reply Last reply
                                    0
                                    • tante@tldr.nettime.orgT tante@tldr.nettime.org

                                      No, OpenAI's new magic models did not autonomously hack Huggingface.

                                      Per OpenAI's PR blog post (https://openai.com/index/hugging-face-model-evaluation-security-incident/😞

                                      "This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber capabilities. We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

                                      So they prompted their model running without guardrails to hack some shit. OpenAI told it to do that, the model didn't do shit autonomously (no LLM ever does anything autonomously, it's always prompted).

                                      The whole story is PR. We know that from Anthopic's Mythos: "Look we have this super secret new model and it's so powerful. We are scared ourselves. And you will soon be able to rent it!"

                                      tbodt@mastodon.socialT This user is from outside of this forum
                                      tbodt@mastodon.socialT This user is from outside of this forum
                                      tbodt@mastodon.social
                                      wrote sidst redigeret af
                                      #43

                                      @tante
                                      - i do not think it's unlikely that the prompt said "get a good score on this benchmark" and the model wildly misinterpreted the goal, as anyone who uses llms sees them do all the time
                                      - even if the prompt was "hack huggingface", the fact that it did that without further prompting is noteworthy.

                                      1 Reply Last reply
                                      0
                                      • mond@mas.toM mond@mas.to

                                        @cm @tante

                                        To me it rather seems that the AI-denial is a cult..

                                        Bergman:

                                        https://youtu.be/KpTZbq-eV38?is=yIA9reuuiFsgwaRO

                                        And with the "humiliations argument" we now have an explanation for it:

                                        https://qummunismus.at/a/article200/

                                        ovrim@wien.rocksO This user is from outside of this forum
                                        ovrim@wien.rocksO This user is from outside of this forum
                                        ovrim@wien.rocks
                                        wrote sidst redigeret af
                                        #44

                                        @mond @cm @tante that's a nice propaganda video and nothing more ...

                                        1 Reply Last reply
                                        0
                                        • pelle@veganism.socialP pelle@veganism.social shared this topic
                                        Svar
                                        • Svar som emne
                                        Login for at svare
                                        • Ældste til nyeste
                                        • Nyeste til ældste
                                        • Most Votes


                                        • Log ind

                                        • Har du ikke en konto? Tilmeld

                                        • Login or register to search.
                                        Powered by NodeBB Contributors
                                        Graciously hosted by data.coop
                                        • First post
                                          Last post
                                        0
                                        • Hjem
                                        • Seneste
                                        • Etiketter
                                        • Populære
                                        • Verden
                                        • Bruger
                                        • Grupper