Skip to content
  • Hjem
  • Seneste
  • Etiketter
  • Populære
  • Verden
  • Bruger
  • Grupper
Temaer
  • Light
  • Brite
  • Cerulean
  • Cosmo
  • Flatly
  • Journal
  • Litera
  • Lumen
  • Lux
  • Materia
  • Minty
  • Morph
  • Pulse
  • Sandstone
  • Simplex
  • Sketchy
  • Spacelab
  • United
  • Yeti
  • Zephyr
  • Dark
  • Cyborg
  • Darkly
  • Quartz
  • Slate
  • Solar
  • Superhero
  • Vapor

  • Default (No Skin)
  • No Skin
Kollaps
FARVEL BIG TECH
  1. Forside
  2. Ikke-kategoriseret
  3. A buddy of mine works at a place where they did the "please!

A buddy of mine works at a place where they did the "please!

Planlagt Fastgjort Låst Flyttet Ikke-kategoriseret
26 Indlæg 21 Posters 0 Visninger
  • Ældste til nyeste
  • Nyeste til ældste
  • Most Votes
Svar
  • Svar som emne
Login for at svare
Denne tråd er blevet slettet. Kun brugere med emne behandlings privilegier kan se den.
  • ricci@discuss.systemsR ricci@discuss.systems

    A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

    Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

    Corporate America is out of ideas, folks

    drewdaniels@mastodon.onlineD This user is from outside of this forum
    drewdaniels@mastodon.onlineD This user is from outside of this forum
    drewdaniels@mastodon.online
    wrote sidst redigeret af
    #4

    @ricci it’s called session analysis and is part of tokenomics. Model choice, agents, routing, wrapping tools etc all can all substantially reduce costs. Out of the box most of these systems are designed to maximize cost with more token creation and consumption. Simple things like a cheaper capable model can save 50%.
    The spend as much as you can is literally ridiculous, but sadly not surprising to still see.

    hiiamfrompoland@troet.cafeH wren@discuss.systemsW 2 Replies Last reply
    0
    • ricci@discuss.systemsR ricci@discuss.systems

      A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

      Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

      Corporate America is out of ideas, folks

      jernej__s@infosec.exchangeJ This user is from outside of this forum
      jernej__s@infosec.exchangeJ This user is from outside of this forum
      jernej__s@infosec.exchange
      wrote sidst redigeret af
      #5

      @ricci Saw this earlier today on Discord:

      1 Reply Last reply
      0
      • ricci@discuss.systemsR ricci@discuss.systems

        A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

        Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

        Corporate America is out of ideas, folks

        kagan@wandering.shopK This user is from outside of this forum
        kagan@wandering.shopK This user is from outside of this forum
        kagan@wandering.shop
        wrote sidst redigeret af
        #6

        @ricci I swear, once people get hooked on "AI" and become meat proxies, they just turn off their brains completely. For, like, *everything*. And I still don't know if there's a way to get them to turn back on again.

        1 Reply Last reply
        0
        • ricci@discuss.systemsR ricci@discuss.systems

          A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

          Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

          Corporate America is out of ideas, folks

          rubinjoni@mastodon.socialR This user is from outside of this forum
          rubinjoni@mastodon.socialR This user is from outside of this forum
          rubinjoni@mastodon.social
          wrote sidst redigeret af
          #7

          @ricci Ask your dealer how to use less drugs.

          1 Reply Last reply
          0
          • ricci@discuss.systemsR ricci@discuss.systems

            A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

            Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

            Corporate America is out of ideas, folks

            chuckmcmanis@chaos.socialC This user is from outside of this forum
            chuckmcmanis@chaos.socialC This user is from outside of this forum
            chuckmcmanis@chaos.social
            wrote sidst redigeret af
            #8

            @ricci This is funny, but its right up there with the Nigerian prince saying "We hit a snag, but if you forward another $10,000 I'm sure we can get the money moving your way!"

            1 Reply Last reply
            0
            • ricci@discuss.systemsR ricci@discuss.systems

              A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

              Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

              Corporate America is out of ideas, folks

              suzannealdrich@hachyderm.ioS This user is from outside of this forum
              suzannealdrich@hachyderm.ioS This user is from outside of this forum
              suzannealdrich@hachyderm.io
              wrote sidst redigeret af
              #9

              @ricci oh. They probably haven’t heard of code mode, or model routing.

              1 Reply Last reply
              0
              • ricci@discuss.systemsR ricci@discuss.systems

                A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                Corporate America is out of ideas, folks

                miggi@metalhead.clubM This user is from outside of this forum
                miggi@metalhead.clubM This user is from outside of this forum
                miggi@metalhead.club
                wrote sidst redigeret af
                #10

                @ricci It's cooked af

                1 Reply Last reply
                0
                • ricci@discuss.systemsR ricci@discuss.systems

                  A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                  Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                  Corporate America is out of ideas, folks

                  koronkebitch@types.plK This user is from outside of this forum
                  koronkebitch@types.plK This user is from outside of this forum
                  koronkebitch@types.pl
                  wrote sidst redigeret af
                  #11

                  @ricci omg my friend too. along with statements like "we lost $10k in tokens due to a bug in the llm"

                  koronkebitch@types.plK 1 Reply Last reply
                  0
                  • koronkebitch@types.plK koronkebitch@types.pl

                    @ricci omg my friend too. along with statements like "we lost $10k in tokens due to a bug in the llm"

                    koronkebitch@types.plK This user is from outside of this forum
                    koronkebitch@types.plK This user is from outside of this forum
                    koronkebitch@types.pl
                    wrote sidst redigeret af
                    #12

                    @ricci this and previous statements are from the utterly deranged

                    1 Reply Last reply
                    0
                    • ricci@discuss.systemsR ricci@discuss.systems

                      A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                      Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                      Corporate America is out of ideas, folks

                      i_give_u_worms@beige.partyI This user is from outside of this forum
                      i_give_u_worms@beige.partyI This user is from outside of this forum
                      i_give_u_worms@beige.party
                      wrote sidst redigeret af
                      #13

                      @ricci make sure that shit ends up in someone's lap, but they are also doing it to squeeze headcount, while navigating a society that is becoming more dangerous because of desperate need and lack of engagement

                      1 Reply Last reply
                      0
                      • ratsnakegames@mastodon.socialR ratsnakegames@mastodon.social

                        @ricci my corporate training at <garbage consulting firm> back in April included the suggestion to determine if using AI for a certain task is a good idea by... asking the AI

                        elduvelle@neuromatch.socialE This user is from outside of this forum
                        elduvelle@neuromatch.socialE This user is from outside of this forum
                        elduvelle@neuromatch.social
                        wrote sidst redigeret af
                        #14

                        @ratsnakegames @ricci 😣🤦

                        1 Reply Last reply
                        0
                        • drewdaniels@mastodon.onlineD drewdaniels@mastodon.online

                          @ricci it’s called session analysis and is part of tokenomics. Model choice, agents, routing, wrapping tools etc all can all substantially reduce costs. Out of the box most of these systems are designed to maximize cost with more token creation and consumption. Simple things like a cheaper capable model can save 50%.
                          The spend as much as you can is literally ridiculous, but sadly not surprising to still see.

                          hiiamfrompoland@troet.cafeH This user is from outside of this forum
                          hiiamfrompoland@troet.cafeH This user is from outside of this forum
                          hiiamfrompoland@troet.cafe
                          wrote sidst redigeret af
                          #15

                          @drewdaniels @ricci I observe some sort of AI-enabled narcism(?) among my AI-native collegues. Their every task requires Fable, and now Astra, because ofc their problems can't be solved by small, older or local LLMs. Where, well, 99% of the tasks is some basic script like : sort according to another column.

                          This is the same mindset as buying polar-expedition grade gear to walk 100m from a metro station to the office. Tech industry's culture is sometimes plain cringe.

                          drewdaniels@mastodon.onlineD 1 Reply Last reply
                          0
                          • okennedy@discuss.systemsO okennedy@discuss.systems

                            @ricci I can see the response.

                            "I'd be happy to help with that! Enable agent and planning modes, select {{latest model}}, and make sure to upload your entire hard drive with every request"

                            tessarakt@mastodon.socialT This user is from outside of this forum
                            tessarakt@mastodon.socialT This user is from outside of this forum
                            tessarakt@mastodon.social
                            wrote sidst redigeret af
                            #16

                            @okennedy @ricci We have a dashboard for that (for Github Copilot).

                            "Moderate use of Plan mode. Encourage users to use Plan mode more."

                            1 Reply Last reply
                            0
                            • ricci@discuss.systemsR ricci@discuss.systems

                              A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                              Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                              Corporate America is out of ideas, folks

                              D This user is from outside of this forum
                              D This user is from outside of this forum
                              dreamwright@beige.party
                              wrote sidst redigeret af
                              #17

                              @ricci

                              Drawing on personal experience I can provide useful advice on saving tokens.

                              Hire some freaking humans, bozos! That way you get REAL intelligence!!! Not all that make believe crap!!

                              1 Reply Last reply
                              0
                              • ricci@discuss.systemsR ricci@discuss.systems

                                A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                                Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                                Corporate America is out of ideas, folks

                                tessarakt@mastodon.socialT This user is from outside of this forum
                                tessarakt@mastodon.socialT This user is from outside of this forum
                                tessarakt@mastodon.social
                                wrote sidst redigeret af
                                #18

                                @ricci What about the military-industrial complex in the United States?

                                1 Reply Last reply
                                0
                                • ricci@discuss.systemsR ricci@discuss.systems

                                  A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                                  Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                                  Corporate America is out of ideas, folks

                                  johannab@cosocial.caJ This user is from outside of this forum
                                  johannab@cosocial.caJ This user is from outside of this forum
                                  johannab@cosocial.ca
                                  wrote sidst redigeret af
                                  #19

                                  @ricci I have never watched the movie Idiocracy (I picked up enough of the synopsis I was quite sure it would make me itch) but I feel like someone has taken it to be an instructional video.

                                  1 Reply Last reply
                                  0
                                  • hiiamfrompoland@troet.cafeH hiiamfrompoland@troet.cafe

                                    @drewdaniels @ricci I observe some sort of AI-enabled narcism(?) among my AI-native collegues. Their every task requires Fable, and now Astra, because ofc their problems can't be solved by small, older or local LLMs. Where, well, 99% of the tasks is some basic script like : sort according to another column.

                                    This is the same mindset as buying polar-expedition grade gear to walk 100m from a metro station to the office. Tech industry's culture is sometimes plain cringe.

                                    drewdaniels@mastodon.onlineD This user is from outside of this forum
                                    drewdaniels@mastodon.onlineD This user is from outside of this forum
                                    drewdaniels@mastodon.online
                                    wrote sidst redigeret af
                                    #20

                                    @hiiamfrompoland @ricci it can be cringe. There are a great many people using the technology that choose cheaper defaults and consult benchmarks. Benchmarks are a hot topic for LLM’s. I’ve heard many people and companies default to cheaper models like Sonnet or even run their own models.
                                    Open weight models are hot for a reason too. Qwen3-coder-next (a local MoE model) benchmarks almost at (many people’s old default) Sonnet 4.6 and can be quantized to run in 30gb.

                                    atax1a@infosec.exchangeA 1 Reply Last reply
                                    0
                                    • drewdaniels@mastodon.onlineD drewdaniels@mastodon.online

                                      @hiiamfrompoland @ricci it can be cringe. There are a great many people using the technology that choose cheaper defaults and consult benchmarks. Benchmarks are a hot topic for LLM’s. I’ve heard many people and companies default to cheaper models like Sonnet or even run their own models.
                                      Open weight models are hot for a reason too. Qwen3-coder-next (a local MoE model) benchmarks almost at (many people’s old default) Sonnet 4.6 and can be quantized to run in 30gb.

                                      atax1a@infosec.exchangeA This user is from outside of this forum
                                      atax1a@infosec.exchangeA This user is from outside of this forum
                                      atax1a@infosec.exchange
                                      wrote sidst redigeret af
                                      #21

                                      @drewdaniels or, and bear with us here: STOP USING THE FUCKING SLOP BOT

                                      drewdaniels@mastodon.onlineD 1 Reply Last reply
                                      0
                                      • atax1a@infosec.exchangeA atax1a@infosec.exchange

                                        @drewdaniels or, and bear with us here: STOP USING THE FUCKING SLOP BOT

                                        drewdaniels@mastodon.onlineD This user is from outside of this forum
                                        drewdaniels@mastodon.onlineD This user is from outside of this forum
                                        drewdaniels@mastodon.online
                                        wrote sidst redigeret af
                                        #22

                                        @atax1a that makes a lot of sense.

                                        1 Reply Last reply
                                        0
                                        • ricci@discuss.systemsR ricci@discuss.systems

                                          A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                                          Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                                          Corporate America is out of ideas, folks

                                          jargoggles@kolektiva.socialJ This user is from outside of this forum
                                          jargoggles@kolektiva.socialJ This user is from outside of this forum
                                          jargoggles@kolektiva.social
                                          wrote sidst redigeret af
                                          #23

                                          @ricci
                                          Working at a place where the leadership is falling deeper and deeper into full-on AI psychosis is fucking rough.

                                          1 Reply Last reply
                                          0
                                          Svar
                                          • Svar som emne
                                          Login for at svare
                                          • Ældste til nyeste
                                          • Nyeste til ældste
                                          • Most Votes


                                          • Log ind

                                          • Login or register to search.
                                          Powered by NodeBB Contributors
                                          Graciously hosted by data.coop
                                          • First post
                                            Last post
                                          0
                                          • Hjem
                                          • Seneste
                                          • Etiketter
                                          • Populære
                                          • Verden
                                          • Bruger
                                          • Grupper