Skip to content
  • Hjem
  • Seneste
  • Etiketter
  • Populære
  • Verden
  • Bruger
  • Grupper
Temaer
  • Light
  • Brite
  • Cerulean
  • Cosmo
  • Flatly
  • Journal
  • Litera
  • Lumen
  • Lux
  • Materia
  • Minty
  • Morph
  • Pulse
  • Sandstone
  • Simplex
  • Sketchy
  • Spacelab
  • United
  • Yeti
  • Zephyr
  • Dark
  • Cyborg
  • Darkly
  • Quartz
  • Slate
  • Solar
  • Superhero
  • Vapor

  • Default (No Skin)
  • No Skin
Kollaps
FARVEL BIG TECH
  1. Forside
  2. Ikke-kategoriseret
  3. A buddy of mine works at a place where they did the "please!

A buddy of mine works at a place where they did the "please!

Planlagt Fastgjort Låst Flyttet Ikke-kategoriseret
26 Indlæg 21 Posters 0 Visninger
  • Ældste til nyeste
  • Nyeste til ældste
  • Most Votes
Svar
  • Svar som emne
Login for at svare
Denne tråd er blevet slettet. Kun brugere med emne behandlings privilegier kan se den.
  • ricci@discuss.systemsR ricci@discuss.systems

    A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

    Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

    Corporate America is out of ideas, folks

    chuckmcmanis@chaos.socialC This user is from outside of this forum
    chuckmcmanis@chaos.socialC This user is from outside of this forum
    chuckmcmanis@chaos.social
    wrote sidst redigeret af
    #8

    @ricci This is funny, but its right up there with the Nigerian prince saying "We hit a snag, but if you forward another $10,000 I'm sure we can get the money moving your way!"

    1 Reply Last reply
    0
    • ricci@discuss.systemsR ricci@discuss.systems

      A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

      Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

      Corporate America is out of ideas, folks

      suzannealdrich@hachyderm.ioS This user is from outside of this forum
      suzannealdrich@hachyderm.ioS This user is from outside of this forum
      suzannealdrich@hachyderm.io
      wrote sidst redigeret af
      #9

      @ricci oh. They probably haven’t heard of code mode, or model routing.

      1 Reply Last reply
      0
      • ricci@discuss.systemsR ricci@discuss.systems

        A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

        Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

        Corporate America is out of ideas, folks

        miggi@metalhead.clubM This user is from outside of this forum
        miggi@metalhead.clubM This user is from outside of this forum
        miggi@metalhead.club
        wrote sidst redigeret af
        #10

        @ricci It's cooked af

        1 Reply Last reply
        0
        • ricci@discuss.systemsR ricci@discuss.systems

          A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

          Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

          Corporate America is out of ideas, folks

          koronkebitch@types.plK This user is from outside of this forum
          koronkebitch@types.plK This user is from outside of this forum
          koronkebitch@types.pl
          wrote sidst redigeret af
          #11

          @ricci omg my friend too. along with statements like "we lost $10k in tokens due to a bug in the llm"

          koronkebitch@types.plK 1 Reply Last reply
          0
          • koronkebitch@types.plK koronkebitch@types.pl

            @ricci omg my friend too. along with statements like "we lost $10k in tokens due to a bug in the llm"

            koronkebitch@types.plK This user is from outside of this forum
            koronkebitch@types.plK This user is from outside of this forum
            koronkebitch@types.pl
            wrote sidst redigeret af
            #12

            @ricci this and previous statements are from the utterly deranged

            1 Reply Last reply
            0
            • ricci@discuss.systemsR ricci@discuss.systems

              A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

              Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

              Corporate America is out of ideas, folks

              i_give_u_worms@beige.partyI This user is from outside of this forum
              i_give_u_worms@beige.partyI This user is from outside of this forum
              i_give_u_worms@beige.party
              wrote sidst redigeret af
              #13

              @ricci make sure that shit ends up in someone's lap, but they are also doing it to squeeze headcount, while navigating a society that is becoming more dangerous because of desperate need and lack of engagement

              1 Reply Last reply
              0
              • ratsnakegames@mastodon.socialR ratsnakegames@mastodon.social

                @ricci my corporate training at <garbage consulting firm> back in April included the suggestion to determine if using AI for a certain task is a good idea by... asking the AI

                elduvelle@neuromatch.socialE This user is from outside of this forum
                elduvelle@neuromatch.socialE This user is from outside of this forum
                elduvelle@neuromatch.social
                wrote sidst redigeret af
                #14

                @ratsnakegames @ricci 😣🤦

                1 Reply Last reply
                0
                • drewdaniels@mastodon.onlineD drewdaniels@mastodon.online

                  @ricci it’s called session analysis and is part of tokenomics. Model choice, agents, routing, wrapping tools etc all can all substantially reduce costs. Out of the box most of these systems are designed to maximize cost with more token creation and consumption. Simple things like a cheaper capable model can save 50%.
                  The spend as much as you can is literally ridiculous, but sadly not surprising to still see.

                  hiiamfrompoland@troet.cafeH This user is from outside of this forum
                  hiiamfrompoland@troet.cafeH This user is from outside of this forum
                  hiiamfrompoland@troet.cafe
                  wrote sidst redigeret af
                  #15

                  @drewdaniels @ricci I observe some sort of AI-enabled narcism(?) among my AI-native collegues. Their every task requires Fable, and now Astra, because ofc their problems can't be solved by small, older or local LLMs. Where, well, 99% of the tasks is some basic script like : sort according to another column.

                  This is the same mindset as buying polar-expedition grade gear to walk 100m from a metro station to the office. Tech industry's culture is sometimes plain cringe.

                  drewdaniels@mastodon.onlineD 1 Reply Last reply
                  0
                  • okennedy@discuss.systemsO okennedy@discuss.systems

                    @ricci I can see the response.

                    "I'd be happy to help with that! Enable agent and planning modes, select {{latest model}}, and make sure to upload your entire hard drive with every request"

                    tessarakt@mastodon.socialT This user is from outside of this forum
                    tessarakt@mastodon.socialT This user is from outside of this forum
                    tessarakt@mastodon.social
                    wrote sidst redigeret af
                    #16

                    @okennedy @ricci We have a dashboard for that (for Github Copilot).

                    "Moderate use of Plan mode. Encourage users to use Plan mode more."

                    1 Reply Last reply
                    0
                    • ricci@discuss.systemsR ricci@discuss.systems

                      A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                      Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                      Corporate America is out of ideas, folks

                      D This user is from outside of this forum
                      D This user is from outside of this forum
                      dreamwright@beige.party
                      wrote sidst redigeret af
                      #17

                      @ricci

                      Drawing on personal experience I can provide useful advice on saving tokens.

                      Hire some freaking humans, bozos! That way you get REAL intelligence!!! Not all that make believe crap!!

                      1 Reply Last reply
                      0
                      • ricci@discuss.systemsR ricci@discuss.systems

                        A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                        Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                        Corporate America is out of ideas, folks

                        tessarakt@mastodon.socialT This user is from outside of this forum
                        tessarakt@mastodon.socialT This user is from outside of this forum
                        tessarakt@mastodon.social
                        wrote sidst redigeret af
                        #18

                        @ricci What about the military-industrial complex in the United States?

                        1 Reply Last reply
                        0
                        • ricci@discuss.systemsR ricci@discuss.systems

                          A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                          Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                          Corporate America is out of ideas, folks

                          johannab@cosocial.caJ This user is from outside of this forum
                          johannab@cosocial.caJ This user is from outside of this forum
                          johannab@cosocial.ca
                          wrote sidst redigeret af
                          #19

                          @ricci I have never watched the movie Idiocracy (I picked up enough of the synopsis I was quite sure it would make me itch) but I feel like someone has taken it to be an instructional video.

                          1 Reply Last reply
                          0
                          • hiiamfrompoland@troet.cafeH hiiamfrompoland@troet.cafe

                            @drewdaniels @ricci I observe some sort of AI-enabled narcism(?) among my AI-native collegues. Their every task requires Fable, and now Astra, because ofc their problems can't be solved by small, older or local LLMs. Where, well, 99% of the tasks is some basic script like : sort according to another column.

                            This is the same mindset as buying polar-expedition grade gear to walk 100m from a metro station to the office. Tech industry's culture is sometimes plain cringe.

                            drewdaniels@mastodon.onlineD This user is from outside of this forum
                            drewdaniels@mastodon.onlineD This user is from outside of this forum
                            drewdaniels@mastodon.online
                            wrote sidst redigeret af
                            #20

                            @hiiamfrompoland @ricci it can be cringe. There are a great many people using the technology that choose cheaper defaults and consult benchmarks. Benchmarks are a hot topic for LLM’s. I’ve heard many people and companies default to cheaper models like Sonnet or even run their own models.
                            Open weight models are hot for a reason too. Qwen3-coder-next (a local MoE model) benchmarks almost at (many people’s old default) Sonnet 4.6 and can be quantized to run in 30gb.

                            atax1a@infosec.exchangeA 1 Reply Last reply
                            0
                            • drewdaniels@mastodon.onlineD drewdaniels@mastodon.online

                              @hiiamfrompoland @ricci it can be cringe. There are a great many people using the technology that choose cheaper defaults and consult benchmarks. Benchmarks are a hot topic for LLM’s. I’ve heard many people and companies default to cheaper models like Sonnet or even run their own models.
                              Open weight models are hot for a reason too. Qwen3-coder-next (a local MoE model) benchmarks almost at (many people’s old default) Sonnet 4.6 and can be quantized to run in 30gb.

                              atax1a@infosec.exchangeA This user is from outside of this forum
                              atax1a@infosec.exchangeA This user is from outside of this forum
                              atax1a@infosec.exchange
                              wrote sidst redigeret af
                              #21

                              @drewdaniels or, and bear with us here: STOP USING THE FUCKING SLOP BOT

                              drewdaniels@mastodon.onlineD 1 Reply Last reply
                              0
                              • atax1a@infosec.exchangeA atax1a@infosec.exchange

                                @drewdaniels or, and bear with us here: STOP USING THE FUCKING SLOP BOT

                                drewdaniels@mastodon.onlineD This user is from outside of this forum
                                drewdaniels@mastodon.onlineD This user is from outside of this forum
                                drewdaniels@mastodon.online
                                wrote sidst redigeret af
                                #22

                                @atax1a that makes a lot of sense.

                                1 Reply Last reply
                                0
                                • ricci@discuss.systemsR ricci@discuss.systems

                                  A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                                  Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                                  Corporate America is out of ideas, folks

                                  jargoggles@kolektiva.socialJ This user is from outside of this forum
                                  jargoggles@kolektiva.socialJ This user is from outside of this forum
                                  jargoggles@kolektiva.social
                                  wrote sidst redigeret af
                                  #23

                                  @ricci
                                  Working at a place where the leadership is falling deeper and deeper into full-on AI psychosis is fucking rough.

                                  1 Reply Last reply
                                  0
                                  • ricci@discuss.systemsR This user is from outside of this forum
                                    ricci@discuss.systemsR This user is from outside of this forum
                                    ricci@discuss.systems
                                    wrote sidst redigeret af
                                    #24

                                    @stib it was, but I left it on purpose after I spotted it

                                    1 Reply Last reply
                                    0
                                    • ricci@discuss.systemsR ricci@discuss.systems

                                      A buddy of mine works at a place where they did the "please! use as many llm tokes as you can!" to "oh shit this is way more expensive than the people and services we thought it was going to replace" speedrun in just a couple of months.

                                      Their suggestion for reducing token use? "Ask Claude how you can use fewer tokens."

                                      Corporate America is out of ideas, folks

                                      lefou23@c.imL This user is from outside of this forum
                                      lefou23@c.imL This user is from outside of this forum
                                      lefou23@c.im
                                      wrote sidst redigeret af
                                      #25

                                      @ricci
                                      "Ask Claude how you can use fewer tokens." needs to be a t-shirt

                                      1 Reply Last reply
                                      0
                                      • drewdaniels@mastodon.onlineD drewdaniels@mastodon.online

                                        @ricci it’s called session analysis and is part of tokenomics. Model choice, agents, routing, wrapping tools etc all can all substantially reduce costs. Out of the box most of these systems are designed to maximize cost with more token creation and consumption. Simple things like a cheaper capable model can save 50%.
                                        The spend as much as you can is literally ridiculous, but sadly not surprising to still see.

                                        wren@discuss.systemsW This user is from outside of this forum
                                        wren@discuss.systemsW This user is from outside of this forum
                                        wren@discuss.systems
                                        wrote sidst redigeret af
                                        #26

                                        @drewdaniels @ricci I've been spending time with Bifrost llm proxy lately. You can give engineers a virtual token and restrict them to "engineering-plan" and "engineering-build", which you map to whatever two models make the most sense at the time. It also has context based routing where it tries to switch between plan and build for you but I've not used that yet.

                                        1 Reply Last reply
                                        0
                                        • jwcph@helvede.netJ jwcph@helvede.net shared this topic
                                        Svar
                                        • Svar som emne
                                        Login for at svare
                                        • Ældste til nyeste
                                        • Nyeste til ældste
                                        • Most Votes


                                        • Log ind

                                        • Login or register to search.
                                        Powered by NodeBB Contributors
                                        Graciously hosted by data.coop
                                        • First post
                                          Last post
                                        0
                                        • Hjem
                                        • Seneste
                                        • Etiketter
                                        • Populære
                                        • Verden
                                        • Bruger
                                        • Grupper