Skip to content
  • Hjem
  • Seneste
  • Etiketter
  • Populære
  • Verden
  • Bruger
  • Grupper
Temaer
  • Light
  • Brite
  • Cerulean
  • Cosmo
  • Flatly
  • Journal
  • Litera
  • Lumen
  • Lux
  • Materia
  • Minty
  • Morph
  • Pulse
  • Sandstone
  • Simplex
  • Sketchy
  • Spacelab
  • United
  • Yeti
  • Zephyr
  • Dark
  • Cyborg
  • Darkly
  • Quartz
  • Slate
  • Solar
  • Superhero
  • Vapor

  • Default (No Skin)
  • No Skin
Kollaps
FARVEL BIG TECH
  1. Forside
  2. Ikke-kategoriseret
  3. "One guy in #Sweden built a #searchengine to fight Google, and it works.

"One guy in #Sweden built a #searchengine to fight Google, and it works.

Planlagt Fastgjort Låst Flyttet Ikke-kategoriseret
swedensearchenginemarginaliasearccrawlerindex
113 Indlæg 93 Posters 1 Visninger
  • Ældste til nyeste
  • Nyeste til ældste
  • Most Votes
Svar
  • Svar som emne
Login for at svare
Denne tråd er blevet slettet. Kun brugere med emne behandlings privilegier kan se den.
  • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

    "One guy in #Sweden built a #searchengine to fight Google, and it works.

    It's called #MarginaliaSearch.

    It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

    What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

    The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

    There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

    It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

    https://marginalia-search.com/

    hammerwell@troet.cafeH This user is from outside of this forum
    hammerwell@troet.cafeH This user is from outside of this forum
    hammerwell@troet.cafe
    wrote sidst redigeret af
    #60

    @eliasulrich Qwant and Ecosia are building their own index too. https://www.eu-searchperspective.com/

    1 Reply Last reply
    0
    • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

      "One guy in #Sweden built a #searchengine to fight Google, and it works.

      It's called #MarginaliaSearch.

      It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

      What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

      The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

      There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

      It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

      https://marginalia-search.com/

      mr_chons@mastodon.worldM This user is from outside of this forum
      mr_chons@mastodon.worldM This user is from outside of this forum
      mr_chons@mastodon.world
      wrote sidst redigeret af
      #61

      @eliasulrich Oh, boy. I like. I like.

      1 Reply Last reply
      0
      • mkalmes@chaos.socialM mkalmes@chaos.social

        @eliasulrich
        cc @hukl

        hukl@chaos.socialH This user is from outside of this forum
        hukl@chaos.socialH This user is from outside of this forum
        hukl@chaos.social
        wrote sidst redigeret af
        #62

        @mkalmes @eliasulrich yeah uruky is using that as one index source 🙂

        mkalmes@chaos.socialM 1 Reply Last reply
        0
        • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

          "One guy in #Sweden built a #searchengine to fight Google, and it works.

          It's called #MarginaliaSearch.

          It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

          What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

          The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

          There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

          It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

          https://marginalia-search.com/

          joost_rekveld@assemblag.esJ This user is from outside of this forum
          joost_rekveld@assemblag.esJ This user is from outside of this forum
          joost_rekveld@assemblag.es
          wrote sidst redigeret af
          #63

          @eliasulrich I use it quite a bit !

          1 Reply Last reply
          0
          • hukl@chaos.socialH hukl@chaos.social

            @mkalmes @eliasulrich yeah uruky is using that as one index source 🙂

            mkalmes@chaos.socialM This user is from outside of this forum
            mkalmes@chaos.socialM This user is from outside of this forum
            mkalmes@chaos.social
            wrote sidst redigeret af
            #64

            @hukl
            Uh, good to know. Thanks for letting me know.

            1 Reply Last reply
            0
            • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

              "One guy in #Sweden built a #searchengine to fight Google, and it works.

              It's called #MarginaliaSearch.

              It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

              What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

              The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

              There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

              It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

              https://marginalia-search.com/

              akk666@norden.socialA This user is from outside of this forum
              akk666@norden.socialA This user is from outside of this forum
              akk666@norden.social
              wrote sidst redigeret af
              #65

              @eliasulrich I love
              https://marginalia-search.com/
              It still finds lots of fandom sites and blogs that are actually readable! Esp. in the niche topics that get buried elsewhere. 🙂
              Heh, it even found my ancient fanfics along with my most recent research paper. ❤️
              #search #fanfic

              1 Reply Last reply
              0
              • devonkiwi@mastodonapp.ukD devonkiwi@mastodonapp.uk

                @eliasulrich Why won't they let us have nice things?
                "The search engine is currently under being hit so aggressively by bots it's interfering with the ability to serve regular search traffic. Emergency anti-scraping measures are enabled. Sorry about the inconvenience."

                schmidt_fu@mstdn.socialS This user is from outside of this forum
                schmidt_fu@mstdn.socialS This user is from outside of this forum
                schmidt_fu@mstdn.social
                wrote sidst redigeret af
                #66

                @Devonkiwi
                This might as well be the link going around on various Mastodon servers and people clicking on it. Sometimes drivers *are* the traffic themselves 🤷
                @eliasulrich

                rdfhrn@hessen.socialR 1 Reply Last reply
                0
                • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                  "One guy in #Sweden built a #searchengine to fight Google, and it works.

                  It's called #MarginaliaSearch.

                  It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                  What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                  The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                  There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                  It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                  https://marginalia-search.com/

                  knowprose@mastodon.socialK This user is from outside of this forum
                  knowprose@mastodon.socialK This user is from outside of this forum
                  knowprose@mastodon.social
                  wrote sidst redigeret af
                  #67

                  @eliasulrich this is great! Thanks!

                  1 Reply Last reply
                  0
                  • tjenwellens@mas.toT tjenwellens@mas.to

                    @woe2you @miramarmike @daveybot @eliasulrich I remember having to tell google about my blog existing. Back when I started it. So kinda seems normal that you'd need to do that again for a search engine that starts from scratch

                    msbellows@c.imM This user is from outside of this forum
                    msbellows@c.imM This user is from outside of this forum
                    msbellows@c.im
                    wrote sidst redigeret af
                    #68

                    @TjenWellens @woe2you @miramarmike @daveybot @eliasulrich When ChatGPT first came out, I asked it about myself and learned a lot of good stuff, including that I've never been a writer (coming as a shock to HuffPost, The Guardian, etc.) but that I *am* on faculty at the local law school (shocking the law school but pleasing my mother no end!). So I guess we're confronted with the choice between search engines that don't know we're alive and search engines that make up imaginary lives for us.

                    mansr@society.oftrolls.comM 1 Reply Last reply
                    0
                    • schmidt_fu@mstdn.socialS schmidt_fu@mstdn.social

                      @Devonkiwi
                      This might as well be the link going around on various Mastodon servers and people clicking on it. Sometimes drivers *are* the traffic themselves 🤷
                      @eliasulrich

                      rdfhrn@hessen.socialR This user is from outside of this forum
                      rdfhrn@hessen.socialR This user is from outside of this forum
                      rdfhrn@hessen.social
                      wrote sidst redigeret af
                      #69

                      @schmidt_fu not in this case https://mastodon.social/@marginalia/117179319967931654 @Devonkiwi @eliasulrich

                      schmidt_fu@mstdn.socialS 1 Reply Last reply
                      0
                      • rdfhrn@hessen.socialR rdfhrn@hessen.social

                        @schmidt_fu not in this case https://mastodon.social/@marginalia/117179319967931654 @Devonkiwi @eliasulrich

                        schmidt_fu@mstdn.socialS This user is from outside of this forum
                        schmidt_fu@mstdn.socialS This user is from outside of this forum
                        schmidt_fu@mstdn.social
                        wrote sidst redigeret af
                        #70

                        @rdfhrn
                        I see.
                        @Devonkiwi @eliasulrich

                        1 Reply Last reply
                        0
                        • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                          "One guy in #Sweden built a #searchengine to fight Google, and it works.

                          It's called #MarginaliaSearch.

                          It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                          What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                          The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                          There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                          It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                          https://marginalia-search.com/

                          harish@hachyderm.ioH This user is from outside of this forum
                          harish@hachyderm.ioH This user is from outside of this forum
                          harish@hachyderm.io
                          wrote sidst redigeret af
                          #71

                          @eliasulrich I love this so, so much. I searched for my own name and found writing of mine from 20 years ago and going down memory lane.

                          1 Reply Last reply
                          0
                          • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                            "One guy in #Sweden built a #searchengine to fight Google, and it works.

                            It's called #MarginaliaSearch.

                            It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                            What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                            The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                            There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                            It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                            https://marginalia-search.com/

                            ve2uwy@mastodon.radioV This user is from outside of this forum
                            ve2uwy@mastodon.radioV This user is from outside of this forum
                            ve2uwy@mastodon.radio
                            wrote sidst redigeret af
                            #72

                            @eliasulrich

                            And it runs on an old UltraSPARC 2 in his hall closet and it runs Solaris 8.

                            Please make this last bit be true.

                            1 Reply Last reply
                            0
                            • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                              "One guy in #Sweden built a #searchengine to fight Google, and it works.

                              It's called #MarginaliaSearch.

                              It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                              What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                              The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                              There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                              It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                              https://marginalia-search.com/

                              csgraves@turtleisland.socialC This user is from outside of this forum
                              csgraves@turtleisland.socialC This user is from outside of this forum
                              csgraves@turtleisland.social
                              wrote sidst redigeret af
                              #73

                              @eliasulrich this is really kind of interesting. Who knew, there was an actual internet out there, if only we cared to look. This makes it much easier.

                              1 Reply Last reply
                              0
                              • tanyakaroli@expressional.socialT tanyakaroli@expressional.social shared this topic
                              • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                                "One guy in #Sweden built a #searchengine to fight Google, and it works.

                                It's called #MarginaliaSearch.

                                It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                                What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                                The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                                There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                                It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                                https://marginalia-search.com/

                                ezrine@piaille.frE This user is from outside of this forum
                                ezrine@piaille.frE This user is from outside of this forum
                                ezrine@piaille.fr
                                wrote sidst redigeret af
                                #74

                                @eliasulrich i swear i cried.

                                1 Reply Last reply
                                0
                                • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                                  "One guy in #Sweden built a #searchengine to fight Google, and it works.

                                  It's called #MarginaliaSearch.

                                  It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                                  What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                                  The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                                  There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                                  It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                                  https://marginalia-search.com/

                                  tiggy@mastodonapp.ukT This user is from outside of this forum
                                  tiggy@mastodonapp.ukT This user is from outside of this forum
                                  tiggy@mastodonapp.uk
                                  wrote sidst redigeret af
                                  #75

                                  @eliasulrich

                                  With just my name, that is fairly common, I was the third entry, about a music award.
                                  Adding my profession didn't lead me to my research and publications but to a snarky piece that references one of my MTIC papers and misrepresents part of our analysis.

                                  1 Reply Last reply
                                  0
                                  • rhube@wandering.shopR rhube@wandering.shop

                                    @64bithero @eliasulrich If it uses Google and an LLM it's part of the problem and not solving anything.

                                    64bithero@mstdn.games6 This user is from outside of this forum
                                    64bithero@mstdn.games6 This user is from outside of this forum
                                    64bithero@mstdn.games
                                    wrote sidst redigeret af
                                    #76

                                    @Rhube @eliasulrich Developing your own crawler can be time consuming and expensive. Bing which (ddg and many others use) is far worse.

                                    Using Google to pull data and then using logic to sort through it isn’t in itself a problem. Heck most SearXng instances I run into still use a Google crawler.

                                    Using an LLM to parse through tons of data can be effective. It’s all about how it’s trained. Websites can be all over the place. Trying to manually code to account for it all will leave huge gaps imo

                                    Overall I’ve been impressed with much of its search results. Plus the code is open source.

                                    rhube@wandering.shopR 1 Reply Last reply
                                    0
                                    • 64bithero@mstdn.games6 64bithero@mstdn.games

                                      @Rhube @eliasulrich Developing your own crawler can be time consuming and expensive. Bing which (ddg and many others use) is far worse.

                                      Using Google to pull data and then using logic to sort through it isn’t in itself a problem. Heck most SearXng instances I run into still use a Google crawler.

                                      Using an LLM to parse through tons of data can be effective. It’s all about how it’s trained. Websites can be all over the place. Trying to manually code to account for it all will leave huge gaps imo

                                      Overall I’ve been impressed with much of its search results. Plus the code is open source.

                                      rhube@wandering.shopR This user is from outside of this forum
                                      rhube@wandering.shopR This user is from outside of this forum
                                      rhube@wandering.shop
                                      wrote sidst redigeret af
                                      #77

                                      @64bithero @eliasulrich Yes it really fucking is a problem. It uses planet-destroying theft machines. It is the precise problem this thread is trying to address. You have completely misunderstood the aim due to your own apparent lack of ethics. Please go talk to some other hacks who are happy destroying the world.

                                      1 Reply Last reply
                                      0
                                      • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                                        "One guy in #Sweden built a #searchengine to fight Google, and it works.

                                        It's called #MarginaliaSearch.

                                        It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                                        What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                                        The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                                        There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                                        It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                                        https://marginalia-search.com/

                                        lunacolon3@blahaj.zoneL This user is from outside of this forum
                                        lunacolon3@blahaj.zoneL This user is from outside of this forum
                                        lunacolon3@blahaj.zone
                                        wrote sidst redigeret af
                                        #78

                                        @eliasulrich@hachyderm.io oh wow, this is actually really really good. i was expecting shitty results but the results are mostly great. the only thing about it i think is weird is that prioritizing text heavy pages means looking up the name of a website will often show a text heavy page from the website before the home page.

                                        1 Reply Last reply
                                        0
                                        • eliasulrich@hachyderm.ioE eliasulrich@hachyderm.io

                                          "One guy in #Sweden built a #searchengine to fight Google, and it works.

                                          It's called #MarginaliaSearch.

                                          It runs its own #crawler and builds its own #index instead of borrowing Bing's. It has no ads, no investors, and no loans.

                                          What it does differently: it ranks for text-heavy, non-commercial pages. Personal blogs. Old university pages.

                                          The weird corners SEO strangled. Every result tells you whether the page uses affiliate links and JavaScript, and you can filter them out.

                                          There's an "explore" mode that just shows you random sites from the index. It's open source under AGPL, so you can host your own copy.

                                          It's keyword-based, so don't type a full question at it. Type two nouns and see where you land. Every #searchengine now shows you the same twelve #monetizedpages."

                                          https://marginalia-search.com/

                                          J This user is from outside of this forum
                                          J This user is from outside of this forum
                                          joetinymind@mastodon.social
                                          wrote sidst redigeret af
                                          #79

                                          @eliasulrich Ran my own Marginalia-style search once — the honest truth is it's not a Google alternative, it's a bookmark. Good for the 40 pages you love, useless for anything new. What I actually missed wasn't another engine, it was the discoverability Marginalia gives a page that would otherwise be unread by anyone. That's the real ask: make small sites findable, don't try to beat Google.

                                          1 Reply Last reply
                                          0
                                          Svar
                                          • Svar som emne
                                          Login for at svare
                                          • Ældste til nyeste
                                          • Nyeste til ældste
                                          • Most Votes


                                          • Log ind

                                          • Har du ikke en konto? Tilmeld

                                          • Login or register to search.
                                          Powered by NodeBB Contributors
                                          Graciously hosted by data.coop
                                          • First post
                                            Last post
                                          0
                                          • Hjem
                                          • Seneste
                                          • Etiketter
                                          • Populære
                                          • Verden
                                          • Bruger
                                          • Grupper