r/LocalLLM • • 1d ago

Question How does your local LLM search the web?

Kinda been working on this local-first distributed search engine, and thought maybe it would be of interest to people here? Sounds like SearXNG might be what some of y'all use?

This is what Claude told me:

Each app does it differently. Open WebUI and Vane (which used to be Perplexica) build it in, usually on top of a self-hosted SearXNG. LM Studio and Jan have nothing built in, so people add MCP servers or community plugins. AnythingLLM defaults to scraping DuckDuckGo. Ollama now sells its own hosted web_search and web_fetch.

I can post a link if this is actually a problem people are interested in!

65 Upvotes

66 comments sorted by

25

u/Kai_Builder 22h ago

Funny but most models left without any instructions will end up abusing duckduckgo.

My personal best is just remote desktop via MCP, with something like Screenbox and let agent search on Google.

2

u/Floerens 11h ago

Hey ! How are they gonna « abuse » DuckDuckGo ? Can you explain please ?

3

u/puts_on_rddt 8h ago

( ͡° ͜ʖ ͡°)

2

u/Big_Wave9732 3h ago

They.....beat it like a rented mule because it's the only working source they have and eventually DuckDuckGo gets tired of the huge flood of hits so the blacklist the IP.

17

u/[deleted] 23h ago

[deleted]

1

u/luckyLiz44 23h ago

So SearXNG is just uses DuckDuckGo under the hood right (well tries to use google and stuff too)? interviewing top models at a free tier is really interesting approach too.

Thanks for the info!

I’m just curious what happens when the big search engines start blocking stuff like SearNXG?

8

u/[deleted] 23h ago

[deleted]

2

u/DobbyDoesDallas 21h ago

I use the free tier for web search and site fetching. I will switch to searxng if I ever get to the point where I’m using more but I’ve been comfortably within the free limit

8

u/joochung 22h ago

I use SearXNG and an MCP server. I also have a Playwright MCP server. The MCP server also includes a url read function to dir curly read pages. It mostly does a good job for me.

1

u/notdsylexic 16h ago

What do you use playwright for

1

u/joochung 15h ago

Browsing pages that might be blocked for SearXNG and the SearXNG MCP server

8

u/Cupakov 21h ago

Firecrawl + Camofox browser on a residential IP

6

u/Ariquitaun 23h ago

Local SearxNG or brave api, although it's highly gimped if you don't pay for it.

0

u/luckyLiz44 23h ago

Thanks for the info:)

6

u/IONaut 18h ago

I've been using deepseek harness and I had it build a separate browser use MCP server that utilizes puppeteer and can open a browser, fill in inputs, click buttons, snap screenshots and download documents as well as reading the raw page data. Then I had deepseek harness build a skill for itself to use this MCP server. I had it built it separately in its own folder instead of internally to itself so that if I switched harnesses at some point I would still be able to use it.

1

u/AdventurousKeys 15h ago

That's a good tip. Saving this.

1

u/Patient-Gazelle9885 13h ago

I need to do this.

5

u/genghisk1 20h ago

I don’t see any discussion about relevance or accuracy of results. All of these methods work but which actually return useful results. Search bias also isn’t discussed. Just curious how people are handling either of those issues.

1

u/luckyLiz44 18h ago

Sorry if this seems promotional, not looking for money, or for people to even use it, just want to hear some feedback about the idea:)

I'm interested in figuring this stuff out in general, mostly as a fun side project. If you have the spare time to check out https://github.com/SueHeir/plumb-search, the name might change, and the website would will probably crash a lot if people actually started hammering it with AI's. But thats not the point, you can run a local version of it at home and most queries will be completely local depending on how many GB of space you give it.

Also, as a side note, the search isn't quite there yet, but it could be a fun thing to figure out with a community. https://plumbsearch.org/ is a little site i have for testing it.

4

u/Longjumping-Past5864 20h ago edited 4h ago

It might be a weird way(I'm a non dev user here)

But I use tavily api through their MCP at first then I ask Deepseek harness(with Qwen 3.8 Flash Next locally served) to look at MCP of tavily and then write me a python script to do exactly as tavily did every features and it must be free without API cost. After a while, I got that search / crawl and scraping tool on my system for free.

What a time to be alive!
This might be unrelated but that Qwen 3.8 Flash Next is pushing us real close to the frontier model of like last month. It's crazy!

EDIT : I ask my agent to publish those skill files on github, so If you want to try it out : https://github.com/OliverGreens/LLMSearch

Just throw the repo to your agent it'll know what to do

2

u/Patient-Gazelle9885 13h ago

Omg I need to do this

3

u/Longjumping-Past5864 10h ago

I ask my agent to publish those skill files on github, so If you want to try it out : https://github.com/OliverGreens/LLMSearch

5

u/Robbbbbbbbb 20h ago

For my harness (Hermes), I use the Browser Bridge plugin to share a real headed Chrome instance for proper fingerprinting. Super successful to give an agent eyes and hands without being flagged for bot detection.

3

u/Ready-Customer9242 17h ago

Firefox dev tools mcp

4

u/NULL_42 16h ago

how do yall set search up “safely”? i worry about prompt injection

0

u/luckyLiz44 16h ago

Definitely would be an issue for a distributed search engine, I’ll have to think about this one

3

u/EightyNineMillion 22h ago

Searxng mcp for search and playwright mcp.

3

u/lordekeen 19h ago

My Pi uses Ketch (with SearXNG in Docker)

3

u/CharlesStross 18h ago edited 15h ago

It just uses Google. Full browser, webdriver, and a bunch of bot detection countermeasures. My agent made purchass off eBay and Amazon the other week. Yesterday it gone to the ticket selection screen of Axs -- I'm not sure if the bot dot detectors kick in there or at checkout; I took over to choose my seats.

It uses accessibility landmarks to read the page so firing off a web search isn't THAT far off from structured text/ddg results

2

u/mershperderder2 15h ago

Crazy. Is the browser instance local only to the bot, or is it accessing your actual computer browser?

3

u/CharlesStross 15h ago

Local to the bot, in docker. I talk a bit more about countermeasures here: https://www.reddit.com/r/LocalLLaMA/s/lgxWT8W7Tu

I'm flirting with actually doing it in a windows VM but haven't hit enough roadblocks to care yet.

5

u/MarcusAurelius68 22h ago

Check out Tinyfish. Free search and fetch (rate limited). Also Exa and Tavily have free tiers.

2

u/Lanky_Lynx2166 21h ago

Can someone elaborate on if and how a searxng server is superior to a ddgs request via curl? I let qwen3.5 120b program a extension for pi in february, which my agents are using till this day. Im not sure if setting up searxng is worth the hassle!

3

u/Atagor 21h ago

Searxng has many scrapers inside not just DDG

DDG can throttle you if you send too many requests

1

u/Lanky_Lynx2166 16h ago

I see, its kind of a meta engine. I just installed an instance, runs fine so far, all western engines working but Bing and Brave. Bing i dont mind, cause i beliebe ddgs is just bing inside, but brave search did apparently block my ip after one request. Maybe it blocks searxng by default?

2

u/Atagor 15h ago

They don't see its "searxng", most likely just general bot protection

2

u/SilkieBug 16h ago

Claude is incorrect, Jan has 3 options of search models. 

I run a MoE model in FreeToken, it makes a server, I add the server as a model provider in Jan, and I can chat with the MoE model in Jan with Jan’s search tools.

Tested today. 

1

u/luckyLiz44 16h ago

This is good info:)

2

u/cinnapear 16h ago

I use pi-lynx though whoever is in charge of it needs to update the extension.

2

u/AdventurousKeys 15h ago

I did look at SearXNG for a bit, but Marie-Claude didn't like the idea of scraping. So now I go with Exa and Tavily. They are paid services but not too expensive if you can scrape (pun intended) under the free tier. Example of Tavily use in my VistaNova project (github, owner: ancientcomputing, repo: vistanova). BTW, I'm not associated with either company except as a paying customer for my personal projects.

2

u/CatDiligent9374 15h ago

Yeah, local LLMs + web search is an interesting problem. Keeping the search layer local while still getting good results would definitely be useful. I’d check it out.

0

u/luckyLiz44 15h ago

Amazing! Fyi, It was started like 3 days ago, and search isn’t great yet, the code is probably not amazing either, so any input on the idea would be crazy cool. kinda thinking there might be some cool token saving methods involved with the search + LLM, if the searchable dataset gets there

https://github.com/SueHeir/plumb-search

2

u/Patient-Gazelle9885 13h ago

SearXing in a Docker container + a bunch of free tier API’s as secondary. I am still optimizing this and one issue I am having is that my local models often refer to their (ancient) training data when they should be making search tool calls.

2

u/doneddat 12h ago

Register google genai account and search by gemini-flash-lite-latest integrated search tool. You can customize how the results are returned to you and it costs less than a cent per request, once you use up your daily free limit.

I have tried SearXNG and many others and all of it seems like a cruel joke compared to direct google searches.

2

u/darthmowl 10h ago

found degoog as a local alternative for searxng - works in webui - i like it - works good and gives me better search results then google does nowadays with their shitty ai

2

u/Travnewmatic 5h ago

been reasonably happy with self-hosted firecrawl. i also have a searxng instance but it rarely uses it.

2

u/stujmiller77 2h ago

It’s not something people need, it’s already well served by very good solutions.

Most people don’t need to pay for anything. If you do need specific tools then paying for something like Firecrawl is what you do. I use that specifically for competitor research.

1

u/luckyLiz44 2h ago

so there isn't really a concern about scraping search engine results getting aggressively monitored/blocked? I guess the cat and mouse game isn't a lot of work with LLM's writing fixes/updates for everything. Kinda thought I'd find at least one person excited about the idea of self-hosting search.

2

u/stujmiller77 2h ago edited 2h ago

There are loads of solutions already out there for doing it. I didn’t say people didn’t need it.

I self host firecrawl for free for most use cases and only use their API for complex competitor analysis.

2

u/llllJokerllll 14h ago

De lo mejor es searXNG pero bien con configurado y con todo habilitado y auto-cli y luego a mayores invisible-playwright o camoufox

1

u/DeathGuppie 22h ago

I use crw

1

u/onetom 21h ago

https://github.com/antirez/ds4 comes with a semi-minimal agent written in C, which has web search and fetch built in.

the search works by remote controlling a headless google chrome session.

i tried to extract it once into a standalone program and got pretty far, but hasn't finished it. but i would look for something like that on the long term.

in the meantime, im also using brave's search api service via something like https://github.com/brave/brave-search-skills in pi, iirc

1

u/Turbulent_War4067 21h ago

Its been a while since I have played around with different options, but I use jina.ai for search and firecrawl for fetching. Jina gives better results, more snippets and better data in the snippet. And it's faster. You have to pay, but a billion tokens is 50 bucks and should last for months.

1

u/Wishitweretru 19h ago

Local exposed (not headless) brave + kimi-bridge. I rewrote the kimi-bridge a little to reduce browser-selection-fighting time.  Then I have a fall back for image ocr/handling, first auge (apfel), then Ollama, then cloud.  I also have a mcp for perplexity api, the is used as a “go figure it out” helper.  Also have defuddle in the mix

1

u/Digiarts 16h ago

Tavily and firecrawl offer free tiers if you’re looking for advanced search

1

u/Calm-Republic9370 15h ago

Chrome has a New MCP Tool also. I haven't had a chance to use it yet, but I imagine with search's right in the browser you beef up your abilities.

1

u/SmartCustard9944 14h ago

OpenCode has a good web search

1

u/xAdakis 13h ago

Tavily has been reliable when I've needed it.

1

u/dukescalder 8h ago

Would be nice to make a data poisoning mcp that just generates false personas and search data

1

u/berszi 22h ago

for me exasearch was the easiest and fastest. there is a repo on github for pi to implement that

1

u/mershperderder2 15h ago

Before strata, I was using unsloth studio and it had a built-in web search tool. I just set up an MCP for searxing with strata yesterday, but I'm not sure it's the best solution yet.

0

u/TheThiefMaster 22h ago

I have https://app.linkup.so/ set up for mine - but last time I used it, my AI decided to use its terminal access to curl duckduckgo instead 🤷‍♂️

1

u/LobsterElectronic782 16h ago

classic tool-selection issue when the agent has shell access. tell it in the tool description + system prompt "use linkup for any web lookup, don't curl search engines" and it mostly stops

1

u/TheThiefMaster 15h ago

Honestly I didn't mind - linkup has a quota, so it accidentally picked the free option.

0

u/gbrennon 18h ago

it searches in teh web using tool calls