Skip to content

Quickstart

The core loop: search → read → store → query. Every command below works immediately after go install.

Terminal window
webx search "concurrent map writes golang panic"

Fused metasearch over 15 providers — weighted RRF + BM25 blending + domain priors. score is normalized 0–1, near-duplicates are dropped (deduped in output). Check provider health anytime with webx doctor.

Terminal window
webx scrape https://go.dev/doc/effective_go

Clean markdown + metadata through the readability→trafilatura→full-page cascade. For JS-heavy pages:

Terminal window
webx scrape https://spa.example.com --auto-render # escalates http → browser automatically
webx scrape https://go.dev --fit "http server" # BM25-filtered markdown — the token-saver
Terminal window
webx ask "how does htmx hx-swap work"

Search → scrape top results → citation-numbered excerpts under a token budget. Add --llm for a synthesized cited answer (needs WEBX_LLM_*).

Terminal window
webx index go.dev
webx query "error handling patterns" # lexical FTS5, offline
webx query --semantic "graceful shutdown" # + embeddings (WEBX_EMBED_MODEL)

index crawls the sitemap (or BFS) into ~/.webx/index.db. --semantic fuses bm25 with chunk-level embedding cosine — Exa-style neural search over your corpus.

Terminal window
webx serve # HTTP API on :8080
curl -X POST localhost:8080/scrape \
-H 'content-type: application/json' \
-d '{"url":"https://go.dev/doc/effective_go"}'

Firecrawl /v1//v2 routes are built in — point a Firecrawl SDK at the same base URL and it works. Full surface: HTTP API.