Discover live websites across TLDs using pure random dictionary words or dynamic semantic themes (Space, Radio, Tech, etc.) with custom skip filters and starred favorites.
.com, .org, .io (3 selected)โผ
Configure parameters above and click 'Run Domain Scanner' to begin.
๐Viewing Historical Snapshot
๐ Seed Words Generated:
Showing all verified domainsTotal Quality Results: 0
No Starred Favorites Yet
Click the star (โ) on any verified domain card to save it here permanently.
Historical domain scans stored in SQLite (0 snapshots)
No History Snapshots Yet
Every completed domain scan is automatically saved here.
โ๏ธ Custom Text Exclusion Filters
Any candidate domain containing these exact phrases in its title, URL, snippet, or body text will be automatically skipped by the scanner.
๐ System Overview & Architecture
Spuggles Domain Hunter is a high-speed asynchronous domain exploration and content intelligence engine. It combines real-time lexicon APIs with a concurrent Python FastAPI backend and a persistent SQLite database to probe, verify, filter, and archive live websites across global top-level domains.
6. Persistent Storage๐พ SQLite Database (/app/data/urlhunter_data.db)Multi-device synchronization for historical scan snapshots, starred favorites, and custom skip rules.
โก Core Engine Components
๐ฒ Dual-Mode Seed GenerationSupports pure random dictionary lookups with exact letter-length constraints (3Lโ8L) and 6 languages, or dynamic semantic topic sampling via the Datamuse Lexicon API that generates fresh vocabulary on every run.
๐ก High-Throughput Async SocketsProbes candidate domains in parallel batches of 35 socket connections using asyncio.gather and connection pools up to 100 concurrent workers with zero browser CORS proxy bottlenecks.
โฑ๏ธ Fast-Fail Timeouts & BudgetingNon-existent and timed-out domains fail fast (connect=1.2s). A strict 18-second time budget guarantees the engine returns quality results immediately without hanging.
๐ก๏ธ Content Extraction & VerificationExtracts page titles, meta descriptions, and clean body text with BeautifulSoup. Strips scripts, styles, headers, and navigation menus to evaluate genuine text content.
๐ก๏ธ Quality & Verification Checklist
HTTP 200 OK Check: Rejects timed-out, unregistered, connection-refused, and HTTP error responses (404, 500, 502).
HTML Validation: Ensures the server returns valid text/html content.
Quality Threshold: Automatically discards blank or empty placeholder pages with fewer than 15 characters of real body text.
Active Exclusion Rules: Evaluates site content in real time against your enabled skip filters, automatically rejecting domain parking, resale, and placeholder pages.
๐พ Persistence & Data Architecture
All history snapshots, starred favorites, and custom exclusion filters are permanently persisted in a dedicated SQLite database (/app/data/urlhunter_data.db) mounted via persistent Docker volumes.
Snapshots Table: Records timestamp, mode, topic, word length, language, seed words used, TLDs searched, and full result payloads.
Favorites Table: Persists bookmarked domains with titles, URLs, and preview snippets.
Filters Table: Stores active and disabled text exclusion rules with instant CRUD access.
๐ TLD Intelligence & Discovery Guide
Curated dossier on all top-level domains available in Spuggles Domain Hunter.