Sosse

Sosse

Selenium search engine for crawling and archiving

410GitHub
Description

Discover Sosse — the Selenium Open Source Search Engine built for powerful web archiving, crawling, and search. Explore all its features and capabilities on the official website. default user password: admin / admin (Admin: Please change your password immediately) Key Features 🌍 Web Page Search: Search the content of web pages, including dynamically rendered ones, with advanced queries. 🕑 Recurring Crawling: Crawl pages at fixed intervals or adapt the rate based on content changes. 🔖 Web Page Archiving: Archive HTML content, adjust links for local use, download required assets, and support dynamic content. 🏷️ Tags: Organize and filter crawled or archived pages using tags for better search and management. 📂 File Downloads: Batch download binary files from web pages. 📡 Webhooks: Integrate with external services using highly flexible webhooks. Connect to proprietary AI platforms \(doc\) or locally hosted solutions to enable advanced data extraction, summarization, auto-tagging, notifications, and more. 🔔 Atom Feeds: Generate content feeds for websites that don’t have them, or receive updates when a new page containing a keyword is published. 🔒 Authentication: The crawler can authenticate to access private pages and retrieve content. 👥 Permissions: Admins can configure crawlers and view statistics, while authenticated users can search or do so anonymously. 👤 Search Features: Includes private search history \(doc\), and external search engine shortcuts , etc.

Screenshots
Screenshot 1
Screenshot 2
Screenshot 3
Screenshot 4
Screenshot 5
Screenshot 6
Screenshot 7
Mobile Screenshots
Mobile Screenshot 1
Mobile Screenshot 2
Mobile Screenshot 3
Mobile Screenshot 4
App Information
Version
0.0.1
Package Size
5.7 KB
Image Size
1.39 GB
Updated
October 10, 2025
Source Code
biolds
Platform Support
PCMobile
Keywords
Sossesosse搜索引擎Selenium抓取网页归档网页