百科.dev
全部条目AI 编程趋势榜开源项目技术资讯提交条目
登录
< 返回工具列表
H

HeadlessX

> 数据库
开源

未被发现的自主托管浏览器自动化平台。由 Camoufox (Firefox) 提供支持,检测率为 0%。专为速度、隐私和可扩展性而构建。

2.1K stars0 点赞0 次浏览
访问官网GitHub

工具介绍

未被发现的自主托管浏览器自动化平台。由 Camoufox (Firefox) 提供支持,检测率为 0%。专为速度、隐私和可扩展性而构建。

--- ## What's next (v2.5) I'm working on HeadlessX v2.5. I know the current dashboard and a lot of the system need a real redesign — cleaner UI, better UX, and a setup that actually fits how people use this with agents. That's what I'm building now. v2.1.2 is still what you run today. v2.5 will replace the playground-style UI and add a proper agent chat that calls operators. Live agentic browser sessions and taught recipes stay in **Helium CLI** — HeadlessX will not ship Interact or YAML workflows. Headed operator work (Google cookie setup, later logins) will still embed **KasmVNC** in the dashboard so keyboard and mouse go into the real browser window. Postgres will ship in Docker only (no Supabase). Platform install will be **`headlessx-setup`**: pick self-host or VPS, auto-detect OS (self-host on Linux/macOS/Windows; VPS Linux only), check Docker/npm/Bun, then on VPS validate DNS before start and print the Dashboard/API links. There is no developer mode. **v2.5 launches [Hexium](https://github.com/HeadlessXLabs/hexium-browser)** (`hexium-browser` / Playwright `launch()` → Chrome 151). Not Camoufox. Not CloakBrowser. v2.5 will live under **[HeadlessXLabs](https://github.com/HeadlessXLabs)**. This repo stays **v2.1** on [@saifyxpro](https://github.com/saifyxpro) until that ship. Watch [HeadlessXLabs](https://github.com/HeadlessXLabs) and the releases here until the move. ### Inspired by ## Overview HeadlessX is a self-hosted scraping platform with a web dashboard, protected API, queue-backed workflows, and a remote MCP endpoint. Current live operator surfaces: - Website operator: scrape, crawl, map, content extraction, screenshots - Google AI Search - Tavily - Exa - YouTube - Queue jobs, logs, API keys, proxy management, and config management - Remote MCP over `/mcp` Important operator setup notes: - Google AI Search requires a one-time `Build Cookies` run in the dashboard before the first search - the saved Google session is kept in the shared persistent browser profile and reused later - the YouTube workspace is active only when `YT_ENGINE_URL` points at a healthy `yt-engine` service ## What Changed In v2.1.2 - Added the published HeadlessX CLI bootstrap flow with `headlessx init`, `start`, `logs`, `stop`, `restart`, `status`, and `doctor` - Upgraded the CLI prompt UX with guided modern setup and login prompts - Added Docker plus Caddy production domain scaffolding under `infra/domain-setup` - Moved local and Docker host defaults to rarer ports to avoid conflicts with common `3000` and `8000` stacks - Refreshed setup, CLI, and self-hosting docs around the current operator-first platform layout ## Sponsors I'm thankful to the people and companies who support HeadlessX. BirdProxies, Swiftproxy, NodeMaven, Mango Proxy, and Thordata have been solid partners — that support is a big part of why this project can keep moving. If you need proxies for scraping or automation, start with them. View
Hey, we built BirdProxies because proxies shouldn't be complicated or overpriced. Fast residential and ISP proxies in 195+ locations, fair pricing, and real support. Try our FlappyBird game on the landing page for free data!

Try BirdProxies now  |  Join the Discord

Swiftproxy — Reliable residential proxies optimized for HeadlessX automation and anti-bot workflows. Access 80M+ rotating residential IPs across 190+ countries with non-expiring traffic, high anonymity, sticky sessions, and free trials. Get 10% off with code PROXY90.

Try Swiftproxy now  |  Contact us

NodeMaven — The most efficient proxy provider for Web Scraping and Automation with the Highest Quality IP on the market.

Why NodeMaven?
  • ZIP targeting
  • 99.9% uptime
  • IP filtering: all proxies have fraud score <97%
  • No KYC required
  • Unique free tools: Proxy Bandwidth Checker, Meta Tag Checker, IP Lookup and others!
Special codes for HeadlessX users:
HEADLESSX35 — 35% off Mobile and Residential Proxies
HEADLESSX40 — 40% off ISP (Static) Proxies

Try NodeMaven now

Mango Proxy — Residential, ISP, datacenter, and mobile proxies for automation, web scraping, SEO, social media management, and other online tasks. HTTP(S) and SOCKS5, rotating and static. Use SAIFYXPRO for 8% off Static ISP Proxies.

Try Mango Proxy now

Thordata — Premium residential proxies with 100M+ real IPs across 195+ countries, geo targeting, 99.99% uptime, and rotating & sticky sessions. Built for headless browsers, scraping tools, and AI workflows so your HeadlessX jobs look like real users. 3-day free trial · 10% off.

Try Thordata now
## Operators ### Coming Soon | Operator | Description | Status | | --- | --- | --- | | Google Maps | Extract business listings, reviews, categories, ratings, contact details, opening hours, and location metadata from Google Maps search results. | Planned | | Twitter / X | Capture profiles, posts, engagement metrics, media, hashtags, and conversation threads from public X pages. | Planned | | LinkedIn | Extract public company and profile data, role details, locations, website links, and business metadata from LinkedIn surfaces. | Planned | | Instagram | Collect public profile data, captions, post metadata, media links, reels references, and engagement signals. | Planned | | Amazon | Extract product listings, seller data, pricing, ratings, reviews, availability, and catalog metadata from Amazon pages. | Planned | | Facebook | Capture public page data, posts, about fields, links, follower counts, and engagement metadata from Facebook pages. | Planned | | Reddit | Extract subreddit, post, comment, author, score, flair, and discussion metadata from Reddit threads and listings. | Planned | | ThomasNet Suppliers Real-Time Scraper | Extract 70+ ThomasNet supplier fields including emails, phone numbers, company data, products, locations, certifications, and more. | Planned | | TLS Appointment Booker | Automate TLS appointment availability checks and booking workflows with support for high-frequency monitoring and retry-safe session handling. | Planned | | GlobalSpec Suppliers Scraper | Extract 200,000+ industrial supplier profiles from GlobalSpec Engineering360 with contact data, business type, product catalogs, specs, and datasheets. | Planned | | ImportYeti Scraper | Extract supplier profiles, shipment records, and trade data from ImportYeti with 60+ fields including HS codes, shipping lanes, carriers, bills of lading, trading partners, and contact info. | Planned | | MakersRow Scraper | Extract 11,600+ US manufacturer profiles from MakersRow with email, phone, address, website, GPS coordinates, capabilities, ratings, gallery images, and business hours. | Planned | ## Agent Surfaces Coming Soon | Surface | Description | Status | | --- | --- | --- | | Web AI Agent (`/web`) | Interactive AI agent workspace inside the dashboard that can use all HeadlessX operators and related workflow actions, including Website, Google AI Search, Tavily, Exa, and YouTube. | Planned | ## Agent Skills You can add the HeadlessX CLI skill to AI coding agents such as Cursor, Claude Code, Warp, Windsurf, OpenCode, OpenClaw, Antigravity, and similar tools that support the `skills` installer flow. ```bash npx skills add https://github.com/saifyxpro/HeadlessX --skill cli ``` This installs the HeadlessX CLI skill from this repository so the agent can use the published `headlessx` command and follow the packaged usage guidance. ## UI Screenshots ### Google AI Search (Recently Tested with Arabic Lang & Region) ### Website ## Proof ### BrowserScan Cloudflare Challenge
### Pixelscan ### Proxy Validation
## Quick Start ### System Requirements | Item | Minimum | Recommended | | --- | --- | --- | | OS | macOS, Linux, or Windows 11 with WSL2 | Ubuntu 22.04+/24.04, Debian 12, or Windows 11 with WSL2 | | CPU | 2 cores | 4+ cores | | RAM | 4 GB | 8-16 GB | | Disk | 10 GB free | 20+ GB SSD | | Network | outbound internet for installs, browser downloads, and APIs | stable broadband | | Docker | Docker Engine + Compose v2 | current stable | ### Mode requirements **Developer** — apps run on the host; **Docker is required**. Postgres and Redis run in Docker Compose. You still need Node.js, pnpm, and Git for `apps/web`, `apps/api`, and the worker. **Production** — the full stack runs in Docker, plus a domain layer (Caddy). You need: - a public server / VPS with Docker + Compose v2 - two hostnames, for example `dashboard.yourdomain.com` and `api.yourdomain.com` - DNS `A` / `AAAA` records pointing both names at the server - ports `80` and `443` open (Caddy issues TLS) - an email for Let's Encrypt (`--caddy-email`) ```bash headlessx init --mode developer headlessx init --mode production --api-domain api.example.com --web-domain dashboard.example.com --caddy-email [email protected] ``` ### Runtime Dependencies - Node.js 22+ and pnpm 10.32.1+ (developer mode; apps on the host) - Git - Docker + Compose v2 (**required for developer and production**) - Python/uv for `yt-engine` - Go for the HTML-to-Markdown sidecar (v2.1.2) PostgreSQL and Redis are not a separate install in either mode — they come from Docker. If your machine does not already use the pinned pnpm release, align it with: ```bash corepack enable corepack use [email protected] ``` ### Practical Sizing No

Issues· 0 开放

查看全部 Issues在 GitHub 打开

暂无开放 Issues,或尚未同步最近议题。

> 标签

TypeScriptautomationautomation-apiautomation-platformbrowser-automation

暂无评论,来聊聊你的看法吧

> 工具信息

发布日期2026年8月1日
最后更新2026年9月17日
分类数据库
定价开源

> 相关工具

P
PostgreSQL
功能强大的开源关系型数据库
R
Redis
内存数据结构存储,常用作缓存与队列
M
MySQL
广泛使用的开源关系型数据库