webclaw
Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.
1,618 tools and skills for search tasks
Sign in with GitHub to unlock author info, star counts, source links, and more.
Sign in with GitHubFast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.
Figranium is a self-hosted, free browser automation and scraping platform built on Playwright, designed to handle everything from simple scraping to complex, human-like browser interactions. It provides a visual task editor, a structured JSON task format, and advanced execution modes that go far beyond traditional scrapers.
Turn X scrolling into an AI-powered weekly digest — no X API, no scraping, just your browser. Scans 65+ AI builders, filters actionable content, exports a poster. Built for content creators.
Dual-engine web search Skill for OpenClaw/Pi: Grok AI + Tavily + FireCrawl
CLI, MCP server, and npm library that turns any website into an API — no docs, no SDK, no browser.
Seedance 2.0 motion graphics skill for Claude Code — liquid glass prompting, Fal AI integration, App Store scraping
🇮🇩 50 Indonesian Government APIs & Data Sources — BPS, OJK, BPJPH, BPOM, Bank Indonesia, IDX, BMKG + MCP servers. Python examples, scraping patterns, and practical gotchas.
OpenClaw plugin for multi-provider web search and extraction: 10 search providers, 5 extract providers, intelligent routing, quality reports, and research mode.
Self-hosted web search skill for AI agents (OpenClaw/Claude Code/Antigravity) via SearxNG
Unified search skill for Clawdbot supporting Serper (Google), Tavily (AI-optimized), and Exa (neural semantic) search providers
Multi-provider web search CLI for AI agents — Brave, Serper, Exa, Jina, Firecrawl, Perplexity, xAI in one Rust binary
Self-hosted private web search plugin for OpenClaw using SearXNG
Scrape any X (Twitter) account and distill the person's thinking patterns into a reusable AI persona skill. Powered by Dokobot.
Free, API-key-free MCP web search server — DuckDuckGo, Bing, Google + optional SerpAPI/Tavily. Works with Claude Desktop, Cursor, and any MCP client.
Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions", or any task requiring programmatic web interaction.
🔍 Agent skill that scrapes any social media post — metadata, comments, transcription & AI video analysis. Drop a link, get everything. Instagram · TikTok · Twitter/X · YouTube
GEO CLI for AI search engine visibility
A bot for scraping data from Facebook Marketplace without API
Anspire Search API 为代理提供实时信息搜索能力。
Web extraction engine with antibot bypass. Scrape, crawl, extract, summarize, search, map, diff, monitor, research, and analyze any URL — including Cloudflare-protected sites. Use when you need reliable web content, the built-in web_fetch fails, or you need structured data extraction from web pages.
npm i -g surf-skill — multi-provider (Tavily + Parallel AI) web skill for AI coding agents. CLI + Node library, auto-fallback, multi-key rotation, cross-OS, batch search.
Convert URLs to Obsidian-compatible Markdown files with one command. A Hermes Agent skill.
Codex skill for SEO, GEO, AI crawler access, schema, llms.txt, and citability audits
Tavily AI search API integration for OpenClaw. Provides web search functionality with AI-powered summarization optimized for RAG and question answering. Use when you need web search and the default web_search tool is not configured or you prefer Tavily's search results.
Resolve a query or URL into compact, LLM-ready markdown using a low-cost cascade: llms.txt first, Exa highlights, Tavily fallback, Firecrawl last.
PageRangers SEO API integration for AI assistants - keyword rankings, SERP analysis, KPIs
Capture a session as a reusable agentskills.io skill file for Claude Code, Cursor, Copilot, Gemini CLI, and other agent tools
Professional web scraper: CSS and XPath selectors, proxy rotation, retry, autopagination, JSON/CSV/SQLite export.
An OpenClaw skill that scrapes YouTube video metadata and downloads subtitles — no YouTube API key required.
从 Web of Science 检索学术文献。当用户需要文献综述数据、论文搜索、研究方向的文献收集时使用。接收 search.json 配置文件,执行 WoS 爬虫搜索,输出临时审核表和导入日志。支持关键词搜索、年份过滤、去重、相关性排序、缺失字段补全。调用 scripts/search_wos.py 执行搜索,scripts/import_review_table.py 导入到主表。当用户提到"文献抓取"、"WoS检索"、"论文搜索"、"综述文献"、"文献库更新"等时触发。
Remonter un domaine expiré de A à Z avec Web Resurrect — création projet, enrichissement SEO (Haloscan + Majestic), scraping Wayback Machine, réécriture, catégorisation AI, publication WordPress et redirections. Fonctionne avec la CLI `wr` (@web-resurrect/cli) ET avec le MCP server `web-resurrect`.
Skill from Scrapeclaw/tiktok-scraper
A DuckDuckGo search skill for the Pi coding agents (and others)
Use this skill for all Thyleads outbound work — drafting B2B cold email campaigns aimed at Indian buyers (D2C founders, growth/marketing/product/HR leaders at Indian SaaS, BFSI, EdTech, Health, Travel, MarTech, HRTech, Series A and Series B companies). India-only geography by design. Triggers include any mention of Indian outbound, India SDR campaign, Smartlead campaign for India, cold emails to Indian companies, campaigns targeting India D2C/Ecom/Fintech/MarTech/HRTech ICP, account scoring or lead scoring for Indian B2B, persona-tailored subject lines for Indian CXOs, ICP refinement from campaign reply data, intent signal scoring (hiring, funding, leadership change, growth, expansion, employee count change), behavioral signals (page changes, pricing updates, careers page deltas), competitive intelligence as buying-window indicator, per-lead signal graph as the unified intelligence layer, signal-to-insight engine with the four-stage pipeline (Signal Collection → Signal Selection → Insight Generation → Email Writing), the deterministic signal selection formula, the signal-to-angle library, timing-aware scoring tied to the Indian fiscal calendar, observation diversity and the three-axis observation model (company-level, lead-level, activity-level), self-learning outbound loops, decisions on which tool to use when (Apollo vs Crustdata vs Coresignal vs Tavily vs Visualping), site-targeted Tavily queries for Indian press and conferences, cold-start campaigns where no TAL or account list exists, building target account lists from scratch, finding decision-makers via Apollo + Sales Nav + Coresignal, contact enrichment waterfalls (Apollo → Hunter → LeadMagic → ZeroBounce), or any task involving the Thyleads Dashboard as the orchestration layer. Produces emails that follow India outbound conventions (subject lines that include the prospect's first name and feel like internal-team communication, attention-catching plain-language first sentences with no marketing jargon, observation-based openers across three axes — company-level / lead-level / activity-level — with batch-level diversity enforcement and signal-graph-driven hook selection via the angle library, 5th-grade reading level English with banned corporate-vocabulary list, conversational warmth and empathy markers tuned for Indian buyer register, average 10-12 word sentences with no US-style staccato, generous mobile-first whitespace with 4-6 paragraphs of 1-3 sentences each, 80-130 word bodies for body 1, three-step sequences with axis rotation, social proof from the seller's roster, generous peer-tone CTAs not vendor imperatives, prose constrained by structured insight briefs not free-form drafting), strict prospect-list exclusions including holding-company sister-brand expansion, a five-layer account scoring model (Fit + Intent + Engagement + Why-Now + Timing/Penalty multipliers) tuned per Thyleads product line (Seed-Series A, Series B, MarTech, HRTech), per-client behavioral watcher infrastructure for prospect and competitor page changes, a per-lead signal graph that consolidates Apollo + Tavily + Coresignal + Crustdata + behavioral signals into one weighted view, tool-intelligence decision trees that pick the minimum tool set per task based on project documents, and a closing-the-loop self-learning system that refines ICP, scoring weights, observation angles, observation axes per persona, signal-to-angle mappings, and subject patterns from each campaign's tagged replies.
淘股吧帖子爬取skill,给 AI agent(AI智能体)使用。
Advanced SearXNG Search Skill
云HIS内网系统爬取工具
OpenClaw Skill: Automated job hunting assistant with multi-channel scraping, AI matching, and Feishu/Lark push
earch flight dates and prices on Dohop using a Playwright-based scraper.
Scrape Product Hunt leaderboard and export top products as JSON. Use when the user wants to scrape Product Hunt, export product rankings, or get a list of top products from today, this week, or this month. Also trigger if the user wants trending products analysis, startup leaderboard data, or wants to compare products across time periods on Product Hunt.
Litminer 是一个面向 AI Agent 的科研文献信息获取 skill。它提供的是一层可复用、可追踪、可验证的文献发现和处理底座,让 Claude Code、Codex 等 Agent 不必只依赖通用 WebSearch/WebFetch 来处理科研检索任务。
Unified web search with automatic multi-provider failover for OpenClaw - Brave/Tavily/DuckDuckGo/Serper/SearchAPI
Transform Reddit posts into Excel business intelligence with keyword search, sentiment, competitor, and action-item analysis
Headless Chrome to scrape your X feed using Docker & cookie injection.
知乎抓取.skill,从知乎收藏夹列表到批量正文与图片爬取,支持断点续传,可选自动分类写入Obsidian知识库。
OpenClaw skill for Indeed job & company scraping via Bright Data Web Scraper API
Claude Code skill for scraping Twitter/X timelines and reconstructing threads
The all-in-one AI productivity accelerator. On device and privacy first with no annoying setup or configuration.
mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local
Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)