浏览器与网页
浏览器控制、网页自动化、抓取和爬取
Skills 列表

scrapy-web-scraping
Expert guidance for building web scrapers and crawlers using the Scrapy Python framework with best practices for spider development, data extraction, and pipeline management.
mindrally
search
通过 Bright Data CLI 搜索网络 — `bdata search` 用于 Google/Bing/Yandex 的 SERP 结果,`bdata discover` 用于基于意图排序的语义结果。当用户需要 SERP 结果、需要 URL 以进行抓取,或希望进行语义网络发现并可选获取页面内容时使用。一旦选定目标 URL,则转交给 `scrape`;当用户需要来自已知平台的结构化数据时,则转交给 `data-feeds`。需要 Bright Data CLI;主动引导安装和登录(如果缺失)。
brightdata
scrape
通过 Bright Data CLI(`bdata scrape`)将网页内容抓取为干净的 Markdown/HTML/JSON。当用户想要获取页面、从 URL 列表中提取内容或爬取分页列表时使用。对于支持的平台(Amazon、LinkedIn、TikTok、Instagram、YouTube、Reddit 等),转交给 `data-feeds`;当需要先发现 URL 时,转交给 `search`。需要 Bright Data CLI;主动引导安装和登录(如果缺失)。
brightdata
wxt-browser-extensions
WXT browser extension performance optimization guidelines. This skill should be used when writing, reviewing, or refactoring WXT browser extension code to ensure optimal performance patterns. Triggers on tasks involving WXT, browser extensions, content scripts, service workers, messaging, and extension APIs.
pproenca
web-scraping
使用Python工具进行网页抓取和数据提取的专家
mindrally
site-launch-checklist
Pre-launch checklist for shipping a new website. Orchestrates analytics setup (GA4, PostHog, Google Search Console, Ahrefs), legal compliance, security headers and audit, SEO and GEO with keyword research validated against Google Trends (robots.txt, sitemaps, llms.txt, AI policy, schema markup, hreflang), copywriting consistency via a TONE.md and a humanizer pass in the matching language, OpenGraph and social previews, full favicon set with manifest, quality gates (Lighthouse, Core Web Vitals, WCAG accessibility, mobile testing), and setup of a weekly SEO agent. Use this skill whenever the user mentions launching a site/app, deploying a domain to production, pre-launch audit, shipping a marketing/docs/SaaS site or lead magnet, or says "checklist for the site", "ready to ship", "before I go live", "audit before launch", "ready for prod", or asks for a site review.
samber
scrape-webpage
Use this when the page-import pipeline needs to fetch a source webpage and prepare it for import/migration to AEM Edge Delivery Services. Covers scraping content, extracting metadata, downloading images, and returning analysis JSON with paths, metadata, cleaned HTML, and local images. Do not invoke directly — called by page-import as a pipeline step.
adobe
mcp-duckgo
Skills for web search and content scraping via DuckDuckGo MCP Server. Used when users need online searching and web scraping.
aahl
zai-tts
Text-to-speech conversion using GLM-TTS service via the `uvx zai-tts` command for generating audio from text. Use when (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, podcast, driving, cooking). (3) Using pre-cloned voices for speech.
aahl
adspower-browser
AdsPower profile operation via adspower-browser CLI. open/launch/start browser or profile, environment, config profile, AdsPower; create/update/delete/list profiles; groups, tags, proxies; kernel download/list; client patch; API check-status. User phrases like open environment or map to commands such as open-browser.
adspower
logfire-ui
Open or return Logfire project pages, live views, trace links, and Explore pages in the Codex browser without querying telemetry first. Use this skill when the user asks to "open in Logfire", "show in the live view", "open Explore", "open the UI", "show in Codex", "use the browser", "give me a link", or asks for a Logfire GUI/browser/live-view presentation of a project, time range, service, span, trace, log, or filter. If "show" or "view" wording is ambiguous, ask whether the user wants a UI view or query analysis.
pydantic
defuddle
Extract clean article content from web pages or local HTML files. Removes clutter (ads, sidebars, nav) and returns readable content with metadata.
joeseesun
internal-linking-optimizer
当用户要求“修复内链”或“查找孤立页面”时使用;映射链接架构、权威流、锚文本和抓取深度,然后提供优先级的源/目标/锚点计划。不适用于外部反向链接——请使用 backlink-analyzer。内链优化/站内架构
aaron-he-zhu
schema-markup-generator
当用户要求“生成schema”时使用;为FAQ、HowTo、Article、Product和LocalBusiness富结果候选创建JSON-LD。不用于标题/元描述标签——请使用meta-tags-optimizer;不用于爬取/索引技术问题——请使用technical-seo-checker。Schema标记/结构化数据
aaron-he-zhu
alert-manager
当用户要求“设置SEO预警”或“排名掉了提醒我”时使用;配置针对未来排名、流量、技术问题和竞争对手变化的阈值通知。不适用于一次性测量或报告——请使用 rank-tracker 或 performance-reporter。SEO预警/排名监控
aaron-he-zhu
technical-seo-checker
当用户要求“检查技术SEO”时使用;审计可抓取性、索引、核心网页指标、robots.txt、站点地图、规范标签、重定向和迁移。不用于页面标签或内容——请使用on-page-seo-auditor。技术SEO/网站速度
aaron-he-zhu
on-page-seo-auditor
当用户要求“审计页面SEO”或“诊断单页排名下降原因”时使用;对标题、元描述、标题结构、关键词布局、链接和图片进行评分,并提供优先级修复建议。不适用于E-E-A-T/发布就绪评分——请使用content-quality-auditor;不适用于爬虫/CWV/索引——请使用technical-seo-checker。页面SEO审计/排名诊断
aaron-he-zhu
audit-website
使用 squirrelscan CLI 对网站进行 SEO、性能、安全、技术、内容及其他 17 个类别的审计,涵盖 240 多条规则。返回针对 LLM 优化的报告,包含健康评分、损坏链接、元标签分析和可操作建议。用于发现和评估网站或 Web 应用的问题与健康状况。
squirrelscan
agent-browser-automation
Headless browser automation CLI for AI agents using native Rust binary with Chrome DevTools Protocol
reason-machines
lightpanda-browser
Expert skill for Lightpanda — the headless browser built in Zig for AI agents and automation. 9x less memory, 11x faster than Chrome. Installation, CLI, CDP server, Playwright/Puppeteer integration, and web scraping.
reason-machines
opencli-web-automation
Turn any website into a CLI using browser session reuse and AI-powered command discovery
reason-machines
posterskill-academic-posters
AI-assisted academic conference poster generation from Overleaf source using Claude Code
reason-machines
karpathy-jobs-bls-visualizer
Research tool for visually exploring BLS Occupational Outlook Handbook data with an interactive treemap, LLM-powered scoring pipeline, and data scraping/parsing utilities.
reason-machines
firecrawl-lead-gen
使用 Firecrawl 浏览器从潜在客户数据库和网络目录中生成结构化的潜在客户列表。用于按角色、公司类型、行业、阶段、地点、技术或其他标准寻找潜在客户,并导出可直接用于 CRM 的 JSON 或 CSV 文件。
firecrawl