Dev & Engineering documentation-scrapinggithub-analysispdf-extractionconflict-detectionmcp-serverast-parsingvector-database

Skill Seekers Skill Builder

Turn documentation, repositories, and PDFs into AI-ready skills while detecting conflicts between documentation and implementation.

FollowSkills review · FSRS-2.0
Not recommended
48/ 100 5-point scale 2.4 / 5
Trust10 / 25 · 2.0/5

The skill names scraping, enhancement, upload, and installation tools that may create external effects, but it requires no user confirmation and does not disclose detailed data flows, sensitive-data handling, permission boundaries, rollback, or pre-upload review. The README mentions API keys and external services but does not complete these controls, so points are deducted. No malware, credential theft, covert exfiltration, or destructive default was evidenced.

Reliability8 / 20 · 2.0/5

SKILL.md gives a source-type routing table and staged workflow, while the repository includes test configuration and Docker test workflows, providing some reproducibility signals. However, tool availability, required parameters, abnormal-input behavior, diagnostic errors, and recovery paths are unspecified. The claim of 40 MCP tools also conflicts with the Docker workflow's 25-tool description, limiting the score under static review.

Adaptability10 / 15 · 3.3/5

Audience, use cases, input patterns, and major source types are clear, covering documentation, GitHub, PDFs, videos, and local directories; the repository also declares Simplified Chinese documentation and international support. Non-fit boundaries, trigger precision, output contracts, and mainland-China network reachability are not established, and the core flow may depend on external sites, MCP, and LLM services, so points are deducted.

Convention8 / 15 · 2.7/5

The skill has concise metadata, trigger guidance, source detection, and a staged workflow. Repository context supplies an MIT license, version metadata, README material, and testing/CI signals. The skill itself lacks installation details, parameter specifications, output examples, FAQs, known limitations, a changelog, and explicit maintenance ownership/update procedure; the inconsistent tool-count description also reduces maintainability.

Effectiveness7 / 15 · 2.3/5

The workflow covers detection, scraping, enhancement, packaging, and vector-database export, and README commands support the claimed core use case. Static evidence does not verify generated-skill completeness, directly usable output formats, conflict-detection accuracy, or actual benefit over manual alternatives. The score therefore uses the static-review ceiling and deducts for missing outcome verification.

Verifiability5 / 10 · 2.5/5

Committed test directories, pytest configuration, and some real Docker CI checks provide auditable evidence. The supplied material does not show that these tests cover the skill-builder MCP paths, nor does it provide third-party execution evidence or independent corroboration. Under the static calibration, the score cannot exceed 5.

Evidence confidence:Low Reviewed Jul 20, 2026 Reviewed revision d3f9b2c93779
Before you use it
  • Before scraping, enhancing, uploading, or installing, require explicit confirmation of scope, destinations, network requests, transmitted data, and credentials.
  • Do not treat the README's '3700+ tests' or performance/quality claims as validation of this skill path; add tests covering source routing, conflict detection, and failure paths.
  • The MCP tool count is inconsistent between 40 and 25; verify the actual server capabilities, dependency versions, and reachability of required external services from mainland-China networks.
See the full review method →

What it does & when to use it

Skill Builder is the skill-builder component of the Skill Seekers repository. It creates AI skills from documentation sites, GitHub repositories, PDFs, videos, codebases, and other knowledge sources. It operates through the Skill Seekers MCP server, which provides scraping, analysis, enhancement, packaging, and export tools. It is intended for developers and AI engineers who need structured, reusable context from technical source material.

Detects the source type from a user request and selects tools such as scrape_docs, scrape_github, scrape_pdf, scrape_video, scrape_codebase, or scrape_generic. It can generate and validate configurations, estimate documentation scope, enhance generated skills, package them for target platforms, and export content to Weaviate, Chroma, FAISS, or Qdrant. GitHub workflows can perform AST analysis, API extraction, and conflict detection between documentation and implementation.

  1. A developer wants to turn a framework's documentation into an AI coding skill.
  2. A team needs to analyze a GitHub repository and compare documented APIs with the actual implementation.
  3. A researcher wants to convert a PDF manual or other knowledge source into structured LLM context.
  4. An AI engineer needs to export processed knowledge to Chroma, FAISS, Qdrant, or Weaviate.

Pros & cons

Pros
  • Supports documentation websites, GitHub repositories, PDFs, videos, codebases, and other source types.
  • Provides AST analysis, API extraction, and documentation-versus-code conflict detection.
  • Supports AI enhancement, platform packaging, and multiple vector-database export targets.
  • Exposes the workflow through natural-language MCP tool calls.
Limitations
  • Requires installation and configuration of the Skill Seekers MCP server for the documented workflow.
  • Some source types require optional dependencies, including video, Jupyter, PPTX, and Notion support.
  • Online scraping requires network access and GitHub workflows may encounter API rate limits.
  • The supplied SKILL.md does not document the MCP server endpoint, authentication flow, or complete parameters for every tool.

How to install

Install Skill Seekers with pip install skill-seekers. For the MCP server, install pip install skill-seekers[mcp] and run python -m skill_seekers.mcp.server_fastmcp; the repository also provides ./setup_mcp.sh for agent configuration. The repository requires Python 3.10 or newer.

How to use

In a compatible client, make a request such as Create an AI skill from facebook/react, Convert this PDF into an LLM-ready skill, or Update the existing skill and detect documentation/code conflicts. The skill selects the appropriate scraper, after which you can request enhancement, packaging, or vector-database export.

FAQ

Does this require paid AI APIs?
The source does not impose one universal cost model. Enhancement or upload workflows may require provider API keys, while the README also documents local-agent enhancement options.
Can it process local files?
Yes. The skill explicitly supports local directories, PDFs, videos, and several other file formats, although some formats require optional dependencies.
Can it find discrepancies between documentation and code?
Yes. GitHub and unified multi-source workflows can report missing implementations, undocumented features, signature mismatches, and description mismatches.

More skills from this repository

All from yusufkaraaslan/Skill_Seekers

Dev & Engineering

Man Page Single-Page Lookup

Quickly find curl options, syntax, examples, and reference links from one generated man page.

Dev & Engineering

Golden Man Command Reference

A compact reference for golden_man Git and Curl commands, options, examples, and cross-references.

Writing & Content

Quiet Feed Empty Feed Skill

Helps agents inspect an Atom feed that currently has no articles.

Dev & Engineering

Skill Seekers Bootstrap

Self-document Skill Seekers as an installable Claude Code skill.

Dev & Engineering

Skill Seekers Builder

Turn documentation, repositories, and other knowledge sources into usable AI skills.

Dev & Engineering

Web Handbook Documentation Skill

Helps agents test and navigate multi-file HTML documentation builds.

Dev & Engineering

Golden EPUB Keyword Classifier

Tests documentation keyword categorization with a compact structural and example summary.

Dev & Engineering

PDF Keyword Categorization Test

A focused skill for validating keyword categorization, document sections, and extracted examples from PDF documentation.

Data & Analysis

Golden Jupyter Notebook Skill

A focused reference for testing, reproducing, and reviewing the Golden_Jupyter analysis notebook.

Dev & Engineering

EPUB Golden Build Guide

A focused reference skill for testing EPUB golden builds and locating document concepts, APIs, examples, and troubleshooting notes.

Dev & Engineering

RSS Engineering Blog Reference

A fixed RSS example for checking feed metadata, articles, tags, and navigation output.

Data & Analysis

Golden Jupyter Directory Notebook Skill

A compact reference for understanding, reproducing, and reviewing Jupyter notebook analysis workflows.

Dev & Engineering

Golden PPTX Keyword Categorization

A focused reference skill for testing keyword categorization across structured presentation content.

Data & Analysis

Jupyter Topics Build Test Skill

A focused skill for validating the golden_jupyter_topics notebook build and its analysis references.

Dev & Engineering

PDF Chapter Classification Skill

Helps agents understand and navigate PDF content by chapter.

Dev & Engineering

Golden Word Documentation Skill

A structured reference skill for testing the Word golden build.

Dev & Engineering

Keyword Categorization Test Docs

A structured HTML documentation reference for testing keyword categorization.

Productivity & Collaboration

Empty Slack Chat Skill

A test skill for validating an empty Slack chat export.

Dev & Engineering

Discord Chat Knowledge Skill

Find solutions, code references, and team decisions in exported Discord conversations.

Writing & Content

Web Handbook Documentation Skill

Quickly navigate structured web documentation for concepts, API references, examples, and troubleshooting.

Related skills