
sitepanda
by hokupod
Sitepanda is designed to scrape websites using the headless browser. The primary goal is to extract the main readable content from web pages and save it as Markdown.
SKILL.md
name: sitepanda description: > Scrape websites with a headless browser and extract main readable content as Markdown. Use this skill when the user asks to retrieve, analyze, or summarize content from a URL or website.
Sitepanda (Web Scraping Tool)
Instructions
-
When the user provides a URL or asks for website content, use Sitepanda to scrape the page.
-
By default, use the following command to scrape a single page:
sitepanda scrape --silent --limit 1
-
If you need to perform recursive scraping (following links), you must ask the user for confirmation before starting, as it may take a long time.
-
Capture the output, which is returned in Markdown format.
-
Read and analyze the extracted content.
-
Respond to the user using only the relevant information from the page.
-
If the content is long, summarize or extract only the necessary sections.
Examples
Example 1
User request: "Please summarize the article at https://example.com/blog/post-123"
Agent behavior:
- Use Sitepanda to scrape the page
- Read the extracted Markdown
- Summarize the main points in the response
Example 2
User request: "What does this documentation page say? https://example.com/docs"
Agent behavior:
- Fetch the page using Sitepanda
- Extract key sections
- Explain the content concisely
スコア
総合スコア
リポジトリの品質指標に基づく評価
SKILL.mdファイルが含まれている
ライセンスが設定されている
100文字以上の説明がある
GitHub Stars 100以上
3ヶ月以内に更新がある
10回以上フォークされている
オープンIssueが50未満
プログラミング言語が設定されている
1つ以上のタグが設定されている
レビュー
レビュー機能は近日公開予定です