All skills
jimliu avatar

/baoyu-url-to-markdown

@d6442e1 official
by Jim Liu 宝玉jimliu/baoyu-skills26k stars
2,896

Fetch any URL and convert to markdown using baoyu-fetch CLI (Chrome CDP with site-specific adapters). Built-in adapters for X/Twitter, YouTube transcripts, Hacker News threads, and generic pages via Defuddle. Handles login/CAPTCHA via interaction wait modes. Use when user wants to save a webpage as markdown.

Use this Skill: https://skilld.dev/gh/jimliu/baoyu-skills/baoyu-url-to-markdown

This session only. Nothing lands on disk.

referencesconfigfirst-time-setup.md

≈615 tokens on demand. Your agent reads this file only when SKILL.md points to it.

First-Time Setup

Overview

When no EXTEND.md is found, guide user through preference setup.

BLOCKING OPERATION: This setup MUST complete before ANY other workflow steps. Do NOT:

  • Start converting URLs
  • Ask about URLs or output paths
  • Proceed to any conversion

ONLY ask the questions in this setup flow, save EXTEND.md, then continue.

Setup Flow

No EXTEND.md found
        |
        v
+---------------------+
| AskUserQuestion     |
| (all questions)     |
+---------------------+
        |
        v
+---------------------+
| Create EXTEND.md    |
+---------------------+
        |
        v
    Continue conversion

Questions

Language: Use user's input language or saved language preference.

Use AskUserQuestion with ALL questions in ONE call:

Question 1: Download Media

header: "Media"
question: "How to handle images and videos in pages?"
options:
  - label: "Ask each time (Recommended)"
    description: "After saving markdown, ask whether to download media"
  - label: "Always download"
    description: "Always download media to local imgs/ and videos/ directories"
  - label: "Never download"
    description: "Keep original remote URLs in markdown"

Question 2: Default Output Directory

header: "Output"
question: "Default output directory?"
options:
  - label: "url-to-markdown (Recommended)"
    description: "Save to ./url-to-markdown/{domain}/{slug}.md"

Note: User will likely choose "Other" to type a custom path.

Question 3: Save Location

header: "Save"
question: "Where to save preferences?"
options:
  - label: "User (Recommended)"
    description: "~/.baoyu-skills/ (all projects)"
  - label: "Project"
    description: ".baoyu-skills/ (this project only)"

Save Locations

Choice Path Scope
User ~/.baoyu-skills/baoyu-url-to-markdown/EXTEND.md All projects
Project .baoyu-skills/baoyu-url-to-markdown/EXTEND.md Current project

After Setup

  1. Create directory if needed
  2. Write EXTEND.md
  3. Confirm: "Preferences saved to [path]"
  4. Continue with conversion using saved preferences

EXTEND.md Template

download_media: [ask/1/0]
default_output_dir: [path or empty]

Modifying Preferences Later

Users can edit EXTEND.md directly or delete it to trigger setup again.

Source: SKILL.md on GitHub

2 warnings16d5 checks · Risk SAFE
  • Gen Agent Trust Hub16d

    This skill fetches web pages and converts them to markdown using a browser-based extraction tool. It includes specific adapters for X (Twitter), YouTube, and Hacker News, and provides a generic fallback for other sites. The skill also handles automated media downloads and manages login sessions locally.

  • Socket16d

    No alerts

  • Snyk16d

    Risk: MEDIUM · 1 issue

  • Runlayer6mo

    4/15 files flagged

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at d6442e1. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub 2 days ago.

Activeupdated 5 months ago
version
1.61.0
Other metadata
metadata
{
  "openclaw": {
    "homepage": "https://github.com/JimLiu/baoyu-skills#baoyu-url-to-markdown",
    "requires": {
      "anyBins": [
        "bun"
      ]
    }
  }
}
  • url-to-markdown
  • web-scraping
  • chrome-cdp
  • markdown
  • baoyu-fetch
  • x-twitter
  • youtube
  • hacker-news
  • login-captcha

README badge

README badge for jimliu/baoyu-skills/baoyu-url-to-markdown

Fetches any URL and converts it to markdown using Chrome DevTools Protocol with site-specific adapters for X, YouTube, Hacker News, and generic pages. Handles login and CAPTCHA interactions via wait modes, and optionally downloads images and videos to local directories.

Generated from the current SKILL.md.

Does this skill require Chrome or Chromium installed?
Yes. The skill uses Chrome DevTools Protocol to fetch and render pages. You can specify a custom Chrome binary with `--browser-path` if the default is not found.
What happens if a page requires login or CAPTCHA?
Use `--wait-for interaction` to pause and let you complete login or CAPTCHA manually, then continue. The skill will wait up to 10 minutes by default (`--interaction-timeout`).
Which websites have built-in adapters?
X (Twitter), YouTube, Hacker News, and generic pages. The skill auto-detects the adapter or you can force one with `--adapter`.
Can this skill download images and videos from a page?
Yes, with `--download-media`. Media files are saved to `imgs/` and `videos/` subdirectories next to the markdown file, and links are rewritten to point to local paths.
Does this skill require Bun?
Yes. Bun is the runtime for the vendored `baoyu-fetch` CLI. If not installed, the skill will prompt you to install it.

Generated from the current SKILL.md. These answers refresh after source changes.