The Problem
Migrating content from legacy CMS platforms means extracting thousands of pages while preserving structure, metadata, and formatting. Manual copy-paste is error-prone and painfully slow.
Use Cases
01Quick Answer
Use CrawlForge crawl_deep (4 credits) to traverse an entire legacy site by following internal links, then extract_text (1 credit per page) to pull clean, readable content stripped of navigation and ads. A single automated run migrates hundreds of pages while preserving structure and metadata, and you can re-run it any time.
02The brief
Migrating content from legacy CMS platforms means extracting thousands of pages while preserving structure, metadata, and formatting. Manual copy-paste is error-prone and painfully slow.
CrawlForge crawl_deep traverses entire sites following internal links, while extract_text pulls clean content from each page. Migrate hundreds of pages in a single automated run.
03In code
1// Crawl legacy site and extract all content for migration2const crawl = await mcp.crawl_deep({3 url: "https://legacy-site.com/blog",4 max_depth: 3,5 follow_links: true,6 include_patterns: ["/blog/*"],7});8 9// Extract clean text from each discovered page10const pages = await Promise.all(11 crawl.urls.map(url =>12 mcp.extract_text({ url, preserve_structure: true })13 )14);15 16console.log(`Migrated ${pages.length} pages`);04The pipeline
05Questions
04
Use CrawlForge crawl_deep to traverse the whole site by following internal links, and extract_text to pull clean, readable content from each page. A single automated run can migrate hundreds of pages while preserving structure and metadata.
crawl_deep returns each page's content, and you can pair it with extract_metadata to keep titles, descriptions, and canonicals. extract_text and extract_content strip boilerplate so you migrate the real content, not navigation and ads.
crawl_deep follows internal links across an entire site or section with depth and page limits you set, so hundreds to thousands of pages in one job. Cost is 4 credits per crawl call plus extraction per page.
crawl_deep is 4 credits and extract_text is 1 credit per page. Migrating a few hundred pages typically costs a few hundred credits — a fraction of the manual copy-paste time, and repeatable if you need to re-run it.
06Keep exploring
Audit your site and competitors for metadata, broken links, content gaps, and ranking opportunities. Returns a structured report across every crawled page.
Collect and structure large-scale web datasets for fine-tuning and training AI models. Crawl entire sites, extract clean text, and export training-ready data.
Start forging
Every new account gets 1,000 free credits. No credit card required.