123FormBuilder Google Sheets

Sitemap page extractor: Discover, clean, and save website URLs to Google Sheets

Streamline your SEO audits by automatically crawling website sitemaps to identify and extract every live content URL while filtering out administrative noise. This intelligent workflow handles nested sitemap structures and robots.txt files to ensure no page is missed during discovery. All cleaned data is then instantly synced to Google Sheets for easy analysis and reporting.

Run this with your team's AI

What This Recipe Does

The Sitemap Page Extractor automation transforms the tedious task of website auditing and content inventory into a streamlined, one-click process. Instead of manually clicking through pages or employing complex crawling software, business users can simply submit a URL to instantly generate a comprehensive list of every live page on a website. This automation parses the site's XML sitemap, processes the data through an intelligent filtering engine, and organizes the results directly into a Google Sheet. This allows marketing teams to quickly assess site architecture, identify outdated content, or prepare for large-scale migrations without technical assistance. By automating the data collection phase, your team can focus on strategic analysis and content optimization rather than manual data entry. The result is a clean, actionable spreadsheet that serves as a single source of truth for your digital footprint, ensuring no page is overlooked during audits or SEO evaluations.

What your team gets

Something anyone can use

Forms and dashboards, so it is not a script only one person understands

It keeps running

Runs on your schedule in the cloud, so it does not stop when a laptop closes

Your other tools can call it

Endpoints, so the rest of your stack can trigger the same work

Accounts connected once

123FormBuilder, Google Sheets connected for the team, not per person

How It Works

  1. 1

    Open the recipe and connect your accounts

    Connect 123FormBuilder and Google Sheets once, in your team cloud, and nobody has to do it again on their own machine

  2. 2

    Tell your own agent what is different about your process

    Claude, ChatGPT, Cursor, whichever your team already uses. It adapts the recipe to how you actually work

  3. 3

    Run it, then leave it running

    It lives in your team cloud, so it keeps going after you close the laptop and every teammate's AI can use it

Who Uses This

Frequently Asked Questions

What information do I need to provide to start the extraction?

You only need to provide the URL of the website or the specific link to the XML sitemap through the simple input form.

Can I filter out specific types of pages from the final report?

Yes, the automation includes filtering logic that can be adjusted to exclude specific URL patterns, such as category pages or internal system links.

Does this require any coding knowledge to operate?

No coding is required. The app provides a user-friendly form interface, and the data is automatically delivered to a standard Google Sheet.

How many pages can this automation handle at once?

The workflow uses batch processing to efficiently handle large sitemaps, ensuring that even extensive websites are documented without hitting performance limits.

Coming from n8n?

This recipe uses nodes like StickyNote, FormTrigger, Set, Code and 5 more. On Runwork, you don't need to learn n8n's workflow syntax. Describe what you want to your own AI agent in plain English.

StickyNote FormTrigger Set Code SplitInBatches HttpRequest If Filter GoogleSheets

Based on n8n community workflow. View original

Related Recipes

Google-gemini Http

Automated YouTube video scheduling & AI metadata generation 🎬

The YouTube Content Intelligence automation transforms how marketing teams monitor and leverage video content. By automating the extraction and processing of YouTube data, this workflow eliminates the manual effort required to track competitors, monitor brand mentions, or gather industry insights. It systematically fetches video details, handles high volumes of data through intelligent batching, and removes duplicates to ensure your dataset remains clean and actionable. Instead of spending hours browsing channels and manually logging information, your team receives a structured feed of intelligence ready for analysis. This automation allows you to stay ahead of market trends, identify high-performing content patterns, and react to new video releases in real-time. By bridging the gap between raw video data and business strategy, you can optimize your content production and competitive positioning with data-driven precision.

See the recipe
Http

Retweet cleanup with scheduling for X/Twitter

Maintaining a professional and focused brand presence on X requires a curated feed that highlights your original insights rather than a cluttered history of old retweets. The Automated Retweet Cleanup for X Accounts allows businesses and personal brands to automatically purge shared content after it has served its purpose. By scheduling regular cleanups, you ensure that your profile remains streamlined and relevant to new followers who are auditing your expertise. This automation systematically scans your account, identifies retweets based on your specific criteria, and removes them without requiring manual oversight. Beyond just tidying your timeline, the workflow includes error handling and Slack notifications to keep you informed of the process. This tool is essential for marketing teams who want to maintain a high-signal-to-noise ratio on social media while protecting their brand's long-term digital footprint.

See the recipe
Http

Extract and merge Twitter (X) threads using TwitterAPI.io

The Twitter Thread Fetcher is a specialized automation designed to solve the problem of fragmented social media content. Instead of manually copying and pasting individual tweets from a long thread, this automation systematically extracts every post in a sequence and merges them into a single, cohesive document. This is particularly valuable for content marketers and researchers who need to archive high-performing social content or repurpose Twitter insights for other platforms like blogs, newsletters, or internal knowledge bases. By automating the retrieval process, businesses can save significant time and ensure that no part of a valuable discussion or announcement is lost. The automation handles the technical complexity of connecting to the Twitter API and restructuring the data, providing a clean, ready-to-use text output that preserves the original narrative flow of the thread.

See the recipe
Http

Auto-post breaking news content using Perplexity AI to X (Twitter)

Maintaining a consistent and authoritative presence on social media is essential for brand growth, yet manually scouring the web for relevant news is a significant time sink. This automation bridges the gap between information gathering and content distribution by leveraging Perplexity AI to research current events and automatically drafting and posting updates to X (Twitter). Instead of spending hours each morning looking for industry trends, this workflow proactively identifies timely topics based on your specific criteria and shares them with your audience. By automating the research-to-publication pipeline, businesses can ensure they remain at the forefront of industry conversations without diverting resources from high-level strategy. This system transforms your social media profile into a real-time information hub, increasing engagement and establishing your brand as a thought leader in your niche. The result is a streamlined marketing operation that delivers consistent value to followers while operating entirely in the background.

See the recipe

Run this with the AI your team already uses

Your agent adapts it, your team cloud keeps it running, and everyone's AI can find it.

Open this recipe in Runwork