Data & Analytics Workflows
Data processing and analytics
Automatic Scraping of Company Information Before Call
Instantly collect and organize detailed company data, allowing sales reps to focus on selling and providing every prospect with personal, in-depth attention. It helps to flag pain points and highlight relevant trends, empowering your team to deliver perfectly customized value propositions every time.
Activepieces$2.99Automate Upwork Job Scraping and Daily Email Reporting
This n8n workflow automates the process of scraping Upwork job listings using Apify, storing data in Google Sheets, and sending daily email reports. It ensures data quality through cleaning and deduplication, providing a streamlined solution for job tracking and market analysis.
n8n$14.99Automate LinkedIn Company URL Extraction and Enrichment
This n8n workflow automates the process of extracting and enriching LinkedIn company URLs from a list of company names in Google Sheets using Bright Data and Google Gemini AI.
n8n$14.99Extract and Analyze URLs from PDF Files Using PDF.co
This workflow extracts all hyperlinks from PDF documents by converting them to HTML using PDF.co, allowing for easy URL analysis and processing.
n8n$9.99Automate LinkedIn Profile Enrichment with Ghost Genius and Google Sheets
Streamline the process of gathering comprehensive LinkedIn profile data using Ghost Genius and Google Sheets. Ideal for sales, recruitment, and marketing teams seeking efficient lead generation and candidate sourcing.
n8n$4.99Collect Company Social Media Profiles with Extruct AI to Google Sheets
**Who's it for:** Sales teams, marketers, and analysts who need to quickly access all the social media and public profile links for any company. **How it works / What it does:** When you enter a company into the form, this workflow automatically searches for and collects all available links to the company's social media accounts, review sites, and public profiles from sources like Crunchbase and Zoominfo. All discovered URLs are added directly to your Google Sheet. **How to set up:** 1. Create an Extruct account at [www.extruct.ai/](https://www.extruct.ai/). 2. Open the Extruct table template, find the table ID in your browser's address bar, and copy it. 3. Make a copy of the provided Google Sheets template to your own Google Drive. 4. In n8n, paste the table ID into the variables node of your flow. 5. Set up Bearer authentication in every HTTP Request node using your Extruct API token (found on the API page in Extruct). 6. In the Google Sheets node, paste the link to your copied template and connect your Google account. 7. Run the flow once to load the fields, then map the output fields to the correct columns in your sheet. 8. Activate the flow and start adding companies via the form. **Requirements:** - Extruct account and API token - Extruct table template - Google account with Google Sheets **How to customize the workflow:** You can add your own columns to the Extruct table and your Google Sheet. Just add the new column in both places and map it in the Google Sheets node in n8n.
n8n$9.99Scrape market data and send it to Slack
Overview Pipedream can leverage both Python and Node.js code in workflows, which can be used to scrape data from websites in real time. This workflow will read market data in real time from a website that publishes unusual asset class pricing, and ...
Pipedream$2.99Send Google Analytics Data to AI to Analyze, Then Save Results in Baserow
## Who's this for? - If you own a website and need to analyze your Google Analytics data - If you need to create an SEO report on which pages are getting the most traffic or how your Google search terms are performing - If you want to grow your site based on suggestions from data   ## Use case Instead of hiring an SEO expert, I run this report weekly. It checks and compares the data from this week to the week before: - Views based on countries - The top performing pages - Google Search Console performance [Watch YouTube tutorial here](https://www.youtube.com/watch?v=KlWFhz9M9g) [Get my SEO A.I. agent system here](https://2828633406999.gumroad.com/l/rumjahn) ## How it works - The workflow gathers Google Analytics data for the past 7 days, then it gathers the data for the week before for comparison. - It does this 3 times to get: views per country, engagement per page, and Google Search Console results for organic search results. - The Google Analytics nodes have already chosen the correct dimensions and metrics. - At the end, it passes the data to openrouter.ai for A.I. analysis. - Finally, it saves to Baserow. ## How to use this - Input your Google Analytics credentials - Input your property ID - Input your Openrouter.ai credentials - Input your Baserow credentials - You will need to create a Baserow database with columns: Name, Country Views, Page Views, Search Report, Blog (name of your blog). Created by [Rumjahn](https://rumjahn.com/)
n8n$14.99Automate AI Access to Transportation Laws and Incentives
Transform the Transportation Laws and Incentives API into a Machine Communication Protocol (MCP) server, enabling seamless AI agent integration for querying laws and incentives.
n8n$9.99Automate Google Places Data Collection and Storage in Google Sheets
Streamline your lead generation by automatically scraping Google Places data using Dumpling AI and storing it in Google Sheets. Ideal for marketers and SEO professionals seeking efficient data collection.
n8n$4.99Automate PostgreSQL Queries and Visualizations with AI-Driven Insights
Leverage AI to query your PostgreSQL database, generate visual insights, and update records seamlessly. This workflow integrates natural language processing to provide multi-KPI insights and auto-generated charts.
n8n$14.99Automate TrustPilot Review Analysis with Bright Data and OpenAI
Streamline the extraction, summarization, and analysis of TrustPilot reviews using Bright Data and OpenAI. This workflow automates data collection and insights generation, enhancing decision-making for product managers and marketing teams.
n8n$14.99Automate Google SERP Extraction and Summarization with Bright Data
Streamline the extraction and summarization of Google Search results using Bright Data and AI tools, delivering structured insights directly to your preferred endpoint.
n8n$9.99Automate YouTube Video Statistics Collection to Google Sheets
Efficiently gather YouTube video statistics such as views, likes, and comments, and store them in Google Sheets for analysis and reporting.
n8n$9.99Automate LinkedIn Company Data Extraction to Airtable
Effortlessly extract and structure detailed company insights from LinkedIn using Airtop, ideal for investors, sales teams, and market researchers.
n8n$4.99Track YouTube Channel's Top Videos in Google Sheets
Automatically fetch and record the most-viewed videos from specified YouTube channels into a Google Sheet for easy tracking and analysis.
n8n$9.99Extract & Summarize B2B Leads from Crunchbase with Bright Data, GPT-4O & Google Sheets
### Who this is for The Crunchbase B2B Lead Discovery Pipeline is designed for sales teams, B2B marketers, business analysts, and data operations teams who need a reliable way to extract, structure, and summarize company information from Crunchbase to fuel lead generation and market intelligence. This workflow is ideal for: 1. **Sales Development Reps (SDRs)** - Needing structured leads from Crunchbase 2. **Marketing Analysts** - Generating segmented outreach lists 3. **Growth Teams** - Identifying trending B2B startups 4. **RevOps Teams** - Automating company research pipelines 5. **Data Teams** - Consolidating insights into Google Sheets for dashboards ### What problem is this workflow solving? Manual extraction of company data from Crunchbase is time-consuming, inconsistent, and often lacks the contextual summary required for sales enablement or growth targeting. This workflow automates the extraction, transformation, summarization, and delivery of Crunchbase company data into structured formats, making it instantly usable for B2B targeting and analysis. It solves: - The difficulty of scaling lead discovery from Crunchbase - The need to summarize raw textual content for quick insights - The lack of integration between web scraping, LLM processing, and storage ### What this workflow does - **Markdown to Textual Data Extractor**: Takes raw scraped markdown from Crunchbase and converts it into readable plain text using a basic LLM chain - **Structured Data Extraction**: Applies a parsing model (OpenAI) to extract structured fields such as company name, funding rounds, industry tags, location, and founding year - **Summarization Chain**: Generates an executive summary from the raw Crunchbase text using a summarization prompt template - **Send to Google Sheets**: Adds the structured data and summary into a Google Sheet for team access and further processing - **Persist to Disk**: Saves both raw and structured data files locally for archiving or further use - **Webhook Notification**: Sends a structured payload to a webhook endpoint (e.g., Slack, CRM, internal tools) with lead insights ### Pre-conditions 1. You need to have a [Bright Data](https://brightdata.com/) account and do the necessary setup as mentioned in the Setup section below. 2. You need to have an OpenAI Account. ### Setup - Sign up at [Bright Data](https://brightdata.com/). - Navigate to Proxies & Scraping and create a new Web Unlocker zone by selecting Web Unlocker API under Scraping Solutions. - In n8n, configure the Header Auth account under Credentials (Generic Auth Type: Header Authentication).  The Value field should be set with the **Bearer XXXXXXXXXXXXXX**. The XXXXXXXXXXXXXX should be replaced by the Web Unlocker token. - In n8n, Configure the Google Sheet Credentials with your own account. Follow this documentation - [Set Google Sheet Credential](https://docs.n8n.io/integrations/builtin/credentials/google/) - In n8n, configure the OpenAi account credentials. - Ensure the URL and Bright Data zone name are correctly set in the **Set URL, Filename and Bright Data Zone** node. - Set the desired local path in the **Write a file** to disk node to save the responses. ### How to customize this workflow to your needs **LLM Prompt Customization**: - Modify the extraction prompt to include additional fields like revenue, social links, leadership team - Adjust summarization tone (e.g., executive summary, sales-focused snapshot or marketing digest) **File Persistence** - Store raw markdown, extracted JSON, and summary text separately for audit/debug **Webhook Notification** - Connect to CRM (e.g., HubSpot, Salesforce) via webhook to automatically create leads - Send Slack notifications to alert sales reps when a new high-potential company is discovered
n8n$14.99Hacker News Job Listing Scraper and Parser
This automated workflow scrapes and processes the monthly Who is Hiring thread from Hacker News, transforming raw job listings into structured data for analysis or integration with other systems. Perfect for job seekers, recruiters, or anyone looking to monitor tech job market trends. ## How it works - Automatically fetches the latest Who is Hiring thread from Hacker News - Extracts and cleans relevant job posting data using the HN API - Splits and processes individual job listings into structured format - Parses key information like location, role, requirements, and company details - Outputs clean, structured data ready for analysis or export ## Set up steps 1. Configure API access to [Hacker News](https://github.com/HackerNews/API) (no authentication required) 2. Follow the steps to get your cURL command from [https://hn.algolia.com/](https://hn.algolia.com/) 3. Set up desired output format (JSON structured data or custom format) 4. Optional: Configure additional parsing rules for specific job listing information 5. Optional: Set up integration with preferred storage or analysis tools The workflow transforms unstructured job listings into clean, structured data following this pattern: - Input: Raw HN thread comments - Process: Extract, clean, and parse text - Output: Structured job listing data This template saves hours of manual work collecting and organizing job listings, making it easier to track and analyze tech job opportunities from Hacker News's popular monthly hiring threads.
n8n$14.99Efficiently Scrape Google Maps Data Using SerpAPI and n8n
Automate the extraction of detailed Google Maps data using SerpAPI and store it in Google Sheets for easy access and analysis.
n8n$14.99Automate Job Tracking from Y Combinator to Google Sheets
Automatically scrape and store the latest job listings from Y Combinator's Jobs page into Google Sheets every 6 hours using n8n and Scrapeless.
n8n$9.99Automate SEO Insights from Matomo Analytics with AI and Store in Baserow
This workflow automates the process of analyzing Matomo analytics data using AI and stores the insights in Baserow. It helps website owners improve visitor engagement by providing actionable SEO recommendations.
n8n$9.99Pull Square Sales Summary Reports for Automated Reporting and Analysis
## Programmatically Pull Square Report Data Into N8N ## What It Does This sub-workflow connects to the Square API and generates a daily sales summary report for all of your Square locations. The report matches the figures displayed in the Square Dashboard > Reports > Sales Summary. It's designed to be reused in other workflows, ideal for reporting, data storage, accounting, or automation. ## Prerequisites To use this workflow, you'll need: - Square API credentials (configured as a Header Auth credential) ## How to Set Up Square Credentials: - Go to Credentials > Create New - Choose Header Auth - Set the Name to Authorization - Set the Value to your Square Access Token (e.g., Bearer <your-api-key>) ## How It Works 1. Trigger: The workflow is triggered as a sub-workflow, requiring a report_date input. 2. Fetch Locations: An HTTP request gets all Square locations linked to your account. 3. Fetch Orders: For each location, an HTTP request pulls completed orders for the specified report_date. 4. Filter Empty Locations: Locations with no sales are ignored. 5. Aggregate Sales Data: A Code node processes the order data and produces a summary identical to Square's built-in Sales Summary report. 6. Output: A cleaned, consistent summary that can be consumed by parent workflows or other nodes. ## Example Use Cases - Automatically store daily sales data in Google Sheets, MySQL, or PostgreSQL for analysis and historical tracking - Automatically send daily email or Slack reports to managers or finance teams - Build weekly/monthly reports by looping over multiple dates - Push sales data into accounting software like QuickBooks or Xero for automated bookkeeping - Calculate commissions or rent payments based on sales volume ## How to Use - Configure both HTTP Request nodes to use your Square API credential. - If you are not in the Toronto/New York timezone, please change the start_at and end_at parameters in the second HTTP node from -05:00 to your local timezone - Use as a sub-workflow inside a main workflow. - Pass a report_date (formatted as YYYY-MM-DD) to the sub-workflow when you call it. ## Customization Options - Add pagination to handle locations with more than 1,000 orders per day. - Expand the workflow to save or send the report output via additional integrations (email, database, webhook, etc.). ## Why It's Useful This workflow saves time, reduces manual report pulling from Square, and enables smarter automation around sales data—whether for operations, finance, or performance monitoring.
n8n$9.99Generate Data Pipeline Blueprints with Claude 3.5, Slack, and Tavily Search
## Architecture Agent ## Overview The Architect Agent listens to Slack messages and generates full data architecture blueprints in response. Powered by Claude 3.5 (Anthropic) for reasoning and design, and heavily for real-time web search, this agent creates production-ready data pipeline scaffolds on-demand - transforming natural language prompts into structured data engineering solutions. ## Capabilities - Understands and interprets user requests from Slack - Designs end-to-end data pipeline architectures using industry best practices. - Outputs include high-level architecture diagrams ## Required Connections To operate correctly, the following integrations must be in place: - Slack API token with permission to read messages and post responses - Heavily API Key for external search functionality - Claude 3.5 API Access via Anthropic Detailed configuration instructions are provided in the workflow. ## Setup time <15 minutes ## Example input: Create a data pipeline orchestrated by Airflow, running on a Docker image. It should connect to a MySQL database, load the data into a PostgreSQL DB (incremental load), and then transform the data into business-oriented tables also in the PostgreSQL database. Create an example setup with raw sales data. ## Customizing this workflow Try saving outputs to Google Drive to store all your architecture blueprints.
n8n$14.99Automate Pinterest Content Scraping with AI and BrightData
This n8n workflow automates the process of scraping Pinterest content based on user-defined keywords using BrightData's API and the Claude Sonnet 4 AI model. It efficiently manages the scraping process, monitors progress, and organizes the extracted data into Google Sheets.
n8n$14.99
Related categories
Custom AI Systems & Services
Our team of experienced AI builders will help build custom AI systems, workflows, and solutions.
Request Custom Work