"dataify-linkedin-company-information-by-url"
Prepare Dataify builder requests for the linkedin.com scraper family rooted at linkedin_company_information_by-url. Use when needs to work with the successfu...
dataify-server
@dataify-server
What This Skill Does
Generates Dataify builder curl requests for scraping LinkedIn company information by URL. It walks the user through selecting a scraper tool, collecting parameter values (user input or saved options), and building the final request with the correct spider name, spider ID, and spider_parameters.
Replaces manually constructing Dataify API requests by automating the parameter collection, tool selection, and curl command generation for LinkedIn company scraping.
When to Use It
- Scrape LinkedIn company profile data by providing a company page URL
- Build a Dataify scraper request for LinkedIn job listings from a job listing URL
- Generate a curl command to scrape LinkedIn job listings by keyword search
- Prepare a Dataify builder request with multiple parameter values zipped into separate rows
- Set up a permanent DATAIFY_API_TOKEN environment variable for repeated scraping tasks
Install
$ openclaw skills install @dataify-server/dataify-linkedin-company-information-by-urlDataify Builder Skill
Use this skill to prepare Dataify builder requests for the scraper family rooted at linkedin_company_information_by-url on linkedin.com.
Workflow
- Check whether
DATAIFY_API_TOKENexists in the environment. - If the token is missing, stop and tell the user to sign in at Dataify Dashboard to obtain it.
- Ask the user to choose exactly one tool from the following Chinese list:
- 通过URL采集 (linkedin_company_information_by-url)
- 通过职位列表URL采集 (linkedin_job_listings_information_by-job-listing-url)
- 通过职位URL采集 (linkedin_job_listings_information_by-job-url)
- 通过关键词采集 (linkedin_job_listings_information_by-keyword)
- Read
references/tool-params.jsonand find the chosen tool bytool_signor Chinese tool name. - For each parameter in the chosen tool:
- If
input_modeisuser_input, ask the user for the value. - If
input_modeisselect, present the saved options to the user.
- If
- Use
scripts/build-dataify-request.pyas the default cross-platform helper. - Use
scripts/build-dataify-request.ps1as the Windows PowerShell helper when needed. - When a selectable parameter has a human-readable Chinese label, keep that label in
spider_parameters. Do not replace it with a code such asHKunless the user explicitly asks for the coded value. - Build
spider_parametersas a JSON array. - If every parameter has only one final value, build one object such as
[{"searchurl":"...","country":"Hong Kong"}]. - If one or more parameters have multiple aligned values, zip them by index and build one object per row. Example:
[{"search_url":"url1","page_turning":"1","max_num":"15"},{"search_url":"url2","page_turning":"1","max_num":"15"}]. - If a parameter has one value while another parameter has multiple values, reuse the single value across every generated row.
- Set
spider_nametolinkedin.com. - Set
spider_idto the selected tool'stool_sign. - Always include
spider_errors=trueandfile_name={{TasksID}}. - Return a curl command for
https://scraperapi.dataify.com/builder.
Set DATAIFY_API_TOKEN
Prefer a permanent environment-variable setup instead of setting the token only for the current terminal session.
Windows PowerShell, permanent for the current user:
[Environment]::SetEnvironmentVariable("DATAIFY_API_TOKEN", "your_token_here", "User")
Then reopen PowerShell. If the current session also needs the token immediately, run:
$env:DATAIFY_API_TOKEN = "your_token_here"
macOS or Linux, permanent for bash:
echo 'export DATAIFY_API_TOKEN="your_token_here"' >> ~/.bashrc
source ~/.bashrc
macOS or Linux, permanent for zsh:
echo 'export DATAIFY_API_TOKEN="your_token_here"' >> ~/.zshrc
source ~/.zshrc
Script usage
Python:
python scripts/build-dataify-request.py --tool-sign <selected_tool_sign> --values-file values.json
PowerShell:
& ".\scripts\build-dataify-request.ps1" -ToolSign "<selected_tool_sign>" -ValuesFile ".\values.json"
The values.json file should contain either one object or an array of objects. Example:
[{"searchurl":"https://www.airbnb.com/s/Greece/homes?...","country":"Hong Kong"}]
Required output shape
Generate a curl command in this form:
curl -X POST 'https://scraperapi.dataify.com/builder' \
-H "Authorization: Bearer $DATAIFY_API_TOKEN" \
-H 'Content-Type: application/x-www-form-urlencoded' \
-d 'spider_name=linkedin.com' \
-d 'spider_id=<selected_tool_sign>' \
-d 'spider_parameters=[{"param":"value"}]' \
-d 'spider_errors=true' \
-d 'file_name={{TasksID}}'
Reference usage
references/tool-params.jsonstores the full saved parameter catalog for every available tool in this scraper family.scripts/build-dataify-request.pyis the portable implementation and should be preferred.scripts/build-dataify-request.ps1mirrors the same behavior for Windows users.- If a parameter has no options, the user must provide the value.
- If a parameter has options, present those options back to the user before building the final request.
- Do not assume
spider_parametersalways contains exactly one object. Multi-value tools may require multiple objects zipped by index. - Use the saved
url_exampleonly as a reference example. Do not assume the user wants the example values unless they explicitly confirm them.
Top skills in this category
Planning with files
@othmanadiManus-style persistent file-based planning for AI coding agents: keeps task_plan.md, findings.md, and progress.md on disk so work survives context loss and /clear. Use when asked to plan out, break down, or organize a multi-step project, research task, or any work requiring 5+ tool calls. Supports a
Web Browsing
@tankebuaaBrowse and summarize websites, extract content from URLs, search the web for information. Use when user asks to visit a website, get webpage content, search...
Multi Search Engine
@gpyangyoujunMulti search engine integration with 16 engines (7 CN + 9 Global). Supports advanced search operators, time filters, site search, privacy engines, and Wolfra...
Agent Browser
@matrixyHeadless browser automation CLI optimized for AI agents with accessibility tree snapshots and ref-based element selection
Nano Pdf
@steipeteEdit PDFs with natural-language instructions using the nano-pdf CLI.