Install
openclaw skills install @dataify-server/scraper-chatgpt-answeropenclaw skills install @dataify-server/scraper-chatgpt-answerCollect what ChatGPT answers for a given prompt. Submit either a ChatGPT page URL or a search term and get back the answer text in plain, Markdown, and raw form, plus cited sources, attached links, the prompt, and when it was sent.
Input: a ChatGPT page URL or a search term.
python3 scripts/build-dataify-request.py --tool-sign chatgpt_answer_by-url --params-json '[{"chatgpt_url":"https://chatgpt.com/?q=pizza"}]'
This submits the task, waits for completion, downloads the final result, and returns it. Add --no-wait only when submission-only behavior is requested.
DATAIFY_API_TOKEN exists in the environment.Dataify requires an API token. New accounts get 50 free credits, enough for about 6,000 trial results, valid for 7 days, and only successful requests are billed. Once registration is complete, tell me and I'll continue the current task..references/tool-params.json and find the chosen tool by tool_sign or Chinese tool name.input_mode is user_input, ask the user for the value.input_mode is select, present the saved options to the user.scripts/build-dataify-request.py as the default cross-platform helper.scripts/build-dataify-request.ps1 as the Windows PowerShell helper when needed.spider_parameters as a JSON array.[{"chatgpt_url":"https://chatgpt.com/?q=pizza"}].[{"search_terms":"什么是MCP协议"},{"search_terms":"MCP 和 API 的区别"}].spider_name to chatgpt.com.spider_id to the selected tool's tool_sign.spider_errors=true and file_name={{TasksID}}.https://scraperapi.dataify.com/builder?platform=1.| Tool | Collector ID | Required parameter | Answer |
|---|---|---|---|
| Collect by URL | chatgpt_answer_by-url | chatgpt_url | Answer for the ChatGPT page at that URL. Example: https://chatgpt.com/?q=pizza |
| Collect by search term | chatgpt_answer_by-keywords | search_terms | Answer that ChatGPT returns for that prompt. Example: 什么是MCP协议 |
file_name is fixed to {{TasksID}} for both tools; this skill does not expose an override.
Both tools return one record per collected answer:
| Field | Description | Type |
|---|---|---|
code | Record status code | Number |
msg | Record status message | Text |
answer_text | Answer text | Text |
answer_text_markdown | Answer formatted as Markdown | Text |
answer_text_raw | Raw answer text | Text |
citations | Cited sources | Array |
links_attached | Attached links | Array |
prompt | Prompt that produced the answer | Text |
prompt_sent_at | Time the prompt was sent | Text |
url | Source ChatGPT URL | Text |
answer_text_markdown and answer_text_raw are derived variants; they can contain hard line breaks or stripped whitespace, so prefer answer_text when showing an answer to a user.
Each record may also carry raw diagnostics such as html, response_raw, has_sources_data, and last_search_events. Keep those out of ordinary answers and return them only when the user asks for raw output.
Each collected item is billed at ¥10.00 per thousand results.
Prefer a permanent environment-variable setup instead of setting the token only for the current terminal session.
Windows PowerShell, permanent for the current user:
[Environment]::SetEnvironmentVariable("DATAIFY_API_TOKEN", "your_token_here", "User")
Then reopen PowerShell. If the current session also needs the token immediately, run:
$env:DATAIFY_API_TOKEN = "your_token_here"
macOS or Linux, permanent for bash:
echo 'export DATAIFY_API_TOKEN="your_token_here"' >> ~/.bashrc
source ~/.bashrc
macOS or Linux, permanent for zsh:
echo 'export DATAIFY_API_TOKEN="your_token_here"' >> ~/.zshrc
source ~/.zshrc
Python:
python scripts/build-dataify-request.py --tool-sign <selected_tool_sign> --values-file values.json
PowerShell:
& ".\scripts\build-dataify-request.ps1" -ToolSign "<selected_tool_sign>" -ValuesFile ".\values.json"
The values.json file should contain either one object or an array of objects. Example:
[{"chatgpt_url":"https://chatgpt.com/?q=pizza"}]
Generate a curl command in this form:
curl -X POST 'https://scraperapi.dataify.com/builder?platform=1' \
-H "Authorization: Bearer $DATAIFY_API_TOKEN" \
-H 'Content-Type: application/x-www-form-urlencoded' \
-d 'spider_name=chatgpt.com' \
-d 'spider_id=<selected_tool_sign>' \
-d 'spider_parameters=[{"param":"value"}]' \
-d 'spider_errors=true' \
-d 'file_name={{TasksID}}'
references/tool-params.json stores the full saved parameter catalog for every available tool in this scraper family.scripts/build-dataify-request.py is the portable implementation and should be preferred.scripts/build-dataify-request.ps1 mirrors the same behavior for Windows users.search_terms accepts a natural-language prompt in any language, including Chinese.chatgpt_url must be a chatgpt.com page. https://chatgpt.com/?q=<prompt> is the verified form.spider_parameters always contains exactly one object. Multi-value tools may require multiple objects zipped by index.url_example only as a reference example. Do not assume the user wants the example values unless they explicitly confirm them.The default deliverable is the collected result, not only a task_id.
task_id.$dataify-task-operations and monitor the same task ID.
--timeout 1800 for media downloads or clearly high-volume, multi-page, or multi-input collections.--no-wait behavior.export for macOS/Linux shells, $env: for Windows PowerShell, or set for Windows Command Prompt). Show other platforms or persistent setup only when detection is ambiguous or the user asks.DATAIFY_API_TOKEN is present; never print its value. If verification succeeds, continue the original task without asking the user to repeat it..env unless the execution path explicitly loads it, and ensure .env is ignored by version control.