

Obtain product basic information, product pricing information, product images, product specification attributes, product SKU information, product rating and review information, review-corresponding SKU information, and store basic information.
This worker collects SHEIN product detail pages through the Core remote browser.
Use targets as an array of product URLs. A string url is also accepted for compatibility. Runtime options are retry_times, retry_delay_seconds, page_timeout_ms, and wait_after_load_ms.
Each target produces one row containing product details, the first SKU, aggregate review information, store information, status, error, and the complete normalized payload in data_json.
The worker extracts window.gbRawData and the goodsDetailSchema script from the live page. It does not depend on shein.html, gbRawData.json, or goodsDetailSchema.json; those files are retained only as local debugging samples.
The endpoint uses BROWSER_WS when present. Otherwise it uses ChromeWs with the optional PROXY_AUTH, matching the other Core customer-site workers. The SHEIN automatic captcha solver is attempted when supported by the connected browser.
Explore more popular scrapers from our marketplace
by mn06pz6r
1688 Product Detail Scraper
by Eric Lozaga
Get reviews and review data from Thumbtack
by mn06pz6r
Retrieve all tweets from a guest based on the user's screen_name or user_id
by scraper
Search by account, date, language, engagement, media type, verification status, or geographic criteria and collect tweet text, author data, likes, replies, reposts, quotes, timestamps, media, and other public metadata. Use CoreClaw for social listening, sentiment analysis, market research, brand monitoring, content research, and large-scale X data workflows.