

获取商品基础信息、商品价格信息、商品图片、商品规格属性、商品 SKU 信息、商品评分及评论信息、评论对应 SKU 信息、店铺基础信息。
This worker collects SHEIN product detail pages through the Core remote browser.
Use targets as an array of product URLs. A string url is also accepted for compatibility. Runtime options are retry_times, retry_delay_seconds, page_timeout_ms, and wait_after_load_ms.
Each target produces one row containing product details, the first SKU, aggregate review information, store information, status, error, and the complete normalized payload in data_json.
The worker extracts window.gbRawData and the goodsDetailSchema script from the live page. It does not depend on shein.html, gbRawData.json, or goodsDetailSchema.json; those files are retained only as local debugging samples.
The endpoint uses BROWSER_WS when present. Otherwise it uses ChromeWs with the optional PROXY_AUTH, matching the other Core customer-site workers. The SHEIN automatic captcha solver is attempted when supported by the connected browser.
探索商店中更多热门采集工具
by mn06pz6r
1688商品详情爬取
by Eric Lozaga
获取Thumbtack的评论及评论数据
by mn06pz6r
根据用户screen_name或user_id游客获取所有推文
by scraper
使用关键词、个人资料、话题标签、高级搜索查询和自定义筛选条件从 X 和 Twitter 提取推文。 按账户、日期、语言、互动数据、媒体类型、认证状态或地理条件进行搜索,并采集推文文本、作者数据、点赞、回复、转发、引用、时间戳、媒体及其他公开元数据。 使用 CoreClaw 进行社交聆听、情感分析、市场研究、品牌监控、内容研究和大规模 X 数据工作流。