

模拟网站访问,用于测试、分析验证和 QA 工作流。输入一个或多个 URL,配置访问行为,并运行可重复的浏览器会话,无需自行搭建流量模拟基础设施。 使用 CoreClaw 网站流量生成器测试分析事件、检查网站在重复访问下的行为、验证页面流程,并模拟跨多个页面的浏览。
URLs to begin the crawler with. The actor generates page views for these pages and, when link discovery is enabled, for the pages it finds linked from them.
How many times should we generate page views for the same page?
Limit the number of pages to prevent the crawler from running too long.
If the page has links, should we also crawl them?
Should we open pages in parallel?
Maximum seconds to wait after loading the page before moving on.
Minimum seconds to wait after loading the page before moving on.
Should we block ads on the page? Reduces bandwidth consumption and improves performance.
Should we block images on the page? Reduces bandwidth consumption and improves performance.
Custom strings that will be used to block requests. For example 'google' will block any script/image/resource whose URL contains 'google'. Useful to block things like chat widgets and scripts that make the website load slower.
Select proxies to be used by your crawler, e.g. {"useApifyProxy": false}.
探索商店中更多热门采集工具
by scraper
从公开的 TikTok 视频中提取评论,并将对话转换为结构化数据。 粘贴一个或多个 TikTok 视频 URL,选择要收集的评论和回复数量,即可获取评论文本、用户名、点赞数、时间戳、回复数以及相关公开数据。 可将结果用于社交聆听、情感分析、营销活动研究、产品反馈、趋势分析或自动化数据工作流。
by scraper
从公开可访问的 Instagram 内容中下载照片、视频、Reels 和 Stories。 粘贴 Instagram 帖子或 Reel URL,或输入公开用户名,即可批量收集媒体。无需登录 Instagram 账户,即可获取直接媒体 URL 和实用的帖子元数据。 使用 CoreClaw 归档你自己的内容、管理已获授权的社交媒体素材、收集公开媒体用于研究,或将 Instagram 媒体提取连接到自动化工作流。
by scraper
无需逐个浏览创业公司资料,即可从 TrustMRR 提取创业公司的营收与增长数据。 以结构化数据形式采集 MRR、历史总营收、月度增长率、分类、排行榜排名、收购挂牌信息、要价、营收倍数、支付服务商和创始人信息。 可按分类、MRR 或收购状态筛选创业公司,并将结果导出,用于市场研究、竞争对手分析、销售线索挖掘和创业公司收购研究。
by scraper
下载公开的 MEGA 文件和文件夹,无需反复使用浏览器变通方法。将公开的 mega.nz链接粘贴到 CoreClaw 中并运行 Worker,即可获取文件,或将完整文件夹打包为 ZIP。 如果你想在开始完整下载前验证链接并查看文件大小,请使用“仅检查”模式。