社区热度榜

data-scraper-agent

面向公开来源的自动数据采集代理,支持定时抓取、LLM 丰富和多种存储落地。

分类
生产力工具
榜单排名
#6
GitHub Stars
219,439

它能做什么

中文摘要

「data-scraper-agent」的原始 GitHub SKILL.md 将它定位为:Build a fully automated AI-powered data collection agent for any public source — job boards, prices, news, GitHub, sports, anything.。它的核心价值是把前端性能、渲染路径、数据获取和包体控制变成可执行检查清单;约束界面层级、组件状态、可访问性和视觉验收细节。文档重点章节包括「When to Activate」「Core Concepts」「The Three Layers」。

为什么推荐

推荐理由

推荐它,是因为它在「生产力工具」分类里热度靠前(GitHub Stars 219,439),同时原始文档信息密度足够高。原文列出的执行要点包括:User wants to scrape or monitor any public website or API;User says "build a bot that checks...", "monitor X for me", "collect data from..."。对经常重复的任务来说,这类 Skill 能把经验沉淀到工作流前端,减少临场猜测。

什么时候用

适用场景

  • 做界面实现、组件重构、视觉验收或设计系统一致性检查
  • 设计数据模型、编写查询、做迁移和排查数据库性能问题
  • 建立测试闭环、回归验证、安全检查和发布前质量门

使用前先看

主要亮点

  • 01

    适合招聘、价格、新闻、赛事等公开页面的持续采集。

  • 02

    可将抓取结果写入 Notion、Sheets、Supabase,便于后续分析。

  • 03

    依赖公开来源与自动化运行,需关注目标站点稳定性。

原始文档

原文摘录

Build a fully automated AI-powered data collection agent for any public source — job boards, prices, news, GitHub, sports, anything.