推荐学习书目
› Learn Python the Hard Way
Python Sites
› PyPI - Python Package Index
› http://diveintopython.org/toc/index.html
› Pocoo
值得关注的项目
› PyPy
› Celery
› Jinja2
› Read the Docs
› gevent
› pyenv
› virtualenv
› Stackless Python
› Beautiful Soup
› 结巴中文分词
› Green Unicorn
› Sentry
› Shovel
› Pyflakes
› pytest
Python 编程
› pep8 Checker
Styles
› PEP 8
› Google Python Style Guide
› Code Style from The Hitchhiker's Guide
SpiderXiantang
V2EX  ›  Python

scrapyd 怎么可以让爬虫定时采集

  •  
  •   SpiderXiantang ·
    xiantang · Aug 14, 2018 · 4758 views
    This topic created in 2975 days ago, the information mentioned may be changed or developed.

    看了一下 scrapyd 的 api 感觉功能好少啊 而且对分布式也没有支持 我现在遇到的问题是是需要采集一家电商网站 然后反复爬取 进行价格的监控 请问下有没有大佬有思路

    5 replies  •  2018-11-14 23:56:16 +08:00
    SpiderXiantang
        1
    SpiderXiantang  
    OP
       Aug 14, 2018
    还有就是如何同时启动分布式爬虫 求思路!
    zzj0311
        2
    zzj0311  
       Aug 14, 2018 via Android
    crontab 了解一下?
    masha
        3
    masha  
       Aug 15, 2018
    分布式可以试试 scrapy-redis
    SpiderXiantang
        4
    SpiderXiantang  
    OP
       Aug 15, 2018
    @masha 我用的 scrapy-redis 但是不知道怎么协同启动爬虫 我需要反复的监控这个网站
    my8100
        5
    my8100  
       Nov 14, 2018
    @SpiderXiantang 如何简单高效地部署和监控分布式爬虫项目 https://v2ex.ih06.com/t/507933
    About   ·   Help   ·   Advertise   ·   Blog   ·   API   ·   FAQ   ·   Privacy   ·   Solana   ·   752 Online   Highest 6679   ·     Select Language
    创意工作者们的社区
    World is powered by solitude
    VERSION: 3.9.8.5 · 27ms · UTC 21:35 · PVG 05:35 · LAX 14:35 · JFK 17:35
    ♥ Do have faith in what you're doing.