scrapyd

2,980 7 7 BSD-3-Clause

1.5.0 (2 Oct 2024) Sep 04 2013 25.1 thousand (month)

Scrapyd is a service for running Scrapy spiders. It allows you to schedule spiders to run at regular intervals and also allows you to run spiders on remote machines. It is built in Python, and it is meant to be used in a server-client architecture, where the scrapyd server runs on a remote machine, and clients can schedule and control spider runs on the server using an HTTP API. With Scrapyd, you can schedule spider runs on a regular basis, schedule spider runs on demand, and view the status of running spiders.

You can also see the logs of completed spiders, and manage spider settings and configurations. Scrapyd also provides an API that allows you to schedule spider runs, cancel spider runs, and view the status of running spiders. You can install the package via pip by running pip install scrapyd and then you can run the package by running scrapyd command in your command prompt. By default, it will start a web server on port 6800, but you can specify a different port using the `--port`` option.

Scrapyd is a good solution if you need to run Scrapy spiders on a remote machine, or if you need to schedule spider runs on a regular basis. It's also useful if you have multiple spiders, and you need a way to manage and monitor them all in one place.

for more web interface see scrapydweb

Example Use

$ scrapyd
$ curl http://localhost:6800/schedule.json -d project=myproject -d spider=spider2

Alternatives / Similar

scrapy

54,211 2.12.0 (9 months ago) Jul 26 2019 compare

scrapydweb

3,218 1.6.0 (6 months ago) Sep 30 2018 compare

gerapy

3,365 0.9.13 (2 years ago) Jul 04 2017 compare

autoscraper

6,638 1.1.14 (3 years ago) Jul 26 2019 compare

gracy

247 1.34.0 (8 months ago) Feb 05 2023 compare

splash

4,122 3.5 (5 years ago) Apr 25 2014 compare

ruia

1,754 0.8.5 (2 years ago) Oct 17 2018 compare

photon

11,149 1.1.9 (6 years ago) Aug 24 2018 compare

dude

428 0.1.3 (2 years ago) Feb 20 2022 compare

Other Languages

colly

23,747 v2.1.0 (5 years ago) May 14 2018 compare

pholcus

7,580 v1.3.4 (5 years ago) Feb 15 2020 compare

geziyor

2,667 2025-02-18 (6 months ago) Jun 06 2019 compare

dataflowkit

676 2025-02-16 (6 months ago) Feb 09 2017 compare

rvest

1,498 1.0.4 (3 years ago) Nov 22 2014 compare

gocrawl

2,039 (4 years ago) Nov 20 2016 compare

ferret

5,716 v0.18.0 (2 years ago) Aug 06 2019 compare

node-crawler

6,733 2.0.2 (1 year, 1 month ago) Sep 10 2012 compare

panther

2,977 v2.2.0 (6 months ago) Jul 17 2018 compare

spidr

813 0.7.2 (6 months ago) Jul 25 2009 compare

wombat

1,316 3.0.0 (3 years ago) Dec 27 2011 compare

ralger

156 2.2.4 (4 years ago) Dec 22 2019 compare

roach

1,384 v3.2.0 (1 year, 4 months ago) Dec 27 2021 compare

ayakashi

213 1.0.0-beta8.4 (2 years ago) Apr 18 2019 compare

phpscraper

554 3.0.0 (1 year, 4 months ago) May 04 2020 compare

php-spider

1,335 v0.7.2 (1 year, 8 months ago) Mar 16 2013 compare

crwlr-crawler

356 v3.2.3 (6 months ago) Apr 18 2022 compare