我已經通過腳本來實現我的蜘蛛一樣,主要的例子:如何改變scrapy用戶代理,而不設置文件
import scrapy
class BlogSpider(scrapy.Spider):
name = 'blogspider'
start_urls = ['https://blog.scrapinghub.com']
def parse(self, response):
for title in response.css('h2.entry-title'):
yield {'title': title.css('a ::text').extract_first()}
next_page = response.css('div.prev-post > a ::attr(href)').extract_first()
if next_page:
yield scrapy.Request(response.urljoin(next_page), callback=self.parse)
我奔跑着:
scrapy runspider myspider.py
如何更改用戶代理,如果我沒有設置或從startproject命令創建?由於這裏指定:
https://doc.scrapy.org/en/latest/topics/settings.html
由於我使用Django,這是Django的settings.py? – Atma
不,這是scrapy的settings.py。 – JkShaw