urllib2.HTTPError：HTTP错误403：禁止 [英] urllib2.HTTPError: HTTP Error 403: Forbidden

查看：85 发布时间：2018/7/9 13:59:22 python http urllib

本文介绍了urllib2.HTTPError：HTTP错误403：禁止的处理方法，对大家解决问题具有一定的参考价值，需要的朋友们下面随着小编来一起学习吧！

问题描述

我正在尝试使用python自动下载历史股票数据。我尝试打开的URL以CSV文件响应，但我无法使用urllib2打开。我之前在几个问题中已经尝试更改用户代理，我甚至尝试接受响应cookie，没有运气。你能帮忙吗？

I am trying to automate download of historic stock data using python. The URL I am trying to open responds with a CSV file, but I am unable to open using urllib2. I have tried changing user agent as specified in few questions earlier, I even tried to accept response cookies, with no luck. Can you please help.

注意：同样的方法适用于雅虎财经。

代码：

import urllib2,cookielib

site= "http://www.nseindia.com/live_market/dynaContent/live_watch/get_quote/getHistoricalData.jsp?symbol=JPASSOCIAT&fromDate=1-JAN-2012&toDate=1-AUG-2012&datePeriod=unselected&hiddDwnld=true"

hdr = {'User-Agent':'Mozilla/5.0'}

req = urllib2.Request(site,headers=hdr)

page = urllib2.urlopen(req)

错误

文件C：\Python27 \lib \ urllib2.py，第527行，http_error_default
引发HTTPError（req.get_full_url（），代码， msg，hdrs，fp）urllib2.HTTPError：HTTP错误403：禁止

File "C:\Python27\lib\urllib2.py", line 527, in http_error_default raise HTTPError(req.get_full_url(), code, msg, hdrs, fp) urllib2.HTTPError: HTTP Error 403: Forbidden

感谢您的帮助

推荐答案

通过添加更多标题，我能够获取数据：

By adding a few more headers I was able to get the data:

import urllib2,cookielib

site= "http://www.nseindia.com/live_market/dynaContent/live_watch/get_quote/getHistoricalData.jsp?symbol=JPASSOCIAT&fromDate=1-JAN-2012&toDate=1-AUG-2012&datePeriod=unselected&hiddDwnld=true"
hdr = {'User-Agent': 'Mozilla/5.0 (X11; Linux x86_64) AppleWebKit/537.11 (KHTML, like Gecko) Chrome/23.0.1271.64 Safari/537.11',
       'Accept': 'text/html,application/xhtml+xml,application/xml;q=0.9,*/*;q=0.8',
       'Accept-Charset': 'ISO-8859-1,utf-8;q=0.7,*;q=0.3',
       'Accept-Encoding': 'none',
       'Accept-Language': 'en-US,en;q=0.8',
       'Connection': 'keep-alive'}

req = urllib2.Request(site, headers=hdr)

try:
    page = urllib2.urlopen(req)
except urllib2.HTTPError, e:
    print e.fp.read()

content = page.read()
print content

实际上，它只适用于这一个额外的标题：

Actually, it works with just this one additional header:

'Accept': 'text/html,application/xhtml+xml,application/xml;q=0.9,*/*;q=0.8',

这篇关于urllib2.HTTPError：HTTP错误403：禁止的文章就介绍到这了，希望我们推荐的答案对大家有所帮助，也希望大家多多支持IT屋！

查看全文

urllib2.HTTPError：HTTP错误403：禁止 [英] urllib2.HTTPError: HTTP Error 403: Forbidden

问题描述

推荐答案

相关文章

Python最新文章

热门教程

热门工具

登录关闭

urllib2.HTTPError：HTTP错误403：禁止 [英] urllib2.HTTPError: HTTP Error 403: Forbidden

问题描述

推荐答案

相关文章

Python最新文章

热门教程

热门工具

登录 关闭

登录关闭