擅长:python、mysql、java
<p>你用的是林克字典。如果不是为了阅读而使用它,请尝试以下代码:</p>
<pre><code> br = mechanize.Browser()
htmltext = br.open(url).read()
articletext = ""
for tag_li in soup.findAll('li', attrs={"data-section":"Op-Ed"}):
for link in tag_li.findAll('a'):
urlnew = urlnew = link.get('href')
brnew = mechanize.Browser()
htmltextnew = brnew.open(urlnew).read()
articletext = ""
soupnew = BeautifulSoup(htmltextnew)
for tag in soupnew.findAll('p'):
articletext += tag.text
print re.sub('\s+', ' ', articletext, flags=re.M)
</code></pre>
<p>注意:<code>re</code>表示正则表达式。为此,您导入<code>re</code>的模块。在</p>