爬取校园新闻首页的新闻

1. 用requests库和BeautifulSoup库，爬取校园新闻首页新闻的标题、链接、正文、show-info。

2. 分析info字符串，获取每篇新闻的发布时间，作者，来源，摄影等信息。

import requests
newsurl='http://news.gzcc.cn/html/xiaoyuanxinwen/'
res = requests.get(newsurl) #返回response对象
res.encoding='utf-8'
from bs4 import BeautifulSoup
soup = BeautifulSoup(res.text,'html.parser')
#循环遍历打印
for news in soup.select('li'):
    if len(news.select('.news-list-title'))>0:
        t=news.select('.news-list-title')[0].text
        d=news.select('.news-list-description')[0].text
        a=news.a.attrs['href']
        print(t,d,a)

# 取出一条新闻的标题、链接、发布时间、来源
print('标题：'+soup.select('.news-list-title')[0].text)
print('链接：'+soup.select('a')[2]['href'])
print('发布时间：'+soup.select('.news-list-info')[0].span.text)
print('来源：'+soup.select('.news-list-info')[0].select('span')[1].text)

结果如下：

相关阅读:
Netty学习笔记——（一）
[feather]StarlingUi框架——组件库及渲染器
[feather]StarlingUi框架——Screen及界面导航
[feather]StarlingUi框架——feather抱怨
[feather]StarlingUi框架——初识feather、界面启动及Ui加载
Shader Language是什么
GPU图形绘制管线总结
ActionScript学习笔记（九）——坐标旋转、角度回弹与台球物理
ActionScript学习笔记（八）——碰撞检测
使用Javascript类库Qrcode处理和生成二维码

原文地址：https://www.cnblogs.com/1103a/p/8718247.html