从webservice获取xml？

1条回答

网友

1楼 · 发布于 2024-09-28 23:27:58

你可以用漂亮的汤和捕捉你想要的标签。下面的代码应该可以让你开始！在

import pandas as pd
import requests
from bs4 import BeautifulSoup

url = "http://degra.wi.pb.edu.pl/rozklady/webservices.php?"

# secure url content
response = requests.get(url).content
soup = BeautifulSoup(response)

# find each tabela_rozklad
tables = soup.find_all('tabela_rozklad')

# for each tabela_rozklad looks like there is 12 nested corresponding   tags
tags = ['dzien', 'godz', 'ilosc', 'tyg', 'id_naucz', 'id_sala',
    'id_prz', 'rodz', 'grupa', 'id_st', 'sem', 'id_spec']

# initialize empty dataframe
df = pd.DataFrame()

# iterate over each tabela_rozklad and extract each tag and append to pandas dataframe
for table in tables:
    all = map(lambda x: table.find(x).text, tags)
    df = df.append([all])

# insert tags as columns
df.columns = tags

# display first 5 rows of table
df.head()

# and the shape of the data
df.shape # 665 rows, 12 columns

# and now you can get to the information using traditional pandas  functionality

# for instance, count observations by rodz
df.groupby('rodz').count()

# or subset only observations where rodz = J
J = df[df.rodz == 'J']

相关问题更多 >

编程相关推荐

热门问题

热门文章

从webservice获取xml？

相关问题 更多 >

编程相关推荐

热门问题

热门文章

相关问题更多 >