Pandas扔铁轨

import pandas as pd import numpy as np pd.set_option('display.height', 1000) pd.set_option('display.max_rows', 500) pd.set_option('display.max_columns', 500) pd.set_option('display.width', 1000) print("reading wolverine xlxs") # defining metadata df_header = ['DisplayName','StoreLanguage','Territory','WorkType','EntryType','TitleInternalAlias', 'TitleDisplayUnlimited','LocalizationType','LicenseType','LicenseRightsDescription', 'FormatProfile','Start','End','PriceType','PriceValue','SRP','Description', 'OtherTerms','OtherInstructions','ContentID','ProductID','EncodeID','AvailID', 'Metadata', 'AltID', 'SuppressionLiftDate','SpecialPreOrderFulfillDate','ReleaseYear','ReleaseHistoryOriginal','ReleaseHistoryPhysicalHV', 'ExceptionFlag','RatingSystem','RatingValue','RatingReason','RentalDuration','WatchDuration','CaptionIncluded','CaptionExemption','Any','ContractID', 'ServiceProvider','TotalRunTime','HoldbackLanguage','HoldbackExclusionLanguage'] df_w01 = pd.read_excel("wolverine_1.xlsx", names = df_header) df_w02 = pd.read_excel("wolverine_2.xlsx", names = df_header) df_w01['version'] = 'OLD' df_w02['version'] = 'NEW' #print(df_w01) df_m_d = pd.concat([df_w01, df_w02], ignore_index = True) first_pass = df_m_d[df_m_d.duplicated(['StoreLanguage','Territory','TitleInternalAlias','LocalizationType','LicenseType','FormatProfile'], keep=False)] first_pass_keep_duplicate = df_m_d[df_m_d.duplicated(['StoreLanguage','Territory','TitleInternalAlias','LocalizationType','LicenseType','FormatProfile'], keep='first')] group_by_1 = first_pass.groupby(['StoreLanguage','Territory','TitleInternalAlias','LocalizationType','LicenseType','FormatProfile']) for i,rows in group_by_1.iterrows(): print("rownumber", i) print (rows) print(first_pass)

2条回答

网友

1楼 · 编辑于 2024-10-01 11:42:13

您的GroupBy对象支持迭代，因此

for i,rows in group_by_1.iterrows():
    print("rownumber", i)
    print (rows)

你需要做点什么

^{pr2}$

然后您可以对每个group执行所需的操作

见the docs

网友

2楼 · 编辑于 2024-10-01 11:42:13

为什么不照建议做并使用apply？比如：

def print_rows(rows):
    print rows

group_by_1.apply(print_rows)

相关问题更多 >

编程相关推荐

热门问题

热门文章