Python - Loading Zip Codes into a DataFrame as Strings?

Question

I'm using Pandas to load an Excel spreadsheet which contains zip code (eg 32771). The zip codes are stored as 5 digit strings in spreadsheet. When they are pulled into a DataFrame using the command...

xls = pd.ExcelFile("5-Digit-Zip-Codes.xlsx")
dfz = xls.parse('Zip Codes')

they are converted into numbers. So '00501' becomes 501.

So my questions are, how do I:

a. Load the DataFrame and keep the string type of the zip codes stored in the Excel file?

b. Convert the numbers in the DataFrame into a five digit string eg "501" becomes "00501"?

Answer 1

As a workaround, you could convert the int s to 0-padded strings of length 5 using Series.str.zfill :

df['zipcode'] = df['zipcode'].astype(str).str.zfill(5)

Demo:

import pandas as pd
df = pd.DataFrame({'zipcode':['00501']})
df.to_excel('/tmp/out.xlsx')
xl = pd.ExcelFile('/tmp/out.xlsx')
df = xl.parse('Sheet1')
df['zipcode'] = df['zipcode'].astype(str).str.zfill(5)
print(df)

yields

  zipcode
0   00501

Answer 2

str(my_zip).zfill(5)

or

print("{0:>05s}".format(str(my_zip)))

are 2 of many many ways to do this

Answer 3

You can avoid panda's type inference with a custom converter, eg if 'zipcode' was the header of the column with zipcodes:

dfz = xls.parse('Zip Codes', converters={'zipcode': lambda x:x})

This is arguably a bug since the column was originally string encoded, made an issue here

Python - Loading Zip Codes into a DataFrame as Strings?

Question

3 answers

solution1
5 2015-10-15 01:00:30

solution2
2 2015-10-15 00:09:48

solution3
2 2015-10-15 00:18:11

Python - Loading Zip Codes into a DataFrame as Strings?

Question

3 answers

solution1 5 2015-10-15 01:00:30

solution2 2 2015-10-15 00:09:48

solution3 2 2015-10-15 00:18:11

solution1
5 2015-10-15 01:00:30

solution2
2 2015-10-15 00:09:48

solution3
2 2015-10-15 00:18:11