熊猫无法识别 Excel 文件

时间:2021-05-26 05:36:48

标签: python excel pandas

我想对存储在 Excel 文件中的文本数据进行可读性分析。我改编的部分代码如下:

import time, datetime     
import pandas as pd     
from textstat.textstat import textstat    
from openpyxl import load_workbook    

ExcelFile = 'Readability.xlsx'
Sheet = 'Raw Data'
Field_ID = 0 

book = load_workbook(ExcelFile)
writer = pd.ExcelWriter(ExcelFile, engine='openpyxl')
writer.book = book
df = pd.read_excel(ExcelFile, sheet_name=Sheet)

运行后出现以下错误:

Traceback (most recent call last):
  File "\\file\UsersR$\rtf13\Home\Desktop\readability_using_textstat.py", line 19, in <module>
    df = pd.read_excel(ExcelFile, sheet_name=Sheet)
  File "C:\Python39\lib\site-packages\pandas\util\_decorators.py", line 299, in wrapper
    return func(*args, **kwargs)
  File "C:\Python39\lib\site-packages\pandas\io\excel\_base.py", line 336, in read_excel
    io = ExcelFile(io, storage_options=storage_options, engine=engine)
  File "C:\Python39\lib\site-packages\pandas\io\excel\_base.py", line 1071, in __init__
    ext = inspect_excel_format(
  File "C:\Python39\lib\site-packages\pandas\io\excel\_base.py", line 965, in inspect_excel_format
    raise ValueError("File is not a recognized excel file")
ValueError: File is not a recognized excel file

此外,Excel 文件在运行代码后最终损坏。我正在使用 pandas 1.2.4、openpyxl 3.0.7 并使用 xlrd 1.2.0(因为更高版本无法使用 .xlsx 文件)。欢迎任何建议。谢谢。

1 个答案:

答案 0 :(得分:0)

pd.ExcelWriter 用于编写pd.DataFrame对象。

df = pd.read_excel(ExcelFile) 
with ExcelWriter(ExcelFile , mode='a') as writer:
    df.to_excel(writer, sheet_name=Sheet)


更新

试试这个:

df = pd.read_excel(r'<path_to_file>\Readability.xlsx', engine='openpyxl') # UPDATED: pip install openpyxl

with ExcelWriter(ExcelFile , mode='a') as writer:
    df.to_excel(writer, sheet_name=Sheet)

here查看更多