我一直在使用tesseract从图像中提取文本的项目。我也在使用python 3.7.7但是我遇到了我无法解决的错误。
tess.pytesseract.tesseract_cmd = r'C:\\Program Files (x86)\\Tesseract-OCR\\tess1\\eng.traineddata'
img = Image.open('C:\\Users\\USER\\PycharmProjects\\selenium\\automation\\screenshot.png')
text = tess.image_to_string(img, lang='eng')
当我运行此程序时出现错误
Traceback (most recent call last):
File "C:/Users/USER/PycharmProjects/selenium/automation/open.py", line 8, in <module>
text = tess.image_to_string(img, lang='eng')
File "C:\Users\USER\PycharmProjects\selenium\venv\lib\site-packages\pytesseract\pytesseract.py", line 360, in image_to_string
}[output_type]()
File "C:\Users\USER\PycharmProjects\selenium\venv\lib\site-packages\pytesseract\pytesseract.py", line 359, in <lambda>
Output.STRING: lambda: run_and_get_output(*args),
File "C:\Users\USER\PycharmProjects\selenium\venv\lib\site-packages\pytesseract\pytesseract.py", line 270, in run_and_get_output
run_tesseract(**kwargs)
File "C:\Users\USER\PycharmProjects\selenium\venv\lib\site-packages\pytesseract\pytesseract.py", line 241, in run_tesseract
raise e
File "C:\Users\USER\PycharmProjects\selenium\venv\lib\site-packages\pytesseract\pytesseract.py", line 238, in run_tesseract
proc = subprocess.Popen(cmd_args, **subprocess_args())
File "C:\Python37\lib\subprocess.py", line 800, in __init__
restore_signals, start_new_session)
File "C:\Python37\lib\subprocess.py", line 1207, in _execute_child
startupinfo)
OSError: [WinError 193] %1 is not a valid Win32 application
请提供合适的解决方案
我认为问题在于tesseract可执行文件需要提供的路径。遵循此link中的答案。这应该可以解决您的问题。