Note: Text extraction from images requires OCR functionality. Basic OCR support is available, but for advanced scenarios, you may need to configure additional OCR providers.
GroupDocs.Parser for Python via .NET supports 50+ document formats across various categories including office documents, PDFs, emails, archives, and images. The library provides comprehensive data extraction capabilities including text, metadata, images, and tables depending on the format.
For specific format support and feature availability, please refer to the detailed tables above or consult the API Reference.
Was this page helpful?
Any additional feedback you'd like to share with us?
Please tell us how we can improve this page.
Thank you for your feedback!
We value your opinion. Your feedback will help us improve our documentation.
On this page
Analyzing your prompt, please hold on...
An error occurred while retrieving the results. Please refresh the page and try again.