Advanced Usage
Leave feedback
On this page
This section describes the advanced features of GroupDocs.Parser for Node.js via Java:
- Loading - load documents from a local disk or a stream, specify the file format, open password-protected documents and control loading of external resources.
- Using OCR to extract a text from images and PDFs - extract a text from images and scanned PDFs with an OCR connector.
- Working with ZIP archives and attachments - iterate through container items and detect their file types.
- Working with templates - define template fields, tables and barcodes to parse documents by a template.
- Working with data extracted by template - read fields and tables from the parsing result.
- Extract data from databases - extract tables from databases via JDBC.
- Logging - receive errors, warnings and events with a logger implemented in JavaScript.
- Working with text - extract text in raw and accurate modes, search text, extract text areas, highlights, structure and formatted text.
- Working with tables - extract tables from documents and pages.
- Working with images - extract images from documents, pages and page areas, and save them to files.
- Working with barcodes - extract barcodes from documents, pages and page areas.
- Working with hyperlinks - extract hyperlinks from documents, pages and page areas.
- Generate previews - render document pages to images.
- Export Data - export extracted data to XML files.
- Loading
- Working with hyperlinks
- Working with tables
- Working with barcodes
- Working with text
- Working with images
- Working with ZIP archives and attachments
- Using OCR to extract a text from images and PDFs
- Extract data from databases
- Working with templates
- Working with data extracted by template
- Logging
- Generate previews
- Export Data
Was this page helpful?
Any additional feedback you'd like to share with us?
Please tell us how we can improve this page.
Thank you for your feedback!
We value your opinion. Your feedback will help us improve our documentation.