Toolbox - Convert Document to hOCR
Stay organized with collections
Save and categorize content based on your preferences.
Convert Document
output from Document AI to an hOCR XML string.
Explore further
For detailed documentation that includes this code sample, see the following:
Code sample
Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License, and code samples are licensed under the Apache 2.0 License. For details, see the Google Developers Site Policies. Java is a registered trademark of Oracle and/or its affiliates.
[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Hard to understand","hardToUnderstand","thumb-down"],["Incorrect information or sample code","incorrectInformationOrSampleCode","thumb-down"],["Missing the information/samples I need","missingTheInformationSamplesINeed","thumb-down"],["Other","otherDown","thumb-down"]],[],[[["\u003cp\u003eThis code demonstrates how to convert a \u003ccode\u003eDocument\u003c/code\u003e object from Document AI into an hOCR XML string.\u003c/p\u003e\n"],["\u003cp\u003eThe process involves utilizing the \u003ccode\u003edocument\u003c/code\u003e module from the \u003ccode\u003egoogle.cloud.documentai_toolbox\u003c/code\u003e library.\u003c/p\u003e\n"],["\u003cp\u003eAuthentication to Document AI requires setting up Application Default Credentials, as detailed in the provided documentation link.\u003c/p\u003e\n"],["\u003cp\u003eThe provided code sample can be found with further details in the Document AI Toolbox client libraries documentation, which can be explored for further information.\u003c/p\u003e\n"],["\u003cp\u003eThe \u003ccode\u003eDocument\u003c/code\u003e object, \u003ccode\u003ewrapped_document\u003c/code\u003e, can be exported into hOCR string format through the \u003ccode\u003eexport_hocr_str\u003c/code\u003e function.\u003c/p\u003e\n"]]],[],null,[]]