processes scanned images or even smartphone photos in JP2 format and creates JP2 documents containing recognized text. To add it to your project, you just need to get Aspose.OCR
Aspose Maven Repository or specify Aspose Maven Repository configuration and install it within your Maven-based project by adding the following configurations to the pom.xml. For Graddle, Ivy, Sbt examples check out our repository .
Package Manager Console Command
PM> Install-Package Aspose.OCR.Cpp
With C++ OCR and just a few lines of code, you can create full-featured application that converts an JP2 image to Searchable PDF document:
- Create an instance of AsposeOcr class
- Call AsposeOCR.asposeocr_page() method
- Pass the JP2 file path as parameter
- AsposeOCR.asposeocr_page returns a String or file of Searchable PDF type
System Requirements
Before running the example, make sure that Microsoft.ML.OnnxRuntime 1.7.0 or above is added to the project. It should be automatically installed if you install Aspose.OCR via NuGet Package Manager.
- NET Standard 2.0+ compatible solution
- Aspose.OCR for .NET referenced in your project.
std::string img_path = "../srcSample.png";
// Prepare buffer for result (in symbols, len_byte = len * sizeof(wchar_t))
const size_t len = 4096;
wchar_t bfr[len] = { 0 };
size_t result = aspose::ocr::page(image_path.c_str(), bfr, len);
//Print result
std::wcout << bfr << L"\n";
JP2 What is JP2 File Format
JPEG 2000 (JP2) is an image coding system and state-of-the-art image compression standard. Designed, using wavelet technology JPEG 2000 can code lossless content in any quality at once. Moreover, without any substantial penalty in coding efficiency, JPEG 2000 have the capability to access and decode the same content efficaciously into a variety of other resolutions and qualities. The code streams in JPEG 2000 is significantly scalable having regions of interest that provide the facility for spatial random access. Possessing Up to 16384 diverse components with the dimensions in terapixels, and precision that can be high as 38 bits/sample.
Read MoreSearchable PDF What is Searchable PDF File Format
Searchable PDF files retain the original scanned image for viewing, as well as OCR text in a hidden layer that can be used for full-text searches within a document or highlighting text for copy and paste operations. Full OCR conversion to PDF, not including the original image, will never retain 100% of the original formatting, especially if the document has many images or a complex layout.
Read More