Solution for Problem 1 by team codesquad for AIDL 2020. Uses ML Kit for OCR and OpenCV for image processing

Last update: Nov 27, 2022

Overview

CodeSquad PS1

Solution for Problem Statement 1 for AIDL 2020 conducted by @unifynd technologies.

Problem

Given images of bills/invoices, the task was to perform the following 3 operations:

Edge detection, cropping, flattening, enhancement of cropped image and compression.
Extracting text from the processed image.
The confidence score for the image to text conversion.

Development

Make sure you have react-native cli & the latest Android SDK installed on your system. To get started with React Native, follow here
To install OpenCV for Android, see here
Clone the github repository and install the dependencies using npm

$ git clone https://github.com/burhanuday/codesquad-PS1
$ cd codesquad-PS1
$ npm install

Move the modified versions of the libraries from the modified_open_source_libs to the node_modules folder. Replace in destination when asked
Run development build (Android SDK and adb tools are required to be installed)

$ npx react-native run-android --no-jetifier
$ npx react-native run-ios

Run the flask server from the flask-server folder

$ python app.py

For Mac

Follow the instructions mentioned on Getting Started on React Native documentation
Download the project zip from here
Edit the sdk.dir statement with the SDK path in the <extracted-folder>/android/local.properties file, for your machine.
If getting this error Could not compile settings file 'android\settings.gradle. First run /usr/libexec/java_home -V which will output something like the following:

Matching Java Virtual Machines (2):
    13.0.1, x86_64:	"Java SE 13.0.1"	/Library/Java/JavaVirtualMachines/jdk-13.0.1.jdk/Contents/Home
    1.8.0_242, x86_64:	"AdoptOpenJDK 8"	/Library/Java/JavaVirtualMachines/adoptopenjdk-8.jdk/Contents/Home

/Library/Java/JavaVirtualMachines/jdk-13.0.1.jdk/Contents/Home

Pick the version you want to be the default (1.8.0_242 the version of AdoptOpenJDK 8) then:

export JAVA_HOME=`/usr/libexec/java_home -v 1.8.0_242`

Run the app with npx react-native run-android --no-jetifier

Screens

Build

Create and then copy a keystore file to android/app

$ keytool -genkey -v -keystore mykeystore.keystore -alias mykeyalias -keyalg RSA -keysize 2048 -validity 10000

Setup your gradle variables in android/gradle.properties

MYAPP_RELEASE_STORE_FILE=mykeystore.keystore
MYAPP_RELEASE_KEY_ALIAS=mykeyalias
MYAPP_RELEASE_STORE_PASSWORD=*****
MYAPP_RELEASE_KEY_PASSWORD=*****

Add signing config to android/app/build.gradle

android {
signingConfigs {
release {
storeFile file(MYAPP_RELEASE_STORE_FILE)
storePassword MYAPP_RELEASE_STORE_PASSWORD
keyAlias MYAPP_RELEASE_KEY_ALIAS
keyPassword MYAPP_RELEASE_KEY_PASSWORD
}
}
buildTypes {
release {
signingConfig signingConfigs.release
}
}
}

Setup your gradle variables in android/gradle.properties

cd android && ./gradlew assembleRelease

Your APK will get generated at: android/app/build/outputs/apk/app-release.apk

Credits

Special thanks to react-native-document-scanner & react-native-perspective-image-cropper

NOTE: We are using heavily modified versions of both these libraries to support our usecase. You can find these modified libraries in the modified_open_source_libs/

Solution for Problem 1 by team codesquad for AIDL 2020. Uses ML Kit for OCR and OpenCV for image processing

Related tags

Overview

CodeSquad PS1

Problem

Development

For Mac

Screens

Build

Credits

Owner

Burhanuddin Udaipurwala

GDB python tool to pretty print and debug c++ xtensor containers

Toolbox for OCR post-correction

Shape Detection - It's a shape detection project with OpenCV and Python.

Handwritten Number Recognition using CNN and Character Segmentation

A post-processing tool for scanned sheets of paper.

The world's simplest facial recognition api for Python and the command line

Text to QR-CODE

Play the Namibian game of Owela against a terrible AI. Built using Django and htmx.

Comparison-of-OCR (KerasOCR, PyTesseract,EasyOCR)

Bu uygulamada Python ve Opencv kullanarak bilgisayar kamerasından yüz tespiti yapıyoruz.

A curated list of papers and resources for scene text detection and recognition

基于Paddle框架的PSENet复现

Image processing using OpenCv

a micro OCR network with 0.07mb params.

Pure Javascript OCR for more than 100 Languages 📖🎉🖥

Smart computer vision application

Markup for note taking

MORAN: A Multi-Object Rectified Attention Network for Scene Text Recognition

An Optical Character Recognition system using Pytesseract/Extracting data from Blood Pressure Reports.

ocroseg - This is a deep learning model for page layout analysis / segmentation.