As a part of 3rd year project for my college, I have built a small tool which helps users to identify different images belonging to 10 categories and extract text from an image that are present on the screen. It is based on a CNN Image Classifier Model and Tesseract Engine.
The user has to crop area of interest on the screen and then the tool will perform the chosen function.