GitHub - open-mmlab/mmocr at 9d62bdf84cc6dffd145881517d1598d0d1910586

Introduction

MMOCR is an open-source toolbox based on PyTorch and mmdetection for text detection, text recognition, and the corresponding downstream tasks including key information extraction. It is part of the open-mmlab project developed by Multimedia Laboratory, CUHK.

The master branch works with PyTorch 1.5.

Documentation: https://mmocr.readthedocs.io/en/latest/.

Major Features

Comprehensive Pipeline

The toolbox supports not only text detection and text recognition, but also their downstream tasks such as key inforation extraction.
Multiple Models

The toolbox supports a wide variety of state-of-the-art models for text detection, text recognition and key information extraction.
Modular Design

The modular design of MMOCR enables users to define their own optimizers, data preprocessors, and model components such as backbones, necks and heads as well as losses. Please refer to getting_started.md for how to construct a customized model.
Numerous Utilities

The toolbox provides a comprehensive set of utilities which can help users assess the performance of models. It includes visualizers which allow visualization of images, ground truths as well as predicted bounding boxes, and a validation tool for evaluating checkpoints during training. It also includes data converters to demostrate how to convert your own data to the annotation files which the toolbox supports.

License

This project is released under the Apache 2.0 license.

Changelog

v1.0 was released on 07/04/2021.

Benchmark and Model Zoo

Please refer to MODEL_ZOO.md for more details.

Installation

Please refer to install.md for installation.

Get Started

Please see getting_started.md for the basic usage of MMOCR.

Contributing

We appreciate all contributions to improve MMOCR. Please refer to contributing.md for the contributing guidelines.

Acknowledgement

MMOCR is an open-source project that is contributed by researchers and engineers from various colleges and companies. We appreciate all the contributors who implement their methods or add new features, as well as users who give valuable feedbacks. We hope the toolbox and benchmark could serve the growing research community by providing a flexible toolkit to reimplement existing methods and develop their own new OCR methods.

Projects in OpenMMLab

MMCV: OpenMMLab foundational library for computer vision.
MMClassification: OpenMMLab image classification toolbox and benchmark.
MMDetection: OpenMMLab detection toolbox and benchmark.
MMDetection3D: OpenMMLab's next-generation platform for general 3D object detection.
MMSegmentation: OpenMMLab semantic segmentation toolbox and benchmark.
MMAction2: OpenMMLab's next-generation action understanding toolbox and benchmark.
MMPose: OpenMMLab's pose estimation toolbox and benchmark.
MMTracking: OpenMMLab video perception toolbox and benchmark.
MMEditing: OpenMMLab image editing toolbox and benchmark.

Name		Name	Last commit message	Last commit date
Latest commit History 27 Commits
configs		configs
demo		demo
docs		docs
mmocr		mmocr
requirements		requirements
resources		resources
tests		tests
tools		tools
.coveragerc		.coveragerc
.gitignore		.gitignore
.gitlab-ci.yml		.gitlab-ci.yml
.pre-commit-config.yaml		.pre-commit-config.yaml
.pylintrc		.pylintrc
.readthedocs.yml		.readthedocs.yml
.travis.yml		.travis.yml
LICENSE		LICENSE
README.md		README.md
requirements.txt		requirements.txt
setup.cfg		setup.cfg
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Introduction

Major Features

License

Changelog

Benchmark and Model Zoo

Installation

Get Started

Contributing

Acknowledgement

Projects in OpenMMLab

About

Releases 20

Packages

Contributors 87

Languages

License

open-mmlab/mmocr

Folders and files

Latest commit

History

Repository files navigation

Introduction

Major Features

License

Changelog

Benchmark and Model Zoo

Installation

Get Started

Contributing

Acknowledgement

Projects in OpenMMLab

About

Topics

Resources

License

Code of conduct

Stars

Watchers

Forks

Releases 20

Packages 0

Contributors 87

Languages

Packages