This is a pytorch re-implementation of EAST: An Efficient and Accurate Scene Text Detector.

Last update: Dec 20, 2022

Overview

EAST: An Efficient and Accurate Scene Text Detector

Description:

This version will be updated soon, please pay attention to this work. The motivation of this version is to build a easy-training model. This version can automatically update best_model by comparing current hmean and the former. At the same time, we can see evaluation info about every sample easily.

1.train
2.predict
3.compress
4.compute Hmean(if Hmean is higher than before, update best_weight.pkl)
5.visualization(blue, green, red)
6.multi-scale test (update soon) multi-scale vis. (vis with score, scales)

Thanks

The version is ported from argman/EAST, from Tensorflow to Pytorch

Check On Website

If you have no confidence of the result of our program, you could use submit.zip to submit on website,then you can see result of every image.

Performance

right -- green || wrong -- red || miss -- blue
recall/precision/hmean for every test image

Introduction

This is a pytorch re-implementation of EAST: An Efficient and Accurate Scene Text Detector. The features are summarized blow:

Only RBOX part is implemented.
A fast Locality-Aware NMS in C++ provided by the paper's author.(g++/gcc version 6.0 + will be ok)
Evalution see here for the detailed results.
Differences from original paper
- Use ResNet-50 rather than PVANET
- Use dice loss (optimize IoU of segmentation) rather than balanced cross entropy
- Use linear learning rate decay rather than staged learning rate decay

Thanks for the author's (@zxytim) help! Please cite his paper if you find this useful.

Installation
Download
Prepare dataset/pretrain
Test
Train
Examples

Installation

Any version of pytorch version > 0.4.0 should be ok.

Download

Pretrained model is not provided temporarily. Web site is updating now, please continue to pay attention

Prepare dataset/pretrain weight

[1]. dataset(you need to prepare for dataset for train and test) suggestions: you could do a soft-link to root_to_this_program/dataset/train/img/*.jpg

-- train ./dataset/train/img/img_###.jpg ./dataset/train/gt/img_###.txt (you need to change name)
-- test ./data/test/img_###.jpg (img only)
-- gt.zip ./result/gt.zip(ICDAR15 gt.zip is avaliable on website

** Note: you can download dataset here

-- ICDAR15
-- ICDAR13

[2]. pretrained

In config.py set resume True and set checkpoint path/to/weight/file
I will provide pretrianed weight soon

[3]. check GPUs and CPUs you can use following to check aviliable gpu, this is for train

watch -n 0.1 nvidia-smi

then, you will see 2,3 is avaliable, modify config.py gpu_ids = [0,1], gpu = 2, and modify run.sh - CUDA_VISIBLE_DEVICES=2,3

Train

If you want to train the model, you should provide the dataset path in config.py and run

sh run.py

** Note: you should modify run.sh to specify your gpu id

If you have more than one gpu, you can pass gpu ids to gpu_list(like gpu_list=0,1,2,3) in config.py

** Note: you should change the gt text file of icdar2015's filename to img_*.txt instead of gt_img_*.txt(or you can change the code in icdar.py), and some extra characters should be removed from the file. See the examples in training_samples/**

Test

By default, we set train-eval process into integer. If you want to use eval independently, you can do it by yourself. Any question can contact me.

Examples

Here are some test examples on icdar2015, enjoy the beautiful text boxes!

This is a pytorch re-implementation of EAST: An Efficient and Accurate Scene Text Detector.

Related tags

Overview

EAST: An Efficient and Accurate Scene Text Detector

Description:

Thanks

Check On Website

Performance

Introduction

Contents

Installation

Download

Prepare dataset/pretrain weight

Train

Test

Examples

Owner

Dejia Song

A version of nrsc5-gui that merges the interface developed by cmnybo with the architecture developed by zefie in order to start a new baseline that is not heavily dependent upon Python processing.

Super Mario Game With Python

SceneCollisionNet This repo contains the code for "Object Rearrangement Using Learned Implicit Collision Functions", an ICRA 2021 paper. For more info

A tool for extracting text from scanned documents (via OCR), with user-defined post-processing.

Qrcode Attendence System with Opencv and Pyzbar

A real-time dolly zoom camera effect

Code for CVPR 2022 paper "SoftGroup for Instance Segmentation on 3D Point Clouds"

POT : Python Optimal Transport

Python Computer Vision Aim Bot for Roblox's Phantom Forces

The official code for the ICCV-2021 paper "Speech Drives Templates: Co-Speech Gesture Synthesis with Learned Templates".

OpenCVを用いたカメラキャリブレーションのサンプルです。2021/06/21時点でPython実装のある3種類(通常カメラ向け、魚眼レンズ向け(fisheyeモジュール)、全方位カメラ向け(omnidirモジュール))について用意しています。

GDB python tool to pretty print and debug c++ xtensor containers

Crop regions in napari manually

Extract tables from scanned image PDFs using Optical Character Recognition.

Fun program to overlay a mask to yourself using a webcam

天池2021"全球人工智能技术创新大赛"【赛道一】：医学影像报告异常检测 - 第三名解决方案

Code for CVPR 2022 paper "Bailando: 3D dance generation via Actor-Critic GPT with Choreographic Memory"

Um simples projeto para fazer o reconhecimento do captcha usado pelo jogo bombcrypto

A post-processing tool for scanned sheets of paper.

Rotational region detection based on Faster-RCNN.