GitHub
github.com › HCIILAB › Scene-Text-Recognition
GitHub - HCIILAB/Scene-Text-Recognition · GitHub
Introduction: The Total-Text dataset [39] contains 1,555 images with 11,459 cropped text instance images. It focuses on curved scene text recognition. Images in Total-Text have more than three different orientations, including horizontal, ...
Starred by 618 users
Forked by 116 users
GitHub
github.com › HCIILAB › Scene-Text-Detection
GitHub - HCIILAB/Scene-Text-Detection · GitHub
Document Analysis and Recognition (ICDAR), 2017 14th IAPR International Conference on. IEEE, 2017, 1: 1429-1434. Paper · Total-Text[74]:Chee C K, Chan C S. Total-text: A comprehensive dataset for scene text detection and recognition.Document Analysis and Recognition (ICDAR), 2017 14th IAPR International Conference on.
Starred by 564 users
Forked by 127 users
TheCVF
openaccess.thecvf.com › content › CVPR2021 › papers › Baek_What_if_We_Only_Use_Real_Datasets_for_Scene_Text_CVPR_2021_paper.pdf pdf
What If We Only Use Real Datasets for Scene Text Recognition?
Figure 2: Examples of two major synthetic datasets. SynthText (ST) [11] is originally generated for scene text
IEEE Xplore
ieeexplore.ieee.org › document › 8270088
Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition | IEEE Conference Publication | IEEE Xplore
Text in curve orientation, despite being one of the common text orientations in real world environment, has close to zero existence in well received scene text datasets such as ICDAR'13 and MSRA-TD500. The main motivation of Total-Text is to fill this gap and facilitate a new research direction for the scene text community.
arXiv
arxiv.org › abs › 1710.10400
[1710.10400] Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition
October 28, 2017 - Abstract page for arXiv paper 1710.10400: Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition
GitHub
github.com › HCIILAB › Scene-Text-Recognition-Recommendations
GitHub - HCIILAB/Scene-Text-Recognition-Recommendations: Papers, Datasets, Algorithms, SOTA for STR. Long-time Maintaining · GitHub
639 test images instances. It is specifically designed to evaluate perspective distorted textrecognition. It is built based on the original SVT dataset by selecting the images at the sameaddress on Google Street View but with different view angles.
Starred by 354 users
Forked by 39 users
Languages Python 96.7% | Shell 3.3%
PubMed Central
pmc.ncbi.nlm.nih.gov › articles › PMC10181526
Lightweight Scene Text Recognition Based on Transformer - PMC
MLT19: A large-scale scene text dataset from multiple Asian countries that includes 10,000 images divided into training, validation, and test sets. The scene text in the dataset includes regular text, digits, letters, and Chinese characters. RCTW17: A scene text recognition dataset from China that includes 8940 images from street and outdoor scenes.
MDPI
mdpi.com › 1424-8220 › 23 › 9 › 4490
Lightweight Scene Text Recognition Based on Transformer
May 5, 2023 - MLT19: A large-scale scene text dataset from multiple Asian countries that includes 10,000 images divided into training, validation, and test sets. The scene text in the dataset includes regular text, digits, letters, and Chinese characters. RCTW17: A scene text recognition dataset from China that includes 8940 images from street and outdoor scenes.
ScienceDirect
sciencedirect.com › science › article › pii › S2215098624002672
Turkish scene text recognition: Introducing extensive real and synthetic datasets and a novel recognition model - ScienceDirect
November 8, 2024 - Existing datasets, regardless of the language, tend to grapple with issues such as limited sample quantity and high noise levels, which considerably restrict the progression and overall efficacy of STR research and applications. Addressing these shortcomings, we introduce the Turkish Scene Text Recognition (TS-TR) dataset, one of the most substantial STR datasets to date, comprising 7288 text instances.
GitHub
github.com › chongyangtao › Awesome-Scene-Text-Recognition › blob › master › README.md
Awesome-Scene-Text-Recognition/README.md at master · chongyangtao/Awesome-Scene-Text-Recognition
A curated list of resources dedicated to scene text localization and recognition - chongyangtao/Awesome-Scene-Text-Recognition
Author chongyangtao
TheCVF
openaccess.thecvf.com › content › ICCV2023 › papers › Jiang_Revisiting_Scene_Text_Recognition_A_Data_Perspective_ICCV_2023_paper.pdf pdf
Revisiting Scene Text Recognition: A Data Perspective
prehensive dataset for scene text detection and recognition.
TheCVF
openaccess.thecvf.com › content_ICCV_2019 › papers › Baek_What_Is_Wrong_With_Scene_Text_Recognition_Model_Comparisons_Dataset_ICCV_2019_paper.pdf pdf
What Is Wrong With Scene Text Recognition Model Comparisons?
searches clearly indicate the training datasets used and com- ... Street View and contains 645 images for evaluation. Many of the images contain perspective projections ... Figure 3: Visualization of an example flow of scene text recognition.
Papers with Code
paperswithcode.com › task › scene-text-recognition
Scene Text Recognition
Kronos, a specialized pre-training framework for financial K-line data, outperforms existing models in forecasting and synthetic data generation through a unique tokenizer and autoregressive pre-training on a large dataset. ... LingBot-Map is a feed-forward 3D foundation model that reconstructs scenes from video streams using a geometric context transformer architecture with specialized attention mechanisms for coordinate grounding, dense geometric cues, and long-range drift correction, achieving stable real-time performance at 20 FPS.
arXiv
arxiv.org › abs › 2403.08007
[2403.08007] IndicSTR12: A Dataset for Indic Scene Text Recognition
March 12, 2024 - It was created specifically for a group of related languages with different scripts. The dataset contains over 27000 word-images gathered from various natural scenes, with over 1000 word-images for each language.
arXiv
arxiv.org › abs › 2103.04400
[2103.04400] What If We Only Use Real Datasets for Scene Text Recognition? Toward Scene Text Recognition With Fewer Labels
June 5, 2021 - Abstract:Scene text recognition (STR) task has a common practice: All state-of-the-art STR models are trained on large synthetic data.
GitHub
github.com › tangzhenyu › Scene-Text-Understanding
GitHub - tangzhenyu/Scene-Text-Understanding: OCR, Scene-Text-Understanding, Text Recognition · GitHub
[2018-arxiv] TextBoxes++: A Single-Shot Oriented Scene Text Detector [Paper] PowerPoint Text Detection and Recognition Dataset 2017
Starred by 379 users
Forked by 116 users
Languages C++ 50.6% | Jupyter Notebook 39.6% | Python 5.7% | Cuda 2.0% | CMake 1.1% | MATLAB 0.4%
GitHub
github.com › Mountchicken › Union14M
GitHub - Mountchicken/Union14M: [ICCV 2023] Code base for Revisiting Scene Text Recognition: A Data Perspective · GitHub
Union14M is a large scene text recognition (STR) dataset collected from 17 publicly available datasets, which contains 4M of labeled data (Union14M-L) and 10M of unlabeled data (Union14M-U), intended to provide a more profound analysis for the ...
Starred by 206 users
Forked by 9 users
Languages Python 99.5% | Shell 0.2% | Dockerfile 0.1% | JavaScript 0.1% | Batchfile 0.1% | Makefile 0.0%