🌐
GitHub
github.com › HCIILAB › Scene-Text-Recognition
GitHub - HCIILAB/Scene-Text-Recognition · GitHub
Introduction: The Total-Text dataset [39] contains 1,555 images with 11,459 cropped text instance images. It focuses on curved scene text recognition. Images in Total-Text have more than three different orientations, including horizontal, ...
Starred by 618 users
Forked by 116 users
🌐
GitHub
github.com › HCIILAB › Scene-Text-Detection
GitHub - HCIILAB/Scene-Text-Detection · GitHub
Document Analysis and Recognition (ICDAR), 2017 14th IAPR International Conference on. IEEE, 2017, 1: 1429-1434. Paper · Total-Text[74]:Chee C K, Chan C S. Total-text: A comprehensive dataset for scene text detection and recognition.Document Analysis and Recognition (ICDAR), 2017 14th IAPR International Conference on.
Starred by 564 users
Forked by 127 users
🌐
Kaggle
kaggle.com › datasets › ipythonx › totaltextstr
Total Text - Scene Text Recognition
June 19, 2020 - TOTAL-TEXT is a word-level based English curve text dataset. It consists of 1555 images with more than 3 different text orientations: Horizontal, Multi-Oriented, and Curved, one of a kind.
🌐
IEEE Xplore
ieeexplore.ieee.org › document › 8270088
Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition | IEEE Conference Publication | IEEE Xplore
Text in curve orientation, despite being one of the common text orientations in real world environment, has close to zero existence in well received scene text datasets such as ICDAR'13 and MSRA-TD500. The main motivation of Total-Text is to fill this gap and facilitate a new research direction for the scene text community.
🌐
arXiv
arxiv.org › abs › 1710.10400
[1710.10400] Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition
October 28, 2017 - Abstract page for arXiv paper 1710.10400: Total-Text: A Comprehensive Dataset for Scene Text Detection and Recognition
🌐
GitHub
github.com › HCIILAB › Scene-Text-Recognition-Recommendations
GitHub - HCIILAB/Scene-Text-Recognition-Recommendations: Papers, Datasets, Algorithms, SOTA for STR. Long-time Maintaining · GitHub
639 test images instances. It is specifically designed to evaluate perspective distorted textrecognition. It is built based on the original SVT dataset by selecting the images at the sameaddress on Google Street View but with different view angles.
Starred by 354 users
Forked by 39 users
Languages   Python 96.7% | Shell 3.3%
🌐
PubMed Central
pmc.ncbi.nlm.nih.gov › articles › PMC10181526
Lightweight Scene Text Recognition Based on Transformer - PMC
MLT19: A large-scale scene text dataset from multiple Asian countries that includes 10,000 images divided into training, validation, and test sets. The scene text in the dataset includes regular text, digits, letters, and Chinese characters. RCTW17: A scene text recognition dataset from China that includes 8940 images from street and outdoor scenes.
🌐
MDPI
mdpi.com › 1424-8220 › 23 › 9 › 4490
Lightweight Scene Text Recognition Based on Transformer
May 5, 2023 - MLT19: A large-scale scene text dataset from multiple Asian countries that includes 10,000 images divided into training, validation, and test sets. The scene text in the dataset includes regular text, digits, letters, and Chinese characters. RCTW17: A scene text recognition dataset from China that includes 8940 images from street and outdoor scenes.
🌐
ScienceDirect
sciencedirect.com › science › article › pii › S2215098624002672
Turkish scene text recognition: Introducing extensive real and synthetic datasets and a novel recognition model - ScienceDirect
November 8, 2024 - Existing datasets, regardless of the language, tend to grapple with issues such as limited sample quantity and high noise levels, which considerably restrict the progression and overall efficacy of STR research and applications. Addressing these shortcomings, we introduce the Turkish Scene Text Recognition (TS-TR) dataset, one of the most substantial STR datasets to date, comprising 7288 text instances.
Find elsewhere
🌐
GitHub
github.com › chongyangtao › Awesome-Scene-Text-Recognition › blob › master › README.md
Awesome-Scene-Text-Recognition/README.md at master · chongyangtao/Awesome-Scene-Text-Recognition
A curated list of resources dedicated to scene text localization and recognition - chongyangtao/Awesome-Scene-Text-Recognition
Author   chongyangtao
🌐
Towards Data Science
towardsdatascience.com › home › latest › scene text detection and recognition using east and tesseract
Scene Text Detection And Recognition Using EAST And Tesseract | Towards Data Science
January 22, 2025 - There are lots of datasets available publicly which can be used for this task, the different datasets with the year of release, Image number, the orientation of text, language and important features are listed below. Image Source – Paper(Scene Text Detection and Recognition)
🌐
TheCVF
openaccess.thecvf.com › content_ICCV_2019 › papers › Baek_What_Is_Wrong_With_Scene_Text_Recognition_Model_Comparisons_Dataset_ICCV_2019_paper.pdf pdf
What Is Wrong With Scene Text Recognition Model Comparisons?
searches clearly indicate the training datasets used and com- ... Street View and contains 645 images for evaluation. Many of the images contain perspective projections ... Figure 3: Visualization of an example flow of scene text recognition.
🌐
Papers with Code
paperswithcode.com › task › scene-text-recognition
Scene Text Recognition
Kronos, a specialized pre-training framework for financial K-line data, outperforms existing models in forecasting and synthetic data generation through a unique tokenizer and autoregressive pre-training on a large dataset. ... LingBot-Map is a feed-forward 3D foundation model that reconstructs scenes from video streams using a geometric context transformer architecture with specialized attention mechanisms for coordinate grounding, dense geometric cues, and long-range drift correction, achieving stable real-time performance at 20 FPS.
🌐
arXiv
arxiv.org › abs › 2403.08007
[2403.08007] IndicSTR12: A Dataset for Indic Scene Text Recognition
March 12, 2024 - It was created specifically for a group of related languages with different scripts. The dataset contains over 27000 word-images gathered from various natural scenes, with over 1000 word-images for each language.
🌐
arXiv
arxiv.org › abs › 2103.04400
[2103.04400] What If We Only Use Real Datasets for Scene Text Recognition? Toward Scene Text Recognition With Fewer Labels
June 5, 2021 - Abstract:Scene text recognition (STR) task has a common practice: All state-of-the-art STR models are trained on large synthetic data.
🌐
GitHub
github.com › tangzhenyu › Scene-Text-Understanding
GitHub - tangzhenyu/Scene-Text-Understanding: OCR, Scene-Text-Understanding, Text Recognition · GitHub
[2018-arxiv] TextBoxes++: A Single-Shot Oriented Scene Text Detector [Paper] PowerPoint Text Detection and Recognition Dataset 2017
Starred by 379 users
Forked by 116 users
Languages   C++ 50.6% | Jupyter Notebook 39.6% | Python 5.7% | Cuda 2.0% | CMake 1.1% | MATLAB 0.4%
🌐
GitHub
github.com › Mountchicken › Union14M
GitHub - Mountchicken/Union14M: [ICCV 2023] Code base for Revisiting Scene Text Recognition: A Data Perspective · GitHub
Union14M is a large scene text recognition (STR) dataset collected from 17 publicly available datasets, which contains 4M of labeled data (Union14M-L) and 10M of unlabeled data (Union14M-U), intended to provide a more profound analysis for the ...
Starred by 206 users
Forked by 9 users
Languages   Python 99.5% | Shell 0.2% | Dockerfile 0.1% | JavaScript 0.1% | Batchfile 0.1% | Makefile 0.0%
🌐
Medium
charmve.medium.com › scene-text-detection-suvery-papers-code-da53b1b475d2
Scene-Text-Detection: Suvery, papers & Code | by Charmve | Medium
November 23, 2020 - Introduction: It contains 18000 images in total. It provides word-level annotation. Compared to MLT, this dataset has 10 languages. It is a more real and complex datasets for scene text detection and recognition..