Scene-Text-Understanding
by tangzhenyu
OCR, Scene-Text-Understanding, Text Recognition
AI summary
Text detection library
A research project focused on developing algorithms and models to accurately detect and recognize text in images and videos from various scenes.
- stars
- 368
- forks
- 112
- watching
- 28
Similar projects
Found by comparing what the projects do, not just their names.
tianzhi0549/ctpn1.3K
Text detector
Detects text in images using a neural network architecture
Text detector
A tool for detecting and translating text from images.
Scene Text Recognizer
A deep learning framework for scene text recognition with rectification and attention mechanisms.
Object detection
An approach to detecting objects in images using multimodal large language models and contextual information
Text classifier
A text classification tool based on LibLinear with support for Chinese tokenize using jieba.
Character encoding library
A C++ library providing iterator and range-based interfaces for encoding and decoding strings in various character encodings.
Text processor
A simple text manipulation library with a fluent interface.
tensorflow/text1.2K
Text processor
Preprocessing and processing tools for text data in machine learning models
NLP model reproducer
Reproduces the results of an ACL 2018 paper on simple word-embedding-based models for natural language processing tasks.
Text analyzer
Language detection library using Bloom filters for speed and memory efficiency.
Object detector
Develops a system to detect, segment, and rank camouflaged objects in images.
Browser detector
A .NET library that identifies characteristics of web browsers by parsing their User-Agent header strings.
Text processing library
A collection of text algorithms and similarity measures
Context-aware detector
A tool for fine-tuning deep neural networks to improve object detection and segmentation capabilities by incorporating contextual information.
Visual Text Understanding
This project enables multi-modal language models to understand and generate text about visual content using referential comprehension.