EDBT 2026 Demo / reviewers in the wild / expert
Ray Smith
dblp:44/5593
· DBLP profile ↗
7ranked-venue papers
2as first author
0since 2021 · last 2013
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Databases, data management, data science and information retrieval · 5 · 2 first-authorArtificial intelligence and machine learning · 4 · 2 first-authorGraphics, computer vision, multimedia, augmented reality and games · 2
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer graphics and multimedia
1 paper |
Visualization and visual analytics · 100% | |
| Human-computer interaction and pervasive computing
1 paper |
User interface design and tools · 100% |
Topics — the 1 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Visualization and visual analytics › focus+context visualization
distortion-oriented views |
0.0 | 1 | 1995 | STAR: A General Architecture for the Support of Distortion Oriented Displays · KDD 1995 |
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2013 | A Simple Equation Region Detector for Printed Document Images in TesseractabstractDetecting equation regions from scanned books has received attention in the document image research community in the past few years. Compared with regular text blocks, equation regions have more complicated layouts so we can not simply use text lines to model them. On the other hand, these regions consist of text symbols that can be reflowed, so that the OCR engines should parse them instead of rasterizing them like image regions. In this paper, we present an equation detector with two major contributions: (i) it is built on a simple algorithm that uses the density of special symbols, such that no additional classifier is required, (ii) it has been built into the open source Tesseract that can be accessed and used by the OCR community. The algorithm is tested on the Google Books database with 1534 entries sampled from books/magazines/newspapers of over thirty languages. And we show that Tesseract performance is improved after enabling the detector. Ray Smith |
ICDAR | 2 |
| 2011 | Limits on the Application of Frequency-Based Language Models to OCRabstractAlthough large language models are used in speech recognition and machine translation applications, OCR systems are "far behind" in their use of language models. The reason for this is not the laggardness of the OCR community, but the fact that, at high accuracies, a frequency-based language model can do more damage than good, unless carefully applied. This paper presents an analysis of this discrepancy with the help of the Google Books n-gram Corpus, and concludes that noisy-channel models that closely model the underlying classifier and segmentation errors are required. Ray Smith |
ICDAR | 1 |
| 2010 | Table detection in heterogeneous documentsabstractDetecting tables in document images is important since not only do tables contain important information, but also most of the layout analysis methods fail in the presence of tables in the document image. Existing approaches for table detection mainly focus on detecting tables in single columns of text and do not work reliably on documents with varying layouts. This paper presents a practical algorithm for table detection that works with a high accuracy on documents with varying layouts (company reports, newspaper articles, magazine pages, ...). An open source implementation of the algorithm is provided as part of the Tesseract OCR engine. Evaluation of the algorithm on document images from publicly available UNLV dataset shows competitive performance in comparison to the table detection module of a commercial OCR system. Faisal Shafait, Ray Smith |
Document Analysis Systems | 2 |
| 2005 | Browsing Texture Image DatabasesabstractThe MPEG-7 standard defines two types of texture features: texture retrieval descriptor (TRD) for retrieval and texture browsing descriptor (TBD)for browsing. The retrieval process is straightforward but it is unclear how one could use TBD for browsing. This paper describes two methods of generating layouts for browsing a texture image database. The layouts are then subject to quantitative and qualitative evaluations. The experiments showed that: (1) only some features of TBD are appropriate for browsing, (2) once the inappropriate features are removed TBD is good for browsing only if the textures are structured, (3) the layouts generated usingTRDare more suitable for browsing. Suryani Lim, Lianping Chen, Guojun Lu, Ray Smith |
MMM | 4 |
| 2001 | Application of Multimedia to the Study of Human Movement
Chris Kirtley, Ray Smith |
Multim. Tools Appl. | 2 |
| 1995 | A simple and efficient skew detection algorithm via text row accumulationabstractAn important part of any document recognition system is detection of skew in the image of a page. This paper presents a new, accurate and robust skew detection algorithm based on a method for finding rows of text in page images. Results of a comparison of the new algorithm against Baird's well-known algorithm on 400 pages show the new algorithm to be more accurate, robust and somewhat faster. In particular, the new algorithm only breaks down at skew angles in excess of 15 degrees, compared to the almost uniform distribution of breakdowns of Baird's algorithm. Ray Smith |
ICDAR | 1 |
| 1995 | STAR: A General Architecture for the Support of Distortion Oriented Displays
Ray Smith |
KDD | 2 |