Ron Sacks-Davis

dblp:s/RonSacksDavis · DBLP profile ↗
← Back
30ranked-venue papers
9as first author
0since 2021 · last 2000
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 27 · 7 first-authorTheory of computation · 3 · 1 first-authorArtificial intelligence and machine learning · 1Applied, interdisciplinary, general and emerging computing · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Databases, data mining, and information retrieval
13 papers
Information retrieval · 56% Database system architecture and tuning · 15% Data models and query languages · 8%

Topics — the 29 heaviest of 32, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Information retrieval › ranking › text ranking
passage ranking
0.011999
Efficient passage ranking for document databases · ACM Trans. Inf. Syst. 1999
Information retrieval › document retrieval
passage retrieval
0.011999
Efficient passage ranking for document databases · ACM Trans. Inf. Syst. 1999
Query processing and optimization
query execution
0.011999
Efficient passage ranking for document databases · ACM Trans. Inf. Syst. 1999
Information retrieval
search engines
0.021993
Searching Large Lexicons for Partially Specified Terms using Compressed Inverted Files · VLDB 1993
An Efficient Indexing Technique for Full Text Databases · VLDB 1992
Information retrieval
indexing
0.021992
An Efficient Indexing Technique for Full Text Databases · VLDB 1992
A Superimposed Coding Scheme Based on Multiple Block Descriptor Files for Indexing Very Large Data Bases · VLDB 1988
Data models and query languages
query language
0.011995
Atlas: A Nested Relational Database System for Text Applications · IEEE Trans. Knowl. Data Eng. 1995
Information retrieval › indexing › signature file
signature file indexing
0.011995
Atlas: A Nested Relational Database System for Text Applications · IEEE Trans. Knowl. Data Eng. 1995
Information retrieval › indexing
text indexing
0.011995
Atlas: A Nested Relational Database System for Text Applications · IEEE Trans. Knowl. Data Eng. 1995
Information retrieval › document retrieval › structured document retrieval
hypertext retrieval
0.011993
Coherent Answers for a Large Structured Document Collection · SIGIR 1993
Information retrieval › indexing
inverted file
0.011993
Searching Large Lexicons for Partially Specified Terms using Compressed Inverted Files · VLDB 1993
Information retrieval
retrieval models
0.011993
Coherent Answers for a Large Structured Document Collection · SIGIR 1993
Information retrieval › document retrieval
structured document retrieval
0.011993
Coherent Answers for a Large Structured Document Collection · SIGIR 1993
Indexing and storage engines
superimposed coding
0.021988
A Superimposed Coding Scheme Based on Multiple Block Descriptor Files for Indexing Very Large Data Bases · VLDB 1988
Multikey Access Methods Based on Superimposed Coding Techniques · ACM Trans. Database Syst. 1987
Information retrieval › indexing › text indexing
full-text indexing
0.011992
An Efficient Indexing Technique for Full Text Databases · VLDB 1992
Data models and query languages › relational model
nested relational model
0.011991
Efficiency of Nested Relational Document Database Systems · VLDB 1991
Information retrieval › query formulation › query generation
boolean query generation
0.011990
Using Syntactic Analysis in a Document Retrieval System that uses Signature Files · SIGIR 1990
Information retrieval
document retrieval
0.011990
Using Syntactic Analysis in a Document Retrieval System that uses Signature Files · SIGIR 1990
Information retrieval
query processing
0.011990
Using Syntactic Analysis in a Document Retrieval System that uses Signature Files · SIGIR 1990
Information retrieval
signature-based retrieval
0.011990
Using Syntactic Analysis in a Document Retrieval System that uses Signature Files · SIGIR 1990
Spatial and temporal data management › spatial databases
geographic database
0.011989
Extending a DBMS for Geographic Applications · ICDE 1989
Spatial and temporal data management
spatial indexing
0.011989
Extending a DBMS for Geographic Applications · ICDE 1989
Spatial and temporal data management › spatial query processing
spatial query optimization
0.011989
Extending a DBMS for Geographic Applications · ICDE 1989
Spatial and temporal data management
spatial query processing
0.011989
Extending a DBMS for Geographic Applications · ICDE 1989
Indexing and storage engines › multidimensional indexing
multiattribute indexing
0.011987
Multikey Access Methods Based on Superimposed Coding Techniques · ACM Trans. Database Syst. 1987
Information retrieval › document processing › document analysis
document structure
0.011993
Coherent Answers for a Large Structured Document Collection · SIGIR 1993
Information retrieval
hashing
0.011984
Recursive Linear Hashing · ACM Trans. Database Syst. 1984
Indexing and storage engines › hash index › dynamic hashing
linear hashing
0.011984
Recursive Linear Hashing · ACM Trans. Database Syst. 1984
Indexing and storage engines › hierarchical index
two-tier index
0.011987
Multikey Access Methods Based on Superimposed Coding Techniques · ACM Trans. Database Syst. 1987
Indexing and storage engines
file organization
0.011984
Recursive Linear Hashing · ACM Trans. Database Syst. 1984

Methods — techniques the papers use, named apart from their topics

DO-TOS algorithm · 0.0superimposed coding · 0.0node similarity · 0.0document structure exploitation · 0.0syntactic analysis · 0.0natural language processing · 0.0SQL extension · 0.0bit inversion · 0.0batch insertion · 0.0simulation · 0.0
YearPublicationVenuePosition
2000 Architecture of a Content Management Server for XML Document Applications
abstract
Describes the data model that is used to implement the SIM content management server (CMS), an SGML/XML-native content server that is designed to support extremely fast data access to and dynamic updating of 100-GByte collections under high loads. This paper describes the requirements for supporting text-intensive applications and for building XML/SGML document management solutions. The SIM CMS employs a data model that is designed to directly support SGML and XML; this model is described, and a comparison with other models based on general-purpose database management systems is made.
Timothy Arnold-Moore, Michael Fuller, Alan J. Kent, Ron Sacks-Davis, Neil Sharman
WISE4
1999 Efficient passage ranking for document databases
abstract
Queries to text collections are resolved by ranking the documents in the collection and returning the highest-scoring documents to the user. An alternative retrieval method is to rank passages, that is, short fragments of documents, a strategy that can improve effectiveness and identify relevant material in documents that are too large for users to consider as a whole. However, ranking of passages can considerably increase retrieval costs. In this article we explore alternative query evaluation techniques, and develop new tecnhiques for evaluating queries on passages. We show experimentally that, appropriately implemented, effective passage retrieval is practical in limited memory on a desktop machine. Compared to passage ranking with adaptations of current document ranking algorithms, our new “DO-TOS” passage-ranking algorithm requires only a fraction of the resources, at the cost of a small loss of effectiveness.
Marcin Kaszkiel, Justin Zobel, Ron Sacks-Davis
ACM Trans. Inf. Syst.3
1998 The Structured Information Manager (SIM)
abstract
No abstract available.
Ron Sacks-Davis, Alan J. Kent
SIGIR1
1997 An Indexing Scheme for Structured Documents and its Implementation
Tuong Dao, Ron Sacks-Davis, James A. Thom
DASFAA2
1996 The Structured Information Manager: A Database System for SGML Documents
Ron Sacks-Davis
VLDB1
1996 Filtered Document Retrieval with Frequency-Sorted Indexes
abstract
Ranking techniques are effective at finding answers in document collections but can be expensive to evaluate. We propose an evaluation technique that uses early recognition of which documents are likely to be highly ranked to reduce costs; for our test data, queries are evaluated in 2% of the memory of the standard implementation without degradation in retrieval effectiveness. Cpu time and disk traffic can also be dramatically reduced by designing inverted indexes explicitly to support the technique. The principle of the index design is that inverted lists are sorted by decreasing within-document frequency rather than by document number, and this method experimentally reduces cpu time and disk traffic to around one third of the original requirement. We also show that frequency sorting can lead to a net reduction in index size, regardless of whether the index is compressed. © 1996 John Wiley & Sons, Inc.
Michael Persin, Justin Zobel, Ron Sacks-Davis
J. Am. Soc. Inf. Sci.3
1995 A Formal Model for Databases of Structured Text
Brian Lowe, Justin Zobel, Ron Sacks-Davis
DASFAA3
1995 Efficient Retrieval of Partial Documents
Justin Zobel, Alistair Moffat, Ross Wilkinson, Ron Sacks-Davis
Inf. Process. Manag.4
1995 Atlas: A Nested Relational Database System for Text Applications
abstract
Advanced database applications require facilities such as text indexing, image storage, and the ability to store data with a complex structure. However, these facilities are not usually included in traditional database systems. In this paper we describe Atlas, a nested relational database system that has been designed for text-based applications. The Atlas query language is TQL, an SQL-like query language with text operators. The query language is supported by signature file text indexing techniques, and by a parser that can be configured for different text formats and even some foreign languages. Atlas can also be used to store images and audio.>
Ron Sacks-Davis, Alan J. Kent, Kotagiri Ramamohanarao, James A. Thom, Justin Zobel
IEEE Trans. Knowl. Data Eng.1
1994 Memory Efficient Ranking
Alistair Moffat, Justin Zobel, Ron Sacks-Davis
Inf. Process. Manag.3
1993 Coherent Answers for a Large Structured Document Collection
abstract
There is a simple method for integrating information retrieval and hypertext. This consists of treating nodes as isolated documents and retrieving them in order of similarity. If the nodes are structured, in particular, if sets of nodes collectively constitute documents, we can do better. This paper shows how the formation of the hypertext, the retrieval of nodes in response to content based queries, and the presentation of the nodes can be achieved in a way that exploits the knowledge encoded as the structure of the documents. The ideas are then exemplified in an SGML based hypertext information retrieval system.
Michael Fuller, Eric Mackie, Ron Sacks-Davis, Ross Wilkinson
SIGIR3
1993 Searching Large Lexicons for Partially Specified Terms using Compressed Inverted Files
Justin Zobel, Alistair Moffat, Ron Sacks-Davis
VLDB3
1992 An Efficient Indexing Technique for Full Text Databases
Justin Zobel, Alistair Moffat, Ron Sacks-Davis
VLDB3
1992 An environment for building graphical user interfaces for nested relational databases
Evan P. Harris, Alan J. Kent, Ron Sacks-Davis
Inf. Syst.3
1991 Querying in a Large Hyperbase
Michael Fuller, Alan J. Kent, Ron Sacks-Davis, James A. Thom, Ross Wilkinson, Justin Zobel
DEXA3
1991 Efficiency of Nested Relational Document Database Systems
Justin Zobel, James A. Thom, Ron Sacks-Davis
VLDB3
1991 Spatial indexing in binary decomposition and spatial bounding
Beng Chin Ooi, Ron Sacks-Davis, Ken J. McDonell
Inf. Syst.2
1990 Using Syntactic Analysis in a Document Retrieval System that uses Signature Files
abstract
Our work involves the study of the extent to which natural language processing techniques aid the automatic indexing and retrieval of documents. In this paper we describe the use of signature files in large text retrieval systems. We show that good performance can be obtained without requiring the significant overheads required for the inverted file technique. We examine the use of syntactic analysis of the text in all stages of retrieval and argue that an initial Boolean query should be performed that provides a subset of documents, which are then ranked. We then give an algorithm for generating such queries, taking into account the syntactic structure of the queries.
Ron Sacks-Davis
SIGIR1
1990 A signature file scheme based on multiple organizations for indexing very large text databases
abstract
A new signature file method for accessing information from large databases containing both formatted and free text data is presented. The new method, called the multiorganizational scheme is proposed for indexing very large databases containing hundreds of thousands or possibly millions of records. With this method, records are grouped into blocks and signatures are formed for each block of records. These signatures are stored in a block descriptor file using a storage device called the bit slice organization. By forming multiple block descriptor files, each based on a possibly different grouping of records into blocks, it is possible to efficiently determine record matches on query. Both computational results based on a mathematical model as well as experimental results using a library database are presented. These results show that the method provides effective access to large text databases. © 1990 John Wiley & Sons, Inc.
Alan J. Kent, Ron Sacks-Davis, Kotagiri Ramamohanarao
J. Am. Soc. Inf. Sci.2
1989 Partial-match Retrieval using Multiple-Key Hashing with Multiple File Copies
Kotagiri Ramamohanarao, John Shepherd 0001, Ron Sacks-Davis
DASFAA3
1989 Extending a DBMS for Geographic Applications
abstract
A method is presented for extending a conventional DBMS (database management system) for geographic applications. The interface language SQL is augmented to allow formulation of queries involving both spatial and nonspatial selection criteria. A novel indexing structure is supported to facilitate query retrieval that is based on spatial proximity. To enable hybrid queries to be evaluated efficiently, an extended optimization strategy is proposed that evaluates minimal implementation effort.>
Beng Chin Ooi, Ron Sacks-Davis, Ken J. McDonell
ICDE2
1988 A Superimposed Coding Scheme Based on Multiple Block Descriptor Files for Indexing Very Large Data Bases
Alan J. Kent, Ron Sacks-Davis, Kotagiri Ramamohanarao
VLDB2
1987 Multikey Access Methods Based on Superimposed Coding Techniques
abstract
Both single-level and two-level indexed descriptor schemes for multikey retrieval are presented and compared. The descriptors are formed using superimposed coding techniques and stored using a bit-inversion technique. A fast-batch insertion algorithm for which the cost of forming the bit-inverted file is less than one disk access per record is presented. For large data files, it is shown that the two-level implementation is generally more efficient for queries with a small number of matching records. For queries that specify two or more values, there is a potential problem with the two-level implementation in that costs may accrue when blocks of records match the query but individual records within these blocks do not. One approach to overcoming this problem is to set bits in the descriptors based on pairs of indexed terms. This approach is presented and analyzed.
Ron Sacks-Davis, Alan J. Kent, Kotagiri Ramamohanarao
ACM Trans. Database Syst.1
1985 Performance of a multi-key access method based on descriptors and superimposed coding techniques
Ron Sacks-Davis
Inf. Syst.1
1984 Recursive Linear Hashing
abstract
A modification of linear hashing is proposed for which the conventional use of overflow records is avoided. Furthermore, an implementation of linear hashing is presented for which the amount of physical storage claimed is only fractionally more than the minimum required. This implementation uses a fixed amount of in-core space. Simulation results are given which indicate that even for storage utilizations approaching 95 percent, the average successful search cost for this method is close to one disk access.
Kotagiri Ramamohanarao, Ron Sacks-Davis
ACM Trans. Database Syst.2
1983 A two level superimposed coding scheme for partial match retrieval
Ron Sacks-Davis, Kotagiri Ramamohanarao
Inf. Syst.1
1982 Applications of Redundant Number Representations to Decimal Arithmetic
abstract
A decimal arithmetic unit is proposed for both integer and floating-point computations. To achieve comparable speed to a binary arithmetic unit, the decimal unit is based on a redundant number representation. With this representation no loss of compactness is made relative to binary coded decimal (BCD) form. In this paper the hardware required for the implementation of the basic operations of addition, subtraction, multiplication and division are described and the properties of floating-point arithmetic based on a redundant number representation are investigated.
Ron Sacks-Davis
Comput. J.1
1981 Hardware Address Translation for Machines with a Large Virtual Memory
Kotagiri Ramamohanarao, Ron Sacks-Davis
Inf. Process. Lett.2
1980 An Alternative Implementation of Variable Step-Size Multistep Formulas for Stiff ODEs
abstract
An alternative technique for the implementation of variable step-size multistep formulas is developed for the numerical solution of ordinary differential equations.Formulas based on this technique have the property that their leading coefficients are constant; thLs LS important for methods which solve stiff systems.In addition, both theoretmal and empirical results indicate that methods based on this techmque have stability properties similar to those of the corresponding variable coefficient implementations.As a particular example, we have nnplemented the backward differentiation formulas m this form, and the numerical results look very promising.
K. R. Jackson, Ron Sacks-Davis
ACM Trans. Math. Softw.2
1980 Fixed Leading Coefficient Implementation of SD-Formulas for Stiff ODEs
abstract
A code based on Enrlght's second-derivative formulas is described for the numermal solution of stiff ODEs A predictor-corrector approach Is taken, and a fixed leadmg coefficmnt technique is used to unplement the variable step-size formulas This strategy is efficient for stiff systems, and the underlying algorithm IS stable and convergent.The code has been designed consistently with the well-known Adams routine STEP, and advantage Is taken of the mmflarlty between Adams' formulas and Ennght's formulas for an effiment lmplementatmn.
Ron Sacks-Davis
ACM Trans. Math. Softw.1