Son Dao

dblp:66/6340 · DBLP profile ↗
← Back
13ranked-venue papers
3as first author
0since 2021 · last 1999
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 13 · 3 first-authorArtificial intelligence and machine learning · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Databases, data mining, and information retrieval
8 papers
Spatial and temporal data management · 40% Data integration and cleaning · 29% Query processing and optimization · 14%
Computer architecture, parallel and distributed computing, and storage systems
2 papers
Parallel and multicore computing · 74% Distributed systems · 26%
Theoretical computer science
1 paper
Graph algorithms and graph theory · 100%

Topics — the 16 heaviest of 20, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Spatial and temporal data management
moving object databases
0.021998
Cost and Imprecision in Modeling the Position of Moving Objects · ICDE 1998
Modeling and Querying Moving Objects · ICDE 1997
Information retrieval › cross-language information retrieval
query translation
0.021995
Translation of Object-Oriented Queries to Relational Queries · ICDE 1995
Construction of a Relational Front-end for Object-Oriented Database Systems · ICDE 1993
Spatial and temporal data management › spatial data model
spatio-temporal data model
0.011997
Modeling and Querying Moving Objects · ICDE 1997
Spatial and temporal data management
spatio-temporal query language
0.011997
Modeling and Querying Moving Objects · ICDE 1997
Query processing and optimization › adaptive query processing
adaptive query optimization
0.011995
The Fittest Survives: An Adaptive Approach to Query Optimization · VLDB 1995
Data integration and cleaning
heterogeneous data sources
0.011995
Toward Scalability and Interoperability of Heterogeneous Information Sources · ICDE 1995
Data integration and cleaning › schema integration
heterogeneous schema integration
0.011995
Applying a Data Miner To Heterogeneous Schema Integration · KDD 1995
Data integration and cleaning › schema mapping
object-relational mapping
0.011995
Translation of Object-Oriented Queries to Relational Queries · ICDE 1995
Query processing and optimization › query optimization › nested query optimization
query unnesting
0.011995
Efficient Processing of Nested Fuzzy SQL Queries · ICDE 1995
Data integration and cleaning
schema integration
0.011995
Applying a Data Miner To Heterogeneous Schema Integration · KDD 1995
Data integration and cleaning
schema matching
0.011995
Applying a Data Miner To Heterogeneous Schema Integration · KDD 1995
Parallel and multicore computing
parallel graph algorithms
0.011994
A Hybrid Transitive Closure Algorithm for Sequential and Parallel Processing · ICDE 1994
Graph algorithms and graph theory › graph algorithms
transitive closure
0.011994
A Hybrid Transitive Closure Algorithm for Sequential and Parallel Processing · ICDE 1994
Data models and query languages › schema management
schema transformation
0.011993
Construction of a Relational Front-end for Object-Oriented Database Systems · ICDE 1993
Machine learning › Optimization for machine learning
evolutionary computation
0.011995
The Fittest Survives: An Adaptive Approach to Query Optimization · VLDB 1995
Distributed systems › distributed system architecture
interoperability
0.011995
Toward Scalability and Interoperability of Heterogeneous Information Sources · ICDE 1995

Methods — techniques the papers use, named apart from their topics

genetic algorithm · 0.0simulation · 0.0blocking technique · 0.0cost-based analysis · 0.0query processing algorithms · 0.0future temporal logic · 0.0relational predicate graph · 0.0query unnesting · 0.0data mining · 0.0OODB predicate graph · 0.0query translation rules · 0.0predicate graph · 0.0
YearPublicationVenuePosition
1999 Sharing Experiences from Scientific Experiments
abstract
The ESP2Net project is developing technologies that enable effective collaborative scientific data sharing to support collaboration among scientists, accelerate production of scientific data products, and improve understanding of the science. We have defined a Scientific Experiment Markup Language (SEML) to capture scientific experiments in hypermedia documents as a basic unit of information sharing. A collection of SEML documents can be viewed as an online electronic experiment logbook that captures the entire experiment experience by including the process and interrelationships between experiments to allow an experiment to be re-created. Complementary means of sharing the experiences from scientific experiments (browsing, searching, dissemination, and mining) are provided by integrating OASIS transparent distributed scientific object access, Conquest dynamic distributed query processing services, and active information dissemination services introduced in semantic multicast.
Greg Kaestle, Eddie C. Shek, Son Dao
SSDBM3
1999 ASSISS: An Active Semi-Structured Scientific Information Sharing System
abstract
The main concept behind our approach to scientific information sharing is to capture scientific experiments in hypermedia documents as the basic unit of information exchange. We defined SEML (Scientific Experiment Markup Language), based on the XML standard that introduces definitions of tags and links for specific and common information describing scientific experiments. A collection of SEML documents can be viewed as an online electronic experiment logbook that captures the entire experiment. The process and inter-relationships between experiments are modeled to allow an experiment to be re-created, perhaps when input data sets change. We have prototyped ASSISS (Active Semi-Structured Information Sharing System), supporting the browsing, searching and dissemination of XML documents. ASSISS extends scalable active data dissemination technologies developed in the Semantic Multicast Project to provide support for collaborative scientific information dissemination. Within the framework, users' information requests are adaptively clustered. The aggregated information processing and dissemination needs of user groups are then satisfied by a network of dynamic content agents that filters and transforms information in the appropriate manner. SEML is used as the vocabulary against which user profiles are specified, and SEML documents capturing an experiment can be proactively multicasted to groups of users who are interested in some aspect of the activity as the experiments are being conducted.
Eddie C. Shek, Greg Kaestle, Son Dao
SSDBM3
1998 Cost and Imprecision in Modeling the Position of Moving Objects
abstract
Consider a database that represents the location of moving objects, such as taxi-cabs (typical query: "retrieve the cabs that are currently within 1 mile of 33 Michigan Ave., Chicago"), or objects in a battle-field. Existing database management systems (DBMSs) are not well equipped to handle continuously changing data, such as the position of moving objects, since data is assumed to be constant unless it is explicitly modified. In this paper, we address position-update policies and imprecision. Assuming that the actual position of a moving object m deviates from the position computed by the DBMS, when should m update its position in the database in order to eliminate the deviation? Furthermore, how can the DBMS provide a bound on the error (i.e. the deviation) when it replies to a query, such as: "what is the current position of m?" We propose a cost-based approach to update policies that answers both questions. We develop several update policies and analyze them theoretically and experimentally.
Ouri Wolfson, Sam Chamberlain, Son Dao, Liqin Jiang, Gisela Mendez
ICDE3
1997 Modeling and Querying Moving Objects
abstract
We propose a data model for representing moving objects in database systems. It is called the Moving Objects Spatio-Temporal (MOST) data model. We also propose Future Temporal Logic (FTL) as the query language for the MOST model, and devise an algorithm for processing FTL queries in MOST.
A. Prasad Sistla, Ouri Wolfson, Sam Chamberlain, Son Dao
ICDE4
1997 A Parallel Scheme Using the Divide-and-Conquer Method
Qi Yang 0011, Son Dao, Clement T. Yu, Naphtali Rishe
Distributed Parallel Databases2
1996 Information Mediation in Cyberspace: Scalable Methods for Declarative Information Networks
Son Dao, Brad Perry
J. Intell. Inf. Syst.1
1995 Toward Scalability and Interoperability of Heterogeneous Information Sources
abstract
Future large and complex information systems create new challenges and opportunities for research and advanced development in data management. A brief description of Hughes research and prototype efforts to meet these challenges is summarized.>
Son Dao
ICDE1
1995 Efficient Processing of Nested Fuzzy SQL Queries
abstract
Fuzzy databases have been introduced to deal with uncertain or incomplete information in many applications. The efficiency of processing fuzzy queries in fuzzy databases is a major concern. We provide techniques to unnest nested fuzzy queries of two blocks in fuzzy databases. We show both theoretically and experimentally that unnesting improves the performance of nested queries significantly. The results obtained in the paper form the basis for unnesting fuzzy queries of arbitrary blocks in fuzzy databases.>
Qi Yang 0011, Chengwen Liu, Clement T. Yu, Son Dao, Hiroshi Nakajima
ICDE5
1995 Translation of Object-Oriented Queries to Relational Queries
abstract
Proposes a formal approach for translating OODB queries to equivalent relational queries. The translation is accomplished through the use of relational predicate graphs and OODB predicate graphs. One advantage of using such a graph-based approach is that we can achieve bidirectional translation between relational queries and OODB queries.>
Clement T. Yu, Weiyi Meng, Won Kim 0001, Gaoming Wang, Tracy Pham, Son Dao
ICDE7
1995 Applying a Data Miner To Heterogeneous Schema Integration
Son Dao, Brad Perry
KDD1
1995 The Fittest Survives: An Adaptive Approach to Query Optimization
Hongjun Lu, Kian-Lee Tan, Son Dao
VLDB3
1994 A Hybrid Transitive Closure Algorithm for Sequential and Parallel Processing
abstract
A new hybrid algorithm is proposed for well-formed path problems including the transitive closure problem. The CPU time for computation is O(ne), and blocking technique is incorporated to reduce the disk I/O cost in disk-resident environment. The new features of the new algorithm are that only parents sets instead of descendant sets are loaded in from disk, and the computation can be parallelized efficiently. Simulation results show that our algorithm is superior to other existing algorithms in sequential computation, and that linear speedup is achieved in parallel computation.>
Qi Yang 0011, Clement T. Yu, Chengwen Liu, Son Dao, Gaoming Wang, Tracy Pham
ICDE4
1993 Construction of a Relational Front-end for Object-Oriented Database Systems
abstract
Proposes a solution for the construction of a relational front-end for object-oriented database systems (OODBs). Rules are provided to transform the structural part of an OODB scheme to an equivalent relational scheme to provide relational users with a relational view of the OODB scheme. A mechanism based on a relational predicate graph and an OODB predicate graph is provided to translate relational queries to OODB queries to allow relational users access to data stored in an OODB database system.>
Weiyi Meng, Clement T. Yu, Won Kim 0001, Gaoming Wang, Tracy Pham, Son Dao
ICDE6