EDBT 2026 Demo / reviewers in the wild / expert
Anant Jhingran
dblp:55/603
· DBLP profile ↗
17ranked-venue papers
6as first author
0since 2021 · last 2006
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Databases, data management, data science and information retrieval · 16 · 6 first-authorSystems, architecture and hardware · 1Applied, interdisciplinary, general and emerging computing · 1
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Databases, data mining, and information retrieval
12 papers |
Query processing and optimization · 39% Database system architecture and tuning · 20% Knowledge graphs · 17% | |
| Computer architecture, parallel and distributed computing, and storage systems
7 papers |
Distributed systems · 34% Storage systems · 24% Cloud and datacenter computing · 20% |
Topics — the 28 heaviest of 36, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Query processing and optimization › materialized view
materialized view selection |
0.1 | 2 | 2004 | A Wavelet Framework for Adapting Data Cube Views for OLAP · IEEE Trans. Knowl. Data Eng. 2004 Dynamic Assembly of Views in Data Cubes · PODS 1998 |
Query processing and optimization
OLAP |
0.0 | 1 | 2004 | A Wavelet Framework for Adapting Data Cube Views for OLAP · IEEE Trans. Knowl. Data Eng. 2004 |
Knowledge graphs
ontology |
0.0 | 1 | 2003 | SemTag and seeker: bootstrapping the semantic web via automated semantic annotation · WWW 2003 |
Knowledge graphs › semantic web
semantic annotation |
0.0 | 1 | 2003 | SemTag and seeker: bootstrapping the semantic web via automated semantic annotation · WWW 2003 |
Data integration and cleaning › entity resolution
block processing |
0.0 | 1 | 2001 | Block Oriented Processing of Relational Database Operations in Modern Computer Architectures · ICDE 2001 |
Query processing and optimization › OLAP
data cube |
0.0 | 1 | 1998 | Dynamic Assembly of Views in Data Cubes · PODS 1998 |
Query processing and optimization › materialized view
view materialization |
0.0 | 1 | 1998 | Dynamic Assembly of Views in Data Cubes · PODS 1998 |
Transaction processing and concurrency control
recovery |
0.0 | 1 | 1997 | Recovery Analysis of Data Sharing Systems under Deferred Dirty Page Propagation Policies · IEEE Trans. Parallel Distributed Syst. 1997 |
Storage systems › data management › database storage
object-oriented database storage |
0.0 | 2 | 1991 | Precomputation in a Complex Object Environment · ICDE 1991 Alternatives in Complex Object Representation: A Performance Perspective · ICDE 1990 |
Indexing and storage engines
multidimensional indexing |
0.0 | 1 | 2004 | A Wavelet Framework for Adapting Data Cube Views for OLAP · IEEE Trans. Knowl. Data Eng. 2004 |
Database system architecture and tuning
parallel database system |
0.0 | 1 | 1995 | An Overview of DB2 Parallel Edition · SIGMOD Conference 1995 |
Query processing and optimization › query optimization
parallel query optimization |
0.0 | 1 | 1995 | An Overview of DB2 Parallel Edition · SIGMOD Conference 1995 |
Database system architecture and tuning › parallel database system
shared-nothing architecture |
0.0 | 1 | 1995 | An Overview of DB2 Parallel Edition · SIGMOD Conference 1995 |
Information retrieval
text analysis |
0.0 | 1 | 2003 | SemTag and seeker: bootstrapping the semantic web via automated semantic annotation · WWW 2003 |
Storage systems
database recovery |
0.0 | 1 | 1992 | Analysis of Recovery in a Database System Using a Write-Ahead Log Protocol · SIGMOD Conference 1992 |
Distributed systems › fault tolerance
failure recovery |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Distributed systems
fault tolerance |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Distributed systems › fault tolerance › failure recovery
log-based recovery |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Distributed systems
replication |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Storage systems
storage reliability |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Query processing and optimization
query result caching |
0.0 | 1 | 1990 | Alternatives in Complex Object Representation: A Performance Perspective · ICDE 1990 |
Database system architecture and tuning › active database
rule system |
0.0 | 1 | 1990 | On Rules, Procedures, Caching and Views in Data Base Systems · SIGMOD Conference 1990 |
Data models and query languages › database views
view support |
0.0 | 1 | 1990 | On Rules, Procedures, Caching and Views in Data Base Systems · SIGMOD Conference 1990 |
Query processing and optimization
query optimization |
0.0 | 1 | 1988 | A Performance Study of Query Optimization Algorithms on a Database System Supporting Procedures · VLDB 1988 |
Query processing and optimization › query optimization
query optimization algorithms |
0.0 | 1 | 1988 | A Performance Study of Query Optimization Algorithms on a Database System Supporting Procedures · VLDB 1988 |
Distributed and cloud data management
data partitioning |
0.0 | 1 | 1995 | An Overview of DB2 Parallel Edition · SIGMOD Conference 1995 |
Performance modeling and evaluation
analytical modeling |
0.0 | 1 | 1992 | Analysis of Recovery in a Database System Using a Write-Ahead Log Protocol · SIGMOD Conference 1992 |
Distributed systems › fault tolerance
high availability |
0.0 | 1 | 1992 | An Efficient Scheme for Providing High Availability · SIGMOD Conference 1992 |
Methods — techniques the papers use, named apart from their topics
sorting · 0.1block-oriented processing · 0.1aggregation expression evaluation · 0.1wavelet decomposition · 0.0greedy algorithm · 0.0disambiguation algorithm · 0.0taxonomy · 0.0qualitative evaluation · 0.0pending update count distribution · 0.0analytical modeling · 0.0queueing model · 0.0asynchronous replication · 0.0LRU buffer replacement · 0.0simulation · 0.0knapsack-based algorithm · 0.0
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2006 | Enterprise Information Mashups: Integrating Information, Simply
Anant Jhingran |
VLDB | 1 |
| 2004 | A Wavelet Framework for Adapting Data Cube Views for OLAPabstractThis article presents a method for adaptively representing multidimensional data cubes using wavelet view elements in order to more efficiently support data analysis and querying involving aggregations. The proposed method decomposes the data cubes into an indexed hierarchy of wavelet view elements. The view elements differ from traditional data cube cells in that they correspond to partial and residual aggregations of the data cube. The view elements provide highly granular building blocks for synthesizing the aggregated and range-aggregated views of the data cubes. We propose a strategy for selectively materializing alternative sets of view elements based on the patterns of access of views. We present a fast and optimal algorithm for selecting a non-expansive set of wavelet view elements that minimizes the average processing cost for supporting a population of queries of data cube views. We also present a greedy algorithm for allowing the selective materialization of a redundant set of view element sets which, for measured increases in storage capacity, further reduces processing costs. Experiments and analytic results show that the wavelet view element framework performs better in terms of lower processing and storage cost than previous methods that materialize and store redundant views for online analytical processing (OLAP). John R. Smith, Chung-Sheng Li, Anant Jhingran |
IEEE Trans. Knowl. Data Eng. | 3 |
| 2003 | SemTag and seeker: bootstrapping the semantic web via automated semantic annotationabstractThis paper describes Seeker, a platform for large-scale text analytics, and SemTag, an application written on the platform to perform automated semantic tagging of large corpora. We apply SemTag to a collection of approximately 264 million web pages, and generate approximately 434 million automatically disambiguated semantic tags, published to the web as a label bureau providing metadata regarding the 434 million annotations. To our knowledge, this is the largest scale semantic tagging effort to date.We describe the Seeker platform, discuss the architecture of the SemTag application, describe a new disambiguation algorithm specialized to support ontological disambiguation of large-scale data, evaluate the algorithm, and present our final results with information about acquiring and making use of the semantic tags. We argue that automated large scale semantic tagging of ambiguous content can bootstrap and accelerate the creation of the semantic web. Stephen Dill, Nadav Eiron, David Gibson, Daniel Gruhl, Ramanathan V. Guha, Anant Jhingran, Tapas Kanungo, Sridhar Rajagopalan, Andrew Tomkins, John A. Tomlin, Jason Y. Zien |
WWW | 6 |
| 2003 | A case for automated large-scale semantic annotation
Stephen Dill, Nadav Eiron, David Gibson, Daniel Gruhl, Ramanathan V. Guha, Anant Jhingran, Tapas Kanungo, Kevin S. McCurley, Sridhar Rajagopalan, Andrew Tomkins, John A. Tomlin, Jason Y. Zien |
J. Web Semant. | 6 |
| 2001 | Block Oriented Processing of Relational Database Operations in Modern Computer ArchitecturesabstractDatabase systems are not well-tuned to take advantage of modern superscalar processor architectures. In particular, the clocks per instruction (CPI) for rather simple database queries are quite poor compared to scientific kernels or SPEC benchmarks. The lack of performance of database systems has been attributed to poor utilization of caches and processor function units as well as higher branching penalties. In this paper, we argue that a block-oriented processing strategy for database operations can lead to better utilization of the processors and caches, generating significantly higher performance. We have implemented the block-oriented processing technique for aggregation expression evaluation and sorting operations as a feature in the DB2 Universal Database (UDB) system. We present results from representative queries on a 30-GB TPC-H (Transaction Processing Council Benchmark H) database to show the value of this technique. Sriram Padmanabhan, Timothy Malkemus, Ramesh C. Agarwal, Anant Jhingran |
ICDE | 4 |
| 2000 | Anatomy of a Real E-Commerce SystemabstractToday's E-Commerce systems are a complex assembly of databases, web servers, home grown glue code, and networking services for security and scalability. The trend is towards larger pieces of these coming together in bundled offerings from leading software vendors, and the networking/hardware being offered through service delivery companies. In this paper we examine the bundle by looking in detail at IBM's WebSphere, Commerce Edition, and its deployment at a major customer site. Anant Jhingran |
SIGMOD Conference | 1 |
| 2000 | A 20/20 Vision of the VLDB-2020?
S. Misbah Deen, Anant Jhingran, Shamkant B. Navathe, Erich J. Neuhold, Gio Wiederhold |
VLDB | 2 |
| 1998 | Dynamic Assembly of Views in Data CubesabstractIn this paper, we present a method for dynamically assembling views in multi-dimensional data cubes in order to more e#ciently support data analysis and querying involving aggregations. The proposed method decomposes the data cubes into an indexed hierarchy of view elements. The view elements di#er from traditional data cube cells in that they correspond to partial and residual aggregations of the data cube. The view elements provide highly granular building blocks for synthesizing the aggregated and rangeaggregated views of the data cubes. We propose a strategy for selecting and materializing the view elements based on the frequency of view access. This allows the dynamic adaptation of the view element sets to patterns of retrieval. We present a fast and optimal algorithm for selecting non-expansive view element sets that minimize the processing costs for generating a population of aggregated views. We also present a greedy algorithm for selecting redundant view element sets in order... John R. Smith, Chung-Sheng Li, Vittorio Castelli, Anant Jhingran |
PODS | 4 |
| 1997 | Interfacing Parallel Applications and Parallel DatabasesabstractThe use of parallel database systems to deliver high performance has become quite common. Although queries submitted to these database systems are executed in parallel, the interaction between applications and current parallel database systems is serial. As the complexity of the applications and the amount of data they access increases, the need to parallelize applications also increases. In this parallel application environment, a serial interface to the database could become the bottleneck in the performance of the application. Hence, parallel database systems should support interfaces that allow the applications to interact with the database system in parallel. We present a taxonomy of such parallel interfaces, namely the Single Coordinator, Multiple Coordinator, Hybrid Parallel, and Pure Parallel interfaces. Furthermore, we discuss how each of these interfaces can be realized and in the process introduce new constructs that enable the implementation of the interfaces. We also qualitatively evaluate each of the interfaces with respect to their restrictiveness and performance impact. Vibby Gottemukkala, Anant Jhingran, Sriram Padmanabhan |
ICDE | 2 |
| 1997 | Recovery Analysis of Data Sharing Systems under Deferred Dirty Page Propagation PoliciesabstractIn a multinode data sharing environment, different buffer coherency control schemes based on various lock retention mechanisms can be designed to exploit the concept of deferring the propagation or writing of dirty pages to disk to improve normal performance. Two types of deferred write policies are considered. One policy only propagates dirty pages to disk at the times when dirty pages are flushed out of the buffer under LRU buffer replacement. The other policy also performs writes at the times when dirty pages are transferred across nodes. The dirty page propagation policy can have significant implications on the database recovery time. In this paper, we provide an analytical modeling framework for the analysis of the recovery times under the two deferred write policies. We demonstrate how these policies can be mapped onto a unified analytic modeling framework. The main challenge in the analysis is to obtain the pending update count distribution which can be used to determine the average numbers of log records and data I/Os needed to be applied during recovery. The analysis goes beyond previous work on modeling buffer hit probability in a data sharing system where only the average buffer composition, not the distribution, needs to be estimated, and recovery analysis in a single node environment where the complexities on tracking the propagation of dirty pages across nodes and the buffer invalidation effect do not appear. Asit Dan, Philip S. Yu, Anant Jhingran |
IEEE Trans. Parallel Distributed Syst. | 3 |
| 1995 | An Overview of DB2 Parallel EditionabstractIn this paper, we describe the architecture and features of DB2 Parallel Edition (PE). DB2 PE belongs to the IBM family of open DB2 client/server database products including DB2/6000, DB2/2, DB2 for HP-UX, and DB2 for the Solaris Operating Environment. DB2 PE employs a shared nothing architecture in which the database system consists of a set of independent logical database nodes. Each logical node represents a collection of system resources including, processes, main memory, disk storage, and communications, managed by an independent database manager. The logical nodes use message passing to exchange data with each other. Tables are partitioned across nodes using a hash partitioning strategy. The cost-based parallel query optimizer takes table partitioning information into account when generating parallel plans for execution by the runtime system. A DB2 PE system can be configured to contain one or more logical nodes per physical processor. For example, the system can be configured to implement one node per processor in a shared-nothing, MPP system or multiple nodes in a symmetric multiprocessor (SMP) system. This paper provides an overview of the storage model, query optimization, runtime system, utilities, and performance of DB2 Parallel Edition. Chaitanya K. Baru, Gilles Fecteau, Ambuj Goyal, Hui-I Hsiao, Anant Jhingran, Sriram Padmanabhan, Walter G. Wilson |
SIGMOD Conference | 5 |
| 1992 | An Efficient Scheme for Providing High AvailabilityabstractReplication at the partition level is a promising approach for increasing availability in a Shared Nothing architecture. We propose an algorithm for maintaining replicas with little overhead during normal failure-free processing. Our mechanism updates the secondary replica in an asynchronous manner: entire dirty pages are sent to the secondary at some time before they are discarded from primary's buffer. A log server node (hardened against failures) maintains the log for each node. If a primary node fails, the secondary fetches the log from the log server, applied it to its replica, and brings itself to the primary's last transaction-consistent state. We study the performance of various policies for sending pages to secondary and the corresponding trade-offs between recovery time and overhead during failure-free processing. Anupam Bhide, Ambuj Goyal, Hui-I Hsiao, Anant Jhingran |
SIGMOD Conference | 4 |
| 1992 | Analysis of Recovery in a Database System Using a Write-Ahead Log ProtocolabstractIn this paper we examine the recovery time in a database system using a Write-Ahead Log protocol, such as ARIES [9], under the assumption that the buffer replacement policy is strict LRU. In particular, analytical equations for log read time, data I/O, log application, and undo processing time are presented. Our initial model assumes a read/write ratio of one, and a uniform access pattern. This is later generalized to include different read/write ratios, as well as a “hot set” model (i.e. x% of the accesses go to y% of the data). We show that in the uniform access model, recovery is dominated by data I/O costs, but under extreme hot-set conditions, this may no longer be true. Furthermore, since we derive anaytical equations, recovery can be analyzed for any set of parameter conditions not discussed here. Anant Jhingran, Pratap Khedkar |
SIGMOD Conference | 1 |
| 1991 | Precomputation in a Complex Object EnvironmentabstractCertain analytical results are established about precomputation in a complex object environment. The concept of 'intervals' is introduced and it is shown that object-identifier caching might be beneficial provided the frequency of updates to the set of subobjects associated with an object is not too high. However, in most cases, procedures with value caching outperformed other forms of representations. Certain performance characteristics of a knapsack-based algorithm are established which can be used to optimally decide which of the precomputed results to cache. Simulation results demonstrate how well this scheme performs. In particular, a binary search strategy seems ideally suited for caching.> Anant Jhingran |
ICDE | 1 |
| 1990 | Alternatives in Complex Object Representation: A Performance PerspectiveabstractWith database systems finding wider use in CAD, office information systems, and logic programming applications, the importance of efficiently representing and manipulating complex objects is growing. In this study a classification of the alternatives for representing complex objects is examined. Consideration is given to the performance aspects of one representation technique based on object identifiers. It is shown that clustering of subobjects with their referencing objects is rarely a good idea. In contrast, it is shown that caching the intermediate results of query processing can yield large benefits.> Anant Jhingran, Michael Stonebraker |
ICDE | 1 |
| 1990 | On Rules, Procedures, Caching and Views in Data Base SystemsabstractThis paper demonstrates that a simple rule system can be constructed that supports a more powerful view system than available in current commercial systems. Not only can views be specified by using rules but also special semantics for resolving ambiguous view updates are simply additional rules. Moreover, procedural data types as proposed in POSTGRES are also efficiently simulated by the same rules system. Lastly, caching of the action part of certain rules is a possible performance enhancement and can be applied to materialize views as well as to cache procedural data items. Hence, we conclude that a rule system is a fundamental concept in a next generation DBMS, and it subsumes both views and procedures as special cases. Michael Stonebraker, Anant Jhingran, Jeffrey Goh, Spyros Potamianos |
SIGMOD Conference | 2 |
| 1988 | A Performance Study of Query Optimization Algorithms on a Database System Supporting Procedures
Anant Jhingran |
VLDB | 1 |