James J. Lu

dblp:35/4982 · DBLP profile ↗
← Back
34ranked-venue papers
16as first author
0since 2021 · last 2015
0000-0001-7888-7412ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 16 · 10 first-authorDatabases, data management, data science and information retrieval · 10 · 3 first-authorTheory of computation · 7 · 3 first-authorApplied, interdisciplinary, general and emerging computing · 4Software engineering, systems software and programming languages · 2Graphics, computer vision, multimedia, augmented reality and games · 2Human-computer interaction and ubiquitous computing · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Databases, data mining, and information retrieval
4 papers
Knowledge graphs · 60% Query processing and optimization · 22% Data models and query languages · 17%
Network and information security
1 paper
Privacy and data protection · 100%
Theoretical computer science
2 papers
Logic in computer science · 84% Computational complexity · 8% Graph algorithms and graph theory · 8%

Topics — the 15 heaviest of 16, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Knowledge graphs
entity linking
0.212013
LinkIT: privacy preserving record linkage and integration via transformations · SIGMOD Conference 2013
Knowledge graphs › entity linking
privacy-preserving record linkage
0.212013
LinkIT: privacy preserving record linkage and integration via transformations · SIGMOD Conference 2013
Privacy and data protection
privacy-preserving record linkage
0.212013
LinkIT: privacy preserving record linkage and integration via transformations · SIGMOD Conference 2013
Data models and query languages
XML
0.112006
XML Evolution: A Two-phase XML Processing Model Using XML Prefiltering Techniques · VLDB 2006
Data models and query languages
object-oriented data model
0.012001
Probabilistic object bases · ACM Trans. Database Syst. 2001
Query processing and optimization
query rewriting
0.012001
Probabilistic object bases · ACM Trans. Database Syst. 2001
Logic in computer science › logic programming
constraint logic programming
0.011996
Hybrid Knowledge Bases · IEEE Trans. Knowl. Data Eng. 1996
Logic in computer science
logic programming
0.011996
Hybrid Knowledge Bases · IEEE Trans. Knowl. Data Eng. 1996
Logic in computer science
nonmonotonic reasoning
0.011996
Hybrid Knowledge Bases · IEEE Trans. Knowl. Data Eng. 1996
Logic in computer science › logic programming
stable model semantics
0.011996
Hybrid Knowledge Bases · IEEE Trans. Knowl. Data Eng. 1996
Query processing and optimization › view maintenance
incremental view maintenance
0.011995
Efficient Maintenance of Materialized Mediated Views · SIGMOD Conference 1995
Query processing and optimization
view maintenance
0.011995
Efficient Maintenance of Materialized Mediated Views · SIGMOD Conference 1995
Transaction processing and concurrency control
consistency
0.012001
Probabilistic object bases · ACM Trans. Database Syst. 2001
Graph algorithms and graph theory › directed graph
AND/OR graph
0.011989
And-Or Graphs Applied to RUE Resolution · IJCAI 1989
Computational complexity › proof complexity
resolution
0.011989
And-Or Graphs Applied to RUE Resolution · IJCAI 1989

Methods — techniques the papers use, named apart from their topics

frequent variable length gram embedding · 0.3differential privacy · 0.3probabilistic disjunction · 0.0probabilistic conjunction · 0.0stable model semantics · 0.0constraint logic programming · 0.0annotated logic programming · 0.0fixpoint operator · 0.0DRed algorithm · 0.0and/or graph · 0.0
YearPublicationVenuePosition
2015 Identification of Venous Thromboembolism from Electronic Medical Records with Information Extraction
Shuai Zheng 0003, Raymund Dantes, James J. Lu, Sheri Chernetsky Tejedor, Michele Beckman, Asha Krishnaswamy, Lisa Richardson, Fusheng Wang 0001
AMIA3
2013 ASLForm: An Adaptive Self Learning Medical Form Generating System
Shuai Zheng 0003, Fusheng Wang 0001, James J. Lu
AMIA3
2013 LinkIT: privacy preserving record linkage and integration via transformations
abstract
We propose to demonstrate an open-source tool, LinkIT, for privacy preserving record Linkage and Integration via data Transformations. LinkIT implements novel algorithms that support data transformations for linking sensitive attributes, and is designed to work with our previously developed tool, FRIL (Fine-grained Record Integration and Linkage), to provide a complete record linkage solution. LinkIT can be also used as a stand-alone secure transformation tool to link string records. The system uses a novel embedding technique based on frequent variable length grams mined from original records with differential privacy, and utilizes a personalized threshold for performing linkage in the embedded space. Compared to the state-of-the-art secure transformation method [16], LinkIT guarantees stronger privacy with better scalability while achieving comparable utility results.
Luca Bonomi, Li Xiong 0001, James J. Lu
SIGMOD Conference3
2012 Enabling ontology based semantic queries in biomedical database systems
abstract
While current biomedical ontology repositories offer primitive query capabilities, it is difficult or cumbersome to support ontology based semantic queries directly in semantically annotated biomedical databases. The problem may be largely attributed to the mismatch between the models of the ontologies and the databases, and the mismatch between the query interfaces of the two systems. To fully realize semantic query capabilities based on ontologies, we develop a system DBOntoLink to provide unified semantic query interfaces by extending database query languages. With DBOntoLink, semantic queries can be directly and naturally specified as extended functions of the database query languages without any programming needed. DBOntoLink is adaptable to different ontologies through customizations and supports major biomedical ontologies hosted at the NCBO BioPortal. We demonstrate the use of DBOntoLink in a real world biomedical database with semantically annotated medical image annotations.
Shuai Zheng 0003, Fusheng Wang 0001, James J. Lu, Joel H. Saltz
CIKM3
2011 Key-Based Problem Decomposition for Relational Constraint Satisfaction Problems
abstract
Constraint satisfaction problems (CSP) are often posed over data residing in relational database systems, which serve as passive data-storage back ends. Several studies have demonstrated a number of important advantages to having database systems capable of natively modelling and solving CSPs. This paper studies the automated decomposition of the input CSP, borrowing another distinctive idea of relational databases, normalization, to improve its solution time. Experimental evaluations for two case studies show the potential benefit of the approach.
James J. Lu, Sebastien Siva
ICTAI1
2009 HIDE: heterogeneous information DE-identification
abstract
While there is an increasing need to share data that may contain personal information, such data sharing must preserve individual privacy without disclosing any identifiable information. A considerable amount of research in the data privacy community has been devoted to formalizing the notion of identifiability with many techniques for anonymization, but is focused exclusively on structured data. On the other hand, efforts on de-identifying medical text documents in the medical informatics community are highly specialized for specific document types or a subset of identifiers. In addition, they rely on simple identifier removal or grouping techniques and do not take advantage of the research developments in the data privacy community. We developed an integrated system, HIDE, for Heterogeneous Information DE-identification including structured and unstructured data utilizing existing anonymization techniques. We demonstrate a prototype of our system and show the effectiveness of our approach through a set of real data augmented with synthesized data.
James J. Gardner, Li Xiong 0001, Kanwei Li, James J. Lu
EDBT4
2009 Thinking about computational thinking
abstract
Jeannette Wing's call for teaching Computational Thinking (CT) as a formative skill on par with reading, writing, and arithmetic places computer science in the category of basic knowledge. Just as proficiency in basic language arts helps us to effectively communicate and in basic math helps us to successfully quantitate, proficiency in computational thinking helps us to systematically and efficiently process information and tasks. But while teaching everyone to think computationally is a noble goal, there are pedagogical challenges. Perhaps the most confounding issue is the role of programming, and whether we can separate it from teaching basic computer science. How much programming, if any, should be required for CT proficiency?
James J. Lu, George Fletcher 0001
SIGCSE1
2008 FRIL: A Tool for Comparative Record Linkage
Pawel Jurczyk, James J. Lu, Li Xiong 0001, Janet D. Cragan, Adolfo Correa
AMIA2
2008 Comparing and Clustering Flow Cytometry Data
abstract
Flow cytometry technique produces large, multi-dimensional datasets of properties of individual cells that are helpful for biomedical science and clinical research. This paper explores an approach for comparing and clustering flow cytometry data. To overcome challenges posed by the irregularities and the high dimensions of the data, we develop a set of data preprocessing techniques to facilitate effective clustering of flow cytometry data files. We present a set of experiments using real data from the Protective Immunity Project (PIP) showing the effectiveness of the approach.
Li Xiong 0001, James J. Lu, Kim M. Gernert, Vicki Stover Hertzberg
BIBM3
2008 A Case Study in Engineering SQL Constraint Database Systems (Extended Abstract)
Sebastien Siva, James J. Lu, Hantao Zhang 0001
ICLP2
2008 Solving SQL Constraints by Incremental Translation to SAT
Robin Lohfert, James J. Lu, Dongfang Zhao 0001
IEA/AIE2
2006 XML Evolution: A Two-phase XML Processing Model Using XML Prefiltering Techniques
Chia-Hsin Huang, Tyng-Ruey Chuang, James J. Lu, Hahn-Ming Lee
VLDB3
2005 Logical Data Independence Reconsidered (Extended Abstract)
James J. Lu
ISMIS1
2003 A case study in the meta-reasoning procedure ND
abstract
A new technique for improving the efficiency of propositional reasoning procedures is presented. The meta-search procedure, ND, is parameterized by a search procedure P and a real number for controlling the way in which P is applied to the given problem. Experiments using SATO on the domain of Crossword Puzzle Construction (CPC) illustrate the potential for ND. Variations of and future experiments with ND are discussed.
James J. Lu, Jeffrey S. Rosenthal, Andrew E. Shaffer
J. Exp. Theor. Artif. Intell.1
2002 Inference for Annotated Logics over Distributive Lattices
James J. Lu, Neil V. Murray, Heydar Radjavi, Erik Rosenthal, Peter Rosenthal
ISMIS1
2001 Probabilistic object bases
abstract
Although there are many applications where an object-oriented data model is a good way of representing and querying data, current object database systems are unable to handle objects whose attributes are uncertain. In this article, we extend previous work by Kornatzky and Shimony to develop an algebra to handle object bases with uncertainty. We propose concepts of consistency for such object bases, together with an NP-completeness result, and classes of probabilistic object bases for which consistency is polynomially checkable. In addition, as certain operations involve conjunctions and disjunctions of events, and as the probability of conjunctive and disjunctive events depends both on the probabilities of the primitive events involved as well as on what is known (if anything) about the relationship between the events, we show how all our algebraic operations may be performed under arbitrary probabilistic conjunction and disjunction strategies. We also develop a host of equivalence results in our algebra, which may be used as rewrite rules for query optimization. Last but not least, we have developed a prototype probabilistic object base server on top of ObjectStore. We describe experiments to assess the efficiency of different possible rewrite rules.
Thomas Eiter, James J. Lu, Thomas Lukasiewicz, V. S. Subrahmanian
ACM Trans. Database Syst.2
2000 Annotated Hyperresolution for Non-horn Regular Multiple-Valued Logics
James J. Lu, Neil V. Murray, Erik Rosenthal
ISMIS1
1999 A Foundation for Hybrid Knowledge Bases
James J. Lu, Neil V. Murray, Erik Rosenthal
FSTTCS1
1998 A Framework for Automated Reasoning in Multiple-Valued Logics
James J. Lu, Neil V. Murray, Erik Rosenthal
J. Autom. Reason.1
1997 Computing Non-Ground Representations of Stable Models
Thomas Eiter, James J. Lu, V. S. Subrahmanian
LPNMR2
1996 Signed Formula Logic Programming: Operational Semantics and Applications (Extended Abstract)
Jacques Calmet, James J. Lu, Maria Rodriguez, Joachim Schü
ISMIS2
1996 Query Processing in Annotated Logic Programming: Theory and Implementation
Sonia M. Leach, James J. Lu
J. Intell. Inf. Syst.2
1996 Logic Programming with Signs and Annotations
abstract
Signed formula is a formalism that has been applied to reasoning about multiple-valued logics. In this paper, the theory of logic programming based on signed formula is developed, and its connection to annotated logic programming investigated. It is shown that a signed formula logic program, together with annotated logic, forms a paraconsistent basis for reasoning about ‘inconsistent’multiplevalued logic programs. A query processing procedure based on signed resolution is introduced.
James J. Lu
J. Log. Comput.1
1996 Hybrid Knowledge Bases
abstract
Deductive databases that interact with, and are accessed by, reasoning agents in the real world (such as logic controllers in automated manufacturing, weapons guidance systems, aircraft landing systems, land-vehicle maneuvering systems, and air-traffic control systems) must have the ability to deal with multiple modes of reasoning. Specifically, the types of reasoning we are concerned with include, among others, reasoning about time, reasoning about quantitative relationships that may be expressed in the form of differential equations or optimization problems, and reasoning about numeric modes of uncertainty about the domain which the database seeks to describe. Such databases may need to handle diverse forms of data structures, and frequently they may require use of the assumption-based nonmonotonic representation of knowledge. A hybrid knowledge base is a theoretical framework capturing all the above modes of reasoning. The theory tightly unifies the constraint logic programming scheme of Jaffar and Lassez (1987), the generalized annotated logic programming theory of Kifer and Subrahmanian (1989), and the stable model semantics of Gelfond and Lifschitz (1988). New techniques are introduced which extend both the work on annotated logic programming and the stable model semantics.
James J. Lu, Anil Nerode, V. S. Subrahmanian
IEEE Trans. Knowl. Data Eng.1
1995 Efficient Maintenance of Materialized Mediated Views
abstract
Integrating data and knowledge from multiple heterogeneous sources -- like databases, knowledge bases or specific software packages -- is often required for answering certain queries. Recently, a powerful framework for defining mediated views spanning multiple knowledge bases by a set of constrained rules was proposed [24, 4, 16]. We investigate the materialization of these views by unfolding the view definition and the efficient maintenance of the resulting materialized mediated view in case of updates. Thereby, we consider two kinds of updates: updates to the view and updates to the underlying sources. For each of these two cases several efficient algorithms maintaining materialized mediated views are given. We improve on previous algorithms like the DRed algorithm [12] and introduce a new fixpoint operator WP which -- opposed to the standard fixpoint operator TP [9] -- allows us to correctly capture the update's semantics without any recomputation of the materialized view.
James J. Lu, Guido Moerkotte, Joachim Schü, V. S. Subrahmanian
SIGMOD Conference1
1994 Computing Annotated Logic Programs
Sonia M. Leach, James J. Lu
ICLP2
1994 Signed Formulas and Fuzzy Operator Logics
James J. Lu, Neil V. Murray, Erik Rosenthal
ISMIS1
1993 Interpreting Disjunctive Logic Programs Based on a Strong Sense of Disjunction
James J. Lu, Monica D. Barback, Lawrence J. Henschen
J. Autom. Reason.1
1993 Completeness Issues in RUE-NRF Deduction: The Undecidability of Viability
James J. Lu, V. S. Subrahmanian
J. Autom. Reason.1
1992 Minimizing Indefinite Information in Disjunctive Deductive Databases
Monica D. Barback, Jorge Lobo 0001, James J. Lu
ICDT3
1992 The Completeness of GP-Resolution for Annotated Logics
James J. Lu, Lawrence J. Henschen
Inf. Process. Lett.1
1990 Automatic Theorem Proving in Paraconsistent Logics: Theory and Implementation
Newton C. A. da Costa, Lawrence J. Henschen, James J. Lu, V. S. Subrahmanian
CADE3
1990 Protected Completions of First-Order General Logic Programs
James J. Lu, V. S. Subrahmanian
J. Autom. Reason.1
1989 And-Or Graphs Applied to RUE Resolution
Vincent J. Digricoli, James J. Lu, V. S. Subrahmanian
IJCAI2