Peter P. Chen

dblp:c/PPChen · also Peter P. S. Chen, Peter Pin-Shan Chen · DBLP profile ↗
← Back
51ranked-venue papers
24as first author
0since 2021 · last 2010
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 35 · 22 first-authorArtificial intelligence and machine learning · 14 · 4 first-authorTheory of computation · 5 · 1 first-authorGraphics, computer vision, multimedia, augmented reality and games · 2Human-computer interaction and ubiquitous computing · 2Applied, interdisciplinary, general and emerging computing · 2Systems, architecture and hardware · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Theoretical computer science
2 papers
Mathematical optimization · 75% Approximation and online algorithms · 25%
Databases, data mining, and information retrieval
5 papers
Data models and query languages · 86% Database system architecture and tuning · 14%

Topics — the 11 heaviest of 14, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Approximation and online algorithms
approximation algorithms
0.112005
Approximating Pseudo-Boolean Functions on Non-Uniform Domains · IJCAI 2005
Mathematical optimization
discrete optimization
0.112005
Approximating Pseudo-Boolean Functions on Non-Uniform Domains · IJCAI 2005
Mathematical optimization › approximation theory
function approximation
0.112005
Approximating Pseudo-Boolean Functions on Non-Uniform Domains · IJCAI 2005
Mathematical optimization › integer programming
pseudo-boolean optimization
0.112005
Approximating Pseudo-Boolean Functions on Non-Uniform Domains · IJCAI 2005
Data models and query languages
entity-relationship model
0.041986
Entity-Relationship Modeling and Fuzzy Databases · ICDE 1986
An Algebra for a Directional Binary Entity-Relationship Model · ICDE 1984
The Entity-Relationship Model - Toward a Unified View of Data · ACM Trans. Database Syst. 1976
Data models and query languages › entity-relationship model
entity-relationship algebra
0.011986
Entity-Relationship Modeling and Fuzzy Databases · ICDE 1986
Database system architecture and tuning › database design
database design tools
0.011977
Design and Performance Tools for Data Base Systems · VLDB 1977
Requirements engineering and software design
database design
0.011976
The Entity-Relationship Model - Toward a Unified View of Data · ACM Trans. Database Syst. 1976
Mathematical optimization
integer programming
0.011980
Optimal Design of Distributed Information Systems · IEEE Trans. Computers 1980
Database system architecture and tuning
database design
0.011975
The Entity-Relationship Model: Toward a Unified View of Data · VLDB 1975
Data models and query languages
relational model
0.011976
The Entity-Relationship Model - Toward a Unified View of Data · ACM Trans. Database Syst. 1976

Methods — techniques the papers use, named apart from their topics

pseudo-boolean function approximation · 0.1integer programming · 0.0fuzzy set theory · 0.0bounded branch and bound · 0.0entity-relationship diagram · 0.0performance modeling · 0.0benchmarking · 0.0
YearPublicationVenuePosition
2010 Transforms of pseudo-Boolean random variables
Guoli Ding, Robert F. Lax, Jianhua Chen 0003, Peter P. Chen, Brian D. Marx
Discret. Appl. Math.4
2009 Thirty Years of ER Conferences: Milestones, Achievements, and Future Directions
Peter P. Chen
ER1
2009 Scheduling for the Dedicated Machine Constraint in Semiconductor Manufacturing Using a Goal-Oriented Approach
abstract
Semiconductor manufacturing with the dedicated machine constraint is a new challenge for scheduling problems. Previous studies have either not taken this constraint into account or proposed heuristic approaches which might not provide the efficient makespan and the efficient time. This paper proposes a goal-oriented approach called the optimization of makespan, load, and execution time (OMLE) approach which considers the above issues. The desired goals find a schedule for machines which simultaneously optimizes the makespan and the load among machines in an efficient time. Experiments are presented to validate the approach.
Huy Nguyen Anh Pham, Arthur M. D. Shr, Peter P. Chen
ICTAI3
2009 Dedicated Machine Constraint Scheduling as a Shortest-Path Problem
abstract
This paper proposes a graph framework to undertake the issue of scheduling for dedicated machine constraint in the semiconductor manufacturing system. By finding the shortest paths in the graph, the framework finds the best scheduling result with an optimal makespan for the system. The framework first constructs a graph based on the current situation and then schedules wafers to machines under the constraint. Experiments are presented to validate the proposed graph framework.
Huy Nguyen Anh Pham, Arthur M. D. Shr, Peter P. Chen
ICTAI3
2008 Empirical Comparison of Greedy Strategies for Learning Markov Networks of Treewidth k
abstract
We recently proposed the Edgewise Greedy Algorithm (EGA) for learning a decomposable Markov network of treewidth k approximating a given joint probability distribution of n discrete random variables. The main ingredient of our algorithm is the stepwise forward selection algorithm (FSA) due to Deshpande, Garofalakis, and Jordan. EGA is an efficient alternative to the algorithm (HGA) by Malvestuto, which constructs a model of treewidth k by selecting hyperedges of order k+1. In this paper, we present results of empirical studies that compare HGA, EGA and FSA-K which is a straightforward application of FSA, in terms of approximation accuracy (measured by KL-divergence) and computational time. Our experiments show that (1) on the average, all three algorithms produce similar approximation accuracy; (2) EGA produces comparable or better approximation accuracy and is the most efficient among the three. (3) Malvestuto's algorithm is the least efficient one, although it tends to produce better accuracy when the treewidth is bigger than half of the number of random variabls; (4) EGA coupled with local search has the best approximation accuracy overall, at a cost of increased computation time by 50 percent.
K. Nunez, Jianhua Chen 0003, Peter P. Chen, Guoli Ding, Robert F. Lax, Brian D. Marx
ICMLA3
2008 Scheduling for Dedicated Machine Constraint Using Integer Programming
abstract
We propose an integer programming (IP) framework to undertake the dedicated photolithography machine constraint in semiconductor manufacturing. The constraint is one of the new challenges set by the process engineer in semiconductor manufacturing due to natural bias of photolithography machines. Previous researches either did not take the constraint into account or the proposed heuristic approach might not efficiently fit the fast-changing market of semiconductor manufacturing. In this paper, the proposed IP framework provides an approach to minimize the production cost in an efficient time. We also present the experiments to validate the approach.
Huy Nguyen Anh Pham, Arthur M. D. Shr, Peter P. Chen, Alan Liu
ICTAI (1)3
2008 Local Soft Belief Updating for Relational Classification
Guoli Ding, Robert F. Lax, Jianhua Chen 0003, Peter P. Chen, Brian D. Marx
ISMIS4
2008 Formulas for approximating pseudo-Boolean random variables
Guoli Ding, Robert F. Lax, Jianhua Chen 0003, Peter P. Chen
Discret. Appl. Math.4
2007 25th International conference on conceptual modeling (ER 2006)
David W. Embley, Antoni Olivé, Peter P. Chen
Data Knowl. Eng.3
2007 Graph-theoretic method for merging security system specifications
Guoli Ding, Jianhua Chen 0003, Robert F. Lax, Peter P. Chen
Inf. Sci.4
2006 Suggested Research Directions for a New Frontier - Active Conceptual Modeling
Peter P. Chen
ER1
2006 A Heuristic Load Balancing Scheduling Method for Dedicated Machine Constraint
Arthur M. D. Shr, Alan Liu, Peter P. Chen
IEA/AIE3
2006 Using a Multiagent Scheduling System for Dedicated Machine Constraint in Semiconductor Manufacturing
abstract
We present a multiagent scheduling (MS) system to tackle the dedicated machine constraint in this paper. The dedicated machine constraint is one of the new issues of the photolithography machinery due to natural bias. Natural bias will impact the alignment of patterns between different photolithography layers. The dedicated machine constraint is the most important challenge to improve productivity and fulfill the request for customers in semiconductor manufacturing today. In this paper, the proposed MS system is based on a resource schedule and execution matrix (RSEM) and keeps the load balancing among photolithography machines during each scheduling step according to the current load among the photolithography machines in the production line. We describe the prototype system including the agents and the coordination strategies in the paper. We also demonstrate the simulation result that validated the proposed MS system.
Alan Liu, Peter P. Chen, Arthur M. D. Shr, Yen-Ru Cheng
SMC2
2006 A Solution for Dedicated Machine Constraint in Semiconductor Manufacturing
abstract
We propose a heuristic load balancing (LB) scheduling approach based on a resource schedule and execution matrix (RSEM) to tackle the dedicated machine constraint for the photolithography process in semiconductor manufacturing. The constraint of having a dedicated machine is one of the new challenges introduced in photolithography machinery due to natural bias. With this dedicated machine constraint, if we randomly schedule the wafer lots to arbitrary photolithography machines at the first photolithography stage, then the load of all photolithography machines might become unbalanced. However, many scheduling policies or modeling methods proposed by previous research for the semiconductor manufacturing production have not discussed this dedicated machine constraint. In this paper, along with providing the LB approach to the issue of the dedicated machine constraint, we also present a novel model-the representation and manipulation methods for the task patterns. The advantage of LB is to easily schedule the wafer lots by simple calculation on a two-dimensional matrix. We present the result of the simulations to validate our approach as well.
Arthur M. D. Shr, Alan Liu, Peter P. Chen
SMC3
2006 Data Warehouse Design to Support Customer Relationship Management Analysis
abstract
CRM is a strategy that integrates concepts of knowledge management, data mining, and data warehousing in order to support an organization’s decision-making process to retain long-term and profitable relationships with its customers. This research is part of a long-term study to examine systematically CRM factors that affect design decisions for CRM data warehouses in order to build a taxonomy of CRM analyses and to determine the impact of those analyses on CRM data warehousing design decisions. This article presents the design implications that CRM poses to data warehousing and then proposes a robust multidimensional starter model that supports CRM analyses. Additional research contributions include the introduction of two new measures, percent success ratio and CRM suitability ratio by which CRM models can be evaluated, the identification of and classification of CRM queries, and a preliminary heuristic for designing data warehouses to support CRM analyses.
Colleen Cunningham, Il-Yeol Song, Peter P. Chen
J. Database Manag.3
2005 Approximating Pseudo-Boolean Functions on Non-Uniform Domains
Robert F. Lax, Guoli Ding, Peter P. Chen, Jianhua Chen 0003
IJCAI3
2005 Efficient Learning of Pseudo-Boolean Functions from Limited Training Data
Guoli Ding, Jianhua Chen 0003, Robert F. Lax, Peter P. Chen
ISMIS4
2005 New bounds for randomized busing
Steven S. Seiden, Peter P. Chen, Robert F. Lax, Jianhua Chen 0003, Guoli Ding
Theor. Comput. Sci.2
2004 Data warehouse design to support customer relationship management analyses
abstract
CRM is a strategy that integrates the concepts of Knowledge Management, Data Mining, and Data Warehousing in order to support the organization's decision-making process to retain long-term and profitable relationships with its customers. In this paper, we first present the design implications that CRM poses to data warehousing, and then propose a robust multidimensional starter model that supports CRM analyses. We then present sample CRM queries, test our starter model using those queries and define two measures (% success ratio and CRM suitability ratio) by which CRM models can be evaluated. We finally introduce a preliminary heuristic for designing data warehouses to support CRM analyses. Our study shows that our starter model can be used to analyze various profitability analyses such as customer profitability analysis, market profitability analysis, product profitability analysis, and channel profitability analysis.
Colleen Cunningham, Il-Yeol Song, Peter P. Chen
DOLAP3
2004 An analysis of additivity in OLAP systems
abstract
Accurate summary data is of paramount concern in data warehouse systems; however, there have been few attempts to completely characterize the ability to summarize measures. The sum operator is the typical aggregate operator for summarizing the large amount of data in these systems. We look to uncover and characterize potentially inaccurate summaries resulting from aggregating measures using the sum operator. We discuss the effect of classification hierarchies, and non-, semi-, and fully- additive measures on summary data, and develop a taxonomy of the additive nature of measures. Additionally, averaging and rounding rules can add complexity to seemingly simple aggregations. To deal with these problems, we describe the importance of storing metadata that can be used to restrict potentially inaccurate aggregate queries. These summary constraints could be integrated into data warehouses, just as integrity constraints and are integrated into OLTP systems. We conclude by suggesting methods for identifying and dealing with non- and semi- additive attributes.
John Horner, Il-Yeol Song, Peter P. Chen
DOLAP3
2004 Editorial introduction by the Editor-in-Chief
Peter P. Chen
Data Knowl. Eng.1
2004 The best expert versus the smartest algorithm
Peter P. Chen, Guoli Ding
Theor. Comput. Sci.1
2003 Generating r-regular graphs
Guoli Ding, Peter P. Chen
Discret. Appl. Math.2
2002 Editorial introduction
Peter P. Chen, Reind P. van de Riet
Data Knowl. Eng.1
1999 ER Model, XML and the Web
Peter P. Chen
ER1
1998 Introduction to the Special Issue Celebrating the 25th Volume of Data & Knowledge Engineering: DKE
Peter P. Chen, Reind P. van de Riet
Data Knowl. Eng.1
1997 Current Issues of Conceptual Modeling: A Summary of Selective Active Research Topics
Peter P. Chen
Conceptual Modeling1
1997 From Ancient Egyptian Language to Future Conceptual Modeling
Peter P. Chen
Conceptual Modeling1
1997 Future Directions of Conceptual Modeling
Peter P. Chen, Bernhard Thalheim, Leah Y. Wong
Conceptual Modeling1
1997 English, Chinese and ER Diagrams
Peter P. Chen
Data Knowl. Eng.1
1996 Efficient Data Retrieval and Manipulation Using Boolean Entity Lattice
Anyuan Yang, Peter P. Chen
Data Knowl. Eng.2
1992 ER vs. OO
Peter P. Chen
ER1
1989 An Integrity System for a Relational Database Architecture
Asuman Dogac, Esen A. Ozkarahan, Peter P. Chen
ER3
1987 Products from Chen & Associates
Peter P. Chen
ER1
1986 The Lattice Structure of Entity Set
Peter P. Chen, Ming-rui Li
ER1
1986 Entity-Relationship Modeling and Fuzzy Databases
abstract
This work describes an integration of an entity/relationship model and fuzzy databases. It outlines a formal model for representing fuzziness in the ER model, and sketchs a version of the ER algebra, adapted to manipulating fuzzy databases.
Arie Zvieli, Peter P. Chen
ICDE2
1985 The Design and Implementation of an Integrity Subsystem for the Relational DBMS RAP
Asuman Dogac, Peter P. Chen, N. Erol
ER2
1985 Mapping Specifications to Formalisms - Panel Session
John F. Sowa, Peter P. Chen, Peter Freeman 0001, Sharon C. Salveter, Roger C. Schank
ER2
1984 An Algebra for a Directional Binary Entity-Relationship Model
abstract
There are many versions of Entity-Relationship (ER) Models. This paper proposes an algebra for a binary ER model with directional relationships. The proposed algebra can be used as the basis of a data manipulation language for an ER database management system.
Peter P. Chen
ICDE1
1983 ER - A Historical Perspective and Future Directions
Peter P. Chen
ER1
1983 English Sentence Structure and Entity-Relationship Diagrams
Peter P. Chen
Inf. Sci.1
1981 Completeness of Query Languages for the Entity-Relationship Model
Paolo Atzeni, Peter P. Chen
ER2
1981 A Preliminary Framework for Entity-Relationship Models
Peter P. Chen
ER1
1981 A Decomposition of Relations Using the Entity-Relationship Approach
Ilchoo Chung, Fumio Nakamura, Peter P. Chen
ER3
1981 Entity-Relationship Model in the ANSI/SPARC Framework
Asuman Dogac, Peter P. Chen
ER2
1980 Optimal Design of Distributed Information Systems
abstract
In this paper a model is developed for the optimization of distributed information systems. Compared with the previous work in this area, the model is more complete, since it considers simultaneously the distribution of processing power, the allocation of programs and databases, and the assignment of communication line capacities. It also considers the return flow of information, as well as the dependencies between programs and databases. In addition, an algorithm, based on the "bounded branch and bound" integer programming technique, has been developed to obtain the optimal solution of the model. The algorithm is more efficient than several existing general nonlinear integer programming algorithms. Also, it avoids some of the disadvantages of heuristic and decomposition algorithms which are used widely in the optimization of computer networks and distributed databases. The algorithm has been implemented in Fortran, and the computation times of the algorithm for several test problems have been found very reasonable.
Peter P. Chen, Jacky Akoka
IEEE Trans. Computers1
1979 Recent Literature on the Entity-Relationship Approach
Peter P. Chen
ER1
1979 Entity-Relationship Diagrams and English Sentence Structure
Peter P. Chen
ER1
1977 Design and Performance Tools for Data Base Systems
Peter P. Chen
VLDB1
1976 The Entity-Relationship Model - Toward a Unified View of Data
abstract
A data model, called the entity-relationship model, is proposed. This model incorporates some of the important semantic information about the real world. A special diagrammatic technique is introduced as a tool for database design. An example of database design and description using the model and the diagrammatic technique is given. Some implications for data integrity, information retrieval, and data manipulation are discussed. The entity-relationship model can be used as a basis for unification of different views of data: the network model, the relational model, and the entity set model. Semantic ambiguities in these models are analyzed. Possible ways to derive their views of data from the entity-relationship model are presented.
Peter P. Chen
ACM Trans. Database Syst.1
1975 The Entity-Relationship Model: Toward a Unified View of Data
abstract
A data model, called the entity-relationship model, which incorporates the semantic information in the real world is proposed. A special diagramatic technique is introduced for exhibiting entities and relationships. An example of data base design and description using the model and the diagramatic technique is given. The implications on data integrity, information retrieval, and data manipulation are discussed.
Peter P. Chen
VLDB1