Dean Jacobs

dblp:96/2901 · DBLP profile ↗
← Back
24ranked-venue papers
7as first author
0since 2021 · last 2013
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 15 · 1 first-authorTheory of computation · 5 · 3 first-authorSoftware engineering, systems software and programming languages · 3 · 3 first-authorApplied, interdisciplinary, general and emerging computing · 1

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Computer architecture, parallel and distributed computing, and storage systems
7 papers
Cloud and datacenter computing · 77% Storage systems · 14% Performance modeling and evaluation · 10%
Databases, data mining, and information retrieval
9 papers
Distributed and cloud data management · 64% Data models and query languages · 18% Data integration and cleaning · 11%

Topics — the 22 heaviest of 28, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Distributed and cloud data management › cloud database
multi-tenant database
0.332011
Extensibility and Data Sharing in evolving multi-tenant databases · ICDE 2011
A comparison of flexible schemas for software as a service · SIGMOD Conference 2009
Multi-tenant databases for software as a service: schema-mapping techniques · SIGMOD Conference 2008
Cloud and datacenter computing › multi-tenancy
tenant placement
0.322013
RTP: robust tenant placement for elastic in-memory database clusters · SIGMOD Conference 2013
Predicting in-memory database performance for automating cluster management tasks · ICDE 2011
Cloud and datacenter computing › cloud service models
software as a service
0.232011
Extensibility and Data Sharing in evolving multi-tenant databases · ICDE 2011
A comparison of flexible schemas for software as a service · SIGMOD Conference 2009
Multi-tenant databases for software as a service: schema-mapping techniques · SIGMOD Conference 2008
Cloud and datacenter computing
cluster resource management and scheduling
0.212013
RTP: robust tenant placement for elastic in-memory database clusters · SIGMOD Conference 2013
Cloud and datacenter computing › cluster resource management and scheduling
cluster resource management
0.112011
Predicting in-memory database performance for automating cluster management tasks · ICDE 2011
Cloud and datacenter computing
multi-tenancy
0.112011
Extensibility and Data Sharing in evolving multi-tenant databases · ICDE 2011
Performance modeling and evaluation › performance prediction
response time estimation
0.112011
Predicting in-memory database performance for automating cluster management tasks · ICDE 2011
Storage systems › file systems
versioning
0.112011
Extensibility and Data Sharing in evolving multi-tenant databases · ICDE 2011
Cloud and datacenter computing
cloud data management
0.112010
Cloudy skies for data management · ICDE 2010
Data models and query languages › schema management
schema evolution
0.112009
A comparison of flexible schemas for software as a service · SIGMOD Conference 2009
Data integration and cleaning
schema mapping
0.112008
Multi-tenant databases for software as a service: schema-mapping techniques · SIGMOD Conference 2008
Storage systems › file systems
transactional file system
0.112005
A high-performance, transactional filestore for application servers · SIGMOD Conference 2005
Data models and query languages
database programming language
0.011996
Heraclitus: Elevating Deltas to be First-Class Citizens in a Database Programming Language · ACM Trans. Database Syst. 1996
Data models and query languages › query language
relational calculus
0.011993
Safety and Translation of Calculus Queries with Scalar Functions · PODS 1993
Database system architecture and tuning
active database
0.011991
Language Constructs for Programming Active Databases · VLDB 1991
Programming languages and type systems
logic programming
0.011990
Type Declarations as Subtype Constraints in Logic Programming · PLDI 1990
Programming languages and type systems › type systems › polymorphism
parametric polymorphism
0.011990
Type Declarations as Subtype Constraints in Logic Programming · PLDI 1990
Programming languages and type systems › type systems › subtyping
subtype constraints
0.011990
Type Declarations as Subtype Constraints in Logic Programming · PLDI 1990
Programming languages and type systems
type checking
0.011990
Type Declarations as Subtype Constraints in Logic Programming · PLDI 1990
Programming languages and type systems
type systems
0.011990
Type Declarations as Subtype Constraints in Logic Programming · PLDI 1990
Programming languages and type systems
language implementation
0.011993
On Implementing a Language for Specifying Active Database Execution Models · VLDB 1993
Programming languages and type systems › language design
language constructs
0.011991
Language Constructs for Programming Active Databases · VLDB 1991

Methods — techniques the papers use, named apart from their topics

elastic scaling algorithms · 0.3main-memory storage · 0.2XOR encoding · 0.2survey · 0.2performance modeling · 0.1benchmarking · 0.1when operator · 0.0first-class deltas · 0.0query translation · 0.0least model semantics · 0.0horn clause semantics · 0.0
YearPublicationVenuePosition
2013 RTP: robust tenant placement for elastic in-memory database clusters
abstract
In the cloud services industry, a key issue for cloud operators is to minimize operational costs. In this paper, we consider algorithms that elastically contract and expand a cluster of in-memory databases depending on tenants' behavior over time while maintaining response time guarantees.
Jan Schaffner, Tim Januschowski, Megan Kercher, Tim Kraska, Hasso Plattner, Michael J. Franklin, Dean Jacobs
SIGMOD Conference7
2011 Strict SLAs for Operational Business Intelligence
abstract
Today, SLAs for SaaS business applications usually lack stringent service level objectives and significant penalties. Moreover, Operational Business Intelligence features of modern business applications, like analytic dashboards, result in mixed workloads which make it even more difficult to predict execution times accurately due to resource contention. In contrast to the traditional three-tier architecture, an architecture for SaaS business applications should combine application and database layer to allow for processing business transactions and queries according to a queuing approach which enables strict SLAs with stringent response time and throughput guarantees. With stricter SLAs it would be easier to compare different cloud offerings with on-premise solutions and thus cloud computing could become more attractive for potential customers.
Michael Seibold, Alfons Kemper, Dean Jacobs
IEEE CLOUD3
2011 Extensibility and Data Sharing in evolving multi-tenant databases
abstract
Software-as-a-Service applications commonly consolidate multiple businesses into the same database to reduce costs. This practice makes it harder to implement several essential features of enterprise applications. The first is support for master data, which should be shared rather than replicated for each tenant. The second is application modification and extension, which applies both to the database schema and master data it contains. The third is evolution of the schema and master data, which occurs as the application and its extensions are upgraded. These features cannot be easily implemented in a traditional DBMS and, to the extent that they are currently offered at all, they are generally implemented within the application layer. This approach reduces the DBMS to a `dumb data repository' that only stores data rather than managing it. In addition, it complicates development of the application since many DBMS features have to be re-implemented. Instead, a next-generation multi-tenant DBMS should provide explicit support for Extensibility, Data Sharing and Evolution. As these three features are strongly related, they cannot be implemented independently from each other. Therefore, we propose FLEXSCHEME which captures all three aspects in one integrated model. In this paper, we focus on efficient storage mechanisms for this model and present a novel versioning mechanism, called XOR Delta, which is based on XOR encoding and is optimized for main-memory DBMSs.
Stefan Aulbach, Michael Seibold, Dean Jacobs, Alfons Kemper
ICDE3
2011 Predicting in-memory database performance for automating cluster management tasks
abstract
In Software-as-a-Service, multiple tenants are typically consolidated into the same database instance to reduce costs. For analytics-as-a-service, in-memory column databases are especially suitable because they offer very short response times. This paper studies the automation of operational tasks in multi-tenant in-memory column database clusters. As a prerequisite, we develop a model for predicting whether the assignment of a particular tenant to a server in the cluster will lead to violations of response time goals. This model is then extended to capture drops in capacity incurred by migrating tenants between servers. We present an algorithm for moving tenants around the cluster to ensure that response time goals are met. In so doing, the number of servers in the cluster may be dynamically increased or decreased. The model is also extended to manage multiple copies of a tenant's data for scalability and availability. We validated the model with an implementation of a multi-tenant clustering framework for SAP's in-memory column database TREX.
Jan Schaffner, Benjamin Eckart, Dean Jacobs, Christian Schwarz 0001, Hasso Plattner, Alexander Zeier
ICDE3
2010 Cloudy skies for data management
abstract
This paper discusses about what cloud computing means for data management from an industrial perspective, illustrate challenges, opportunities, technologies and impact. We will also ask where cloud computing stands on the hype curve, if we are reinventing the wheel with some technologies, and outline important research questions the data management community should address to move cloud computing to the next level.
David Campbell, Brian Cooper, Dean Jacobs, Ashok Joshi, Volker Markl, Srinivas Narayanan
ICDE3
2009 Principles for Inconsistency
Shel Finkelstein, Dean Jacobs, Rainer Brendle
CIDR2
2009 A comparison of flexible schemas for software as a service
abstract
A multi-tenant database system for Software as a Service (SaaS) should offer schemas that are flexible in that they can be extended different versions of the application and dynamically modified while the system is on-line. This paper presents an experimental comparison of five techniques for implementing flexible schemas for SaaS. In three of these techniques, the database "owns" the schema in that its structure is explicitly defined in DDL. Included here is the commonly-used mapping where each tenant is given their own private tables, which we take as the baseline, and a mapping that employs Sparse Columns in Microsoft SQL Server. These techniques perform well, however they offer only limited support for schema evolution in the presence of existing data. Moreover they do not scale beyond a certain level. In the other two techniques, the application "owns" the schema in that it is mapped into generic structures in the database. Included here are XML in DB2 and Pivot Tables in HBase. These techniques give the application complete control over schema evolution, however they can produce a significant decrease in performance. We conclude that the ideal database for SaaS has not yet been developed and offer some suggestions as to how it should be designed.
Stefan Aulbach, Dean Jacobs, Alfons Kemper, Michael Seibold
SIGMOD Conference2
2008 Implementing Software as a Service
abstract
Summary form only given. This paper describes basic architectures and best practices for implementing software as a service application. In this context, achieving good margins requires making careful engineering trade-offs between adding features and lowering total cost of ownership. Achieving good system utilization requires that businesses share resources using either virtual machines (OS-level virtualization) or multi-tenancy (application-level virtualization). While multi-tenancy achieves greater levels of consolidation, it limits the kinds of features that can be provided. Thus the ideal level for virtualization depends on the characteristics of the application and its users.
Dean Jacobs
EDOC1
2008 Multi-tenant databases for software as a service: schema-mapping techniques
abstract
In the implementation of hosted business services, multiple tenants are often consolidated into the same database to reduce total cost of ownership. Common practice is to map multiple single-tenant logical schemas in the application to one multi-tenant physical schema in the database. Such mappings are challenging to create because enterprise applications allow tenants to extend the base schema, e.g., for vertical industries or geographic regions. Assuming the workload stays within bounds, the fundamental limitation on scalability for this approach is the number of tables the database can handle. To get good consolidation, certain tables must be shared among tenants and certain tables must be mapped into fixed generic structures such as Universal and Pivot Tables, which can degrade performance.
Stefan Aulbach, Torsten Grust, Dean Jacobs, Alfons Kemper, Jan Rittinger
SIGMOD Conference3
2005 The Geek-Tones: An Experiment in Distributed, Real-time Musical Integration
Michael J. Carey 0001, Dean Jacobs, Leonard J. Seligman
CIDR2
2005 A high-performance, transactional filestore for application servers
abstract
There is a class of data, including messages and business workflow state, for which conventional monolithic databases are less than ideal. Performance and scalability of Application Server systems can be dramatically increased by distributing such data across transactional filestores, each of which is bound to a server instance in a cluster. This paper describes a high-performance, transactional filestore that has been developed for the BEA WebLogic Application ServerTM and benchmarks it against a database. The filestore uses a novel, platform-independent disk scheduling algorithm to minimize the latency of small, synchronous writes to disk.
Bill Gallagher, Dean Jacobs, Anno Langen
SIGMOD Conference2
2003 Distributed Computing with BEA WebLogic Server
Dean Jacobs
CIDR1
1996 Heraclitus: Elevating Deltas to be First-Class Citizens in a Database Programming Language
abstract
Traditional database systems provide a user with the ability to query and manipulate one database state, namely the current database state. However, in several emerging applications, the ability to analyze “what-if” scenarios in order to reason about the impact of an update (before committing that update) is of paramount importance. Example applications include hypothetical database access, active database management systems, and version management, to name a few. The central thesis of the Heraclitus paradigm is to provide flexible support for applications such as these by elevating deltas , which represent updates proposed against the current database state, to be first-class citizens. Heraclitus[Alg,C] is a database programming language that extends C to incorporate the relational algebra and deltas. Operators are provided that enable the programmer to explicitly construct, combine, and access deltas. Most interesting is the when operator, that supports hypothetical access to a delta: the expression E when σ yields the value that side effect free expression E would have if the value of delta expression σ were applied to the current database state. This article presents a broad overview of the philosophy underlying the Heraclitus paradigm, and describes the design and prototype implementation of Heraclitus[Alg, C]. A model-independent formalism for the Heraclitus paradigm is also presented. To illustrate the utility of Heraclitus, the article presents an in-depth discussion of how Heraclitus[Alg, C] can be used to specify, and thereby implement, a wide range of execution models for rule application in active databases; this includes both prominent execution models presented in the literature, and more recent “customized” execution models with novel features.
Shahram Ghandeharizadeh, Richard Hull 0001, Dean Jacobs
ACM Trans. Database Syst.3
1993 Safety and Translation of Calculus Queries with Scalar Functions
abstract
Article Free Access Share on Safety and translation of calculus queries with scalar functions Authors: Martha Escobar-Molano View Profile , Richard Hull View Profile , Dean Jacobs View Profile Authors Info & Claims PODS '93: Proceedings of the twelfth ACM SIGACT-SIGMOD-SIGART symposium on Principles of database systemsAugust 1993Pages 253–264https://doi.org/10.1145/153850.153909Published:01 August 1993Publication History 19citation245DownloadsMetricsTotal Citations19Total Downloads245Last 12 Months9Last 6 weeks1 Get Citation AlertsNew Citation Alert added!This alert has been successfully added and will be sent to:You will be notified whenever a record that you have chosen has been cited.To manage your alert preferences, click on the button below.Manage my AlertsNew Citation Alert!Please log in to your account Save to BinderSave to BinderCreate a New BinderNameCancelCreateExport CitationPublisher SiteeReaderPDF
Martha Escobar-Molano, Richard Hull 0001, Dean Jacobs
PODS3
1993 On Implementing a Language for Specifying Active Database Execution Models
Shahram Ghandeharizadeh, Richard Hull 0001, Dean Jacobs, Jaime Castillo, Martha Escobar-Molano, Shih-Hui Lu, Junhui Luo, Chiu Tsang
VLDB3
1992 Implementation of Delayed Updates in Heraclitus
Shahram Ghandeharizadeh, Richard Hull 0001, Dean Jacobs
EDBT3
1991 Language Constructs for Programming Active Databases
Richard Hull 0001, Dean Jacobs
VLDB2
1990 Multiple Specialization of Logic Programs with Run-Time Test
Dean Jacobs, Anno Langen, William H. Winsborough
ICLP1
1990 Type Declarations as Subtype Constraints in Logic Programming
abstract
This paper presents a type system for logic programs that supports parametric polymorphism and subtypes. This system follows most knowledge representation and object-oriented schemes in that subtyping is name-based, i.e., τ1 is considered to be a subtype of τ2 iff it is declared as such. We take this as a fundamental principle in the sense that type declarations have the form of subtype constraints. Types are assigned meaning by viewing such constraints as Horn clauses that, together with a few basic axioms, define a subtype predicate. This technique provides a (least) model for types and, at the same time, a sound and complete proof system for deriving subtypes. Using this proof system, we define well-typedness conditions which ensure that a logic program/query respects a set of predicate types. We prove that these conditions are consistent in the sense that every atom of every resolvent produced during the execution of a well-typed program is consistent with its type.
Dean Jacobs
PLDI1
1990 Compatibility Problems in the Development of Algebraic Module Specifications
Hartmut Ehrig, Werner Fey, Horst Hansen, Michael Löwe, Dean Jacobs, Francesco Parisi-Presicce
Theor. Comput. Sci.5
1989 Algebraic Software Development Concepts for Module and Configuration Families
Hartmut Ehrig, Werner Fey, Horst Hansen, Michael Löwe, Dean Jacobs
FSTTCS5
1988 Compilation of Logic Programs for Restricted And-Parallelism
Dean Jacobs, Anno Langen
ESOP1
1988 Corrections to "A Synthesis of Several Sorting Algorithms" by J. Darlington
Dean Jacobs, Martin Feather
Acta Informatica1
1985 General Correctness: A Unification of Partial and Total Correctness
Dean Jacobs, David Gries
Acta Informatica1