VLDB 2026 Research / reviewers in the wild / expert
Qinyi Wu
dblp:11/260
· DBLP profile ↗
9ranked-venue papers
5as first author
0since 2021 · last 2017
0009-0009-5555-9315ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Databases, data management, data science and information retrieval · 3 · 3 first-authorSystems, architecture and hardware · 2Software engineering, systems software and programming languages · 2 · 1 first-authorArtificial intelligence and machine learning · 1 · 1 first-authorHuman-computer interaction and ubiquitous computing · 1 · 1 first-author
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Human-computer interaction and pervasive computing
1 paper |
Collaborative and social computing · 100% | |
| Software engineering, system software, and programming languages
2 papers |
Services computing and microservices · 34% Concurrent programming · 34% Program synthesis and code generation · 25% | |
| Databases, data mining, and information retrieval
1 paper |
Indexing and storage engines · 77% Transaction processing and concurrency control · 23% |
Topics — the 8 heaviest of 8, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Indexing and storage engines
persistent data structure |
0.1 | 1 | 2010 | A partial persistent data structure to support consistency in real-time collaborative editing · ICDE 2010 |
Collaborative and social computing
collaborative editing |
0.1 | 1 | 2010 | A partial persistent data structure to support consistency in real-time collaborative editing · ICDE 2010 |
Collaborative and social computing › collaborative editing
consistency maintenance |
0.1 | 1 | 2010 | A partial persistent data structure to support consistency in real-time collaborative editing · ICDE 2010 |
Services computing and microservices
business process management |
0.1 | 1 | 2007 | Categorization and Optimization of Synchronization Dependencies in Business Processes · ICDE 2007 |
Concurrent programming › synchronization
process synchronization |
0.1 | 1 | 2007 | Categorization and Optimization of Synchronization Dependencies in Business Processes · ICDE 2007 |
Program synthesis and code generation › generative programming
modular code generation |
0.1 | 1 | 2005 | Clearwater: extensible, flexible, modular code generation · ASE 2005 |
Transaction processing and concurrency control › versioning
version storage |
0.0 | 1 | 2010 | A partial persistent data structure to support consistency in real-time collaborative editing · ICDE 2010 |
Programming languages and type systems
domain-specific languages |
0.0 | 1 | 2005 | Clearwater: extensible, flexible, modular code generation · ASE 2005 |
Methods — techniques the papers use, named apart from their topics
view synchronization · 0.2dependency optimization · 0.1dataflow programming · 0.1XSLT · 0.1XML-weaving · 0.1XML · 0.1
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2017 | The Millibottleneck Theory of Performance Bugs, and Its Experimental VerificationabstractThe performance of n-tier web-facing applications often suffer from response time long-tail problem. With relatively low resource utilization (less than 50%) and the majority of requests returning within a few milliseconds, a non-negligible num-ber of normally short requests may take seconds to return. We propose the millibottleneck theory of performance bugs (that lead to long-tail problems). Several case studies have confirmed the millibottlenecks (that last a few tens to hundreds of milliseconds) as causal agents of long requests. A concrete example (garbage collection) illustrates the experimental verification of millibottlenecks. An open source fine-grain monitoring toolkit is being devel-oped to facilitate the experimental research on millibottlenecks. Calton Pu, Josh Kimball, Chien-An Lai, Jack Li 0001, Junhee Park, Qingyang Wang 0001, Deepal Jayasinghe, PengCheng Xiong, Simon Malkowski, Qinyi Wu, Gueyoung Jung, Younggyun Koh, Galen S. Swint |
ICDCS | 11 |
| 2010 | Elusive vandalism detection in wikipedia: a text stability-based approachabstractThe open collaborative nature of wikis encourages participation of all users, but at the same time exposes their content to vandalism. The current vandalism-detection techniques, while effective against relatively obvious vandalism edits, prove to be inadequate in detecting increasingly prevalent sophisticated (or elusive) vandal edits. We identify a number of vandal edits that can take hours, even days, to correct and propose a text stability-based approach for detecting them. Our approach is focused on the likelihood of a certain part of an article being modified by a regular edit. In addition to text-stability, our machine learning-based technique also takes into account edit patterns. We evaluate the performance of our approach on a corpus comprising of 15000 manually labeled edits from the Wikipedia Vandalism PAN corpus. The experimental results show that text-stability is able to improve the performance of the selected machine-learning algorithms significantly. Qinyi Wu, Danesh Irani, Calton Pu, Lakshmish Ramaswamy |
CIKM | 1 |
| 2010 | Modeling and implementing collaborative editing systems with transactional techniquesabstractMany collaborative editing systems have been developed for coauthoring documents. These systems generally have different infrastructures and support a subset of interactions found in collaborative environments. In this paper, we propose a transactional framework with two advantages. First, the frame Qinyi Wu, Calton Pu |
CollaborateCom | 1 |
| 2010 | A partial persistent data structure to support consistency in real-time collaborative editingabstractCo-authored documents are becoming increasingly important for knowledge representation and sharing. Tools for supporting document co-authoring are expected to satisfy two requirements: 1) querying changes over editing histories; 2) maintaining data consistency among users. Current tools support either limited queries or are not suitable for loosely controlled collaborative editing scenarios. We address both problems by proposing a new persistent data structure-partial persistent sequence. The new data structure enables us to create unique character identifiers that can be used for associating meta-information and tracking their changes, and also design simple view synchronization algorithms to guarantee data consistency under the presence of concurrent updates. Experiments based on real-world collaborative editing traces show that our data structure uses disk space economically and provides efficient performance for document update and retrieval. Qinyi Wu, Calton Pu, João Eduardo Ferreira |
ICDE | 1 |
| 2010 | Towards Flexible Event-Handling in Workflows through Data StatesabstractDespite recent advances in many real-time and workflow management systems (WFMS), event-handling is still a manual or semi-automated task. The integration of automated event processing with workflows remains an open research challenge to both academic and industrial communities. In this work, we propose a concrete approach that logs interactions between workflow component activities in the form of data states that accurately and efficiently store necessary information for event-handling. Our approach (called WED-flow) explicitly represents various dependencies and constraints of a WFMS in sophisticated data states. Due to the availability of this large amount of historic information, our approach is able to support a flexible event-handling in WFMS. In this paper we present definitions for workflow management systems that incorporate events, and characterize such systems using the WED-flow approach. We also present a scientific workflow example in genetic testing to illustrate the advantages of integrating events with workflow through the WED-flow approach. João Eduardo Ferreira, Qinyi Wu, Simon Malkowski, Calton Pu |
SERVICES | 2 |
| 2007 | Categorization and Optimization of Synchronization Dependencies in Business ProcessesabstractThe current approach for modeling synchronization in business processes relies on sequencing constructs, such as sequence, parallel etc. However, sequencing constructs obfuscate the true source of dependencies in a business process. Moreover, because of the nested structure and scattered code that results from using sequencing constructs, it is hard to add or delete additional constraints without over-specifying necessary constraints or invalidating existing ones. We propose a dataflow programming approach in which dependencies are explicitly modeled to guide activity scheduling. We first give a systematic categorization of dependencies: data, control, service and cooperation. Each dimension models dependency from its own point of view. Then we show that dependencies of various kinds can be first merged and then optimized to generate a minimal dependency set, which guarantees high concurrency and minimal maintenance cost for process execution. Qinyi Wu, Calton Pu, Akhil Sahai, Roger S. Barga |
ICDE | 1 |
| 2006 | DSCWeaver: Synchronization-Constraint Aspect Extension to Procedural Process Specification LanguagesabstractBPEL is emerging as an open-standards language for Web service composition. However, its procedural style can lead to inflexible and tangled code for managing a crosscutting aspect - synchronization constraints that define permissible sequences of execution for activities in a process. In this paper, we present DSCWeaver, a tool that enables a synchronization-aspect extension to BPEL. It uses DSCL, a synchronization expression language, to specify constraints. DSCL has the desirable features of declarative syntax, fine granularity, and validation support. A designer can use DSCL to describe and validate the synchronization behavior and rely on DSCWeaver to generate BPEL code. We demonstrate the advantages of our approach in a service deployment process and evaluate its performance using two metrics: lines of code (LoC) and places to visit (PtV). Evaluation results show that our approach can effectively reduce development effort of process designers while providing performance competitive to un-woven BPEL code Qinyi Wu, Calton Pu, Akhil Sahai, Roger S. Barga, Gueyoung Jung |
ICWS | 1 |
| 2005 | Comparison of Approaches to Service DeploymentabstractIT today is driven by the trend of increasing scale and complexity. Utility and Grid computing models, PlanetLab, and traditional data centers, are reaching the scale of thousands of computers. Installed software consists of dozens of interdependent applications and services. As the complexity and scale of these systems continues to grow, it becomes increasingly difficult to administer and manage them. At the same time, the service deployment technologies are still based on scripts and configuration files with minimal ability to express dependencies, to document and to verify configurations. This results in hard-to-use and erroneous system configurations. Language- and model-based tools, such as SmartFrog and Radia, are proposed for addressing these deployment challenges, but it is unclear whether they are beneficial over traditional solutions. In this paper, we quantitatively compare manual, script-, language-, and model-based deployment solutions as a function of scale, complexity, and susceptibility to change. We also qualitatively compare them in terms of expressiveness and barrier to first use. We demonstrate that script-based solutions are well matched for large scale deployments, language-based for services of large complexity, and model-based for dynamic changes to the design. Finally, we offer a table summarizing rules of thumb regarding which solution to use in which case, subject to deployment needs. Vanish Talwar, Qinyi Wu, Calton Pu, Wenchang Yan, Gueyoung Jung, Dejan S. Milojicic |
ICDCS | 2 |
| 2005 | Clearwater: extensible, flexible, modular code generationabstractDistributed applications typically interact with a number of heterogeneous and autonomous components that evolve independently. Methodical development of such applications can benefit from approaches based on domain-specific languages (DSLs). However, the evolution and customization of heterogeneous components introduces significant challenges to accommodating the syntax and semantics of a DSL in addition to the heterogeneous platforms on which they must run. In this paper, we address the challenge of implementing code generators for two such DSLs that are flexible (resilient to changes in generators or input formats), extensible (able to support multiple output targets and multiple input variants), and modular (generated code can be re-written). Our approach, Clearwater, leverages XML and XSLT standards: XML supports extensibility and mutability for in-progress specification formats, and XSLT provides flexibility and extensibility for multiple target languages. Modularity arises from using XML meta-tags in the code generator itself, which supports controlled addition, subtraction, or replacement to the generated code via XML-weaving. We discuss the use of our approach and show its advantages in two non-trivial code generators: the Infopipe Stub Generator (ISG) to support distributed flow applications, and the Automated Composable Code Translator to support automated distributed application deployment. As an example, the ISG accepts as input an XML description and generates output for C, C++, or Java using a number of communications platforms such as sockets and publish-subscribe. Galen S. Swint, Calton Pu, Gueyoung Jung, Wenchang Yan, Younggyun Koh, Qinyi Wu, Charles Consel, Akhil Sahai, Koichi Moriyama |
ASE | 6 |