VLDB 2026 Research / reviewers in the wild / expert
Di Zhong
dblp:27/2358
· DBLP profile ↗
20ranked-venue papers
8as first author
0since 2021 · last 2018
0000-0001-9645-9112ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 15 · 8 first-authorArtificial intelligence and machine learning · 2Databases, data management, data science and information retrieval · 2Security and privacy · 1Software engineering, systems software and programming languages · 1Applied, interdisciplinary, general and emerging computing · 1
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Software engineering, system software, and programming languages
1 paper |
Debugging and program repair · 44% Program analysis · 44% Concurrent programming · 13% | |
| Network and information security
1 paper |
Cryptographic primitives and cryptanalysis · 50% Cryptographic protocols and secure computation · 50% | |
| Computer graphics and multimedia
3 papers |
Multimedia analysis and retrieval · 92% Image and video processing · 8% |
Topics — the 11 heaviest of 13, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Program analysis
dynamic analysis |
0.3 | 1 | 2018 | Finding broken promises in asynchronous JavaScript programs · Proc. ACM Program. Lang. 2018 |
Debugging and program repair
fault localization |
0.3 | 1 | 2018 | Finding broken promises in asynchronous JavaScript programs · Proc. ACM Program. Lang. 2018 |
Cryptographic primitives and cryptanalysis
authenticated encryption |
0.2 | 1 | 2016 | Efficient Deniably Authenticated Encryption and Its Application to E-Mail · IEEE Trans. Inf. Forensics Secur. 2016 |
Concurrent programming › concurrency models
asynchronous programming |
0.1 | 1 | 2018 | Finding broken promises in asynchronous JavaScript programs · Proc. ACM Program. Lang. 2018 |
Multimedia analysis and retrieval
video retrieval |
0.0 | 2 | 1997 | VideoQ: An Automated Content Based Video Search System Using Visual Cues · ACM Multimedia 1997 Video Parsing, Retrieval and Browsing: An Integrated and Content-Based Solution · ACM Multimedia 1995 |
Multimedia analysis and retrieval › sports video analysis
sports video event detection |
0.0 | 1 | 2001 | Real-time personalized sports video filtering and summarization · ACM Multimedia 2001 |
Multimedia analysis and retrieval › event detection
video event detection |
0.0 | 1 | 2001 | Real-time personalized sports video filtering and summarization · ACM Multimedia 2001 |
Content delivery and video streaming
adaptive video streaming |
0.0 | 1 | 2001 | Real-time personalized sports video filtering and summarization · ACM Multimedia 2001 |
Multimedia analysis and retrieval › multimedia browsing
video browsing |
0.0 | 1 | 1995 | Video Parsing, Retrieval and Browsing: An Integrated and Content-Based Solution · ACM Multimedia 1995 |
Image and video processing › real-time image processing
real-time video processing |
0.0 | 1 | 2001 | Real-time personalized sports video filtering and summarization · ACM Multimedia 2001 |
Multimedia analysis and retrieval › video content analysis › video structure analysis
video parsing |
0.0 | 1 | 1995 | Video Parsing, Retrieval and Browsing: An Integrated and Content-Based Solution · ACM Multimedia 1995 |
Methods — techniques the papers use, named apart from their topics
promise graph · 0.3dynamic analysis · 0.3random oracle model · 0.2multi-resolution content analysis · 0.1compressed-domain analysis · 0.1video object segmentation · 0.0video editing · 0.0spatio-temporal attributes · 0.0content-based indexing · 0.0
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2018 | Finding broken promises in asynchronous JavaScript programsabstractRecently, promises were added to ECMAScript 6, the JavaScript standard, in order to provide better support for the asynchrony that arises in user interfaces, network communication, and non-blocking I/O. Using promises, programmers can avoid common pitfalls of event-driven programming such as event races and the deeply nested counterintuitive control ow referred to as “callback hell”. Unfortunately, promises have complex semantics and the intricate control– and data- ow present in promise-based code hinders program comprehension and can easily lead to bugs. The promise graph was proposed as a graphical aid for understanding and debugging promise-based code. However, it did not cover all promise-related features in ECMAScript 6, and did not present or evaluate any technique for constructing the promise graphs. In this paper, we extend the notion of promise graphs to include all promise-related features in ECMAScript 6, including default reactions, exceptions, and the synchronization operations race and all. Furthermore, we report on the construction and evaluation of PromiseKeeper, which performs a dynamic analysis to create promise graphs and infer common promise anti-patterns. We evaluate PromiseKeeper by applying it to 12 open source promise-based Node.js applications. Our results suggest that the promise graphs constructed by PromiseKeeper can provide developers with valuable information about occurrences of common anti-patterns in their promise-based code, and that promise graphs can be constructed with acceptable run-time overhead. Saba Alimadadi, Di Zhong, Magnus Madsen, Frank Tip |
Proc. ACM Program. Lang. | 2 |
| 2016 | Accelerating mathematical knot simulations with R on the webabstractMany geometric problems of interest to mathematical visualization applications involve changing structures requiring mathematical simulations, such as the moves that transform one knot into an equivalent knot. In this paper, we propose to explore a unique paradigm that makes use of self-deformable object models embedded in mathematical space, supplemented by energy-driven relaxation and user-defined motion constraints to guide the simulation through the configuration space. Furthermore, we exploit web-based user interfaces and parallelization to accelerate mathematical simulations, and to extract key moments where successive terms in the sequence differ by one critical change to represent and analyze various mathematical evolutions. The presented work leverages the nature of the interrelationship between mathematics and computer science, especially computer graphics, graph algorithms, user interfaces, and accelerated computation. Juan Lin 0001, Di Zhong, Yiwen Zhong, Hui Zhang 0006 |
IEEE BigData | 2 |
| 2016 | Efficient Deniably Authenticated Encryption and Its Application to E-MailabstractConfidentiality and authentication are two main security goals in secure electronic mail (e-mail). Pretty good privacy (PGP) and secure/multipurpose internet mail extensions (S/MIME) are two famous secure e-mail solutions. Both PGP and S/MIME use digital envelope to provide message confidentiality and digital signature to provide message authentication. However, these methods have the following two weaknesses: 1) digital signature provides non-repudiation evidence of sender that is not desired in some e-mail applications and 2) efficiency is low, since these methods use two kinds of public key cryptographic primitives: public key encryption and digital signature. To overcome the above two weaknesses, we introduce a new concept called deniably authenticated encryption that can achieve confidentiality, integrity, and deniable authentication in a logical single step. We first propose a deniably authenticated encryption scheme and prove its security in the random oracle model. Then, we design a secure e-mail protocol using the proposed deniably authenticated encryption scheme. The deniable authentication property protects senders' privacy. Fagen Li, Di Zhong, Tsuyoshi Takagi |
IEEE Trans. Inf. Forensics Secur. | 2 |
| 2007 | Enabling MPEG-7 structural and semantic descriptions in retrieval applicationsabstractAbstract The MPEG‐7 standard supports the description of both the structure and the semantics of multimedia; however, the generation and consumption of MPEG‐7 structural and semantic descriptions are outside the scope of the standard. This article presents two research prototype systems that demonstrate the generation and consumption of MPEG‐7 structural and semantic descriptions in retrieval applications. The active system for MPEG‐4 video object simulation (AMOS) is a video object segmentation and retrieval system that segments, tracks, and models objects in videos (e.g., person, car) as a set of regions with corresponding visual features and spatiotemporal relations. The region‐based model provides an effective base for similarity retrieval of video objects. The second system, the Intelligent Multimedia Knowledge Application (IMKA), uses the novel MediaNet framework for representing semantic and perceptual information about the world using multimedia. MediaNet knowledge bases can be constructed automatically from annotated collections of multimedia data and used to enhance the retrieval of multimedia. Ana B. Benitez, Di Zhong, Shih-Fu Chang |
J. Assoc. Inf. Sci. Technol. | 2 |
| 2004 | Real-time view recognition and event detection for sports video
Di Zhong, Shih-Fu Chang |
J. Vis. Commun. Image Represent. | 1 |
| 2001 | MPEG-7 MDS Content Description Tools and Applications
Ana B. Benitez, Di Zhong, Shih-Fu Chang, John R. Smith |
CAIP | 2 |
| 2001 | Long-term moving object segmentation and tracking using spatio-temporal consistencyabstractThe success of object-based media representation and description (e.g., MPEG-4 and -7) depends largely on effective object segmentation tools. We expand our previous work on automatic video region tracking and develop a robust-moving objects detection system. In our system, we first utilize innovative methods of combining color and edge information in improving the object motion estimation results. Then we use the long-term spatio-temporal constraints to achieve reliable object tracking over long sequences. Our extensive experiments demonstrate excellent results in handling challenging cases in general domains (e.g., stock footage) including depth-varying multi-layer background and fast camera motion. Di Zhong, Shih-Fu Chang |
ICIP (2) | 1 |
| 2001 | Structure Analysis of Sports Video Using Domain ModelsabstractIn this paper, we present an effective framework for scene detection and structure analysis for sports videos, using tennis and baseball as examples. Sports video can be characterized by its predictable temporal syntax, recurrent events with consistent features, and a fixed number of views. Our approach combines domain-specific knowledge, supervised machine learning techniques, and automatic feature analysis at multiple levels. Real time processing performance is achieved by utilizing compressed-domain processing techniques. High accuracy in view recognition is achieved by using compressed-domain global features as prefilters and object-level refined analysis in the latter verification stage. Applications include high-level structure browsing/navigation, highlight generation, and mobile Di Zhong, Shih-Fu Chang |
ICME | 1 |
| 2001 | Real-time personalized sports video filtering and summarizationabstractWe demonstrate a real-time fully automated software system for filtering important events in sports video. Events represent occurrences of actions or state changes in video content. In the current prototype, we demonstrate detection of pitching in baseball and serving in tennis. For wireless video applications, we propose and apply a unique notion of content-based adaptive streaming, in which video encoding rate and media modality is dynamically varied according to the event filtering results. Our system includes an event detection module, an adaptive encoding module, and a buffer management module for adaptive streaming. We achieve the real-time performance by exploring compresseddomain techniques and multi-stage multi-resolution contentanalysis processes. Di Zhong, Shih-Fu Chang |
ACM Multimedia | 1 |
| 1999 | Region Feature Based Similarity Searching of Semantic Video ObjectsabstractNew video representations based on semantic objects (e.g., MPEG-4) provide great potential for content-based video searching. In this paper we present an efficient query model for similarity searching of video objects based on localized region features and spatial-temporal structures. The query model is barred on an existing framework on region-level video query, but addresses several new issues like right spatio-temporal relationships effective performance of the proposed approach. Di Zhong, Shih-Fu Chang |
ICIP (2) | 1 |
| 1999 | Searching and Editing MPEG-Compressed Video in a Distributed Online Environment
Horace J. Meng, Di Zhong, Shih-Fu Chang |
Multim. Syst. | 2 |
| 1999 | An integrated approach for content-based video object segmentation and retrievalabstractObject-based video data representations enable unprecedented functionalities of content access and manipulation. We present an integrated approach using region-based analysis for semantic video object segmentation and retrieval. We first present an active system that combines low-level region segmentation with user inputs for defining and tracking semantic video objects. The proposed technique is novel in using an integrated feature fusion framework for tracking and segmentation at both region and object levels. Experimental results and extensive performance evaluation show excellent results compared to existing systems. Building upon the segmentation framework, we then present a unique region-based query system for semantic video object. The model facilitates powerful object search, such as spatio-temporal similarity searching at multiple levels. Di Zhong, Shih-Fu Chang |
IEEE Trans. Circuits Syst. Video Technol. | 1 |
| 1998 | AMOS: An Active System for MPEG-4 Video Object SegmentationabstractObject segmentation and tracking is a fundamental step for many digital video applications. In this paper, we present an active system (AMOS) which combines low level automatic region segmentation with an active method for defining and tracking high-level semantic video objects. The system contains two stages: an initial object segmentation stage where user input in the starting frame is used to create a semantic object; and an object tracking stage where underlying regions of the semantic object are tracked and grouped through successive frames. Experiments with different types of videos show very good performance. Di Zhong, Shih-Fu Chang |
ICIP (2) | 1 |
| 1998 | A fully automated content-based video search engine supporting spatiotemporal queriesabstractThe rapidity with which digital information, particularly video, is being generated has necessitated the development of tools for efficient search of these media. Content-based visual queries have been primarily focused on still image retrieval. In this paper, we propose a novel, interactive system on the Web, based on the visual paradigm, with spatiotemporal attributes playing a key role in video retrieval. We have developed innovative algorithms for automated video object segmentation and tracking, and use real-time video editing techniques while responding to user queries. The resulting system, called VideoQ , is the first on-line video search engine supporting automatic object-based indexing and spatiotemporal queries. The system performs well, with the user being able to retrieve complex video clips such as those of skiers and baseball players with ease. Shih-Fu Chang, William Chen 0001, Horace J. Meng, Hari Sundaram, Di Zhong |
IEEE Trans. Circuits Syst. Video Technol. | 5 |
| 1997 | Spatio-Temporal Video Search Using the Object-Based Video RepresentationabstractObject-based video representation provides great promises for new search and editing functionalities. Feature regions in video sequences are automatically segmented, tracked, and grouped to form the basis for content-based video search and higher levels of abstraction. We present a new system for video object segmentation and tracking using feature fusion and region grouping. We also present efficient techniques for spatio-temporal video query based on the automatically segmented video objects. Di Zhong, Shih-Fu Chang |
ICIP (1) | 1 |
| 1997 | VideoQ: An Automated Content Based Video Search System Using Visual CuesabstractThe rapidity with which digitat information, particularly video, is being generated, has necessitated the development of tools for efficient search of these media.Content based visual queries have been primarily focussed on still image retrieval.In this papel; we propose a novel, real-time, interactive system on the Web, based on the visual paradigm, with spatio-temporal attributesplaying a key role in video retrieval.We have developed algorithms for automated video object segmentation and tracking and use real-time video editing techniques while responding to user queries.The resulting system pe$orms well, with the user being able to retrieve complex video clips such as those of skiers, baseball players, with ease.Penlli~iollto m&e digitnlhrd copies of ail or pa11 ofthis iilfiterinl for personal or clmsroom use is granted without fee provided tht (112 copiare not made or distributed for profit or commercial ndvrmtnge.the COPYridtt notice, tile title ofthe publicntion and its date appear.and notice is given tl]ntcopyrigllt is by permission ofthe ACM, hC.TO Copy othWh& to republisl), to post on servers or to redistribule to lists, requires specific peniiission nnd/or fee.ACM Multimedia 97 ,~en/f/e I~TfJ.~/lingfO17iJX?I Shih-Fu Chang, William Chen 0001, Horace J. Meng, Hari Sundaram, Di Zhong |
ACM Multimedia | 5 |
| 1997 | A distributed system for editing and browsing compressed video over the networkabstractWe present a new framework for distributed video editing and browsing over the network. This framework uses a distributed client-server model including a server engine for content analysis/editing, and clients for interactive controls of video browsing/editing. It utilizes several unique features, including compressed-domain video manipulation, multi-resolution video access, content based video browsing/retrieval, and a distributed network architecture. We have developed a complete functional prototype, CVEPS, which includes a compressed domain video analysis and editing engine, and a JAVA based user interface. It has been incorporated into a World Wide Web application, WebClip, for editing compressed video over the Web. Horace J. Meng, Di Zhong, Shih-Fu Chang |
MMSP | 2 |
| 1997 | WebClip: a WWW video editing/browsing systemabstractWebClip is a complete working prototype for editing/browsing MPEG-1 and MPEG-2 compressed video distributively over the World Wide Web. It uses a general system architecture to store, retrieve, and edit MPEG-1 or MPEG-2 compressed video over the network. It uses an innovative distributed network support architecture. It also uses our unique CVEPS (Compressed video editing, parsing, and search) technologies. Horace J. Meng, Di Zhong, Shih-Fu Chang |
MMSP | 2 |
| 1997 | An integrated system for content-based video retrieval and browsing
HongJiang Zhang, Jianhua Wu 0002, Di Zhong, Stephen W. Smoliar |
Pattern Recognit. | 3 |
| 1995 | Video Parsing, Retrieval and Browsing: An Integrated and Content-Based SolutionabstractNo abstract available. HongJiang Zhang, Chien Yong Low, Stephen W. Smoliar, Di Zhong |
ACM Multimedia | 4 |