Shijun Liu

dblp:04/5268 · DBLP profile ↗
← Back
144ranked-venue papers
1as first author
75since 2021 · last 2026
0000-0002-4108-1391ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Human-computer interaction and ubiquitous computing · 34 · 9 since 2021Software engineering, systems software and programming languages · 27 · 11 since 2021Computer networks · 23 · 20 since 2021Systems, architecture and hardware · 17 · 12 since 2021Applied, interdisciplinary, general and emerging computing · 17 · 9 since 2021Artificial intelligence and machine learning · 13 · 1 first-author · 8 since 2021Databases, data management, data science and information retrieval · 11 · 6 since 2021Graphics, computer vision, multimedia, augmented reality and games · 3 · 2 since 2021
YearPublicationVenuePosition
2026 DHMRec: Collaboration-Guided Multimodal Disentanglement and Hierarchical Fusion for Recommendation
abstract
Multimodal recommender systems have emerged as a pivotal paradigm for harnessing diverse data modalities to deliver personalized services. Contemporary research predominantly focuses on integrating heterogeneous modality information through graph learning. However, these approaches face two key challenges: (1) the inherent complexity of modalities, characterized by entangled redundant signals and noise; and (2) the challenge of effectively integrating multimodal representations, each of which may exert varying degrees of influence on users' preferences. To address these challenges, we propose a novel Collaboration-Guided Multimodal Disentanglement and Hierarchical Fusion for Recommendation (DHMRec), which simultaneously achieves intra-modal denoising disentanglement and inter-modal hierarchical fusion. Specifically, we introduce a collaboration-related modality disentanglement module to distinguish between modality-common and modality-specific features. Then, through multi-view graph learning to capture both item-item dependencies and user-item interaction patterns. Additionally, we implement hierarchical fusion between the disentangled multimodal features and ID embeddings using a positive-negative attention-aware fusion module and an interaction distribution-based alignment module. Extensive experiments on three benchmarks demonstrate that our DHMRec surpasses various state-of-the-art baselines, highlighting its effectiveness in intra-modal disentanglement and multimodal features fusion.
Xiaohan Zhan, Yuliang Shi, Jihu Wang, Shijun Liu, Fanyu Kong 0002
AAAI4
2026 A risk assessment framework for online transactions via Graph Neural Networks and efficient probabilistic prediction
Jicai Chang, Xuejing Fu, Li Pan 0001, Shijun Liu
Eng. Appl. Artif. Intell.5
2026 Hexagonal grid-based representation and generative prediction method for citywide traffic accident situations in urban area
Xueshen Li, Guangxu Mei, Shijun Liu, Sheharyar Khan, Li Pan 0001
Expert Syst. Appl.3
2026 A reinforcement learning-based approach for scheduling ML training tasks in heterogeneous Kubernetes clusters
Shuwei Dong, Bingbing Zheng, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.4
2026 Optimization-based hybrid offloading framework for IoMT in edge-cloud healthcare systems
Sheharyar Khan, Shijun Liu, Li Pan 0001, Guangxu Mei
Future Gener. Comput. Syst.2
2026 A heterogeneous node representation and uncertainty handling approach under local edge-cloud architectures
Haoran Shi 0002, Ying Li 0136, Shijun Liu, Li Pan 0001
Inf. Softw. Technol.4
2026 GT-MARL: Graph- and Transformer-Enhanced Multiagent Reinforcement Learning for Cloud-Edge Collaborative Scheduling
abstract
Scheduling across cloud and edge systems must handle job heterogeneity, dynamic arrivals, and coupled resource constraints. To address the above issues, we present GT-MARL, a reinforcement learning framework with graph and Transformer enhancements for joint task selection and resource allocation across multiple clusters. At each decision epoch, GT-MARL encodes a heterogeneous graph over tasks and resources with Hierarchical Attention Network (HAN) to capture structural dependencies, applies a temporal Transformer to model backlog evolution and delayed interactions, and produces factorized decisions for each cluster through a Transformer-guided task selection head and discrete resource allocation with a Multi-Layer Perceptron (MLP) head. Training follows centralized training with decentralized execution (CTDE) using a centralized critic and masked action spaces. Experiments on both synthetic and trace-replay workloads show that GT-MARL consistently achieves the best p95 job completion time (JCT), while keeping mean JCT competitive with representative heuristic and learning baselines. Specifically, GT-MARL obtains 35.56 mean JCT and 70.76 p95 JCT on the synthetic workload, and 144.21 mean JCT and 264.47 p95 JCT on the trace-replay workload. On the synthetic workload, this corresponds to a 26.3% p95 JCT reduction relative to the best mean baseline with only a 3.0% increase in mean JCT. Additional queueing and service decomposition across workload intensities indicates that the tail-latency gain mainly comes from alleviating queue accumulation rather than shortening the inherent service time of tasks. GT-MARL provides a favorable tradeoff between mean latency and tail latency for scheduling under SLO constraints across cloud and edge systems.
Kaiyuan Qi, Li Pan 0001, Shijun Liu
IEEE Internet Things J.4
2026 DynaAdaFL: A GNN-Assisted MARL Framework for Client Hyperparameter Adaptive Optimization in Edge Federated Learning for Carbon Emission Prediction
Li Pan 0001, Shijun Liu, Weiping Li 0002
IEEE Internet Things J.3
2026 MCIRP: A multi-granularity cross-modal interaction model based on relational propagation for Multimodal Named Entity Recognition with multiple images
Yongheng Mu, Lixu Shao, Shijun Liu, Feng Li 0030, Guangxu Mei
Inf. Process. Manag.5
2026 Learning reward functions via GNNs for multi-agent task placement in edge-cloud LLM services
abstract
Deploying Large Language Models (LLMs) on edge–cloud collaborative clusters leverages cloud elasticity and edge low-latency to mitigate resource bottlenecks. Such deployment, however, necessitates coordinated task offloading at the edge and resource scaling in the cloud. Deep Reinforcement Learning (DRL) methods like Proximal Policy Optimization (PPO) are unable to simultaneously learn strategies for both task offloading and resource scaling. While recent studies have introduced Multi-Agent Reinforcement Learning (MARL) methods for joint optimization, they are limited by the need for extensive expertise in specific cluster configurations to formalize the reward function, making them non-scalable. This paper proposes FGPPO, a Graph Neural Network (GNN)-based distributed multi-agent task placement strategy for the joint optimization of task offloading and cloud resource scaling. FGPPO comprises distributed agents responsible for task offloading and resource scaling, respectively, and employs a GNN-based reward model for distributed reward assignment. The proposed reward model leverages a heterogeneous graph-attention mechanism to holistically integrate service quality, cost, and interdependencies among multiple resource nodes. Experimental results show that FGPPO outperforms the baselines across various real-world traces and cluster configurations. The average cost of FGPPO is 1.25 times that of the offline greedy method, while it significantly reduces the SLA violation rate. These findings indicate that the proposed method holds promise for optimizing costs for LLM providers while ensuring user-side service quality.
Hao Yang 0065, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2025 How LLMs Aid in Domain Modeling: Opportunities and Challenges
abstract
As the complexity of business and scenarios contin-ues to grow, the traditional, inefficient, and cumbersome domain modeling process can no longer adapt to the rapid iteration requirements. In recent years, technological breakthroughs in generative artificial intelligence (Generative AI), particularly in large language models (LLMs), present novel opportunities to enhance domain modeling efficiency. While LLMs demonstrate baseline capabilities in information extraction, their potential for constructing complex domain-specific models remains underexplored. This study investigates how LLMs can facilitate automated domain modeling tasks and support domain modeling education. Through systematic experimentation, we evaluate the impacts of LLM fine-tuning techniques and prompt engineering strategies on model outputs, while comparing two distinct generation modes: indirect model construction via domain element extraction versus direct domain model generation. Our empirical results demonstrate that LLMs exhibit significant potential in supporting high-efficiency domain modeling, with fine-tuning techniques and indirect generation modes yielding superior out-comes. Furthermore, we illustrate LLMs' utility as pedagogical tools for domain modeling education while identifying critical limitations and implementation risks that warrant consideration in both practical applications and future research.
Haoran Shi 0002, Shijun Liu, Li Pan 0001
SSE2
2025 A Format-Character Cooperative Recognition Method for Ancient Chinese Books
abstract
This paper proposes a computer-aided collaborative text recognition algorithm for Chinese ancient books, utilizing synergistic techniques in layout analysis and text recognition to effectively address issues of complex formats and image degradation in ancient book texts. First, the DP-LinkNet architecture is used for adaptive binarization, enhancing the clarity of textual content while reducing noise. In the layout analysis, a “character-based column determination“ strategy is employed in conjunction with the DBNet detection algorithm to precisely locate text regions and filter out non-text elements. Next, the text recognition stage applies a “column-based character determination“ strategy using the SVTR LCNet model, which combines a convolutional network with a sparse Transformer mechanism to ensure strong performance in resource-limited environments. Experimental results demonstrate that this method improves character integrity, recognition accuracy, and processing efficiency compared to traditional approaches. This research provides an effective tool for the digital preservation and automated analysis of ancient books, aiding cultural heritage preservation and supporting further academic research.
Xianya Fu, Shijun Liu
CSCWD3
2025 Zero-Sum vs. Positive-Sum: Effects of Inter-Team Competition Modes and Haptic Feedback on Team Flow in Multi-Team VR
abstract
Virtual reality (VR), particularly through multi-user VR systems, enhances collaboration and immersion. However, multi-team VR systems (MTVR), which allow several teams to interact, either by cooperating or competing, are less studied. Drawing inspiration from concepts in game theory, specifically the positive-sum game (PSG) and zero-sum game (ZSG), this study investigated the impact of inter-team competition modes at the team level and haptic feedback at the individual level on team flow and individual flow for the first time. An experimental MTVR that supports inter-team competition in a two-person vs two-person pattern was implemented first. Then, a$2 \times 2$within-subject experiment was conducted to examine the effects of different inter-team competition modes (PSG/ZSG) and haptic feedback (on/off). The results based on 188 participants indicate that the PSG mode leads to significantly higher levels of collective ambition and improved team relationships compared to the ZSG mode. Furthermore, providing haptic feedback can significantly enhance awareness of shared and personal task goals, resulting in more cautious performance during teamwork.
Qianqian Xiong, Yan Hu 0003, Yulong Bian, Juan Liu 0008, Yichen Hong, Chao Zhou 0012, Wei Gai, Shijun Liu, Chenglei Yang
ISMAR10
2025 EAGA-Net: a novel simulation-based grasping detection dataset and network with efficient adaptability of gripper attribute
Yibang Zhou, Yuhao Ouyang, Shijun Liu, Xiangkai Li, Jun Nan, Xiancheng Ji, Jianjun Yi
Adv. Eng. Informatics3
2025 Long-term river flow forecasting: An integrated deep learning model with multi-scale feature extraction
Qian Li 0003, Shijun Liu, Li Pan 0001
Expert Syst. Appl.3
2025 A Lyapunov Optimization-Based Online Algorithm for Scheduling Cloud-Edge Collaborative Real-Time Video Stream Analytics Tasks
abstract
With the rapid development of cities and the increase in the number of motor vehicles, traditional intelligent traffic video analytics systems face a significant challenge due to soaring computational demands and limited network transmission resources. Today’s widely used cloud-based traffic video analytics system, which is generally reliant on transmitting all video data to centralized cloud servers, suffer from high latency during network fluctuations and inability to respond promptly to urban traffic management. This paper introduces a cloud-edge collaborative framework for traffic video analytics. Within this framework, edge servers serve as intermediaries between video sources and the cloud center. The framework prioritizes the offloading of computing tasks to nodes located near the video sources, rationally leveraging the limited computing and network resources to mitigate network transmission load. We develop a measurement-based analytical model to describe the trade-offs among network latency, inference latency, and analytics accuracy in edge-based real-time video analytics system. As a core element of our approach, we present a resolution selection and bandwidth allocation algorithm based on Lyapunov optimization and heuristic search, designed to dynamically adjust video resolution and distribute bandwidth between edge and cloud servers to balance latency and accuracy without requiring future information. Experiments on a cloud-edge collaborative video analytics system demonstrate the algorithm’s effectiveness in substantially enhancing accuracy and responsiveness.
Xiulin Li, Li Pan 0001, Shijun Liu
IEEE Internet Things J.4
2025 An Online Algorithm for Inference Service Scheduling Using Combinations of Server-Based and Serverless Instances in Cloud Environments
abstract
With the continuous development in the field of machine learning, there is an increasing demand for the cloud-based machine learning inference services, which are latency-sensitive tasks, such as the service requests from the Internet of Things (IoT) devices. These inference services are generally accompanied by fluctuations and uncertainties, so they often require vastly varied numbers of servers at different time spots. As a result, how to dynamically and rationally schedule cloud servers for inference services has become an important issue. Alibaba Cloud currently provides its serverless instances called elastic container instance (ECI), and due to the advantages of pay-as-you-go billing and second-level elasticity, they are well-suited for handling bursty or fluctuating workloads. At the same time, Alibaba Cloud’s subscription-based elastic compute service (ECS) instances can be used for steady-state workloads. Our objective is to dynamically combine these two types of instances to deal with inference service requests. In this article, we propose a deterministic online algorithm that can rationally schedule these two types of instances to optimize costs without requiring knowledge of future workloads. We prove that the proposed online algorithm achieves a competitive ratio of no more than 2 compared to the optimal offline algorithm. Through simulation experiments, we demonstrate that our algorithm outperforms three benchmarks, which are all-reserved algorithm, all-on-demand algorithm, and traditional online algorithms that only use ECS instances. Our algorithm exhibits superiority across various workloads and can significantly reduces costs in most cases.
Li Pan 0001, Shijun Liu, Kaiyuan Qi
IEEE Internet Things J.3
2025 An RL-Based Cost-Effective Two-Layer Scaling Strategy for Multiregional Heterogeneous and Time-Varying Cloud Instances
abstract
Fueled by the advancements in 5G networks, Internet of Things (IoT) technologies are undergoing rapid development. Utilizing the computational power and elasticity of cloud computing, deploying IoT applications as Software-as-a-Service (SaaS) can reduce costs while enhancing scalability. However, in recent years, large-scale cloud downtime caused by natural disasters and network attacks, which seriously affects SaaS applications, has been widely observed. Deploying SaaS applications across geo-dispersed instances in a distributed way can effectively avoid service termination caused by cloud downtime. Generally, SaaS applications face dynamic traffic, which can lead to variations in service qualities and potentially fail to meet user demands. Scaling cloud instances according to the real traffic could save costs for service providers while ensuring service qualities. Considering the lack of effectiveness of reactive-threshold-based scaling strategies, many works suggest proactive scaling strategies based on reinforcement learning (RL). However, most RL-based scaling strategies are cursed by dimensionality when facing dynamic traffic. Moreover, heterogeneous and time-varying multiregional instances often lead to a complex state space which intensifies the curse of dimensionality. To address the above issues, in this article, we propose an RL-based two-layer scaling strategy that addresses the challenge of cloud resource scaling in complex state spaces through hierarchical decision-making. We formulate the scaling of multiregional heterogeneous instances as a bi-objective optimization problem and discuss the Markov property of time-varying variables, including time-varying instances and dynamic traffic. Experimental results based on two real-world workload datasets show that our two-layer scaling strategy achieves more fine-grained, cost-effective scaling while maintaining service quality, compared to the baselines.
Hao Yang 0065, Li Pan 0001, Shijun Liu
IEEE Internet Things J.3
2025 Caching or re-computing: Online cost optimization for running big data tasks in IaaS clouds
Xiankun Fu, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2025 A profit-effective function service pricing approach for serverless edge computing function offloading
Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2025 Balancing function performance and cluster load in serverless computing: A reinforcement learning solution
Menglin Zhou, Bingbing Zheng, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.4
2025 A Fair and Efficient Resource Allocation Algorithm for Cloud Rendering Jobs
abstract
Service level agreements (SLAs) formulated by cloud rendering service providers and users are varied, as users may have diverse performance requirements for their own jobs. This leads to a complex issue that cloud resources need to be allocated to rendering jobs in an appropriate and effective manner to satisfy users' diverse SLAs. To address this issue, in this paper, we propose a novel fair and efficient resource allocation algorithm, which aims to maximize the execution efficiency of rendering service applications while satisfying users' diverse SLAs. Firstly, to satisfy users' diverse SLAs, we propose a rigorous definition ofweighted acceleration ratio fairness, whose guiding principle is that the execution speed of a rendering job should be proportional to its weight determined by users' SLAs. Then, under the guidance of the proposed principle of acceleration ratio fairness, we formulate a new algorithm to fairly allocate resources to rendering jobs. Lastly, to improve execution efficiency and coordinate efficiency and fairness, we propose a fair and efficient resource allocation algorithm with relaxing fairness in resource competitive and non-competitive situations separately for rendering service applications. With extensive experiments that involve real rendering application workloads, we validate the effectiveness of our algorithms in improving execution efficiency and satisfying users' diverse SLAs.
Xiulin Li, Li Pan 0001, Shijun Liu, Xiangxu Meng
IEEE Trans. Serv. Comput.3
2025 RECaching: Cost-Effective Edge Caching for Cloud Storage With Differentiated Regional Workloads
abstract
For cloud storage, edge caching not only reduces the latency of delivering data to requesters due to closer delivery distances, but also potentially brings cost-savings to cloud users due to cheaper edge resources. However, cost-effective caching requires the exact knowledge of future requests, which is hard for cloud users to obtain in advance. It raises a risk of incurring more costs due to unpopular redundant replicas if data is blindly cached at the edge. To overcome this challenge, in this paper, we propose an online cost-effective caching algorithm based on reinforcement learning to dynamically make decisions of edge caching and caching lifespan online without any knowledge of the future. Further, we propose several mechanisms to enhance the cost performance of the proposed algorithm, to make up for possible mistakes made at the beginning of the learning and avoid cost-ineffective periods in caching lifespans. We then analyze the performance of caching decisions given by the proposed algorithm with enhancements. Finally, we conduct extensive simulations driven by real-world traces under prevalent pricing schemes of both the cloud and the edge, which reveals that significant cost-savings and superiority over benchmark algorithms can be achieved.
Li Pan 0001, Shijun Liu
IEEE Trans. Serv. Comput.3
2025 Online Traffic Allocation for Video Service Providers in Cloud-Edge Cooperative Systems
abstract
Currently, with the popularity of live video applications, VSPs (video service providers) begin to use cloud servers to enhance user experience and reduce operational costs. In this paper, we consider VSPs leveraging a cloud-edge cooperative model to deliver video services for cost reduction. Since bandwidth costs make up a significant portion of VSPs' operating expenses, we mainly consider bandwidth cost optimizations in traffic allocation. In addition, the QoE (quality of experience) is also very important, while the latency has a larger impact on QoE. Thus our traffic allocation approach aims to strike a fine balance between minimizing bandwidth cost and bounding the latency experienced by clients. Such a trade-off is difficult to optimize with some prevailing bandwidth billing methods such as the$95^{th}$percentile bandwidth billing. We quantify such a trade-off by constructing a linear bandwidth cost optimization problem. We first describe the offline version of the optimization problem, and then design an online greedy algorithm that considers minimizing the current bandwidth cost at each time slot. By applying the Lyapunov optimization framework, we design another online algorithm based on the original greedy one. We prove that the time average delay achieved by our online algorithm is smaller than the upper bound we set when certain conditions are satisfied. Through extensive simulation experiments, we show that the proposed online algorithm can significantly reduce both the bandwidth cost and the time average delay of clients.
Li Pan 0001, Shijun Liu
IEEE Trans. Serv. Comput.3
2024 Q-scheduler: Optimize Job Scheduling in Hadoop with Reinforcement Learning
abstract
Hadoop, as an open-source implementation of the MapReduce paradigm, is increasingly being used in both industry and academia for large-scale data processing. Yarn, one of the core components of the second-generation Hadoop, manages cluster resources and job scheduling. Minimizing the total completion time of a set of MapReduce jobs is a point worth exploring in terms of Yarn’s performance. Hadoop’s default schedulers, including first-in-first-out (FIFO), Fair, and Capacity, do not consider the characteristics and preferences of job resource demand, resulting in insufficient resource utilization. Therefore, in this paper, a new job scheduler named Q-scheduler is proposed. It uses reinforcement learning (RL) to accumulate scheduling experience autonomously based on the Fair scheduler. Specifically, the proposed scheduler consists of a Classifier and a Decider. The Classifier classifies jobs through similarity measurement, and the Decider, as an agent with a Q-Table, considers the execution order of different job classes and updates the state-action values of the Q-Table to learn optimal scheduling. The experimental results show that Q-scheduler can reduce the total completion time of the job set and improve resource utilization.
Li Pan 0001, Shijun Liu
CSCWD3
2024 An Improved Genetic Optimization Algorithm for Scheduling Serverless Application Jobs in Public Cloud Environments
abstract
Due to the advantages of ease of management, high elasticity, and low prices, serverless computing has gained more and more momentum in recent years. Serverless computing allows developing and deploying applications in the form of Function-as-a-Service (FaaS). Application service providers can benefit greatly from deploying and running their applications in public serverless FaaS cloud platforms. Currently, public FaaS cloud platforms generally offer multiple computation resource configurations for application service providers to select for deploying and executing their functions with a pay-per-use billing. Application service providers can select a function instance with a higher processing capability to achieve a shorter job completion time, but it also incurs a higher cost, while service providers often have budget constraints for running their jobs. Thus, when scheduling serverless application jobs in public FaaS cloud platforms, application service providers need to trade off between the job completion time and the cost, which is generally an NP-hard problem. To address these issues, in this paper we propose an improved genetic optimization algorithm for scheduling serverless application jobs in public FaaS clouds, which can achieve the minimization of the job completion time within the given budgets. Specially, during the evolution process of our genetic scheduling algorithm, we use a greedy strategy to search for optimal schedules. Through experiments running on both synthetic and real-world data, we validate the efficiency and effectiveness of our proposed genetic optimization based scheduling algorithm in minimizing the job completion time within budget constraints.
Lingjie Pei, Li Pan 0001, Shijun Liu
CSCWD3
2024 Research on Multi-Model Fusion for Multi-Indicator Collaborative Anomaly Prediction in IoT Devices
abstract
In the era of Industry 4.0, the widespread application of Internet of Things (IoT) technology enables us to monitor the operational status of production equipment through sensors and creates new requirements for equipment anomaly prediction and analysis. It can shift maintenance tasks from passive to proactive, reducing downtime and repair costs associated with equipment failures, and ensuring production safety and efficiency. However, the main challenge in industrial IoT applications lies in acquiring a sufficient amount of anomaly data. To address this issue, we propose a method that analyzes normal operating data of equipment to reveal trends in equipment status, combining multiple-model fusion prediction and multi-index coordinated decision-making. Firstly, a multi-model fusion strategy is adopted for prediction, integrating models such as XGBoost, LightGBM, and LSTM to enhance the accuracy of equipment attribute prediction. Secondly, through a multi-index coordinated approach, considering combinations of multiple indicators, an adaptive dynamic threshold rule is formulated based on the residuals between predicted and actual values to achieve early warning of equipment anomalies. Finally, we validate the effectiveness of this method using a mine ventilation fan as an example. The experiment demonstrated that this method can detect equipment anomalies one day earlier than traditional threshold alarm methods, achieving proactive equipment maintenance.
Donghao Wang, Tengjiang Wang, Shijun Liu, Li Pan 0001
CSCWD4
2024 An Online Mechanism for Market-Oriented Service Provisioning in Cloud Environments
abstract
Currently, there are a number of users pursuing to purchase professional services for executing their jobs. Meanwhile, a service provider generally purchases on-demand or reserved instances from public IaaS (Infrastructure-as-a-Service) cloud platforms to elastically run users’ submitted jobs and then charges users for job-execution services accordingly. In the process of service provisioning, to maximize social welfare, service providers should systematically and economically decide their instance purchase, job scheduling, and service pricing schemes. For making optimal decisions, several challenges need to be addressed, including the online arrival of users, the NP-hardness of the problem, and the possible misreporting of selfish users. Faced with the above challenges, we design an online auction mechanism which can help service providers provisioning better services to achieve optimal social welfare, without requiring future information. Through strict theoretical analysis, we prove that the designed mechanism can guarantee truthfulness and individual rationality, run in polynomial time, and satisfy budget balance, while achieving competitive social welfare. The extensive evaluations on the basis of both synthetic and realistic Google cluster datasets demonstrate the effectiveness and efficiency of the designed mechanism.
Bingbing Zheng, Li Pan 0001, Shijun Liu, Kexian Sun
ISPA3
2024 To store or not: Online cost optimization for running big data jobs on the cloud
Xiankun Fu, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2024 Performance analysis of parallel composite service-based applications in clouds
Xiulin Li, Li Pan 0001, Shijun Liu, Xiangxu Meng
Future Gener. Comput. Syst.4
2024 A Q-learning based auto-scaling approach for provisioning big data analysis services in cloud environments
Shihao Song, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2024 Faster or Cheaper: A Q-learning based cost-effective mixed cluster scaling method for achieving low tail latencies
Hao Yang 0065, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2024 A Meta-Model Architecture and Elimination Method for Uncertainty Modeling
abstract
Uncertainty exists widely in various fields, especially in industrial manufacturing. From traditional manufacturing to intelligent manufacturing, uncertainty always exists in the manufacturing process. With the integration of rapidly developing intelligent technology, the complexity of manufacturing scenarios is increasing, and the postdecision method cannot fully meet the needs of the high reliability of the process. It is necessary to research the pre‐elimination of uncertainty to ensure the reliability of process execution. Here, we analyze the sources and characteristics of uncertainty in manufacturing scenarios and propose a meta‐model architecture and uncertainty quantification (UQ) framework for uncertainty modeling. On the one hand, our approach involves the creation of a meta‐model structure that incorporates various strategies for uncertainty elimination (UE). On the other hand, we develop a comprehensive UQ framework that utilizes quantified metrics and outcomes to bolster the UE process. Finally, a deterministic model is constructed to guide and drive the process execution, which can achieve the purpose of controlling the uncertainty in advance and ensuring the reliability of the process. In addition, two typical manufacturing process scenarios are modeled, and quantitative experiments are conducted on a simulated production line and open‐source data sets, respectively, to illustrate the idea and feasibility of the proposed approach. The proposed UE approach, which innovatively combines the domain modeling from the software engineering field and the probability‐based UQ method, can be used as a general tool to guide the reliable execution of the process.
Haoran Shi 0002, Shijun Liu, Li Pan 0001
IET Softw.2
2024 A DRL-Based Real-Time Video Processing Framework in Cloud-Edge Systems
abstract
Nowadays, the Internet is rapidly evolving toward the future of the Internet of Things (IoT), where billions or even trillions of edge devices may be interconnected. The proliferation of network cameras and the advancement of IoT technologies have provided broader opportunities for data collection and utilization. In the past, the massive real-time videos generated by network cameras were mostly transmitted over the network to the cloud for analysis. However, due to network speed limitations, the latency incurred by uploading all videos to the cloud makes it difficult to meet the real-time requirements of video analysis. While edge computing significantly reduces latency, the computational capabilities of edge devices are limited, making it difficult to handle large amounts of real-time video data. In this article, we introduce a real-time video processing framework called DeepVA, which utilizes cloud-edge collaboration technology to reduce latency in real-time video processing and enhance the accuracy of analysis. The DeepVA framework incorporates the DRLVA video frame distribution algorithm based on deep reinforcement learning (DRL), which dynamically determines whether to distribute video frames for processing at the cloud or edge. To evaluate the performance of the proposed DRLVA algorithm, we first verify that it is superior to several other DRL-based distribution algorithms on the Gym environment. We also evaluate the performance of DeepVA on the MOT2015 data set, MOTSynth data set, and real campus surveillance videos. The experiments show that our DeepVA outperforms both cloud-only and edge-only solutions in terms of reducing latency and improving accuracy.
Xiankun Fu, Li Pan 0001, Shijun Liu
IEEE Internet Things J.3
2024 Service Provisioning Based on Edge-Cloud Collaboration: A Two-Timescale Online Scheduling Algorithm
abstract
With the development of 5G network and edge computing technology, service providers can deploy applications on the edge cloud close to the users to improve service quality and efficiency. However, due to the limited number of edge resources and the dynamic change of user requests over time, it remains challenging for service providers to attain optimal resource procurement and job processing decisions at an edge data center. The computing capacity of public cloud resources is generally unlimited, but it is difficult to achieve low-latency performance similar to the edge resources and job transmission also requires high network bandwidth costs. In this article, we consider service providers using the edge-cloud collaborative mode to take advantage of the low-latency characteristics of edge resources and improve their scalability. Service providers purchase resources at the cloud and edge, respectively, to deploy applications, so as to satisfy dynamic service requests with different service quality requirements. In order to optimize the cost of service providers while ensuring the quality of their services, we propose a two-timescale online scheduling algorithm based on the Lyapunov optimization, which makes optimal decisions without knowing future information about users’ jobs. By determining the number of edge resources at a large time scale and the number of cloud resources at a small time scale, our algorithm can avoid frequent application deployments at edge while achieving rapid service response in a cost-effective manner. Rigorous theoretical analysis and extensive experiments based on both the synthetic and real-world data verify the effectiveness of our proposed algorithm.
Yuxiao Qi, Li Pan 0001, Shijun Liu
IEEE Internet Things J.3
2024 An Online Algorithm Based on Replication for Using Spot Instances in IaaS Clouds
Li Pan 0001, Shijun Liu
J. Comput. Sci. Technol.3
2024 An online bi-objective scheduling algorithm for service provisioning in cloud computing
Yuxiao Qi, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2024 An online cost optimization approach for edge resource provisioning in cloud gaming
Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2024 Caching or not: An online cost optimization algorithm for geodistributed data analysis in cloud environments
Weitao Yang, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2024 Dynamic Recommendation Based on Graph Diffusion and Ebbinghaus Curve
abstract
Nowadays, many dynamic recommendations still suffer from the insufficiency of finding user online interest evolving patterns because of those complicated interactions. In general, each interaction is usually impacted by multiple underlying reasons, which needs us to open the “box” of each interaction instance instead of simply treating them as a pair-wise link. Besides, different users usually perform differently for their long-term and short-term tastes, leaving traditional sequential models far from personalized. In this article, we propose a novel recommendation model based on Graph Diffusion and Ebbinghaus Curve. Specifically, to explore the underline reasons for different interactions, we explore an underlying sub-graph for each interaction and find important reasoning paths within the sub-graph via a well-designed graph diffusion method. To capture users’ personalized strategies on long-term and short-term tastes, we are inspired by the Ebbinghaus Curve, which can naturally describe users’ memory patterns, and design an effective neural network to process users’ evolving behaviors. We conduct extensive experiments on four real-world datasets and the results further confirm the superiority of our model compared with existing state-of-the-art baselines.
Zhihong Cui, Xiangguo Sun, Hongxu Chen 0002, Li Pan 0001, Li-Zhen Cui 0001, Shijun Liu, Guandong Xu
IEEE Trans. Comput. Soc. Syst.6
2024 Enhancing Temporal Knowledge Graph Alignment in News Domain With Box Embedding
abstract
In many fields, such as social networks and recommendation systems with high time requirements, fake news and false information are often released in real time, impacting on people’s daily life. Entity alignment (EA) in temporal knowledge graph (TKG) can fuse the information contained in entities by finding equivalent entities, thus helping to determine the regular pattern of disinformation under time change. The existing methods either ignore the use of temporal attributes’ information and structural information or the modeling of that is insufficient, which has become a major obstacle to the further and wider application of TKG EA. In this article, we put forward a new idea of training for the processing of time attributes and relational structure information, to further enhance the ability in the EA process of TKGs. By forming box embedding matrix and name embedding matrix, and adaptively fusing the above information, we propose a new TKG EA solution. We carry out comparative experiments on standard news media and social media datasets collected from the real world, which validates the effectiveness of our proposal.
Shihao Hou, Weiyi Zhong, Xiaoran Zhao 0001, Yuwen Liu 0003, Yihong Yang, Shijun Liu, Li Pan 0001
IEEE Trans. Comput. Soc. Syst.7
2024 LIHAN: A Lattice-Guided Incomplete Heterogeneous Information Network Embedding Model for Node Classification
abstract
Real-world heterogeneous information networks (HINs) are modeled as heterogeneous graphs, in which features and structures are often incomplete. Existing models employ manual imputation or dynamic adjustment to populate the incomplete data. However, there are some limitations in incomplete heterogeneous graph representation learning: 1) using populated data may lose content and high-level interaction information of HINs, even lead to negative impacts on the performance of downstream tasks; and 2) existing models fail to utilize the high-order heterogeneous structures in original incomplete network data. To resolve the above issues, in this article, we proposed a lattice-based incomplete heterogeneous structural attention network (LIHAN) for learning incomplete heterogeneous node embeddings. LIHAN first constructs characteristic lattice and structure lattice by mining characteristic sets and structure sets according to the partial order relations in between. Then, an improved lattice-based heterogeneous dual-attention mechanism is used to learn the heterogeneous node representations. Extensive node classification experiments are conducted on five open datasets to verify the superior performance of the proposed LIHAN model over the state-of-the-art models. Experimental results illustrate that LIHAN outperforms other methods on the micro-F1 and macro-F1 in node classification tasks. Moreover, experiments on different levels of lattices and the parameter sensitivity analysis shows the great stability during the process of experiments.
Guangxu Mei, Li Pan 0001, Qian Li 0003, Feng Li 0030, Shijun Liu
IEEE Trans. Comput. Soc. Syst.6
2024 NUS: Noisy-Sample-Removed Undersampling Scheme for Imbalanced Classification and Application to Credit Card Fraud Detection
abstract
Since minority samples are substantially less common than majority samples, many industrial applications, such as credit card fraud detection (CCFD) and defective part identification, call for imbalanced classification. The performance of a classifier tends to suffer from the noisy samples in majority or minority classes. This work proposes a new undersampling scheme, called a clustering-based noisy-sample-removed undersampling scheme (NUS) for imbalanced classification. The majority class samples are first clustered. The distance of the majority class sample from the cluster center that is furthest away is used as the radius to build a hypersphere, with each cluster’s center assumed to be a spherical center. We determine the Euclidean distance between the center of a cluster and each minority sample to find whether they are in the hypersphere or not. Afterward, we exclude noisy samples from the minority class. The noisy samples of majority classes are removed by using the same procedure. Second, we propose an NUS, which combines noisy sample removal with undersampling techniques. Finally, to prove the effectiveness of NUS, we integrate NUS with the basic classifiers random forest (RF), decision tree (DT), and logistics regression (LR). We conduct their comparison with seven undersampling, oversampling, and noisy-sample-removed methods. This work performs experiments on 13 public and three real transaction datasets related to e-commerce. The results show that NUS plays a positive role in promoting existing classifiers’ performance.
Honghao Zhu, MengChu Zhou, Guanjun Liu, Yu Xie 0019, Shijun Liu
IEEE Trans. Comput. Soc. Syst.5
2024 Collaborative Storage for Tiered Cloud and Edge: A Perspective of Optimizing Cost and Latency
abstract
Edge storage is emerging as a novel storage paradigm, which offers the advantage of low latency and low cost, but has the disadvantage of limiting the scope of services to a certain area. In contrast, cloud storage offers anywhere services but has disadvantages in terms of latency and cost. In this paper, a collaborative storage scheme is proposed to leverage their complementary advantages. Considering the impact of collaboration on the cloud and the edge, we propose collaborative optimization models for cost and latency. To address the challenge of requiring future information when performing optimization for cost and latency, we first transform the long-term optimization into individual optimizations in each time slot using the Lyapunov optimization technique, based on which our algorithm decides whether a replica should be created at the edge and which cloud tier a replica should be migrated to. Then, we prove that the proposed algorithm can obtain near-optimal costs and guaranteed latencies. Finally, we conduct extensive simulations driven by real-world traces, and show that our algorithm can achieve a trade-off between cost and latency and outperform other benchmark algorithms.
Li Pan 0001, Shijun Liu
IEEE Trans. Mob. Comput.3
2024 DGERCL: A Dynamic Graph Embedding Approach for Root Cause Localization in Microservice Systems
abstract
Root cause localization in microservice systems refers to finding the root cause that causes system anomalies using system information. Many methods construct a graph structure and perform random walk on it to localize the root cause. This is not suitable for larger systems due to the high computational overhead. Besides, the constructed graph is usually static which mismatches with evolving metrics. Different metrics also contribute differently to determining root cause. To address these challenges, we have developed DGERCL, a novel method that employs dynamic graph embedding to localize root causes in microservice systems. We construct a dynamic graph where nodes, edges, and features correspond to microservices, invocations, and metrics. DGERCL first gets invocation information by aggregating node embedding and features via a trainable structure. An LSTM then processes invocation information to update node embedding. We also propose a neighbor information aggregation method to enrich structure information and a self-attention-inspired mechanism to leverage the importance of metrics for better mining metrics information. Finally, a classifier maps node embedding learned by LSTM to possibilities belonging to root cause. We conduct comprehensive experiments on two microservice benchmarks. Our model achieves good results which demonstrates the effectiveness of DGERCL.
Qian Li 0003, Shijun Liu, Li Pan 0001
IEEE Trans. Serv. Comput.4
2024 SpotDAG: An RL-Based Algorithm for DAG Workflow Scheduling in Heterogeneous Cloud Environments
abstract
As increasingly complex functions are implemented in applications, directed acyclic graphs (DAGs) are widely used to model the inter-dependencies between individual functions. Cloud-based data processing platforms need to consider the complex topology of DAGs and arbitrary deadlines given by users for job scheduling, leading to an NP-hard decision-making problem. Leveraging spot instances in data processing platforms can achieve significant cost savings, but the unpredictable interruption of spot instances makes the problem of VM scaling and job scheduling more difficult. In this paper, a Reinforcement Learning (RL) based approach called SpotDAG is proposed to solve the auto-scaling problem for jobs modeled as DAGs on a data processing platform where spot instances are introduced. SpotDAG makes cluster scaling and job scheduling decisions at the same time by mapping its output to several meta-policies. This paper introduces the self-attention mechanism for feature extraction to help the intelligent agent learn faster. A mask layer after the output of the proposed RL-based algorithm circumvents illegal actions to ensure that a job is completed by its deadline. Extensive experimental results show that the proposed approach can significantly reduce the cost of instances for data processing platforms while ensuring that jobs are completed in time.
Liduo Lin, Li Pan 0001, Shijun Liu
IEEE Trans. Serv. Comput.3
2024 Open knowledge base canonicalization with multi-task learning
Huang Peng, Weixin Zeng, Xiang Zhao 0002, Shijun Liu, Li Pan 0001
World Wide Web (WWW)5
2023 Dynamic Communications Network Linking Prediction by Disseminating Event Embedding
abstract
Communication networks represent communication between entities like social networks and microservice call graphs of microservice systems. Link prediction is useful in various communication network service systems such as predicting the relation between two services. Continuous-time dynamic graph (CTDG) is one form of representing temporal information in communication networks that treats them as a set of events occurring over time. Dynamic graph embedding for CTDG handles these events and disseminates event information to other nodes to get node embedding. Though dynamic graph embedding is suitable for making link prediction in communication networks, graph embedding for CTDG faces challenges such as how to model the event information dissemination process including information decaying over distance and influence of time information. To cope with this issue, we propose a CTDG-based dynamic graph embedding framework for dynamic communication networks link prediction called CTDGNN (Continuous-Time Dynamic Graph Neural Networks). In particular, we propose a self-adaptive information dissemination strategy based on node importance to update node embedding by disseminating event information. Finally, extensive numerical experiments on three real-world communication network datasets validate the effectiveness of our proposed model compared to other related methods.
Qian Li 0003, Zhihong Cui, Shijun Liu, Li Pan 0001
ICWS4
2023 Deep Temporal State Perception Toward Artificial Cyber-Physical Systems
abstract
Cyber–physical systems (CPS), as the cornerstone of smart city, has been attracting great interest from academia and industry. It aims to monitor/control physical components via communication and computation while ensuring effectiveness, intelligence, and security. The related research has pointed that the state perception on physical device is the prerequisite for boosting overall CPS performance. Toward this end, we present an effective deep temporal perception network to achieve classification-based state detection. Namely, we first design a multifeature encoding network for multiview time-series representation. Concretely, on the one hand, we utilize two piecewise aggregate representation strategies to obtain the key temporal trends; on the other hand, we adopt a temporal symbolic representation strategy to capture the necessary contextual semantic correlations. Thereafter, we develop a comprehensive representation enhancement module to improve feature comprehension capability and thus boosting the overall performance and interpretability. Corresponding comparison experiments, ablation studies, and data visualization analyses on benchmark data sets have verified the effectiveness of our model.
Shaokun Wang, Fan Liu 0008, Hongyun Fan, Yupeng Hu 0003, Shijun Liu
IEEE Internet Things J.7
2023 Event-based incremental recommendation via factors mixed Hawkes process
Zhihong Cui, Xiangguo Sun, Li Pan 0001, Shijun Liu, Guandong Xu
Inf. Sci.4
2023 Heterogeneous graphlets-guided network embedding via eulerian-trail-based representation
Guangxu Mei, Siyuan Ye, Shijun Liu, Li Pan 0001, Qian Li 0003
Inf. Sci.3
2023 An online service provisioning strategy for container-based cloud brokers
Xingjia Li, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2023 A rating prediction model with cross projection and evolving GCN for bitcoin trading network
Li Pan 0001, Shijun Liu
Pers. Ubiquitous Comput.3
2023 RLTiering: A Cost-Driven Auto-Tiering System for Two-Tier Cloud Storage Using Deep Reinforcement Learning
abstract
The cloud storage boom has prompted providers to offer two storage tiers, i.e., hot and cold tiers, which are respectively purpose-built to provide the lowest cost for frequent and infrequent access patterns. However, for cloud users, it is non-trivial to determine cost-effective tiers because it is hard to obtain future access patterns in advance and is difficult to predict them exactly. The lack of future information poses a risk of increasing costs instead of saving costs. This is not the only challenge encountered when it comes to cost optimization. In this article, we take Amazon S3 as an example to analyze the pricing of two-tier cloud storage and derive several major challenges faced by cost optimization. Then, assuming a priori knowledge of future access patterns, we propose an optimal offline algorithm based on dynamic programming to determine cost-effective tiers for each time slot. Further, to handle online workload arrivals, we formulate the problem using Markov decision processes and propose RLTiering based on deep reinforcement learning. Eventually, the cost performance of RLTiering is evaluated based on real-world traces and prevalent Amazon S3 pricing, and the results show that it achieves significant cost-savings.
Li Pan 0001, Shijun Liu
IEEE Trans. Parallel Distributed Syst.3
2023 A DRL-based online VM scheduler for cost optimization in cloud brokers
Xingjia Li, Li Pan 0001, Shijun Liu
World Wide Web (WWW)3
2023 How the four-nodes motifs work in heterogeneous node representation?
Siyuan Ye, Qian Li 0003, Guangxu Mei, Shijun Liu, Li Pan 0001
World Wide Web (WWW)4
2022 Job scheduling for big data analytical applications in clouds: A taxonomy study
Youyou Kang, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2022 Effeclouds: A cost-effective cloud-of-clouds framework for two-tier storage
Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2022 A Lyapunov optimization-based online scheduling algorithm for service provisioning in cloud computing
Yuxiao Qi, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2022 An online algorithm for optimally releasing multiple on-demand instances in IaaS clouds
Xurui Song, Li Pan 0001, Shijun Liu
Future Gener. Comput. Syst.3
2022 Heterogeneous graph embedding by aggregating meta-path and meta-structure through attention mechanism
Guangxu Mei, Li Pan 0001, Shijun Liu
Neurocomputing3
2022 A Cost-Effective Framework for Running Industrial Big Data Analysis Applications in Public Clouds
abstract
Nowadays, the improvement of the data acquisition capability of the Industrial Internet of Things (IIoT) systems has brought about higher data throughput. Users can deploy industrial big data analysis applications on cloud platforms in a pay-as-you-go way to deal with the unstable data generation and data analysis workloads in the IIoT scenario. To procure stable and flexible computing resources, as a traditional pay-as-you-go cloud instance procurement option, on-demand instances are widely used, but their expensive prices also significantly increase users’ cost burden. To reduce the data analysis cost, in this article, we establish a per-job cost-effective framework adopting spot instances, on-demand instances, and cloud storage for industrial big data analysis applications. In our framework, we take advantage of spot instances, which are computing instances provided at low prices under a pay-as-you-go model, to achieve cost savings. However, using spot instances carries the risk of being interrupted. Therefore, we propose to use a checkpointing mechanism to back up intermediate results to cloud storage to reduce the potential loss caused by spot instance interruptions. Considering the time sensitivity of industrial big data analysis applications, we use on-demand instances as alternative computing resources after spot instance interruptions to ensure that users’ jobs can be completed without high time latencies. Evaluation results show that our framework can achieve cost savings as well as minimize time latencies for users’ jobs.
Liduo Lin, Li Pan 0001, Shijun Liu
IEEE Internet Things J.3
2022 Fully convolutional networks with shapelet features for time series classification
Cun Ji, Yupeng Hu 0003, Shijun Liu, Li Pan 0001, Bo Li 0103, Xiangwei Zheng 0001
Inf. Sci.3
2022 Learning to make auto-scaling decisions with heterogeneous spot and on-demand instances via reinforcement learning
Liduo Lin, Li Pan 0001, Shijun Liu
Inf. Sci.3
2022 A survey of resource provisioning problem in cloud brokers
Xingjia Li, Li Pan 0001, Shijun Liu
J. Netw. Comput. Appl.3
2022 Randomized online edge service renting: Extending cloud-based CDN to edge environments
Zizhe Jin, Li Pan 0001, Shijun Liu
Knowl. Based Syst.3
2022 An online algorithm for scheduling big data analysis jobs in cloud environments
Youyou Kang, Li Pan 0001, Shijun Liu
Knowl. Based Syst.3
2022 A cost-driven online auto-scaling algorithm for web applications in cloud environments
Wen Si, Li Pan 0001, Shijun Liu
Knowl. Based Syst.3
2022 Ontology Guided Sparse Tensor Factorization for joint recommendation with hierarchical relationships
Hao Liu 0026, Xiutao Shi, Guangxi Li, Shijun Liu, Li Pan 0001
Pers. Ubiquitous Comput.4
2022 Reinforced KGs reasoning for explainable sequential recommendation
Zhihong Cui, Hongxu Chen 0002, Li-Zhen Cui 0001, Shijun Liu, Xueyan Liu 0001, Guandong Xu, Hongzhi Yin
World Wide Web4
2021 Amazon Spot Instance Price Prediction with GRU Network
abstract
The Amazon cloud platform sells its idle resources to cloud users as Spot Instances, which provide an ultra-low discount compared to the price of on-demand instances. Unlike the pricing strategy of the on-demand and reserved instances that use fixed price, the price of Spot Instances is dynamically changed, which introduce an interesting research topic of price prediction. In this paper, we firstly analyze the actual price distribution on a 90 days Amazon spot price history data downloaded from the Amazon Cloud platform, by using the parameter k-AMSE to represent the price fluctuation of a spot instance which can reflect recent data fluctuations better than MSE(mean square error). Then, We analyzed the factors that affect the price fluctuation and presented a prediction model based on the GRU(Gated Recurrent Unit) network. We compare the proposed algorithm with others and evaluate it with RMSE (root mean square error) measurement. The experiment results show that the GRU network approach can perform over others with an accuracy rate of 1.58e-3.
Dawei Kong, Shijun Liu, Li Pan 0001
CSCWD2
2021 Paper Recommendation Based on Author-paper Interest and Graph Structure
abstract
The recommendation system can recommend information to users efficaciously, which helps many users to obtain information in different fields. The paper recommendation is a research topic to provide authors with personalized papers of interest. However, most existing approaches equally treat title and abstract as the input to learn the representation of a paper, ignoring the author's interest and structure information of the academic network. In the paper recommendation system, authors and papers and the interaction of their information have a crucial impact on the efficiency and accuracy of the recommendations. However, most recommendation systems are usually designed based only on users. Therefore, we propose a method based on the author's periodic interest and academic graph network structure to obtain as much effective information as possible to recommend papers. Extensive offline experiments on large-scale real data show that our method outperforms the representative baselines.
Hao L, Shijun Liu, Li Pan 0001
CSCWD2
2021 Online Cost-effective Edge Service Renting for Content Providers in Cloud and Edge Environments
abstract
For solving the problem of high bandwidth costs and service delays faced by content service providers (CSPs), edge computing services can be used as a supplement to existing cloud data centers for building more efficient Content Delivery Networks (CDNs). When there are a large number of requests for a content in one certain area, the content service provider can choose to rent an edge computing service near this area to lower the bandwidth cost for content delivery. But if the requests then drop after that, additional costs will be incurred instead due to the edge service renting. Therefore, it is necessary to dynamically decide whether to rent an edge service according to the request arrival situations in the future, but the future is often difficult to predict. For dealing with this problem, we propose an online edge service renting approach, as well as a corresponding request redirection algorithm, which can help content service providers save bandwidth cost significantly, while without requiring any knowledge about the future. Through theoretical analysis, we prove that the cost achieved by our online algorithm won't exceed 2 - α times compared to the optimal offline algorithm, where α is the bandwidth discount between edge and cloud services. Finally, by conducting extensive simulations with both real-world and synthetic data, we verify that our online edge service renting approach can effectively save costs for CSPs.
Zizhe Jin, Li Pan 0001, Shijun Liu
ICWS3
2021 Market-oriented online bi-objective service scheduling for pleasingly parallel jobs with variable resources in cloud environments
Bingbing Zheng, Li Pan 0001, Shijun Liu
J. Syst. Softw.3
2021 Keep Hot or Go Cold: A Randomized Online Migration Algorithm for Cost Optimization in STaaS Clouds
abstract
Storage-as-a-Service clouds generally offer both hot and cold storage tiers with different pricing options. Hot tiers provide a higher storage price but a lower access price, and vice versa for cold tiers. Many studies show that those user-generated data generally receive relatively high access frequency in the early period of their lifetimes while the overall trend of accesses is downward. Thus, when such kinds of data are hosted in clouds, they can be stored in hot tiers initially and then migrated to cold tiers for optimizing costs. However, the cost ofmigrationis non-negligible, and the number of accesses may then unexpectedly increase after migration, which indicates that a rash migration will incur more costs instead of cost-savings. For making optimal migration decisions, future data access curves are needed, but it is generally very hard to predict them precisely. To solve this problem, in this paper we propose a randomized online algorithm to optimize costs for those user-generated data stored in clouds, without requiring any future information. We show theoretically that the proposed algorithm can achieve a guaranteed competitive ratio of$1 + \frac {2(1-\lambda)}{e - 3 + 2\lambda + \lambda /\alpha }$, and it can be easily extended with prediction windows when short-term predictions are reliable. Eventually, we validate the effectiveness of our proposed algorithms through simulations driven by real-world video-visiting traces collected from a well-known video-sharing website.
Li Pan 0001, Shijun Liu
IEEE Trans. Netw. Serv. Manag.3
2020 Bidding Strategy Based on Adaptive Differential Evolution Algorithm for Dynamic Pricing IaaS Instances
Dawei Kong, Guangze Liu, Li Pan 0001, Shijun Liu
CollaborateCom (1)4
2020 A Big Service with Network Represent Learning for Quantified Flight Delay Prediction
abstract
An air traffic network is a special and complex Spatio-temporal network. What makes it unique is that multi-data sources-including airports, airlines and air routes-spatial dependence and strong temporal dependence in a dynamic environment. In this paper, we use big service to predict the flight departure delay time in air traffic networks. In the local services layer, we use graph sequences to model the Spatiotemporal network from multi-data sources, what is, using graphs to model the spatial dependence, and using sequences to model the temporal dependence. In the domain-oriented services layer, we use graph neural network to embed the graph sequence. We validate the method on an air Spatiotemporal network. Then, we use the embedding to estimate the departure delay time of the flight based on real-time conditions. In the demand-oriented services layer, we design a weighted cross entropy loss function and use a special evaluation to predict the flight departure delay time by the embedding in the domain-oriented services layer. Evaluated through a series of experiments on a real-world data set, we show that the method produces an effective result on the Spatio-temporal network which is substantially better than state-of-the-art alternative task: flight delay estimation. And it performs well in predicting the departure delay time with a total accuracy of 0.87.
Guangxu Mei, Lei Bian, Hongwu Tang, Diansheng Wang, Li Pan 0001, Shijun Liu
ICWS7
2020 ADARC: An anomaly detection algorithm based on relative outlier distance and biseries correlation
abstract
Summary The application of anomaly detection to data monitoring is a fundamental requirement of the public service systems of a smart city. Many detection methods have been proposed for identifying anomalous situations, including methods based on periodicity or biseries correlations. However, the detection results of these methods are not ideal. Thus, we present a new anomaly detection algorithm for time series based on the relative outlier distance (ROD) and biseries correlations. The proposed algorithm detects outliers based on the ROD and identifies abnormal points and change points based on biseries correlations. Experimental results show that our method achieves better recall and F1‐measure scores than various time series–based techniques while maintaining a high level of precision.
Cun Ji, Xiunan Zou, Shijun Liu, Li Pan 0001
Softw. Pract. Exp.3
2019 Infer Latent Privacy for Attribute Network in Knowledge Graph
abstract
The information of the real world is stored as triplets (head entity, relation, tail entity) in knowledge graphs. They are extremely useful resources for many intelligent applications but suffer from incompleteness. This paper proposes a knowledge graph representation model to infer latent privacy based on the existing data in attribute network. In our model, considering the nodes are heterogeneous, we classify the nodes into attribute nodes and entity nodes. In order to protect the privacy of entities, we don't follow the previous methods to learn and store the feature embedding of each entity in knowledge graph. Our model focuses in capturing the restriction patterns of attribute nodes, which is safe when merging data from various sources. Given a triplet (entity node, relation, attribute node), firstly, we get the embedding of the entity node by using a sophisticated way to utilize all the information of the node, not only the node connections but also the external text information. Then, we infer the attribute node for the entity node in a certain relation. Finally, we calculate the probability that the triplet is exist. In experiments, we evaluate our model on the tasks of triplet classification and link prediction. Evaluation results show that our approach outperforms the state-of-the-art methods with an accuracy rate of 90.0% in the task of triplet classification on the person attribute knowledge graph FB13. Besides, our model reaches promising performance by MeanRank =5.10, Hits@l = 35.14% and Hits@5=64.94% in the task of conference prediction on the academic network DBLP.
Zeyuan Cui, Li Pan 0001, Shijun Liu, Li-Zhen Cui 0001
IEEE BigData3
2019 SGNN: A Graph Neural Network Based Federated Learning Approach by Hiding Structure
abstract
Networks are general tools for modeling numerous information with features and complex relations. Network Embedding aims to learn low-dimension representations for vertexes in the network with rich information including content information and structural information. In recent years, many models based on neural network have been proposed to map the network representations into embedding space whose dimension is much lower than that in original space. However, most of existing methods have the following limitations: 1) they are based on content of nodes in network, failing to measure the structure similarity of nodes; 2) they cannot do well in protecting the privacy of users including the original content information and the structural information. In this paper, we propose a similarity-based graph neural network model, SGNN, which captures the structure information of nodes precisely in node classification tasks. It also takes advantage of the thought of federated learning to hide the original information from different data sources to protect users' privacy. We use deep graph neural network with convolutional layers and dense layers to classify the nodes based on their structures and features. The node classification experiment results on public data sets including Aminer coauthor network, Brazil and Europe flight networks indicate that our proposed model outperforms state-of-the-art models with a higher accuracy.
Guangxu Mei, Shijun Liu, Li Pan 0001
IEEE BigData3
2019 Selecting Superior Candidates from a Suitable Set: A Selective Extraction Algorithm for Accelerating Shapelet Discovery in Time Series Data
abstract
A serious challenge that confronts shapelet-based algorithms for time series classification is finding optimal shapelets in a short time. Representative shapelet-discovery algorithms find shapelets by evaluating the qualities of candidates extracted from the subsequences. One of the main difficulties is the large amount of time consumed, due to the excessive number of shapelet candidates. To address the above problem, in this paper we propose a fast and interpretable candidate-extraction algorithm to accelerate the process of shapelet discovery. The proposed algorithm utilizes a time series subclass splitting technique to sample time series dataset first. Then, an IDP (Important Data Point)-based selective-extraction strategy is used to extract shapelet candidates. The generated candidates have significant improvements in quality and reductions in quantity. Furthermore, the shapelet candidates generated are more interpretable. To test the effectiveness of the shapelet candidates generated, we transform the original time series and use an off-the-shelf attribute-selection technique to select optimal shapelets from candidates. We then evaluate the proposed algorithm through extensive experiments. The results demonstrate that the proposed algorithm makes significant improvements in accuracy, compared with baselines. Meanwhile, the time consumption is also greatly reduced.
Shijun Liu, Li Pan 0001, Cun Ji, Chenglei Yang
CSCWD2
2019 An Online Mechanism for Purchasing IaaS Instances and Scheduling Pleasingly Parallel Jobs in Cloud Computing Environments
abstract
Nowadays, many users select to outsource their job executions to service clouds. These users often have heterogeneous demands while they dynamically arrive at the clouds. For reducing the costs and risks, lots of service cloud operators purchase on-demand instances from public IaaS clouds and provide professional services elastically to users. However, without knowing the future information, it is hard for cloud operators to optimally determine the instance purchasing as well as job scheduling and pricing schemes. In order to achieve maximum social welfare, this paper targets to design an auction mechanism which executes in an online fashion for service clouds, with unique features of job-oriented users, pleasingly parallel jobs and soft deadline constraints. Such a mechanism ought to run in polynomial time, provide truthfulness guarantee, satisfy individual rationality and budget balance, and achieve competitive social welfare. Nevertheless, when designing mechanisms there are a few significant challenges, including the difficulty for finding optimal solution, the strategic behaviours of selfish users with private information and the online arrivals of users. Facing these challenges, we leverage the idea of proportional sharing and propose an online mechanism which is proven to achieve all desired properties. The efficiency of the proposed mechanism is validated by both theoretical analysis and extensive simulations which use both synthetic data and Google's job traces.
Bingbing Zheng, Li Pan 0001, Shijun Liu, Lu Wang 0007
ICDCS3
2019 An Online Algorithm for Selling Your Reserved IaaS Instances in Amazon EC2 Marketplace
abstract
In cloud platforms such as Amazon EC2, users can reserve IaaS instances rather than buy on-demand ones to save cost. But it would incur the waste of reservations if there are few demands arriving after reserving instances. Currently, there is a reserved instance marketplace launched by Amazon EC2 cloud, where users can sell their unused instances for avoiding such waste of unused reservations. But for users, it is difficult to make the decision to sell their instances optimally without knowing any information for future demands, for it would incur the extra cost when there are new demands arriving after selling their reservations. For solving this problem, an online selling algorithm is proposed in this paper to guide cloud users in selling reserved instances in Amazon EC2 marketplace. We prove theoretically that our online algorithm Aβ can guarantee a bounded competitive ratio of 2T/β, whose value is specific to the type of reserved instances. Taking the i3.large instance provided by Amazon EC2 as an example, the competitive ratio is 3.36 under its pricing rules for 1-year term. Finally, via extensive experiments using workload data collected from actual applications, we verify our online algorithm's effectiveness and demonstrate that it is much more cost effective to cloud users in IaaS platforms.
Shengsong Yang, Li Pan 0001, Shijun Liu
ICWS3
2019 Mixed Reality Storytelling Environments Based on Tangible User Interface: Take Origami as an Example
abstract
This paper presents a mixed reality storytelling system, which takes handicrafts as tangible interaction tools. Via the system, users can learn handicraft and then use the handicraft pieces to design, create and tell stories with HoloLens iteratively. In order to overcome the limitations of HoloLens gesture interaction, the system uses hand tracking with Kinect to implement a touch-like effect on the desktop. User study shows that our system has good usability, and it is welcomed by users. In addition, it can stimulate users interest in handicraft and storytelling, and even promote parent-child interaction effectively.
Nianmei Zhou, Wei Gai, Juan Liu 0008, Yulong Bian, Shijun Liu, Li-Zhen Cui 0001, Chenglei Yang
VR7
2019 A fast shapelet selection algorithm for time series classification
Cun Ji, Shijun Liu, Chenglei Yang, Li Pan 0001, Lei Wu 0002, Xiangxu Meng
Comput. Networks3
2019 A just-in-time shapelet selection service for online time series classification
Cun Ji, Li Pan 0001, Shijun Liu, Chenglei Yang, Xiangxu Meng
Comput. Networks4
2018 A 2D Transform Based Distance Function for Time Series Classification
Cun Ji, Xiunan Zou, Yupeng Hu 0003, Shijun Liu
CollaborateCom4
2018 A Market-Oriented Heuristic Algorithm for Scheduling Parallel Applications in Big Data Service Platform
abstract
Big Data analytics service platform delivers a new type of public cloud offerings, through which end users can outsource their job executions by using a group of professional Big Data processing services in a pay-per-use way. Different from other type of cloud services, parallel jobs dominate the domain of data processing services, whose execution time can be varied greatly with different runtime configurations, such as different degrees of parallelism. In such a market-oriented environment, scheduling jobs from end users efficiently to optimize the Big Data analytics service platform's revenue is a more challenging task. In this paper, we propose a market-oriented heuristic algorithm for scheduling parallel jobs in a Big Data analytics service platform with admission control to optimize the platform operator's revenue. The proposed scheduling heuristic takes into account not only the dynamic revenue gained from accomplishing a job within a specific runtime as well as the consumption of resources needed for running it to achieve this given runtime, but also the potential loss it causes to the system by running this job instead of other waiting jobs currently in the system. We also propose a collaborative filtering based approach to quickly and accurately predict the execution time of parallel jobs running in a Big Data analytics service platform. We have conducted extensive experiments and simulations based on workload data derived from the real-world data analytics service platform and parallel applications. We show that our scheduler can outperform the other scheduling algorithms used for comparison, which are based on classical heuristics from literature, thereby fully evaluating the effectiveness of our market-oriented heuristic scheduling algorithm.
Qingshi Shao, Shijun Liu, Li Pan 0001, Chenglei Yang, Tingting Niu
COMPSAC (1)2
2018 An Approach to Web Service Organization Based on Hypergraph Clustering
abstract
With the rapidly growing number of web services, discovery and selection for numerous services under the dynamic and large-scale environment of web services is becoming a crucial task, how to organization these services becomes a challenging and hot issue. In this paper, we mainly investigate the method of improving service organization, and aims to present a new method, which can increase service discovering efficiency and improve the service system reliability. We introduce atomic service, abstract service and cluster service to describe web service from different levels, and propose a clustering-based service organization framework, which manages Web services hierarchically. We apply hypergraph to describe the relationship of web services and put forward the hypergraph-based service clustering method. First, we construct a weighted hypergraph with web services, and then we obtain the cluster of web services through hypergraph partitioning algorithm Based on the hierarchical service organization framework, the algorithms of dynamic service matchmaking, discovery, and replacement will be performed efficiently. We develop the service management system to manage our services hierarchically in our cloud service platform.
Lei Wu 0002, Shijun Liu, Minggang He
CSCWD3
2018 Network-Constrained Tensor Factorization for Personal Recommendation in an Enterprise Network
abstract
While standard product recommendation systems have proven to be useful for e-commerce, they mainly rely on some prior information about users and products, such as ratings and intrinsic properties of products as well as profile attributes of users. In e-commerce settings, however, a more complete understanding of the demands of customers and the enterprise network constructed by suppliers and manufacturers can be utilized to improve the quality of product recommendations. Moreover, user ratings may be very sparse in some domains. Standard approaches suffer from such data sparsity and neglect to account for important additional dependencies that can be taken into consideration. This motivates us to design a new recommendation model, which incorporates information of network into rating prediction. In this paper, we propose a network-constrained tensor factorization approach, which imposes network constraints as regularization terms on tensor non-negative factorization to improve the accuracy of prediction. To solve the network-constrained regularization problem in our model, we use the Alternating Direction Method of Multipliers (ADMM) method. Experiment results on real-world dataset demonstrate that our approach outperforms other state-of-the-art baselines.
Xiutao Shi, Zhouchonghao Wu, Li Pan 0001, Lei Wu 0002, Shijun Liu, Yuliang Shi
CSCWD5
2018 Social Media vs. News Media: Analyzing Real-World Events from Different Perspectives
Yafang Wang, Zeyuan Cui, Shijun Liu, Gerard de Melo
DEXA (2)5
2018 To Sell or Not To Sell: Trading Your Reserved Instances in Amazon EC2 Marketplace
abstract
Recently, Amazon EC2 offers a reserved instance marketplace, where cloud users can sell their idle reserved instances varying in contract lengths and pricing options for avoiding the waste of their unused reservations. However, without knowing the future demands, it is hard for users to determine how to sell instances optimally, for it would incur more cost if new demands arrive after selling their reserved instances. For dealing with this problem, in this paper we first propose three online selling algorithms to guide cloud users in making decisions whether or not to sell their reservations in Amazon EC2 marketplace while guaranteeing competitive ratios. We prove theoretically that the three proposed online algorithms can guarantee bounded competitive ratios, whose values are specific to the type of reserved instances under consideration. Specifically, for all standard instances (Linux, US East) for 1-year terms in Amazon EC2, compared with a benchmark optimal offline algorithm, our algorithm A3T/4 can achieve a ratio of 2-α-a/4 in managing instance purchasing cost, where α is the entitled discount due to reservation and a is the selling discount specified by the user who sells its reservations. Finally, through extensive experiments based on workload data collected from real-world applications, we validate the effectiveness of our online instance selling algorithms by showing that it can bring significant cost savings to cloud users compared with always keeping their reservations in Amazon EC2 reserved instance marketplace.
Shengsong Yang, Li Pan 0001, Qingyang Wang 0001, Shijun Liu
ICDCS4
2018 QoS Optimization of Service Clouds Serving Pleasingly Parallel Jobs
Xiulin Li, Li Pan 0001, Shijun Liu, Yuliang Shi, Xiangxu Meng
ICSOC3
2018 A Truthful Mechanism for Optimally Purchasing IaaS Instances and Scheduling Parallel Jobs in Service Clouds
Bingbing Zheng, Li Pan 0001, Dong Yuan 0001, Shijun Liu, Yuliang Shi, Lu Wang 0007
ICSOC4
2018 Performance Analysis of Service Clouds Serving Composite Service Application Jobs
abstract
Performance analysis is important for service clouds serving composite service application jobs containing parallelizable tasks, for optimizing the degree of parallelism (DOP) and resource allocation schemes could improve performance obviously. In this paper, we describe a novel tandem queuing network with a parallel multi-station multi-server system as an analytical model for service clouds serving composite service application jobs. We design a partition method (termed the 'pleasing partition') to help us propose an analytical model for parallelizable service which is the vital fraction of composite service. After that, we could obtain a complete probability distribution of response time, waiting time and other important performance metrics calculated by our proposed analytical model. Thus, to use this model, cloud operators could determine proper job configurations and resource allocation schemes, for achieving specific QoS (Quality of Service). Extensive simulations are conducted to validate that our analytical model has high accuracy in predicting performance metrics of composite service application jobs.
Xiulin Li, Shijun Liu, Li Pan 0001, Yuliang Shi, Xiangxu Meng
ICWS2
2018 Subscription or Pay-as-You-Go: Optimally Purchasing IaaS Instances in Public Clouds
abstract
In public clouds such as Amazon EC2, there are two main pricing models in purchasing Infrastructure-as-a-Service (IaaS) instances: the pay-as-you-go model and the subscription model. For these two options, users can dynamically combine them to provide services for demands to save their instance acquisition costs. Making optimal decisions toward the purchase of IaaS instances generally requires prior knowledge of future demands; however, it is difficult for users to predict all future workloads accurately. To deal with this problem, online reservation algorithms have been proposed to guide users in reserving instances. However, existing online algorithms do not conform to the pricing rules currently used in public cloud platforms. Therefore, we put forward a new online reserving algorithm for instance in accordance with the pricing policies used in most public IaaS offerings. Specifically, in this study, we use Amazon EC2 as an example to illustrate our algorithm. Through theoretical analysis, we prove that the cost of the proposed algorithm Aβin this paper is not greater than 2-1/β times of the optimal offline algorithm, where β>1 is a critical point in the online reservation algorithm proposed in this paper. Via extensive experimental simulations using both synthetic and actual workload datasets, we demonstrated that the online algorithm Aβis much more cost effective for cloud users than always paying-as-you-go in public IaaS markets.
Shengsong Yang, Li Pan 0001, Qingyang Wang 0001, Shijun Liu
ICWS4
2018 A Truthful Mechanism for Scheduling and Pricing Pleasingly Parallel Jobs in a Service Cloud
abstract
As more and more users outsource their job executions to service clouds, effective job scheduling and pricing models are needed to solve resource and service competitions between users. Considering the particularity of scheduling and pricing problems in a service cloud whose goal is generally social welfare maximization, current commonly used models, such as fixed-pricing schemes, have obvious shortcomings and thus are unfeasible. Therefore, in this paper, we propose a randomized mechanism to schedule and charge job executions in service clouds. Our proposed mechanism can schedule jobs in a flexible way to achieve approximate social welfare maximization while guaranteeing non-preemption. Flexibility means the number of instances which are allocated to a job can be changed over time. The mechanism is truthful in expectation, computationally efficient and individually rational. The theoretical analysis shows that our mechanism can achieve an expected social welfare approximation ratio α, which can be 2 in some situations. Extensive simulations show that our proposed mechanism can efficiently solve the job scheduling problem in service clouds.
Bingbing Zheng, Li Pan 0001, Dong Yuan 0001, Shijun Liu
ICWS4
2017 An Experimental Study of the Impact of vCPU Provisioning on the Performance of a 2-Tier Application Running in Cloud
abstract
Leveraging Virtual Machine (VM) technologies to host multiple Web applications on the same physical machine can improve the resource utilization and thus save a cloud provider's provisioning cost. By allocating and scheduling virtual CPU (vCPU) resources for running VMs, a hosted Web application may achieve varying performances. Thus, when facing an end user with a specific SLA (Service Level Agreement) requirement, a cloud provider needs to decide how many vCPUs to provision for the target SaaS application to meet the user's end-to-end performance requirement while saving cost. However, it is a non-trivial task to economically determine an optimal resource configuration to meet an end user's SLA requirement. Accurate performance analytic models based on traditional modeling techniques such as queuing systems are difficult to construct for web applications. In this paper, we describe our experience in studying the impact of vCPU provisioning on the performance of 2-tier Web applications, through benchmarking a 2-tier web application in the context of provisioning vCPUs to SaaS applications in a cloud environment. From a cloud provider's perspective, we focus on a generic approach for SaaS benchmarking, which can help to study the impact of vCPU allocations on a Web application's performance and make the optimal vCPU allocation decisions to meet end users' QoS requirements while saving provisioning costs. Besides, based on our benchmark experimental results, we also propose an adaptive controller with a vCPU allocation optimization algorithm which can automatically adjust the vCPU allocations to meet the end users' workload requirements.
Li Pan 0001, Qingyang Wang 0001, Shijun Liu, Dahui Chen
CLOUD4
2017 Predicting hospital readmission from longitudinal healthcare data using graph pattern mining based temporal phenotypes
abstract
The rapidly increasing availability of healthcare data from multiple heterogeneous sources has spearheaded the adoption of data-driven approaches for improved clinical research, decision making, and patient management. The patient healthcare data are usually longitudinal and can be expressed as medical event sequences, where the events include clinical diagnosis, medications, laboratory reports, etc. Because healthcare data has both longitudinal and heterogeneous attributes, analyzing healthcare data is an inherently difficult challenge. In this paper, we propose a hospital readmission prediction method using temporal phenotypes, namely the Tephe. Specifically, each patient's medical event sequence is first represented by a temporal graph, which captures temporal relationships of the medical events in each event sequence and makes the raw data more intuitive. Based on graph pattern mining, we define more significant frequent subgraphs as temporal phenotypes. This enables us to better understand the disease evolving patterns and treatment approach. In addition, we designed an improved greedy algorithm to find the optimal expression coefficient of frequent subgraphs for each patient. Finally, based on the optimal expression coefficient of the frequent subgraph, random forests are used to perform prediction tasks. The experimental results show that our proposed method is more accurate in the prediction tasks compared with the baselines.
Xiangzhen Xu, Li-Zhen Cui 0001, Shijun Liu, Hui Li 0048, Lei Liu 0003, Yongqing Zheng
BIBM3
2017 A Reinforcement Learning Based Workflow Application Scheduling Approach in Dynamic Cloud Environment
Daniel Kudenko, Shijun Liu, Li Pan 0001, Lei Wu 0002, Xiangxu Meng
CollaborateCom3
2017 Integrating supply and demand chains in personalized recommendation via chain-coupled tensor factorization
abstract
Standard recommender systems usually rely only on past user ratings as well as optional profiles of customers and products. In e-commerce settings, however, a more complete understanding of the corresponding bi-directional impact between the demands of customers and the supply capabilities of providers can be the key to success. This motivates us to design a recommendation model that explicitly reflects the supply and demand chains. We propose a Multi-relational Coupled Tensor and Matrix Factorization model, which jointly models user ratings as well as supply chain relationships for product recommendation. In addition, our model can predict the links between suppliers and manufacturers. We design an algorithm based on the Alternating Direction Method of Multipliers (ADMM) technique. Experiments on real-world datasets find that the proposed model outperforms traditional methods.
Qianyu Jiang, Xiutao Shi, Guangxi Li, Bin Liu 0022, Lei Wu 0002, Shijun Liu
CSCWD6
2017 Distributed ACO based on a crowdsourcing model for multiobjective problem
abstract
MOP (Multiobjective Optimization Problem) is a prevailing research field for its well-modeling on the decision-making dilemma in the real world. We present a distributed ACO (Ant Colony Optimization) algorithm based on a crowdsourcing model, with a few innovative strategies as enhancement, for continuous MOPs. The original MOP is expected to be decomposed into multiple single-objective subtasks, which are then distributed to the individuals on the network, and the non-dominated front for the MOP is constructed by aggregating the solutions from the crowd. Finally, an experiment on the KUR test problem illustrates that our approach is practical.
Li Pan 0001, Shijun Liu
CSCWD3
2017 A collaborative filtering based approach to performance prediction for parallel applications
abstract
Parallel application jobs account for a large population in current domain of cloud computing and Big Data processing services, whose execution time can be varied greatly with different runtime configurations. For efficiently scheduling resources and services to run parallel jobs, the ability to quickly and accurately estimate the performance of parallel applications is critical. Analytic predictive models based on traditional modeling techniques such as queuing systems are difficult to construct for parallel applications, due to the high complexity lying in the structures of parallel application models. Furthermore, due to the heterogeneity of resources computing capacities with a scalable computing environment such as a cloud computing platform, performance analytic and prediction becomes increasingly difficult for parallel applications. To address this problem, in this paper we propose a collaborative filtering based approach to quickly and accurately predict the execution time of parallel applications running in heterogenous resources. Particularly, we use the widely used Apache Spark platform as the running framework for parallel applications, and propose a bounds-based performance model to improve the prediction accuracy. Through extensive simulations and experiments on real Spark clusters and two large-scale machine learning applications as well as the simple but classic WordCount sample application, we show that the proposed Collaborative Filtering based approach and bounds-based performance model can accurately estimate the performance of parallel applications.
Qingshi Shao, Li Pan 0001, Shijun Liu
CSCWD3
2017 Dynamic-priority based profit-driven scheduling in mobile cloud computing
abstract
Mobile cloud computing is now emerging as a promising way to enlarge the capabilities of mobile devices by computation offloading. One of the critical challenges faced by the mobile cloud providers today is how to increase the profitability of their cloud services. In this paper, we deal with the problem of scheduling parallelizable computation jobs offloaded by mobile users in public cloud to maximize cloud providers' profit. We propose an efficient dynamic-priority based and profit-driven scheduling mechanism. We first compute the inherent profitability of a job based on the marginal effect theory and use it as the initial priority. Then, we update the priorities of waiting jobs based on the proposed time-dependent dynamic-priority calculation model as the elapse of time. The results of numerical experiments and simulations show that our approach are efficient in scheduling these kind of jobs in cloud data centers external to mobile devices and considerable profit improvements can be achieved by our proposed dynamic-priority based scheduling mechanism.
Li Pan 0001, Lei Wu 0002, Shijun Liu, Xiangxu Meng
CSCWD4
2017 Performance Analysis of Cloud Computing Centers Serving Parallelizable Rendering Jobs Using M/M/c/r Queuing Systems
abstract
Performance analysis is crucial to the successful development of cloud computing paradigm. And it is especially important for a cloud computing center serving parallelizable application jobs, for determining a proper degree of parallelism could reduce the mean service response time and thus improve the performance of cloud computing obviously. In this paper, taking the cloud based rendering service platform as an example application, we propose an approximate analytical model for cloud computing centers serving parallelizable jobs using M/M/c/r queuing systems, by modeling the rendering service platform as a multi-station multi-server system. We solve the proposed analytical model to obtain a complete probability distribution of response time, blocking probability and other important performance metrics for given cloud system settings. Thus this model can guide cloud operators to determine a proper setting, such as the number of servers, the buffer size and the degree of parallelism, for achieving specific performance levels. Through extensive simulations based on both synthetic data and real-world workload traces, we show that our proposed analytical model can provide approximate performance prediction results for cloud computing centers serving parallelizable jobs, even those job arrivals follow different distributions.
Xiulin Li, Li Pan 0001, Jiwei Huang, Shijun Liu, Yuliang Shi, Calton Pu
ICDCS4
2017 Automated Performance Evaluation for Multi-tier Cloud Service Systems Subject to Mixed Workloads
abstract
In multi-tier cloud service systems, performance evaluation relies on numerous experiments in order to collect key metrics such as resources usage. The approach may result in highly time-consuming in practice. In this paper, we propose an automated framework for performance tracking, data management and analysis to minimize human intervention in multi-tier cloud service systems. The framework support fine-grained analysis of the mixed workloads through the Discrete-time Markov-modulated Poisson process (DMMPP). A general multi-tier application is theoretically formulated as a queueing network to evaluate the performance. The effectiveness of the model has been validated through extensive experiments conducted in the RUBiS benchmark system.
Xudong Zhao 0004, Jiwei Huang, Lei Liu 0003, Shijun Liu, Calton Pu, Li-Zhen Cui 0001
ICDCS4
2017 Nash Equilibrium and Decentralized Pricing for QoS Aware Service Composition in Cloud Computing Environments
abstract
QoS aware service composition necessitates an effective pricing mechanism in regulating service providers in public cloud computing environments. However, due to the fact that service providers are usually autonomous, strategic and self-motivated, it is far from trivial to deal with the pricing issues between them. In this paper we formulate a non-cooperative service pricing game to understand the performance of a QoS aware service composition model, for which multiple providers strategically bid how to provide and price their elementary services and establish the Nash equilibrium as the final service composition scheme. We also develop a proportional revenue division rule to incentivize elementary service providers to contribute in improving the QoS of the final composite service delivered to end users. Concerning privacy conservation, we develop a decentralized and recursive bidding algorithm, allowing service providers to reach an equilibrium without disclosing their private information. Through theoretical analysis, we show that a Nash equilibrium exists in a QoS aware service composition game. Through extensive simulations, we show that the proposed recursive bidding process can converge quickly to a Nash equilibrium service composition scheme, and its efficiency is generally high.
Li Pan 0001, Bo An 0001, Shijun Liu, Li-Zhen Cui 0001
ICWS3
2017 Selling Reserved Instances through Pay-as-You-Go Model in Cloud Computing
abstract
Current Infrastructure-as-a-Service (IaaS) clouds offer both on-demand and reservation instance purchasing options. Users can combine these two options dynamically to serve time-varying demands while minimizing their instance acquisition costs. However, when future demands are unknown, it is far from trivial for cloud users to make optimal instance purchasing decisions. To deal with this problem, a carefully designed online algorithm can be employed to guide users in acquiring instances without any prior knowledge of future demands while guaranteeing a competitive ratio. In this paper, we propose an instance reselling model, in which a cloud user can temporarily rent out its idle reserved instances to other users through pay-as-you-go model. We also design online instance acquisition strategies which achieve a better competitive ratio than previous methods. Through extensive simulations based on both synthetic data and real-world traces, we show that our online algorithm under the proposed reselling model can outperform previous models and achieve significant cost savings.
Dong Yuan 0001, Li Pan 0001, Shijun Liu, Xiangxu Meng
ICWS4
2017 Real-Time Soft Resource Allocation in Multi-Tier Web Service Systems
abstract
Soft resource allocation is an important factor of system configuration which plays a critical role in guaranteeing the performance of multi-tier web service systems. There is a tradeoff between real-time performance and resource consumption, and thus the real-time adjustment of soft resource allocation in response to dynamic workload is quite challenging. In this paper, we propose a real-time soft resource allocation method that integrates both model-based analysis and real-time optimization. Specifically, a multi-tier web service system is firstly formulated by a queueing network model, and theoretical analyses are provided. Then, an optimization approach for real-time soft resource allocation is designed by applying sliding window techniques, in order to cope with dynamic workloads and performance demands. Based on the RUBiS benchmark system, model parameters are obtained by measurements and the efficacy of our approach is finally validated.
Xudong Zhao 0004, Jiwei Huang, Lei Liu 0003, Yuliang Shi, Shijun Liu, Calton Pu, Li-Zhen Cui 0001
ICWS5
2017 Link prediction by exploiting network formation games in exchangeable graphs
abstract
In social network analysis, we often need to predict new links, given some available evidence. This may, for instance, enable us to study user behavior and infer likely new interactions in the near future. Recently, a family of algorithms based on exchangeable graphs has proven effective for link prediction. The network is modeled as an exchangeable array, whose entries can flexibly be traced back to random function priors (e.g., block models, Gaussian Processes). Unfortunately, the burdensome computational complexity of these methods inhibit their application to even just moderate-scale networks. In this paper, we present a novel online training algorithm based on local Gaussian processes on subgraphs, which successfully overcomes this challenge. Moreover, we address the sparsity problem of links in social networks by presenting an improved algorithm based on network formation games. The network formation games we design also shed light on the ambiguity of missing links - not observed vs. non-existing. We evaluate our method against state-of-the-art algorithms on real-world datasets, demonstrating both the effectiveness and the efficiency of our method.
Yafang Wang, Bin Liu 0022, Lirong He, Shijun Liu, Gerard de Melo, Zenglin Xu
IJCNN5
2016 A Self-Evolving Method of Data Model for Cloud-Based Machine Data Ingestion
abstract
In the case of a cloud-based remote control system such as SCADA (Supervisory Control and Data Acquisition) that enables users to collect data from cloud-connected machines deployed anywhere at any time. However, machine data models may not be updated in a timely manner after the devices are upgrades or modified. This leads to mismatches between the machine data and data models. A key obstacle of the matching is that the machines can be modified. To address this, we present a self-evolving method for machine data model. We give the description of the evolution of machine data models and the self-evolving method for the models in details. The method detects the conflicts between the machine data and models, and transfer or derive models if necessary. Our method can thus facilitate the evolution of machine data models and ensure that every machine in the cloud corresponds to the correct machine data model automatically. At last, we present two case studies to validate our method.
Cun Ji, Shijun Liu, Chenglei Yang, Li-Zhen Cui 0001, Li Pan 0001, Lei Wu 0002
CLOUD2
2016 SpatialGraphx: A Distributed Graph Computing Framework for Spatial and Temporal Data at Scale
abstract
Development of the Smart City has produced much data with attributions of timestamp and location, but in some applications like investigation of the large bomb explosion in New York, the government takes precedence to investigate the relation data from New York city rather than the whole Country, which prompts us to do some research works in computing partial graph more fast. So we propose SpatialGraphx, a graph parallel computing framework supporting direct and fast partial graph construction and partial graph computation. Leveraging the spatial and temporal attributions of data, SpatialGraphx presents two extensions on the partial graph construction by building a spatio-temporal tree index and on the computation by a new location-based partition strategy. Using mobile network's data with hundred million edges, we demonstrate SpatialGraphx can support direct and fast partial graph construction and enables efficient partial graph analysis for spatial and temporal data. And compared to original Graphx, the improvement of SpatialGraphx is 3x to several orders of magnitude for large enough dataset.
Zongfei Lu, Yang Liu 0165, Shanqing Guo, Xin-Shun Xu, Shijun Liu
COMPSAC6
2016 A piecewise linear representation method based on importance data points for time series data
abstract
With the development of intelligent manufacturing technology, it can be foreseen that time series data generated by smart devices will raise to an unprecedented level. For time series with high amount, high dimension and renewal speed characteristics, resulting in difficult data mining and presentation on the original time series data. This paper presented a piecewise linear representation based on importance data points for time series data, which called PLR_IDP for short. The method finds importance data points by calculating the fitting error of single point and piecewise, and then represents time series approximately by linear composed of the importance data points. Results from theoretical analysis and experiments show that PLR_IDP reduces the dimensionality, holds the main characteristic with small fitting error of segments and single points.
Cun Ji, Shijun Liu, Chenglei Yang, Lei Wu 0002, Li Pan 0001, Xiangxu Meng
CSCWD2
2016 A cost-optimal service selection approach for collaborative workflow execution in clouds
abstract
Today, there has been a strong demand of distributed collaboration in design and manufacturing, due to the acceleration of economic globalization and the popularity of virtual enterprises (VE) model. Because of the characteristics of cloud computing, such as elasticity and on-demand computing, it is promising to deploy and execute collaborative workflows that contain multiple tasks and services such as Computer-Aided Design (CAD) software components on cloud resources for supporting collaboration across enterprises. Specifically, how to cost-effectively select appropriate services to execute workflows within deadlines while without violating multiple constraints becomes an important issue. In this paper, through investigating the practical requirements of collaborative design workflow, we first formulate the issue of the cost-optimal cloud service selection for collaborative workflow executions as a multi-dimensional optimization problem with multiple constraints. Then we propose an effective approach based on genetic algorithms to address this problem for obtaining near-optimal solutions. Based on workload data derived from real-world systems, we conduct experiments which show that our approach outperforms traditional greedy algorithms in finding better solutions and it also provides real-time performance guarantees in real-world cloud computing environments.
Li Pan 0001, Dong Yuan 0001, Shijun Liu, Lei Wu 0002, Xiangxu Meng
CSCWD4
2016 An Optimal and Iterative Pricing Model for Multiclass IaaS Cloud Services
Li Pan 0001, Shijun Liu, Lei Wu 0002, Li-Zhen Cui 0001, Dong Yuan 0001
ICSOC3
2016 Integrating Theoretical Modeling and Experimental Measurement for Soft Resource Allocation in Multi-tier Web Systems
abstract
Soft resources, which are system software components that use hardware or synchronize the use of hardware, are playing a critical role in the performance of multi-tier web systems, and thus it is quite important to tune the soft resource allocation for using the limited hardware resources to obtain maximum effectiveness. In this paper, we integrate both theoretical and experimental studies to the soft resource allocation problem. Specifically, we apply the queueing network model for formulating multi-tier web systems, and conduct experimental measurements based on the RUBiS benchmark system to obtain precise model parameters. Quantitative analysis is carried out, based on which an optimization model as well as an algorithm are put forward for soft resource allocation. The efficacy of our approach is validated by both theoretical analyses and experimental results.
Yuliang Shi, Jiwei Huang, Xudong Zhao 0004, Lei Liu 0003, Shijun Liu, Li-Zhen Cui 0001
ICWS5
2016 Profit Based Two-Step Job Scheduling in Clouds
Li Pan 0001, Shijun Liu, Lei Wu 0002, Xiangxu Meng
WAIM (2)3
2016 Real Time Prediction on Revisitation Behaviors of Short-Term Type Commodities
Xiangzhen Xu, Jinghua Fu, Yuliang Shi, Shijun Liu, Li-Zhen Cui 0001
WISE (1)4
2016 A Sub Chunk-Confusion Based Privacy Protection Mechanism for Association Rules in Cloud Services
abstract
In cloud computing services, according to the customized privacy protection policy by the tenant and the sub chunk-confusion based on privacy protection technology, we can partition the tenant’s data into many chunks and confuse the relationships among chunks, which makes the attacker cannot infer tenant’s information by simply combining attributes. But it still has security issues. For example, with the amount of data growing, there may be a few hidden association rules among some attributes of the data chunks. Through these rules, it is possible to get some of the privacy information of the tenant. To address this issue, the paper proposes a privacy protection mechanism based on chunk-confusion privacy protection technology for association rules. The mechanism can detect unidimensional and multidimensional attributes association rules, hide them by adding fake data, re-chunking and re-grouping, and then ensure the privacy of tenant’s data. In addition, this mechanism also provides evaluation formulas. They filter detected association rules, remove the invalid and improve system performance. They also evaluate the effect of privacy protection. The experimental evaluation proves that the mechanism proposed in this paper can better protect the data privacy of tenant and has feasibility and practicality in real world applications.
Yuliang Shi, Zhongmin Zhou, Li-Zhen Cui 0001, Shijun Liu
Int. J. Softw. Eng. Knowl. Eng.4
2015 Construction of Semantic Collocation Bank Based on Semantic Dependency Parsing
Shijun Liu, Yanqiu Shao, Lijuan Zheng
PACLIC1
2014 Finding Optimized Deployment Strategy for Multitenant Services by Iterative Staging
abstract
A serious challenge that confronts multi-tenant service systems is finding an optimized deployment strategy according to their business scale and operating characteristics. The tenants want to rent high performance services and services providers demand minimizing cost at the same time of meeting the requirements of tenants. But there are often contradictions between high performance and low cost. Therefore, in order to balance the contradiction, this paper proposes a staging-based optimized deployment method. This method performs iterative optimization based on customized workload generation, continuously emulation and evaluation in a benchmark suite. We demonstrate our method by a case study on a multi-tenant Supplier Business Management (SBM) service, as well as evaluate the capability of our benchmark suite through two sets of experiments. Results from these experiments characterized the relationship between workloads and performance, which can help find optimized deployment strategies for multi-tenant applications. In the case study on multi-tenant SBM service system, we gain an optimized strategy that satisfies the requirement of tenants and makes the maximum use of the resources, which can give useful recommendations in real service instances deployment stage.
Jizun Liu, Ze-yu Di, Shijun Liu, Calton Pu, Lei Wu 0002, Li Pan 0001
APSCC3
2014 Data Organization Patterns for Cloud Enterprise Applications
abstract
With the popularity of cloud computing and SaaS, enterprises are willing to rent various cloud services and use them in their daily work. However, on-premise applications are still widely used inside enterprises. That is to say, to achieve the goal of business collaboration, a cloud service usually need to obtain business data from multiple sources, which include the relevant on-premise applications and other cloud services. In this paper, we study and summarize data organization and management patterns for cloud enterprise applications from different aspects. Based on these patterns, we propose an innovative cloud-based data organization model (CEDM). Compared with traditional ones, it highlights the characters of resource sharing and reuse. Besides, this model is more convenient for tackling complex data synchronization issues in cloud environment.
Lei Wu 0002, Shijun Liu, Li Pan 0001, Xiangxu Meng
APSCC3
2012 A Novel QoS-Aware Service Composition Approach Based on Path Decomposition
abstract
QoS-aware Service Composition is to build new services by orchestrating a set of atomic services, and ensure the new services to satisfy certain QoS constraints. However, the current methods can't be able to address the problem efficiently in the situation that the new service is comprised of multiple tasks, the structure of its execution path is complicated, and the number of corresponding candidate service is huge. Therefore, in this paper, a novel QoS-aware service composition approach based on path decomposition (SCP) is proposed, which adopts the Case-Based Reasoning and Genetic Algorithm. In order to enhance the cases' reusability and matching flexibility, the entire execution plan is decomposed into fine-grained fragments before storing to the Case Library. When resolve the emerging service composition problem, through reusing existing cases, the execution path is adjusted to downgrade the problem size and reduce the complexity. For the adjusted execution path, the Genetic Algorithm is used to form an execution plan meeting user's requirement. A large number of experiments verify the validity of our approach.
Lei Wu 0002, Shijun Liu
APSCC3
2012 A cooperative construction approach for SaaS applications
abstract
Recently, much of the attention of SaaS has focused on the new business model that on-demand software enables. This trend requires the agile construction of applications which the long development cycle such as waterfall software-development model could not satisfy. This paper proposes a cooperative construction approach based on multi-level abstraction for SaaS applications, in which the tenant instance of a SaaS application will be abstracted into different levels according to its customization requirements, and then be constructed from different components and through different constructing routes with the collaboration of different participants in a SaaS system. We discuss the multi-level components model, the roles and the cooperative work, the multi-level construction method in details. We also introduce a case study of a SaaS application for supply business in automobile industry with the discussion of agile construction approaches.
Qian Li 0003, Shijun Liu
CSCWD2
2011 A Hybrid Approach to Placement of Tenants for Service-Based Multi-tenant SaaS Application
abstract
In a service-based multi-tenant SaaS application, the number of servers on which Web service instances are deployed are limited, and tenants share the same application and services. With the purpose of lowering cost of ownership by high economies of scale, we must solve the problem that how to optimally place tenants with end users to maximize the total number of tenants without violating their Service Level Agreement (SLA). This paper proposes a hybrid approach to solve placement of tenants which is called Tenant Placement Strategy (TPS). The TPS uses a combination of resource consumption estimation model, service selection with genetic algorithm (GA), case-based reasoning (CBR) and heuristic approach. CBR is proposed for matching existing execution plans which are generated by GA. In order to fully use all types of resources of the servers, a heuristic approach is proposed for selecting the optimal execution plan based on the distance of the tenant resources consumption vector and the server residual resource vector. The results of simulated experiments show that the strategy proposed in this paper is effective in placing tenants.
Enfeng Yang, Yong Zhang 0051, Lei Wu 0002, Shijun Liu
APSCC5
2011 An Integrated Multi-channel Messaging Model supporting for business collaboration
abstract
As the competitions among enterprises become more and more serious, and the production processes become increasingly complex, finding an efficient way to realize collaboration between different processes of industries is necessary for enterprises. Therefore, in order to support business collaboration, we propose an Integrated Multi-channel Messaging Model (IM3) in this paper, which can be used as a stand-alone web-service in many different business systems. The main feature of our model of work is that it can integrate e-mail, SMS and instant message formats into a general purpose message format, and then send it out according to the requirements of users. The general purpose message format is transformed by the use of XSLT (extensible style language transformation), and is also extensible to more message formats. Besides, our model of work has functions of providing message alerts, the coordination between various business processes, the support of business process execution etc. It is an efficient model that provides an open, shared, distributed but centralized collaborative working environment.
Fuhou Liang, Shijun Liu, Xiangxu Meng
CSCWD2
2010 A service-oriented architecture of virtual enterprise for manufacturing industry
abstract
Manufacturing industry is now being challenged with globalization and grid technology provides a reliable and distributed computing solution. In this paper, firstly, a service-oriented architecture of virtual enterprise (VE) is proposed. Then three main modules, which are project management, partner management and business process management, are described how to construct a VE based on manufacturing grid. The project management controls the projects distributed in different enterprises, the partner management maps roles in entity enterprises to that of VE, and the business process management inspects the VE process flows coming from different entity enterprises.
Xiangxu Meng, Shijun Liu
CSCWD3
2010 LBVS: A Load Balancing Strategy for Virtual Storage
abstract
Cloud Storage is an important part of Cloud Computing, and it provides a way to achieve large scale storage architecture. And virtual storage is a strategy for Cloud Storage. A load balancing virtual storage strategy (LBVS) is proposed in this paper. The contribution of this strategy is that, it provides a large scale net data storage model and Storage as a Service model based on Cloud Storage. Three layers architecture is used to achieve storage virtualization, and two load balancing modules are used to balance systems load. This paper describes details of LBVS strategy, the model of virtual storage as well as how to implement the LBVS with iRODS and two load balancing algorithms.
Hao Liu 0026, Shijun Liu, Xiangxu Meng, Yong Zhang 0051
ICSS2
2009 Towards high level SaaS maturity model: Methods and case study
abstract
This paper introduces a software as a service (SaaS) application which is designed and delivered in high level maturity model. In order to realize the configurability, metadata is used to define all the variability points of the application. Meanwhile, JMX is used to manage the metadata so that changed metadata can be hot deployed immediately during runtime. Scalability is discussed both in application layer and data layer. Integration requirements and roadmap between SaaS application and on-premise applications are introduced. What's more, service surrogate extended from SCA and a message engine are used to meet the integration requirements in business processes.
Yong Zhang 0051, Shijun Liu, Xiangxu Meng
APSCC2
2009 Digital media service oriented digital museum Grid
abstract
In this paper we design and implement a digital media service oriented digital museum platform based on Grid. We construct the whole system based on the Open Grid Service Architecture (OGSA), which supports flexibly constructing and deploying various digital media applications. The system shield the heterogeneity of resources via data access service so that the enormous dispersed resources can be effectively accessed and integrated. We also propose an on-demand application constructing method based on service composition. Using this method, we can implement digital media applications by dynamic service binding. Moreover, we implement digital media service applications based on this platform, such as 3D virtual museum walkthrough system.
Xiangxu Meng, Shijun Liu, Hai Guo, Chenglei Yang
CSCWD3
2008 Dynamic Reliable Service Routing in Enterprise Service Bus
abstract
The enterprise service bus (ESB) is the core department of the service oriented architecture (SOA) which takes charge of managing mass services in the SOA. And the message routing among services is a very important mechanism for service communication in ESB, it is the main function provided by ESB. At the present, there have already been several patterns of message routing in ESB, but they only support static configurable routing, and they also cannot ensure the reliability of message routing, so this paper presents a new pattern of message routing. By integrating the service discovery engine and designing dynamic routing component, we solve the limitation of existent static configurable routing and accomplish dynamically reliable message routing.
Shijun Liu, Lei Wu 0002
APSCC2
2008 Generating Associated Relation between Documents
abstract
Traditional text mining techniques have weak ability to provide associated relations with rich semantics that is a foundation of the intelligent browsing of topics, discovery of semantic community and precise personalized recommendation in current Web and knowledge Grid, etc. In this paper we propose an algorithm to generate and calculate the associated relations and their strengths between documents within a domain. Each document is represented by a bag of words and their weights. We first build domain knowledge background based on the association rules at keyword level, and then we apply those association rules to generate and calculate the documents' semantic relations and their strengths at document level, which effectively shorten the semantic gap from keyword semantics to document semantics. Experimental results show that our proposed method is feasible and able to discover interesting facts within a domain.
Xiangfeng Luo, Guoning Liang, Shijun Liu
HPCC3
2008 Research on the UI Integration Architecture of Service System
abstract
As researches of service science have gradually come into practical use, architecture of practical service system has arisen as an important problem. To solve this problem, the paper proposes a kind of service system integration architecture based on UI (user interface) integration technology. Important modules of this architecture and their constructing process are particularly introduced. In the end of this paper, the main characteristics of this architecture are enumerated to summarize the whole architecture.
Shijun Liu, Xiangxu Meng
ICC2
2007 A Visual Lens Toolkit for Mobile Devices
abstract
Current smart mobile devices such as Pocket PCs have become common accepted daily-life personal assistants. Such mobile devices provide much less support for visualization compared to desktop computers. When developing mobile visualization applications, one should take the limited screen size, currently available interaction tools and relatively poor computation power into consideration. We describe a visual lens toolkit integrated with Focus+Context technique and it can be incorporated into various applications on Pocket PCs. Our technique creates a distorted display view of target object in a high resolution focus area that is smoothly embedded into the distorted context. We have tested the toolkit in two applications as stylistic lens and sketch lens separately and received fine experiment results. In remote line drawing of 3D model system, the lens toolkit embedded into the client side mainly serve as stylistic lens which provide the users with different presentation styles of the model according to the current viewpoint and focus. In mobile sketch system, the lens toolkit performs like a refinement filter which adoptively turns the loose sketch or SVG file into finer image. Both tested applications.
Yuezhu Huang, Xiangxu Meng, Chenglei Yang, Shijun Liu
APSCC4
2007 The Research and Implementation of Turning Conference Management System into a Service
abstract
As the need for application service becomes larger and larger, turning software into a service has been addressed as a matter of urgency. There are two kinds of application delivery models, ASP and SaaS. This paper describes a method for turning software into a service using the ASP model, and use conference management system as an example to illustrate the idiographic process. According to the pertinences, software systems can be classified to two types. One type is user-independent, and the other is organization-centralized. This paper describes how to turn the organization-centralized software to a large-granularity application service.
Shijun Liu, Xiangxu Meng
APSCC2
2007 A Service-Oriented, Scalable Approach to Grid-Enabling of Manufacturing Resources
Lei Wu 0002, Xiangxu Meng, Shijun Liu
CDVE3
2007 Research on Semantic Attributes Based Partner Selection in Virtual Manufacturing Enterprises
abstract
Selection of correct partner is one of the most crucial factors in organizing virtual enterprises. Approaches of partner selection do not consider the semantic features in building a virtual enterprise. In this paper, we describe a system that can select partners effectively through semantic attributes with naive Bayesian algorithm. We show that semantic features can be successfully extracted by applying naive Bayesian technology and build a knowledge base of potential partners in manufacturing enterprises. An example of machine tool industry is used to describe the approach.
Xiangxu Meng, Shijun Liu
CSCWD3
2007 Ontology-based Resource Description in Manufacturing Grid
abstract
The exploration in the grid technology and semantic web service promotes the development of modern manufacturing technology. Grid applications need resources and services to be discovered quickly and efficiently. The greatest obstacle we have, however, lies in the difficulties in appropriately describing resources and services. To easily find the needed services, ontology is used to describe resources and services in our Manufacturing Grid (MG) architecture. We extend UDDI in order to enhance the precision of service discovering through embedding service category ontology into service descriptions. In addition, the extended UDDI can support service publishing and service retrieving. We built a prototype to implement resource description based on ontology in MG.
Xiangxu Meng, Shijun Liu, Ran Yuan
CSCWD3
2007 Service-oriented integration of industrial simulation codes
abstract
The service-oriented integration approach is presented to integrate many geographically and logically distributed, heterogeneous industrial simulation codes. The hierarchy is decomposed into four layers: the infrastructure layer, concrete web service layer, business service layer and application layer. The paper presents an approach to encapsulate industrial simulation codes into services following Web Service Resource Framework (WSRF) specification which can make them offer their services and functionality in a standard web service environment. To hide the complexity, changeability and the technical details (such as web services' WSDL address) of concrete web services for business users, these concrete web services are virtualized as business services. Three kinds of virtualization patterns are presented in the paper. Engineers can compose these business services in a high level to complete virtual product development. At last, we give a use case to validate our method and put forward the future work.
Lei Wu 0002, Xiangxu Meng, Shijun Liu, Yuchang Jiao, Lei Meng 0001
CSCWD3
2006 A BPEL4WS-based Composite Service Modeling Solution in Manufacturing Grid
abstract
In manufacturing grid, a lot of applications and resources in manufacturing enterprises are encapsulated as Web services and a cooperative work environment is provided. Applying Web service composition and workflow technologies, services are integrated into a composite service to represent a cooperative business process. Therefore the enterprise cooperation can be implemented by executing and monitoring the composite service. According to the characteristics of manufacturing, this paper presents a composite service modeling solution based on BPEL4WS and discusses how this solution support BPEL4WS during the whole modeling process. The solution includes dynamic service discovery and service selection to support a flexible binding with the partners' services. This paper also describes how to execute and monitor the composite service
Lei Duan, Shijun Liu, Changhe Tu, Xiangxu Meng
APSCC2
2006 Dynamic Manufacturing Job Management in Manufacturing Grid
abstract
The complexities of manufacturing resources greatly increase the difficulty of managing them. In this paper, with the help of dynamic manufacturing job management which is designed to invoke various types of services, a feasible approach is presented. According to the different mechanisms of service invoking, the invoking engine is divided into three categories: composite service invoking engine, Web service invoking engine and OGSA-DAI service invoking engine. Furthermore, through the use of composite service invoking engine, not only Web services but also flows which are composed of several manufacturing services can be invoked and controlled in a uniform way
Xiangxu Meng, Shijun Liu, Ruyue Ma, Lei Wu 0002
CSCWD3
2005 An approach for flexible RBAC workflow system
abstract
With the fast increase of electronic commerce, more and more enterprises and organizations are facilitating their business processes by workflow. To protect information secure and meet the requirement of frequent business changes over time, the security and flexibility become two of the most important aspects that attract attention both from academy and industry. Many related research work on flexible workflow and secure RBAC model were presented respectively. Unfortunately, the analysis and implementation of enforcing RBAC into Web-based flexible workflow have not been mentioned The intention of this paper is to extend RBAC framework further to flexible workflow to support the security, flexibility and expansibility of organization business. A model and its corresponding mechanism are introduced for establishment, dynamical customization and run-time management of the RBAC workflow. A practical system for Property Right Exchange (PRES) based on this model is implemented.
Yuqing Sun 0001, Xiangxu Meng, Shijun Liu
CSCWD (1)3
2005 Resource organizing in the manufacturing grid
abstract
In the paper, we are devoted to resolve the problem of resource organizing in manufacturing grid. A manufacturing grid (MG) architecture is first introduced which adopts the thoughts of peer-to-peer system and Web service. Then the issue to organize the resource in the manufacturing grid is mainly introduced. To describe the method of resource organizing, we introduce processes of resource publishing, resource searching and resource transfer or duplication when the hub joins in or exits from the grid are introduces seriatim. A distributed hashing algorithm is also presented. Then a prototype system of manufacturing grid is built to validate the way to organize resources.
Yexin Tong, Xiangxu Meng, Shijun Liu, Lei Wu 0002
CSCWD (2)4
2005 A trading supported manufacturing resource sharing model for manufacturing grid
abstract
With the development of grid technology, manufacturing grid (MG) has been proposed to standardize networked manufacturing modes in industry. As manufacturing resources are complicated, heterogeneous, geographically distributed, and owned by different enterprises, manufacturing resource sharing is much more complicated in MG. Resource owners and users have respective requirements, and then MG should provide capability to trade between them. This paper presents a trading supported manufacturing resources sharing model. Organized in a P2P pattern, Hubs, computer nodes deployed core services required, make up the backbone of the model and responsible for aggregating manufacturing resource. In the model, the resource selection process includes two phases: resource discovering and resource trading. Firstly, discover all the potential appreciate resources by multi-keys hashing; Secondly, gain the best appreciate resources by resource trading to make the benefits of both sides best. This model supports many kinds of economic models. The resources owners and users can use any one or some of these models or even combinations of them in meeting their objectives. The procedure of the competitive bidding trade and the method of evaluating tenders are discussed in detail.
Lei Wu 0002, Xiangxu Meng, Shijun Liu, Yexin Tong
CSCWD (1)3