VLDB 2026 Research / reviewers in the wild / expert
Yong-Goo Shin
dblp:165/2150
· DBLP profile ↗
17ranked-venue papers
2as first author
12since 2021 · last 2025
0000-0002-2189-3886ORCID · verified
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 13 · 1 first-author · 12 since 2021Graphics, computer vision, multimedia, augmented reality and games · 4 · 1 first-authorApplied, interdisciplinary, general and emerging computing · 1
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | A Novel Generator With Auxiliary Branch for Improving GAN PerformanceabstractThe generator in the generative adversarial network (GAN) learns image generation in a coarse-to-fine manner in which earlier layers learn the overall structure of the image and the latter ones refine the details. To propagate the coarse information well, recent works usually build their generators by stacking up multiple residual blocks. Although the residual block can produce a high-quality image as well as be trained stably, it often impedes the information flow in the network. To alleviate this problem, this brief introduces a novel generator architecture that produces the image by combining features obtained through two different branches: the main and auxiliary branches. The goal of the main branch is to produce the image by passing through the multiple residual blocks, whereas the auxiliary branch is to convey the coarse information in the earlier layer to the later one. To combine the features in the main and auxiliary branches successfully, we also propose a gated feature fusion module (GFFM) that controls the information flow in those branches. To prove the superiority of the proposed method, this brief provides extensive experiments using various standard datasets including CIFAR-10, CIFAR-100, LSUN, CelebA-HQ, AFHQ, and tiny-ImageNet. Furthermore, we conducted various ablation studies to demonstrate the generalization ability of the proposed method. Quantitative evaluations prove that the proposed method exhibits impressive GAN performance in terms of Inception score (IS) and Frechet inception distance (FID). For instance, the proposed method boosts the FID and IS scores on the tiny-ImageNet dataset from 35.13 to 25.00 and 20.23 to 25.57, respectively. Seung Park, Yong-Goo Shin |
IEEE Trans. Neural Networks Learn. Syst. | 2 |
| 2025 | Rethinking Image Skip Connections in StyleGAN2abstractVarious models based on StyleGAN have gained significant traction in the field of image synthesis, attributed to their robust training stability and superior performances. Within the StyleGAN framework, the adoption of image skip connection is favored over the traditional residual connection. However, this preference is just based on empirical observations; there has not been any in-depth mathematical analysis on it yet. To rectify this situation, this brief aims to elucidate the mathematical meaning of the image skip connection and introduce a groundbreaking methodology, termed the image squeeze connection, which significantly improves the quality of image synthesis. Specifically, we analyze the image skip connection technique to reveal its problem and introduce the proposed method which not only effectively boosts the GAN performance but also reduces the required number of network parameters. Extensive experiments on various datasets demonstrate that the proposed method consistently enhances the performance of state-of-the-art models based on StyleGAN. We believe that our findings represent a vital advancement in the field of image synthesis, suggesting a novel direction for future research and applications. Seung Park, Yong-Goo Shin |
IEEE Trans. Neural Networks Learn. Syst. | 2 |
| 2024 | Prognostic prediction of sepsis patient using transformer with skip connected token for tabular data
Jee-Woo Choi, Min-Uk Yang, Jae-Woo Kim, Yoon Mi Shin, Yong-Goo Shin, Seung Park |
Artif. Intell. Medicine | 5 |
| 2024 | A unified framework to stereotyped behavior detection for screening Autism Spectrum Disorder
Cheol-Hwan Yoo, Jang-Hee Yoo, Moon-Ki Back, Woo-Jin Wang, Yong-Goo Shin |
Pattern Recognit. Lett. | 5 |
| 2024 | Conditional Convolution Projecting Latent Vectors on Condition-Specific SpaceabstractDespite rapid advancements over the past several years, the conditional generative adversarial networks (cGANs) are still far from being perfect. Although one of the major concerns of the cGANs is how to provide the conditional information to the generator, there are not only no ways considered as the optimal solution but also a lack of related research. This brief presents a novel convolution layer, called the conditional convolution (cConv) layer, which incorporates the conditional information into the generator of the generative adversarial networks (GANs). Unlike the most general framework of the cGANs using the conditional batch normalization (cBN) that transforms the normalized feature maps after convolution, the proposed method directly produces conditional features by adjusting the convolutional kernels depending on the conditions. More specifically, in each cConv layer, the weights are conditioned in a simple but effective way through filter-wise scaling and channel-wise shifting operations. In contrast to the conventional methods, the proposed method with a single generator can effectively handle condition-specific characteristics. The experimental results on CIFAR, LSUN, and ImageNet datasets show that the generator with the proposed cConv layer achieves a higher quality of conditional image generation than that with the standard convolution layer. Min-Cheol Sagong, Yoon-Jae Yeo, Yong-Goo Shin, Sung-Jea Ko |
IEEE Trans. Neural Networks Learn. Syst. | 3 |
| 2023 | Effective shortcut technique for generative adversarial networks
Seung Park, Cheol-Hwan Yoo, Yong-Goo Shin |
Appl. Intell. | 3 |
| 2023 | Image generation with self pixel-wise normalization
Yoon-Jae Yeo, Min-Cheol Sagong, Seung Park, Sung-Jea Ko, Yong-Goo Shin |
Appl. Intell. | 5 |
| 2022 | Generative residual block for image generation
Seung Park, Yong-Goo Shin |
Appl. Intell. | 2 |
| 2022 | PConv: simple yet effective convolutional layer for generative adversarial network
Seung Park, Yoon-Jae Yeo, Yong-Goo Shin |
Neural Comput. Appl. | 3 |
| 2022 | Generative convolution layer for image generation
Seung Park, Yong-Goo Shin |
Neural Networks | 2 |
| 2022 | Simple Yet Effective Way for Improving the Performance of GANabstractIn adversarial learning, the discriminator often fails to guide the generator successfully since it distinguishes between real and generated images using silly or nonrobust features. To alleviate this problem, this brief presents a simple but effective way that improves the performance of the generative adversarial network (GAN) without imposing the training overhead or modifying the network architectures of existing methods. The proposed method employs a novel cascading rejection (CR) module for discriminator, which extracts multiple nonoverlapped features in an iterative manner using the vector rejection operation. Since the extracted diverse features prevent the discriminator from concentrating on nonmeaningful features, the discriminator can guide the generator effectively to produce images that are more similar to the real images. In addition, since the proposed CR module requires only a few simple vector operations, it can be readily applied to existing frameworks with marginal training overheads. Quantitative evaluations on various data sets, including CIFAR-10, CelebA, CelebA-HQ, LSUN, and tiny-ImageNet, confirm that the proposed method significantly improves the performance of GAN and conditional GAN in terms of the Frechet inception distance (FID), indicating the diversity and visual appearance of the generated images. Yoon-Jae Yeo, Yong-Goo Shin, Seung Park, Sung-Jea Ko |
IEEE Trans. Neural Networks Learn. Syst. | 2 |
| 2021 | PEPSI++: Fast and Lightweight Network for Image InpaintingabstractAmong the various generative adversarial network (GAN)-based image inpainting methods, a coarse-to-fine network with a contextual attention module (CAM) has shown remarkable performance. However, due to two stacked generative networks, the coarse-to-fine network needs numerous computational resources, such as convolution operations and network parameters, which result in low speed. To address this problem, we propose a novel network architecture called parallel extended-decoder path for semantic inpainting (PEPSI) network, which aims at reducing the hardware costs and improving the inpainting performance. PEPSI consists of a single shared encoding network and parallel decoding networks called coarse and inpainting paths. The coarse path produces a preliminary inpainting result to train the encoding network for the prediction of features for the CAM. Simultaneously, the inpainting path generates higher inpainting quality using the refined features reconstructed via the CAM. In addition, we propose Diet-PEPSI that significantly reduces the network parameters while maintaining the performance. In Diet-PEPSI, to capture the global contextual information with low hardware costs, we propose novel rate-adaptive dilated convolutional layers that employ the common weights but produce dynamic features depending on the given dilation rates. Extensive experiments comparing the performance with state-of-the-art image inpainting methods demonstrate that both PEPSI and Diet-PEPSI improve the qualitative scores, i.e., the peak signal-to-noise ratio (PSNR) and structural similarity (SSIM), as well as significantly reduce hardware costs, such as computational time and the number of network parameters. Yong-Goo Shin, Min-Cheol Sagong, Yoon-Jae Yeo, Seung-Wook Kim 0002, Sung-Jea Ko |
IEEE Trans. Neural Networks Learn. Syst. | 1 |
| 2020 | Self-Attentive Normalization for Automated Gleason Grading SystemabstractRecently, convolutional neural networks (CNNs)- based automated Gleason grading system for prostate cancer has been widely researched. However, these systems still need further improvement to achieve pathologist-level performance. To this end, this paper introduces a novel self-attentive normalization (SAN) which is the first work to employ the attention mechanism for the automated Gleason grading system. Unlike conventional normalization techniques, e.g. batch normalization and instance normalization, which learn a single affine transformation, the proposed method can learn the elementwise affine transformation to focus on more informative regions of the feature map. Since SAN requires a small number of extra learning parameters, it can be integrated into existing automated Gleason grading systems seamlessly with negligible overheads. Extensive quantitative evaluations show that, by applying SAN to various CNN architectures, the diagnostic accuracy can be significantly improved. For instance, we raise VGG-16's diagnostic accuracy from 73.99% to 79.16% on the Harvard Dataverse. Hong-Kyu Shin, Sung-Hoo Hong, Yeong-Jin Choi, Yong-Goo Shin, Seung Park, Sung-Jea Ko |
TENCON | 4 |
| 2020 | Simple Yet Effective Way for Improving the Performance of Depth Map Super-ResolutionabstractIn depth map super-resolution (SR), a high-resolution color image plays an important role as guidance for preventing blurry depth boundaries. However, excessive/deficient use of the color image features often causes performance degradation such as texture-copying/edge-smoothing in flat/boundary areas. To alleviate these problems, this letter presents a simple yet effective method for enhancing the performance of the SR without requiring significant modifications to the original SR network. To this end, we present a self-selective concatenation (SSC), which is a substitute for the conventional feature concatenation. In the upsampling layers of the SR network, the SSC extracts spatial and channel attention from both color and depth features such that color features can be selectively used for depth SR. Specifically, the SSC learns to use sufficient color features for rendering sharp depth boundaries, whereas their effects are reduced in smooth regions to prevent texture-copying. The proposed SSC can be included in any existing SR networks that have the encoder-decoder structure. The experimental results show that the proposed method can further improve the performances of existing SR networks in terms of the root mean squared error and peak signal-to-noise ratio. Yoon-Jae Yeo, Min-Cheol Sagong, Yong-Goo Shin, Seung-Won Jung, Sung-Jea Ko |
IEEE Signal Process. Lett. | 3 |
| 2020 | Simple Yet Effective Way for Improving the Performance of Lossy Image CompressionabstractLossy image compression methods with deep neural network (DNN) include a quantization process between encoder and decoder networks as an essential part to increase the compression rate. However, the quantization operation impedes the flow of gradient and often disturbs the optimal learning of the encoder, which results in distortion in the reconstructed images. To alleviate this problem, this paper presents a simple yet effective way that enhances the performance of lossy image compression without imposing training overhead or modifying the original network architectures. In the proposed method, we utilize an auxiliary branch called a shortcut which directly connects the encoder and decoder. Since the shortcut does not include the quantization process, it supports the optimal learning of the encoder by flowing the accurate gradient. Furthermore, to assist the decoder which should handle additional feature maps obtained via the shortcut, we also propose a residual refinement unit (RRU) following the quantizer. The experimental results show that the image compression network trained with the proposed method remarkably improves the performance in terms of peak signal-to-noise ratio (PSNR), structural similarity (SSIM), and multi-scale structural similarity (MS-SSIM). Yoon-Jae Yeo, Yong-Goo Shin, Min-Cheol Sagong, Seung-Wook Kim 0002, Sung-Jea Ko |
IEEE Signal Process. Lett. | 2 |
| 2020 | Unsupervised Deep Contrast Enhancement With Power Constraint for OLED DisplaysabstractVarious power-constrained contrast enhance-ment (PCCE) techniques have been applied to an organic light emitting diode (OLED) display for reducing the pow-er demands of the display while preserving the image qual-ity. In this paper, we propose a new deep learning-based PCCE scheme that constrains the power consumption of the OLED displays while enhancing the contrast of the displayed image. In the proposed method, the power con-sumption is constrained by simply reducing the brightness a certain ratio, whereas the perceived visual quality is pre-served as much as possible by enhancing the contrast of the image using a convolutional neural network (CNN). Furthermore, our CNN can learn the PCCE technique without a reference image by unsupervised learning. Ex-perimental results show that the proposed method is supe-rior to conventional ones in terms of image quality assess-ment metrics such as a visual saliency-induced index (VSI) and a measure of enhancement (EME).1. Yong-Goo Shin, Seung Park, Yoon-Jae Yeo, Min-Jae Yoo, Sung-Jea Ko |
IEEE Trans. Image Process. | 1 |
| 2019 | PEPSI : Fast Image Inpainting With Parallel Decoding NetworkabstractRecently, a generative adversarial network (GAN)-based method employing the coarse-to-fine network with the contextual attention module (CAM) has shown outstanding results in image inpainting. However, this method requires numerous computational resources due to its two-stage process for feature encoding. To solve this problem, in this paper, we present a novel network structure, called PEPSI: parallel extended-decoder path for semantic inpainting. PEPSI can reduce the number of convolution operations by adopting a structure consisting of a single shared encoding network and a parallel decoding network with coarse and inpainting paths. The coarse path produces a preliminary inpainting result with which the encoding network is trained to predict features for the CAM. At the same time, the inpainting path creates a higher-quality inpainting result using refined features reconstructed by the CAM. PEPSI not only reduces the number of convolution operation almost by half as compared to the conventional coarse-to-fine networks but also exhibits superior performance to other models in terms of testing time and qualitative scores. Min-Cheol Sagong, Yong-Goo Shin, Seung-Wook Kim 0002, Seung Park, Sung-Jea Ko |
CVPR | 2 |