Kiyoshi Honda

dblp:72/2302 · DBLP profile ↗
← Back
61ranked-venue papers
8as first author
7since 2021 · last 2023
0000-0002-7725-5031ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Graphics, computer vision, multimedia, augmented reality and games · 47 · 4 first-author · 5 since 2021Artificial intelligence and machine learning · 36 · 4 first-author · 3 since 2021Software engineering, systems software and programming languages · 10 · 4 first-author · 2 since 2021Applied, interdisciplinary, general and emerging computing · 4 · 1 first-author · 1 since 2021
YearPublicationVenuePosition
2023 Frequency Patterns of Individual Speaker Characteristics at Higher and Lower Spectral Ranges
Ju Zhang 0001, Yujie Chi, Kiyoshi Honda, Jianguo Wei
INTERSPEECH5
2023 Transvelar Nasal Coupling Contributing to Speaker Characteristics in Non-nasal Vowels
Yujie Chi, Kiyoshi Honda, Jianguo Wei
INTERSPEECH4
2022 Compressing Transformer-Based ASR Model by Task-Driven Loss and Attention-Based Multi-Level Feature Distillation
abstract
The current popular knowledge distillation (KD) methods effectively compress the transformer-based end-to-end speech recognition model. However, existing methods fail to utilize complete information of the teacher model, and they distill only a limited number of blocks of the teacher model. In this study, we first integrate a task-driven loss function into the decoder’s intermediate blocks to generate task-related feature representations. Then, we propose an attention-based multi-level feature distillation to automatically learn the feature representation summarized by all blocks of the teacher model. Under the 1.1M parameters model, the experimental results on the Wall Street Journal dataset reveal that our approach achieves a 12.1% WER reduction compared with the baseline system.
Yongjie Lv, Longbiao Wang, Meng Ge, Sheng Li 0010, Chenchen Ding, Lixin Pan, Yuguang Wang 0003, Jianwu Dang 0001, Kiyoshi Honda
ICASSP9
2022 Vocal-Tract Area Functions with Articulatory Reality for Tract Opening
Ju Zhang 0001, Jianguo Wei, Kiyoshi Honda, Tatsuya Kitamura
INTERSPEECH4
2021 Preliminary Literature Review of Machine Learning System Development Practices
abstract
To guide practitioners and researchers to design and research Machine Learning (ML) system development processes, we conduct a preliminary literature review on ML system development practices. We identified seven papers and two other papers determined in an ad-hoc review. Our findings include emphasized phases in ML system developments, frequently described ML-specific practices, and tailored traditional practices.
Yasuhiro Watanabe, Hironori Washizaki, Kazunori Sakamoto, Daisuke Saito, Kiyoshi Honda, Naohiko Tsuda, Yoshiaki Fukazawa, Nobukazu Yoshioka
COMPSAC5
2021 Portable Photoglottography for Monitoring Vocal Fold Vibrations in Speech Production
abstract
Photoglottography (PGG) is an effective method to monitor vocal fold vibrations via measuring light transmission across the glottis. The difficulty in operation however limits its wide use in speech studies. This paper is to realize a portable PGG (P-PGG) module with an audio interface to record glottal and speech waveforms simultaneously with ease. Near Infrared (NIR) LEDs are driven as a light source and an extremely high-gain photodetector circuit is employed. The output PGG signal is subsequently band-pass filtered and amplified for recording. The whole system is minimized, battery powered with well-controlled heat radiation. In experiments, P-PGG, EGG and microphone are worn together by speakers. The results verify that the NIR lighting P-PGG is successful at recording complete information of glottal cycles in comparison to EGG. Thus, the P-PGG is an effective method for investigating phonation types and consonant-vowel interactions in speech.
Yujie Chi, Kiyoshi Honda, Jianguo Wei
ICASSP2
2021 Data-Driven Persona Retrospective Based on Persona Significance Index in B-to-B Software Development
abstract
Business-to-Business (B-to-B) software development companies develop services to satisfy their customers’ requirements. Developers should prioritize customer satisfaction because customers greatly influence agile software development. However, satisfying current customer’s requirements may not fulfill actual users or future customers’ requirements because customers’ requirements are not always derived from actual users. To reconcile these differences, developers should identify conflicts in their strategic plan. This plan should consider current commitments to end users and their intentions as well as employ a data-driven approach to adapt to rapid market changes. A persona models an end user representation in human-centered design. Although previous works have applied personas to software development and proposed data-driven software engineering frameworks with gap analysis between the effectiveness of commitments and expectations, the significance of developers’ commitment and quantitative decision-making are not considered. Developers often do not achieve their business goal due to conflicts. Hence, the target of commitments should be validated. To address these issues, we propose Data-Driven Persona Retrospective (DDR) to help developers plan future releases. DDR, which includes the Persona Significance Index (PerSI) to reflect developers’ commitments to end users’ personas, helps developers identify a gap between developers’ commitments to personas and expectations. In addition, DDR identifies release situations with conflicts based on PerSI. Specifically, we define four release cases, which include different situations and issues, and provide a method to determine the release case based on PerSI. Then we validate the release cases and their determinations through a case study involving a Japanese cloud application and discuss the effectiveness of DDR.
Yasuhiro Watanabe, Hironori Washizaki, Yoshiaki Fukazawa, Kiyoshi Honda, Masahiro Taga, Akira Matsuzaki, Takayoshi Suzuki
Int. J. Softw. Eng. Knowl. Eng.4
2020 Retrieving Vocal-Tract Resonance and anti-Resonance From High-Pitched Vowels Using a Rahmonic Subtraction Technique
abstract
Vocal tract resonances give rise to core spectral information of speech signals. Linear prediction and cepstral methods are widely used for this purpose. However, both approaches are prone to fail as the fundamental frequency (F0) rises. In this study, a new cepstral method is developed combined with a refined rahmonic subtraction technique (RS-CEPS) to extract spectral envelopes excited by glottal noise sources. A vowel synthesis system based on 3D-printed solid vocal tract models is used to obtain reference transfer functions for accuracy verification. A series of stable vowels /a/ was synthesized for a wide F0 range. By analyzing the synthetic vowels, the results showed that the RS-CEPS yields accurate estimates of resonance-peak and anti-resonance frequencies in comparison to those from the conventional methods. The RS-CEPS is simple and stable, offering a potential for expanding speech analysis applications.
Kiyoshi Honda, Jianguo Wei
ICASSP2
2020 Investigation of Effectively Synthesizing Code-Switched Speech Using Highly Imbalanced Mix-Lingual Data
Shaotong Guo, Longbiao Wang, Sheng Li 0010, Ju Zhang 0001, Yuguang Wang 0003, Jianwu Dang 0001, Kiyoshi Honda
ICONIP (1)8
2020 Regional Resonance of the Lower Vocal Tract and its Contribution to Speaker Characteristics
abstract
S.1391-1395
Lin Zhang 0054, Kiyoshi Honda, Jianguo Wei, Seiji Adachi
INTERSPEECH2
2019 Glottographic and Aerodynamic Analysis on Consonant Aspiration and Onset F0 in Mandarin Chinese
abstract
Stop consonants in Mandarin Chinese are all voiceless at word-initial positions only showing aspirated and unaspirated distinctions. Between the two phonation types, voice onset time (VOT) shows a clear contrast in duration, whereas voice onset fundamental frequency (onset F0) does not, as seen in previous studies. This study reports our first use of improved instrumentation techniques to record speech sound, oral airflow and glottal activity to examine consonant aspiration and onset F0. Glottal abduction, adduction and vibration cycles are monitored by a new external photo-glottographic system (ePGG) refined for better signal quality to detect accurate voice onset. Experimental data on Mandarin stops obtained from two male subjects suggests that consonant aspiration results in a large variation of VOT and oral airflow at voice onset. The onset F0 shows individual variation of falling and rising contours, and it is higher in aspirated stops than in unaspirated ones.
Yujie Chi, Kiyoshi Honda, Jianguo Wei
ICASSP2
2019 Inappropriate Usage Examples in Web API Documentations
abstract
Application Programming Interfaces (APIs) are common in software development to reuse other products. Although the documentation allows API consumers to learn about API usages, it can be unreliable. Here, we investigate the characteristics of inappropriate usage examples in web API documentation by extracting and comparing OpenAPI Specifications from usage example-response pairs. About 65.5% of the endpoints have some form of inappropriate usage examples. Furthermore, mismatches are classified into four categories: undocumented keys pattern, dynamic keys pattern, unreturned keys pattern, and type mismatched pattern. Our results suggest that the number of keys in the response is correlated with the number of mismatches. These findings should assist both API providers and consumers who deal with unreliable documentation in web APIs.
Masaki Hosono, Susumu Tokumoto, Supasit Monpratarnchai, Hironori Washizaki, Kiyoshi Honda, Hiromasa Nagumo, Hisanobu Sonoda, Yoshiaki Fukazawa, Kazuki Munakata, Takao Nakagawa, Yusuke Nemoto
ICSME5
2019 Acoustic and Articulatory Study of Ewe Vowels: A Comparative Study of Male and Female
Kowovi Comivi Alowonou, Jianguo Wei, Wenhuan Lu, Kiyoshi Honda, Jianwu Dang 0001
INTERSPEECH5
2019 Individual Difference of Relative Tongue Size and its Acoustic Effects
Chongke Bi, Kiyoshi Honda, Wenhuan Lu, Jianguo Wei
INTERSPEECH3
2018 An Empirical Study on the Reliability of the Web API Document
abstract
The importance of APIs in software development, especially web APIs, has increased Developers read documentation, which is available on the internet, and use the corresponding APIs in their products. However, documentation occasionally contains mistakes. Such mistakes can confuse developers or lead to defects that lower the quality of the product. In this paper, we investigate the reliability of web APIs by extracting and comparing OpenAPI specifications from both the documentations and the results of the API calls. Almost half of the documentations are somehow unreliable. Mismatches between documentation and the response can be categorized into four types: 1) Undocumented Keys, 2) Dynamic Keys, 3) Unreturned Keys, and 4) Type Mismatched. This study will help developers design more reliable products.
Masaki Hosono, Hironori Washizaki, Yoshiaki Fukazawa, Kiyoshi Honda
APSEC4
2018 Tongue Segmentation with Geometrically Constrained Snake Model
Zhihua Su, Jianguo Wei, Qiang Fang 0003, Jianrong Wang, Kiyoshi Honda
INTERSPEECH5
2018 Study of articulators' contribution and compensation during speech by articulatory speech recognition
Jianguo Wei, Jingshu Zhang, Qiang Fang 0003, Wenhuan Lu, Kiyoshi Honda, Xugang Lu
Multim. Tools Appl.6
2018 Tooth visualization in vowel production MR images for three-dimensional vocal tract modeling
Ju Zhang 0001, Kiyoshi Honda, Jianguo Wei
Speech Commun.2
2017 Generalized Software Reliability Model Considering Uncertainty and Dynamics: Model and Applications
abstract
Today’s development environment has changed drastically; the development periods are shorter than ever and the number of team members has increased. Consequently, controlling the activities and predicting when a development will end are difficult tasks. To adapt to changes, we propose a generalized software reliability model (GSRM) based on a stochastic process to simulate developments, which include uncertainties and dynamics such as unpredictable changes in the requirements and the number of team members. We assess two actual datasets using our formulated equations, which are related to three types of development uncertainties by employing simple approximations in GSRM. The results show that developments can be evaluated quantitatively. Additionally, a comparison of GSRM with existing software reliability models confirms that the approximation by GSRM is more precise than those by existing models.
Kiyoshi Honda, Hironori Washizaki, Yoshiaki Fukazawa
Int. J. Softw. Eng. Knowl. Eng.1
2016 Effects of Subglottal-Coupling and Interdental-Space on Formant Trajectories During Front-to-Back Vowel Transitions in Chinese
Shuanglin Fan, Kiyoshi Honda, Jianwu Dang 0001
INTERSPEECH2
2016 Audio-visual speech recognition integrating 3D lip information obtained from the Kinect
Jianrong Wang, Ju Zhang 0001, Kiyoshi Honda, Jianguo Wei, Jianwu Dang 0001
Multim. Syst.3
2015 Vocal responses to frequency modulated composite sinewaves via auditory and vibrotactile pathways
abstract
Feedback control mechanisms for speaking have been examined using the transformed auditory feedback (TAF) technique. Previous studies have shown that speakers demonstrate fundamental frequency (F0) changes when they monitor their voice with artificial alterations of F0. However, those studies underestimate the role of vibrotactile information involved in feedback F0 control. This pilot study aims at exploring whether and how vibrotactile information from the larynx influences vowel F0. Participants in our experiment were asked to sustain vowel with their F0 adjusted to composite sinewave stimuli, which were given via auditory and vibrotactile channels using a headset on the ears or a bone-conduction transducer on the larynx. Results revealed the greater compensatory responses to combined vibrotactile-auditory stimuli than to the responses to auditory-only stimuli. The effect of vibrotactile stimuli on feedback F0 adjustment was also observed with the shorter latency of the responses.
Kiyoshi Honda, Jianwu Dang 0001, Jianguo Wei
ICASSP2
2015 Combined cine- and tagged-MRI for tracking landmarks on the tongue surface
Honghao Bao, Wenhuan Lu, Kiyoshi Honda, Jianguo Wei, Qiang Fang 0003, Jianwu Dang 0001
INTERSPEECH3
2015 Measuring oral and nasal airflow in production of Chinese plosive
Yujie Chi, Kiyoshi Honda, Jianguo Wei, Jianwu Dang 0001
INTERSPEECH2
2014 Predicting Time Range of Development Based on Generalized Software Reliability Model
abstract
Development environments have changed drastically, development periods are shorter than ever and the number of team members has increased. These changes have led to difficulties in controlling the development activities and predicting when the development will end. Especially, quality managers try to control software reliability and project managers try to estimate the end of development for planning developing term and distribute the manpower to other developments. In order to assess recent software developments, we propose a generalized software reliability model (GSRM) based on a stochastic process, and simulate developments that include uncertainties and dynamics. We also compare our simulation results to those of other software reliability models. Using the values of uncertainties and dynamics obtained from GSRM, we can evaluate the developments in a quantitative manner. Additionally, we use equations to define the uncertainty regarding the time required to complete a development, and predict whether or not a development will be completed on time. We compare GSRM with an existing model using two old actual datasets and one new actual dataset which we collected, and show that the approximation curve generated by GSRM is about 12% more precise than that generated by the existing model. Furthermore, GSRM can narrow down the predicted time range in which a development will end to less than 40% of that obtained by the existing model.
Kiyoshi Honda, Hidenori Nakai, Hironori Washizaki, Yoshiaki Fukazawa, Ken Asoh, Kazuyoshi Takahashi, Kentarou Ogawa, Maki Mori, Takashi Hino, Yosuke Hayakawa, Yasuyuki Tanaka, Shinichi Yamada, Daisuke Miyazaki
APSEC (1)1
2014 Initial Industrial Experience of GQM-Based Product-Focused Project Monitoring with Trend Patterns
abstract
It is important for project stakeholders to identify the states of projects and quality of products. Although metrics are useful for identifying them, it is difficult for project stake-holders to select appropriate metrics and determine the purpose of measuring metrics. We propose an approach that defines the measured metrics by GQM method to identify tendency in projects and products based on Trend Pattern. Additionally, we implement a tool as a Jenkins Plug in to visualize an evaluation results based on GQM method. We perform an industrial case study, which objects are two software development projects. In our industrial case study, we can identify the problem that product contains. As our future work, we will adopt our approach and GQM Plug in to software development project continuously to assess their effectiveness in the long term.
Hidenori Nakai, Kiyoshi Honda, Hironori Washizaki, Yoshiaki Fukazawa, Ken Asoh, Kazuyoshi Takahashi, Kentarou Ogawa, Maki Mori, Takashi Hino, Yosuke Hayakawa, Yasuyuki Tanaka, Shinichi Yamada, Daisuke Miyazaki
APSEC (2)2
2014 Continuous Product-Focused Project Monitoring with Trend Patterns and GQM
abstract
It is important for project stakeholders to identify the states of projects and quality of products. Although metrics are useful for identifying them, it is difficult for project stakeholders to select appropriate metrics and determine the purpose of measuring metrics. We propose an approach that defines the measured metrics by GQM method, and supports identifying tendency in projects and products based on Trend Pattern. Additionally, we implement a tool as a Jenkins Plug in which to visualizes an evaluation results based on GQM method. We perform an experiment with OSS and industrial case study with two software development projects. In our experiment, we can identify the problem and project tendency. In our industrial case study, we can also identify the problem that project contains. As our future work, we will adopt our approach and GQM Plug in to software development project continuously to assess their effectiveness in the long term.
Hidenori Nakai, Kiyoshi Honda, Hironori Washizaki, Yoshiaki Fukazawa, Ken Asoh, Kazuyoshi Takahashi, Kentarou Ogawa, Maki Mori, Takashi Hino, Yosuke Hayakawa, Yasuyuki Tanaka, Shinichi Yamada, Daisuke Miyazaki
APSEC (2)2
2014 Predicting Release Time Based on Generalized Software Reliability Model (GSRM)
abstract
Development environments have changed drastically, development periods are shorter than ever and the number of team members has increased. Especially in open source software (OSS), a large number of developers contribute to OSS. OSS has difficulties in predicting or deciding when it will be released. In order to assess recent software developments, we proposed a generalized software reliability model (GSRM) based on a stochastic process, and compared GSRM with other models. In this paper, we focus on the release dates of OSS and the growth of faults (issues).
Kiyoshi Honda, Hironori Washizaki, Yoshiaki Fukazawa
COMPSAC1
2014 Detection of speaker individual information using a phoneme effect suppression method
Songgun Hyon, Jianwu Dang 0001, Hongcui Wang, Kiyoshi Honda
Speech Commun.5
2013 An MRI-based acoustic study of Mandarin vowels
Yuguang Wang 0003, Jianwu Dang 0001, Jianguo Wei, Hongcui Wang, Kiyoshi Honda
INTERSPEECH6
2013 A Generalized Software Reliability Model Considering Uncertainty and Dynamics in Development
Kiyoshi Honda, Hironori Washizaki, Yoshiaki Fukazawa
PROFES1
2012 Correlation between vocal tract length, body height, formant frequencies, and pitch frequency for the five Japanese vowels uttered by fifteen male speakers
Hiroaki Hatano, Tatsuya Kitamura, Hironori Takemoto, Parham Mokhtari, Kiyoshi Honda, Shinobu Masaki
INTERSPEECH5
2010 Guest Editorial
Bruce Denby, Tanja Schultz, Kiyoshi Honda
Speech Commun.3
2010 Silent speech interfaces
Bruce Denby, Tanja Schultz, Kiyoshi Honda, Thomas Hueber, J. M. Gilbert, Jonathan S. Brumberg
Speech Commun.3
2007 A Family of Quadratic Snakes for Road Extraction
Ramesh Marikhu, Matthew N. Dailey, Stanislav S. Makhanov, Kiyoshi Honda
ACCV (1)4
2007 An MRI analysis of the extrinsic tongue muscles during vowel production
Sayoko Takano, Kiyoshi Honda
Speech Commun.2
2006 Communication Between Speech Production and Perception Within the Brain-Observation and Simulation
Jianwu Dang 0001, Masato Akagi, Kiyoshi Honda
J. Comput. Sci. Technol.3
2006 Spatiotemporal Fusion of Rice Actual Evapotranspiration With Genetic Algorithms and an Agrohydrological Model
abstract
Monitoring of water consumption in irrigation systems has become increasingly important for water managers in the actual trend of integrated water management. Low spatial resolution (LSR) satellite remote sensing has already proven the capacity of monitoring evapotranspiration (ETa) over large areas at high temporal frequencies, by which monitoring for large irrigation systems can be satisfied. However, smaller pixel size is still required for more local management while keeping the return period within a few days. High spatial resolution (HSR) satellite imagery is indeed available for calculation of ETa and has already been used in many studies. However, its practical return period is a major drawback to its implementation for monitoring irrigation systems. This paper is perusing into the use of genetic algorithms to assimilate parameters of an agrohydrological model called soil-water-air-plant for each of the pixels of HSR images contained into one single pixel of an LSR multitemporal image. The methodology developed and experimented here is trying to take advantage of the spatial content of HSR images and the temporal content of LSR images by fusing them by the process of data assimilation
Yann H. Chemin, Kiyoshi Honda
IEEE Trans. Geosci. Remote. Sens.2
2005 Vocal tract area function inversion by linear regression of cepstrum
abstract
Vocal tract data from 3D cine-MRI are used together with synchronised acoustics to evaluate a linear regression model for inversion. The first two principal components of vocalic area functions are predicted with correlations 0.99 and 0.97 respectively, from 24 FFT-cepstra measured in the frequency band 0-4 kHz. This best regression model together with the two component representation yields mean absolute errors of 0.37 cm 2 in section area and 0.15 cm in vocal tract length.
Parham Mokhtari, Tatsuya Kitamura, Hironori Takemoto, Kiyoshi Honda
INTERSPEECH4
2004 An experimental method for measuring transfer functions of acoustic tubes
abstract
This work proposes an experimental method for direct measurement of transfer functions of acoustic tubes. The method obtains a pressure-to-velocity transfer function from measurement of input volume velocity and output pressures of a target tube. Steady sinusoidal waves from 100 Hz to 5 kHz with a 10-Hz increment were used as a source signal. Experimental results compared with transmission line simulations indicate the following: (1) transfer functions obtained from the measurements agree well with those from transmission line simulations; (2) differences between the resonant frequencies obtained from the measurements and simulations with a uniform tube are less than 2.6 %. These results show conclusive evidence that the proposed method permits accurate measurements of transfer functions of acoustic tubes.
Tatsuya Kitamura, Satoru Fujita, Kiyoshi Honda, Hironori Nishimoto
INTERSPEECH3
2003 Consideration of muscle co-contraction in a physiological articulatory model
abstract
Physiological models of the speech organs must consider cocontraction of the muscles, a common phenomenon taking place during articulation. This study investigated cocontraction of the tongue muscles using the physiological articulatory model that replicates midsagittal regions of the speech organs to simulate articulatory movements during speech [1,2]. The relation between the muscle force and tongue movement obtained by the model simulation indicated that each muscle drives the tongue towards an equilibrium position (EP) corresponding to the magnitude of the activation forces. Contributions of the muscles to the tongue movement were evaluated by the distance between the equilibrium positions. Based on the EPs and the muscle contributions, an invariant mapping (the EP map) was established to function the connection of a spatial location to a muscle force. Cocontractions between agonist and antagonist muscles were simulated using the EP maps. The simulations demonstrated that coarticulation with multiple targets could be compatibly realized using the co-contraction mechanism. The implementation of the co-contraction mechanism enables relatively independent control over the tongue tip and body.
Jianwu Dang 0001, Kiyoshi Honda
INTERSPEECH2
2003 Translation and rotation of the cricothyroid joint revealed by phonation-synchronized high-resolution MRI
Sayoko Takano, Kiyoshi Honda, Shinobu Masaki, Yasuhiro Shimada, Ichiro Fujimoto
INTERSPEECH2
2002 Investigation of coarticulation based on electromagnetic articulographic data
Jianwu Dang 0001, Masaaki Honda, Kiyoshi Honda
INTERSPEECH3
2000 Improvement of a physiological articulatory model for synthesis of vowel sequences
Jianwu Dang 0001, Kiyoshi Honda
INTERSPEECH2
2000 Observation of laryngeal control for voicing and pitch change by magnetic resonance imaging technique
Kiyoshi Honda, Shinobu Masaki, Yasuhiro Shimada
INTERSPEECH1
1998 Speech production of vowel sequences using a physiological articulatory model
abstract
This report describes the development of a physiologically-based articulatory model, which consists of the tongue, mandible, hyoid bone and vocal tract wall. These organs are represented in a quasi-3D shape to replicate a midsagittal layer with a thickness of 2 cm for tongue tissue and 3 cm for tract wall. The geometry of these organs and muscles are extracted from volumetric MR images of a male speaker. Both the soft and rigid structures are represented by mass-points and viscoelastic springs for connective tissue, where the springs for bony organs are set to extremely large stiffness. This design is suitable to compute soft tissue deformations and rigid organ displacements simultaneously using a single algorithm, and thus reduces computational complexities of the simulation. A novel control method is developed to produce dynamic actions of the vocal
Jianwu Dang 0001, Kiyoshi Honda
ICSLP2
1998 An MRI study on the relationship between oral cavity shape and larynx position
Kiyoshi Honda, Mark K. Tiede
ICSLP1
1998 A pressure sensitive palatography: application of new pressure sensitive sheet for measuring tongue-palatal contact pressure
Masahiko Wakumoto, Shinobu Masaki, Kiyoshi Honda, Toshikazu Ohue
ICSLP3
1996 An improved vocal tract model of vowel production implementing piriform resonance and transvelar nasal coupling
Jianwu Dang 0001, Kiyoshi Honda
ICSLP2
1996 Subglottal pressure and final lowering in English
abstract
Quantitative models of intonation in a variety of languages typically specify a long-range downtrend across the sentence that provides a declining backdrop for the steeper rises and falls of more local pitch events such as accents and word tones.Several studies of this "declination" in English and several other languages have isolated a component of somewhat steeper decline that covers only the last few centiseconds of "lab speech" utterances.Other studies suggest that this "final lowering" may be particular to utterances with a "declarative intonation" pattern, and that it is associated particularly with the ends of discourse units.Thus final lowering seems to be associated pragmatically with a sense of fading off or finality.To see whether final lowering can be attributed in part to a fading off of subglottal pressure, we examined the two measures together in two databases of utterances that varied in intonation contour.To minimize confounds from more local pitch specifications, we looked at the relationship between final lowering and subglottal pressure only in the intonational "tail" -i.e., the portion of the contour after the last pitch accent.The slope of the decline of the subglottal pressure varied as a function of the phonological specification of the tones in the tail.Utterances with declarative intonation or with any other contour sharing the phonological specification of a low tone at the end of the tail consistently showed a decline in subglottal pressure, whereas utterances with "yes-no question intonation" or any other contour sharing the phonological specification of a final high tone showed lesser declines or even increases.
Rebecca Herman, Mary E. Beckman, Kiyoshi Honda
ICSLP3
1996 Human palate and related structures: their articulatory consequences
abstract
The vowel space reflects the right-angled shape of the vocal tract, and many consonants exploit the palatal wall.These two facts suggest the importance of the geometry of peripheral structure in speech production.In this study, the relationship between geometry and articulatory variation was examined using a database of English and Japanese speakers.The geometry of each speaker's vocal tract was defined by a quadrilateral bounded by the palatal plane and other rigid structures.This quadrilateral, whose area we refer to as the articulatory (or A) space, provides indices of pharyngeal distance, lower facial height, mandibular position and inclination, and head rotation.The A-spaces of different speakers vary in size and form: the speakers with longer pharyngeal distance tend to have shorter lower facial height.There is also significant variation among speakers in the degree of inclination of the mandibular symphysis.Qualitative comparisons suggested that speakers' vowel articulations adapt to the form of their respective A-space, while consonant articulations seem to be independent of the A-space.
Kiyoshi Honda, Shinji Maeda, Michiko Hashi, Jim Dembowski, John R. Westbury
ICSLP1
1994 Investigation of the acoustic characteristics of the velum for vowels
Jianwu Dang 0001, Kiyoshi Honda
ICSLP2
1994 Global pitch range and the production of low tones in English intonation
Donna Erickson, Kiyoshi Honda, Hiroyuki Hirai, Mary E. Beckman, Seiji Niimi
ICSLP2
1994 A physiological model of speech production and the implication of tongue-larynx interaction
Kiyoshi Honda, Hiroyuki Hirai, Jianwu Dang 0001
ICSLP1
1994 Estimation of temporal processing unit of speech motor programming for Japanese words based on the measurement of reaction time
Shinobu Masaki, Kiyoshi Honda
ICSLP2
1992 Neural network modeling of speech motor control
Makoto Hirayama, Eric Vatikiotis-Bateson, Mitsuo Kawato, Kiyoshi Honda
ICSLP4
1992 The articulatory dynamics of running speech: gestures from phonemes?
Eric Vatikiotis-Bateson, Makoto Hirayama, Kiyoshi Honda, Mitsuo Kawato
ICSLP3
1992 Physiologically Based Speech Synthesis
Makoto Hirayama, Eric Vatikiotis-Bateson, Kiyoshi Honda, Yasuharu Koike, Mitsuo Kawato
NIPS3
1990 A study on respiratory and glottal controls in six western singing qualities: airflow and intensity measurement of professional singing
Jo Estill, Noriko Kobayashi, Kiyoshi Honda, Yuki Kakita
ICSLP3
1990 Sequential control model of speech articulation in producing word utterance
Naoki Kusakawa, Kiyoshi Honda, Yuki Kakita
ICSLP2
1986 Simultaneous high-speed digital recording of vocal fold vibration and speech signal
abstract
A new method for the high-speed digital recording of the vocal fold vibration is presented. The method employs a solid endoscope, a solid-state image sensor and a digital image memory. About 100 frames of continuous image data with 50 × 50 picture elements can be stored in the image memory at one time. Image recording at a rate of 2000 frames per second was achieved using a small light source of a 250 W halogen lamp. The method enables to record the speech signal without special consideration on the camera noises, and also to observe the recorded image on the spot through CRT display. An example of the image data with simultaneously recorded speech and EEG signals are presented comparing the male and female voices, and the chest voice, falsetto and breathy voice.
Shigeru Kiritani, Kiyoshi Honda, Hiroshi Imagawa, Hajime Hirose
ICASSP2