# Kazuyoshi Yoshii Source: https://hello.cv/kazuyoshiyoshii ## Work ### Associate Professor | Kyoto University ### Senior Lecturer | Kyoto University ### Senior Researcher | National Institute of Advanced Industrial Science and Technology (AIST) ### Researcher | National Institute of Advanced Industrial Science and Technology (AIST) ## Education ### Kyoto University | Ph.D. ### Kyoto University | B.S. ### Kyoto University | B.E. ## Publications ### Joint Music Segmentation and Clustering Based on Self-Attentive Contrastive Learning of Multifaceted Self-Similarity Representation IEEE Transactions on Audio, Speech and Language Processing journal-article ### Unsupervised Pitch-Timbre-Variation Disentanglement of Monophonic Music Signals Based on Random Perturbation and Re-entry Training Apsipa Transactions on Signal and Information Processing journal-article ### Joint Music Segmentation and Clustering Based on Self-Attentive Contrastive Learning of Multifaceted Self-Similarity Representation IEEE Transactions on Audio, Speech and Language Processing journal-article ### Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection ArXiv journal-article ### NEURAL STEERER: NOVEL STEERING VECTOR SYNTHESIS WITH A CAUSAL NEURAL FIELD OVER FREQUENCY AND DIRECTION Ieee International Conference on Acoustics, Speech, and Signal Processing Workshops, Icasspw journal-article ### Neural Blind Source Separation and Diarization for Distant Speech Recognition ArXiv journal-article ### End-to-End Singing Transcription Based on CTC and HSMM Decoding with a Refined Score Representation Apsipa Transactions on Signal and Information Processing journal-article ### JOINT AUDIO SOURCE LOCALIZATION AND SEPARATION WITH DISTRIBUTED MICROPHONE ARRAYS BASED ON SPATIALLY-REGULARIZED MULTICHANNEL NMF 16TH INTERNATIONAL WORKSHOP ON ACOUSTIC SIGNAL ENHANCEMENT (IWAENC) journal-article ### NEURAL STEERER: NOVEL STEERING VECTOR SYNTHESIS WITH A CAUSAL NEURAL FIELD OVER FREQUENCY AND DIRECTION ArXiv journal-article ### DOA-Aware Audio-Visual Self-Supervised Learning for Sound Event Localization and Detection Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA) conference-paper ### TIME-DOMAIN AUDIO SOURCE SEPARATION BASED ON GAUSSIAN PROCESSES WITH DEEP KERNEL LEARNING IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) conference-paper ### Neural Fast Full-Rank Spatial Covariance Analysis for Blind Source Separation ArXiv journal-article ### Joint Separation and Localization of Moving Sound Sources Based on Neural Full-Rank Spatial Covariance Analysis IEEE Signal Processing Letters journal-article ### Generalized Fast Multichannel Nonnegative Matrix Factorization Based on Gaussian Scale Mixtures for Blind Source Separation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Generalized Fast Multichannel Nonnegative Matrix Factorization Based on Gaussian Scale Mixtures for Blind Source Separation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Impact of chronological age on efficacy and safety of fluoropyrimidine plus bevacizumab in older non-frail patients with metastatic colorectal cancer: a combined analysis of individual data from two phase II studies of patients aged >75 years Japanese Journal of Clinical Oncology journal-article ### JOINT LOCALIZATION AND SYNCHRONIZATION OF DISTRIBUTED CAMERA-ATTACHED MICROPHONE ARRAYS FOR INDOOR SCENE ANALYSIS 16TH INTERNATIONAL WORKSHOP ON ACOUSTIC SIGNAL ENHANCEMENT (IWAENC) journal-article ### Weakly-Supervised Neural Full-Rank Spatial Covariance Analysis for a Front-End System of Distant Speech Recognition Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Unsupervised Disentanglement of Timbral, Pitch, and Variation Features From Musical Instrument Sounds With Random Perturbation Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA) conference-paper ### A rare case of fibrohistiocytic hepatic inflammatory pseudotumor with cholecystocholangitis showing positive IgG4 staining Clinical Journal of Gastroenterology journal-article ### Autoregressive Moving Average Jointly-Diagonalizable Spatial Covariance Analysis for Joint Source Separation and Dereverberation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Autoregressive Moving Average Jointly-Diagonalizable Spatial Covariance Analysis for Joint Source Separation and Dereverberation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Computationally-Efficient Overdetermined Blind Source Separation Based on Iterative Source Steering IEEE Signal Processing Letters journal-article ### DNN-FREE LOW-LATENCY ADAPTIVE SPEECH ENHANCEMENT BASED ON FRAME-ONLINE BEAMFORMING POWERED BY BLOCK-ONLINE FASTMNMF 16TH INTERNATIONAL WORKSHOP ON ACOUSTIC SIGNAL ENHANCEMENT (IWAENC) journal-article ### Development of a continuum robot enhanced with distributed sensors for search and rescue ROBOMECH Journal journal-article ### Direction-Aware Adaptive Online Neural Speech Enhancement with an Augmented Reality Headset in Real Noisy Conversational Environments IEEE International Conference on Intelligent Robots and Systems conference-paper ### Direction-Aware Adaptive Online Neural Speech Enhancement with an Augmented Reality Headset in Real Noisy Conversational Environments ArXiv journal-article ### Direction-Aware Joint Adaptation of Neural Speech Enhancement and Recognition in Real Multiparty Conversational Environments ArXiv journal-article ### Direction-Aware Joint Adaptation of Neural Speech Enhancement and Recognition in Real Multiparty Conversational Environments Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Elliptically Contoured Alpha-Stable Representation for MUSIC-Based Sound Source Localization European Signal Processing Conference conference-paper ### FLOW-BASED FAST MULTICHANNEL NONNEGATIVE MATRIX FACTORIZATION FOR BLIND SOURCE SEPARATION IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Audio-to-score singing transcription based on a CRNN-HSMM hybrid model APSIPA Transactions on Signal and Information Processing journal-article ### Global Structure-Aware Drum Transcription Based on Self-Attention Mechanisms Signals journal-article ### Global Structure-Aware Drum Transcription Based on Self-Attention Mechanisms Signals journal-article ### A Real-Time Drum-Wise Volume Visualization System for Learning Volume-Balanced Drum Performance Entertainment Computing, Icec journal-article ### AUTOREGRESSIVE FAST MULTICHANNEL NONNEGATIVE MATRIX FACTORIZATION FOR JOINT BLIND SOURCE SEPARATION AND DEREVERBERATION IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Alpha-Stable Autoregressive Fast Multichannel Nonnegative Matrix Factorization for Joint Speech Enhancement and Dereverberation Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Fast Multichannel Correlated Tensor Factorization for Blind Source Separation European Signal Processing Conference conference-paper ### Gamma Process FastMNMF for Separating an Unknown Number of Sound Sources European Signal Processing Conference conference-paper ### Neural Full-Rank Spatial Covariance Analysis for Blind Source Separation IEEE Signal Processing Letters journal-article ### Neural Full-Rank Spatial Covariance Analysis for Blind Source Separation IEEE Signal Processing Letters journal-article ### PITCH-TIMBRE DISENTANGLEMENT OF MUSICAL INSTRUMENT SOUNDS BASED ON VAE-BASED METRIC LEARNING IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Peak Identification and Quantification by Proteomic Mass Spectrogram Decomposition Journal of Proteome Research journal-article ### Robust Auditory Functions Based on Probabilistic Integration of MUSIC and CGMM IEEE Access journal-article ### Semi-supervised Multichannel Speech Separation Based on a Phone- and Speaker-Aware Deep Generative Model of Speech Spectrograms European Signal Processing Conference conference-paper ### A Flow-Based Deep Latent Variable Model for Speech Spectrogram Modeling and Enhancement other ### A Flow-Based Deep Latent Variable Model for Speech Spectrogram Modeling and Enhancement other ### Statistical learning and estimation of piano fingering Information Sciences journal-article ### A Flow-Based Deep Latent Variable Model for Speech Spectrogram Modeling and Enhancement IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### A Flow-Based Deep Latent Variable Model for Speech Spectrogram Modeling and Enhancement IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Flow-Based Independent Vector Analysis for Blind Source Separation IEEE Signal Processing Letters journal-article ### Adaptive Neural Speech Enhancement with a Denoising Variational Autoencoder Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Bayesian Melody Harmonization Based on a Tree-Structured Generative Model of Chord Sequences and Melodies IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Bayesian Singing Transcription Based on a Hierarchical Generative Model of Keys, Musical Notes, and F0 Trajectories IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Effects of chemotherapy on quality of life and night-time sleep of colon cancer patients Journal of Medical Investigation journal-article ### Fast Multichannel Nonnegative Matrix Factorization With Directivity-Aware Jointly-Diagonalizable Spatial Covariance Matrices for Blind Source Separation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Fast Multichannel Nonnegative Matrix Factorization With Directivity-Aware Jointly-Diagonalizable Spatial Covariance Matrices for Blind Source Separation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Flow-Based Independent Vector Analysis for Blind Source Separation IEEE Signal Processing Letters journal-article ### Self-supervised Neural Audio-Visual Sound Source Localization via Probabilistic Spatial Modeling IEEE International Conference on Intelligent Robots and Systems conference-paper ### Self-supervised Neural Audio-Visual Sound Source Localization via Probabilistic Spatial Modeling ArXiv journal-article ### Semi-Supervised Neural Chord Estimation Based on a Variational Autoencoder With Latent Chord Labels and Features IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Unsupervised Speech Enhancement Based on Multichannel NMF-Informed Beamforming for Noise-Robust Automatic Speech Recognition IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### DEEP BAYESIAN UNSUPERVISED SOURCE SEPARATION BASED ON A COMPLEX GAUSSIAN MIXTURE MODEL IEEE 27TH INTERNATIONAL WORKSHOP ON MACHINE LEARNING FOR SIGNAL PROCESSING journal-article ### Fast Multichannel Source Separation Based on Jointly Diagonalizable Spatial Covariance Matrices European Signal Processing Conference conference-paper ### Fast Multichannel Source Separation Based on Jointly Diagonalizable Spatial Covariance Matrices ArXiv journal-article ### Phase II study of S-1 on alternate days plus bevacizumab in patients aged ≥ 75 years with metastatic colorectal cancer (J-SAVER) International Journal of Clinical Oncology journal-article ### Probabilistic Sequential Patterns for Singing Transcription 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference, APSIPA ASC 2018 - Proceedings conference-paper ### Semi-Supervised Multichannel Speech Enhancement With a Deep Speech Prior IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Semi-Supervised Multichannel Speech Enhancement With a Deep Speech Prior IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Sequential Generation of Singing F0 Contours from Musical Note Sequences Based on WaveNet 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference, APSIPA ASC 2018 - Proceedings conference-paper ### Unsupervised Speech Enhancement Based on Multichannel NMF-Informed Beamforming for Noise-Robust Automatic Speech Recognition ArXiv journal-article ### Unsupervised Speech Enhancement Based on Multichannel NMF-Informed Beamforming for Noise-Robust Automatic Speech Recognition IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Unsupervised Speech Enhancement Based on Multichannel NMF-Informed Beamforming for Noise-Robust Automatic Speech Recognition IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Bayesian Multichannel Speech Enhancement with a Deep Speech Prior 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference, APSIPA ASC 2018 - Proceedings conference-paper ### UNSUPERVISED BEAMFORMING BASED ON MULTICHANNEL NONNEGATIVE MATRIX FACTORIZATION FOR NOISY SPEECH RECOGNITION IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Bayesian Multichannel Audio Source Separation Based on Integrated Source and Spatial Models IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Bayesian Multichannel Audio Source Separation Based on Integrated Source and Spatial Models IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Bayesian Multichannel Audio Source Separation Based on Integrated Source and Spatial Models IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### An end-to-end approach to joint social signal detection and automatic speech recognition ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Bayesian Multichannel Speech Enhancement with a Deep Speech Prior Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA) conference-paper ### Chord-aware automatic music transcription based on hierarchical Bayesian integration of acoustic and language models APSIPA Transactions on Signal and Information Processing journal-article ### Correlated Tensor Factorization for Audio Source Separation ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Engagement recognition from listener’s behaviors in spoken dialogue using a latent character model Transactions of the Japanese Society for Artificial Intelligence journal-article ### Engagement Recognition from Listener’s Behaviors in Spoken Dialogue Using a Latent Character Model Transactions of the Japanese Society for Artificial Intelligence journal-article ### Generative statistical models with self-emergent grammar of chord sequences Journal of New Music Research journal-article ### Independent Low-Rank Tensor Analysis for Audio Source Separation European Signal Processing Conference conference-paper ### Independent low-rank tensor analysis for audio source separation European Signal Processing Conference conference-paper ### Multi-party Interactions by Quizmaster Robot in Speech-Based Jeopardy! Like Games Proceedings - 2017 International Conference on Computational Science and Computational Intelligence, CSCI 2017 conference-paper ### Phase II study of S-1 on alternate days combined with bevacizumab in elderly patients (aged ≥75 years) with metastatic colorectal cancer (mCRC) Journal of Clinical Oncology journal-article ### STATISTICAL SPEECH ENHANCEMENT BASED ON PROBABILISTIC INTEGRATION OF VARIATIONAL AUTOENCODER AND NON-NEGATIVE MATRIX FACTORIZATION IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Statistical Speech Enhancement Based on Probabilistic Integration of Variational Autoencoder and Non-Negative Matrix Factorization ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Speech Enhancement Based on Bayesian Low-Rank and Sparse Decomposition of Multichannel Magnitude Spectrograms IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### Speech Enhancement Based on Bayesian Low-Rank and Sparse Decomposition of Multichannel Magnitude Spectrograms IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Speech Enhancement Based on Bayesian Low-Rank and Sparse Decomposition of Multichannel Magnitude Spectrograms IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Statistical Speech Enhancement Based on Probabilistic Integration of Variational Autoencoder and Non-Negative Matrix Factorization ArXiv journal-article ### Statistical piano reduction controlling performance difficulty APSIPA Transactions on Signal and Information Processing journal-article ### Towards Complete Polyphonic Music Transcription: Integrating Multi-Pitch Detection and Rhythm Quantization ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Unsupervised beamforming based on multichannel nonnegative matrix factorization for noisy speech recognition ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Bayesian multichannel nonnegative matrix factorization for audio source separation and localization ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### BAYESIAN MULTICHANNEL NONNEGATIVE MATRIX FACTORIZATION FOR AUDIO SOURCE SEPARATION AND LOCALIZATION IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Bayesian multichannel nonnegative matrix factorization for audio source separation and localization 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### Combined multi-channel NMF-based robust beamforming for noisy speech recognition Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH conference-paper ### Combined Multi-channel NMF-based Robust Beamforming for Noisy Speech Recognition Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Combined Multi-Channel NMF-Based Robust Beamforming for Noisy Speech Recognition Interspeech 2017 conference-paper ### Design of UAV-Embedded Microphone Array System for Sound Source Localization in Outdoor Environments Sensors journal-article ### Development of Microphone-Array-Embedded UAV for Search and Rescue Task IEEE International Conference on Intelligent Robots and Systems conference-paper ### A diagonal plus low-rank covariance model for computationally efficient source separation IEEE International Workshop on Machine Learning for Signal Processing, MLSP conference-paper ### Infinite probabilistic latent component analysis for audio source separation IEEE International Workshop on Machine Learning for Signal Processing, MLSP conference-paper ### Influence of Different Impulse Response Measurement Signals on MUSIC-Based Sound Source Localization Journal of Robotics and Mechatronics journal-article ### Layout Optimization of Cooperative Distributed Microphone Arrays Based on Estimation of Source Separation Performance Journal of Robotics and Mechatronics journal-article ### Layout optimization of cooperative distributed microphone arrays based on estimation of source separation performance Journal of Robotics and Mechatronics journal-article ### Low Latency and High Quality Two-Stage Human-Voice-Enhancement System for a Hose-Shaped Rescue Robot Journal of Robotics and Mechatronics journal-article ### Low latency and high quality two-stage human-voice-enhancement system for a hose-shaped rescue robot Journal of Robotics and Mechatronics journal-article ### Note Value Recognition for Piano Transcription Using Markov Random Fields IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### Note Value Recognition for Piano Transcription Using Markov Random Fields IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Real-Time Human-Voice Enhancement for a Hose-Shaped Rescue Robot Based on Multi-Channel Low-Rank Sparse Decomposition The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec) journal-article ### Rhythm Transcription of Polyphonic Piano Music Based on Merged-Output HMM for Multiple Voices IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### Rhythm Transcription of Polyphonic Piano Music Based on Merged-Output HMM for Multiple Voices IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### SEMI-BLIND SPEECH ENHANCEMENT BASED ON RECURRENT NEURAL NETWORK FOR SOURCE SEPARATION AND DEREVERBERATION IEEE 27TH INTERNATIONAL WORKSHOP ON MACHINE LEARNING FOR SIGNAL PROCESSING journal-article ### Semi-Blind speech enhancement basedon recurrent neural network for source separation and dereverberation IEEE International Workshop on Machine Learning for Signal Processing, MLSP conference-paper ### Simultaneous identification and localization of still and mobile speakers based on binaural robot audition Journal of Robotics and Mechatronics journal-article ### Sound Source Localization and Separation and Self-Localization Using Asynchronous Distributed Microphone Arrays The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec) journal-article ### Audio-Visual Beat Tracking Based on a State-Space Model for a Robot Dancer Performing with a Human Dancer Journal of Robotics and Mechatronics journal-article ### Audio-Visual Beat Tracking Based on a State-Space Model for a Robot Dancer Performing with a Human Dancer Journal of Robotics and Mechatronics journal-article ### Variational Bayesian Multi-channel Robust NMF for Human-voice Enhancement with a Deformable and Partially-occluded Microphone Array European Signal Processing Conference conference-paper ### Variational Bayesian multi-channel robust NMF for human-voice enhancement with a deformable and partially-occluded microphone array European Signal Processing Conference conference-paper ### Rhythm transcription of MIDI performances based on hierarchical Bayesian modelling of repetition and modification of musical note patterns European Signal Processing Conference conference-paper ### Songle Widget: A web-based development framework for making animation and physical devices synchronized with music The Journal of the Acoustical Society of America journal-article ### Musical Similarity and Commonness Estimation Based on Probabilistic Generative Models of Musical Elements International Journal of Semantic Computing journal-article ### 3D Posture Estimation for a Hose-shaped Rescue Robot using a Microphone and Accelerometer Array The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec) journal-article ### A unified Bayesian model of time-frequency clustering and low-rank approximation for multi-channel source separation European Signal Processing Conference conference-paper ### A Unified Bayesian Model of Time-frequency Clustering and Low-rank Approximation for Multi-channel Source Separation European Signal Processing Conference conference-paper ### A unified Bayesian model of time-frequency clustering and low-rank approximation for multi-channel source separation 2016 24th European Signal Processing Conference (EUSIPCO) conference-paper ### Online Localization of Multiple Sound Sources and Multiple Robots with Asynchronous Microphone Arrays The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec) journal-article ### Online Simultaneous Localization and Mapping of Multiple Sound Sources and Asynchronous Microphone Arrays IEEE International Conference on Intelligent Robots and Systems conference-paper ### Online simultaneous localization and mapping of multiple sound sources and asynchronous microphone arrays IEEE International Conference on Intelligent Robots and Systems conference-paper ### Propensity score-matched study of laparoscopic and open surgery for colorectal cancer in rural hospitals Journal of Gastroenterology and Hepatology journal-article ### Rhythm transcription of MIDI performances based on hierarchical Bayesian modelling of repetition and modification of musical note patterns 2016 24th European Signal Processing Conference (EUSIPCO) conference-paper ### STUDENT'S T MULTICHANNEL NONNEGATIVE MATRIX FACTORIZATION FOR BLIND SOURCE SEPARATION IEEE International Workshop on Acoustic Signal Enhancement (IWAENC) conference-paper ### Singing Voice Separation and Vocal F0 Estimation Based on Mutual Combination of Robust Principal Component Analysis and Subharmonic Summation IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### Singing Voice Separation and Vocal F0 Estimation Based on Mutual Combination of Robust Principal Component Analysis and Subharmonic Summation IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Songle Widget: Making Animation and Physical Devices Synchronized with Music Videos on the Web Proceedings - 2015 IEEE International Symposium on Multimedia, ISM 2015 conference-paper ### Sound-based online localization for an in-pipe snake robot SSRR 2016 - International Symposium on Safety, Security and Rescue Robotics conference-paper ### Sound-based Online Localization for an In-pipe Snake Robot IEEE International Symposium on Safety, Security, and Rescue Robotics (SSRR) conference-paper ### Sound-based online localization for an in-pipe snake robot 2016 IEEE International Symposium on Safety, Security, and Rescue Robotics (SSRR) conference-paper ### Sparse learning for music signal analysis Journal of the Institute of Electronics, Information and Communication Engineers journal-article ### Student's T nonnegative matrix factorization and positive semidefinite tensor factorization for single-channel audio source separation ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Student's T nonnegative matrix factorization and positive semidefinite tensor factorization for single-channel audio source separation 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### Student's t multichannel nonnegative matrix factorization for blind source separation 2016 International Workshop on Acoustic Signal Enhancement, IWAENC 2016 conference-paper ### Student's t multichannel nonnegative matrix factorization for blind source separation 2016 IEEE International Workshop on Acoustic Signal Enhancement (IWAENC) conference-paper ### Tree-structured probabilistic model of monophonic written music based on the generative theory of tonal music ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Tree-structured probabilistic model of monophonic written music based on the generative theory of tonal music 2016 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### Identification and localization of one or two concurrent speakers in a binaural robotic context Proceedings - 2015 IEEE International Conference on Systems, Man, and Cybernetics, SMC 2015 conference-paper ### Human-voice enhancement based on online RPCA for a hose-shaped rescue robot with a microphone array SSRR 2015 - 2015 IEEE International Symposium on Safety, Security, and Rescue Robotics conference-paper ### Musical Similarity and Commonness Estimation Based on Probabilistic Generative Models Proceedings - 2015 IEEE International Symposium on Multimedia, ISM 2015 conference-paper ### A feedback framework for improved chord recognition based on NMF-based approximate note transcription ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### A feedback framework for improved chord recognition based on NMF-based approximate note transcription 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### A music performance assistance system based on vocal, harmonic, and percussive source separation and content visualization for music audio signals Proceedings of the 12th International Conference in Sound and Music Computing, SMC 2015 conference-paper ### A score-informed piano tutoring system with mistake detection and score simplification Proceedings of the 12th International Conference in Sound and Music Computing, SMC 2015 conference-paper ### Audio-Visual Beat Tracking Based on a State-Space Model for a Music Robot Dancing with Humans IEEE International Conference on Intelligent Robots and Systems conference-paper ### Audio-visual beat tracking based on a state-space model for a music robot dancing with humans IEEE International Conference on Intelligent Robots and Systems conference-paper ### Audio-visual beat tracking based on a state-space model for a music robot dancing with humans 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) conference-paper ### Automatic singing voice to music video generation via mashup of singing video clips Proceedings of the 12th International Conference in Sound and Music Computing, SMC 2015 conference-paper ### Bayesian Integration of Sound Source Separation and Speech Recognition: A New Approach to Simultaneous Speech Recognition Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH journal-article ### Bayesian integration of sound source separation and speech recognition: A new approach to simultaneous speech recognition Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH conference-paper ### CHALLENGES IN DEPLOYING A MICROPHONE ARRAY TO LOCALIZE AND SEPARATE SOUND SOURCES IN REAL AUDITORY SCENES IEEE International Conference on Acoustics, Speech, and Signal Processing conference-paper ### Challenges in deploying a microphone array to localize and separate sound sources in real auditory scenes ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Human-Voice Enhancement based on Online RPCA for a Hose-shaped Rescue Robot with a Microphone Array IEEE International Symposium on Safety, Security, and Rescue Robotics (SSRR) conference-paper ### Human-voice enhancement based on online RPCA for a hose-shaped rescue robot with a microphone array 2015 IEEE International Symposium on Safety, Security, and Rescue Robotics (SSRR) conference-paper ### Recognition of in-field frog chorusing using Bayesian nonparametric microphone array processing AAAI Workshop - Technical Report conference-paper ### Singing voice analysis and editing based on mutually dependent F0 estimation and source separation ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Singing voice analysis and editing based on mutually dependent F0 estimation and source separation 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### Unified inter- and intra-recording duration model for multiple music audio alignment 2015 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics, WASPAA 2015 conference-paper ### Unified inter- and intra-recording duration model for multiple music audio alignment 2015 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) conference-paper ### A robot quizmaster that can localize, separate, and recognize simultaneous utterances for a fastest-voice-first quiz game IEEE-RAS International Conference on Humanoid Robots conference-paper ### Identification and Localization of One or Two Concurrent Speakers in a Binaural Robotic Context 2015 IEEE International Conference on Systems, Man, and Cybernetics conference-paper ### Toward a quizmaster robot for speech-based multiparty interaction Advanced Robotics journal-article ### Microphone-accelerometer based 3D posture estimation for a hose-shaped rescue robot IEEE International Conference on Intelligent Robots and Systems conference-paper ### Microphone-Accelerometer Based 3D Posture Estimation for a Hose-shaped Rescue Robot IEEE International Conference on Intelligent Robots and Systems conference-paper ### Microphone-accelerometer based 3D posture estimation for a hose-shaped rescue robot 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) conference-paper ### Musical Similarity and Commonness Estimation Based on Probabilistic Generative Models 2015 IEEE International Symposium on Multimedia (ISM) conference-paper ### Optimizing the Layout of Multiple Mobile Robots for Cooperative Sound Source Separation IEEE International Conference on Intelligent Robots and Systems conference-paper ### Optimizing the layout of multiple mobile robots for cooperative sound source separation IEEE International Conference on Intelligent Robots and Systems conference-paper ### Optimizing the layout of multiple mobile robots for cooperative sound source separation 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) conference-paper ### Posture estimation of hose-shaped robot by using active microphone array Advanced Robotics journal-article ### Development of a robot quizmaster with auditory functions for speech-based multiparty interaction 2014 IEEE/SICE International Symposium on System Integration, SII 2014 conference-paper ### Cultivating vocal activity detection for music audio signals in a circulation-type crowdsourcing ecosystem ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### A robot quizmaster that can localize, separate, and recognize simultaneous utterances for a fastest-voice-first quiz game 2014 IEEE-RAS International Conference on Humanoid Robots conference-paper ### Cultivating vocal activity detection for music audio signals in a circulation-type crowdsourcing ecosystem 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### A sound-based online method for estimating the time-varying posture of a hose-shaped robot 12th IEEE International Symposium on Safety, Security and Rescue Robotics, SSRR 2014 - Symposium Proceedings conference-paper ### Development of a robot quizmaster with auditory functions for speech-based multiparty interaction 2014 IEEE/SICE International Symposium on System Integration conference-paper ### AutoMashUpper: Automatic creation of multi-song music mashups IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### AutoMashUpper: Automatic Creation of Multi-Song Music Mashups IEEE/ACM Transactions on Audio, Speech, and Language Processing journal-article ### Nonparametric bayesian dereverberation of power spectrograms based on infinite-order autoregressive processes IEEE/ACM Transactions on Audio Speech and Language Processing journal-article ### Timbre replacement of harmonic and drum components for music audio signals ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Timbre replacement of harmonic and drum components for music audio signals 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### Vocal timbre analysis using latent Dirichlet allocation and cross-gender vocal timbre similarity ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Vocal timbre analysis using latent Dirichlet allocation and cross-gender vocal timbre similarity 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### A nested infinite Gaussian mixture model for identifying known and unknown audio events International Workshop on Image Analysis for Multimedia Interactive Services conference-paper ### A nested infinite Gaussian mixture model for identifying known and unknown audio events 2013 14th International Workshop on Image Analysis for Multimedia Interactive Services (WIAMIS) conference-paper ### Infinite kernel linear prediction for joint estimation of spectral envelope and fundamental frequency ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Infinite kernel linear prediction for joint estimation of spectral envelope and fundamental frequency 2013 IEEE International Conference on Acoustics, Speech and Signal Processing conference-paper ### Infinite positive semidefinite tensor factorization for source separation of mixture signals 30th International Conference on Machine Learning, ICML 2013 conference-paper ### Nested iGMM recognition and multiple hypothesis tracking of moving sound sources for mobile robot audition IEEE International Conference on Intelligent Robots and Systems conference-paper ### Posture Estimation of Hose-Shaped Robot using Microphone Array Localization IEEE International Conference on Intelligent Robots and Systems conference-paper ### A Nonparametric Bayesian Multipitch Analyzer Based on Infinite Latent Harmonic Allocation IEEE Transactions on Audio, Speech and Language Processing journal-article ### A Nonparametric Bayesian Multipitch Analyzer Based on Infinite Latent Harmonic Allocation IEEE Transactions on Audio, Speech and Language Processing journal-article ### Infinite composite autoregressive models for music signal analysis Proceedings of the 13th International Society for Music Information Retrieval Conference, ISMIR 2012 conference-paper ### PodCastle and Songle: Crowdsourcing-based web services for spoken document retrieval and active music listening 2012 Information Theory and Applications Workshop, ITA 2012 - Conference Proceedings conference-paper ### PodCastle and songle: Crowdsourcing-based web services for spoken document retrieval and active music listening 2012 Information Theory and Applications Workshop conference-paper ### PodCastle and songle: Crowdsourcing-based web services for retrieval and browsing of speech and music content CEUR Workshop Proceedings conference-paper ### PodCastle and songle: Crowdsourcing-based web services for spoken content retrieval and active music listening CrowdMM 2012 - Proceedings of the 2012 ACM Workshop on Crowdsourcing for Multimedia, Co-located with ACM Multimedia 2012 conference-paper ### Unsupervised music understanding based on nonparametric Bayesian models ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Unsupervised music understanding based on nonparametric Bayesian models 2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) conference-paper ### A vocabulary-free infinity-gram model for nonparametric bayesian chord progression analysis Proceedings of the 12th International Society for Music Information Retrieval Conference, ISMIR 2011 conference-paper ### Songle: A web service for active music listening improved by user contributions Proceedings of the 12th International Society for Music Information Retrieval Conference, ISMIR 2011 conference-paper ### Timbre and melody features for the recognition of vocal activity and instrumental solos in polyphonic music Proceedings of the 12th International Society for Music Information Retrieval Conference, ISMIR 2011 conference-paper ### Infinite latent harmonic allocation: A nonparametric bayesian approach to multipitch analysis Proceedings of the 11th International Society for Music Information Retrieval Conference, ISMIR 2010 conference-paper ### Continuous PLSI and smoothing techniques for hybrid music recommendation Proceedings of the 10th International Society for Music Information Retrieval Conference, ISMIR 2009 conference-paper ### MusicCommentator: Generating Comments Synchronized with Musical Audio Signals by a Joint Probabilistic Model of Acoustic and Textual Features Entertainment Computing – ICEC 2009 other ### A robot listens to music and counts its beats aloud by separating music from counting voice 2008 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS conference-paper ### A robot singer with music recognition based on real-time beat tracking ISMIR 2008 - 9th International Conference on Music Information Retrieval conference-paper ### A robot uses its own microphone to synchronize its steps to musical beats while scatting and singing 2008 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS conference-paper ### An efficient hybrid music recommender system using an incrementally trainable probabilistic generative model IEEE Transactions on Audio, Speech and Language Processing journal-article ### An efficient hybrid music recommender system using an incrementally trainable probabilistic generative model IEEE Transactions on Audio, Speech and Language Processing journal-article ### An Efficient Hybrid Music Recommender System Using an Incrementally Trainable Probabilistic Generative Model IEEE Transactions on Audio, Speech, and Language Processing journal-article ### Analysis-and-manipulation approach to pitch and duration of musical instrument sounds without distorting timbral characteristics Proceedings - 11th International Conference on Digital Audio Effects, DAFx 2008 conference-paper ### Automatic chord recognition based on probabilistic integration of chord transition and bass pitch estimation ISMIR 2008 - 9th International Conference on Music Information Retrieval conference-paper ### Music thumbnailer: Visualizing musical pieces in thumbnail images based on acoustic features ISMIR 2008 - 9th International Conference on Music Information Retrieval conference-paper ### A biped robot that keeps steps in time with musical beats while listening to music with its own ears IEEE International Conference on Intelligent Robots and Systems conference-paper ### A biped robot that keeps steps in time with musical beats while listening to music with its own ears 2007 IEEE/RSJ International Conference on Intelligent Robots and Systems conference-paper ### Drum sound recognition for polyphonic audio signals by adaptation and matching of spectrogram templates with harmonic structure suppression IEEE Transactions on Audio, Speech and Language Processing journal-article ### Drum sound recognition for polyphonic audio signals by adaptation and matching of spectrogram templates with harmonic structure suppression IEEE Transactions on Audio, Speech and Language Processing journal-article ### Drumix: An Audio Player with Real-time Drum-part Rearrangement Functions for Active Music Listening IPSJ Digital Courier journal-article ### Improving efficiency and scalability of model-based music recommender system based on incremental training Proceedings of the 8th International Conference on Music Information Retrieval, ISMIR 2007 conference-paper ### An error correction framework based on drum pattern periodicity for improving drum sound detection ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings conference-paper ### Hybrid collaborative and content-based music recommendation using probabilistic model with latent user preferences ISMIR 2006 - 7th International Conference on Music Information Retrieval conference-paper ### INTER:D: A drum sound equalizer for controlling volume and timbre of drums IET Seminar Digest conference-paper ## Source Read this profile on Hello.cv: https://hello.cv/kazuyoshiyoshii Create your free profile at https://hello.cv