STOCHASTIC MARKOV RECURRENT NEURAL NETWORK FOR SOURCE SEPARATION

Full metadata record

DC Field	Value	Language
dc.contributor.author	Chien, Jen-Tzung	en_US
dc.contributor.author	Kuo, Che-Yu	en_US
dc.date.accessioned	2019-10-05T00:09:45Z	-
dc.date.available	2019-10-05T00:09:45Z	-
dc.date.issued	2019-01-01	en_US
dc.identifier.isbn	978-1-4799-8131-1	en_US
dc.identifier.issn	1520-6149	en_US
dc.identifier.uri	http://hdl.handle.net/11536/152935	-
dc.description.abstract	Monaural source separation based on recurrent neural network is learned to characterize the sequential patterns in source signals based on dynamic states which are propagated through time. The hidden states are assumed to be deterministic along a single path where a shared long short-term memory (LSTM) is used. Such assumptions may not faithfully reflect the randomness and the variety of temporal features in mixed signals. To strengthen the capability of LSTM in source separation, we propose a stochastic Markov LSTM where the regression from the mixed signal to its source signals is learned with a stochastic indicator of Markov state which selects the state-dependent LSTM for signal separation at each time. A set of LSTMs is discovered to capture the structural diversity of temporal signals or the stochastic trajectory of state transitions for sequential prediction. A new state machine is constructed to learn the complicated latent semantics in heterogeneous and structural mappings between mixed signals and source signals. The Gumbel-softmax sampling is implemented in the backpropagation algorithm with discrete Markov states. Experiments on speech enhancement illustrate the merit of the proposed stochastic Markov LSTM in terms of short-term objective intelligibility measure of the separated speech.	en_US
dc.language.iso	en_US	en_US
dc.subject	Source separation	en_US
dc.subject	deep sequential learning	en_US
dc.subject	stochastic transition	en_US
dc.subject	Markov state	en_US
dc.subject	latent variable model	en_US
dc.title	STOCHASTIC MARKOV RECURRENT NEURAL NETWORK FOR SOURCE SEPARATION	en_US
dc.type	Proceedings Paper	en_US
dc.identifier.journal	2019 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH AND SIGNAL PROCESSING (ICASSP)	en_US
dc.citation.spage	8072	en_US
dc.citation.epage	8076	en_US
dc.contributor.department	電機工程學系	zh_TW
dc.contributor.department	Department of Electrical and Computer Engineering	en_US
dc.identifier.wosnumber	WOS:000482554008062	en_US
dc.citation.woscount	0	en_US
Appears in Collections:	Conferences Paper

APA	Chien, J., & Kuo, C. (2019). STOCHASTIC MARKOV RECURRENT NEURAL NETWORK FOR SOURCE SEPARATION. WOS:000482554008062.
Bibtex	@article{Chien2019STOCHASTIC, title={STOCHASTIC MARKOV RECURRENT NEURAL NETWORK FOR SOURCE SEPARATION}, author={Chien, Jen-Tzung and Kuo, Che-Yu}, journal={WOS:000482554008062}, year={2019}, url={https://ir.lib.nycu.edu.tw/handle/11536/152935?mode=full}, }