IEEE Transactions on Audio, Speech and Language Processing

IEEE Transactions on Audio, Speech and Language Processing

T. D. Abhayapala and A. Gupta, ?Spherical Harmonic Analysis of Wavefields Using,? Audio, Speech, and Language Processing, IEEE Transactions on, vol. 18, no ...

[View/Download]




 Investigation of IMU&Elevoc Submission for the Short-Duration ...

Investigation of IMU&Elevoc Submission for the Short-Duration ...

T. D. Abhayapala and A. Gupta, ?Spherical Harmonic Analysis of Wavefields Using,? Audio, Speech, and Language Processing, IEEE Transactions on, vol. 18, no ...

[View/Download]




 lter representation of speech with a VAE - Research web sites - Inria

lter representation of speech with a VAE - Research web sites - Inria

Because speech with con- strained lexical content is harder to collect, often TD models are fine-tuned from a TI model using a small target phrase dataset.

[View/Download]




 Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

This index covers all technical items?papers, correspondence, reviews, etc.?that appeared in this periodical during 2021, and items from previous.

[View/Download]




 Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

This index covers all technical items?papers, correspondence, reviews, etc.?that appeared in this periodical during 2021, and items from previous.

[View/Download]




 Multi-channel audio source separation using multiple ... - Hal-Inria

Multi-channel audio source separation using multiple ... - Hal-Inria

We find that the ambient reverberation and noise distort the trigger and break its connection to the implanted backdoor, thus making digital attacks fail.

[View/Download]




 Dynamic Stream Weighting for Turbo-Decoding-Based Audiovisual ...

Dynamic Stream Weighting for Turbo-Decoding-Based Audiovisual ...

Abstract?This paper addresses the problem of speech separa- tion and enhancement from multichannel convolutive and noisy.

[View/Download]




 On the influence of transfer function noise on sound zone control in ...

On the influence of transfer function noise on sound zone control in ...

Févotte et al., Sparse linear regression with structured priors and application to denoising of musical audio, IEEE TASLP, 2007. ... TD-PSOLA (Moulines and ...

[View/Download]




 Chargé de recherche CRCN - Marco Dinarelli

Chargé de recherche CRCN - Marco Dinarelli

(1 C ) ' soit elli p tique sur . Il existe alors deux f onctions ? et ... JL KMONQPSRUTWV X YSV Z P R\[ Il su ffi t d 'a pp liquer le corollaire III .

[View/Download]




 Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

Time-Contrastive Learning Based Deep Bottleneck Features for Text ...

Abstract?This paper addresses the problem of audio source recovery from multichannel noisy convolutive mixture for source separation and speech enhancement, ...

[View/Download]




 Far-Field End-to-End Text-Dependent Speaker Veri?cation based ...

Far-Field End-to-End Text-Dependent Speaker Veri?cation based ...

In this paper, we propose the time-domain real-valued gener- alized Wiener filter (TD-GWF) as an alternative to frequency- domain conventional beamformers for ...

[View/Download]




 Deep Sentence Embedding Using Long Short-Term Memory Networks

Deep Sentence Embedding Using Long Short-Term Memory Networks

More appropriate models that use deformations (Tf or Td) have been tried with promising results. Better overall results. (10.64 dB) are obtained ...

[View/Download]




 Late-Reverberation Synthesis Using Interleaved Velvet-Noise ...

Late-Reverberation Synthesis Using Interleaved Velvet-Noise ...

We propose to use the stream weights of audio and video streams to maximize a discriminative cost function in each time frame and TD- iteration.

[View/Download]




 A Complete Bibliography of IEEE/ACM Transactions on Audio ...

A Complete Bibliography of IEEE/ACM Transactions on Audio ...

The author and publisher of this book have used their best efforts in preparing this book. These efforts include the development, research, and testing of ...

[View/Download]




 Direct Generation of Speech from Facial Electromyographic Signals

Direct Generation of Speech from Facial Electromyographic Signals

Abstract?Sound zones are valuable in scenarios where mul- tiple people are present in the same room but want to listen to individual audio content without ...

[View/Download]