Back to search

Article

CiwGAN and fiwGAN: Encoding information in acoustic data to model lexical learning with Generative Adversarial Networks

2020-04-21

Abstract excerpt

<p>How can deep neural networks encode information that corresponds to words in human speech into raw acoustic data? This paper proposes two neural network architectures for modeling unsupervised lexical learning from raw acoustic inputs: ciwGAN (Categorical InfoWaveGAN) and fiwGAN (Featural InfoWaveGAN). These combine Deep Convolutional GAN architecture for audio data (WaveGAN; Donahue et al., 2019) with the info...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
30719058-ad12-5a5d-be08-c8bae202b37b
DOI
10.31234/osf.io/mwb5u
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
CiwGAN and fiwGAN: Encoding information in acoustic data to model lexical learning with Generative Adversarial NetworksDOI 10.31234/osf.io/mwb5u
Select a neighboring publication to make it the new centre.