Výsledky vyhledávání - "Goldwater, Sharon"

Report

Orthogonality and isotropy of speaker and phonetic information in self-supervised speech representations

Autor: Mohamed, Mukhtar, Liu, Oli Danyi, Tang, Hao, Goldwater, Sharon

Self-supervised speech representations can hugely benefit downstream speech technologies, yet the properties that make them useful are still poorly understood. Two candidate properties related to the geometry of the representation space have been hyp

Externí odkaz: http://arxiv.org/abs/2406.09200

Zobrazit plný text záznamu

Report

Estimating the Level of Dialectness Predicts Interannotator Agreement in Multi-dialect Arabic Datasets

Autor: Keleg, Amr, Magdy, Walid, Goldwater, Sharon

On annotating multi-dialect Arabic datasets, it is common to randomly assign the samples across a pool of native Arabic speakers. Recent analyses recommended routing dialectal samples to native speakers of their respective dialects to build higher-qu

Externí odkaz: http://arxiv.org/abs/2405.11282

Zobrazit plný text záznamu

Report

A predictive learning model can simulate temporal dynamics and context effects found in neural representations of continuous speech

Autor: Liu, Oli Danyi, Tang, Hao, Feldman, Naomi, Goldwater, Sharon

Speech perception involves storing and integrating sequentially presented items. Recent work in cognitive neuroscience has identified temporal and contextual characteristics in humans' neural encoding of speech that may facilitate this temporal proce

Externí odkaz: http://arxiv.org/abs/2405.08237

Zobrazit plný text záznamu

Report

ALDi: Quantifying the Arabic Level of Dialectness of Text

Autor: Keleg, Amr, Goldwater, Sharon, Magdy, Walid

Transcribed speech and user-generated text in Arabic typically contain a mixture of Modern Standard Arabic (MSA), the standardized language taught in schools, and Dialectal Arabic (DA), used in daily communications. To handle this variation, previous

Externí odkaz: http://arxiv.org/abs/2310.13747

Zobrazit plný text záznamu

Report

Acoustic Word Embeddings for Untranscribed Target Languages with Continued Pretraining and Learned Pooling

Autor: Sanabria, Ramon, Klejch, Ondrej, Tang, Hao, Goldwater, Sharon

Acoustic word embeddings are typically created by training a pooling function using pairs of word-like units. For unsupervised systems, these are mined using k-nearest neighbor (KNN) search, which is slow. Recently, mean-pooled representations from a

Externí odkaz: http://arxiv.org/abs/2306.02153

Zobrazit plný text záznamu

Report

Self-supervised Predictive Coding Models Encode Speaker and Phonetic Information in Orthogonal Subspaces

Autor: Liu, Oli, Tang, Hao, Goldwater, Sharon

Self-supervised speech representations are known to encode both speaker and phonetic information, but how they are distributed in the high-dimensional space remains largely unexplored. We hypothesize that they are encoded in orthogonal subspaces, a p

Externí odkaz: http://arxiv.org/abs/2305.12464

Zobrazit plný text záznamu

Report

Prosodic features improve sentence segmentation and parsing

Autor: Nielsen, Elizabeth, Goldwater, Sharon, Steedman, Mark

Parsing spoken dialogue presents challenges that parsing text does not, including a lack of clear sentence boundaries. We know from previous work that prosody helps in parsing single sentences (Tran et al. 2018), but we want to show the effect of pro

Externí odkaz: http://arxiv.org/abs/2302.12165

Zobrazit plný text záznamu

Report

Analyzing Acoustic Word Embeddings from Pre-trained Self-supervised Speech Models

Autor: Sanabria, Ramon, Tang, Hao, Goldwater, Sharon

Given the strong results of self-supervised models on various tasks, there have been surprisingly few studies exploring self-supervised representations for acoustic word embeddings (AWE), fixed-dimensional vectors representing variable-length spoken

Externí odkaz: http://arxiv.org/abs/2210.16043

Zobrazit plný text záznamu

Report

Cross-linguistically Consistent Semantic and Syntactic Annotation of Child-directed Speech

Autor: Szubert, Ida, Abend, Omri, Schneider, Nathan, Gibbon, Samuel, Mahon, Louis, Goldwater, Sharon, Steedman, Mark

This paper proposes a methodology for constructing such corpora of child directed speech (CDS) paired with sentential logical forms, and uses this method to create two such corpora, in English and Hebrew. The approach enforces a cross-linguistically

Externí odkaz: http://arxiv.org/abs/2109.10952

Zobrazit plný text záznamu

Report

On the Difficulty of Segmenting Words with Attention

Autor: Sanabria, Ramon, Tang, Hao, Goldwater, Sharon

Word segmentation, the problem of finding word boundaries in speech, is of interest for a range of tasks. Previous papers have suggested that for sequence-to-sequence models trained on tasks such as speech translation or speech recognition, attention

Externí odkaz: http://arxiv.org/abs/2109.10107

Zobrazit plný text záznamu

Vyhledávací nástroje:

Upřesnit hledání