Library

Subject
Tags

2,287 matches · Phonetics and Phonology Research

#
001LiveLingo Voice Translation Benchmarks 2026OpenAlexPaperRon Villomo6 days ago
002StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech SynthesisOpenAlexPaperKaicheng Luo, X Z Gong et al.Jul 22
003WDYW.01: Tone Restoration for Standard YorùbáOpenAlexPaperTosin Sina AkereleJul 20
004Roxi-Duplex: Low-Resource Indian-English Adaptation of a Full-Duplex Speech-to-Speech Model with Synthetic Two-Channel Training DataOpenAlexPaperJ A, J AJul 19
005Roxi-Duplex: Low-Resource Indian-English Adaptation of a Full-Duplex Speech-to-Speech Model with Synthetic Two-Channel Training DataOpenAlexPaperJ A, J AJul 19
006SLT 2026 REAL-TSE Challenge: Real-world Target Speaker Extraction from Conversational RecordingsOpenAlexPaperShuai Wang, Zihan Qian et al.Jul 16
007DIVERSE: a corpus of cisgender female speakers for phonetic and forensic researchOpenAlexPaperAlice Paver, Chloe Patman et al.Jul 15
008Malayalam SK-ARJOpenAlexPaperSomdev Kar, ashitha rachel jacobJul 15
009Rethinking Speech Foundation Model Fine-tuning: Better SFT or Better Match?OpenAlexPaperWangjin Zhou, Yizhou Zhang et al.Jul 15
010The role of temporal information in automatic speech and voice recognitionOpenAlexPaperErdem Baha Topbas, Srikanth Madikeri et al.Jul 15
011Towards explainable linguistic phonetic methods to authenticate allegedly deepfake audioOpenAlexPaperBen Gibb-Reid, Jessica Wormald et al.Jul 15
012Synchronized Three-Dimensional Vocal-Tract Motion for Speech Synchronization via Joint-Embedding Predictive Architecture AlignmentOpenAlexPaperSheng Li, Takahiro ShinozakiJul 13
013An Objective Intelligibility Metric Evaluation on Spanish SpeechOpenAlexPaperIván López-Espejo, Jesper JensenJul 12
014ICPHS 2027OpenAlexPaperNeda Mousavi, Pauline Larrouy-MaestriJul 11
015Silent speech with ultrasoundOpenAlexPaperVadims Casecnikovs, Gimran Abdullin et al.Jul 7
016Silent speech with ultrasoundOpenAlexPaperVadims Casecnikovs, Gimran Abdullin et al.Jul 7
017The Receiver-Limited Floor: Rate-Distortion Bounds on Serial Decoding ThroughputOpenAlexPaperGrant Lavell Whitmer IIIJul 7
018Enhancing EFL Learners’ Pronunciation: A GenAI-Powered Approach to Weak Forms for Students at Arab Open UniversityOpenAlexPaperShaimaa Mohamed Helal, Hassan Saleh Mahdi et al.Jul 5
019Measuring Word Length in Japanese: Mora or Syllable?OpenAlexPaperWenchao LiJul 4
020QuaSR: Quality-Aware Sample Reweighting for Pacific Indigenous Speech RecognitionOpenAlexPaperYishun Li, Y H Xiao et al.Jul 4
021Protocol for a Systematic Literature Review: Visual Speech Recognition for Bengali (Bangla) and AssameseOpenAlexPaperAbdul Hasib Uddin, Abu Shamim Mohammad ArifJul 3
022The Receiver-Limited Floor: Rate-Distortion Bounds on Serial Decoding ThroughputOpenAlexPaperGrant Lavell Whitmer IIIJul 3
023CALLHOME American English Second EditionOpenAlexPaperAlexandra Canavan, David Graff et al.Jul 1
024Speaker-Conditioned Neural Speech Synthesis for Marathi and English Using Guided Attention on the vVISWa DatasetOpenAlexPaperDiksha R. Pawar and Dr. Pravin L. YannawarJul 1
025From Acoustic Diagnosis to Generative Interaction: An Integrated Model of Digital Tools in College English Speaking InstructionOpenAlexPaperYijing ZhangJun 30
026An Italian corpus of clinical speech dataOpenAlexPaperLaura TagliaferroJun 29
027PIAS: Transparent Multimodal Assessment of Presentation DeliveryOpenAlexPaperNina Hosseini-Kivanani, Nafiseh Taghva et al.Jun 29
028PIAS: Transparent Multimodal Assessment of Presentation DeliveryOpenAlexPaperNina Hosseini-Kivanani, Nafiseh Taghva et al.Jun 29
029Preserving Speech-to-Text LLM Capabilities in Speech-to-Speech GenerationOpenAlexPaperYuxuan Hu, Hüseyin Ozan Ҫirkinoğlu et al.Jun 29
030The Multilingual Evaluation Paradox: Ramoju Potential Across Languages, Scripts, and 2.37 Billion SpeakersOpenAlexPaperSurya Sai Brahmendra Ramoju, Sri Likhitha Anuganti et al.Jun 29
031The Multilingual Evaluation Paradox: Ramoju Potential Across Languages, Scripts, and 2.37 Billion SpeakersOpenAlexPaperSurya Sai Brahmendra Ramoju, Sri Likhitha Anuganti et al.Jun 29
032Learning Latent Representations with Progressive Hypothesis Space ExpansionOpenAlexPaperJonathan Charles ParamoreJun 26
033Enhancing Zero-shot Whisper ASR for Under-resourced Javanese and Sundanese via Generative Fusion Decoding and Logit CalibrationOpenAlexPaperAgung Santosa, Asril Jarin et al.Jun 24
034Fine-Tuning VGG19 with Mel Spectrograms for Amazigh Spoken Digit RecognitionOpenAlexPaperHossam Boulal, Abdelkader Benzirar et al.Jun 24
035Acoustic Landmark Detector based on Conformer and HuBERTOpenAlexPaperMateo Cámara, José Luis Blanco et al.Jun 22
036Enhancing Arabic Diacritization Using BERT with BiLSTM and BiGRU HeadsOpenAlexPaperAhmed Abdelhamed Elgayar, Abdul-Hadi Nabi Ahmed et al.Jun 22
037Enhancing Hausa Words Lemmatization Through Feature EngineeringOpenAlexPaperAbba Bello Muhammad, Rasheed Abubakar Rasheed et al.Jun 22
038ProsoCodec: Prosody-Oriented Speech Codec for Voice ConversionOpenAlexPaperJ H Choi, Ji-Hoon Kim et al.Jun 20
039Using Phonological-Level Wav2Vec2 for Mandarin Automatic Mispronunciation Detection and DiagnosisOpenAlexPaperJinghao Chen, Mostafa Shahin et al.Jun 20
040COMPUTATIONAL MODEL OF KAZAKH VOWEL–CONSONANT HARMONYOpenAlexPaperZhansaya Segizbayeva, Marek MiłoszJun 19
041DisSpeech: Low-Resource Controllable Mandarin Stuttered Speech Synthesis for ASR AugmentationOpenAlexPaperYao LuJun 19
042Standard and Dagestani Russian in automatic speech recognition: weaknesses of ASR systemsOpenAlexPaperA. Katsnelson, O. LyashevskayaJun 19
043To Shorten or to Lengthen? Output Strategies of ASR Models in Atypical Speech RecognitionOpenAlexPaperAnastasia Kolmogorova, Ekaterina Yavshits et al.Jun 19
044Exploring Pre-training Benefits on Phoneme Addition through Fine-tuning in Speech SynthesisOpenAlexPaperMasato Murata, Koichi Miyazaki et al.Jun 18
045Interpreting Content and Speaker Characteristics in Factorised Self-Supervised SubspacesOpenAlexPaperKyle Janse van Rensburg, Herman KamperJun 18
046Language Geometry: A Unified Geometric Framework for Understanding Large Language ModelsOpenAlexPaperXiaobo LiJun 18
047Quantifying Punctuation Patterns in Chinese Language for Language Service ApplicationsOpenAlexPaperJarosław Kwapień, Jakub Dec et al.Jun 18
048Self-Supervised and Semi-Supervised Learning for Nepali ASR with Limited Labeled DataOpenAlexPaperRajesh Raskoti, Kobid KarkeeJun 18
049Transcript-Free Flow-Matching Text-to-Speech via Speech Feature ConditioningOpenAlexPaperSooHwan Eom, Hee Suk Yoon et al.Jun 18
050ОЦЕНКА ПРОСОДИЧЕСКОЙ ИНФОРМАЦИИ В САМООБУЧАЮЩИХСЯ ПРЕДСТАВЛЕНИЯХ РЕЧИ: СРАВНИТЕЛЬНЫЙ АНАЛИЗ HUBERT И WAVLMOpenAlexPaperА С ЛаринJun 18
051ОЦЕНКА ПРОСОДИЧЕСКОЙ ИНФОРМАЦИИ В САМООБУЧАЮЩИХСЯ ПРЕДСТАВЛЕНИЯХ РЕЧИ: СРАВНИТЕЛЬНЫЙ АНАЛИЗ HUBERT И WAVLMOpenAlexPaperА С ЛаринJun 18
052Responsible ASR: Overcoming Challenges of Foundational Models in Narrow-Band and Low-Resource SettingsOpenAlexPaperTejas Godambe, Nutan Choudhary et al.Jun 17
053Zero-Shot Performance of Compact Whisper Variants on Urdu-English Code-Mixed Speech: An Empirical Investigation Using WER and Switch Point AnalysisOpenAlexPaperAbdullah Haroon Mohammed haroonJun 17
054Zero-Shot Performance of Compact Whisper Variants on Urdu-English Code-Mixed Speech: An Empirical Investigation Using WER and Switch Point AnalysisOpenAlexPaperAbdullah Haroon Mohammed haroonJun 17
055Next-Turn: Duration-Aware Streaming Endpoint Detection via Time-to-Next-Speech-Onset PredictionOpenAlexPaperTristan Tsoi, Jiajun Deng et al.Jun 16
056CraBERT: Efficient Phoneme Encoder Pre-Training via Cascade Fusion of Subword Representations for Text-to-SpeechOpenAlexPaperD Yang, Yuki Saito et al.Jun 15
057Joycent: Diffusion-based Accent TTS without Accented Phone PredictionOpenAlexPaperXintong Wang, Ye WangJun 15
058Bridging the SEA Gap: An Initial Benchmark for Neural Audio Codec-Synthesized Speech Deepfakes in South-East Asian LanguagesOpenAlexPaperOrchid Chetia Phukan, Girish et al.Jun 14
059Automated Vietnamese text-to-speech: A survey of artificial intelligence techniquesOpenAlexPaperTrương Thanh Thảo, Tran Thanh Nam et al.Jun 13
060CSE-Guided Linguistically Constrained Morphological Segmentation for TurkmenOpenAlexPaperUalsher Tukeyev, Dina Amirova et al.Jun 13

Showing 60 of 2,287 documents · scroll for more