A Speech to Text Transcription Approach based on Romanian Corpus

Authors

DOI:

https://doi.org/10.70594/

Keywords:

speech recognition, formant energy, algorithmic method, neural network

Abstract

Automatic speech segmentation has many applications in speech processing and phonetics, e.g., in automatic speech recognition and automatic annotation of speech corpora. In both processes of training and evaluation of speech recognition systems large aligned speech-to-text corpora are needed. Once aligned, identification of phonemes could be based on samples that are picked-up inbetween phonemes' boundaries. Because manual segmentation is costly and extremely time consuming, automatic methods of alignment are searched for. In this paper, we propose a simple, yet efficient, method for speech to text recognition based on a machine learning approach, using a Romanian speech corpus.

Author Biographies

  • Andrei Scutelnicu, Alexandru Ioan Cuza University of Iasi, Romania & Romanian Academy, Iași, Romania

    Faculty of Computer Science
    Alexandru Ioan Cuza University of Iasi, Romania
    Institute of Computer Science
    Romanian Academy, Iași, Romania

  • Mihaela Onofrei , Alexandru Ioan Cuza University of Iasi, Romania & Romanian Academy, Iași, Romania

    Faculty of Computer Science
    Alexandru Ioan Cuza University of Iasi, Romania
    Institute of Computer Science
    Romanian Academy, Iași, Romania

  • Anca Diana Bibiri , Alexandru Ioan Cuza University of Iași, Romania 

    Alexandru Ioan Cuza University of Iași, Romania 

  • Mircea Hulea , Gheorghe Asachi University of Iași, Romania

    Faculty of Automatic Control and Computer Engineering
    Gheorghe Asachi University of Iași, Romania

Downloads

Published

2025-07-28

Issue

Section

Table of Contents

Similar Articles

11-20 of 430

You may also start an advanced similarity search for this article.