Search results for: text-to-speech transcription

Search results for: text-to-speech transcription

results on page:
embed this view on your website

Towards Effective Processing of Large Text Collections
Publication
- J. Szymański
- H. Krawczyk
- Year 2012
In the article we describe the approach to parallelimplementation of elementary operations for textual data categorization.In the experiments we evaluate parallel computations ofsimilarity matrices and k-means algorithm. The test datasets havebeen prepared as graphs created from Wikipedia articles relatedwith links. When we create the clustering data packages, wecompute pairs of eigenvectors and eigenvalues for visualizationsof...
Machine Learning and Text Analysis in an Artificial Intelligent System for the Training of Air Traffic Controllers
Publication
- T. Shmelova
- Y. Sikirda
- N. Rizun
- V. Lazorenko
- V. Kharchenko
- Year 2020
This chapter presents the application of new information technology in education for the training of air traffic controllers (ATCs). Machine learning, multi-criteria decision analysis, and text analysis as the methods of artificial intelligence for ATCs training have been described. The authors have made an analysis of the International Civil Aviation Organization documents for modern principles of ATCs education. The prototype...

Full text available to download
Study on Speech Transmission under Varying QoS Parameters in a OFDM Communication System
Publication
- M. Zamłyńska
- P. Falkowski-Gilski
- G. Debita
- B. Miedziński
- Year 2021
Although there has been an outbreak of multiple multimedia platforms worldwide, speech communication is still the most essential and important type of service. With the spoken word we can exchange ideas, provide descriptive information, as well as aid to another person. As the amount of available bandwidth continues to shrink, researchers focus on novel types of transmission, based most often on multi-valued modulations, multiple...

Full text to download in external service
Database of speech and facial expressions recorded with optimized face motion capture settings
Publication
- A. Czyżewski
- M. Kawaler
- JOURNAL OF INTELLIGENT INFORMATION SYSTEMS - Year 2019
The broad objective of the present research is the analysis of spoken English employing a multiplicity of modalities. An important stage of this process, discussed in the paper, is creating a database of speech accompanied with facial expressions. Recordings of speakers were made using an advanced system for capturing facial muscle motion. A brief historical outline, current applications, limitations and the ways of capturing face...

Full text available to download
Transfer learning in imagined speech EEG-based BCIs
Publication
- J. S. Garcia Salinas
- L. Villaseñor-Pineda
- C. A. Reyes-Garćia
- A. A. Torres-García
- Biomedical Signal Processing and Control - Year 2019
The Brain–Computer Interfaces (BCI) based on electroencephalograms (EEG) are systems which aim is to provide a communication channel to any person with a computer, initially it was proposed to aid people with disabilities, but actually wider applications have been proposed. These devices allow to send messages or to control devices using the brain signals. There are different neuro-paradigms which evoke brain signals of interest...

Full text available to download
Estimation of the short-term predictor parameters of speech under noisy conditions
Publication
- M. Kuropatwinski
- W. Kleijn
- M. Kuropatwiński
- IEEE Transactions on Audio Speech and Language Processing - Year 2006
Full text to download in external service
Improvement of speech intelligibility in the presence of noise interference using the Lombard effect and an automatic noise interference profiling based on deep learning
Publication
- K. Kąkol
- Year 2023
The Lombard effect is a phenomenon that results in speech intelligibility improvement when applied to noise. There are many distinctive features of Lombard speech that were recalled in this dissertation. This work proposes the creation of a system capable of improving speech quality and intelligibility in real-time measured by objective metrics and subjective tests. This system consists of three main components: speech type detection,...

Full text available to download
Estimation of the excitation variances of speech and noise AR-models for enhanced speech coding
Publication
- M. Kuropatwinski
- W. Kleijn
- M. Kuropatwiński
- Year 2001
Full text to download in external service
The role of Snail1 transcription factor in colorectal cancer progression and metastasis
Publication
- M. Brzozowa
- M. Michalski
- G. Wyrobiec
- A. Piecuch
- A. Dittfeld
- M. Harabin-Słowińska
- D. Boroń
- R. Wojnicz
- Contemporary Oncology/Współczesna Onkologia - Year 2015
Full text to download in external service
Multiple transcriptional factors regulate transcription of the rpoE gene in Escherichia coli under different growth conditions and when the lipopolysaccharide biosynthesis is defective.
Publication
- G. Klein-Raina
- A. Stupak
- D. Biernacka
- P. Wojtkiewicz
- B. Lindner
- S. Raina
- JOURNAL OF BIOLOGICAL CHEMISTRY - Year 2016
The RpoE sigma factor is essential for the viability of Escherichia coli. RpoE regulates extracytoplasmic functions including lipopolysaccharide (LPS) translocation and some of its non-stoichiometric modifications. Transcription of the rpoE gene is positively autoregulated by EσE and by unknown mechanisms that control the expression of its distally located promoter(s). Mapping of 5′ ends of rpoE mRNA identified five new transcriptional...

Full text available to download
Subjective Quality Evaluation of Speech Signals Transmitted via BPL-PLC Wired System
Publication
- P. Falkowski-Gilski
- G. Debita
- M. Habrych
- B. Miedziński
- P. Jedlikowski
- B. Polnik
- J. Wandzio
- X. Wang
- Year 2020
The broadband over power line – power line communication (BPL-PLC) cable is resistant to electricity stoppage and partial damage of phase conductors. It maintains continuity of transmission in case of an emergency. These features make it an ideal solution for delivering data, e.g. in an underground mine environment, especially clear and easily understandable voice messages. This paper describes a subjective quality evaluation of...

Full text to download in external service
Noise profiling for speech enhancement employing machine learning models
Publication
- K. Kąkol
- G. Korvel
- B. Kostek
- Journal of the Acoustical Society of America - Year 2022
This paper aims to propose a noise profiling method that can be performed in near real-time based on machine learning (ML). To address challenges related to noise profiling effectively, we start with a critical review of the literature background. Then, we outline the experiment performed consisting of two parts. The first part concerns the noise recognition model built upon several baseline classifiers and noise signal features...

Full text available to download
Intra-subject class-incremental deep learning approach for EEG-based imagined speech recognition
Publication
- J. S. Garcia Salinas
- A. A. Torres-García
- C. A. Reyes-Garćia
- L. Villaseñor-Pineda
- Biomedical Signal Processing and Control - Year 2023
Brain–computer interfaces (BCIs) aim to decode brain signals and transform them into commands for device operation. The present study aimed to decode the brain activity during imagined speech. The BCI must identify imagined words within a given vocabulary and thus perform the requested action. A possible scenario when using this approach is the gradual addition of new words to the vocabulary using incremental learning methods....

Full text to download in external service
Intelligent processing of stuttered speech.
Publication
- A. Czyżewski
- A. Kaczmarek
- JOURNAL OF INTELLIGENT INFORMATION SYSTEMS - Year 2003
W artykule zaprezentowano kilka metod analizy i automatycznego zliczania potknięć artykulacyjnych, związanych z jąkaniem się, opartych na wykorzystaniu algorytmów uczących się sztucznych sieci neuronowych i zbiorów przybliżonych.
Text categorization with semantic commonsense knowledge: First results
Publication
- P. Majewski
- J. Szymański
- Year 2008
Do przetwarzania tekstów typowo wykorzystuje się reprezentacjeBOW. Podejście takie nie daje jednak dobrych rezultatów w sytuacjigdy podobne dokumenty nie współdzielą ze sobą słów.W artykule zaprezentowano podejście do konstrukcji funkcjijądra dla klasyfikatorów SVM opartego na zewnętrznej bazie wiedzyo pojęciach językowych.
External Validation Measures for Nested Clustering of Text Documents
Publication
- K. Draszawka
- J. Szymański
- Year 2011
Abstract. This article handles the problem of validating the results of nested (as opposed to "flat") clusterings. It shows that standard external validation indices used for partitioning clustering validation, like Rand statistics, Hubert Γ statistic or F-measure are not applicable in nested clustering cases. Additionally to the work, where F-measure was adopted to hierarchical classification as hF-measure, here some methods to...
Towards facts extraction from text in Polish language
Publication
- T. M. Boiński
- A. Chojnowski
- Year 2017
Natural Language Processing (NLP) finds many usages in different fields of endeavor. Many tools exists allowing analysis of English language. For Polish language the situation is different as the language itself is more complicated. In this paper we show differences between NLP of Polish and English language. Existing solutions are presented and TEAMS software for facts extraction is described. The paper shows also evaluation of...

Full text available to download
Quality Evaluation of Speech Transmission via Two-way BPL-PLC Voice Communication System in an Underground Mine
Publication
- P. Falkowski-Gilski
- G. Debita
- Archives of Acoustics - Year 2023
In order to design a stable and reliable voice communication system, it is essential to know how many resources are necessary for conveying quality content. These parameters may include objective quality of service (QoS) metrics, such as: available bandwidth, bit error rate (BER), delay, latency as well as subjective quality of experience (QoE) related to user expectations. QoE is expressed as clarity of speech and the ability...

Full text available to download
Expression of Selected Epithelial-Mesenchymal Transition Transcription Factors in Endometrial Cancer
Publication
- P. Sadłecki
- J. Jóźwicki
- P. Antosik
- M. Walentowicz-Sadłecka
- Biomed Research International - Year 2020
Full text to download in external service
Comparison of Language Models Trained on Written Texts and Speech Transcripts in the Context of Automatic Speech Recognition
Publication
- S. Dziadzio
- A. Nabożny
- A. Smywiński-Pohl
- B. Ziółko
- Year 2015
Full text to download in external service
EXAMINING INFLUENCE OF VIDEO FRAMERATE AND AUDIO/VIDEO SYNCHRONIZATION ON AUDIO-VISUAL SPEECH RECOGNITION ACCURACY
Publication
- Year 2014
The problem of video framerate and audio/video synchronization in audio-visual speech recognition is considered. The visual features are added to the acoustic parameters in order to improve the accuracy of speech recognition in noisy conditions. The Mel-Frequency Cepstral Coefficients are used on the acoustic side whereas Active Appearance Model features are extracted from the image. The feature fusion approach is employed. The...
EXAMINING INFLUENCE OF VIDEO FRAMERATE AND AUDIO/VIDEO SYNCHRONIZATION ON AUDIO-VISUAL SPEECH RECOGNITION ACCURACY
Publication
- Year 2014
The problem of video framerate and audio/video synchronization in audio-visual speech recogni-tion is considered. The visual features are added to the acoustic parameters in order to improve the accuracy of speech recognition in noisy conditions. The Mel-Frequency Cepstral Coefficients are used on the acoustic side whereas Active Appearance Model features are extracted from the image. The feature fusion approach is employed. The...
Wieloznaczność w języku i tekście [Ambiguity in language and text]
Publication
- K. Wojan
- PROGRESS. JOURNAL OF YOUNG RESEARCHERS - Year 2017
Full text to download in external service
Representation of hypertext documents based on terms, Links and text compressibility
Publication
- J. Szymański
- W. Duch
- LECTURE NOTES IN COMPUTER SCIENCE - Year 2010
Opisano metody reprezentacji dokumentów tekstowych oparte na słowach, wzajemnych powiązaniach i metodach kompresji. Dokonano ich oceny w oparciu o klasyfikator SVM.
Integration of speech enhancement and coding techniques
Publication
- M. Kuropatwinski
- D. Leckschat
- K. Kroschel
- A. Czyzewski
- M. Kuropatwiński
- Year 1999
Full text to download in external service
A system for multitask noisy speech enhancement.
Publication
- A. Czyżewski
- A. Kaczmarek
- J. Kotus
- A. Pawlik
- A. Rypulak
- P. Żwan
- Year 2004
W artykule przedstawiono ogolną charakterystyke opracowanego systemu rejestracji i rekonstrukcji mowy. Artykuł zawiera opis składników systemu, ktory jest oprogramowaniem zawierającym zaawansowane narzędzia służące poprawie zrozumiałości mowy. Zaimplementowane narzędzia systemu umożliwiają wyszukiwanie nagrań dźwiękowych i ich obróbkę przy pomocy zaimplementowanych pluginów. W artykule przedstawione wykorzystane w systemie algorytmy...
Multitask Noisy Speech Enhancement System
Publication
- A. Czyżewski
- J. Kotus
- G. Szwoch
- M. Dziubiński
- A. Rypulak
- A. Pawlik
- Year 2005
W referacie opisano Wielozadaniowy System Poprawy Jakości Sygnału Mowy. Jest to wyspecjalizowany pakiet oprogramowania przeznaczony do rejestrowania sygnału mowy i do poprawy jego jakości oraz zrozumiałości mowy, przy użyciu zaawansowanych procedur cyfrowego przetwarzania sygnału. Pakiet oprogramowania składa się z programów: Rejestrator, Przeglądarka oraz Rekonstruktor. Oprogramowanie to może być użyte w przypadkach, gdy zrozumiałość...
Novel approaches to wideband speech coding
Publication
- M. Kulesza
- A. Czyżewski
- Year 2008
Dwie metoda kodowania szerokopasmowego mowy zostały zaprezentowane. W pierwszej metodzie wykorzystano algorytm kompresji i ekspansji czasowej sygnału mowy, pozwalający na kodowanie szerokopasmowe sygnału mowy z wykorzystaniem ustandaryzowanych kodeków. Metoda ta jest przewidziana do zastosowania w adaptacyjnych algorytmach kodowania mowy. Drugie z proponowanych rozwiazan dotyczy nowej metody estymacji obwiedni widma sygnalu mowy...

Full text to download in external service
Broadband interference in speech reinforcement systems
Publication
- H. Lasota
- R. Mazurek
- Year 2008
Artykuł podejmuje niedoceniany problem wpływu liczby i rozkładu głośników w systemach nagłośnienia, na jakość przekazu głosowego, czyli na zrozumiałość mowy w audytoriach. Superpozycji przesuniętych w czasie szerokopasmowych sygnałów o tym samym kształcie i lekko różnych wielkościach, które docierają do słuchacza z licznych spójnych źródeł, towarzyszy zjawisko interferencji prowadzące do głębokiej modyfikacji odbieranych sygnałów...
Difference in Perceived Speech Signal Quality Assessment Among Monolingual and Bilingual Teenage Students
Publication
- P. Falkowski-Gilski
- Year 2021
The user perceived quality is a mixture of factors, including the background of an individual. The process of auditory perception is discussed in a wide variety of fields, ranging from engineering to medicine. Many studies examine the difference between musicians and non-musicians. Since musical training develops musical hearing and other various auditory capabilities, similar enhancements should be observable in case of bilingual...

Full text to download in external service
The Phytoestrogen Genistein Modulates Lysosomal Metabolism and Transcription Factor EB (TFEB) Activation
Publication
- M. Moskot
- S. Montefusco
- J. Jakóbkiewicz-Banecka
- P. Mozolewski
- A. Węgrzyn
- D. Di
- G. Węgrzyn
- D. Medina
- A. Ballabio
- M. Gabig-Cimińska
- Journal of Biological Chemistry - Year 2014
Full text to download in external service
Language material for English audiovisual speech recognition system developmen . Materiał językowy do wykorzystania w systemie audiowizualnego rozpoznawania mowy angielskiej
Publication
- A. Czyżewski
- B. Kostek
- T. Ciszewski
- D. Majewicz
- Year 2013
The bi-modal speech recognition system requires a 2-sample language input for training and for testing algorithms which precisely depicts natural English speech. For the purposes of the audio-visual recordings, a training data base of 264 sentences (1730 words without repetitions; 5685 sounds) has been created. The language sample reflects vowel and consonant frequencies in natural speech. The recording material reflects both the...
Estimation of time-frequency complex phase-based speech attributes using narrow band filter banks
Publication
- K. Abratkiewicz
- K. Czarnecki
- D. Fourer
- F. Auger
- Year 2017
In this paper, we present nonlinear estimators of nonstationary and multicomponent signal attributes (parameters, properties) which are instantaneous frequency, spectral (or group) delay, and chirp-rate (also known as instantaneous frequency slope). We estimate all of these distributions in the time-frequency domain using both finite and infinite impulse response (FIR and IIR) narrow band filers for speech analysis. Then, we present...

Full text available to download
Application of colour image segmentation for localization and extraction text from images
Publication
- M. Pazio
- K. Cisowski
- Year 2005
W otaczającym nas świecie informacja tekstowa odgrywa wielką rolę. W postaci tekstowej podawane są: nazwy ulic, nazwy sklepów i instytucji, opisy przedmiotów np. tytuły książek, opakowań itp. Jednocześnie współczesne programy komputerowe służące do rozpoznawania tekstu (OCR) ''nie radzą sobie'' z analizą obrazów otrzymanaych za pomocą kamer. Segmentacja obrazu z następującą kontekstową analizą parametrów segmentów może dostarczyć...
Application of dynamic time warping and cepstrograms to text-dependent speaker verification
Publication
- A. Kaczmarek
- M. Staworko
- Year 2009
This work provides a description of an automatic speaker verification (ASV) system. In particular, it documents the evolution of all individual stages of the proposed ASV system design from the phase of preprocessing to an operational decision making system. The aim of this research was to achieve the system of the best safety and ease of use in view of users. The objective estimation of this target has been accomplished by assessing...
Speech recognition system for hearing impaired people.
Publication
- P. Dalka
- A. Czyżewski
- Year 2005
Praca przedstawia wyniki badań z zakresu rozpoznawania mowy. Tworzony system wykorzystujący dane wizualne i akustyczne będzie ułatwiał trening poprawnego mówienia dla osób po operacji transplantacji ślimaka i innych osób wykazujących poważne uszkodzenia słuchu. Active Shape models zostały wykorzystane do wyznaczania parametrów wizualnych na podstawie analizy kształtu i ruchu ust w nagraniach wideo. Parametry akustyczne bazują na...
Transient detection algorithms for speech coding applications
Publication
- G. Szwoch
- M. Kulesza
- A. Czyzewski
- Journal of the Acoustical Society of America - Year 2006
Full text to download in external service
Comprehensive Evaluation of Statistical Speech Waveform Synthesis
Publication
- T. Merritt
- B. Putrycz
- A. Nadolski
- T. Ye
- D. Korzekwa
- W. Dolecki
- T. Drugman
- V. Klimkov
- A. Moinet
- A. Breen... and 3 others
- Year 2018
Full text to download in external service
New generation speech aid for stuttering people
Publication
- P. Odya
- A. Czyżewski
- Year 2008
Współczesne Cyfrowe Procesory Sygnałowe (ang. DSP) mają niewielkie wymiary, ale są w stanie re-alizować złożone algorytmy. Ich dodatkową zaletą jest łatwość wymiany oprogramowania, a co za tym idzie łatwość zmiany dziedziny zastosowań. Wykorzystując możliwości procesów stało się możliwe budowanie miniaturowych protez słuchu i mowy. W referacie skupiono się na zagadnieniach związanych z projekto-wanie i implementacją algorytmów...

Full text available to download
New generation speech aid for stuttering people
Publication
- P. Odya
- A. Czyżewski
- Archives of Acoustics - Year 2008
Współczesne Cyfrowe Procesory Sygnałowe (ang. DSP) mają niewielkie wymiary, ale są w stanie re-alizować złożone algorytmy. Ich dodatkową zaletą jest łatwość wymiany oprogramowania, a co za tym idzie łatwość zmiany dziedziny zastosowań. Wykorzystując możliwości procesów stało się możliwe budowanie miniaturowych protez słuchu i mowy. W referacie skupiono się na zagadnieniach związanych z projekto-wanie i implementacją algorytmów...

Full text available to download
Influence of modulation detection threshold on speech intelligibility
Publication
- K. Leo
- ACTA PHYSICA POLONICA A - Year 2011
Full text available to download
Ontology-based text convolution neural network (TextCNN) for prediction of construction accidents
Publication
- S. Donghui
- L. Zhigang
- J. Zurada
- A. Manikas
- J. Guan
- P. Weichbroth
- KNOWLEDGE AND INFORMATION SYSTEMS - Year 2024
The construction industry suffers from workplace accidents, including injuries and fatalities, which represent a significant economic and social burden for employers, workers, and society as a whole.The existing research on construction accidents heavily relies on expert evaluations,which often suffer from issues such as low efficiency, insufficient intelligence, and subjectivity.However, expert opinions provided in construction...

Full text to download in external service
Plug-in to Eclipse environment for VHDL source code editor with advanced formatting of text
Publication
- B. Niton
- K. Pozniak
- R. Romaniuk
- R. S. Romaniuk
- Year 2011
Full text to download in external service
New markers for regulation of transcription and macromolecule metabolic process in porcine oocytes during in vitro maturation
Publication
- M. Brązert
- W. Kranc
- M. Nawrocki
- P. Sujka‑Kordowska
- A. Konwerska
- M. Jankowski
- I. Kocherova
- P. Celichowski
- M. Jeseta
- K. Ożegowska... and 9 others
- Molecular Medicine Reports - Year 2020
Full text to download in external service
Hyponastic Leaves 1 Interacts with RNA Pol II to Ensure Proper Transcription of MicroRNA Genes
Publication
- D. Bielewicz
- J. Dolata
- M. Bajczyk
- L. Szewc
- T. Gulanicz
- S. Bhat
- A. Karlik
- M. Jozwiak
- A. Jarmolowski
- Z. Szweykowska-Kulinska
- PLANT AND CELL PHYSIOLOGY - Year 2023
Full text to download in external service
The effect of cabbage juice and its active components on transcription factor Nrf2 in breast cell lines
Publication
- V. Krajka-Kuźniak
- H. Szaefer
- B. Licznerska
- A. Bartoszek-Pączkowska
- W. Bear-Dubowska
- Acta Biochimica Polonica - Year 2009
Obniżenie zagrożenia chorobami nowotworowymi w przypadku diety bogatej w kapustę wiąże się ze zdolnością bioaktywnych związków tego warzywa do indukowania enzymów ochronnych 2. fazy będacych pod kontrolą czynnika transkrypcyjnego Nrf2. Zbadano wpływ soków z kapusty na poziom Nrf2 w komórkach unieśmiertelnionych i nowotworowych linii odpowiednio MCF10A i MCF-7. W obu przypadkach obserwowano wzrost poziomu tego czynnika oraz jego...

Full text to download in external service
System of speech signal processing and visualisation for linguistic purposes
Publication
- K. Wojan
- Archives of Acoustics - Year 2005
Digital analysis of ethnic speech – extraction of information code
Publication
- K. Wojan
- Archives of Acoustics - Year 2003
On the EM algorithm for the estimation of speech AR parameters in noise
Publication
- M. Kuropatwinski
- B. Kleijn
- M. Kuropatwiński
- Year 2014
Full text to download in external service
New approach to localization of clicks in archive speech signals.
Publication
- M. Niedźwiecki
- A. Sobociński
- Year 2004
Przedstawiono problem lokalizacji zniekształceń impulsowych w archiwalnych sygnałach mowy. Pokazano, że detekcja oparta na dwuzakresowym modelu autoregresyjnym i przetwarzanie dwukierunkowe pozwala uzyskać znaczącą poprawę działania w stosunku do istniejących metod lokalizacji zniekształceń.

Search

Filters

Catalog

Category

Year

Options

Search results for: text-to-speech transcription