Filters
total: 2574
displaying 1000 best results Help
Search results for: SPECH PROCESSING
-
Analytical and legislative challenges of sewage sludge processing and management
PublicationThis article presents the most popular methods of sewage sludge management and analytical techniques which could be a powerful tool in designing new sewage sludge management methods. Chemical analysis is also described as a vital point at the subsequent stages of technological processes control and sewage sludge quality assessment. It is also an instrument essential to maintaining control of processed sewage sludge introduced to...
-
Towards Understanding the Health Aspects of the Processing of Lignocellulosic Fillers
PublicationHealth and safety issues should be addressed during the development and investigation of the industrial processes. In order to develop a sustainable process and fully evaluate its benefits and drawbacks for its optimization, it is crucial to determine its impact on the surrounding environment. This study aimed to assess the emission of volatile organic compounds during the modification of lignocellulosic fillers with passive dosimetry....
-
Food Bioactive Ingredients Processing Using Membrane Distillation
PublicationSeparation processes are an important part of today’s food industries, especially in the case of specific bioactive components due to their health benefits. In general, processing of bioactive food ingredients assumes the introduction of integrated system directed to their separation, fractionation, and recovery. Recently, membrane distillation (MD) has been considered as an alternative membrane-based separation and concentration...
-
Digital Processing of Frequency–Pulse Signal in Measurement System
PublicationThe work presents the issue of the use of multichannel measurement systems of sensors processing input value to impulse signal frequency. The frequency impulse signal obtained from such sensors is often required to be processed at the same time with a voltage signal which is obtained from other sensors used in the same measurement system. In such case, it is usually necessary to sample the output signals from all sensors in the...
-
Influence of Processing Conditions on Crystal Structure of Bi6Fe2Ti3O18 Ceramics
PublicationAim of the present research was to apply a solid state reaction route to fabricate Aurivillius-type ceramics described with the formula Bi6Fe2Ti3O18 (BFTO) and reveal the influence of processing conditions on its crystal structure. Pressureless sintering in ambient air was employed and the sintering temperatures were 850 and 1080 °C. It was found that the fabricated BFTO ceramics were multiphase ones. They consisted of two Bim+1Fem-3Ti3O3m+3...
-
Vident-lab: a dataset for multi-task video processing of phantom dental scenes
Open Research DataWe introduce a new, asymmetrically annotated dataset of natural teeth in phantom scenes for multi-task video processing: restoration, teeth segmentation, and inter-frame homography estimation. Pairs of frames were acquired with a beam splitter. The dataset constitutes a low-quality frame, its high-quality counterpart, a teeth segmentation mask, and...
-
Lecture Notes in Business Information Processing
Journals -
Time-domain prosodic modifications for text-to-speech synthesizer
PublicationAn application of prosodic speech processing algorithms to Text-To-Speech synthesis is presented. Prosodic modifications that improve the naturalness of the synthesized signal are discussed. The applied method is based on the TD-PSOLA algorithm. The developed Text-To-Speech Synthesizer is used in applications employing multimodal computer interfaces.
-
Evaluation and Irony in Text in the Light of Speech Act Theory
Publication -
A Method of Real-Time Non-uniform Speech Stretching
PublicationDeveloped method of real-time non-uniform speech stretching is presented.The proposed solution is based on the well-known SOLA algorithm(Synchronous Overlap and Add). Non-uniform time-scale modification isachieved by the adjustment of time scaling factor values in accordance with thesignal content. Dependently on the speech unit (vowels/consonants), instantaneousrate of speech (ROS), and speech signal presence, values of the scalingfactor...
-
Visual Lip Contour Detection for the Purpose of Speech Recognition
PublicationA method for visual detection of lip contours in frontal recordings of speakers is described and evaluated. The purpose of the method is to facilitate speech recognition with visual features extracted from a mouth region. Different Active Appearance Models are employed for finding lips in video frames and for lip shape and texture statistical description. Search initialization procedure is proposed and error measure values are...
-
Investigations of speech signal parameters with regard to articulation influences
PublicationW pracy zostało podjęte zagadnienie parametryzacji sygnału mowy w kontekście ekstrakcji cech biometrycznych. Analizowane parametry to parametry cepstralne (cepstrum liniowe i mel-cepstrum, czyli MFCC), parametry liniowej predykcji (LPC) oraz momenty widmowe i parametr F0. Zastosowano analize w krótkich stałych segmentach sygnału z zastosowaniem dużego zakładkowania, tzw. ''implicite segmentation''. Umożliwiło to zaobserwowanie...
-
New approach to localization of clicks in archive speech signals.
PublicationPrzedstawiono problem lokalizacji zniekształceń impulsowych w archiwalnych sygnałach mowy. Pokazano, że detekcja oparta na dwuzakresowym modelu autoregresyjnym i przetwarzanie dwukierunkowe pozwala uzyskać znaczącą poprawę działania w stosunku do istniejących metod lokalizacji zniekształceń.
-
Detection of dialogue in movie soundtrack for speech intelligibility enhancement
PublicationA method for detecting dialogue in 5.1 movie soundtrack based on interchannel spectral disparity is presented. The front channel signals (left, right, center) are analyzed in the frequency domain. The selected partials in the center channel signal, which yield high disparity with left and right channels, are detected as dialogue. Subsequently, the dialogue frequency components are boosted to achieve increased dialogue intelligibility....
-
Advanced speech archiving and restoration system for aviation applications
PublicationW referacie przedstawiono opracowany System Rejestracji I Rekonstrukcji Mowy dla potrzeb lotnictwa. System ten umożliwia jednoczesny zapis, archiwizację i poprawę zrozumiałości sygnału mowy pochodzącego z wielu różnych kanałów komunikacji radiowej. Głównym celem systemu jest rejestracja i rekonstrukcja komunikatów słownych wymienianych drogą radiową pomiędzy pilotem samolotu a stacją kontroli lotów - jest to niezwykle istotne w...
-
Application of hybrid signals processors to speech and hearing aids
PublicationDzięki postępowi w technice Cyfrowych Procesorów Sygnałowych (ang. DSP) stało się możliwe budowanie miniaturowych protez słuchu i mowy. Mimo niewielkich wymiarów procesory te są w stanie wykonywać złożone algorytmy. Ich dodatkową zaletą jest łatwość zmiany oprogramowania, a co za tym idzie łatwość zmiany dziedziny zastosowań. W pracy skupiono się na zagadnieniach związanych z projektowanie i implementacją algorytmów mających zastosowanie...
-
Digital analysis of ethnic speech – extraction of information code
Publication -
On the EM algorithm for the estimation of speech AR parameters in noise
Publication -
Automatic Image and Speech Recognition Based on Neural Network
Publication -
Audiovisual speech recognition for training hearing impaired patients
PublicationPraca przedstawia system rozpoznawania izolowanych głosek mowy wykorzystujący dane wizualne i akustyczne. Modele Active Shape Models zostały wykorzystane do wyznaczania parametrów wizualnych na podstawie analizy kształtu i ruchu ust w nagraniach wideo. Parametry akustyczne bazują na współczynnikach melcepstralnych. Sieć neuronowa została użyta do rozpoznawania wymawianych głosek na podstawie wektora cech zawierającego oba typy...
-
Examining Influence of Distance to Microphone on Accuracy of Speech Recognition
PublicationThe problem of controlling a machine by the distant-talking speaker without a necessity of handheld or body-worn equipment usage is considered. A laboratory setup is introduced for examination of performance of the developed automatic speech recognition system fed by direct and by distant speech acquired by microphones placed at three different distances from the speaker (0.5 m to 1.5 m). For feature extraction from the voice signal...
-
An audio-visual corpus for multimodal automatic speech recognition
Publicationreview of available audio-visual speech corpora and a description of a new multimodal corpus of English speech recordings is provided. The new corpus containing 31 hours of recordings was created specifically to assist audio-visual speech recognition systems (AVSR) development. The database related to the corpus includes high-resolution, high-framerate stereoscopic video streams from RGB cameras, depth imaging stream utilizing Time-of-Flight...
-
Comparison of various speech time-scale modificartion methods
PublicationThe objective of this work is to investigate the influence of the different time-scale modification (TSM) methods on the quality of the speech stretched up using the designed non-uniform real-time speech time-scale modification algorithm (NU-RTSM). The algorithm provides a combination of the typical TSM algorithm with the vowels, consonants, stutter, transients and silence detectors. Based on the information about the content and...
-
Acoustic Sensing Analytics Applied to Speech in Reverberation Conditions
PublicationThe paper aims to discuss a case study of sensing analytics and technology in acoustics when applied to reverberation conditions. Reverberation is one of the issues that makes speech in indoor spaces challenging to understand. This problem is particularly critical in large spaces with few absorbing or diffusing surfaces. One of the natural remedies to improve speech intelligibility in such conditions may be achieved through speaking...
-
Detecting Lombard Speech Using Deep Learning Approach
PublicationRobust Lombard speech-in-noise detecting is challenging. This study proposes a strategy to detect Lombard speech using a machine learning approach for applications such as public address systems that work in near real time. The paper starts with the background concerning the Lombard effect. Then, assumptions of the work performed for Lombard speech detection are outlined. The framework proposed combines convolutional neural networks...
-
Improving Objective Speech Quality Indicators in Noise Conditions
PublicationThis work aims at modifying speech signal samples and test them with objective speech quality indicators after mixing the original signals with noise or with an interfering signal. Modifications that are applied to the signal are related to the Lombard speech characteristics, i.e., pitch shifting, utterance duration changes, vocal tract scaling, manipulation of formants. A set of words and sentences in Polish, recorded in silence,...
-
Ranking Speech Features for Their Usage in Singing Emotion Classification
PublicationThis paper aims to retrieve speech descriptors that may be useful for the classification of emotions in singing. For this purpose, Mel Frequency Cepstral Coefficients (MFCC) and selected Low-Level MPEG 7 descriptors were calculated based on the RAVDESS dataset. The database contains recordings of emotional speech and singing of professional actors presenting six different emotions. Employing the algorithm of Feature Selection based...
-
Transfer learning in imagined speech EEG-based BCIs
PublicationThe Brain–Computer Interfaces (BCI) based on electroencephalograms (EEG) are systems which aim is to provide a communication channel to any person with a computer, initially it was proposed to aid people with disabilities, but actually wider applications have been proposed. These devices allow to send messages or to control devices using the brain signals. There are different neuro-paradigms which evoke brain signals of interest...
-
Quantum Information Processing
Journals -
Results of tests on speech intelligibility in reverberant conditions
Open Research DataThe dataset contains the results of tests that aimed to provide a relationship between the rate of speech (RoS) and reverberation conditions characterized by the Speech Transmission Index (STI).
-
Artur Gańcza dr inż.
PeopleI received the M.Sc. degree from the Gdańsk University of Technology (GUT), Gdańsk, Poland, in 2019. I am currently a Ph.D. student at GUT, with the Department of Automatic Control, Faculty of Electronics, Telecommunications and Informatics. My professional interests include speech recognition, system identification, adaptive signal processing and linear algebra.
-
WSEAS Transactions on Signal Processing
Journals -
Journal of Signal and Image Processing
Journals -
Journal of Food Processing & Technology
Journals -
Foundations and Trends in Signal Processing
Journals -
Journal of Information Processing Systems
Journals -
Scientific and Technical Information Processing
Journals -
Optoelectronics Instrumentation and Data Processing
Journals -
Eurographics Symposium on Geometry Processing
Journals -
Lasers in Manufacturing and Materials Processing
Journals -
IEEE SIGNAL PROCESSING LETTERS
Journals -
MULTIDIMENSIONAL SYSTEMS AND SIGNAL PROCESSING
Journals -
Food Processing: Techniques and Technology
Journals -
Material Design and Processing Communications
Journals -
Springer Topics in Signal Processing
Journals -
Food Production Processing and Nutrition
Journals -
Theory and Practice of Meat Processing
Journals -
INTERNATIONAL JOURNAL OF MINERAL PROCESSING
Journals -
JOURNAL OF CERAMIC PROCESSING RESEARCH
Journals -
SIGNAL PROCESSING-IMAGE COMMUNICATION
Journals