Wyniki wyszukiwania dla: VISUAL SPEECH RECOGNITION

Wyniki wyszukiwania dla: VISUAL SPEECH RECOGNITION

wyników na stronę:
osadź ten widok na swojej stronie

Filtry

wszystkich: 1419

wyczyść wszystkie filtry niedostępne

wyświetlamy 1000 najlepszych wyników Pomoc

Parameters optimization in medicine supporting image recognition algorithms
Publikacja
- A. Brzeski
- Rok 2011
In this paper, a procedure of automatic set up of image recognition algorithms' parameters is proposed, for the purpose of reducing the time needed for algorithms' development. The procedure is presented on two medicine supporting algorithms, performing bleeding detection in endoscopic images. Since the algorithms contain multiple parameters which must be specified, empirical testing is usually required to optimise the algorithm's...
Accelerometer-based Human Activity Recognition and the Impact of the Sample Size
Publikacja
- Rok 2014
The presented study focused on the recognition of eight user activities (e.g. walking, lying, climbing stairs) basing on the measurements from an accelerometer embedded in a mobile device. It is assumed that the device is carried in a specific location of the user’s clothing. Three types of classifiers were tested on different sizes of the samples. The influence of the time window (the duration of a single trial) on selected activities...

Pełny tekst do pobrania w serwisie zewnętrznym
Comparison of selected off-the-shelf solutions for emotion recognition based on facial expressions
Publikacja
- Rok 2016
The paper concerns accuracy of emotion recognition from facial expressions. As there are a couple of ready off-the-shelf solutions available in the market today, this study aims at practical evaluation of selected solutions in order to provide some insight into what potential buyers might expect. Two solutions were compared: FaceReader by Noldus and Xpress Engine by QuantumLab. The performed evaluation revealed that the recognition...

Pełny tekst do pobrania w serwisie zewnętrznym
Automatic Singing Voice Recognition EmployingNeural Networks and Rough Sets
Publikacja
- Rok 2008
Celem badań jest automatyczne rozpoznawanie głosów śpiewaczych w kategorii rodzaju i jakości technicznej śpiewu. W artykule opisano stworzoną bazę danych głosów, która zawiera próbki głosu śpiewaków profesjonalnych i amatorskich. W dalszej części opisano parametry zdefiniowane w oparciu o zjawiska biomechaniczne w narządzie głosu podczas śpiewania. W oparciu o stworzone macierze parametrów wytrenowano i porównano automatyczne klasyfikatory...
Cross-Lingual Knowledge Distillation via Flow-Based Voice Conversion for Robust Polyglot Text-to-Speech
Publikacja
- D. Piotrowski
- R. Korzeniowski
- A. Falai
- S. Cygert
- K. Pokora
- G. Tinchev
- Z. Zhang
- K. Yanagisawa
- Rok 2023
In this work, we introduce a framework for cross-lingual speech synthesis, which involves an upstream Voice Conversion (VC) model and a downstream Text-To-Speech (TTS) model. The proposed framework consists of 4 stages. In the first two stages, we use a VC model to convert utterances in the target locale to the voice of the target speaker. In the third stage, the converted data is combined with the linguistic features and durations...

Pełny tekst do pobrania w serwisie zewnętrznym
Systematic Literature Review for Emotion Recognition from EEG Signals
Publikacja
- P. A. Leszczełowska
- N. Dawidowska
- Rok 2022
Researchers have recently become increasingly interested in recognizing emotions from electroencephalogram (EEG) signals and many studies utilizing different approaches have been conducted in this field. For the purposes of this work, we performed a systematic literature review including over 40 articles in order to identify the best set of methods for the emotion recognition problem. Our work collects information about the most...

Pełny tekst do pobrania w serwisie zewnętrznym
Systematic Literature Review for Emotion Recognition from EEG Signals
Publikacja
- P. A. Leszczełowska
- N. Dawidowska
- Rok 2022
Researchers have recently become increasingly interested in recognizing emotions from electroencephalogram (EEG) signals and many studies utilizing different approaches have been conducted in this field. For the purposes of this work, we performed a systematic literature review including over 40 articles in order to identify the best set of methods for the emotion recognition problem. Our work collects information about the most...

Pełny tekst do pobrania w portalu
Automatic recognition of therapy progress among children with autism
Publikacja
- A. Kołakowska
- A. Landowska
- A. Anzulewicz
- K. Sobota
- Scientific Reports - Rok 2017
The article presents a research study on recognizing therapy progress among children with autism spectrum disorder. The progress is recognized on the basis of behavioural data gathered via five specially designed tablet games. Over 180 distinct parameters are calculated on the basis of raw data delivered via the game flow and tablet sensors - i.e. touch screen, accelerometer and gyroscope. The results obtained confirm the possibility...

Pełny tekst do pobrania w portalu
Visual Attention Distribution Based Assessment of User's Skill in Electronic Medical Record Navigation
Publikacja
- T. Kocejko
- J. Wtorek
- K. Goforth
- K. Moidu
- Journal of Medical Imaging and Health Informatics - Rok 2015
Currently, the most precise way of reflecting the skills level is an expert’s subjective assessment. In this paper we investigate the possibility of the use of eye tracking data for scalar quantitative and objective assessment of medical staff competency in EMR system navigation. According to the experiment conducted by Yarbus the observation process of particular features is associated with thinking. Moreover, eye tracking is...

Pełny tekst do pobrania w serwisie zewnętrznym
Local Texture Pattern Selection for Efficient Face Recognition and Tracking
Publikacja
- M. Smiatacz
- J. Rumiński
- Advances in Intelligent Systems and Computing - Rok 2015
This paper describes the research aimed at finding the optimal configuration of the face recognition algorithm based on local texture descriptors (binary and ternary patterns). Since the identification module was supposed to be a part of the face tracking system developed for interactive wearable computer, proper feature selection, allowing for real-time operation, became particularly important. Our experiments showed that it is...

Pełny tekst do pobrania w serwisie zewnętrznym
Place attachment, place identity, and visual pollution sensitivity.
Dane Badawcze
- M. Jaśkiewicz
- Uniwersytet Gdański
The data include individual responses on the following scales (1) place attachment, (2) place identity, and (3) visual pollution sensitivity. Each line represents responses obtained from one participant and his or her demographic characteristics. like gender, age, and education level.
Feasibility Study for Food Intake Tasks Recognition Based on Smart Glasses
Publikacja
- M. Biallas
- A. Andrushevich
- R. Kistler
- A. Klapproth
- K. Czuszyński
- A. Bujnowski
- Journal of Medical Imaging and Health Informatics - Rok 2015
In this exploratory study 13 adult test subjects have performed different food intake tasks while wearing a three axis accelerometer mounted at a temple of glasses. Two different algorithms for task recognition have been applied and compared. The retrospective data processing leads to better task recognition results when the frequency range of 50 Hz to 100 Hz is analysed within accelerometer signal recordings. A straightforward...

Pełny tekst do pobrania w serwisie zewnętrznym
Fuzzy rule-based dynamic gesture recognition employing camera & multimedia projector
Publikacja
- M. Lech
- B. Kostek
- Rok 2010
In the paper the system based on camera and multimedia projector enabling a user to control computer applications by dynamic hand gestures is presented. The main objective is to present the gesture recognition methodology which bases on representing hand movement trajectory by motion vectors analyzed using fuzzy rule-based inference. The approach was engineered in the system developed with J2SE and C++ / OpenCV technology. OpenCV...

Pełny tekst do pobrania w serwisie zewnętrznym
Novel approaches to wideband speech coding
Publikacja
- M. Kulesza
- A. Czyżewski
- Rok 2008
Dwie metoda kodowania szerokopasmowego mowy zostały zaprezentowane. W pierwszej metodzie wykorzystano algorytm kompresji i ekspansji czasowej sygnału mowy, pozwalający na kodowanie szerokopasmowe sygnału mowy z wykorzystaniem ustandaryzowanych kodeków. Metoda ta jest przewidziana do zastosowania w adaptacyjnych algorytmach kodowania mowy. Drugie z proponowanych rozwiazan dotyczy nowej metody estymacji obwiedni widma sygnalu mowy...

Pełny tekst do pobrania w serwisie zewnętrznym
Integration of speech enhancement and coding techniques
Publikacja
- M. Kuropatwinski
- D. Leckschat
- K. Kroschel
- A. Czyzewski
- M. Kuropatwiński
- Rok 1999
Pełny tekst do pobrania w serwisie zewnętrznym
Broadband interference in speech reinforcement systems
Publikacja
- H. Lasota
- R. Mazurek
- Rok 2008
Artykuł podejmuje niedoceniany problem wpływu liczby i rozkładu głośników w systemach nagłośnienia, na jakość przekazu głosowego, czyli na zrozumiałość mowy w audytoriach. Superpozycji przesuniętych w czasie szerokopasmowych sygnałów o tym samym kształcie i lekko różnych wielkościach, które docierają do słuchacza z licznych spójnych źródeł, towarzyszy zjawisko interferencji prowadzące do głębokiej modyfikacji odbieranych sygnałów...
A system for multitask noisy speech enhancement.
Publikacja
- A. Czyżewski
- A. Kaczmarek
- J. Kotus
- A. Pawlik
- A. Rypulak
- P. Żwan
- Rok 2004
W artykule przedstawiono ogolną charakterystyke opracowanego systemu rejestracji i rekonstrukcji mowy. Artykuł zawiera opis składników systemu, ktory jest oprogramowaniem zawierającym zaawansowane narzędzia służące poprawie zrozumiałości mowy. Zaimplementowane narzędzia systemu umożliwiają wyszukiwanie nagrań dźwiękowych i ich obróbkę przy pomocy zaimplementowanych pluginów. W artykule przedstawione wykorzystane w systemie algorytmy...
Multitask Noisy Speech Enhancement System
Publikacja
- A. Czyżewski
- J. Kotus
- G. Szwoch
- M. Dziubiński
- A. Rypulak
- A. Pawlik
- Rok 2005
W referacie opisano Wielozadaniowy System Poprawy Jakości Sygnału Mowy. Jest to wyspecjalizowany pakiet oprogramowania przeznaczony do rejestrowania sygnału mowy i do poprawy jego jakości oraz zrozumiałości mowy, przy użyciu zaawansowanych procedur cyfrowego przetwarzania sygnału. Pakiet oprogramowania składa się z programów: Rejestrator, Przeglądarka oraz Rekonstruktor. Oprogramowanie to może być użyte w przypadkach, gdy zrozumiałość...
Visual Traffic Noise Monitoring in Urban Areas
Publikacja
- A. Czyżewski
- P. Dalka
- International Journal of Multimedia and Ubiquitous Engineering - Rok 2007
The paper presents an advanced system for railway and road traffic noise monitoring in metropolitan areas. This system is a functional part of a more complex solution designed for environmental monitoring in cities utilizing analyses of sound, vision and air pollution, based on a ubiquitous computing approach. The system consists of many autonomous, universal measuring units and a multimedia server, which gathers, processes and...

Pełny tekst do pobrania w serwisie zewnętrznym
Modeling pragmatics for visual modeling language evaluation
Publikacja
- A. Bobkowska
- Rok 2005
Podczas oceny użyteczności języków modelowania wizualnego istnieje potrzeba uwzględnienia ich pragmatyki. Języki modelowania wizualnego mogą być stosowane w różnym kontekście, co powoduje różnice w wymaganiach, które są im stawiane. Jawny opis kontekstu użycia ułatwia precyzyjną ocenę. Pragmatyka składa się ze zbioru profili, które opisują konkretne konteksty użycia. W referacie podjęto próbę zastosowania modeli zadań do opisu...
Difference in Perceived Speech Signal Quality Assessment Among Monolingual and Bilingual Teenage Students
Publikacja
- P. Falkowski-Gilski
- Rok 2021
The user perceived quality is a mixture of factors, including the background of an individual. The process of auditory perception is discussed in a wide variety of fields, ranging from engineering to medicine. Many studies examine the difference between musicians and non-musicians. Since musical training develops musical hearing and other various auditory capabilities, similar enhancements should be observable in case of bilingual...

Pełny tekst do pobrania w serwisie zewnętrznym
Contextual Knowledge to Enhance Workplace Hazard Recognition and Interpretation in a Cognitive Vision Platform
Publikacja
- C. De
- C. Sanin
- E. Szczerbicki
- Rok 2018
The combination of vision and sensor data together with the resulting necessity for formal representations builds a central component of an autonomous Cyber Physical System for detection and tracking of laborers in workplaces environments. This system must be adaptable and perceive the environment as automatically as possible, performing in a variety of plants and scenes without the necessity of recoding the application for each...

Pełny tekst do pobrania w portalu
JOURNAL OF MOLECULAR RECOGNITION

Czasopisma

ISSN: 0952-3499 , eISSN: 1099-1352
COMPUTER SPEECH AND LANGUAGE

Czasopisma

ISSN: 0885-2308 , eISSN: 1095-8363
SEMINARS IN SPEECH AND LANGUAGE

Czasopisma

ISSN: 0734-0478 , eISSN: 1098-9056
Speech and Language Technology

Czasopisma

ISSN: 1895-0434
Speech Language and Hearing

Czasopisma

ISSN: 1361-3286 , eISSN: 2050-5728
Quarterly Journal of Speech

Czasopisma

ISSN: 0033-5630 , eISSN: 1479-5779
SpringerBriefs in Speech Technology

Czasopisma

ISSN: 2191-737X , eISSN: 2191-7388
Audiology and Speech Research

Czasopisma

ISSN: 2635-5019 , eISSN: 2635-5027
Voice and Speech Review

Czasopisma

ISSN: 2326-8263 , eISSN: 2326-8271
Theory of recognition in a historical perspective. Axel Honneth's Anerkennung: Eine europäische Ideengeschichte
Publikacja
- A. Karalus
- Archiwum Historii Filozofii i Myśli Społecznej - Rok 2019
The article discusses Honneth excursion into the realm of the history of ideas. This time Honneth decides to laser it on the notion of "recognition" in three different cultural areas and three different traditions: French, English, and German. The article discusses Honneth's persepctive and attempts at finding the common thread that would link three aforementioned traditions.

Pełny tekst do pobrania w portalu
A review of emotion recognition methods based on keystroke dynamics and mouse movements
Publikacja
- A. Kołakowska
- Rok 2013
The paper describes the approach based on using standard input devices, such as keyboard and mouse, as sources of data for the recognition of users’ emotional states. A number of systems applying this idea have been presented focusing on three categories of research problems, i.e. collecting and labeling training data, extracting features and training classifiers of emotions. Moreover the advantages and examples of combining standard...

Pełny tekst do pobrania w serwisie zewnętrznym
Visual and auditory attention stimulator for assisting pedagogical therapy . Stymulator uwagi wzrokowej i słuchowej do wspomagania terapii pedagogicznej
Publikacja
- Rok 2015
Visual and auditory attention stimulator provides a system developed in order to improve reading skills using simultaneous presentation of text in its visual form and in transformed auditory form accompanied by related movie material. The described research employed 40 children at the age of 8 13 years having difficulties in learning of reading, who were diagnosed as having developmental dyslexia. It was shown that application...
Speaker Recognition Using Convolutional Neural Network with Minimal Training Data for Smart Home Solutions
Publikacja
- M. Wang
- T. Sirlapu
- A. Kwaśniewska
- M. Szankin
- M. Bartscherer
- R. Nicolas
- Rok 2018
With the technology advancements in smart home sector, voice control and automation are key components that can make a real difference in people's lives. The voice recognition technology market continues to involve rapidly as almost all smart home devices are providing speaker recognition capability today. However, most of them provide cloud-based solutions or use very deep Neural Networks for speaker recognition task, which are...

Pełny tekst do pobrania w serwisie zewnętrznym
Michał Lech dr inż.

Osoby

Michał Lech was born in Gdynia in 1983. In 2007 he graduated from the faculty of Electronics, Telecommunications and Informatics of Gdansk University of Technology. In June 2013, he received his Ph.D. degree. The subject of the dissertation was: “A Method and Algorithms for Controlling the Sound Mixing Processes with Hand Gestures Recognized Using Computer Vision”. The main focus of the thesis was the bias of audio perception caused...
Marek Blok dr hab. inż.

Osoby

Marek Blok w 1994 roku ukończył studia na kierunku Telekomunikacja wydziału Elektroniki Politechniki Gdańskiej i uzyskał tytuł mgra inżyniera. Doktorat w zakresie telekomunikacji uzyskał w 2003 roku na Wydziale Elektroniki, Telekomunikacji i Informatyki Politechniki Gdańskiej. W 2017 roku uzyskał stopień naukowy dra habilitowanego w dyscyplinie telekomunikacja. Jego zainteresowania badawcze ukierunkowane są na telekomunikacyjne...
Journal of Pattern Recognition Research

Czasopisma

ISSN: 1558-884X
Pattern Recognition and Image Analysis

Czasopisma

ISSN: 1054-6618
Graph Representation Integrating Signals for Emotion Recognition and Analysis
Publikacja
- SENSORS - Rok 2021
Data reusability is an important feature of current research, just in every field of science. Modern research in Affective Computing, often rely on datasets containing experiments-originated data such as biosignals, video clips, or images. Moreover, conducting experiments with a vast number of participants to build datasets for Affective Computing research is time-consuming and expensive. Therefore, it is extremely important to...

Pełny tekst do pobrania w portalu
Examining Classifiers Applied to Static Hand Gesture Recognition in Novel Sound Mixing System
Publikacja
- Advances in Intelligent Systems and Computing - Rok 2013
The main objective of the chapter is to present the methodology and results of examining various classifiers (Nearest Neighbor-like algorithm with non-nested generalization (NNge), Naive Bayes, C4.5 (J48), Random Tree, Random Forests, Artificial Neural Networks (Multilayer Perceptron), Support Vector Machine (SVM) used for static gesture recognition. A problem of effective gesture recognition is outlined in the context of the system...

Pełny tekst do pobrania w serwisie zewnętrznym
A Review of Emotion Recognition Methods Based on Data Acquired via Smartphone Sensors
Publikacja
- SENSORS - Rok 2020
In recent years, emotion recognition algorithms have achieved high efficiency, allowing the development of various affective and affect-aware applications. This advancement has taken place mainly in the environment of personal computers offering the appropriate hardware and sufficient power to process complex data from video, audio, and other channels. However, the increase in computing and communication capabilities of smartphones,...

Pełny tekst do pobrania w portalu
Piotr Szczuko dr hab. inż.

Osoby

Katedra Systemów Multimedialnych

Dr hab. inż. Piotr Szczuko w 2002 roku ukończył studia na Wydziale Elektroniki, Telekomunikacji i Informatyki Politechniki Gdańskiej zdobywając tytuł magistra inżyniera. Tematem pracy dyplomowej było badanie zjawisk jednoczesnej percepcji obrazu cyfrowego i dźwięku dookólnego. W roku 2008 obronił rozprawę doktorską zatytułowaną "Zastosowanie reguł rozmytych w komputerowej animacji postaci", za którą otrzymał nagrodę Prezesa Rady...
New Aspects of Virtual Sound Source Localization Research—Impact of Visual Angle and 3-D Video Content on Sound Perception
Publikacja
- B. Kunka
- B. Kostek
- JOURNAL OF THE AUDIO ENGINEERING SOCIETY - Rok 2013
The influence of image on virtual sound source localization, called the “image proximity effect” or the “ventriloquism effect”, is a well known phenomenon. This paper focuses on other aspects related to this effect, namely the impact of the visual angle of the presented object and 3D video content on sound perception. The research conducted confirmed that the visual angle of the presented object determines the image proximity effect...

Pełny tekst do pobrania w portalu
Applicability of Emotion Recognition and Induction Methods to Study the Behavior of Programmers
Publikacja
- M. Wróbel
- Applied Sciences-Basel - Rok 2018
Recent studies in the field of software engineering have shown that positive emotions can increase and negative emotions decrease the productivity of programmers. In the field of affective computing, many methods and tools to recognize the emotions of computer users were proposed. However, it has not been verified yet which of them can be used to monitor the emotional states of software developers. The paper describes a study carried...

Pełny tekst do pobrania w portalu
Estimation of time-frequency complex phase-based speech attributes using narrow band filter banks
Publikacja
- K. Abratkiewicz
- K. Czarnecki
- D. Fourer
- F. Auger
- Rok 2017
In this paper, we present nonlinear estimators of nonstationary and multicomponent signal attributes (parameters, properties) which are instantaneous frequency, spectral (or group) delay, and chirp-rate (also known as instantaneous frequency slope). We estimate all of these distributions in the time-frequency domain using both finite and infinite impulse response (FIR and IIR) narrow band filers for speech analysis. Then, we present...

Pełny tekst do pobrania w portalu
Recognition of environmentally important ions
Publikacja
- N. Łukasik
- E. Wagner-Wysiecka
- V. Hubscher-Bruder
- M. Bocheńska
- S. Michel
- Logistyka - Rok 2013
..
High frequency oscillations are associated with cognitive processing in human recognition memory
Publikacja
- M. T. Kucewicz
- J. Cymbalnik
- J. Matsumoto
- B. H. Brinkmann
- M. R. Bower
- V. Vasoli
- V. Sulc
- F. Meyer
- W. Marsh
- S. M. Stead
- G. A. Worrell
- Brain: A Journal of Neurology - Rok 2014
High frequency oscillations are associated with normal brain function, but also increasingly recognized as potential biomarkers of the epileptogenic brain. Their role in human cognition has been predominantly studied in classical gamma frequencies (30-100 Hz), which reflect neuronal network coordination involved in attention, learning and memory. Invasive brain recordings in animals and humans demonstrate that physiological oscillations...

Pełny tekst do pobrania w portalu
Influence of Thermal Imagery Resolution on Accuracy of Deep Learning based Face Recognition
Publikacja
- Rok 2019
Human-system interactions frequently require a retrieval of the key context information about the user and the environment. Image processing techniques have been widely applied in this area, providing details about recognized objects, people and actions. Considering remote diagnostics solutions, e.g. non-contact vital signs estimation and smart home monitoring systems that utilize person’s identity, security is a very important factor....

Pełny tekst do pobrania w portalu
Deep Learning: A Case Study for Image Recognition Using Transfer Learning
Publikacja
- S. Erpolat Tasabat
- O. Aydin
- Rok 2021
Deep learning (DL) is a rising star of machine learning (ML) and artificial intelligence (AI) domains. Until 2006, many researchers had attempted to build deep neural networks (DNN), but most of them failed. In 2006, it was proven that deep neural networks are one of the most crucial inventions for the 21st century. Nowadays, DNN are being used as a key technology for many different domains: self-driven vehicles, smart cities,...

Pełny tekst do pobrania w serwisie zewnętrznym

Wyszukiwarka

Filtry

Katalog

Wyniki wyszukiwania dla: VISUAL SPEECH RECOGNITION

Michał Lech dr inż.

Marek Blok dr hab. inż.

Piotr Szczuko dr hab. inż.