Geri Dön

Improved microphone array design with statistical speaker identification methods

İstatiksel ses tanıma metodları ile gelişmiş mikrofon dizisi tasarımı

  1. Tez No: 433000
  2. Yazar: KADİR ERDEM DEMİR
  3. Danışmanlar: DOÇ. DR. MUSTAFA TANER ESKİL, PROF. DR. MUSTAFA KARAMAN
  4. Tez Türü: Yüksek Lisans
  5. Konular: Bilgisayar Mühendisliği Bilimleri-Bilgisayar ve Kontrol, Computer Engineering and Computer Science and Control
  6. Anahtar Kelimeler: Ses işleme, Speech processing
  7. Yıl: 2016
  8. Dil: İngilizce
  9. Üniversite: Işık Üniversitesi
  10. Enstitü: Fen Bilimleri Enstitüsü
  11. Ana Bilim Dalı: Bilgisayar Mühendisliği Ana Bilim Dalı
  12. Bilim Dalı: Belirtilmemiş.
  13. Sayfa Sayısı: Belirtilmemiş.

Özet

Mikrofon dizilerinin kazanc dizinin boyutlar n b uy ut urek art r labilir fakat kazanc art rmak i cin sens or eklemek cok maliyetlidir. Bu nedenle e ger ortamda yeterince alan olsa bile algoritma kar s kl g n art rarak kazanc art rma tercih edilir. Spektral dizi i sleme methodlar nda, odaklan lmak istenen ki sinin ve g ur ult un un bulundu gu posizyonlar n bilinmesi b uy uk avantaj sa glar. Geleneksel metodlar bu problemi istatiksel olmayan y ontemlerle c ozmeye cal s r. Ayr ca ses tan ma metodlar n n performanslar g ur ult u oran n y uksek oldu gu ortamlarda azal r. Bu gibi ortamlarda, mikrofon dizilerinin kullan lmas ses sinyalinin kalitesini art r r. Bu nedenlerde dolay , mikrofon dizileri ve ses tan ma metodlar birbirlerine katk sa glarlar. Bu cal smam zda, mikrofon dizisi sistemi ve ses tan ma sistemi tek bir sistemin par calar olarak tasarlanm st r. Mikrofon dizisi kullanarak ses tan ma sisteminin do grulu gu art l rken ses tan ma sisteminin sonu clar kullan larakta mikrofon dizisinin kazanc art r lm st r. Ses tan ma sistemi uygulumas nda Fusion ve N-Gram temel frekans y ontemleri onerilmi stir Geli smi s mikrofon tasar m n g osterebilmek i cin simulasyon ortam konu smac lar n odan n herhangi bir yerine eklenebilice gi bir simulasyon ortam geli stirilmi stir. Simulasyon ortam nda deneyler sonu cu onerilen metodlar n geleneksel metodlar ust un oldu gu g ozlemlenmi stir.

Özet (Çeviri)

Conventional microphone array implementations aim to lock onto a source with given location and if required, tracking it. This implementation is straightforward when the location or the path of the source and interference are provided. It becomes a challenge to detect the intended source when multiple unknown sources exist in the same environment. Performance of speaker identi cation degrades drastically when the speech signal is severely distorted by additive noise and reverberation. In such environments, microphone arrays are often utilized as a means of improving the quality of captured speech signals. Both microphone array and speaker identi cation are mature elds. The advances of these two distinct elds can be combined into one system that maximizes gain on the intended speaker, which is the topic of this thesis. We utilize microphone array methods to improve the accuracy of speaker identi cation in a cocktail party environment. When the source and interferences are localized microphone array can be tuned to further reduce noise and increase the gain. In this thesis we developed a robust simulation environment to demonstrate the proposed improved microphone array design with statistical speaker identi cation. This is an open source implementation in which users can assign speakers anywhere in the room. We proposed two features; fusion based, and computationally e cient N-Gram for speaker identi cation. We demonstrated that the proposed features and the algorithm that leverages the synergy of microphone array processing and speaker identi cation methods outperforms conventional algorithms.

Benzer Tezler

  1. Fourıer bölgesinde çapraz ilinti yöntemi ile ses kaynağının konumunun fpga kullanarak tespiti

    Sound source localization in the fourier region with cross correlation method using fpga

    MERVE ÖZTÜRK LAFCI

    Yüksek Lisans

    Türkçe

    Türkçe

    2019

    Elektrik ve Elektronik MühendisliğiGazi Üniversitesi

    Elektrik-Elektronik Mühendisliği Ana Bilim Dalı

    DOÇ. DR. HASAN ŞAKİR BİLGE

  2. Bir Türkçe sesli ifade tanıma sisteminin kural tabanlı tasarımı ve gerçekleştirimi

    Rule based design and implementation of a speech recognition system for Turkish language

    ERHAN MENGÜŞOĞLU

    Yüksek Lisans

    Türkçe

    Türkçe

    1999

    Bilgisayar Mühendisliği Bilimleri-Bilgisayar ve KontrolHacettepe Üniversitesi

    Bilgisayar Bilimleri Ana Bilim Dalı

    YRD. DOÇ. DR. HARUN ARTUNER

  3. Performance evaluation of real-time noisy speech recognition for mobile devices

    Mobil cihazlarda gerçek zamanlı gürültülü konuşma tanıma performans değerlendirilmesi

    YASER YURTCAN

    Yüksek Lisans

    İngilizce

    İngilizce

    2019

    Bilim ve TeknolojiOrta Doğu Teknik Üniversitesi

    Bilişim Sistemleri Ana Bilim Dalı

    DOÇ. DR. BANU GÜNEL KILIÇ

  4. FPGA tabanlı şifreli kablosuz haberleşme sistemi

    FPGA based encrypted wireless communication system

    ILGAZ AZ

    Yüksek Lisans

    Türkçe

    Türkçe

    2014

    Elektrik ve Elektronik Mühendisliğiİstanbul Teknik Üniversitesi

    Ana Bilim Dalı (disiplinlerarası)

    DOÇ. DR. GÖKHAN İNALHAN

  5. Ego noise estimation for robot audition

    Başlık çevirisi yok

    GÖKHAN İNCE

    Doktora

    İngilizce

    İngilizce

    2011

    Makine MühendisliğiTokyo Instıtute Of Technology

    PROF. JUNİCHİ IMURA