Skip to content
akaturk Academic measurement

Article detail · 2008

Spoken Term Detection for Turkish Broadcast News

OpenAlex Citations 79 Top 10% Percentile 97.9% FWCI 7.05
Year
2008
Type
conference-paper

Data source split

  • OpenAlex OpenAlex enrichment (abstract, citations, topics)

Abstract

OpenAlex · English

In this paper, we present a baseline spoken term detection (STD) system for Turkish broadcast news. The agglutinative structure of Turkish causes a high out-of-vocabulary (OOV) rate and increases word error rate (WER) in automatic speech recognition. Several approaches are attempted to reduce this negative effect on the STD system. Sub-word units are used to handle the OOV queries and lattice-based indexing is used to obtain different operating points and handle high WER cases. A recently proposed method for setting term specific thresholds is also evaluated and extended to allow us to choose an operating point suitable for our needs. Best results are obtained by using a cascade of word and sub-word lattice indices with term-thresholding.

Topics

Citations

OpenAlex cited_by_count. Not a WoS or Scopus citation count; those sources have no separate column here.

79 citations

OpenAlex cited_by_count (cache / database)

Authors

No author information.