REFAppendix
References
Full bibliographic entries for every citation referenced inline on this site.
[K1]
Koçak, C. & Ulas, H. B. (2024). Implementation of a Whisper Architecture-Based Turkish Automatic Speech Recognition (ASR) System and Evaluation of the Effect of Fine-Tuning with a Low-Rank Adaptation (LoRA) Adapter on Its Performance
Electronics (MDPI), 13(21), 4227
https://www.mdpi.com/2079-9292/13/21/4227
DOI: 10.3390/electronics13214227
[K2]
Mozilla Common Voice — Türkçe derlem (2024). Türkçe konuşma derlemi (Corpus 15.0)
Mozilla Common Voice
[K3]
Codersera / Promptquorum (2026). faster-whisper / whisper.cpp / OpenAI Whisper karşılaştırması ve int8 niceleme etkileri
Technical comparison
https://codersera.com/blog/faster-whisper-vs-whisper-cpp-speech-to-text-2026/
[K4]
NVIDIA-AI-IOT (2024). whisper_trt — Jetson platformunda TensorRT ile hızlandırma
GitHub
[K5]
Asano, Y., Hassan, S., Sharma, P., Sicilia, A., Atwell, K., Litman, D. & Alikhani, M. (2025). Contextual ASR Error Handling with LLMs Augmentation for Goal-Oriented Conversational AI
arXiv:2501.06129
[K6]
Pidiylab (2024). Raspberry Pi üzerinde Piper TTS ve çevrimdışı TTS motorlarının karşılaştırması
Technical comparison
[K7]
Casanova, E. vd. (2024). XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model
arXiv:2406.04904
[K8]
Shulitskyi, I. (2023). Fonetik eşleştirme algoritmaları: Soundex, Metaphone, Double Metaphone ve dile özgü türevleri
Technical survey
https://medium.com/@ievgenii.shulitskyi/phonetic-matching-algorithms-50165e684526
[K9]
Wang, X., Liu, Y., Li, J., Miljanic, V., Zhao, S. & Khalil, H. (2022). Towards Contextual Spelling Correction for Customization of End-to-end Speech Recognition Systems
arXiv:2203.00888
[K10]
He, J., Yang, Z. & Toda, T. (2023). ED-CEC: Improving Rare Word Recognition using ASR Postprocessing Based on Error Detection and Context-Aware Error Correction
Proc. ASRU 2023, arXiv:2310.05129
[K11]
He, J. & Toda, T. (2025). PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding
IEEE/ACM TASLP, arXiv:2506.11064
[K12]
ACL Findings EMNLP 2025 (2025). Retrieval-Augmented Contextual ASR via Decoder-State Guided Retrieval
ACL Findings EMNLP 2025
[K18]
TÜBİTAK (2026). TÜBİTAK 2026-2028 Öncelikli Ar-Ge ve Yenilik Konuları
TÜBİTAK
https://tubitak.gov.tr/tr/kurumsal/politikalar/tubitak-2026-2028-oncelikli-ar-ge-ve-yenilik-konulari
[K19]
Radford, A., Kim, J. W., Xu, T., Brockman, G., McLeavey, C. & Sutskever, I. (2022). Robust Speech Recognition via Large-Scale Weak Supervision (Whisper)
arXiv:2212.04356
[K20]
Baevski, A., Zhou, H., Mohamed, A. & Auli, M. (2020). wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
arXiv:2006.11477
[K21]
Hu, E. J., Shen, Y., Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L. & Chen, W. (2021). LoRA: Low-Rank Adaptation of Large Language Models
arXiv:2106.09685
[K22]
Gholami, A., Kim, S., Dong, Z., Yao, Z., Mahoney, M. W. & Keutzer, K. (2021). A Survey of Quantization Methods for Efficient Neural Network Inference
arXiv:2103.13630
[K23]
Bisani, M. & Ney, H. (2004). Bootstrap Estimates for Confidence Intervals in ASR Performance Evaluation
IEEE ICASSP 2004
[K24]
Macháček, D., Dabre, R. & Bojar, O. (2023). Turning Whisper into a Real-Time Transcription System
IJCNLP-AACL 2023, arXiv:2307.14743
[K25]
Dean, J. & Barroso, L. A. (2013). The Tail at Scale
Communications of the ACM, 56(2), 74-80
[K26]
Levenshtein, V. I. (1966). Binary Codes Capable of Correcting Deletions, Insertions and Reversals
Soviet Physics Doklady, 10(8), 707-710
[K27]
Winkler, W. E. (1990). String Comparator Metrics and Enhanced Decision Rules in the Fellegi-Sunter Model of Record Linkage
Proceedings of the Section on Survey Research Methods, ASA
[K28]
Burkhard, W. A. & Keller, R. M. (1973). Some Approaches to Best-Match File Searching
Communications of the ACM, 16(4), 230-236
[K29]
Damerau, F. J. (1964). A Technique for Computer Detection and Correction of Spelling Errors
Communications of the ACM, 7(3), 171-176