Conceptio
›
speech-recognition
Topic
speech-recognition
Knowledge-graph topic
· documents ABOUT speech-recognition across the archive
22
Documents about speech-recognition
Documents about speech-recognition
Unsupervised Speech Recognition with N-Skipgram and Positional Unigram Matching
#433629
KOASAS: KAIST Open Access Self-Archiving System
Whisper-Aware LLM: Self-Supervised Uncertainty Learning for Robust Whispered Speech Recognition
#444691
arXiv CS
Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation
#809825
arXiv (OAI Expanded)
Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation
#819821
arXiv (All)
Curriculum-Based Noise Adaptation for Phoneme-to-Text Reconstruction in Visual Speech Recognition
#1034816
arXiv (All)
Deep learning-based bimodal speech and facial expression recognition of miners' unsafe emotions.
#193406
NCBI PubMed Central
Context-Dependent Pre-Trained Deep Neural Networks for Large-Vocabulary Speech Recognition
#303010
Semantic Scholar
LI-TTA: Language Informed Test-Time Adaptation for Automatic Speech Recognition
#433630
KOASAS: KAIST Open Access Self-Archiving System
Conjoint Audio-to-Spikes Encoding and Processing for Efficient Neuromorphic Speech Recognition
#622545
arXiv CS
A 1D-CNN with advanced data augmentation for robust speech emotion recognition.
#659806
NCBI PubMed Central
SISER: Speaker-Invariant Speech Emotion Recognition with Entropy-Based Adversarial Training
#921088
arXiv (OAI Expanded)
SISER: Speaker-Invariant Speech Emotion Recognition with Entropy-Based Adversarial Training
#925324
arXiv (All)
MedWER: A Reproducible, Model-Free Evaluation Protocol for Medical Speech Recognition
#972776
arXiv (All)
AI-Mediated Input Transformation in Medical English: Speech Recognition and Transcript Reliability
#989250
OSF
Leveraging Fine-grained Error Correction in Korean Speech Recognition for Consultation Services
#996702
arXiv (All)
Leveraging Unlabeled Audio-Visual Data in Speech Emotion Recognition using Knowledge Distillation
#974203
arXiv (All)
Multi-datasets for different keyboard key sound recognition
#199471
OpenAlex
Socio-technical risks of clinical speech-to-text systems: Transparency, privacy, and reliability challenges in AI-driven documentation.
#136067
PubMed
Detecting Alzheimer’s Disease through Chinese Speech: A Contrast of Manual and Automatic Linguistic Feature Extraction
#368857
OSF
Measuring Equality in Machine Learning Security Defenses: A Case Study in Speech Recognition
#133948
Semantic Scholar
An effective two-stage key frame extraction method for speech-visual emotion recognition.
#342486
NCBI PubMed Central
Implementation of the Standard I-vector System for the Kaldi Speech Recognition Toolkit
#448961
Infoscience
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
#611275
arXiv (OAI Expanded)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
#612752
arXiv (All)
Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems
#808901
arXiv (OAI Expanded)
Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems
#818893
arXiv (All)
Backdoor Attacks on Speech Emotion Recognition via TTS-Generated Poisoning
#821925
arXiv (All)
SpeakPay: Domain-Adaptive LoRA Fine-Tuning of Whisper for Low-Resource Nepali Financial Speech Recognition
#822167
arXiv (All)
Dual-Form ASR: Semantics-Aware Inverse Text Normalization for Chinese Speech Recognition
#921048
arXiv (OAI Expanded)
Dual-Form ASR: Semantics-Aware Inverse Text Normalization for Chinese Speech Recognition
#925284
arXiv (All)
AVSRBench: A Multi-Condition AVSR Benchmark
#997256
arXiv (All)
Replication Data for "The Mason-Alberta Phonetic Segmenter: A forced alignment system based on deep neural networks and interpolation"
#377866
George Mason University Dataverse Dataverse OAI Archive
Convolutional Neural Networks for Sentence Classification
#6459
OpenAlex
The effects of intelligibility on conflict resolution and monitoring during speech recognition in noise.
#161821
NCBI PubMed Central
KANWhisper: leveraging learnable activation functions for interpretable and efficient arabic automatic speech recognition.
#371730
NCBI PubMed Central
Short Forms and Computerized Adaptive Tests With Monosyllabic Words Can Efficiently Measure Speech Recognition.
#472763
NCBI PubMed Central
Associations Between Self-Reported Workload and Measures of Speech Recognition in Adults Across the Lifespan.
#670937
NCBI PubMed Central
amphion/Emilia-Dataset
#786572
Hugging Face Datasets
FedEmoNet: Privacy-preserving federated learning with TCN-Transformer fusion for cross-corpus speech emotion recognition.
#161873
NCBI PubMed Central
RTP Payload Format for European Telecommunications Standards Institute (ETSI) European Standard ES 201 108 Distributed Speech Recognition Encoding
#193180
IETF RFCs
Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation
#214314
Papers With Code
A large-scale Wi-Fi channel state information dataset for contactless human speech recognition.
#389016
NCBI PubMed Central
ADAMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
#436428
KOASAS: KAIST Open Access Self-Archiving System
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
#609765
arXiv (All)
RoboGesture: Real-Time Semantic-aligned Co-Speech Gestures Generation for Humanoid Interaction
#788937
arXiv (OAI Expanded)
RoboGesture: Real-Time Semantic-aligned Co-Speech Gestures Generation for Humanoid Interaction
#792941
arXiv (All)
When Speech Meets Lips: Interpretable Audio-Visual Synchronization for L2 Pronunciation Assessment
#974876
arXiv (All)
WhisTLE: Deeply Supervised, Text-Only Domain Adaptation for Small Pretrained Speech Recognition Transformers
#984069
arXiv (All)
GestureFAR: Streaming Co-Speech Gesture Generation with Flow Autoregression
#1035972
arXiv (All)
Estimating Hearing Sensitivity Using Speech Recognition Thresholds in PART
#151618
OSF
← Previous
Page 3 of 5
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.