
Closed
Posted
I’m building an Amharic speech-recognition system that reliably understands short voice commands on both mobile devices and desktop computers. To reach production quality I need help gathering or generating a robust command-focused dataset, training a low-latency model, and packaging everything so it can run efficiently on-device. Here’s what I need from you: • Collect or record clean Amharic voice-command audio with good speaker and acoustic diversity, then label each clip in UTF-8 Amharic text. • Pre-process the corpus (16 kHz, mono, noise-reduced) and create train/validation/test splits. • Train and fine-tune an acoustic and language model—Kaldi, Wav2Vec 2.0, Whisper, or similar—so it recognises common control phrases instantly. • Optimise and export an inference-ready model that runs offline on an average Android handset as well as standard PC CPUs. • Provide all training scripts, configuration files, and usage instructions in a Git repository so I can reproduce and extend the work. Acceptance criteria 1. ≤ 10 % word-error rate on a held-out test set of 1–5 second commands. 2. Startup latency under 500 ms and per-command recognition under 150 ms on target hardware. 3. Complete, well-commented source code, data splits, and README delivered. If you have prior experience with Amharic or other low-resource languages and can meet the above metrics, I’d love to collaborate.
Project ID: 40662240
5 proposals
Active 3 days ago
Location: Ethiopia
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs
5 freelancers are bidding on average $6 USD/hour for this job

Dear Client, I am a native Amharic speaker with extensive, hands-on experience in Amharic speech data collection, precise voice recording, and UTF-8 audio transcription. I can help you build and pre-process a clean, high-quality Amharic voice command dataset for your speech-recognition system. How I can assist with your project: Clean Data Collection & Recording: I can record clear Amharic voice command clips across diverse acoustic environments and varied speakers. Accurate UTF-8 Labeling: Every audio clip will be accurately transcribed in UTF-8 Amharic text to ensure clean training inputs. Audio Pre-processing: I will format audio files according to your exact specifications (16\text{ kHz}, mono, noise-reduced) and organize them into balanced train/validation/test splits. Reliable Delivery: I pay close attention to audio quality and guidelines to help you meet your target metrics (\le 10\% WER). As a native speaker based in Ethiopia, I bring native accuracy to Amharic linguistic details and command phrasing. I am ready to start immediately and commit 40 hours per week to complete this project efficiently. Looking forward to working with you! Best regards, Tariku Mamo
$5 USD in 40 days
0.0
0.0

Hello! I am a native Amharic speaker with a strong academic background and practical experience in data processing, voice transcription, and AI data collection. I am very interested in this Amharic Voice Command project and can ensure high-accuracy, clean, and well-structured results following all your guidelines. I am ready to start immediately and deliver quality work on time. Let's discuss further!
$5 USD in 50 days
0.0
0.0

Hi there, I am a Software Engineering graduate with a strong foundation in Python, machine learning concepts, and audio processing. While I am early in my professional career, I have extensive academic experience building end-to-end AI projects, handling datasets, and writing clean, documented code. I understand the complexity of this Amharic speech-recognition project, especially the challenges of low-resource languages. Here is my structured plan to help you achieve your goals: 1. Data Collection & Labeling: I will help organize the collection of diverse Amharic voice commands and strictly label them in UTF-8 Amharic text. I will ensure the audio is standardized to 16kHz mono. 2. Model Fine-Tuning: I will research and experiment with pre-trained models like Wav2Vec 2.0 or Whisper, focusing on fine-tuning them specifically for your command list to achieve the <10% Word Error Rate. 3. Optimization for On-Device Use: I will use quantization and pruning techniques to compress the model so it runs under 500ms latency on an Android device and standard PCs. 4. Clean Delivery: I will provide all training scripts, configuration files, and a clear README file in a Git repository so you can easily reproduce the work. I am a fast learner, highly responsive, and dedicated to delivering quality work on time. I am eager to take on this challenge and help you build a reliable Amharic voice AI. Looking forward to collaborating with you!
$12 USD in 40 days
0.0
0.0

Hello, Your project caught my attention because accurate Amharic speech recognition requires understanding natural Ethiopian Amharic pronunciation, accents, and everyday speech patterns. I am a native Amharic speaker and can help with testing and evaluating short Amharic voice commands, identifying recognition errors, and providing accurate feedback to improve the system. I can carefully follow your guidelines and deliver consistent results. I am also comfortable working with mobile-based tasks and audio recordings. If needed, I can provide sample recordings or complete a small test task first so you can evaluate the quality of my work. Could you please tell me approximately how many voice commands or hours of recordings you need processed? Thank you.
$5 USD in 40 days
0.0
0.0

I am Abdulkerim Seid Endris, an ideal candidate for your Amharic Voice Command AI project. Being a native Ethiopian English and Amharic speaker, I possess an innate understanding of the language patterns that can contribute significantly to gathering or generating a robust command-focused dataset. Furthermore, my MSc in Computer Engineering equips me with the technical expertise necessary to preprocess your audio corpus, train the acoustic and language model, and optimize the inference-ready model for offline use. My skills in Python, AI Development, and AI Training Data development would come into play here. I employ such skills to ensure minimal word-error rate for speech recognition on short voice commands, lightning-fast startup latency and per-command recognition times on different target hardwares. Plus, you can rely on me to deliver all training scripts, configuration files, and usage instructions in a comprehensive Git repository. In conclusion, my proven track record in audio transcription, speech data annotation, and linguistic evaluation sits well within the parameters of your project. Combined with my strong listening abilities and written accuracy, I am poised to help build an Amharic speech-recognition system that meets production quality standards on mobile devices and desktop computers. Feel free to leverage my bilingual prowess as I offer not only proficiency in Amharic but also Ethiopian English for any possible future needs!
$5 USD in 40 days
0.0
0.0

East Gojjam, Ethiopia
Member since Aug 21, 2026
₹750-1250 INR / hour
₹1500-12500 INR
₹600-1500 INR
min ₹2500 INR / hour
$2-8 USD / hour
$15-25 USD / hour
€2-3 EUR / hour
$250-750 USD
₹12500-37500 INR
₹12500-37500 INR
₹1500-12500 INR
$750-1500 USD
$2000-6000 HKD
€30-250 EUR
$100-250 USD
$250-750 USD
$15-80 USD / hour
$2-8 USD / hour
₹1500-12500 INR
$2-8 USD / hour