Turn this role into an interview — a resume and cover letter built around what this employer wants.
Kaya by Chemin Sdn Bhd is hosting a short-term remote transcription project focused on Cantonese-English audio with healthcare terminology. You will listen, transcribe, segment, label speakers, and add basic annotations while maintaining high accuracy. Fluency in Cantonese and English is required, with Traditional Chinese writing ability.
Work is remote with flexible hours, about 4 hours daily, over a 2-month period (July 20 – September 2026). Onboarding pre-assessment is required.
Kaya by Chemin Sdn Bhd is a community for high-performing data annotators who play an integral role in shaping the future of machine learning and artificial intelligence.
Kaya offers a collaborative environment where ambitious annotators can thrive. It is a tight-knit community that supports members' professional growth and helps them build a long-term career in data labelling and AI.
Join Kaya and start contributing to impactful AI projects.
We are looking for fluent Cantonese speakers and writers who have a healthcare background or are familiar with medical terminology to join a short-term remote project helping train AI language systems. The recordings in this project are in Cantonese-English (code-switching), meaning speakers naturally switch between Cantonese and English within the same conversation. As such, fluency in both Cantonese and English is required.
Your task in this project is to listen to Cantonese-English audio recordings and transcribe them accurately. Many of these recordings contain healthcare-related conversations, making familiarity with medical terminology essential for producing high-quality transcriptions. Your work will help AI better understand authentic Cantonese-English speech across different speakers, accents, and speaking styles.
If you're comfortable working with audio, detail-oriented, and can commit a few hours a day, this project is for you.
Bonus point: Previous experience in transcription, audio annotation, data labelling, linguistics, healthcare-related language work, or AI-related projects.