Lacandon Speech to Text: A Complete Guide
Lacandon Speech to Text: Preserving a Mayan Language with AI
The Lacandon Language Today
Lacandon (Lakant'un) is a Mayan language belonging to the Yucatecan branch, closely related to Yucatec Maya and Mopan. It is spoken by around 1,000 people in the Lacandon Jungle of Chiapas, Mexico, with a few speakers in Guatemala. The language is written in a Latin-based orthography that became standardized in the late 20th century. Despite its small speaker population, Lacandon carries a rich oral tradition of myths, prayers, and historical narratives that are at risk of disappearing.
Why Accurate Speech-to-Text for Lacandon Matters
Transcribing Lacandon audio by hand is impractical at scale. Linguists, community organizations, and educators need a tool that can convert spoken words to text reliably. Automatic speech transcription for Lacandon enables:
- Quick creation of text archives from recordings of elders.
- Subtitle generation for videos in Lacandon.
- Support for language learning and dictionaries.
- Accessibility for deaf community members through captions.
Speechyou's Lacandon model delivers over 85% word accuracy on clear recordings, making these tasks feasible for the first time.
Specific Transcription Challenges in Lacandon
Unique Phonology
Lacandon features ejective consonants (p', t', k', tz') that are rare globally. Many ASR systems confuse them with plain stops. Speechyou's model uses fine-grained acoustic features to distinguish them.
Vowel Length Distinction
Short and long vowels are phonemic: for example, k'ak' (fire) vs. k'aak' (strong). Our system is trained on vowel duration cues.
Limited Data
With few transcribed hours available, building an ASR model from scratch is impossible. We leverage transfer learning from other Mayan languages and active learning to improve with each user upload.
Use Cases for Lacandon Transcription
- Oral history projects: Non-profits record stories from the last fluent speakers. Speechyou transcribes these files, creating searchable text that can be published or studied.
- Educational content: Teachers create bilingual Lacandon-Spanish subtitles for video lessons, helping children learn to read their ancestral language.
- Community radio: DJs upload pre-recorded shows and generate timestamped text for online archives.
- Linguistic research: Academics quickly analyze syntactic patterns without manual transcription.
- Film subtitling: Documentaries about the Lacandon people are subtitled in their own language, validating its role in the media landscape.
How Speechyou Helps
Speechyou is the only commercial speech-to-text platform offering a dedicated Lacandon model. You can upload audio or video and select 'Lacandon' from the language menu. The system processes the file and outputs text with punctuation, speaker labels, and timestamps. You can then export SRT or VTT subtitles or copy the plain text.
The platform handles the glottalized sounds and vowel length with higher accuracy than general models like Whisper. We also provide a confidence score per word, so you can quickly review uncertain sections. The first 30 minutes are free, and the Solo plan includes unlimited transcription, making it affordable for community projects.
Getting Started
To transcribe Lacandon audio, follow these steps:
- Sign up at Speechyou.com.
- Upload your audio or video file (MP3, WAV, MP4, etc.).
- Choose 'Lacandon' as the language.
- Start transcription and wait for the text to appear.
- Edit if needed, then download subtitles or text.
Your contribution also helps improve the model for future users. Each corrected transcript is used (anonymously) to retrain the system, boosting accuracy for everyone.
Conclusion
Lacandon speech to text is no longer a distant possibility. Speechyou brings modern AI to an endangered language, empowering speakers and researchers to document, preserve, and share their linguistic heritage. Try it today and give Lacandon a lasting digital voice.







