Fetching the latest programs, projects, and workspace data.

A free/open-source machine translation platform
Showing 5 of 56 projects. Click any project card for scope, mentors, and proposal studio.
Mentors: Student: Chaitanya Gambali
The aim is to construct bidirectional dictionaries for a language pair, given a pair of parallel corpora - i.e., the same content in two different languages using a single script that does the job for the user in a single step, unlike the current method that requires multiple steps to do the job. The final deliverable is a Python script that does the job in one step.
Mentors: Student: Pedro Manicardi
This project aims to add the Capitalization Handling Module for the es-pt language pair, as well as to create more rules expanding the tool. After that, capitalization will be moved to monolingual modules to improve system performance. As deliveries, will be done: 1. Uptate the language pair with the tool/ 2. Expansion in capitalization restoration/ 3. Documentation explaining how to use/ 3. Sets of tests ensuring the efficiency and opetarion of the capitalization handling module.
Mentors: Student: Hari Krishna Reddy
Developing a spell-checking interface for Apertium's web tools to enhance user experience and accessibility. The project aims to update the existing interface, integrate Voikko and Libvoikko support, and improve error detection and correction across multiple languages
Mentors: Student: Natasha Singh
I will be working on creating a morphological dictionary for Kumaoni language and then use it for implementing a morphological analyzer. There are very few digital texts on Kumaoni language and no digital dictionary for it is available.
Mentors: Student: Robert Pugh
My idea is to develop a language pair for Highland Puebla Nahuatl (`azz`) and Western Sierra Puebla Nahuatl (`nhi`). Both are endangered variants of Nahuatl (`nah`, but different branches), and are in contact where they are spoken. `azz` is a higher-resource language with publications and government materials, whereas `nhi` is substantially more endangered with only some short stories available. There are monolingual repositories for both, so I'm interested in developing the bilingual dictionaries and transfer rules. This would really nice to have, especially since speakers of both languages often live in the same or neighboring areas, and there is interest in being able to translate materials from one variant to another.