We're building a small contract team to take Kwaghfan from corpus to deployed product.
Own participant recruitment, consent, and recording logistics for our Tiv data collection on the ground in Benue. You'll schedule and run recording sessions, obtain and log consent, capture audio to spec, and keep the metadata clean. You should be organized, trustworthy handling consent and payments, and comfortable with a smartphone recorder and simple offline mobile forms (ODK/KoBoToolbox).
Transcribe recorded Tiv audio into accurate, standard-orthography text, preserving tone diacritics where they occur naturally. You should be genuinely literate in written Tiv, careful, and reliable. No computer science background is needed. You won't transcribe your own recordings.
Review collected items against a fixed quality checklist and accept or reject each with a reason code. You should have strong command of standard written and spoken Tiv. Older speakers are especially valued for judging the authenticity of folktales and proverbs. You're paid for every item you review, whether or not it's accepted.
Support data cleaning, processing, evaluation, and documentation of the Tiv corpus. Ideal for final-year students or recent graduates with Python skills and a genuine interest in NLP and African language technology. You'll learn the full pipeline hands-on and contribute to published research. This is a fixed three-month program that begins after a successful data-collection pilot.
The field and language roles require native Tiv. For the intern roles, Tiv is a strong advantage but not required.
Tell us a little about your background, skills, and language competence, it takes about five minutes. We review every application and reach out to shortlisted candidates with next steps.