Guide to the five files in the language corpus section, covering text and speech datasets, natural language processing, lexicography, descriptive grammars, and language pedagogy and policy.
A comprehensive scholarly survey of documented Yorùbá lexical databases, parallel machine translation corpora, named entity benchmarks, and speech datasets.
A comprehensive scholarly survey of Yorùbá computational linguistics, covering automatic diacritic restoration, machine translation, speech recognition, speech synthesis, and community benchmarks.
The historical development of Yoruba dictionary-making from nineteenth-century missionary foundations through modern computational lexical databases.
A scholarly analysis of major reference grammars, foundational syntactic debates, and dialectological classifications in Yorùbá linguistics.
An analysis of Yorùbá language pedagogy, instructional materials, second-language tone acquisition research, and mother-tongue education policy in Nigeria.