Tiny Aya Vision
Extending Tiny Aya, a 3.35B multilingual model in 70+ languages, with lightweight visual capabilities through parameter-efficient fusion.
Open Science Community
Research notes, technical essays, and personal stories from the Cohere Labs Community
Research notes, stories, and ideas from people shaping AI together—transparent, collaborative, and community-led.
Extending Tiny Aya, a 3.35B multilingual model in 70+ languages, with lightweight visual capabilities through parameter-efficient fusion.
A student's journey from a frustrating ChatGPT session to building an open-source framework for deciding what an LLM should actually remember.
We built NILEAGI-SUB to measure Swahili understanding where people actually use it. Our first sector is education: 2,569 school questions, eight compact open models, and a ranking that size alone cannot explain.
Quality data for under-resourced languages is hard to acquire. It is even harder to get code in these languages. Through research, it showed that introducing code in training a model has proved to be a well rounded data source to improve performance. How can we utilize non-English code to strengthen the performance of these multilingual models, especially where data scarcity exists? We started with the question:- what if the code you train a language model on was written in a language other than English?
How a Luganda word-review task became a startup, a Gold Award, and a lesson about what really keeps gates closed for African builders, and what opens them.