IT green

Ethical Multilingual AI Data Platform

Description

Karya builds datasets, evaluations and language resources for artificial intelligence while structuring the work so that people from lower-income and linguistically diverse communities can participate in the AI economy. The Bengaluru organisation has created conversational speech and other datasets across Indian languages and also offers enterprise data services. Founder Manu Chopra has positioned the model around better compensation and broader participation in data creation, challenging the low-wage outsourcing model common in AI annotation. The innovation is both technical and organisational: high-quality multilingual data is produced alongside a system intended to distribute more of the economic value of AI work to the communities generating it.

Problem Addressed

AI systems need high-quality Indian-language datasets, but data work is often poorly paid and under-representative.

How It Works

Karya organises distributed data collection, speech recording, annotation and evaluation workflows.

Key Differentiator / Impact

Combines multilingual AI-data infrastructure with a worker-impact model.

Location
Video
Comment

Add new comment

Restricted HTML

  • You can align images (data-align="center"), but also videos, blockquotes, and so on.
  • You can caption images (data-caption="Text"), but also videos, blockquotes, and so on.
Business Info

Contact Form

Background body show when use boxed layout
You can upload image for /themes/gavias_lozin/images/patterns
x
x
x
x
x
x
x
x
x
x
x
x
x
x
x
Color skins
Body layout