K

Karya

KARYAA5 · direct accessManage supplier listing

Indian-language AI datasets + custom collection: conversational speech datasets spanning 22 official Indian languages, egocentric work/life datasets for embodied AI, and the Samiksha multilingual eval benchmark (6 Indian languages, 17 models, 4 domains); plus domain transcription, localized translation, and multimodal dataset creation services

A sample dataset retrieved from the supplier.

No sample attached yet

Request one and the supplier can attach it here.

Request sample

Universe, instruments and categories.

Categories
multilingual data collection/annotation services + off-the-shelf Indian-language datasets
Regions
APAC
Sample tickers
TCS.NSINFY.NSWIPRO.NSHCLTECH.NSTECHM.NS

Karya has not added their own details yet. Not yet on file:

  • Dataset
  • Data dictionary
  • Coverage
  • Provenance
  • Rights & data handling