LIFEWOOD
Finalizing091
Type B — Horizontal LLM Data

Type 
Horizontal LLM Data

Comprehensive AI data solutions that cover the entire spectrum from data collection and annotation to model testing. Creating multimodal datasets for deep learning, large language models.

Voice content spans 6 project types and 9 data domains across 23 countries

25,400 valid hours of annotated multilingual speech data for large language model training

01

01 / 03
01

TARGET

Target

Collection
Target

Capture and transcribe recordings from native speakers from 23 different countries (Netherlands, Spain, Norway, France, Germany, Poland, Russia, Italy, Japan, South Korea, Mexico, UAE, Saudi Arabia, Egypt, etc.). Voice content involves 6 project types and 9 data domains. A total of 25,400 valid hours durations.

23 Countries6 Project Types9 Data Domains25,400 Valid Hours