In this role, you'll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input.
Requirements
Proven expertise in big data engineering with hands-on experience building and maintaining large-scale data pipelines.
Advanced proficiency in Python for data processing, automation, and integration.
Deep understanding of relational and NoSQL databases, including optimization and management techniques.
Experience with distributed data processing frameworks (e.g., Hadoop, Spark, Flink).
Strong foundation in data modeling, ETL processes, and data warehousing principles.
This role is posted on our partner platform. When you click Apply, you'll go to the posting, where the application, interview, skill validation, and onboarding all happen. lehico is an independent site that surfaces these opportunities — we don't process applications or guarantee acceptance.