Data Engineer - Equipo de LLMs Training - IT
Mercado LibreWhat You'll Do
Build and maintain data ingestion and processing pipelines at scale for training LLMs. Implement data curation processes such as quality, deduplication, scoring, and diversification with autonomy in engineering decisions. Scale the generation of synthetic data using open source and commercial models. Collaborate with Data Scientists and ML engineers to translate training needs into concrete pipelines.
What We're Looking For
Experience in text data engineering using Python, SQL, BigQuery, and large-scale ETL pipelines. Apply engineering best practices focused on testing, versioning, and reproducibility. Demonstrate autonomy in managing data pipelines from start to finish. Interest in the LLM and deep learning ecosystem to understand the impact of data on models.
What We Offer
Hybrid work modality. Opportunity to work on challenging and innovative projects that have a real impact on e-commerce and financial services in Latin America.
Key Skills & Technologies
Additional Information
Experience Level
Mid-Level
Job Language
Spanish
Employment Type
Full-time
Work Mode
Hybrid