Skip to main content

Reply

Data Engineer - GenAI

Chicago, Illinoisfull timemidAdded today

About this role

Develop and maintain data pipelines and infrastructure to support generative AI applications. This role focuses on building scalable data systems, optimizing data workflows, and ensuring data quality for AI model training and deployment in a Chicago-based environment.

What you'll do

  • Design and build ETL/ELT pipelines for AI model training datasets
  • Optimize data infrastructure for high-volume, low-latency processing
  • Implement data quality monitoring and validation frameworks
  • Collaborate with ML engineers to prepare and transform training data
  • Manage data storage solutions and database performance
  • Document data lineage and maintain data governance standards

What they're looking for

  • ETL/ELT pipeline development
  • SQL and database optimization
  • Python or Scala programming
  • Cloud platforms (AWS, GCP, or Azure)
  • Apache Spark or similar big data frameworks
  • Data modeling and schema design
  • Version control and CI/CD practices
  • Generative AI concepts and LLM workflows
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Reply

Reply builds enterprise software solutions across identity and access management, cloud infrastructure, and mobile/web applications, with a focus on client relationships and secure, modern technology stacks. The company is hiring software engineers, developers, and infrastructure specialists in areas including iOS development, C# and React, IAM platforms, and DevOps/Kubernetes.

View all jobs at Reply

Likely interview questions

  • Describe your experience building ETL pipelines for machine learning workflows.
  • How have you optimized data processing for large-scale AI model training?