Skip to main content

Twitch

Software Engineer, Data Platform

San Francisco, CAFrom $160kmidAdded today

About this role

Twitch's Data Platform team seeks a Software Engineer to develop and operate data infrastructure handling 200+ billion events daily. You'll build scalable tools for data warehousing, pipelines, and analytics while collaborating with peers to advance the platform's capabilities.

What you'll do

  • Develop new features and capabilities for data warehouses and real-time pipelines
  • Enhance reliability, flexibility, and scalability of existing data tools and systems
  • Collaborate on architectural decisions as the platform scales to petabyte volumes
  • Integrate AI tooling to improve engineering workflows
  • Review and shape peer contributions through code review and technical collaboration
  • Operate and maintain highly available distributed systems

What they're looking for

  • Distributed systems design and fundamentals
  • Data modeling and pipeline architecture
  • API design and scalability patterns
  • Concurrency and parallelism
  • Python or Go programming
  • AWS cloud services
  • High-availability systems
  • AI-assisted coding tools

Benefits

  • Medical, dental, vision, and disability insurance
  • 401(k) matching
  • Flexible PTO
  • Maternity and parental leave
  • Amazon employee discount
  • Mental health support and EAP
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Twitch

Twitch builds a platform for streaming and creator communities, offering features around user safety, memberships, and commerce to support millions of concurrent users. The company is hiring Software Engineers across multiple teams to develop scalable systems, moderation tools, and features that enhance experiences for creators and viewers.

Website
twitch.tv
View all jobs at Twitch

Likely interview questions

  • Describe your experience designing or scaling data pipelines to handle high-throughput event streams.
  • How would you approach improving the reliability of a distributed system handling petabyte-scale data?