Skip to main content

Andromeda

Developer Relations Engineer

North America Remote / San Francisco, CA (Remote)fulltimemidAdded today

About this role

Andromeda seeks a Developer Relations Engineer to translate internal knowledge into scalable artifacts—benchmarks, documentation, and code—that help customers successfully run distributed AI workloads on the platform. This is an engineering-focused role reporting to the Head of Product, combining hands-on coding with community building and direct engagement with customer infrastructure challenges.

What you'll do

  • Write production-grade code including reference implementations, integration examples, and tools that reduce friction for new users
  • Create technical content: benchmarks, performance analysis, architecture guides, and postmortems backed by reproducible measurements
  • Develop and maintain comprehensive documentation covering quickstarts, orchestration (Slurm, Kubernetes), storage, and troubleshooting
  • Engage with developer community through public channels, GitHub, office hours, and technical problem-solving
  • Partner with solutions and product engineering teams on customer onboardings and incidents, translating learnings into public artifacts
  • Advocate internally for product improvements and missing platform capabilities based on real-world workload feedback

What they're looking for

  • Python development for production systems
  • Distributed training and large-scale inference operations and debugging
  • PyTorch, NCCL, and CUDA-adjacent tooling
  • Container technologies and orchestration (Kubernetes, Slurm)
  • Go or Bash scripting
  • LLM serving frameworks (vLLM or similar)
  • Technical writing and documentation
  • Infrastructure troubleshooting and performance analysis
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Andromeda

Andromeda is an AI infrastructure startup that builds hardware and software platforms for large-scale AI training and inference, operating a compute marketplace that routes jobs across multiple cloud providers. The company is hiring Solutions Engineers, Site Reliability Engineers, and Infrastructure Platform Engineers to deliver customer solutions, manage Kubernetes clusters, and design core orchestration systems.

View all jobs at Andromeda

Likely interview questions

  • Describe a time you debugged a multi-node distributed training failure—what was wrong and how did you track it down?
  • Walk us through how you would create a reference implementation for a popular orchestration framework on a heterogeneous compute platform.