Baseten
AI Inference Engineer
About this role
Baseten is hiring an AI Inference Engineer to work as a Forward Deployed Engineer, partnering with high-growth AI companies to architect and deploy production AI applications on their platform. You'll own the full customer journey from initial exploration through production deployment, blending software engineering with product management and customer success.
What you'll do
- Develop and maintain production software systems using Python and general-purpose programming languages
- Partner with customers to design, implement, and deploy AI solutions end-to-end across sales, implementation, and expansion phases
- Translate ambiguous business goals into well-specified technical projects and rapidly deliver tested services
- Optimize AI/ML projects and contribute to platform feature development and product specifications
- Own products and customer projects end-to-end as engineer, project manager, and product manager
- Navigate ambiguity, exercise sound technical judgment, and maintain accountability for delivered outcomes
What they're looking for
- Python programming in production environments
- General-purpose programming languages
- AI/ML pipeline design and model deployment
- Technical communication and documentation
- Full-stack project ownership and execution
- Problem specification and requirements translation
- Performance optimization for ML systems
- Customer collaboration and stakeholder management
Benefits
- Competitive compensation with meaningful equity
- 100% medical, dental, and vision insurance coverage for employee and dependents
- Flexible work arrangements
- Exposure to cutting-edge AI companies and production-scale deployments
- Growth in both technical depth and product/business acumen
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Baseten
Baseten builds an AI inference platform that enables companies to deploy and manage machine learning models in production at scale. The company is hiring software engineers, infrastructure specialists, and product engineers to develop observability systems, frontend experiences, reliability infrastructure, developer tools, and enterprise deployment solutions.
- Website
- baseten.com
Likely interview questions
- Describe your experience deploying machine learning models to production. What were the key challenges you faced around latency, cost, or reliability?
- Tell us about a time you had to translate vague business requirements into a concrete technical specification. How did you approach the problem?