Skip to main content

Databricks

AI Engineer — GTM Analytics

United StatesFrom $201.7kmidAdded today

About this role

Databricks seeks an AI Engineer to build intelligence and agentic workflows for their GTM Analytics platform, which serves 6,000+ field users. You'll develop Genie-powered agents, LLM-driven features, and applications using Python, SQL, and AI coding tools while working closely with senior engineers in a high-ownership, fast-shipping environment.

What you'll do

  • Build and ship AI-powered features including Genie analytics agents and LLM-assisted workflows
  • Develop agentic workflows with prompts, tool integrations, retrieval, and evaluation mechanisms
  • Write Python and SQL against Databricks lakehouse (Unity Catalog, Delta) and Postgres services
  • Leverage AI coding tools as core workflow and help build reusable skills and harnesses
  • Measure and improve AI output accuracy through eval sets, scorecards, and regression testing
  • Collaborate across data engineers, app engineers, and business stakeholders from concept to production

What they're looking for

  • Python programming
  • SQL and database design
  • LLMs and generative AI
  • Agentic AI systems and prompt engineering
  • RAG, knowledge graphs, and retrieval systems
  • Databricks and lakehouse architecture
  • AI evaluation and quality metrics
  • Postgres and backend services
Apply on the employer's site

Opens the official application on the employer’s site. No login required.

Databricks

Databricks builds a unified data and AI platform that combines database systems, distributed computing, and generative AI capabilities across multi-cloud infrastructure. The company is hiring software engineers, applied AI engineers, and web engineers to develop core database engines, ML/AI features, inference systems, and user-facing products.

View all jobs at Databricks

Likely interview questions

  • Describe a time you built an LLM-based application or agent system—what was the architecture and how did you handle hallucinations or accuracy issues?
  • How do you approach evaluating and measuring the quality of AI-generated outputs in production?