Skip to main content

OneBrief

Outcome Engineer - Early in Career Professional

United States | Remote (Remote)$120k–$200kfulltimemidAdded 1 month ago

About this role

Onebrief is seeking an early career Outcome Engineer to help shape the future of software development by transforming AI capabilities into tangible applications. This role involves collaborating on the development of multi-agent systems and automated governance processes within a fully remote environment, catering primarily to military staff.

What you'll do

  • Architect multi-agent systems for collaborative task execution
  • Implement automated governance and policy engines
  • Engineer context and memory for agents
  • Build evaluation frameworks for deterministic output validation
  • Design self-healing loops for system reliability
  • Prototype and deploy new tools for internal teams and customers

What they're looking for

  • Strong software engineering skills
  • AI research and application
  • Experience in game or simulation engineering
  • Product management expertise
  • Understanding of platform and site reliability engineering
  • Interdisciplinary collaboration

Benefits

  • Fully remote work environment
  • Opportunity to work with military staff
  • Impactful role in a rapidly growing company
  • Work in a diverse and collaborative team
  • Engagement in cutting-edge technology
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

OneBrief

OneBrief builds AtomEngine, an AI-powered cloud-based simulation platform designed for military training and operational exercises in classified and live environments. The company is hiring Software Development Engineers in Test, Gameplay Engineers, and Customer Success Engineers to support quality assurance, platform development, and direct customer deployment and support.

View all jobs at OneBrief

Likely interview questions

  • Walk us through a recent project where you built or experimented with agentic AI systems or multi-agent workflows. What challenges did you encounter, and how did you approach them?
  • How do you think about evaluating whether an AI agent is actually working correctly versus just producing plausible-looking outputs? What metrics or frameworks would you use?