Staff AI Engineer building reliable, observable agentic AI workflows for Acquia’s enterprise digital experience platform.
Architecting multi-agent systems with Python, LangGraph, Temporal, Pydantic, and LangFuse.
Responsibilities
Design, build, and ship production-grade agentic AI workflows across the Acquia DXP
Write and ship production AI code daily as an active contributor
Architect stateful, cyclic, multi-agent workflows using LangGraph, Temporal, and Pydantic
Own AI observability through LangFuse, including tracing, prompt versioning, evaluation, and performance benchmarking
Set engineering standards for agent design patterns, RAG, prompt management, context optimization, and tool-calling strategies
Partner with product and platform teams on AI architectures meeting enterprise SLA, security, and compliance requirements
Evaluate and adopt emerging LLM providers, orchestration frameworks, and agentic stack improvements
Mentor engineers through code reviews, pairing sessions, and design discussions
Represent Acquia’s AI capabilities in customer architectural reviews, technical discovery, and roadmap conversations
Requirements
8+ years of software engineering experience
3+ years of production experience with AI agents
Hands-on expertise with LangGraph, Temporal, and Pydantic
Hands-on expertise with LangFuse, including tracing, evaluation, prompt management, and dataset-driven testing
Proficiency with agent harness frameworks such as LangChain, LlamaIndex, or CrewAI
Deep Python proficiency
Strong engineering fundamentals in testing, CI/CD, and architecture
Cloud AI deployment experience with AWS, Azure, or GCP
Experience with containerization and inference cost management
Knowledge of RAG architecture, vector databases, embedding models, and retrieval strategies
B.S. in Computer Science or equivalent practical experience
Enterprise SaaS or CMS experience, including familiarity with Drupal-based DXP
Experience with AI-assisted coding tools such as Copilot, Cursor, or Claude
Familiarity with persistent agent runtimes such as OpenClaw and Hermes Agent
LLM fine-tuning or model evaluation experience
Understanding of human-in-the-loop, interrupt-driven agents, and enterprise design
Strong communication skills for presenting AI system design to engineers and C-suite stakeholders
Senior individual-contributor track record with high-quality code and system designs
AI fluency, orchestration mindset, radical adaptability, builder mentality, and intellectual humility
Benefits
Competitive healthcare coverage
Wellness programs
Take it when you need it time off
Parental leave
Recognition programs
Inclusive, transparent, efficient, and educational interview experience
Technical lead building safe, evaluable LLM agents and AI infrastructure for OpenLoop’s telehealth platform. Setting architecture, observability, retrieval, and model operations direction.
Senior software engineer evaluating AI coding agents such as Codex and Claude Code. Providing rigorous written and video feedback on engineering quality.
Senior software engineer evaluating AI coding agents for G2i. Assessing engineering judgment, explanations, and trustworthiness across Codex, Claude Code, and Cursor.
Senior engineer evaluating AI coding agents such as Codex, Claude Code, and Cursor. Providing rigorous written and video feedback on engineering judgment, reasoning, and interaction quality.
Forward Deployed Engineer building full - stack AI solutions on AWS and Azure for Huron’s consulting clients. Partnering daily with business users to deliver tested features.
Staff AI Security Engineer securing EQ Bank’s enterprise AI, machine learning, and cloud platforms. Building guardrails, controls, automation, and detection capabilities for Canada’s Challenger Bank.
Staff Software Engineer owning OAuth, authorization, and agent delegation systems. Building governed identity infrastructure for Redpanda’s enterprise AI data platform.
Senior software engineer evaluating AI coding agents for G2i’s engineering team. Assessing reasoning, explanations, and engineering judgment in Codex, Claude Code, and Cursor interactions.
Senior software engineer evaluating AI coding agents such as Codex, Claude Code, and Cursor. Providing rigorous written and video feedback on engineering quality.
Senior engineer evaluating Codex, Claude Code, and Cursor interactions for G2i. Providing rigorous written and video feedback on AI - generated engineering work.