Senior DevOps Engineer designing and implementing GenAI solutions at Financeit, driving innovation in financial operations. Collaborating with teams to leverage AI in solving complex problems.
Responsibilities
Design, develop, and implement GenAI solutions for various financial applications, including personalized recommendations, risk assessment, fraud detection, and automated reporting.
Design and implement AI Agents, utilizing rigorous harness engineering to build the necessary guardrails, memory management, and tool orchestration for safe and reliable execution.
Process and analyze large datasets of structured and unstructured data.
Architect and scale dynamic context retrieval systems (RAG) and semantic search infrastructures, leveraging tools like AWS Bedrock and LangGraph to securely ground AI outputs in enterprise data
Develop and refine advanced prompting strategies for LLMs.
Test, evaluate, and analyze the performance of LLM and other GenAI models.
Collaborate closely with engineering teams to deploy and maintain GenAI models in production environments, including containerization, CI/CD pipelines, and cloud infrastructure management.
Communicate effectively with business stakeholders.
Stay up-to-date with the latest advancements in GenAI research and development, including areas like Agentic AI.
Requirements
Bachelor’s degree in Computer Science, Software Engineering, or a related field
5 years+ of experience in DevOps, Software Development, or AI/ML development, with a proven track record of building and deploying sophisticated GenAI applications
Deep understanding of GenAI models, architectures and agentic AI concepts
Extensive experience of LLM architectures, prompt engineering, fine-tuning, evaluation, and deploying agents (Agentcore stack)
Expert Python skills, Vector Databases (Qdrant, Pinecone, pgvector), and RAG pipelines using LlamaIndex and LangGraph
Experience in MLOps, containerization (Docker/Kubernetes), CI/CD, and cloud infrastructure (AWS, Azure, or GCP)
Analytical problem-solver with a proven ability to effectively communicate and collaborate across departments and drive projects forward
**Strongly Preferred Experience: **
Experience with financial data and applications
Familiarity with chatbot development frameworks and best practices, including conversational AI design and natural language understanding (NLU)
Experience leading or contributing to complex data science or AI/ML projects in a fast-paced environment
Experience with data visualization and reporting tools
Experience with SQL databases
Benefits
An award-winning culture with a collaborative & inclusive team.
Competitive pay and performance-based bonus:
Annual Base Salary: $130,000 - 145,000
Annual Bonus: 20%
Committed to flexible work arrangements, offering hybrid workplace options.
Comprehensive medical, dental and vision coverage + Lifestyle Account.
RRSP Matching and Parental Leave Top UP Program.
In office massage, meditation & workout sessions.
Virtual events such as Lunch & Learns, company parties, fun team activities and charity initiatives.
Senior DevOps Engineer building secure, scalable Azure platforms for CARET’s legal and accounting practice - management software. Leading infrastructure, Kubernetes, CI/CD, security, observability, and reliability initiatives.
Senior SRE operating Kubernetes and cloud infrastructure for Penn Entertainment’s sports betting and media platforms. Leading migrations, automation, observability, and incident response across regulated production services.
DevOps Engineer operating multi - cloud Kubernetes infrastructure for InfluxData’s time - series platform. Automating operations and supporting highly available distributed services.
Senior Reliability Engineer improving embedded protection, control, and software products for utility grids. Leading reliability testing, failure analysis, and modernization initiatives for resilient energy systems.
Senior Reliability Engineer improving embedded grid automation reliability for utility - scale energy systems. Leading testing, failure analysis, KPIs, and modernization initiatives with utilities.
Site Reliability Expert leading observability and SRE for Valtech, an experience innovation company. Improving reliability across cloud - native, microservices - based environments.
Staff DevOps Engineer owning reliable, scalable infrastructure for Nexxa’s AI systems. Supporting machine learning workloads across heavy - industry operations.
Staff SRE securing and scaling IAM systems at RBC, a Canadian bank. Designing resilient infrastructure, automating operations, and leading incident response.