DevOps Administrator II managing technical infrastructure for global complex web technology solutions. Collaborating with team members in a hybrid work setup while enhancing automation and maintaining cloud systems.
Responsibilities
Deploys, configures, manages, and performs ongoing maintenance of technical infrastructure covering both Managed Hostings and AWS instances.
Provides ongoing improvements in automation of technical processes and deployments.
Provides subject matter expertise on the design and architecture of ongoing product evolutions.
Monitors infrastructure performance and backups, and responds to incidents.
Identifies gaps and drives efficiencies with the DevOps team and processes.
Provides, maintains, and manages the appropriate release policy, processes, standards, and procedures.
Assists the development team with the preparation of releases for production.
Creates or improves the automated deployment processes, techniques, and tools.
Troubleshoots and resolves technical operational issues related to IT Infrastructure and software behavior.
Develops and maintains infrastructure documentation including network diagrams, disaster recovery plans, and infrastructure dependencies.
Participates in the yearly audit process in the collection and distribution of audit information.
Requirements
Post-secondary diploma or degree in Computer Science, Engineering or a related discipline from an accredited institution.
7 years of experience as a Systems Administrator or DevOps Administrator.
Hands-on experience operating production workloads on AWS (ECS, compute, networking, IAM, storage).
Strong fundamentals in Linux server administration, web servers (NGINX, Apache), and containerization with Docker.
Practical experience with infrastructure-as-code (Terraform, AWS CDK).
CI/CD pipelines (Bitbucket Pipelines, Buildkite, or similar).
Configuration management (Ansible, Chef, or Puppet).
Code quality and static analysis (SonarQube).
Git-based workflows.
Proficiency with scripting languages (Bash / Python or similar).
Experience with APM and infrastructure monitoring (e.g., New Relic).
Vulnerability scanning and log/threat monitoring (Alert Logic, or similar).
Incident response and implementing automated validation of disaster recovery and business continuity plans.
Excellent communication skills both verbal and written.
Strong problem solving skills.
Strong organizational and multitasking skills to manage multiple priorities and deadlines.
Experience implementing test automations covering DRP and BC plans.
AWS Certified SysOps Administrator or AWS Certified DevOps Engineer desired.
Experience in database administration and use (MySQL / Redis / SQLite, or similar) desired.
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
Mozilla release engineer optimizing Firefox build, test, and deployment pipelines at global scale. Improving developer experience, maintaining automation, and responding to critical service outages without an on - call rotation.
DevOps Manager overseeing releases, enterprise tooling, and incident response for Delta Controls, a building - automation solutions manufacturer. Establishing standards across global product teams and offices.