About the role
We are looking for a Senior DevOps Engineer to join a small, experienced team.
You will work across a diverse enterprise environment that combines established platforms, modern cloud services, and ongoing modernization initiatives. The role requires strong technical foundations, excellent troubleshooting skills, and the ability to quickly understand complex systems and dependencies.
This is a hands-on senior position for someone who is comfortable taking ownership, driving technical initiatives from analysis through delivery, and collaborating effectively across teams.
Your responsibilities
- Drive DevOps initiatives from investigation and planning through implementation.
- Troubleshoot complex issues across applications, operating systems, infrastructure, networks, and databases.
- Build and maintain Infrastructure as Code and configuration management solutions.
- Develop and improve CI/CD pipelines, deployment automation, and release processes.
- Support and enhance Microsoft Azure, Windows, and Linux environments.
- Improve monitoring, logging, alerting, and operational visibility.
- Support incident response, root-cause analysis, and production troubleshooting.
- Improve system security, reliability, scalability, and maintainability.
- Coordinate technical dependencies and communicate risks, decisions, and progress clearly.
- Create and maintain technical documentation, runbooks, and architecture diagrams.
- Participate in operational support and on-call activities where required.
Your profile
- At least 5 years of experience in DevOps, Platform Engineering, Systems Engineering, Site Reliability Engineering, or a similar role.
- Strong analytical and troubleshooting skills across operating systems, infrastructure, applications, and networks.
- Strong practical knowledge of Windows Server and Linux.
- Ability to take ownership and drive technical initiatives from analysis through delivery.
- Comfortable working independently and collaborating within a small team.
- Ability to communicate technical risks, decisions, dependencies, and progress clearly.
- Practical understanding of Agile delivery and a shared DevOps culture.
- Good spoken and written English.
Technical skills
We are looking for strong practical experience in the core areas below and working knowledge across the wider technology landscape. We do not expect expert-level knowledge of every listed technology.
- Programming: Python, Go, or JavaScript
Ability to create maintainable automation, tools, and integrations. - Scripting: Bash and PowerShell
Experience automating operational, deployment, and administrative tasks. - Operating systems: Windows Server and Linux
Understanding of services, processes, permissions, filesystems, package management, certificates, logs, and performance troubleshooting. - Networking: OSI model, TCP/IP, DNS, HTTP/HTTPS, TLS, routing, proxies, firewalls, and load balancing
Ability to diagnose connectivity, certificate, name-resolution, routing, and protocol-related issues. - Web servers and proxies: IIS, NGINX, and HAProxy
Understanding of web hosting, reverse proxies, routing, TLS termination, health checks, and load balancing. - Version control: Git with platforms such as GitHub, GitLab, Bitbucket, or Azure Repos
Experience with branching, pull requests, code reviews, repository management, and release workflows. - Databases: Microsoft SQL Server and Azure SQL
Ability to troubleshoot connectivity, authentication, permissions, backup, recovery, and performance issues. - Cloud: Microsoft Azure
Practical knowledge of infrastructure, identity, networking, storage, monitoring, security, governance, and cost awareness. - Infrastructure as Code and configuration management: Terraform, Ansible, or Puppet
Experience managing repeatable infrastructure and configuration changes through code. - Containers: Docker; Kubernetes knowledge is welcome but not required
Understanding of images, registries, networking, storage, secrets, resource management, container orchestration, and troubleshooting. - CI/CD: Azure DevOps and GitHub Actions
Experience with build, test, quality checks, security scanning, artifact management, deployments, approvals, and rollback workflows. - Observability: Azure Monitor, Prometheus, and Grafana
Experience with metrics, logs, alerts, dashboards, service health, and production troubleshooting. - Security and reliability:
Understanding of identity and access management, secrets and certificate management, vulnerability management, high availability, backup and recovery, and resilient system design. - Code quality: SonarQube or similar tools
Familiarity with automated quality and security controls integrated into delivery pipelines.