Huxley
Website:
huxley.com
Job details:
About the Team
We are a global technology organization responsible for building, operating, and scaling large-scale digital platforms that serve millions of users. Our engineering teams are distributed across multiple regions and work closely together in a highly collaborative, fast-paced environment focused on reliability, automation, and continuous innovation.
As part of the Infrastructure & Platform Engineering team, you will help design, automate, and maintain critical production systems supporting a broad portfolio of online services.
The Opportunity
We are looking for an experienced Senior DevOps Engineer who thrives in complex production environments and enjoys solving large-scale infrastructure challenges. The ideal candidate combines strong cloud and infrastructure expertise with a passion for automation, reliability, observability, and operational excellence.
You will play a key role in modernizing infrastructure, improving platform stability, driving automation initiatives, and enabling engineering teams to deliver high-quality software efficiently and securely.
Key Responsibilities
- Manage and support mission-critical production and non-production environments.
- Design, implement, and optimize CI/CD pipelines to accelerate software delivery.
- Automate operational processes and eliminate repetitive manual tasks.
- Build and maintain highly available, scalable infrastructure across cloud and on-premise environments.
- Monitor application and infrastructure performance and proactively address issues before they impact users.
- Troubleshoot complex infrastructure, networking, middleware, and application problems.
- Improve platform observability through enhanced monitoring, logging, alerting, and tracing solutions.
- Lead infrastructure-as-code initiatives using modern automation frameworks.
- Collaborate with software engineers, QA teams, security teams, and product stakeholders throughout the software lifecycle.
- Drive cloud adoption and migration initiatives.
- Contribute to system architecture discussions and technology evaluations.
- Implement best practices related to reliability, security, scalability, and operational efficiency.
- Participate in incident response and production support activities when required.
Required Qualifications
- 5+ years of experience in DevOps, Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering.
- Hands-on experience managing large-scale production systems with high availability requirements.
- Strong Linux administration and troubleshooting skills.
- Practical experience with cloud platforms such as GCP, AWS, or Azure.
- Experience provisioning infrastructure using Terraform or similar Infrastructure-as-Code tools.
- Solid experience with container technologies including Docker and Kubernetes.
- Hands-on experience with configuration management and automation tools such as Ansible, Chef, or equivalent.
- Strong CI/CD experience using Jenkins, GitHub Actions, GitLab CI, or similar technologies.
- Experience implementing monitoring and observability solutions using Prometheus, Grafana, OpenTelemetry, or related tools.
- Ability to analyze and resolve performance, scalability, and reliability issues.
- Strong understanding of networking concepts including VPCs, load balancing, DNS, routing, and security controls.
- Excellent communication and collaboration skills.
Preferred Qualifications
- Experience supporting hybrid cloud and on-premise infrastructure environments.
- Strong scripting or programming experience using Python, Shell, Go, or similar languages.
- Experience with large-scale cloud migration projects.
- Knowledge of application security, network security, secrets management, and web security principles.
- Familiarity with HTTP protocols, distributed systems, and microservices architectures.
- Experience implementing SRE practices such as SLIs, SLOs, error budgets, and incident management.
- Experience with logging, tracing, and advanced observability platforms.
- Security certifications or cloud certifications are advantageous.
Click on Apply to know more.