Website:
fegmo.com
Job details:
Company Description Fegmo is an AI-first, agentic platform that helps brands, retailers, and sellers create, manage, and distribute product content across multiple channels. The company focuses on turning product stories into engines for business growth through autonomous AI agents. Fegmo’s mission is rooted in helping people and organizations protect and realize their dreams by elevating how their products are presented. Team members work with advanced AI technologies in a fast-evolving product content ecosystem, collaborating closely with customers who value innovation and scalability.
Role Description
We’re Hiring — DevOps Engineer
Experience: 4–6 Years | Immediate Hiring | GCP | Kubernetes | Cost Optimization
We are looking for a hands-on DevOps Engineer with 4–6 years of experience to help us build, operate, and scale reliable production infrastructure on Google Cloud Platform (GCP).
The ideal candidate should have strong practical experience with GCP, Kubernetes, CI/CD, infrastructure automation, monitoring, and cloud cost optimization.
Key Responsibilities☁️ GCP & Infrastructure
- Hands-on experience working with GCP production environments.
- Good understanding of:
- Compute Engine
- GKE
- Cloud SQL
- Cloud Storage
- VPC & networking
- Load Balancing
- IAM
- Cloud Monitoring & Logging
- Deploy, maintain, troubleshoot, and scale production infrastructure.
- Understand high availability, scalability, backups, and disaster recovery concepts.
Kubernetes / GKE
- Hands-on experience with Kubernetes, preferably GKE.
- Experience with:
- Deployments
- Services
- ConfigMaps & Secrets
- Ingress
- Resource requests/limits
- HPA
- Cluster autoscaling
- Node pools
- Troubleshoot common Kubernetes issues involving CPU, memory, networking, pods, and deployments.
Infrastructure as Code
- Practical experience with Terraform or equivalent IaC tools.
- Create and maintain reusable infrastructure configurations.
- Manage infrastructure across development, staging, and production environments.
- Follow version-controlled infrastructure practices.
CI/CD
- Experience building and maintaining CI/CD pipelines using tools such as:
- GitHub Actions
- GitLab CI
- Jenkins
- ArgoCD
- Automate build, test, deployment, rollback, and release processes.
- Implement safe deployment strategies and minimize production downtime.
Monitoring & Observability
- Experience with tools such as:
- Prometheus
- Grafana
- GCP Cloud Monitoring
- Cloud Logging
- ELK
- Create useful dashboards and alerts.
- Monitor application and infrastructure health.
- Participate in production incident investigation and RCA.
Cost Monitoring & Optimization
This is an important part of the role.
- Monitor GCP infrastructure and service costs.
- Identify unnecessary or excessive cloud spending.
- Help implement cloud cost optimization / FinOps practices.
- Optimize:
- Compute resources
- GKE workloads
- Persistent disks
- Cloud SQL
- Storage
- Network usage and egress
- Logging and monitoring costs
- Implement resource rightsizing, autoscaling and scheduling where appropriate.
- Create cost dashboards and alerts to track cloud spending.
Required Skills
- 4–6 years of DevOps / Cloud / Infrastructure experience.
- Strong hands-on experience with GCP.
- Good production experience with Kubernetes/GKE.
- Working knowledge of Terraform.
- Strong Linux fundamentals.
- Good understanding of networking concepts:
- TCP/IP
- DNS
- HTTP/HTTPS
- Load Balancing
- VPC
- Firewalls
- Experience with CI/CD pipelines.
- Experience with monitoring and troubleshooting production systems.
- Scripting knowledge in Bash, Python, or similar.
- Strong debugging and problem-solving skills.
Good to Have
- GCP Professional Cloud DevOps Engineer certification.
- GCP Professional Cloud Architect certification.
- Experience with Helm.
- Experience with ArgoCD/GitOps.
- Experience with Prometheus/Grafana.
- Experience with OpenTelemetry.
- Experience with Docker.
- Experience with security scanning and vulnerability management.
- Basic understanding of SRE concepts such as SLI, SLO and SLA.
What We're Looking For
We're looking for someone who can own production infrastructure, not just execute predefined deployment tasks.
You should be comfortable answering questions such as:
- Why is a Kubernetes pod continuously restarting?
- Why did our GCP bill increase by 30%?
- How would you scale a GKE application during traffic spikes?
- How would you troubleshoot a production latency issue?
- How would you safely deploy a new version without downtime?
- How would you reduce cloud infrastructure costs without impacting reliability?
Hiring
Immediate Hiring
If you enjoy working with GCP, Kubernetes, automation, production systems, monitoring, and cloud cost optimization, we'd love to hear from you.
Click on Apply to know more.