Job Type: Full-Time
Interview Type: Video or In Person
Duration: Long Term
Work Preference: Onsite
Experience: 4+Years
Work Location: Gondia, MH
Roles & Responsibilities
- Design, deploy, and manage scalable cloud infrastructure on Amazon Web Services (primary), with exposure to Microsoft Azure as a plus.
- Work with core AWS services including:
- Amazon EC2, Amazon EKS, Amazon S3, Amazon RDS
- Networking via Amazon VPC, Elastic Load Balancing, Amazon Route 53
- Security using AWS Identity and Access Management and AWS Secrets Manager
- Manage and operate Kubernetes environments (EKS) with hands-on experience in Pods, Deployments, Services, Secrets, Ingress, and Persistent Volume Claims (PVC).
- Deploy and manage open source and containerized applications using Helm, EC2, Docker and Kubernetes.
- Build and maintain CI/CD pipelines using:
- Jenkins, ArgoCD, Bitbucket
- AWS native services such as AWS CodePipeline and AWS CodeBuild
- Implement Infrastructure as Code (IaC) using Terraform or CloudFormation.
- Implement monitoring, logging, and alerting using:
Prometheus, Grafana & Amazon CloudWatch
- Troubleshoot production issues, perform root cause analysis, and ensure high availability and performance.
- Optimize infrastructure for cost, scalability, and operational efficiency.
Required Skills
- 4+ years of hands-on experience with Amazon Web Services.
- Strong understanding of cloud fundamentals: compute, storage, networking, IAM, and security best practices.
- Hands-on experience with Kubernetes (Pods, Deployments, Services, Ingress, Secrets, PVC).
- Experience with CI/CD tools (Jenkins / ArgoCD / Bitbucket / AWS CodePipeline) and Linux.
- Experience with containerization using Docker and Kubernetes.
- Experience with Infrastructure as Code using Terraform.
- Monitoring and observability experience using Amazon CloudWatch, Prometheus and Grafana.
- Scripting experience using Python/Bash for automation.
Good to Have
- Experience with code quality and security tools like SonarQube and Snyk.
- Experience with caching technologies such as Redis / Amazon ElastiCache.
- Experience with Helm for Kubernetes deployments.
- Exposure to multi-cloud environments (Azure preferred).
Preferred Skills
- Experience with advanced observability platforms like Datadog
- Exposure to aws services like Amazon SQS, EventBridge
- Exposure to LLMOps tools and platforms (e.g., Arize Phoenix, Langfuse or similar)
- Experience working in high-availability, production-grade environments
- Understanding of DevSecOps practices and secure CI/CD pipelines