Job description
Project Role : DevOps Architect
Project Role Description : Assess, design, and engineer DevOps solutions in complex technology environments. Design and build DevOps technical solutions for cloud-native, cloud-enabled, and on-prem infrastructure with microservices or traditional application patterns.
Must have skills : Cloud Automation DevOps
Good to have skills : NA
Minimum 7.5 year(s) of experience is required
Educational Qualification : 15 years full time education
Summary:
We re looking for a Infrastructure Auotmation and DevOps Engineer who combines software development expertise with strong DevOps and cloud infrastructure skills. You ll be responsible for designing, developing, deploying, and maintaining scalable web applications while implementing CI/CD pipelines, automating infrastructure, and ensuring high system reliability and performance. You ll work closely with developers, QA engineers, and operations teams to streamline deployment processes and enhance development workflows.
Roles & Responsibilities:
IaC Development
1) Design, develop, and maintain Infrastructure Automation using IaC principals and DevOps Skills.
2) Develop Predictive automation using Monitoring and Observability tools.
DevOps & Cloud Infrastructure
3) Build and maintain CI/CD pipelines using tools such as GitHub Actions, GitLab CI, Jenkins, or CircleCI.
4) Manage infrastructure as code (IaC) using Terraform, CloudFormation, or Ansible, Pulumi and CrossPlane.
5) Deploy and monitor applications in cloud environments (AWS, Azure, or GCP).
6) Set up automated monitoring, logging, and alerting (Prometheus, Grafana, ELK Stack, Datadog).
7) Enhance system reliability, security, and scalability through automation and best practices.
Operations & Collaboration
8) Participate in code reviews and DevOps strategy discussions.
9) Troubleshoot and resolve production issues with a focus on uptime and performance.
10) Implement security best practices and assist with compliance standards.
11) Mentor junior developers and contribute to improving development workflows.
Professional & Technical Skills:
DevOps & Automation Tools
1) CI/CD Pipelines: Jenkins, GitHub Actions, GitLab CI, CircleCI, Travis CI, ArgoCD
2) Infrastructure as Code (IaC): Terraform, Ansible, AWS CloudFormation, Pulumi
3) Configuration Management: Chef, Puppet, Ansible, SaltStack
4) Monitoring & observability: Splunk, DataDog, Prometheus.
Cloud & Infrastructure
5) Cloud Platforms: AWS, Azure, Google Cloud Platform (GCP)
6) Compute & Networking: EC2, Lambda, VPC, Load Balancers, Route 53, DNS, VPNs
7) Storage & Databases: S3, RDS, DynamoDB, PostgreSQL, MySQL, MongoDB
8) Containerization: Docker, Podman
9) Container Orchestration: Kubernetes (K8s), Helm, OpenShift
System Administration & Networking
10) Operating Systems: Linux (Ubuntu, CentOS, Alpine), Bash scripting, basic Windows Server knowledge
11) Networking Fundamentals: TCP/IP, DNS, SSL/TLS, HTTP/HTTPS, VPN, firewalls
12) Security & Compliance: IAM, Secrets Management (Vault, AWS Secrets Manager), SSH, encryption, vulnerability scanning
Observability & Incident Response
13) Monitoring: Application and system performance metrics
14) Alerting & Incident Management: PagerDuty, Opsgenie, Slack integrations
15) Logging: Centralized log management (ELK, Fluentd, Splunk)
Bonus / Emerging Skills
16) Serverless architectures (AWS Lambda, Azure Functions)
17) GitOps (Flux, ArgoCD)
18) AIOps and observability automation
19) SRE principles (Service Level Indicators/Objectives)
Cloud Observability Skills: Multi-cloud monitoring, Cloud cost optimization, Cloud performance analysis, Cloud-native troubleshooting, Cloud security observability
Tools: Azure Monitor, Azure Log Analytics, Azure Application Insights, AWS CloudWatch, AWS X-Ray, GCP Operations Suite, Datadog Cloud Monitoring, Dynatrace Cloud Automation
Skills: Network flow analysis, Packet analytics, WAN observability, SD-WAN monitoring, Network performance management, Network anomaly detection
Tools: ThousandEyes, SolarWinds NPM, Cisco AppDynamics Network Monitoring, Riverbed
Observability: SCOM, SolarWinds, Datadog Infrastructure Monitoring, Dynatrace Infrastructure, VMware Aria Operations, Turbonomic
Additional Information:
1) Strong problem-solving and debugging skills
2) Collaboration with developers, QA, and IT teams
3) Understanding agile methodologies (Scrum, Kanban)
4) Continuous learning and adapting to new tools and cloud services
5) Flexible to work in 24*7 environment
This job post has been translated by AI and may contain minor differences or errors.