Jobbie
โ† Discover jobs
Weekday AI

DevOps Engineer

Industry Technology & Software

Bengaluru, KA, IndiaPosted 2d ago

Job description

๐—ง๐—ต๐—ถ๐˜€ ๐—ฟ๐—ผ๐—น๐—ฒ ๐—ถ๐˜€ ๐—ณ๐—ผ๐—ฟ ๐—ผ๐—ป๐—ฒ ๐—ผ๐—ณ ๐˜๐—ต๐—ฒ ๐—ช๐—ฒ๐—ฒ๐—ธ๐—ฑ๐—ฎ๐˜†'๐˜€ ๐—ฐ๐—น๐—ถ๐—ฒ๐—ป๐˜๐˜€

๐—ฆ๐—ฎ๐—น๐—ฎ๐—ฟ๐˜† ๐—ฟ๐—ฎ๐—ป๐—ด๐—ฒ: ๐—ฅ๐˜€ ๐Ÿญ๐Ÿฎ๐Ÿฏ๐Ÿด๐Ÿฌ๐Ÿฌ๐Ÿฌ - ๐—ฅ๐˜€ ๐Ÿฎ๐Ÿฌ๐Ÿฒ๐Ÿฐ๐Ÿฌ๐Ÿฌ๐Ÿฌ (๐—ถ๐—ฒ ๐—œ๐—ก๐—ฅ ๐Ÿญ๐Ÿฎ.๐Ÿฏ๐Ÿด-๐Ÿฎ๐Ÿฌ.๐Ÿฒ๐Ÿฐ ๐—Ÿ๐—ฃ๐—”)

Experience: 2+ yrs

Location: Bengaluru, Karnataka, India

Job Type: Full-time

We are looking for a hands-on and technically strongย  DevOps Engineer ย to build, maintain, and improve reliable, scalable, and secure cloud infrastructure and deployment environments. The role will focus onย  Linux, AWS, Prometheus, and Grafana Cloud , with responsibility for infrastructure automation, monitoring, observability, deployment processes, system reliability, and production support.

The ideal candidate will have strong troubleshooting skills, a practical understanding of cloud infrastructure, and the ability to work closely with software engineering and other technical teams to improve application reliability and operational efficiency.

Requirements

KEY RESPONSIBILITIES - Design, deploy, configure, and maintain scalableย  AWS cloud infrastructure ย across development, staging, and production environments. - Administer and troubleshootย  Linux-based servers and systems , including performance, availability, security, and resource utilisation. - Support cloud services across compute, networking, storage, databases, IAM, and other AWS components. - Implement and maintain infrastructure automation and configuration-management practices. - Build and maintain reliableย  CI/CD pipelines ย to automate application build, testing, deployment, and release processes. - Configure and manageย  Prometheus ย for infrastructure and application monitoring, metrics collection, and alerting. - Develop and maintainย  Grafana Cloud dashboards , visualisations, alerts, and observability solutions. - Monitor system health, application performance, resource utilisation, availability, and service-level indicators. - Investigate production incidents, identify root causes, and implement permanent corrective actions. - Troubleshoot Linux, networking, application deployment, infrastructure, and cloud-related issues. - Improve system reliability through automation, proactive monitoring, capacity planning, and performance optimisation. - Implement appropriate security controls across AWS infrastructure, Linux systems, access management, and deployment environments. - Collaborate with software engineers, QA, architects, and other technical teams to improve deployment and operational processes. - Maintain infrastructure documentation, operational runbooks, monitoring standards, and troubleshooting procedures. - Support backup, disaster recovery, high-availability, and business-continuity requirements. - Identify opportunities to reduce operational overhead through automation and standardisation. - Participate in production releases, incident response, maintenance activities, and continuous improvement initiatives. - Stay current with AWS services, DevOps practices, cloud-native technologies, observability tools, and infrastructure automation.

WHAT MAKES YOU A GREAT FIT - 2+ years of professional experience ย in DevOps, Cloud Engineering, Site Reliability Engineering, Infrastructure Engineering, or a related role. - Strong hands-on experience administering and troubleshootingย  Linux environments . - Good practical experience withย  AWS cloud services ย and cloud infrastructure management. - Strong understanding of AWS compute, networking, storage, IAM, monitoring, and security concepts. - Hands-on experience withย  Prometheus ย for metrics collection, monitoring, and alerting. - Practical experience withย  Grafana Cloud , including dashboards, visualisations, alerts, and observability. - Experience building and maintainingย  CI/CD pipelines ย and automated deployment workflows. - Understanding of infrastructure-as-code and configuration-management practices. - Good knowledge of networking fundamentals, DNS, HTTP/HTTPS, TCP/IP, load balancing, and security concepts. - Strong troubleshooting and root-cause analysis skills across infrastructure and application environments. - Understanding of system reliability, availability, scalability, monitoring, and performance optimisation. - Experience with scripting or automation usingย  Bash, Python, or similar technologies . - Familiarity with Git and modern software development and deployment workflows. - Exposure to Docker, Kubernetes, or other containerisation technologies will be an advantage. - Strong understanding of DevOps principles, automation, observability, and production operations. - Excellent communication and collaboration skills with the ability to work effectively with cross-functional engineering teams. - Proactive mindset with strong ownership of infrastructure reliability, operational excellence, and continuous improvement.