As an Associate in Technology Operations Engineering, you will monitor, maintain, and support high-availability production environments, troubleshoot runtime issues, and automate operational tasks.
Provide hands-on runtime operational support for business-critical applications to achieve target Mean Time to Restore (MTTR) service goals.
Participate actively in incident response management and Root Cause Analysis (RCA) to boost system stability and application resiliency.
Collaborate with software engineering, cloud, and infrastructure teams to identify and resolve performance bottlenecks, scalability issues, and system faults.
Design, configure, and maintain monitoring, logging, and alerting solutions to proactively address potential runtime failures.
Support application deployment and continuous monitoring across test, integration, and production environments using CI/CD pipelines like Jenkins.
Automate routine operational tasks and manual deployment workflows using Python, Bash scripting, Linux tools, and Ansible.
Drive operational automation and ensure automated test scripts are executed for newly released product features.
Implement resiliency standards, disaster recovery practices, and enterprise tooling to improve operational agility and reduce system toil.
Provide hands-on runtime operational support for business-critical applications to achieve target Mean Time to Restore (MTTR) service goals.
Participate actively in incident response management and Root Cause Analysis (RCA) to boost system stability and application resiliency.
Collaborate with software engineering, cloud, and infrastructure teams to identify and resolve performance bottlenecks, scalability issues, and system faults.
Design, configure, and maintain monitoring, logging, and alerting solutions to proactively address potential runtime failures.
Support application deployment and continuous monitoring across test, integration, and production environments using CI/CD pipelines like Jenkins.
Automate routine operational tasks and manual deployment workflows using Python, Bash scripting, Linux tools, and Ansible.
Drive operational automation and ensure automated test scripts are executed for newly released product features.
Implement resiliency standards, disaster recovery practices, and enterprise tooling to improve operational agility and reduce system toil.
Skills & Eligibility
Bachelor’s degree in Computer Science, Computer Engineering, Information Technology, Computer Applications (B.E / B.Tech / B.Sc / BCA), or equivalent technical experience.
Working knowledge of cloud infrastructure (AWS, Azure, or Google Cloud), distributed computing, and containerization platforms.
Practical experience with monitoring, logging, and incident management frameworks in production environments.
Strong programming and scripting skills to automate operational workflows using Python, Bash, Linux, or Ansible.
Hands-on object-oriented programming foundation in at least one language such as Java or Python.
Understanding of system integration concepts, including REST APIs and real-time/batch data integration layers.
Hands-on exposure to Continuous Integration / Continuous Delivery (CI/CD) environments using tools like Git, Maven, and Jenkins.
Basic experience with relational and NoSQL database management systems such as DB2, Redis, Postgres, or Couchbase.
Public Cloud certifications (AWS / Azure / GCP) or IT network/security certifications are considered an added advantage.
Strong troubleshooting and analytical thinking skills focused on rapidly diagnosing runtime issues.
Clear technical communication skills to translate operational concepts for product managers and cross-functional teams.
Collaborative mindset with a willingness to learn, innovate, and drive continuous operational improvements.
Bachelor’s degree in Computer Science, Computer Engineering, Information Technology, Computer Applications (B.E / B.Tech / B.Sc / BCA), or equivalent technical experience.
Working knowledge of cloud infrastructure (AWS, Azure, or Google Cloud), distributed computing, and containerization platforms.
Practical experience with monitoring, logging, and incident management frameworks in production environments.
Strong programming and scripting skills to automate operational workflows using Python, Bash, Linux, or Ansible.
Hands-on object-oriented programming foundation in at least one language such as Java or Python.
Understanding of system integration concepts, including REST APIs and real-time/batch data integration layers.
Hands-on exposure to Continuous Integration / Continuous Delivery (CI/CD) environments using tools like Git, Maven, and Jenkins.
Basic experience with relational and NoSQL database management systems such as DB2, Redis, Postgres, or Couchbase.
Public Cloud certifications (AWS / Azure / GCP) or IT network/security certifications are considered an added advantage.
Strong troubleshooting and analytical thinking skills focused on rapidly diagnosing runtime issues.
Clear technical communication skills to translate operational concepts for product managers and cross-functional teams.
Collaborative mindset with a willingness to learn, innovate, and drive continuous operational improvements.
Note: This job is posted on external sites. Joblit shares the listing for convenience and does not take responsibility for third-party content.