Site Reliability Engineer

Infosys

Bengaluru, INonsitePosted Jul 2, 2026
Posting intelligenceActively listed

Skills

datadoggopythonopenaiazurerubygooglecloudawsml

About the role

Key Responsibilities:

Roles and Responsibilities

Design and implement the lifecycle of services from conception to inception including system design build and deployment

Develop software solutions to enable operability of large scale distributed systems capable of handling millions of transactions and petabytes of data

Manage capacity and performance to help scale the infrastructure both on public and private clouds around the world

Define and implement standards and best practices related to System Architecture Deployment metrics operational tasks

Support services through activities such as monitoring availability system health and incident response

Improve system performance application delivery and efficiency through automation process refinement postmortem reviews and in depth configuration analysis

Engage in Communications across all areas of the organization

Troubleshooting and monitoring production systems to ensure the highest uptimes are maintained

Support and improve upon existing high availability architecture solutions as well as manage the operational activity

Integrate Generative AI GenAI and AIOps tools to automate incident detection root cause analysis and resolution workflows e

g

self healing scripts intelligent runbooks reducing manual toil and accelerating response times

Apply Prompt Engineering techniques to enhance interactions with AI based observability and automation platforms improving accuracy and efficiency of AI responses

Leverage platform specific AI capabilities e

g

AWS Bedrock Azure OpenAI GCP Vertex AI to architect intelligent SRE solutions tailored to cloud environments

Design implement and maintain AI ML driven monitoring and alerting systems to proactively detect anomalies and predict potential failures enabling preemptive remediation

Develop and train machine learning models using operational telemetry logs metrics events traces to support predictive analytics and intelligent automation

Evaluate and deploy AIOps platforms e

g

Moogsoft Dynatrace Splunk BigPanda Datadog Elastic to enhance observability reduce noise and accelerate incident resolution

Experience in one or more high level programming languages like Python or Ruby or GoLang and familiar with Object Oriented Programming

Preferred Skills:

Technology->DevOps->Site Reliability Engineering(SRE),Technology->AI-AI Engineering->AI/ML Solution Architecture and Design->generative ai,Technology->DevOps->DevOps Architecture Consultancy

Questions about this role

Click "Apply with AI Applyd" above and you are done. Your resume is rewritten for this advert, the screening questions are answered, and it is submitted on Infosys's own hiring system. No retyping your history, no fourteen tabs, no evening lost.

Compensation for DevOps / SRE roles in India varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for India medians across recent openings.

You never touch the form - the application is filled and submitted for you on Infosys's own hiring system. It is not marked sent when we press submit. It is marked sent when a confirmation from their system arrives at the address we apply with, and your dashboard shows which stage each application is at until then.

Twelve applicant tracking systems have a real apply path: Workday, Greenhouse, Lever, Ashby, Workable, iCIMS, Personio, Recruitee, Teamtailor, Rippling, Breezy and SmartRecruiters. Your application goes in on the employer's own hiring system, never into an aggregator queue.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.