Applications Support Senior Analyst - Assistant Vice President

Citi · Pune · 5–8 yrs experience · Posted 2026-07-18

Apply on the company site · Get a referral for this role

Citi salary & ratings · Citi interview process · More live openings

About the role

The SRE Observability Specialist is a hands-on expert, delivering the future of Observability across Services Technology. This role is a part of a central SRE enablement team within Services Production, working closely with SREs, developers, and platform teams to embed telemetry, implement SLOs, and build meaningful visualizations for key production flows — particularly in critical Payments Business.
Responsibilities: - Translate Organization strategy into an actionable delivery plan in partnership with Services Products, Operations & Engineering function, delivering incremental, high-value milestones.
- Understand Critical Business Services functional scope and translate into End-to-End monitoring solutions.
- Deliver against the observability roadmap for Services Technology by building scalable, reusable telemetry solutions.
- Periodic review and analyze application monitoring TOIL and collaborate with stakeholders and remediate them as per organization goal.
- Create and maintain dashboards and visualizations for critical client journeys, including real-time flows across Payments.
- Understanding of SRE principles, DevOps, CI/CD pipeline and tools
Qualifications: - of Grafana/Splunk and MCP will be additional value.
- Hands on experience with modern observability tools (e.g., Prometheus, Grafana, Loki, Artemis, Tempo).
- Contribute to the delivery of reusable automation solutions and reliability frameworks to support LoB teams and reduce operational burden.
- Work with SREs across Services to identify automation opportunities, share reusable tooling, and improve adoption of production engineering standards.
- Maintain and contribute to the central production engineering backlog using Jira, helping prioritize and organize workstreams aligned to Services Technology needs.
- Support the development of utilities, scripts, and templates to improve production support effectiveness and reduce manual intervention.
- Engage in the Services Production Engineering Forum, collaborating with other engineers to share knowledge and resolve common challenges.
- Hands-on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.
- Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments.
- Proven ability to troubleshoot integration issues and support observability across hybrid platforms (on-prem, cloud, containers).
- building dashboards aligned to business outcomes and incident workflows, especially in critical flows like payments.
- Recommended Qualifications: 5-8 years experience in an SRE Role installing, configuring or supporting business applications.
- with some programming languages and willingness/ability to learn.
- Advanced execution capabilities and ability to adjust quickly to changes and re-prioritization
- Effective written and verbal communications including ability to explain technical issues in simple terms that non-IT staff can understand.
- Demonstrated analytical skills
- Issue tracking and reporting using tools
- Good all-round technical skills
- Effectively share information with other support team members and with other technology teams
- Ability to plan and organize workload
- Consistently demonstrates clear and concise written and verbal communication skills
- Ability to communicate appropriately to relevant stakeholde
- Bachelor’s/University degree or equivalent experience

Qualifications

- of Grafana/Splunk and MCP will be additional value.
- Hands on experience with modern observability tools (e.g., Prometheus, Grafana, Loki, Artemis, Tempo).
- Contribute to the delivery of reusable automation solutions and reliability frameworks to support LoB teams and reduce operational burden.
- Work with SREs across Services to identify automation opportunities, share reusable tooling, and improve adoption of production engineering standards.
- Maintain and contribute to the central production engineering backlog using Jira, helping prioritize and organize workstreams aligned to Services Technology needs.
- Support the development of utilities, scripts, and templates to improve production support effectiveness and reduce manual intervention.
- Engage in the Services Production Engineering Forum, collaborating with other engineers to share knowledge and resolve common challenges.
- Hands-on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.
- Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments.
- Proven ability to troubleshoot integration issues and support observability across hybrid platforms (on-prem, cloud, containers).
- building dashboards aligned to business outcomes and incident workflows, especially in critical flows like payments.
- Recommended Qualifications: 5-8 years experience in an SRE Role installing, configuring or supporting business applications.
- with some programming languages and willingness/ability to learn.
- Advanced execution capabilities and ability to adjust quickly to changes and re-prioritization
- Effective written and verbal communications including ability to explain technical issues in simple terms that non-IT staff can understand.
- Demonstrated analytical skills
- Issue tracking and reporting using tools
- Good all-round technical skills
- Effectively share information with other support team members and with other technology teams
- Ability to plan and organize workload
- Consistently demonstrates clear and concise written and verbal communication skills
- Ability to communicate appropriately to relevant stakeholde
- Bachelor’s/University degree or equivalent experience

Responsibilities

- Translate Organization strategy into an actionable delivery plan in partnership with Services Products, Operations & Engineering function, delivering incremental, high-value milestones.
- Understand Critical Business Services functional scope and translate into End-to-End monitoring solutions.
- Deliver against the observability roadmap for Services Technology by building scalable, reusable telemetry solutions.
- Periodic review and analyze application monitoring TOIL and collaborate with stakeholders and remediate them as per organization goal.
- Create and maintain dashboards and visualizations for critical client journeys, including real-time flows across Payments.
- Understanding of SRE principles, DevOps, CI/CD pipeline and tools

More openings at Citi