3 months ago
Hyderābād, IndiaStaff+

Responsibilities

  • Design, configure, and enhance monitoring and instrumentation across critical network locations and platforms.
  • Analyze network telemetry, logs, and performance data to identify trends, risks, and root causes.
  • Lead troubleshooting for complex connectivity and performance issues using topology and visualization tools.
  • Build and maintain advanced dashboards, health views, and executive-ready visualizations.
  • Support off-hours and weekend implementations and disaster recovery tests as required.
  • Apply observability, high availability, disaster recovery, governance, risk, scalability, and resilience principles to supported platforms.
  • Collaborate with Network Engineering on system design and implementation to embed observability from project inception.
  • Mentor Associates, review their work, and provide technical guidance and best practices.
  • Document system designs and supporting materials to improve transparency and reduce future maintenance costs.
  • Participate in technical decision-making and define future-state technical architecture based on changing business requirements.
  • Monitor system metrics and ensure solutions meet performance and enterprise technology requirements.

Requirements

  • Minimum of 6+ years of related experience.
  • Bachelor’s degree preferred or equivalent experience.
  • Strong knowledge of network, server, application performance, cloud, event management, and synthetic monitoring observability tools.
  • Strong knowledge of OpenShift container environments and Docker images and application hosting.
  • Hands-on experience with SNMP, SSH, HTTPS, TACACS, LDAP, YANG, REST APIs, YAML, and JSON.
  • Hands-on experience with Python, Django, and Ansible.
  • Solid understanding of infrastructure automation, CI/CD pipelines, GitHub, and image repositories.
  • Hands-on experience with DevOps tools and editor tools including VSCode, PyCharm, KIRO, and IntelliJ.
  • Ability to collaborate effectively across diverse technical and business stakeholder groups.

Benefits

  • Competitive compensation including base pay and an annual incentive.
  • Comprehensive health and life insurance and well-being benefits based on location.
  • Pension and retirement benefits.
  • Paid time off, personal and family care, and other leaves of absence.
  • Flexible hybrid work model with 3 days onsite and 2 days remote; onsite days include Tuesdays, Wednesdays, and a third team- or employee-specific day.
  • Professional development investment and a supportive internal community.

Tech Stack

AnsibleDjangoDockerOpenShiftPython

Categories

Depository Trust & Clearing Corporation (DTCC)

About Depository Trust & Clearing Corporation (DTCC)

5,001-10,000 employees
Contact me