1 day ago
Bengaluru, IndiaSenior / Staff+
Responsibilities
- Own and evolve relational, analytical/columnar, object-storage, and caching data stores, including schema design, migrations, and retention.
- Design and operate streaming and queueing systems that move data from ingestion to queryable form.
- Build and scale stream processors, writers, schedulers, distributed worker fleets, and other data-processing pipelines.
- Optimize query and write latency, indexing, sharding, throughput, capacity, reliability, and cost efficiency.
- Own quotas, rate limiting, multi-tenant isolation, backup and restore, replication, retention, deletion, migration, and backfill operations.
- Build load-testing, profiling, benchmarking, observability, and telemetry tooling for platform-scale validation and optimization.
- Develop, test, maintain, and debug production services and internal tooling in Python and/or Go.
- Support monitoring, on-call operations, root-cause analysis, architecture direction, design reviews, code reviews, documentation, and technical mentoring.
Requirements
- Bachelor’s degree with 7+ years of related experience, or a master ’s degree with 4+ years, or a PhD with 1+ year, in Computer Science, Software Engineering, or a related field.
- Strong backend software engineering experience building and operating highly scalable, reliable, production-grade data-intensive services and platforms.
- Experience designing and solving scalability, availability, performance, reliability, and fault-tolerance problems in large-scale distributed systems.
- Hands-on experience operating and tuning an OLTP database and an analytical, columnar, or time-series store, including schema design, indexing, query tuning, and migration safety.
- Production experience with streaming or queueing systems, including partitioning, consumer-group semantics, delivery guarantees, backlog handling, and dead-letter handling.
- Experience scaling data-processing compute through queue consumers, asynchronous workers, or streaming jobs, including concurrency tuning, batching, backpressure, and sustained-load throughput.
- Strong Python proficiency and solid experience with at least one additional backend language such as Go, Java, or C++.
- Demonstrated performance-engineering experience with profiling, benchmarking, load testing, capacity decisions, system design, and cost optimization.
- Experience with Kubernetes, containerized environments, and public cloud platforms such as AWS or GCP.
- Ability to independently design, develop, debug, test, and maintain software with minimal guidance.
- Preferred qualifications include operating ClickHouse or comparable columnar stores at scale, capacity modeling or FinOps, multi-tenant isolation, disaster recovery, data retention compliance, observability instrumentation, performance tooling, distributed task frameworks, stream processing, distributed query engines, model-serving infrastructure, cloud/SaaS and air-gapped deployments, and leading medium-sized features or projects.
- Strong communication skills and the ability to explain complex performance, capacity, and cost findings to engineering and product partners.
Tech Stack
Apache FlinkApache KafkaAWSC++ClickHouseGoGoogle BigQueryGoogle Cloud PlatformJavaKubernetesMySQLPostgreSQLPrestoPythonRabbitMQSnowflake
Categories
BackendData Engineering
About Cisco
Cisco designs and sells networking, security, and collaboration platforms for enterprises, service providers, and governments, spanning routers and switches, Wi‑Fi, firewalls, zero‑trust, observability, and cloud-managed IT (Meraki) plus Webex. Its business model mixes hardware, software subscriptions, and support/consulting services. Founded in 1984 and headquartered in San Jose, California, Cisco is a public company traded on Nasdaq and serves customers across data centers, campuses, and service provider networks.
