11 days ago
Pune, IndiaStaff+
Responsibilities
- Design and implement high-performance data path components for the Data Ocean platform.
- Develop intelligent caching, cache quotas, eviction policies, lifecycle controls, and object storage capabilities.
- Build scalable services for data movement across object, file, and block storage environments.
- Implement metadata management, object tracking, lookup, and on-demand retrieval services.
- Optimize data transfer pipelines and storage services for throughput, latency, scalability, and reliability.
- Develop production-quality C/C++ multithreaded systems software on Linux.
- Debug complex issues involving performance, concurrency, data consistency, and large-scale storage workloads.
- Participate in architecture discussions, technical design reviews, and code reviews.
- Collaborate with senior engineers, architects, QA, performance engineering, release teams, and cross-functional global engineering groups.
- Contribute to engineering best practices, automation, and continuous improvement.
Requirements
- Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or a related technical field.
- 8+ years of experience developing systems software, storage software, or distributed systems.
- Strong expertise in C/C++ and production-grade software development.
- Strong Linux systems programming experience.
- Experience with multithreading, concurrency, synchronization, and performance tuning.
- Understanding of storage systems, filesystems, or distributed storage architectures.
- Experience with object storage concepts and APIs, including S3.
- Ability to debug complex software issues and optimize performance-critical systems.
- Preferred: experience with distributed storage systems or high-performance data platforms.
- Preferred: knowledge of filesystem internals, metadata management, and namespace services.
- Preferred: experience with intelligent caching, data tiering, replication, or data mobility solutions.
- Preferred: familiarity with storage protocols, networking technologies, cloud storage integration, AI infrastructure, HPC workloads, or large-scale data processing environments.
Benefits
- Opportunity to work on foundational storage technologies supporting next-generation AI, HPC, and enterprise workloads.
- Opportunity to design and build high-performance distributed systems software and solve challenging scalability and performance problems.
- Collaboration with a global engineering team delivering advanced data management solutions.