4 hours ago
Bengaluru, IndiaSenior
Responsibilities
- Build backend APIs and scalable data pipelines using Python and PySpark.
- Work with data lakehouse and warehouse technologies including Iceberg, Delta Lake, Snowflake, and Databricks.
- Orchestrate workflows with Airflow and optimize big data frameworks.
- Manage infrastructure as code and support monitoring, logging, reliability, scalability, and cost efficiency.
- Design large-scale distributed data architectures and solve complex data integration challenges with internal teams and customers.
Requirements
- At least 5 years of experience in software engineering, data engineering, or infrastructure roles.
- Strong Python skills and experience building scalable data pipelines from scratch.
- Hands-on experience with Apache Iceberg or Delta Lake and Snowflake or Databricks.
- Expertise with workflow orchestration tools such as Airflow or Luigi.
- Experience with big data frameworks including Spark or Hadoop.
- Familiarity with monitoring and analytics tools such as Prometheus, Grafana, ELK, or Datadog.
- Experience designing scalable, reliable, and cost-efficient distributed data systems.
- Strong problem-solving, communication, and customer-facing skills.
- Experience with Terraform or other infrastructure-as-code tools, cloud platforms, Docker, Kubernetes, and data-processing security and privacy practices is preferred.
Benefits
- Competitive salary, meaningful equity, and performance bonus for top performers.
- 401(k) with company match, comprehensive health coverage, and unlimited PTO.
- Daily catered meals in the Mountain View office.
- Support for research, publication, and conference participation.
- The role is based in the Mountain View office; no remote or hybrid arrangement is stated.
Tech Stack
Apache AirflowApache HadoopApache SparkAWSAzureDatabricksDatadogDockerGoogle Cloud PlatformGrafanaKubernetesPrometheusPythonSnowflakeTerraform
Categories
BackendData Engineering
About Granica
Granica builds AI data infrastructure and a continuous optimization product (Crunch) for enterprises running petabyte- to exabyte-scale lakehouse environments. It sells enterprise software that reduces storage and compute spend through distributed execution, workload optimization, scheduling, and query performance, with approximately $200K in annualized value per petabyte reported within weeks. Founded in 2023 and headquartered in Mountain View, California, the company is privately held.
