almost 5 years ago
Base Salary
$100k - $165k/yr
Responsibilities
- Own the architecture, design, development, deployment, and operations of the backend API stack.
- Design and implement large-scale data-processing pipelines for images, video, text, audio, neural networks, and human annotations.
- Build tools, tests, metrics, and dashboards that accelerate model-training development.
- Collaborate with frontend engineers to integrate backend systems.
- Influence architecture with emphasis on security, scalability, reliability, and performance.
- Set up and maintain monitoring, metrics, reporting, observability, and alerting systems.
- Develop systems for deploying, scaling, and monitoring software.
- Analyze complex scalability, reliability, performance, and security problems.
- Own projects and improve engineering efficiency and operations.
- Contribute clear technical documentation and promote software-development best practices.
Requirements
- Bachelor’s degree in Computer Science, Physics, Computer Engineering, or Electrical Engineering, or exceptional skills or practical software engineering experience.
- Strong knowledge of at least one programming language related to data engineering, with Python required.
- Solid backend experience building RESTful Python applications.
- Knowledge of Python backend frameworks including Flask, Django, or Tornado; Django is preferred.
- Experience with SQL, relational and non-relational databases, and systems such as SQL Server, MySQL, and Redis.
- Experience with Docker and container orchestration technologies such as Kubernetes.
- Ability to create efficient, reliable, scalable, and maintainable software architecture.
- Hands-on experience with CI/CD pipelines.
- Experience with unit, integration, and functional testing across end-to-end development processes.
- Understanding of logging, monitoring, and alerting systems.
- AWS EC2, S3, and RDS experience is a plus.
- ElasticSearch or other scalable search-system experience is a plus.
- Real-time data-processing experience is a plus.
- Experience with AWS SQS, Spark, Hadoop/MapReduce, NumPy, pandas, scikit-learn, PostgreSQL, data stores, indexers, and cluster-management software is preferred.
- Exceptional oral and written communication skills.
Benefits
- Medical, dental, and vision insurance
- Flexible parental leave
- Flexible time off and paid holidays
- Generous equity and 401(k) plan
- Growth potential and rapid advancement for high-impact team members
- Full-time role located in San Francisco or Palo Alto, California
Tech Stack
Apache HadoopApache SparkAWSDjangoDockerElasticsearchFlaskKubernetesMicrosoft SQL ServerMySQLNumPyPandasPostgreSQLPythonRedisscikit-learnSQL
