2 months ago
Responsibilities
- Debug and triage issues across Linux-based sensor nodes and edge appliances deployed at customer sites.
- Use SSH to diagnose, patch, and recover field hardware with limited remote access and incomplete information.
- Own site bring-ups end to end and restore systems to operation.
- Build and maintain fleet-management systems for OTA updates, device health tracking, remote diagnostics, and lifecycle management.
- Eliminate recurring incidents through tooling, pre-deployment checks, root-cause processes, and automation.
- Collaborate with embedded systems and platform teams on reliability and deployment requirements.
- Design and implement logging, metrics, alerting, and telemetry across edge devices and AWS cloud infrastructure.
- Develop runbooks and incident-response procedures and participate in on-call rotations.
Requirements
- Strong Linux systems administration experience and comfort working over SSH in production environments.
- Experience working with edge or on-premises hardware alongside cloud infrastructure.
- Solid networking fundamentals, including DNS, firewalls, VPNs, subnets, and secure remote access.
- Scripting or programming experience with Python, Go, or Bash for operational tooling.
- Familiarity with containerization; Docker experience is relevant and Kubernetes is a plus.
- Embedded systems experience, including reading firmware logs and understanding hardware-software boundaries.
- AWS infrastructure, IAM, networking, and observability tooling experience is a strong plus.
- Rust or C experience is advantageous for reading and reasoning about low-level firmware code.
About Specter
Specter delivers real-time data and insights on private companies, enabling investors to make informed, confident decisions. Harness the power of live data and AI-driven analysis to outsmart the competition and make confident decisions in private markets.
