
Senior Linux Infrastructure Engineer
tastytrade2 days ago
Base Salary
$140k - $180k/yr
Responsibilities
- Diagnose and tune Linux production performance across CPU scheduling, NUMA placement, memory, page cache, disk and filesystem I/O, and network behavior.
- Lead incident troubleshooting, write postmortems, organize follow-up work, and prevent recurring production issues.
- Build and maintain reviewed, tested, version-controlled configuration management code using Salt, Ansible, Chef, or Puppet.
- Build, deploy, and operate containerized services on Kubernetes or Nomad, including scheduling, resource limits, health checks, rollouts, and failure diagnosis.
- Automate operational work using Bash and Python.
- Operate DNS, DHCP, TCP/IP, VLANs, routing, and TLS-related network services and troubleshoot application, host, and network boundaries.
- Configure and troubleshoot Nginx and HAProxy and support Redis and RabbitMQ in production.
- Provision and maintain virtual machines across VMware, Xen, or KVM, including capacity planning, host maintenance, and live migration.
- Administer Vault secrets, dynamic credentials, policies, and rotation.
- Maintain Elastic Stack log aggregation and Nagios, CheckMK, or Icinga alerting.
- Use Git and pull requests for infrastructure changes and participate in peer review.
- Create runbooks, design proposals, and incident write-ups and support change management.
Requirements
- 6+ years of experience in a Linux systems, infrastructure, or SRE role.
- Expert-level Linux knowledge, including performance analysis with tools such as perf, strace, ss, iostat, and bpftrace or equivalents.
- Strong understanding of processes, signals, systemd, cgroups, namespaces, filesystems, memory exhaustion, and file descriptors.
- Production-scale experience authoring and maintaining declarative configuration management code with Salt, Ansible, Chef, or Puppet.
- Production experience with Kubernetes or Nomad and container orchestration operations.
- Strong Bash skills and working Python knowledge.
- Foundational networking experience across VLANs, routing, DNS, DHCP, TCP behavior, and TLS.
- Experience with Nginx or HAProxy for proxying and load balancing.
- Virtualization experience with at least one of VMware, Xen, or KVM.
- Daily experience with Git and peer review, including keeping changes reviewable.
- Excellent written communication for runbooks, design proposals, and incident write-ups.
- Ability to become productive quickly in unfamiliar technical areas.
- Preferred: primary production ownership of Redis or RabbitMQ.
- Preferred: Vault administration including policies, authentication methods, and rotation at scale.
- Preferred: Elastic Stack operations involving index lifecycle, mappings, and cluster tuning.
- Preferred: experience in regulated environments or systems where downtime has direct revenue impact.
- Preferred: colocation or bare-metal experience involving hardware lifecycle, remote hands, and capacity planning.
Benefits
- Hybrid work arrangement in Chicago, Illinois.
- Base salary range of $140,000-$180,000, with a discretionary performance bonus of 10-12% of base salary.
- Stock purchase options, medical, vision, dental, and 401(k) benefits.
- 20 paid vacation days plus an additional paid vacation day during the employee's birthday month.
- 10 paid sick days.
- Gym membership reimbursement and an in-building gym.
- Commuter benefits and a shuttle to and from Metra.
- Pet insurance and wellness and mental health programs.
- Charitable donation matching and two paid volunteer days off.
- Daily catered lunch and an office kitchen with snacks and beverages.