Summary
A multi-billion-dollar investment firm is looking for a Senior System Administrator to own the reliability of the Enterprise Linux infrastructure that its research and trading teams depend on. This is a hands-on ownership role on a small, senior team, not a seat in a large IT organization.
Company Information:
A multi-billion-dollar, multi-strategy investment firm with a two-decade track record, running fundamental equities, quantitative, and volatility strategies from offices across the U.S.
Job Description:
- Take ownership of uptime and reliability for the core Enterprise Linux footprint, spanning physical and virtual infrastructure
- Install, patch, and maintain Linux systems on bare-metal hardware and virtualized platforms
- Administer and support key infrastructure services, including database systems, storage platforms, and virtualization layers
- Help internal teams make effective use of containerization platforms and shared infrastructure
- Identify and drive process improvements that reduce manual, repetitive operational work
- Build and maintain automation with proper guardrails, including version control, review, and observability
- Own and evolve the monitoring stack so issues are caught early and resolved quickly
- Partner with research, engineering, security, and compliance teams to align infrastructure with operational and regulatory needs
- Participate in on-call support rotations and incident response
- Be available for scheduled maintenance windows, which may occur evenings or weekends
Requirements / Qualifications:
- Roughly 8+ years running Linux/UNIX systems in on-premise production environments
- Hands-on experience with enterprise services such as relational databases (PostgreSQL, MariaDB, MySQL), storage platforms (Vast, NetApp, ZFS), and virtualization (Proxmox, Hyper-V)
- Working familiarity with containerization (Docker, Podman)
- Experience with configuration management and infrastructure-as-code tooling (Ansible, Terraform, Puppet)
- Solid scripting ability in Python and Bash
- Experience building or improving monitoring and observability systems (Prometheus, Zabbix, Grafana)
- Comfort with source control (Git), code review, and disciplined change-management practices
- Strong troubleshooting instincts across the stack, from hardware and kernel up through applications
- Clear communication skills across technical and non-technical teams
