Our client, a leading global financial technology organization, is seeking a Linux Platform Operations Engineer to join a high-performing infrastructure team responsible for supporting mission-critical trading environments. This role offers the opportunity to work with large-scale Linux platforms, low-latency systems, automation technologies, and modern infrastructure solutions in a highly collaborative and performance-driven environment.
Key Responsibilities
Linux Infrastructure & Platform Management
- Design, build, and maintain enterprise Linux server and storage infrastructure across global environments.
- Manage Linux platform configurations, operating system lifecycle activities, patching, and system upgrades.
- Support configuration management and infrastructure standardization initiatives.
- Execute production, disaster recovery, and certification environment changes.
Incident Management & Troubleshooting
- Act as an escalation point for critical Linux infrastructure incidents.
- Perform root cause analysis and drive issue resolution to improve platform stability.
- Collaborate with global engineering and operations teams during incident response activities.
- Produce post-incident documentation and remediation plans.
Performance & Reliability Engineering
- Monitor and optimize system performance, availability, and capacity.
- Support low-latency infrastructure environments with a focus on Linux kernel performance, networking, and system resilience.
- Partner with Engineering and SRE teams on scaling and performance initiatives.
Automation & DevOps
- Develop and maintain automation solutions using Python and Shell scripting.
- Identify opportunities to streamline operational processes and reduce manual effort.
- Contribute to infrastructure-as-code and DevOps initiatives.
- Support implementation of automated monitoring, alerting, and self-healing capabilities.
Monitoring & Observability
- Build and maintain operational dashboards and reporting solutions.
- Monitor Linux platform health, patch compliance, capacity utilization, and system availability.
- Analyze telemetry and operational data to identify performance trends and operational risks.
Project Delivery
- Lead Linux infrastructure workstreams for strategic technology initiatives.
- Work closely with Engineering, Network, Security, and Operations teams to deliver infrastructure projects.
- Document technical requirements and ensure successful project execution.
Operational Support
- Participate in an on-call rotation supporting critical production systems.
- Support maintenance activities, disaster recovery exercises, and infrastructure testing when required.
Required Skills & Experience
- 5+ years of experience managing Linux infrastructure within large-scale, highly available enterprise environments.
- Strong understanding of Linux operating systems, kernel fundamentals, networking, storage, and performance tuning.
- Experience with configuration management tools such as Salt, Puppet, or similar technologies.
- Hands-on scripting and automation experience using Python and/or Shell.
- Experience supporting both bare-metal and virtualized/cloud-based environments.
- Exposure to monitoring and observability tools such as Grafana, Prometheus, or equivalent platforms.
- Strong troubleshooting, analytical, and problem-solving abilities.
- Excellent communication skills and ability to work effectively within global teams.
Interested? Please drop your updated resume to *************