For over 17 years, Trilyon has been a leader in global workforce solutions, specializing in Cloud Technology, AI/ML, Software Development, Technical Writing, and Digital Transformation. We partner with top companies to deliver high-quality talent in engineering, IT, and emerging technologies. For additional information or to view all of our job opportunities, please visit our website https://trilyonservices.com/careers/
We are seeking a Lead Linux Platform Engineer to join our team. This role will involve owning enterprise Linux platform strategy, architecture, automation, security, reliability, lifecycle management, and continuous improvement across on-premises and cloud environments. The ideal candidate will have experience in enterprise Linux, SUSE Linux, automation, Infrastructure as Code, cloud platforms, virtualization, observability, security hardening, and platform engineering, with a passion for building secure, scalable, automated, and highly reliable infrastructure.
Job Title: Lead Linux Platform Engineer
Location: Remote
Duration: 6 Months Contract
Job Description:
The Lead Linux Platform Engineer will serve as a senior technical leader and platform owner responsible for the design, standardization, reliability, automation, security, and lifecycle management of enterprise Linux infrastructure. The Lead Linux Platform Engineer will operate as a technical authority across Linux platforms, hybrid infrastructure, cloud environments, automation, observability, security, virtualization, storage, and core infrastructure services.
The Lead Linux Platform Engineer will work closely with infrastructure, cloud, security, network, storage, database, application, and vendor teams to deliver resilient and scalable solutions. This role goes beyond traditional system administration by driving platform engineering practices, Infrastructure as Code adoption, automation, operational excellence, and systemic resolution of complex incidents.
Responsibilities: Platform Ownership & Strategy
- Own the enterprise Linux platform roadmap, including operating system lifecycle management, patching, standardization, modernization, and technical debt reduction.
- Define and enforce Linux engineering standards, build patterns, configuration standards, and operational best practices.
- Establish automation-first and reliability-driven operating models for Linux platforms.
- Continuously evaluate platform scalability, resiliency, availability, performance, and cost efficiency.
- Develop platform standards that support consistent deployment and operational management across enterprise environments.
- Establish lifecycle strategies for operating systems, infrastructure components, and supporting technologies.
- Identify opportunities to modernize legacy Linux environments and transition workloads to scalable cloud and hybrid architectures.
Architecture & Engineering
- Design, implement, and maintain scalable and resilient Linux-based solutions across on-premises and cloud environments.
- Serve as a technical authority for enterprise compute, virtualization, storage integration, and core Linux services.
- Lead technical design reviews, architecture assessments, and engineering decisions for Linux platforms.
- Develop technical standards and reference architectures for enterprise Linux environments.
- Partner with application, database, network, storage, and cloud teams to support modern platform architectures.
- Support containerized and hybrid cloud architectures where appropriate.
- Evaluate new infrastructure technologies and recommend solutions that improve reliability, scalability, security, and operational efficiency.
Automation & Modernization
- Lead the design and implementation of infrastructure automation using Ansible, Terraform, Bash, Python, and related technologies.
- Drive adoption of Infrastructure as Code (IaC) and reusable automation modules.
- Develop automated provisioning, configuration management, patching, compliance, and operational workflows.
- Eliminate manual and error-prone processes through orchestration, automation, and self-service capabilities.
- Champion platform engineering and DevOps practices across infrastructure teams.
- Build reusable automation frameworks and standardized deployment patterns.
- Integrate infrastructure automation into CI/CD and DevOps workflows where appropriate.
- Develop automated health checks and remediation processes to improve platform reliability.
Reliability, Performance & Observability
- Own Linux platform availability, performance, capacity, scalability, and recoverability in alignment with business requirements and SLAs.
- Establish and maintain monitoring, logging, alerting, and observability standards.
- Implement and support enterprise monitoring platforms such as Datadog, Dynatrace, Nagios, and SolarWinds.
- Serve as a technical escalation lead during major incidents and complex production issues.
- Conduct detailed root cause analysis and drive permanent corrective and preventive actions.
- Identify recurring incidents and develop systemic solutions to eliminate operational issues.
- Establish platform health metrics, service-level indicators, and performance benchmarks.
- Conduct capacity planning and performance tuning to ensure infrastructure can support current and future business demands.
- Ensure backup, recovery, business continuity, and disaster recovery strategies are implemented and regularly validated.
- Participate in an on-call rotation and lead complex incident response and resolution activities.
Security & Compliance
- Embed security-by-design principles into Linux platform architecture and operations.
- Partner with cybersecurity teams to implement OS hardening, vulnerability management, patching, and compliance controls.
- Define and maintain Linux security baselines and configuration standards.
- Own patching governance and ensure timely remediation of security vulnerabilities.
- Support security audits and ensure configurations are defensible and properly documented.
- Ensure Linux platforms comply with enterprise security policies, regulatory requirements, and governance standards.
- Implement access controls, privileged access practices, and secure configuration standards.
- Collaborate with security teams to investigate and remediate infrastructure-level vulnerabilities.
Virtualization, Storage & Core Services
- Engineer and support Linux workloads running on VMware vSphere and cloud platforms.
- Partner with storage teams on SAN/NAS integration, LUN provisioning, RAID, storage performance, and capacity management.
- Support and maintain Linux-based infrastructure services including DNS, DHCP, LDAP, and NTP.
- Ensure Linux platforms are properly aligned with compute, storage, network, and virtualization architectures.
- Troubleshoot complex compute, storage, networking, and virtualization issues.
- Support high-availability infrastructure and enterprise application environments.
- Perform infrastructure performance tuning and capacity analysis across compute and storage layers.
Cloud & Container Platforms
- Support Linux platforms deployed within hybrid and cloud environments.
- Work with cloud engineering teams to design and operate Linux workloads on AWS, Azure, and GCP.
- Support containerized workloads using Docker, Kubernetes, and OpenShift.
- Promote cloud-native and platform engineering principles where appropriate.
- Help establish standardized Linux deployment patterns across cloud environments.
- Support migration and modernization initiatives involving Linux workloads and cloud infrastructure.
Cross-Functional Leadership
- Act as a senior technical escalation point for complex Linux platform issues.
- Collaborate with infrastructure, network, storage, database, application, security, cloud, and vendor teams.
- Translate technical risks, dependencies, and trade-offs into clear business impact for technical and non-technical stakeholders.
- Lead complex technical initiatives from architecture and design through implementation, validation, and operational handover.
- Coordinate cross-functional activities during major infrastructure changes and platform upgrades.
- Communicate project status, risks, technical decisions, and recommendations to leadership.
Mentorship & Capability Building
- Mentor and guide Linux engineers and infrastructure professionals.
- Establish repeatable, documented, and audit-ready operating practices.
- Promote a culture of automation, reliability, security, operational excellence, and continuous improvement.
- Conduct knowledge-sharing sessions and technical workshops for engineering teams.
- Establish engineering standards and encourage consistent adoption across teams.
- Provide technical guidance during complex troubleshooting and architectural decision-making.
Required Qualifications:
- Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent practical experience.
- 8 10+ years of progressive experience engineering and operating enterprise Linux platforms.
- Deep expertise with SUSE Linux, RHEL, or equivalent enterprise Linux distributions.
- Advanced proficiency with Linux command-line administration and scripting.
- Strong experience with Bash and Python; Perl experience is a plus.
- Proven experience designing and operating automation and Infrastructure as Code solutions.
- Strong hands-on experience with Ansible, Terraform, Puppet, or similar automation technologies.
- Strong experience with VMware vSphere or comparable virtualization technologies.
- Solid understanding of networking fundamentals including TCP/IP, DNS, DHCP, NTP, and VPN.
- Hands-on experience integrating Linux systems with SAN, NAS, RAID, and enterprise storage environments.
- Experience with enterprise backup and recovery technologies such as Veeam or Rubrik.
- Experience building, maintaining, and distributing RPMs and Linux packages.
- Experience implementing enterprise monitoring and observability platforms.
- Strong knowledge of Linux security hardening, vulnerability management, patching, and compliance.
- Demonstrated ability to lead complex technical initiatives and take ownership of platform outcomes.
- Strong communication, collaboration, leadership, and stakeholder management skills.
- Ability to participate in an on-call rotation and lead complex incident response activities.
Preferred Qualifications:
- RHCE, LPIC-3, or equivalent advanced Linux certification.
- Experience operating Linux platforms across hybrid or public cloud environments.
- Experience with AWS, Microsoft Azure, or Google Cloud Platform (GCP).
- Experience with Docker, Kubernetes, or OpenShift.
- Experience working within Platform Engineering or Site Reliability Engineering (SRE) operating models.
- Experience supporting high-availability enterprise applications, including Oracle Single Instance and Oracle RAC.
- Experience with storage array-level configuration, performance tuning, and capacity planning.
- Experience implementing DevOps, CI/CD, and Infrastructure as Code practices.
- Experience with enterprise configuration management and automated compliance tooling.
Key Skills:
SUSE Linux, Red Hat Enterprise Linux (RHEL), Linux Administration, Linux Platform Engineering, Linux Architecture, Linux CLI, Bash, Python, Perl, Ansible, Terraform, Puppet, Infrastructure as Code (IaC), Automation, Configuration Management, Platform Engineering, DevOps, SRE, VMware vSphere, Virtualization, AWS, Azure, GCP, Docker, Kubernetes, OpenShift, SAN, NAS, RAID, LUN Provisioning, Storage Management, Storage Performance Tuning, DNS, DHCP, LDAP, NTP, TCP/IP, VPN, Networking, Veeam, Rubrik, RPM Packaging, Linux Package Management, Datadog, Dynatrace, Nagios, SolarWinds, Monitoring, Logging, Observability, Performance Management, Capacity Planning, High Availability, Disaster Recovery, Backup and Recovery, Business Continuity, Linux Security Hardening, Vulnerability Management, Patch Management, Security Baselines, Compliance, Incident Management, Root Cause Analysis, Problem Management, Technical Escalation, Cloud Infrastructure, Hybrid Cloud, Automation Frameworks, CI/CD, Technical Leadership, Architecture, Infrastructure Modernization, Operational Excellence, Continuous Improvement, Mentoring, Stakeholder Management.
Why Join Us?
- Trilyon, Inc., offers a comprehensive benefits package.
- Opportunities for growth and professional development.
- Collaborative and inclusive company culture.
Equal Employment Opportunity (EEO) Statement:
Trilyon, Inc., is an Equal Opportunity Employer committed to diversity, equity, and inclusion. We do not discriminate based on race, color, religion, gender, gender identity, sexual orientation, national origin, age, disability, veteran status, or any other protected status under applicable laws. Our diverse team drives innovation, competitiveness, and creativity, enhancing our ability to effectively serve our clients and communities. This commitment to diversity makes us stronger and more adaptable.