Senior Platform Engineer – AWS / Kubernetes
Job Summary
We are seeking an experienced Senior Platform Engineer to design, develop, and maintain cloud and containerized platforms supporting application development and deployment.
The ideal candidate will have strong hands-on experience with AWS, Kubernetes, Amazon EKS, Terraform, GitHub Actions, GitOps, Linux, Bash/PowerShell, and cloud automation. This role will focus on building scalable deployment solutions, infrastructure automation, platform reliability, security, monitoring, and operational efficiency across hybrid cloud environments.
The successful candidate will also contribute to technical design, knowledge sharing, mentoring, and project leadership.
Roles & Responsibilities
Create and maintain custom GitHub Actions for building and deploying applications.
Develop scripts and automation modules to deploy applications to target platforms.
Design and develop frameworks for backup and restore of applications, containers, and platform resources.
Work with Kubernetes to develop solutions and controllers that can automatically identify and remediate issues within containerized clusters.
Develop and manage Infrastructure as Code (IaC) using Terraform following industry best practices.
Build infrastructure solutions that support effective monitoring, reporting, and observability.
Design and implement GitOps workflows for infrastructure and application deployments.
Support and manage AWS EKS environments and integrate EKS with on-premises infrastructure.
Create and manage AWS resources using the AWS Console and appropriate automation tools.
Troubleshoot and resolve issues across Linux, Kubernetes, AWS, containers, and application deployment environments.
Implement and maintain container orchestration and deployment strategies.
Support multi-cluster Kubernetes environments using appropriate management and deployment tools.
Implement monitoring, observability, and centralized logging solutions across hybrid environments.
Help optimize cloud costs for hybrid cloud workloads.
Implement security best practices for cloud accounts and platforms while maintaining efficient deployment processes.
Support compliance, security, and operational requirements for hybrid cloud deployments.
Contribute to disaster recovery planning and platform resiliency.
Collaborate with development, infrastructure, security, operations, and other teams to improve platform capabilities.
Participate in technical discussions, design reviews, and knowledge-transfer sessions.
Mentor junior engineers/developers on technical design, implementation, and best practices.
Take a project leadership role on medium to large-scale technical projects when required.
Test solutions, troubleshoot issues, learn new technologies, and iterate quickly.
Proactively identify problems and develop pragmatic and scalable solutions.
Communicate effectively with technical and non-technical stakeholders across the organization.
Mandatory Skills & Qualifications
Candidates must have:
5–7 years of relevant Platform Engineering, Cloud Engineering, DevOps, or related experience.
Strong hands-on experience with AWS and AWS resource management.
Strong experience with Amazon EKS and Kubernetes.
Experience integrating AWS EKS with on-premises infrastructure.
Strong experience with Terraform and Infrastructure as Code (IaC).
Experience implementing GitOps workflows for infrastructure and application deployments.
Hands-on experience with GitHub Actions and CI/CD automation.
Strong scripting experience using Bash and/or PowerShell.
Strong troubleshooting and debugging experience in Linux environments.
Experience with container platforms and container orchestration.
Experience developing or implementing automated deployment solutions.
Understanding of cloud security and securing AWS accounts and infrastructure.
Strong problem-solving and analytical skills.
Ability to work effectively across development, infrastructure, security, and operations teams.
Strong verbal and written communication skills.
Nice-to-Have Skills
Experience with Talos Linux / Talos Kubernetes.
Experience with multi-cluster management tools such as Fleet, Rancher, or Argo CD.
Experience with service mesh technologies for cross-cluster communication.
Experience with Grafana and Prometheus.
Experience with the ELK Stack and centralized logging.
Knowledge of centralized logging architectures across hybrid cloud environments.
Experience with cloud cost optimization strategies.
Experience with hybrid cloud security and compliance requirements.
Knowledge of disaster recovery planning for hybrid infrastructure.
Experience with data and process analysis and data modeling.
Transportation, logistics, or technology industry experience.
Experience serving as a technical/project lead on medium to large projects.
Preferred Technical Environment
The role may involve working with:
AWS
Amazon EKS
Kubernetes
Talos
Terraform
GitHub Actions
GitOps
Fleet
Rancher
Argo CD
Bash
PowerShell
Linux
Grafana
Prometheus
ELK Stack
Service Mesh technologies
Certifications
Required
AWS Certified Developer
AWS Certified Technician
Preferred
AWS Certified Solutions Architect
Certified Kubernetes Administrator (CKA)
Certified Kubernetes Application Developer (CKAD)
HashiCorp Certified: Terraform Associate
Linux Foundation Certified System Administrator (LFCS)
Key Competencies
Proactive and pragmatic problem solver.
Strong analytical and troubleshooting skills.
Ability to work quickly and efficiently.
Ability to test, learn, and iterate solutions rapidly.
Strong collaboration and knowledge-sharing skills.
Ability to mentor junior team members.
Strong organizational and project leadership skills.
Effective communication across multiple channels and organizational levels.
Ability to work independently while collaborating effectively with cross-functional teams.
#L1-RB1