Cloud Operations Engineer - Infrastructure
TP-Link
Irvine, CA
See who TP-Link has hired for this role
See who TP-Link has hired for this role
Description
ABOUT US:
Headquartered in the United States, TP-Link Systems Inc. is a global provider of reliable networking devices and smart home products, consistently ranked as the world’s top provider of Wi-Fi devices. The company is committed to delivering innovative products that enhance people’s lives through faster, more reliable connectivity. With a commitment to excellence, TP-Link serves customers in over 170 countries and continues to grow its global footprint.
We believe technology changes the world for the better! At TP-Link Systems Inc, we are committed to crafting dependable, high-performance products to connect users worldwide with the wonders of technology.
Embracing professionalism, innovation, excellence, and simplicity, we aim to assist our clients in achieving remarkable global performance and enable consumers to enjoy a seamless, effortless lifestyle.
Key Responsibilities
REQUIRED QUALIFICATIONS
Salary range: $100-150K
Please, no third-party agency inquiries, and we are unable to offer visa sponsorships at this time.
ABOUT US:
Headquartered in the United States, TP-Link Systems Inc. is a global provider of reliable networking devices and smart home products, consistently ranked as the world’s top provider of Wi-Fi devices. The company is committed to delivering innovative products that enhance people’s lives through faster, more reliable connectivity. With a commitment to excellence, TP-Link serves customers in over 170 countries and continues to grow its global footprint.
We believe technology changes the world for the better! At TP-Link Systems Inc, we are committed to crafting dependable, high-performance products to connect users worldwide with the wonders of technology.
Embracing professionalism, innovation, excellence, and simplicity, we aim to assist our clients in achieving remarkable global performance and enable consumers to enjoy a seamless, effortless lifestyle.
Key Responsibilities
- Design, build, and maintain reliable, scalable, and secure cloud-native infrastructure platforms supporting large-scale production workloads.
- Operate and optimize multi-account AWS environments, ensuring infrastructure is secure, repeatable, and auditable through Infrastructure as Code tools such as Terraform.
- Manage production Kubernetes clusters, including provisioning, upgrades, autoscaling, networking, observability, capacity planning, and day-to-day operations.
- Build and operate Kubernetes ecosystem components such as CRDs, Helm, HPA, Cluster Autoscaler, CoreDNS, and Cluster API.
- Operate and improve GitOps-based deployment workflows using tools such as FluxCD or ArgoCD.
- Manage and enhance Istio service mesh capabilities, including traffic routing, service discovery, resilience, security, and service-to-service communication.
- Define and improve reliability practices, including SLOs, Error Budgets, monitoring, alerting, incident response, and post-mortems.
- Participate in a scheduled on-call rotation to support production cloud infrastructure and Kubernetes platforms.
- Troubleshoot complex production issues across cloud infrastructure, Kubernetes, Linux systems, networking, and distributed services.
- Drive automation for infrastructure provisioning, configuration management, CI/CD pipelines, observability, and operational workflows using Terraform, Go, Python, or similar technologies.
- Collaborate with application engineering, architecture, security, and platform teams to improve infrastructure reliability, scalability, and operational efficiency.
REQUIRED QUALIFICATIONS
- Bachelor’s degree or above in Computer Science, Software Engineering, Information Technology, or a related field.
- 2+ years of hands-on experience in cloud infrastructure, Kubernetes operations, platform engineering, SRE, or related areas.
- Strong knowledge of AWS services, including EKS, IAM, VPC, EC2, S3, and related networking and security capabilities.
- Hands-on experience operating Kubernetes in production environments, including cluster architecture, workload orchestration, networking, autoscaling, and troubleshooting.
- Familiarity with Kubernetes ecosystem tools such as CRDs, Helm, Cluster API, HPA, Cluster Autoscaler, and CoreDNS.
- Experience with GitOps tools such as FluxCD or ArgoCD.
- Solid Linux administration and troubleshooting skills, including systemd, networking, and performance analysis.
- Experience with CI/CD pipelines and infrastructure automation using Terraform, Go, Python, or similar tools.
- Good understanding of reliability engineering practices, including SLOs, incident response, monitoring, alerting, and post-mortems.
- Strong problem-solving skills and ability to diagnose and resolve complex infrastructure issues in distributed systems.
- Good communication skills and ability to collaborate effectively with cross-functional engineering teams.
- Willingness to participate in a scheduled on-call rotation.
- Experience with NVIDIA device plugins, GPU scheduling, or GPU workload operations in Kubernetes environments.
- Experience with additional public cloud platforms such as Azure or Alibaba Cloud.
- Kubernetes certifications such as CKA, CKAD, or CKS are a plus.
Salary range: $100-150K
- Free snacks and drinks
- Fully paid medical, dental, and vision insurance (partial coverage for dependents)
- Contributions to 401k funds
- Bi-annual reviews, and annual pay increases
- Health and wellness benefits, including free gym membership
- Quarterly team-building events
Please, no third-party agency inquiries, and we are unable to offer visa sponsorships at this time.
-
Seniority level
Entry level -
Employment type
Full-time -
Job function
Engineering and Information Technology -
Industries
Consumer Electronics
Referrals increase your chances of interviewing at TP-Link by 2x
See who you knowGet notified about new Cloud Engineer jobs in Irvine, CA.
Sign in to create job alertSimilar jobs
People also viewed
-
Site Reliability Engineer
Site Reliability Engineer
-
Systems Engineer- Automation
Systems Engineer- Automation
-
Infrastructure Engineer
Infrastructure Engineer
-
Site Reliability Engineer III
Site Reliability Engineer III
-
IT Systems Administrator / AWS Infrastructure Specialist
IT Systems Administrator / AWS Infrastructure Specialist
-
Site Reliability Engineer II
Site Reliability Engineer II
-
DevOps Systems Engineer
DevOps Systems Engineer
-
Senior Platform Engineer
Senior Platform Engineer
-
Development Ops Engineer
Development Ops Engineer
-
Site Reliability Engineer
Site Reliability Engineer
Similar Searches
Explore top content on LinkedIn
Find curated posts and insights for relevant topics all in one place.
View top content