Job Overview
Job description
- Location:
- United States - Remote
- Work arrangement:
- Remote
Role Summary
ScaleOps is redefining autonomous cloud and AI infrastructure. We're on a mission to free DevOps and platform engineers from manual resource management so they can focus on innovation, not tuning resources. The results: maximized performance and a reduction of cloud costs by up to 80%.
As the category leader in Autonomous Cloud and AI Infrastructure Resource Management, we're trusted by leading enterprises including Adobe, Wiz, Epic Games, Northwestern Mutual, Coinbase, DocuSign, and Fortune 100 companies to autonomously manage their most critical production environments.
Backed by Insight Partners, Lightspeed Venture Partners, and other leading VCs with over $210M in funding, ScaleOps is the leading player in a massive and growing market. We are building the autonomous infrastructure management platform that will power the next decade of enterprise compute.
Responsibilities
- Own ScaleOps' infrastructure end-to-end — our self-hosted product, its installation and onboarding flows, our SaaS platform, and AI infrastructure
- Manage ScaleOps' cloud infrastructure across AWS, GCP, and Azure — networking, security, SSO, and compute
- Work closely with customers on deployments and troubleshooting
- Collaborate closely with backend, product, and R&D teams to support rapid feature delivery without compromising reliability
- Identify and eliminate operational toil through automation and tooling improvements
- Maintain security and compliance best practices across the infrastructure stack
Requirements
- 5+ years of hands-on experience in infrastructure, platform engineering, or SRE roles in high-scale distributed systems
- Hands-on experience with at least two major cloud providers (AWS, GCP, or Azure)
- Solid understanding of networking, security groups, IAM, SSO/OIDC, and cloud-native security principles
- A strong ownership mentality — you don't just flag problems, you fix them
- Full professional fluency in Hebrew and English
- (Nice to have) Deep expertise with Kubernetes, Helm, and Go — you understand the internals and can build on them, not just deploy
- (Nice to have) Experience with cloud cost optimization, resource management, or working at a company that sells infrastructure tooling
- Role:
- Infra Engineer
Company profile
ScaleOps
scaleops.comScaleOps provides autonomous resource management for cloud and AI infrastructure, keeping Kubernetes and GPU workloads performant under any load while reducing cloud costs. Its platform acts on full workload context in real time, continuously allocating and scaling CPU, GPU and memory for agents, inference models and applications in production, with no manual tuning, implementation or code changes. ScaleOps says customers can cut Kubernetes costs by up to 80%. Founded in 2022 and led by founder and CEO Yodar Shafrir, the company is backed by investors including Insight Partners.
- Headquarters
- Tel Aviv, Israel
- Founded
- 2022
- Founders
- Yodar Shafrir