InternFlow

Greenhouse ·

full-time

Senior Cloud Infrastructure Engineer

Scaleai · San Francisco, CA; New York, NY

Scale is a data and infrastructure provider that supports companies building large‑scale generative AI models, serving customers such as OpenAI and Microsoft. The Platform team builds the core abstractions and cloud foundation for these products. In this role the intern will oversee the full lifecycle of Scale’s AWS environment, handling routine patching, version upgrades, and system maintenance to keep services highly available. Daily work includes collaborating with the security team to find and remediate vulnerabilities, monitoring resource usage with tools like CloudWatch or Datadog, and applying cost‑ and performance‑optimisation techniques. The candidate will automate repetitive tasks using Bash or Python scripts and infrastructure‑as‑code tools such as Terraform, AWS Systems Manager, and Atlantis, and will assist in incident response and health reviews. The position suits someone with several years of hands‑on AWS experience, strong scripting ability, familiarity with Kubernetes and security hardening, and a proactive mindset toward automation and reliability.

aws infrastructureterraformkubernetescloud securityautomation scriptingincident responsecloud optimizationmonitoringsenior cloud engineer

Education

  • Level: Bachelor's degree
Cloud Infrastructure Operations6-6 yrs experience

Key Skills

AWSEC2VPCIAMS3RDSEKSTerraformAtlantisKubernetesBashPythonDatadogCloudWatchAzure

Experience required

6-6 years experience
0 yrs10+ yrs

About the Role Scale is powering the generative AI wave by providing data and infrastructure for large‑scale foundation models. The Platform team builds core abstractions and infrastructure for rapid product iteration. Responsibilities Manage end‑to‑end AWS infrastructure lifecycle including patching, version upgrades, and maintenance. Collaborate with Security Team for vulnerability identification, prioritization, and remediation. Monitor and optimize AWS resource utilization for performance, reliability, and cost. Automate operational tasks using scripting and tools such as Bash, Python, Terraform, Atlantis, and AWS Systems Manager. Provide incident response, troubleshooting, and participate in health reviews. Ensure security compliance and proactive hardening of cloud environments. Requirements 6+ years of experience with core AWS services (EC2, VPC, IAM, S3, RDS, EKS) and multi‑account management. Proven track record of large‑scale patching programs and OS/middleware version upgrades. Proficiency with AWS Systems Manager, Terraform, Atlantis, Kubernetes. Scripting experience in Bash and Python for automation. Experience with vulnerability scanning tools and cloud security hardening. Familiarity with monitoring tools like Datadog or CloudWatch. Bonus: experience with Azure or Google Cloud Platform. Eligibility

Skills for this role

AWSEC2VPCIAMS3RDSEKSTerraformAtlantisKubernetesBashPythonDatadogCloudWatchAzure

Required skills

AWSEC2VPCIAMS3RDSEKSTerraformAtlantisKubernetesBashPythonDatadogCloudWatchAzure

Also mentioned in the listing

aws infrastructurecloud securityautomation scriptingincident responsecloud optimizationmonitoringsenior cloud engineer

Typical process for this type of role

A general guide — the exact steps for this specific listing may vary; check the original posting for details.

  1. 1ApplicationSubmit your resume through the apply link.
  2. 2ScreeningRecruiter reviews your background against the role.
  3. 3AssessmentA technical test, assignment, or coding round, depending on the role.
  4. 4Interview(s)One or more rounds with the hiring team.
  5. 5OfferOffer letter with compensation and start date.

Before you apply

0/4

Senior Cloud Infrastructure Engineer at Scaleai: frequently asked questions

Who can apply for the Senior Cloud Infrastructure Engineer role at Scaleai?
The listing asks for Bachelor's degree; 6-6 years of experience.
What skills does the Senior Cloud Infrastructure Engineer role require?
The listing highlights AWS, EC2, VPC, IAM, S3, RDS, EKS, Terraform, Atlantis, Kubernetes. Show each of these in a project or past role on your resume.
What is the salary for this role?
Scaleai has not stated compensation in the listing. Check the original posting or ask during the application process.
Is the Senior Cloud Infrastructure Engineer position remote, hybrid or onsite?
The listing gives San Francisco, CA; New York, NY as the location and does not state a work mode.
What is the application deadline?
Scaleai has not listed a fixed deadline, so apply early in case the opening is filled.
How do I apply for the Senior Cloud Infrastructure Engineer role?
Use the Apply button on this page. It opens the original listing on job-boards.greenhouse.io, where you submit your application with the company.

More at Scaleai

Other jobs at Scaleai

See all 21 openings at Scaleai

Explore Related Placements


// similar opportunities

You might also like