IT Systems Engineer: Cloud Infrastructure
About the Team:
The IT Engineering at CoreWeave designs, builds, and operates the systems that enable our employees and acquired organizations to work securely and efficiently at scale. The team partners closely with Security, Business Development, PeopleOps, and Corporate functions to deliver resilient, standardized, and automation-first IT solutions that support CoreWeave’s rapid growth.
About the Role:
CoreWeave is hiring an IT Systems Engineer (Cloud Infrastructure) to build and operate the GCP, Azure, and AWS platforms that power our internal IT and enterprise systems. The role covers cloud infrastructure, enterprise IT, and the operational reliability of the platforms those systems run on.
You will work hands-on across our internal cloud foundations: landing zones, hybrid connectivity to data centers, and the infrastructure behind identity, collaboration, IT operations, and internal tooling.
The work is execution-heavy. You will also mentor junior engineers and contribute to team-level technical decisions.
What you'll do
Cloud foundations and architecture (GCP, Azure and AWS)
Provision and manage cloud resources within established landing zones, including compute, storage, IAM, and workload configuration.
Manage IAM, service accounts, Workload Identity Federation, and Secret Manager for internal workloads, aligned with security baselines established by the cloud council.
Build and maintain Infrastructure as Code using Terraform, including GCP-specific modules for project factory and IAM.
Partner with Security and Networking teams to implement approved zero-trust and defense-in-depth patterns for internal cloud services
Contribute to cloud migrations and integrations tied to company growth, including acquisitions and platform consolidations.
Infrastructure you'll own
You will build and operate the cloud infrastructure behind CoreWeave's internal IT platforms and enterprise tooling.
Core enterprise platforms in GCP projects, multi-tenant Azure, and AWS
Shared internal services: VDI, bastion hosts, internal APIs, job runners, and internal tooling
GCP-native workloads on GKE, Cloud Run, or Compute Engine
Operations, reliability, and automation
Maintain the reliability of internal cloud environments, including performance, capacity, cost, and security posture.
Set up and maintain monitoring, alerting, and logging across GCP, Azure, and AWS environments.
Implement, test, and maintain backup, disaster recovery, and failover configurations based on defined recovery requirements.
Identify bottlenecks and propose fixes across performance, security, and scalability.
Build automation for common workflows using Cloud Build, GitHub Actions, or similar pipelines.
Contribute to CI/CD pipelines for infrastructure deployment and management.
Troubleshoot production issues, participate in incident response, and contribute to root-cause analysis and corrective actions.
Mentorship and collaboration
Support and mentor less-experienced engineers through troubleshooting, documentation, code reviews, and knowledge sharing.
Be a reliable technical resource for the team.
Work with engineers across teams on system integrations.
Explain technical ideas and trade-offs to engineers and non-engineers alike.
What you bring
This role operates across GCP, Azure, and AWS. GCP is where most new work lands, so that's where we look for the most depth.
Required qualifications
Required Qualifications
7–10 years of experience as a Systems, Cloud, or Infrastructure Engineer supporting production environments.
Strong hands-on GCP experience, including networking, IAM, compute, storage, security, and observability in multi-project environments.
Working experience with at least one additional public cloud platform, preferably Azure or AWS, and the ability to apply equivalent infrastructure concepts across cloud providers.
Strong hands-on experience using Terraform to provision and manage cloud infrastructure, including reusable modules, remote state, plan review, troubleshooting, and CI/CD-based deployment.
Working knowledge of cloud networking, including VPCs/VNets, subnets, routing, DNS, firewalls, load balancers, peering, VPNs, and private connectivity.
Experience managing IAM, service accounts, workload identities, secrets, and least-privilege access.
Experience using Git-based development workflows, pull requests, code reviews, and automated validation.
Experience operating Linux and Windows workloads in the cloud, including deployment, configuration, patching, hardening, and troubleshooting.
Experience implementing monitoring, logging, alerting, dashboards, and operational runbooks for cloud infrastructure and workloads.
Proficiency in at least one scripting language, such as Python, Bash, or PowerShell.
Demonstrated ability to troubleshoot production issues across infrastructure, networking, IAM, operating systems, and application dependencies.
Clear communication of technical decisions, risks, and trade-offs to technical and non-technical stakeholders.
Preferred Qualifications
Experience using AI-assisted development tools to accelerate infrastructure engineering, scripting, troubleshooting, and documentation.
Familiarity with agentic AI concepts and experience building AI-enabled, end-to-end automation or operational workflows.
Experience integrating AI models or agents with enterprise systems, APIs, cloud services, and automation platforms.
Wondering if you’re a good fit? We believe in investing in our people and value candidates who bring unique backgrounds—even if you don’t check every box. Here are a few qualities we’ve found make someone successful on our team. If some of this describes you, we’d love to talk.
You're deeply hands-on with cloud platforms—especially GCP—and enjoy building infrastructure as code with Terraform across multi-cloud environments.
You have strong problem-solving skills with a proactive, analytical mindset and thrive troubleshooting production issues across infrastructure, networking, and IAM.
You have excellent communication skills, a demonstrated ability to mentor others, and enjoy collaborating across teams in a fast-paced environment.
Why CoreWeave?
At CoreWeave, we work hard, have fun, and move fast! We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:
Be Curious at Your Core
Act Like an Owner
Empower Employees
Deliver Best-in-Class Client Experiences
Achieve More Together
We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!
The base salary range for this role is $182,000 to $242,000. The starting salary will be determined by job-related knowledge, skills, experience, and the market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).