Senior Site Reliability Engineer

Number of employees

Redwood City, United States

Posted on: 2026-09-14

Category: energy

Ready to make this your next chapter?

Let Gridcare know you found them on WorkInGreen. It helps more companies post climate jobs here.

Employment type:

Full time

Remote?

Yes

Experience required:

Senior

Salary

Salary not provided

About the company:

GridCARE is an AI-native platform company that accelerates power delivery to large-scale AI infrastructure, including data centers, by reducing the typical years-long time-to-energize delay down to months. The company partners with data centers and electric utilities to unlock latent capacity within the existing power grid, avoiding the need to build new generation assets. By applying advanced AI analytics and large-scale data processing, GridCARE identifies underutilized grid capacity and enables faster, more efficient interconnection for high-demand compute workloads such as AI inference and sovereign low-latency computing. GridCARE's mission is closely tied to sustainability and climate goals by maximizing the utilization of existing grid infrastructure rather than requiring costly and carbon-intensive new construction. The company highlights that the US grid operates at less than 32% utilization even in its most congested locations, representing a significant opportunity to serve growing AI energy demand without proportional increases in new energy infrastructure. This approach supports a more efficient and lower-impact path to scaling AI factories. What makes GridCARE unique is its combination of AI-driven grid analytics, utility partnerships, and a platform specifically designed for large flexible loads. The company has demonstrated concrete results, including unlocking over 400 MW of capacity for Portland General Electric years ahead of schedule, securing 150 MW for AI Fabrik, and identifying 650 MW of untapped capacity on a New York utility network. GridCARE claims to create over $25 billion in value per gigawatt of accelerated power capacity delivered.

About Us

GridCARE is a leading venture-backed startup solving the most critical constraint in AI’s growth trajectory: immediate access to power. As demand for computing skyrockets, access to energy has become the defining bottleneck in the AI infrastructure race. While leading tech companies invest billions in speculative, long-term solutions that may take decades to arrive, GridCARE’s pioneering physics-based generative AI platform unlocks gigawatts of hidden capacity in today’s electric grid — enabling hyperscalers, data center developers, and utilities to power AI infrastructure years sooner than conventional approaches and without costly upgrades.

Founded at Stanford’s Doerr School of Sustainability and backed by leading climate-tech and deep-tech investors, GridCARE has assembled a world-class team spanning power systems, AI, and infrastructure.

At GridCARE, you will:

⚡ Work at the intersection of AI, energy, and infrastructure — the foundation of the next industrial revolution.

🤝 Partner with hyperscalers, developers, and utilities on high-impact, real-world deployments.

🌎 Help shape a more abundant, efficient, and resilient energy future for the digital era.

🚀 Join a company defining a new category — capacity acceleration for AI.

💰 Receive competitive compensation, equity, and benefits in a fast-growth, mission-driven environment.

Learn more about GridCARE:

Job Description

We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the availability our customers (utilities, data center operators) require.

Responsibilities

  • Design and operate infrastructure on AWS using Terraform and Kubernetes

  • Build monitoring, alerting, and observability (Prometheus, Grafana, Datadog, or similar) with meaningful SLOs/SLIs

  • Automate away toil — deployment pipelines, capacity management, self-healing systems

  • Partner with engineering on architecture reviews to catch reliability and scalability risks before they ship

  • Manage database and data pipeline reliability for large-scale, real-time grid data processing

  • Drive security and compliance best practices across infrastructure

Qualifications

Required

  • 5+ years in SRE, DevOps, or infrastructure engineering roles

  • Deep experience with Kubernetes, Terraform/IaC, and cloud platforms (AWS Preferred)

  • Strong scripting/programming ability (Python, Bash)

  • Observability Experience (Prometheus, Grafana, Datadog)

  • Track record of running on-call for production systems and leading incident response

  • Experience with CI/CD pipelines (Github Actions) and infrastructure automation

  • Experience with Gitops concepts and tooling (ArgoCD/Flux)

  • Solid understanding of networking, distributed systems, and database reliability

  • Comfortable operating in a fast-moving startup environment with ambiguity

Preferred

  • Experience with data-intensive or real-time processing systems

  • Background in energy, climate tech, or critical infrastructure

  • Experience scaling infrastructure through hypergrowth

  • On-Prem Kubernetes Deployment Experience

  • Windows Server Administration Experience

What We Offer

  • Competitive salary, performance bonus, and equity.

  • Comprehensive health, dental, and vision coverage.

  • Lunch provided three days a week in office.

  • Hybrid schedule for local employees: 3 days in office for collaboration, 2 days remote for focused work.

  • Access to leading academic, industry, and government partners in the AI-energy ecosystem.

  • A mission-driven team focused on shaping the future of the energy transition.

Salary Range

$180,000-$230,000 Total

Join us in tackling one of the most important infrastructure challenges of our time — enabling the energy foundation for the age of AI.

Get job alerts

Receive new climate job opportunities matching your preferences.

Not quite the right fit? Keep looking.

More climate roles that match your skills and values.

View all jobs
Powerline logo
United States
Number of employees

41

Dollar sign

$140,000.00 - $205,000.00

Full time
Energy
WeaveGrid logo
United States
Powerline logo
United States
Number of employees

41

Dollar sign

$140,000.00 - $205,000.00

Full time
Energy
Powerline logo
United States

16 Energy jobs at Gridcare

Gridcare is hiring Senior Full Stack Software Engineer,Principal Power Systems Engineer,Senior Site Reliability Engineer , and more.

Hiring at Gridcare? Feature these jobs →

View all jobs at Gridcare
Gridcare logo
United States
Gridcare logo
United States
Gridcare logo
United States
Gridcare logo
United States
Gridcare logo
United States
Gridcare logo
United States
Number of employees

Dollar sign

$175,000.00 - $240,000.00

Full time
Energy