GridCARE logo

Senior Platform Engineer

Posted about 11 hours ago

RemoteRedwood CitySE170k - 195k USD

Senior Software Engineer, Platform

Location: Redwood City, CA (3 days/week in office)

Job Type: Full-time · Hybrid

About Us

GridCARE is the pioneer of Power Acceleration — a new category solving the most critical constraint in AI's growth trajectory: immediate access to power. As demand for computing skyrockets, access to energy has become the defining bottleneck in the AI infrastructure race. While leading tech companies invest billions in speculative, long-term solutions that may take decades to arrive, GridCARE's Energize™ platform uses physics-based AI to identify and activate gigawatts of hidden capacity in today's electric grid — enabling hyperscalers, data center developers, and utilities to power AI infrastructure years sooner than conventional approaches and without costly upgrades.

Founded at Stanford's Doerr School of Sustainability, GridCARE has assembled a world-class team spanning power systems, AI, and infrastructure, and is backed by leading climate-tech and deep-tech investors.

At GridCARE, you will:

⚡ Work at the intersection of AI, energy, and infrastructure — the foundation of the next industrial revolution.

🤝 Partner with hyperscalers, developers, and utilities on high-impact, real-world deployments.

🌎 Help shape a more abundant, efficient, and resilient energy future for the digital era.

🚀 Join a company defining a new category — Power Acceleration for AI.

💰 Receive competitive compensation, equity, and benefits in a fast-growth, mission-driven environment.

Learn more about GridCARE:

The Role

We’re looking for a Senior Software Engineer, Platform to build the shared software foundations that let GridCARE turn grid intelligence into products: secure APIs, tenant-aware authorization, integrations with job orchestration systems, and reusable application services.

Working with our tech lead, product engineers, data owner, and power systems team, you’ll take these capabilities from architecture through production adoption—making them easy for developers and AI agents to discover, integrate, and operate. You’ll independently resolve implementation choices, build on managed services, and own rollout and production behavior, with the tech lead guiding platform architecture. You’ll raise the engineering bar through design reviews, code review, and mentoring.

Responsibilities

  • Build the API platform. Use managed gateway and identity services to support browser applications and machine clients. Own the integration code, trusted identity and tenant context, request validation, and clear, versioned API contracts.

  • Partner on job orchestration. Work with data and power systems engineers to integrate job workflows with shared platform services. Build the API and access-control interfaces that connect products to these workflows, preserving tenant and study context through job submission, status, and result access.

  • Make authorization and tenant isolation dependable. Implement shared access controls across APIs, services, jobs, and data interfaces. Enforce resource ownership and study boundaries, including for internal users authorized to work with multiple customers, and define how permission changes affect ongoing work.

  • Make the platform easy to build on. Create reusable APIs, libraries, application templates, machine-readable contracts, and executable examples. Help developers and AI agents discover capabilities, compose workflows, and recover from errors.

  • Make systems dependable in production. Instrument services, investigate failures and performance bottlenecks, and partner with SRE on deployment, observability, and recovery.

  • Drive delivery and adoption. Turn ambiguous needs into incremental releases, make pragmatic build-versus-buy decisions, and help existing products adopt shared authentication, authorization, and execution patterns. Evolve interfaces safely and work with data and power systems engineers on integration boundaries.

  • Build reliable usage metering. Capture durable API and job events, attribute usage to the right tenant and principal, and handle retries and deduplication so usage records remain accurate.

Qualifications

Required

  • 5+ years of relevant software engineering experience, with a track record of building shared backend capabilities for a multi-tenant product and owning their rollout and operation in production.

  • Strong Python engineering skills and experience building production APIs and services with clear interfaces, tests, and maintainable data models.

  • Hands-on experience building shared APIs, workflow services, or identity and access capabilities on managed services, with ownership of integration code, policies, and production behavior.

  • Strong distributed systems fundamentals: you can reason about concurrency, partial failures, retries, idempotency, consistency, and backpressure.

  • Practical experience implementing multi-tenant authorization: resource ownership, fine-grained permissions, trusted context across service boundaries, and access controls that hold through asynchronous execution.

  • Experience with relational databases, queues, cloud services, and diagnosing production behavior through logs, metrics, and traces.

  • Sound technical judgment and the ability to independently turn an unclear requirement into a reliable system adopted by other engineers. You can work with a tech lead on architecture and coordinate delivery across product, data, and domain teams.

  • Clear communication, thoughtful code review, and an interest in mentoring teammates. You can explain tradeoffs and work constructively across teams.

Required

  • Experience with FastAPI, Pydantic, OpenAPI, PostgreSQL, or the wider Python service ecosystem.

  • Experience with durable workflow or job orchestration systems such as Temporal, Prefect, or Airflow.

  • Experience with OAuth2/OIDC, Auth0 or similar identity providers, service identities, and relationship-based authorization systems such as OpenFGA.

  • Familiarity with AWS, Kubernetes, object storage, and OpenTelemetry; experience with durable event processing or usage metering.

  • Experience supporting computationally intensive, geospatial, time-series, or AI workloads.

  • Effective use of AI coding tools, with disciplined review and testing of generated code.

What We Offer

  • Competitive salary, performance bonus, and equity.

  • Comprehensive health, dental, and vision coverage.

  • Lunch provided three days a week in office.

  • Hybrid schedule: 3 days in office for collaboration, 2 days remote for focused work.

  • Access to leading academic, industry, and government partners in the AI-energy ecosystem.

  • A mission-driven team focused on shaping the future of the energy transition.

Location

  • This is a hybrid, in-office role based in Redwood City, CA. Employees work onsite three days/week, Tu-Thurs.

Salary Range

$170,000 - $195,000 Total

Join us in tackling one of the most important infrastructure challenges of our time — enabling the energy foundation for the age of AI.

About GridCARE

GridCARE is the pioneer of Power Acceleration, a new approach to delivering power for AI infrastructure. The company’s Energize™ platform uses advanced AI models, grid simulations, and real-time system intelligence to identify and activate latent capacity across existing power infrastructure. By coordinating utilities, AI infrastructure developers, and flexible energy resources, GridCARE enables large AI factories to secure and activate power in months instead of years. GridCARE is helping power the next generation of AI infrastructure while preserving grid reliability and accelerating economic growth.

51Energy Technology
Apply smarter with Jobr

Jobr aggregates jobs directly from company career portals — no middlemen. Our team applies on your behalf with AI-tailored resumes, reviewed by a human before submission.

Direct from company career pages
AI-personalised cover letters
Human review before every submit
Application tracking & follow-ups