Staff Backend Software Engineer, Developer Productivity Engineering
Posted about 22 hours ago
About Tapestry
Tapestry is Alphabet’s moonshot for the electric grid, operating at the convergence of energy infrastructure and advanced AI. Born at X (the innovation engine behind Waymo, Verily, and Google Brain), Tapestry builds computational and analytical platforms that make the world’s power grids visible, predictable, and resilient.
We provide AI-driven planning and simulation tools that allow system operators, utilities, and planners globally to operate more efficiently and integrate clean energy at scale. Tapestry currently collaborates with key partners across the U.S., U.K., Chile, New Zealand, Australia, and Brazil.
About the Role
We are seeking a Staff Software Engineer / Technical Lead to co-anchor the technical direction, systems architecture, and engineering standards for our Infrastructure and Developer Productivity team.
In this role, you will partner alongside another L6 Tech Lead and the Engineering Manager to set our multi-year platform roadmap. Our systems must process massive amounts of power-grid data, run heavy scientific simulations, and support emerging AI workflows. You will treat our internal infrastructure as a product—building reliable, self-service tools for our developers, automating our delivery pipelines, and creating secure environments where engineers and automated agents can build and test software quickly and safely.
How You Will Make 10X Impact
- Core Cloud & Compute Platform: Architect and operate scalable, multi-tenant Kubernetes clusters on Google Cloud Platform (GCP). Ensure efficient compute scheduling and autoscaling across mixed hardware workloads (CPUs, GPUs, and TPUs) supporting simulation and machine learning models. Design resilient networking, IAM boundaries, and secure multi-project cloud environments.
- Developer Velocity & CI/CD: Design fast, hermetic, and automated build, test, and release pipelines that reduce cycle times for product engineers. Provide reliable, on-demand testing environments so teams can validate changes safely before production. Build automated deployment and rollback mechanisms with clear canary verification.
- Infrastructure as Code & Security: Drive declarative, reproducible cloud infrastructure using modern Infrastructure-as-Code (such as Terraform). Implement automated policy checks, secret management, and container vulnerability scanning into standard deployment workflows.
- AI Tooling & Execution Sandboxes: Design secure, isolated sandbox environments that allow automated AI tools and agents to safely run tests and inspect code. Identify high-leverage opportunities to automate repetitive developer workflows using modern AI tools.
- Technical Co-Leadership & Mentorship: Partner with your fellow L6 Tech Lead to split architectural ownership, guide system designs, and run engineering design reviews. Mentor mid-level and senior engineers (L4/L5), raising the technical bar for code reviews, testing, and system design.
- Reliability & Observability: Define team-wide standards for metrics, logs, and distributed tracing to ensure high visibility into production health. Partner with data and ML teams to establish SLAs/SLOs, lead disaster recovery exercises, and run blameless post-mortems.
What You Should Have
- Technical Breadth & Depth: 8+ years of production experience in Infrastructure, Site Reliability Engineering, DevOps, or Developer Productivity/Platform Engineering.
- Technical Leadership: 2+ years serving as a formal Tech Lead or Staff Engineer directing the technical roadmap, architectural designs, and execution for a multi-pod or multi-team engineering surface.
- Modern Cloud & Orchestration: Advanced production-grade expertise with Kubernetes (GKE), container networking, and multi-tenant cloud architectures (Google Cloud Platform preferred).
- Infrastructure as Code (IaC): Deep architectural experience with modern declarative tools (Terraform, Pulumi, or similar) managing complex multi-environment cloud footprints.
- CI/CD & Developer Experience Primitives: Proven track record building large-scale, automated build/test/release pipelines (e.g., Tekton, GitHub Actions, Argo Workflows, Bazel) designed around self-service internal developer platforms.
- Platform-as-a-Product Mindset: Demonstrated ability to interview internal engineering stakeholders, quantify developer friction points, and deliver platforms that measurably increase overall deployment frequency and reduce MTTR.
Preferred Qualifications
- AI & Agentic Workloads: Practical experience designing infrastructure, sandboxes, and execution runtimes specifically geared toward AI agents, LLM evaluations, or high-performance GPU orchestration.
- Data Platform Adjacency: Hands-on architectural exposure to supporting large-scale data platforms (e.g., BigQuery, Spark, Kafka, Ray) or complex distributed simulation environments.
- Alphabet Ecosystem: Working knowledge of Google-internal infrastructure primitives (Borg, Monarch, Spanner, Piper/Blaze) or experience operationalizing an X moonshot into an independent production environment.
- Tier-1 Tech Platform Experience: Background engineering high-throughput platform tooling or developer infrastructure at companies operating at high engineering scale (e.g., Netflix, Snowflake, LinkedIn, Datadog).
Our Values
- Take charge: We take initiative and own outcomes that move the mission forward.
- Transform with purpose: We build solutions that solve real problems and create meaningful impact.
- Be a Tapestry, not a thread: We collaborate across diverse skills and perspectives to achieve more than we can individually.
- Always fine-tune: We stay curious, seek feedback, and refine our understanding as we learn.
- Stay grounded: We listen openly, value different perspectives, and stay focused on what matters most.
What we offer
A culture that supports growth, ownership, and meaningful impact, along with:
- Competitive salary and equity
- Medical, dental, and vision coverage
- Generous PTO and flexible hybrid work model
- 401(k) with employer contribution
- Professional development
- The ability to work on important real-world problems within an Alphabet-backed environment
The US base salary range for this full-time position is $207,000 - $290,000 + bonuses + equity + benefits. Our salary ranges are determined by role, level, and location. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific salary range for your location during the hiring process.
Please note that the compensation details listed in US role postings reflect the base salary only, and do not include bonus, or benefits.
Similar Software Engineer jobs(8)
Our global house of brands unites the magic of Coach and Kate Spade New York. By intertwining different people and ideas, we push ourselves in our work and expand the bounds of possibility. Learn about our iconic brands: tapestry.com/our-brands We’ve grown by finding people dedicated to the dream all over the world. We hold ourselves to high standards in every material and process, and we embrace difference by design because diverse perspectives are at the heart of creativity. We find brilliance in the intersections—of beauty and function, of heritage and innovation, of accessibility and aspiration—which is how we bring together magic and logic in our craft. Find out about our people and employer priorities: tapestry.com/responsibility/our-people The result is that we stand taller together, elevating the best in our people and brands. We use our collective strengths to move our customers and empower our communities, to make the fashion industry sustainable, and to build a house that’s equitable, inclusive, and diverse. Individually, our brands are iconic. Together, we can stretch what’s possible. See our values and commitments to support our people, communities and planet: tapestry.com/responsibility __ Please Be Advised - Recruitment Scams: Tapestry and its brands will only reach out to interview, make an offer of employment or conduct onboarding activities for candidates who have applied through our careers site. If you find a job posting on a third-party job site, such as LinkedIn, please know that a legitimate posting will direct you to our careers site to apply. When interviewing for a position, the candidate experience will include live interaction, such as a video conference or phone call, with a Recruiter and/or company employee(s). Be aware of suspicious recruitment activity. If you think you are a victim of an employment scam, please visit the Federal Trade Commission website: https://www.consumer.ftc.gov/articles/0243-job-scams
Key team members
Jobr aggregates jobs directly from company career portals — no middlemen. Our team applies on your behalf with AI-tailored resumes, reviewed by a human before submission.