$209K to $270K a year as published by the employer
Applications are completed on the employer's own site.
Design and build compute platform infrastructure supporting production services, batch jobs, and ML training workloads for real-time battery asset operations.
The Role
We're looking for a Senior Software Engineer to join our Platform team and build the foundational infrastructure that powers Gridmatic. Our platform challenges are shaped by the nature of energy markets: forecasts and trading decisions run on tight schedules, battery dispatch commands must execute reliably in real time, and ML models need to train and deploy continuously as new data arrives.
What you'll do:
Design and build our compute platform, creating infrastructure that supports production services, batch jobs, and ML training workloads
Work on real-time systems for operating battery assets, where reliability directly impacts both revenue and grid stability
Establish patterns, tooling, and best practices that help teams across Gridmatic run services reliably
Make architectural decisions that shape how we build software as we grow
What we're looking for:
Significant experience building and operating production infrastructure on a public cloud platform (we run on GCP, but AWS or Azure experience translates well)
Hands-on experience with Kubernetes (we run on GKE and it's foundational to our platform)
Proficiency in Python
Experience with infrastructure-as-code tools like Terraform
Either already know Go or have experience with a similar systems language (C++, Java, Rust) and are excited to work in Python and Go day-to-day
Systems thinking—understanding how components interact, where failures can cascade, and how to build for efficiency, scalability, and resilience
Clear communication, whether writing a design doc, reviewing code, or explaining a complex system to someone new to it
Opinions about how to build reliable infrastructure, informed by experience with what works and what doesn't at scale
Nice to have:
Background in ML infrastructure or training pipelines
Experience with workflow orchestration tools (Flyte, Temporal, Airflow, or similar)
Familiarity with observability tooling (we use Grafana + Google Cloud Monitoring)
Experience with real-time systems
Strong Linux fundamentals
Prior work in domains where latency and reliability have direct business consequences
Taking care of you today:
Protecting your future for you and your family:
Pay range $209,000—$270,000 USD
and 4 more
Gridmatic is an AI-powered energy company focused on decarbonizing the grid by optimizing energy supply, demand, and transactions.