Company Description
Duracell Power Center, a Duracell Authorized Licensee, focuses on empowering homeowners to self-power their homes with clean, affordable, and reliable energy. The company designs, refines, and manufactures smart energy storage solutions that benefit both residents and the environment. Its flagship product, the Duracell Power Center, is a highly flexible home battery and solar storage solution from a trusted global power brand. Featuring industry-leading Lithium Iron Phosphate batteries, these systems are modular, easy to install, and compatible with both new and existing home energy setups. Team members collaborate on advanced energy technologies that directly support the transition to sustainable home power.
Role Description
The Senior Platform Engineer - Data Reliability is a full-time, on-site role based in San Jose, CA. This role is responsible for designing, building, and maintaining data platforms that ensure high reliability, scalability, and performance for energy storage and monitoring systems. Day-to-day tasks include developing and optimizing software components, implementing robust data pipelines, managing infrastructure and database environments, and monitoring system health and data integrity. The engineer will troubleshoot complex issues across applications, infrastructure, and data layers, and implement automation to improve reliability and observability. Collaboration with cross-functional teams in engineering, product, and operations is expected to align platform capabilities with business and customer needs.
Qualifications
- Candidates should possess strong Programming and Software Development skills for building and maintaining reliable platform services.
- Candidates should possess solid Infrastructure skills for designing, deploying, and operating scalable, cloud or hybrid environments.
- Candidates should possess robust Databases skills for modeling, managing, and optimizing data storage and retrieval.
- Candidates should possess advanced Troubleshooting skills for diagnosing and resolving issues across software, infrastructure, and data systems.
- Experience with data reliability, observability, and monitoring tools (e.g., logging, metrics, tracing) is highly beneficial.
- Background in distributed systems, data pipelines, and API-driven architectures is preferred.
- Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
- Ability to work collaboratively on-site in San Jose, CA and communicate clearly with technical and non-technical stakeholders.
The role
This hire becomes the second platform engineer on a system that currently has one. The job is to make the data we show customers provably true and the production runtime reliably boring - and to own both.
Areas of ownership
- Ingestion truthfulness. Multiple vendor APIs, each with its own auth quirks, rate limits, and freshness semantics. Owns the pipelines, the freshness/coverage metrics, and surfacing \"how current is this, really\" to the product.
- Alert lifecycle. Alerts that open, resolve, and mean something - deduplicated correctly in a multi-tenant system.
- Scheduled-job runtime. Everything periodic runs as declared infrastructure (EventBridge/ECS via CDK), observed and alarmed.
- Deploy discipline. Pinned deploys from reviewed, merged code; migration gates; a meaningful staging environment; CI kept green and required.
- Incident response. First responder for pipeline/vendor incidents, with the technical lead as escalation - and makes the alarms themselves trustworthy.
- Second reviewer. Substantial backend changes get two sets of eyes, in both directions.
First 90 days
- Day 30: Freshness measurement productionized; on the alarm rotation; operations runbook co-written; reviewing pull requests.
- Day 60: Alert lifecycle working end-to-end; all periodic jobs scheduled as code; running deploys; CI enforced.
- Day 90: Owns pipeline and runtime outright, with escalation support only.
Qualifications — must have
- 5+ years backend engineering with real production ownership
- Python async services (FastAPI/SQLAlchemy or equivalent) and PostgreSQL at scale - time-series experience a strong plus
- AWS containers and infrastructure-as-code (ECS/Fargate, CloudWatch; CDK or Terraform)
- Experience integrating and debugging third-party APIs that misbehave
- Fluent, critical use of AI coding tools - generates fast and verifies; runs the test before trusting the claim
- Evidence-first habits - asks for the denominator, distrusts green dashboards
Nice to have
- Kafka; metric/alarm design; solar/energy or IoT fleet domain; high-velocity migration hygiene
- Medical, dental, and vision
- 401(k)
- Paid Time Off (PTO)
- Paid company holidays
Salary Range
$170,000 to $200,000 yearly, based on skills and experience