Background
The energy transition toward a carbon-free electric grid poses considerable challenges. The growing integration of renewable energy requires high-performance storage solutions and robust grid balancing mechanisms. Batteries make it possible to integrate renewable energy into the grid, but their operation and optimization are complex. In fact, a battery can generate revenue through various mechanisms, each with its own specific characteristics, but none of these mechanisms, taken in isolation, guarantees the system’s economic viability. To ensure profitability and accelerate the deployment of these technologies, it is necessary to operate and optimize the battery within a comprehensive framework.
StackEase addresses this challenge by developing algorithms that maximize net revenue from batteries based on market prices and the physical constraints of each site. The services we offer include consulting—generating long-term revenue forecasts—and operations—optimizing the assets entrusted to us in real time.
Operations involve defining the orders that correspond to the most efficient use of each battery, and then submitting them to the electricity markets. Depending on market acceptance, we then control the battery through our locally installed edge server. We began operations this year and are looking to scale them up.
About StackEase
StackEase is a deep-tech spin-off from INRIA (the French National Institute for Research in Computer Science and Control). Based in Paris and Marseille, we bring together talented individuals who are passionate about technological innovation. Our values are innovation, operational excellence, customer satisfaction, meritocracy, and a commitment to sustainability.
Missions
Manage and improve the reliability of our AWS infrastructure
- Operating and optimizing our ECS clusters on EC2: diagnosing issues related to instances, containers, and available resources.
- Manage and anticipate capacity requirements (autoscaling, memory, storage, availability) and analyze infrastructure performance.
- Manage network components: VPCs, subnets, routing, DNS, VPNs, and security groups.
- Manage related services: RDS, S3, CloudWatch, IAM, Secrets Manager, etc.
- Strengthen infrastructure and network security: segmentation, flow control, access management, configuration hardening, and application of the principle of least privilege.
- Help with infrastructure sizing and AWS cost optimization.
Improving Operations, Resilience, and Observability
- Define and evolve our observability stack (metrics, logs, dashboards, alerts) to quickly detect, understand, and diagnose performance degradations and incidents.
- Diagnose production incidents, conduct root cause analyses, and implement corrective actions to prevent their recurrence.
- Define and improve backup, restoration, and disaster recovery mechanisms to enhance the platform's resilience and service continuity.
- Document operational procedures, runbooks, and diagnostic tools.
Industrialize the platform and deployments
- Maintain and improve our Infrastructure as Code (AWS CDK / TypeScript).
- Maintain and improve our GitHub Actions pipelines.
- Develop automation scripts and tools, primarily in Bash and Python, to automate recurring tasks and reduce the need for manual intervention.
- Improve the mechanisms for deploying, rolling back, and managing environments.
- Maintain and upgrade our self-hosted Apache Airflow instance on ECS.
Develop the platform's services and integrations
- Develop and maintain the Python APIs and services necessary for the platform's operation and management.
- Design integrations with our clients’ and partners’ information systems: APIs, file transfers, and other automated interfaces.
- Develop jobs and workflows to orchestrate internal and external processing, including asynchronous processing on AWS.
- Ensure the robustness of these integrations: authentication, error handling, retries, timeouts, idempotence, logs, monitoring, and testing.
- Help define the interfaces and data exchange protocols with our partners.
Desired Qualifications
You have at least 5 years of experience working on infrastructure, systems, DevOps, or a related field.
Above all, we are looking for someone with a solid foundation in cloud computing, systems, and networking, who is also comfortable with software development in Python. You will be responsible for developing automation tools as well as services and integrations with external systems. Experience with IoT or edge computing is a plus, but not required.
You will define best practices and necessary changes, while remaining directly involved in their implementation and in resolving any issues that arise during production.
Key Skills
- Proficiency in Linux and system diagnostics.
- Strong knowledge of networking : TCP/IP, DNS, routing, VPN, firewall.
- Significant experience in AWS.
- Proficiency in Docker and containerized environments.
- Significant experience operating a container orchestration platform in a production environment; experience withAWS ECS on EC2 is particularly valued.
- Proficiency inInfrastructure as Code.
- Proficiency in Git and collaborative development workflows.
- Proficiency in Python and the ability to develop services, APIs, and tools that are maintainable in production, using best practices for testing and code quality.
- Experience developingHTTP/REST APIs and integrating external services.
- A solid understanding of the challenges involved insystem integration : authentication, data and file exchange, asynchronous processing, error handling, retries, timeouts, and idempotence.
- Proficiency in system scripting, particularly in Bash.
- Experience in monitoring, observability, diagnostics, and incident management in production.
- A solid understanding of issues related to system, network, and cloud security.
Desirable Skills
- Experience IoT / Edge Computing and management of Linux-based equipment fleets.
- AWS IoT Core.
- MQTT and asynchronous/event-driven messaging architectures.
- AWS CDK / TypeScript.
- GitHub Actions.
- Apache Airflow.
- PostgreSQL / RDS.
- Grafana or other observability tools.
Benefits
- Competitive compensation.
- An opportunity to work on challenging projects that have a significant impact on the energy sector.
- A collaborative work environment that fosters innovation and professional growth.
- Remote work up to 2 days a week.
We look forward to receiving your application and a letter explaining why you’d like to join us at jobs@stackease.ai. If you don’t think you have all the required skills, please feel free to apply anyway—we’ll review your application.