Our Product Development (PD) Platform Operations team is responsible for the 24x7 uptime, stability, and performance of our internally customized deployment of the 3DX PLM (Product Lifecycle Management) platform — including its underlying infrastructure, network, and application layers. This platform is mission-critical to our engineering and manufacturing organizations.
We are looking for a Senior SRE & Monitoring Developer who combines deep hands-on Elastic Stack expertise with strong site reliability engineering practices to help us proactively detect, diagnose, and prevent issues before they impact our internal customers.
Elastic Platform Ownership: Maintain deep expertise in Elastic, including cluster management, performance tuning, index/shard optimization, and Fleet & APM management. Conduct deep dives into Elastic query performance and resource consumption, specific to the Dassault 3DX product line. Manage index lifecycle policies (ILM) to balance performance, cost, and retention requirements
Proactive Monitoring & Alerting: Design, implement, and manage comprehensive monitoring and alerting systems across our platforms, with a specific focus on Elastic clusters. Define key metrics (SLIs), thresholds (SLOs), and escalation procedures to proactively identify and address potential issues before they become incidents. Develop and maintain a suite of automated health checks for critical endpoints, APIs, infrastructure components, and network paths
Author advanced ES|QL (Elasticsearch Query Language) queries to aggregate, transform, and analyze log, metric, and trace data for performance analysis, capacity planning, and root cause investigations.
Utilize KQL (Kibana Query Language) to create efficient search filters, Discover saved views, dashboard controls, and alerting rule conditions. Translate ad-hoc operational and diagnostic questions into performant ES|QL pipelines that can be operationalized into Kibana dashboards or scheduled alerts.
Design and build advanced custom Kibana visualizations using Vega and Vega-Lite for complex use cases beyond out-of-the-box Kibana Lens capabilities (e.g., SLO burn-rate tracking, multi-layered service performance visuals, topology maps). Develop and maintain operations, executive, and incident dashboards combining standard Kibana panels with custom Vega visual specifications.
Performance Analysis & Optimization: Conduct regular performance analysis of platforms, identifying bottlenecks and implementing optimizations to improve responsiveness, scalability, and resource utilization. This includes deep dives into Elastic query performance & resource consumption – specific to the Dassault 3DX product line
Automation & Tooling: Build and maintain scripts/tools (Python, Bash, or Go) to streamline health checks, monitoring integrations, and routine operational tasks
Collaboration & Enablement: Work closely with infra teams to ensure the reliability and scalability of applications that interact with Elastic. Provide guidance to engineering teams on Elastic best practices, query optimization, and observability instrumentation

We don't just make history -- we make the future. Ford put the world on wheels over a century ago, and our teams are re-inventing icons and creating groundbreaking connected and electric vehicles for the next century. We believe in serving our customers, our communities, and the world. If you do, too, come move the world and make the future with us.
Ford is a global company with shared ideals and a deep sense of family. From our earliest days as a pioneer of modern transportation, we have sought to make the world a better place – one that benefits lives, communities and the planet. We are here to provide the means for every person to move and pursue their dreams, serving as a bridge between personal freedom and the future of mobility. In that pursuit, our 186,000 employees around the world help to set the pace of innovation every day.
Privacy Policy: https://www.ford.com/help/privacy/