Join the dynamic Global Operations Excellence (GOE) team, where innovation meets impact. We're at the forefront of developing cutting-edge solutions that drive efficiency and intelligence across Siemens Digital Industries Software. As part of our team, you'll contribute to exciting projects enhancing our operational insights and automation, streamlining processes across teams. We foster a collaborative environment where creativity is encouraged, and your contributions directly shape the future of our digital landscape.
Your responsibilities will include management of our active platform initiatives including our Command Center, providing full visibility of our environment and services as a single pane of glass. Also, our full-stack AI infrastructure operations platform including a chat bot and our end-to-end budget tracking system amongst other projects. You will be responsible for the health, design, implementation, and production reliability of the platforms.
Key Responsibilities
As a Full Stack Developer, your responsibilities will include:
• Platform accountability: Take technical accountability for maintaining, evolving, and operating production platform initiatives, with readiness to absorb additional initiatives in the future.
• Backend Services: Design, implement, and operate platform components using
Python, with an emphasis on reliability, performance, and long-term maintainability.
• AI Agent Development: maintain AI agents using Python, working with LLM APIs, agent
frameworks, prompt design, tools, guardrails, and fallback flows. Design agents that
support infrastructure workflows such as VM provisioning, troubleshooting, incident triage, and operational recommendations.
• APIs & Integrations: Build and evolve APIs and integration services for internal and external consumers. Integrate with cloud, virtualization, monitoring, ITSM, and infrastructure systems.
• Full-Stack Command Center Development: Maintain web applications using React, Next.js, FastAPI, and TypeScript.
• System Design: Contribute to architecture discussions, evaluating trade-offs between simplicity, scalability, and operational complexity.
• Monitoring, Logging & Alerting: Improve system operability by enhancing logging, metrics, alerting, observability, and deployment workflows.
• Automation & CI/CD: Maintain, and optimize CI/CD pipelines to automate delivery, testing, and deployment processes.
• Code Quality: Participate in code reviews and design discussions, maintaining high standards for code quality, testing, and documentation.
• AI Risk Management: Monitor agent accuracy, failures, latency, cost, and user
feedback. Manage AI risks such as hallucinations, unsafe recommendations, prompt injection, and unauthorized actions.
• Production Support: Own incident triage, troubleshooting, and resolution for all
Platforms.
Skills Required (Non-Technical)
• Proven ability to own and deliver complex features from design through production.
• Strong verbal and written communication skills, with the ability to articulate complex technical concepts clearly.
• Ability to work effectively in a collaborative, global, agile environment.
Skills Required (Technical)
Python Backend & Platform Engineering
• Strong Python development experience at a senior level.
• Experience building APIs using FastAPI.
• Experience designing backend services for agent execution, workflows, infrastructure actions, and integrations.
• Experience building background jobs, queues, schedulers, and automation workflows.
• Ability to write reliable, testable, maintainable code and debug production issues.
AI Agent & LLM Application Development
• Experience building and maintaining AI agents using Python.
• Working knowledge of LLM APIs, agent frameworks, prompt design, tools, guardrails, and fallback flows.
• Experience using internal documentation, runbooks, logs, tickets, and knowledge bases for retrieval-augmented responses.
React / TypeScript Full-Stack Development
• Experience with React and Next.js, including server-side rendering, API routes, and modern frontend architecture patterns.
• Strong focus on developer experience and maintainable, scalable frontend codebases.
• Experience integrating frontend applications with Python backend APIs.
Infrastructure & Cloud Operations
• Experience with cloud platforms, primarily AWS.
• Understanding of virtualization platforms, infrastructure APIs, and infrastructure-as-code concepts.
• Experience with monitoring, logging, alerting, and operational runbooks.
Cloud & DevOps
• Experience with CI/CD pipelines.
• Experience with containerization and orchestration using Docker and Kubernetes.
• Solid understanding of distributed systems and cloud-native architectures.
Education & Experience

Siemens AG (Berlin and Munich) is a leading technology company focused on industry, infrastructure, mobility, and healthcare. The company’s purpose is to create technology to transform the everyday, for everyone. By combining the real and the digital worlds, Siemens empowers customers to accelerate their digital and sustainability transformations, making factories more efficient, cities more livable, and transportation more sustainable. A leader in industrial AI, Siemens leverages its deep domain know-how to apply AI – including generative AI – to real-world applications, making AI accessible and impactful for customers across diverse industries. Siemens also owns a majority stake in the publicly listed company Siemens Healthineers, a leading global medical technology provider pioneering breakthroughs in healthcare. For everyone. Everywhere. Sustainably. In fiscal 2025, which ended on September 30, 2025, the Siemens Group generated revenue of €78.9 billion and net income of €10.4 billion. As of September 30, 2025, the company employed around 318,000 people worldwide on the basis of continuing operations.