This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Python Developer based in United States.
This role sits within a Site Reliability Engineering team focused on automating operations across a large-scale global network infrastructure.
You will design and build software that reduces operational toil, improves reliability, and enables infrastructure teams to work more efficiently.
The position combines software engineering, network automation, observability, and cloud-native technologies in a highly interconnected environment.
You will work closely with architecture, engineering, and operations teams to identify bottlenecks and develop scalable solutions.
A key focus will be strengthening network telemetry, automation, monitoring, alerting, and operational workflows through robust software systems.
You will contribute to reliable development practices, including CI/CD, test-driven development, infrastructure as code, and iterative system design.
This is an opportunity to work on complex global systems where your engineering contributions directly support the reliability and scalability of critical digital infrastructure.
Accountabilities:
- Design and build systems that automate operational workflows and provide deep insights into global network infrastructure and data.
- Identify and eliminate operational bottlenecks caused by inefficient processes, manual work, and technical debt.
- Develop software solutions that reduce operational toil and improve the efficiency and reliability of infrastructure operations.
- Generate telemetry data and strengthen observability across complex network infrastructure.
- Design and maintain robust, self-sustaining, and maintainable code with a focus on CI/CD, test-driven development, and iterative engineering practices.
- Support and deploy cloud-native and third-party software using infrastructure-as-code principles.
- Contribute to the design and implementation of full-stack systems using appropriate tools and frameworks for each technical challenge.
- Maintain and enhance network telemetry and automation platforms, including source-of-truth systems, metric agents, time-series databases, dashboards, and alerting systems.
- Apply software development lifecycle best practices consistently to improve engineering efficiency and system quality.
- Develop a strong understanding of complex, globally interconnected systems and proactively identify opportunities to improve their reliability.
- Collaborate closely with architecture, engineering, and operations teams to identify high-impact improvements and deliver scalable technical solutions.
- Strong software engineering skills with an interest in network automation, infrastructure reliability, and operational tooling.
- Experience contributing to full-stack software systems and selecting appropriate tools and frameworks for technical requirements.
- Strong programming experience, particularly with Python or comparable software development technologies.
- Understanding of software development lifecycle practices, including CI/CD, automated testing, test-driven development, and iterative design.
- Experience or strong familiarity with observability concepts, including telemetry, metrics, dashboards, monitoring, and alerting.
- Understanding of network infrastructure, automation, or complex interconnected systems is highly relevant to the role.
- Experience with cloud-native technologies and infrastructure-as-code approaches is preferred.
- Ability to identify operational inefficiencies and translate them into practical, scalable software solutions.
- Strong problem-solving skills and the ability to work collaboratively across engineering, architecture, and operations teams.
- Ability to understand complex systems and proactively contribute to improving their reliability, maintainability, and scalability.
- Strong focus on writing robust, maintainable software and continuously improving engineering practices.
- Remote work model in the United States.
- Flexible approach to working from home, in an office, or through a combination of both, depending on role and individual needs.
- Benefits supporting health, well-being, financial security, and life beyond work.
- Opportunity to work on highly distributed global infrastructure and complex networking technologies.
- Collaborative environment alongside architecture, engineering, and operations professionals.
- Opportunity to contribute directly to automation, observability, reliability, and scalability initiatives supporting critical digital services.

