Role overview
About this role
At IBM Finance & Operations, we are the backbone of IBM’s transformation driving efficiency, transparency, and smart decision-making across the business. Our teams provide the insight and discipline that guide strategy, ensure financial strength, and enable IBM to invest in innovation and growth. Working in Finance & Operations means combining analytical skills with collaboration and curiosity. You’ll partner with colleagues across functions and geographies, using data, technology, and process excellence to create solutions that improve performance and deliver measurable impact. IBM offers continuous learning, career development, and a culture that values diverse perspectives. Join us and be part of a global team that keeps IBM moving forward, while building your own future in a dynamic and evolving environment. The Associate Software Engineer, SRE is responsible for following detailed instructions and standardized procedures to support the reliability and performance of services. This role focuses on executing routine tasks and resolving issues of minimal complexity, ensuring operational stability. The Associate Software Engineer, SRE works on defined assignments, applying systems and software engineering principles in a structured environment. While contributing to Service Level Objectives (SLOs), the Associate Software Engineer, SRE will gain exposure to the core concepts of site reliability, automation, and monitoring, and will work under guidance to develop skills in managing distributed systems. Implement well-defined scripts and procedures to assist in maintaining the reliability and availability of services. Contribute to routine code updates and participate in peer reviews as part of a structured development process. Monitor service performance and assist in responding to incidents according to established procedures. Collaborate with team members to share insights and learn from more experienced engineers. Engage in scheduled on-call rotations, addressing basic incidents and escalating complex issues as needed. Assist in the post-incident review process, documenting findings and contributing to incremental improvements. Follow detailed instructions to maintain and enhance systems, contributing to ongoing system health. • Software and Systems Knowledge: Exposure to software and systems engineering principles, with an understanding of how to design, build, and maintain reliable and resilient systems. • Problem Analysis and Resolution: Experience working with problem determination methodologies, analyzing complex issues, and applying technical expertise to resolve problems affecting system performance and reliability. • System Design and Architecture: Exposure to system design and architecture principles, with an understanding of how to design and build well-engineered information systems and ecosystems. • Testing and Deployment: Experience working with testing and deployment methodologies, ensuring seamless deployment and minimal disruption to services. • Reliability and Resiliency: Exposure to reliability and resiliency principles, with an understanding of how to analyze business needs and provide recommendations for enhancing system reliability and resiliency. • Cloud Computing Knowledge: Exposure to cloud computing platforms and technologies, with an understanding of how to design, build, and maintain reliable and resilient cloud-based systems. • Scripting and Automation: Experience working with scripting languages and automation tools, analyzing complex issues, and applying technical expertise to resolve problems affecting system performance and reliability. • IT Service Management: Exposure to IT service management principles, with an understanding of how to design and build well-engineered information systems and ecosystems that meet business requirements and reliability standards.