Site Reliability Engineer
crisil
Job Description
- Designs, develops, and oversees the implementation of advanced cloud-native solutions, ensuring alignment with business objectives and adherence to best practices while managing end-to-end technical delivery of projects.
- Demonstrates expertise in containerized environments and CI/CD pipelines to optimize deployment efficiency and system reliability.
- Leads the evaluation and integration of emerging cloud technologies and infrastructure patterns, driving innovation and continuous improvement to enhance system performance, scalability, and user experience across distributed environments.
- Executes software development tasks and projects requiring advanced problem-solving, decision-making, and strategic thinking, with particular focus on cloud optimization and burst-compute scenarios.
- Leverages advanced and creative skills to resolve complex software development and cloud infrastructure challenges, fostering cross-functional collaboration to identify and implement innovative solutions.
- Demonstrates deep familiarity with observability concepts and has successfully implemented monitoring, logging, and tracing solutions to enhance system visibility and operational insights.