About
Rajat Bansal is a Site Reliability Engineer (SRE) with over three years of hands-on experience in designing, building, and managing scalable, reliable, and efficient systems. In his current role, he specializes in cloud infrastructure management across platforms such as AWS and GCP, ensuring seamless system performance and long-term reliability. His expertise spans automation, observability, and CI/CD pipelines, where he leverages tools like Datadog, Prometheus, and Kubernetes to deliver high availability and smooth deployments.
Known for his problem-solving mindset, Rajat thrives in environments that demand both technical precision and innovative thinking. He has a proven ability to diagnose and resolve complex infrastructure challenges, optimize system performance, and drive automation across workflows. His focus has consistently been on creating solutions that reduce manual intervention, improve reliability, and enable faster delivery cycles—making him a valuable asset in modern DevOps and cloud-native ecosystems.
Beyond his technical strengths, Rajat is deeply passionate about continuous learning and knowledge sharing. He has contributed to the developer community through course creation on Udemy, delivering content on Prometheus, Computer Vision, and Machine Learning, which have been highly rated by learners globally. This reflects not only his expertise but also his commitment to making technology accessible and impactful. With his strong foundation in cloud technologies, programming (Python), monitoring systems, and automation practices, Rajat continues to grow as a versatile professional who is always eager to learn, build, and contribute to meaningful systems that make a difference.
Top Skills: Amazon Web Services (AWS) • Prometheus.io • Python • Kubernetes • CI/CD
"Site Reliability Engineer 2"
