Actively Looking for Work
Site Reliability Engineer with 2+ years of experience supporting large-scale enterprise and cloud-hosted infrastructure environments, including an environment of 5,000+ servers with a strong focus on reliability, availability, incident response, monitoring, and operational stability.
In my current SRE-focused experience, I have worked with Linux-based infrastructure, supporting day-to-day production operations and ensuring systems remain stable, available, and performant. I have hands-on experience with incident management, change management, infrastructure monitoring, alerting, troubleshooting, and operational support in enterprise environments. My experience includes working with monitoring and observability platforms such as Dynatrace, where I have worked with infrastructure and application-related monitoring, alerts, incident investigation, and identifying potential reliability issues.
I have practical experience with Linux administration and LDAP administration, along with troubleshooting system-level issues and supporting enterprise infrastructure. I am comfortable working from the command line, investigating logs, analyzing system behavior, troubleshooting services, and following structured operational processes during incidents and changes.
Alongside my production support and SRE experience, I have been developing my skills across modern cloud and DevOps technologies. I have hands-on knowledge of AWS, Docker, and Jenkins, and I am actively expanding my capabilities in cloud infrastructure, containerization, CI/CD, automation, and Kubernetes-based deployments.
My current learning and project focus is on production-grade Kubernetes and Microsoft Azure Kubernetes Service (AKS). I am following a CKA-oriented learning path and building practical Kubernetes skills around deployments, services, networking, configuration, storage, troubleshooting, security, scaling, and cluster operations. My goal is to combine this Kubernetes knowledge with my existing SRE background and move further into DevOps, SRE tooling, infrastructure automation, and cloud-native engineering.
I am particularly interested in roles where I can work at the intersection of reliability, cloud infrastructure, automation, observability, and platform engineering. I enjoy solving infrastructure problems, investigating incidents, improving operational processes, and learning how systems can be designed and operated more reliably at scale.
My technical areas of interest and experience include:
I bring a combination of real-world production infrastructure experience, troubleshooting ability, reliability-focused thinking, and continuous hands-on learning. Rather than limiting myself to a single technology, I am focused on understanding how infrastructure, applications, monitoring, automation, containers, and cloud platforms work together to build reliable production systems.
My long-term goal is to become a strong Cloud/SRE/Platform Engineer capable of designing, automating, monitoring, troubleshooting, and operating highly available infrastructure and cloud-native workloads at scale.