Hi, I'm Yazan.
I'm an AIOps Engineer. I specialize in fusing AI Agent Engineering with MLOps to transform infrastructure, turning manual operations into intelligent, self-healing systems powered by SRE and DevOps best practices.
Over the past 6+ years, I've architected and scaled cloud platforms across high-profile government institutions, Middle Eastern retail conglomerates, and fast-paced tech startups.
Beyond code and cloud platforms, I am an active contributor to open-source software, spending the last 4+ years sharing deep technical knowledge as a writer for the Fedora Project.
Professional Background
Senior DevOps Engineer
Architecting and owning the cloud infrastructure for an enterprise data governance SaaS platform, delivering secure, highly available deployments aligned with SDAIA & NDMO framework specifications to high-profile government and institutional clients across Saudi Arabia.
Terraform, ensuring scalable, repeatable, and auditable environments aligned with SDAIA (Saudi Data & AI Authority) and NDMO (National Data Management Office) regulations.Jenkins as a central automation server and engineered modular Ansible playbooks, eliminating manual ops toil and standardizing multi-client configuration management.Senior DevOps Engineer
Contracted to modernize DevOps practices at one of the Middle East's largest retail conglomerates, owning mobile release pipelines and cloud infrastructure automation end-to-end.
Fastlane, automating full build, test, and release workflows to App Store and Google Play — significantly reducing release friction.Terraform and Ansible, guaranteeing consistency across development, staging, and production environments.Helm charts, standardizing rollouts and eliminating manual intervention in production.Senior DevOps Engineer
Owned the complete DevOps lifecycle at a fast-scaling social media platform, building the infrastructure foundation to deploy code from commit to production with high velocity and zero downtime.
Jenkins, reducing manual deployment overhead and enabling low-risk microservice releases.Python, Bash, and Ansible that accelerated deployment cycles and eliminated human error.Zabbix for infrastructure metrics and ELK Stack for MongoDB cluster log analysis for proactive incident detection.Site Reliability Engineer
DevOps Engineer
Ansible for system & service configuration management.Extracurricular Activities
Technical Writer
Authoring and publishing deep-dive technical articles on container runtime engines (Podman), Linux system administration, and CLI workflows for the global Fedora developer community.
Working with Me
A brief blueprint of how I operate, communicate, and solve infrastructure engineering challenges.
Doc-First & Async Communication
I believe in writing things down. I document architecture decisions (ADRs), SRE processes, and runbooks to reduce context-switching and enable team members to move forward autonomously.
Toil Reduction & Automation
If a manual task has to be performed twice, it belongs in an automation script or declarative IaC configuration. Automation is key to eliminating operational friction.
Actionable Alerting Over Noise
Alerts should point to real customer impact and have clear runbooks. I strive to design high-fidelity, actionable monitoring setups to prevent pager fatigue.
Blameless Post-Mortems
Outages are learning opportunities. I advocate for blameless post-mortems that investigate systemic gaps in architecture or processes, not individual mistakes.
Elsewhere
Let's work together
Have a project in mind? Let's discuss how I can help with AIOps, AI Agent Engineering, MLOps, or cloud infrastructure.