Habtamu Asfaw

Site Reliability Engineering (Sre) @Dotdash Meredith

Seattle, WA, US
MOBILE NUMBERS
+91 *********19

Signup · Get unlimited contacts

WORK HISTORY

Nov 2021 — Present

Site Reliability Engineering (Sre) @Dotdash Meredith

View department →

Seattle, WA, US

Implement and automate DevOps practices to reduce the level of incidents and improve reliability and scalability. Reporting bugs to the Core development team and getting involved in debugging if it is a production outage, responsible for debugging and fixing the infrastructure issues. Develop typical measurement metrics, for Error Budgets, SLOs (Service Level Objectives), SLIs (Service Level Indicators), and SLAs (Service Level Agreements). Conduct the Post-Incident reviews to identify the root cause and document the findings to provide feedback to the core development team. Record and document results and compare to expected results. Observability tools: Prometheus and Grafana for collecting and visualizing the different metrics, incident alert tools (VictorOps, PageDuty), Ansible, Docker, and Kubernetes for container orchestration, cloud platform AWS, GCP, Azure, JIRA, GitHub. Focus on resilience, scaling, reliability, uptime, and robustness. Ensuring reliability - getting systems back to steady state as quickly as possible

EDUCATION

2017 — 2018

University of Washington

PCE

2018 — 2019

University of Washington

PCE

2003 — 2006

Bahir Dar University

Bsc

2024 — 2024

Coding Dojo

Full-Stack Developer Certificate

2016 — 2017

University of Washington

PCE

ABOUT HABTAMU ASFAW

As a multi-cloud environment technical expert with over 7 years of experience, I have a strong background in system engineering, building, delivering, and managing cloud solutions across various platforms including MS Azure, AWS, and GCP. I possess extensive knowledge of IaaS, PaaS, and SaaS platforms and am well-versed in using CI/CD stacks and DevOps tools such as Git, Bitbucket, Jenkins, TeamCity, Octopus, ProGet, CircleCI, and VSTS. My expertise also includes using Python scripting language and cloud-native technologies such as Docker container and Kubernetes. I am experienced with configuration management systems like Ansible and Infrastructure as Code (IaC) tools like TerraForm and AWS infrastructure deployment through Cloud formation. In addition, I have a strong background in end-to-end observability and visibility for business-critical systems with log ingestion, metrics, and traces. I am skilled at driving the adoption of best practices in monitoring, alerting, automation, and site reliability to ensure the highest level of performance and stability. Overall, my technical expertise and hands-on experience make me a valuable asset in any cloud-related project. CERTIFICATES: AWS Certified Solution Architect - Associate (SAA-C02) 2020 • Candidate ID # AWS00••••07 Python Programming, 2019 • UW Professional & Continuing Education AWS Certified SysOps Administrator - Associate (SOA-C00) 2018 • Registration # 270477 Cloud Architecture on Amazon Web Services Cert, 2018 • UW Professional & Continuing Education Administering Large Unix and Linux Systems, 2017 • UW Professional & Continuing Education Qualified/ Ethical Hacking Certification (Q/EH). 2014 • Cisco Certified Network Professional (CCNP) 2014 • 642-813 SWITCH Cisco Certified Network Associate (CCNA). 2014 • Cisco Id: CSCO12••••73 Microsoft Windows ServerTM 2003,(MCSA). 2009 • MCP ID: 69•••47

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.

Habtamu Asfaw — Site Reliability Engineering (Sre) at Dotdash Meredith in Seattle, WA, US | Unifers