Infrastructure & reliability assessment
A focused review of your environment to uncover operational risks, performance bottlenecks, recurring failures, and opportunities to improve.
Specialized DevOps, SRE, cloud, and observability consulting for companies looking to improve reliability, automate operations, and gain deeper visibility into their environments.
Services
Hands-on support across cloud infrastructure, automation, observability, and production reliability.
A focused review of your environment to uncover operational risks, performance bottlenecks, recurring failures, and opportunities to improve.
Metrics, logs, traces, dashboards, and actionable alerts that make systems easier to understand and incidents faster to resolve.
Repeatable infrastructure, delivery pipelines, and operational workflows designed to make changes safer and more consistent.
Architecture and usage reviews that reduce unnecessary cloud spend without trading away performance or reliability.
Design, improvement, and troubleshooting of clusters and workloads for dependable day-to-day production operations.
Service-level objectives, reliability indicators, incident practices, and continuous improvement grounded in operational data.
How it works
Every engagement begins with the current environment and business goals, then moves through clear priorities, controlled delivery, and an effective handover.
Step 1
I learn how the environment works today, where the risks are, and what the business needs to achieve.
Step 2
I define the priorities, technical approach, responsibilities, risks, and expected deliverables.
Step 3
I implement the solution in controlled stages, with validation, documentation, and clear communication throughout.
Step 4
I review the outcome, share the operational knowledge, and make sure the team is ready to own the solution.
About
I am a DevOps Engineer and SRE with more than three years of hands-on experience building, operating, and improving production environments. My work spans cloud infrastructure, automation, Kubernetes, CI/CD, and observability.
I work with AWS, Terraform, Docker, Kubernetes, Helm, GitHub Actions, Jenkins, and Argo CD, alongside observability tooling such as Prometheus, Grafana, Loki, Elastic Stack, and OpenTelemetry. I also troubleshoot complex issues, investigate incidents, and work closely with development and infrastructure teams.
My background in Civil Engineering shaped a structured approach to planning, risk, cost, and problem-solving. At FSO Cloud Consulting, I bring that discipline to reliable, automated solutions tailored to each environment.
3+ years in DevOps and SRE
Hands-on production experience
AWS, Terraform, Kubernetes, CI/CD, and GitOps
Observability, troubleshooting, and reliability
Incident investigation and cross-team collaboration
A structured approach to planning, risk, and cost
Projects and technical cases
Technical case studies focused on sound architecture, reliable operations, automation, and measurable outcomes.
A React SPA served from a private S3 origin through CloudFront, with managed TLS and DNS in Route 53.
Challenge
Deliver a low-latency SPA on a custom domain without exposing the origin or breaking client-side routing.
Solution
A private S3 origin protected by OAC, fronted by CloudFront, ACM, and Route 53, with a dedicated SPA routing fallback.
Outcome
Fast global delivery, HTTPS by default, reliable client-side routes, and no direct public access to the bucket.
Centralized metrics, logs, and operational signals designed to speed up incident investigation.
Coming soonRepeatable, standardized cloud environments delivered through infrastructure as code and automation.
Coming soonArticles and knowledge
Field notes, guides, and practical perspectives on DevOps, SRE, cloud, observability, and modern production systems.
A practical view of metrics, logs, traces and visibility across distributed systems.
Turning technical indicators into reliability objectives aligned with business needs.
Optimizing cloud environments without compromising performance, stability or growth.
Technologies and capabilities
Technology choices are driven by context, architecture, business goals, and the team's operational maturity.
Approach
Tools are a means, not the goal. I use them to improve reliability, operational efficiency, and the team's ability to evolve the platform.
Platforms used to build and operate secure, scalable and highly available environments.
Technologies for application runtime, orchestration and consistent delivery.
Tools for metrics, logs, traces, availability, alerting and incident investigation.
Provisioning, configuration and automation with consistency and traceability.
Automated integration, validation, software delivery and change control.
More than tools
Architecture, automation, and observability working together to support reliable production systems.
Contact
Tell me about your environment and the challenges you are facing. I can help identify practical ways to improve automation, observability, reliability, and cloud efficiency.
Chat on WhatsApp