$ connecting to prod-cluster...
Open to Freelance · Contract · Remote · Full-Time

Muhammad Software Engineer

I build the infrastructure that never sleeps. Architecting cloud-native platforms across AWS, GCP & Azure - from Kubernetes clusters humming at 99.99% uptime to FinOps strategies that cut costs by 40%.

muhammad@prod-k8s ~ zsh
EXPLORE ↓
Performance Metrics

Engineering Impact

- Platform Uptime SLA SRE Gold Standard
- Cloud Cost Reduction FinOps Certified Results
- Years Experience Enterprise Grade
- Engineers Led Team Leadership
✦ PERFECT PCI-DSS Audit Score 100% Compliance · Zero Gaps
- Microservices Deployed Production Grade
- K8s Clusters Managed EKS · GKE · AKS
- Clients Served Global · Cross-Industry
- Incident Response On-Call · Always Ready
Toolchain

The Arsenal

☁️ Cloud Platforms
AWS (EC2, S3, RDS, VPC, Lambda) GCP (GKE, Compute Engine) Azure (AKS, VMs)
⚙️ Compute & Orchestration
Docker Kubernetes (EKS/GKE/AKS) EC2 Helm ArgoCD Istio
🏗️ Infrastructure as Code
Terraform Terragrunt CloudFormation Ansible
🚀 CI/CD & Automation
Jenkins GitHub Actions GitLab CI
📊 Observability & SRE
Prometheus Grafana Datadog ELK Stack SLO / SLI / Error Budgets
🔐 Security & Compliance
Vault Trivy OPA PCI-DSS
💻 Programming & Scripting
Python Node.js Bash Go
🧠 MLOps & Data Systems
PyTorch TensorFlow MLflow Airflow Kafka
Work History

Where I've Built Things

Nixense Vixion
Senior Software Engineer (DevOps / SRE / Platform / MLOps)
2023 - Present
  • Led cloud, SRE and platform engineering across multi-cloud (AWS, GCP, Azure) environments supporting production workloads.
  • Designed and operated Kubernetes-based ML infrastructure for scalable training, inference and GPU workloads (TensorFlow, PyTorch).
  • Built end-to-end MLOps platform with CI/CD, GitOps (ArgoCD, Helm), model registry and automated rollback pipelines.
  • Implemented observability stack (Prometheus, Grafana, Datadog, ELK), improving reliability and reducing incident detection time.
  • Established SRE practices (SLIs, SLOs, error budgets) improving production stability through structured incident response and RCA.
  • Architected multi-region infrastructure using Terraform and Ansible for high availability and fault tolerance.
  • Delivered FinOps optimization reducing cloud spend by ~30–40% while maintaining SLA targets.
  • Implemented PCI-DSS compliant infrastructure with successful external audit outcomes.
AWSGCPAzure KubernetesArgoCDTerraform MLOpsFinOpsVault PrometheusGrafanaPCI-DSS
InvoZone
DevOps Engineer
2021 - 2023
  • Built and scaled CI/CD pipelines using Jenkins, GitHub Actions and GitLab CI for large-scale microservices systems.
  • Migrated legacy systems to Kubernetes (EKS) using blue/green and canary deployments with zero downtime.
  • Implemented full observability stack (Prometheus, Grafana, ELK, Datadog), improving system visibility and operational efficiency.
  • Defined and enforced SLOs, SLIs and error budgets across production systems to improve reliability maturity.
  • Automated infrastructure provisioning using Terraform, CloudFormation and Ansible for reproducible environments.
  • Strengthened platform security using Vault, Trivy, OPA and CIS-hardened Kubernetes workloads.
  • Supported microservices and event-driven architectures using Node.js and Golang for scalable distributed systems.
  • Led production incident response and change management, improving MTTR and system resilience.
EKSJenkinsGitHub Actions DatadogOPAPCI-DSSISO 27001
MailMunch
DevOps Automation Engineer
2020 - 2021
  • Designed and deployed AWS cloud infrastructure (EC2, ECS, RDS, CloudFront) for high-traffic production workloads.
  • Introduced Infrastructure as Code (Terraform), reducing provisioning time by 70% and eliminating configuration drift.
  • Implemented centralized observability using ELK stack and APM tools for system monitoring and performance tracking.
  • Automated CI/CD deployment workflows, improving reliability, release velocity and operational efficiency across environments.
AWSTerraformECS ELK StackCloudFront
Caramel Tech Studios
DevOps Engineer
2016 - 2020
  • Managed hybrid cloud infrastructure (on-prem + cloud) for enterprise fintech and healthcare systems with high availability requirements.
  • Containerized applications using Docker and Kubernetes, improving scalability, deployment speed and workload portability.
  • Built infrastructure automation using Ansible and Bash for patching, backups and infrastructure lifecycle management.
  • Improved system reliability through monitoring, incident response and disaster recovery (DR) planning.
DockerLinux AnsibleBashAWS
Caramel Tech Studios
Software Quality Assurance Engineer
2016 - 2018
  • Built automated test frameworks for web, mobile and API systems supporting enterprise-grade applications.
  • Performed load and stress testing on distributed messaging systems (EMQX, RabbitMQ) using JMeter and mzBench.
  • Automated regression and functional testing using Cypress, Postman and JMeter, improving test coverage and execution efficiency.
  • Validated APIs and backend systems for performance, reliability and production readiness in distributed environments.
  • Managed QA lifecycle using Jira-based workflows ensuring defect tracking, test traceability and release quality control.
JMeterCypressPostman API Testing
Black Box Labz
Software Quality Assurance Engineer
2015 - 2016
  • Performed functional and non-functional testing for web and mobile applications including e-commerce platforms.
  • Created structured test plans and test cases covering unit, integration and system-level validation.
  • Ensured requirement traceability and quality documentation for release validation and audit readiness.
  • Reported and tracked defects using Jira and Asana, improving coordination between QA and development teams.
QA TestingTest CasesJira AgileManual Testing
Core Expertise

What I Excel At

Multi-Cloud Architecture
Designing resilient, cost-optimised infrastructure across AWS, GCP and Azure with cross-cloud failover, federation and unified observability.
AWSExpert
GCPAdvanced
AzureAdvanced
Kubernetes & Platform Engineering
Production-grade K8s platforms (EKS, GKE, AKS) with GitOps, service mesh, RBAC, multi-tenancy and custom operators.
KubernetesExpert
ArgoCD / GitOpsExpert
HelmExpert
Security & Compliance
Zero-trust architectures, secrets management, Policy-as-Code and full regulatory compliance with PCI-DSS and ISO 27001 - zero audit findings.
PCI-DSSExpert
Vault / Secrets MgmtExpert
ISO 27001Advanced
Observability & SRE
Full-stack observability with Prometheus, Grafana, Datadog. SLO/SLI/error-budget frameworks, incident command and chaos engineering.
Prometheus / GrafanaExpert
DatadogAdvanced
CI/CD & Automation
End-to-end pipeline design with GitHub Actions, Jenkins and ArgoCD. Progressive delivery, canary releases, feature flags and release automation at scale.
GitHub ActionsExpert
JenkinsExpert
FinOps & MLOps
Cloud cost governance delivering 40% savings. MLOps pipelines covering model training automation, registry management, serving infrastructure and drift detection.
FinOpsAdvanced
MLOpsAdvanced
Social Proof

What They Say

"

Muhammad transformed our infrastructure from a liability into a competitive advantage. His ability to architect systems that are simultaneously resilient, cost-efficient and audit-ready is genuinely rare. The 99.99% uptime we achieved wasn't luck - it was his engineering discipline.

JR
James R.
VP Engineering, Fintech SaaS
"

We'd been struggling with cloud costs spiralling out of control. Muhammad came in, did a thorough FinOps analysis and within 6 months had cut our bill by 40% without touching a single feature. He also built a culture of cost-consciousness in the team that outlasted his engagement.

SA
Sarah A.
CTO, Healthcare SaaS Platform
"

The PCI-DSS certification was a make-or-break moment for our fintech product. Muhammad led the entire compliance engineering effort - from threat modelling to controls implementation. Zero findings on the final audit. Exceptional work under real pressure.

MK
Michael K.
Head of Engineering, DaRemit
Notable Clients & Partners

Trusted by Industry Leaders

Academic Background

Education & Certifications

Machine Learning
Stanford University

Supervised/unsupervised learning, neural networks and model evaluation. Applied to MLOps pipeline design, model lifecycle management and production ML infrastructure on cloud-native systems.

Neural Networks Model Training MLOps Foundations
Data Analysis with R Programming
Johns Hopkins University

Statistical analysis, R programming and reproducible research. Used for SLI/SLO observability, FinOps cost modeling and data-driven infrastructure decision-making.

R Programming Statistical Analysis Data Visualisation
Economic & Business Studies
National University of Singapore

Business economics, strategy and financial analysis. Applied to FinOps optimization, cloud cost governance and translating infrastructure decisions into business value.

Business Strategy Economics FinOps Thinking
Get In Touch

Let's Build Together

Open to every kind of challenge.

Freelance project? Contract engagement? Part-time consulting? Full-time role? Remote, hybrid, anywhere in the world - if you're building something ambitious and need a cloud engineer who can deliver, let's talk.

Freelance Contract Part-Time Full-Time Remote Short / Long Projects
Send a Message
Message received! I'll get back to you hours.
Back to top