Skip to content
View Codeprojectingfuture's full-sized avatar

Block or report Codeprojectingfuture

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Typing SVG

πŸ‘‹ Hi, I'm Victor U

πŸ’» Site Reliability Engineer | DevOps Engineer
πŸ“ Based in Houston, Texas, United States
πŸ“§ Harrimerri10@gmail.com | πŸ“ž 713-480-9493
🌎 Open to Remote Roles


πŸš€ About Me

I’m a Site Reliability Engineer and DevOps Engineer with 8+ years of cloud and production engineering experience, specializing in reliable, scalable, and observable infrastructure across AWS, GCP, and Azure.

My experience spans Kubernetes, Terraform, Linux, CI/CD, observability, infrastructure automation, incident response, and production reliability. At T-Mobile, I work across production readiness, SLI/SLO and error-budget ownership, incident response, release support, capacity planning, recovery planning, and reliability improvements.

I’m comfortable troubleshooting complex production issues across Kubernetes, Linux, networking, cloud infrastructure, TypeScript services, and Ruby on Rails applications, tracing failures across multiple layers until the underlying issue is understood.

Earlier at Salesforce, I built reusable infrastructure and delivery automation that reduced standard environment provisioning from several days to under one hour, shortened release time by approximately 60%, and contributed to performance improvements of more than 20% in page-load and service response times.

I’m passionate about building resilient production platforms, automating operational work, improving observability, and turning recurring incidents into durable engineering improvements.


🧠 Technical Skills

☁️ Cloud Platforms

AWS Microsoft Azure Google Cloud Oracle Cloud DigitalOcean

AWS: EC2, VPC, IAM, S3, RDS, ECS, EKS, ALB, Auto Scaling, CloudWatch


☸️ Kubernetes & Containers

Kubernetes Docker Helm Amazon EKS

  • Kubernetes cluster operations
  • Amazon EKS
  • Docker
  • Helm
  • Readiness and liveness probes
  • Rolling deployments
  • Scaling and resource management
  • Pod troubleshooting
  • Image-pull and configuration troubleshooting
  • Kubernetes events and production diagnostics

πŸ—οΈ Infrastructure as Code & Automation

Terraform Ansible Python Bash

  • Terraform
  • Reusable Terraform modules
  • Infrastructure as Code
  • Ansible
  • Python automation
  • Bash utilities
  • Environment standardization
  • Automated health checks
  • Log collection and validation
  • Repeatable infrastructure changes

πŸ”„ CI/CD & Release Engineering

GitHub Actions Jenkins Git GitHub

  • GitHub Actions
  • Jenkins
  • CI/CD pipelines
  • Build and deployment automation
  • Release validation
  • Rollback planning
  • Production release support
  • Deployment troubleshooting

πŸ“Š Observability & Monitoring

Prometheus Grafana Datadog Splunk CloudWatch


πŸ›‘οΈ SRE & Reliability


🐧 Linux & Networking

Linux Ubuntu Red Hat CentOS


πŸ’» Programming & Application Technologies

TypeScript Ruby Ruby on Rails Python Bash


πŸ’Ό Professional Experience

πŸ“‘ Site Reliability Engineer

T-Mobile
February 2024 – Present | Remote, United States

Production Readiness & Service Reliability

  • Partner with development and platform teams before production changes to review architecture, service dependencies, capacity, observability, rollback plans, and operational risk.
  • Own SLI, SLO, and error-budget reviews, using availability, latency, error rates, resource pressure, and incident trends to guide release decisions.
  • Review production dashboards and signals around releases and maintenance windows, including health probes, saturation, traffic behavior, dependency health, and recovery readiness.
  • Stay engaged after deployments to identify regressions and reliability issues early.

Incident Response & Deep Troubleshooting

  • Investigate live production incidents across AWS, GCP, Azure, Kubernetes, Linux, TypeScript, Ruby on Rails, DNS, TLS, and network dependencies.
  • Trace metrics, logs, Kubernetes events, application behavior, and host signals to identify failures across infrastructure and application layers.
  • Troubleshoot Kubernetes issues including unhealthy readiness/liveness probes, failed pods, image-pull errors, configuration problems, downstream dependencies, and CPU or memory limits.
  • Participate in root-cause and post-incident reviews, separating immediate recovery actions from longer-term engineering improvements.
  • Convert recurring failure patterns into improvements to alerts, runbooks, configuration, and deployment practices.

Automation, Resilience & Technical Leadership

  • Build Python and Bash utilities for health checks, log collection, maintenance, and routine validation.
  • Use Ansible for repeatable server changes that are safer and easier to review than one-off manual updates.
  • Review backup results, failover procedures, capacity trends, access changes, patching plans, and maintenance risks.
  • Mentor engineers through production troubleshooting, observability, Kubernetes operations, and reliability reviews.
  • Document operational reasoning and troubleshooting approaches to improve knowledge sharing across engineering teams.
  • Work across AWS, GCP, and Azure production environments using common operational practices for access, monitoring, deployment safety, and recovery.

☁️ Cloud / DevOps Engineer

Salesforce
March 2018 – February 2024 | Remote, United States

Cloud Infrastructure & Infrastructure as Code

  • Provisioned AWS environments using Terraform across VPC, EC2, IAM, S3, RDS, load balancing, Auto Scaling, and monitoring.
  • Created reusable infrastructure modules that reduced standard environment provisioning from several days to less than one hour.
  • Built reusable Terraform and Ansible patterns for development, staging, and production environments.
  • Supported cloud networking and access patterns including VPC design, subnets, routing, security controls, load balancing, IAM, DNS, and TLS.
  • Troubleshot connectivity and permission issues spanning application and infrastructure boundaries.

CI/CD, Containers & Release Engineering

  • Built GitHub Actions and Jenkins pipelines for build, testing, and deployment.
  • Automated repeated release processes, shortening release time by approximately 60%.
  • Containerized three web applications using Docker.
  • Supported application deployments through Amazon ECS and Kubernetes.
  • Worked with developers during releases and production troubleshooting to distinguish application defects from environment, dependency, and platform issues.

Linux, Observability & Operations

  • Administered Ubuntu, CentOS, and RHEL systems, including patching, systemd services, storage, logs, SSL/TLS, user access, and baseline hardening.
  • Supported business-critical applications through planned maintenance and production incidents.
  • Implemented CloudWatch, Prometheus, and Grafana monitoring for host and application health.
  • Used observability data to identify performance bottlenecks, contributing to improvements of more than 20% in page-load and service response times.
  • Maintained infrastructure code, deployment notes, runbooks, and environment documentation in version control.

🧩 Selected Engineering Work

☸️ Kubernetes Reliability Environment

  • Built a multi-node Kubernetes environment using EKS/kops and Helm.
  • Added readiness and liveness checks.
  • Implemented rolling deployments and scaling rules.
  • Built Prometheus and Grafana dashboards for monitoring.
  • Practiced recovery from failed pods, resource pressure, and configuration changes.

πŸ—οΈ Reusable Terraform Infrastructure

  • Created reusable Terraform modules for:
    • VPC
    • Public and private subnets
    • Application Load Balancer
    • Auto Scaling
    • IAM
    • EC2
    • RDS
  • Used variables and outputs to separate environment-specific configuration while avoiding duplicated infrastructure code.

πŸ”„ Resilience & Recovery Practice

  • Designed backup validation checks and recovery procedures.
  • Created controlled failure scenarios to expose service dependencies.
  • Validated recovery behavior across infrastructure components.
  • Turned findings into clearer operational documentation and runbooks.

πŸ… Certifications

  • Linux Foundation Certified System Administrator (LFCS)
  • Certified Kubernetes Administrator (CKA)
  • Certified Kubernetes Security Specialist (CKS)
  • AWS Certified Solutions Architect – Associate

πŸŽ“ Education

πŸŽ“ Continuing Education – Cloud Engineering

Texas State University

Hands-on coursework covering:

  • Linux Administration
  • AWS
  • Kubernetes
  • Docker
  • Terraform
  • CI/CD
  • Python
  • Bash
  • Networking
  • Security
  • Monitoring
  • Automation

πŸŽ“ Business Administration & Management Coursework

University of Debrecen

Completed two years of undergraduate coursework.
No degree awarded.


πŸ“ˆ Reliability Engineering Workflow

Production Readiness
        ↓
SLIs / SLOs / Error Budgets
        ↓
Observability & Monitoring
        ↓
Incident Response
        ↓
Root-Cause Analysis
        ↓
Automation & Reliability Engineering
        ↓
Resilience & Recovery

πŸ› οΈ Engineering Philosophy

Reliable systems are built through strong engineering practices, clear observability, thoughtful automation, and continuous learning from production failures.

I focus on making infrastructure and operations:

  • πŸ”Ή Repeatable
  • πŸ”Ή Observable
  • πŸ”Ή Automated
  • πŸ”Ή Resilient
  • πŸ”Ή Reviewable
  • πŸ”Ή Easier to operate

πŸ“Š GitHub Stats


πŸ† GitHub Trophies

Victor's Trophies


πŸ’¬ Random Dev Quote


🌟 Let's Connect

πŸ’Ό Open to remote Site Reliability Engineering and DevOps opportunities.

πŸ“ Houston, Texas | 🌎 Remote, United States

πŸ“§ Harrimerri10@gmail.com

πŸ“ž 713-480-9493

πŸ”— GitHub: Codeprojectingfuture

Pinned Loading

  1. azure-aks-platform-engineering azure-aks-platform-engineering Public

    Reusable Azure AKS platform and learning environment covering Terraform, GitOps, DevSecOps, observability, platform engineering, and AIOps workflows.

    HCL

  2. kubernetes-devsecops-platform kubernetes-devsecops-platform Public

    End-to-end Kubernetes DevSecOps platform demonstrating Terraform, Ansible, Vault, Argo CD, security controls, and cloud-native observability.

    HCL

  3. kubernetes-disaster-recovery-ha kubernetes-disaster-recovery-ha Public

    Production-style Kubernetes disaster recovery and high-availability platform using GitOps, observability, chaos testing, and automated infrastructure provisioning.

    Go Template

  4. sre-observability-lab sre-observability-lab Public

    Practical observability lab for multi-service systems featuring SLO math, burn-rate alerts, Grafana dashboards, runbooks, request correlation, and chaos scenarios.

    Go

  5. sre-reliability-chaos-lab sre-reliability-chaos-lab Public

    Self-contained SRE sandbox with Go services, SLO-based alerting, chaos engineering, error budgets, supply-chain security, and automated remediation.

    Go

  6. DevSecOps-Lab DevSecOps-Lab Public

    End-to-end DevSecOps Task Manager demonstrating automated testing, security scanning, containerization, monitoring, and cloud deployment.

    TypeScript 9