About Me
I am a systems, site reliability, and automation engineer with more than 20 years of experience building and operating scalable Linux infrastructure.
I have spent my career designing and managing private clouds, storage platforms, and automation workflows for high-traffic, high-uptime environments. My background includes OpenStack, Ceph, NetApp, Docker Swarm, Ansible, Terraform, MAAS, and observability stacks built around Grafana and related tooling.
I have led platform modernization efforts for large organizations, including containerizing legacy services, improving deployment consistency, and building GitOps-style workflows for faster and safer releases. I enjoy solving infrastructure problems that require both architectural thinking and hands-on implementation.
A major part of my work has been operating at scale, from multi-petabyte storage environments to large server fleets and high-volume production systems. I have also built provisioning, testing, and deployment pipelines that reduce manual effort and improve reliability across development, QA, staging, and production.
I value practical engineering, clear operational visibility, and systems that are easy for teams to adopt and maintain. I have worked closely with developers, architects, and operations teams to improve uptime, performance, and delivery speed.
I am open to opportunities where I can contribute deep Linux infrastructure expertise, cloud and storage architecture, and automation leadership to complex production environments.
Skills
PythonSQLKubernetesLinuxCI CDTerraformPHPAnsibleRedisGrafanaBASHPrometheusObservabilityVirtualizationDNSCloud ArchitectureIdentity And Access ManagementSite Reliability EngineeringSystem AdministrationSREGitOpsNginxOpenStackCachingApacheELKKVMLoad BalancingCephInfluxDBDocker SwarmSaltVMware ESXiNFSVagrantCPanelMAASFirewall AdministrationIcingaVictoriaMetricsCouchDBOpenLDAPTelegrafStorage EngineeringNetApp
Tech Stack & Tools
Analytics
Application Hosting
Data Stores
Development
Monitoring
Experience
Architected next-generation on-prem infrastructure for online game backend services with a focus on performance, reliability, and visibility. Containerized services and configurations to reduce config drift, then tested, promoted, and deployed them through SwarmCD. Automated ingress, load balancing, vhost routing, and SSL management with Traefik. Replaced Jenkins plus Fabric or Ansible deployments with CephFS shares and replicated the environment across stage, test, and local development.
Built modular custom build images, tools, and workflows for deploying and configuring highly specific systems. Created provisioning images and tools for vCMTS appliances, used Canonical MAAS for discovery and PXE tasks, and built in-memory images with Ansible and HashiCorp Packer. Applied OSTree for the testing and installation framework and ran testing and configuration processes in Docker containers.
Supported a CDN team initiative to migrate manually managed Akamai rules to configuration-as-code. Codified CDN configuration with the principal architect, rebuilt utilities for shaping and validating web traffic, and mentored an intern delivering a CDN purge CLI tool and AWS Lambda function using Git, Pyenv, Poetry, PyPI/Homebrew packaging, and Ansible.
Provided custom OpenStack and automation solutions for high-value clients. Built Ansible Tower and Foreman pipelines to provision systems at scale, helped guide OpenStack reliability testing, and created tools for containerized deployment, testing, management, and monitoring. Automated hardware checks, advised on PXE and firmware remediation, and collaborated on Ceph cluster management and performance tuning.
Guided the creation of new API, services, and ETL tools to migrate application data from an obsolete backend with minimal customer impact. Redesigned the mobile backend across AWS EC2/S3, OpenStack, and KVM/Vagrant, built services in Node.js, Express.js, Redis, Percona XtraDB Cluster, and Nginx, and implemented monitoring with Icinga, ELK, and Telegraf/InfluxDB/Grafana. Deployed Salt for configuration and scaling.
Architected and developed tools for an on-prem multi-site OpenStack cluster deployed on Cisco UCS and NetApp hardware. Oversaw migration of bare-metal services into a virtualized IaaS environment, operated an Ubuntu OpenStack private cloud, and shared responsibility for 22 PB of storage across 72 NetApp appliances. Built custom storage balancing tools and rebalanced FlexVols ahead of projected data ingest.
Provided large-scale architecture and automation expertise to modernize the application stack and improve uptime and performance. Supported Associated Content and Yahoo Contributor Network across 340 RHEL5 servers, architected a Git-based deployment system with tiered deploys and rollback, and worked with developers on performance and capacity planning as the platform grew internationally.
Mentored junior admins and support staff to improve ticket resolution times and customer support experience. Supported more than 250 high-availability CentOS servers, redesigned Kickstart provisioning for near zero-touch builds, and reduced hands-on repairs by deploying serial console servers and improving network boot and switching design.
Built, racked, cabled, and configured web servers for multiplayer game hosting and web hosting services. Migrated more than 500 websites and email accounts from a non-cPanel environment to cPanel/WHM systems while overseeing PHP upgrades and customer hosting operations.
Education
No education data available.
This professional hasn’t added portfolio projects yet.
This professional hasn’t listed any services yet.