Senior Site Reliability Engineer & Platform Developer

Joe Hernandez

7+ years across regulated and high-growth environments (medical devices, enterprise SaaS, financial services), owning cloud infrastructure across 120+ production services, incident response, and the automation that eliminates manual toil. An SRE who writes Go and Python tooling, standardizes Azure infrastructure with Terraform, and builds Next.js applications.

7+
Years SRE & Cloud
120+
Production Services
99.99%
Multi-Region Uptime
$12K+/mo
Cloud Spend Saved

Featured Projects & Open Source

All projects & case studies →

Career Journey & Impact

Senior Site Reliability & Infrastructure Engineer

April 2025 – Present
Enterprise SaaS & Regulated Financial Technology
  • Architected multi-cluster AKS environments with modular Terraform, maintaining 99.99% availability, and built Go and Python self-healing tooling that eliminated 12+ hours per week of toil.
  • Established SLIs/SLOs and error budget policies to gate deployments during high-burn periods, and engineered CI/CD and DevSecOps pipelines with policy-as-code and vulnerability scanning.

DevOps & Platform Infrastructure Engineer

July 2022 – April 2025
Enterprise SaaS
  • Scaled multi-region AKS infrastructure across 600+ Linux nodes with Terraform and Ansible, shipping zero-downtime blue/green and canary rollouts via Helm.
  • Centralized secrets management via HashiCorp Vault and Azure Key Vault, enforcing least-privilege RBAC, and led postmortems as part of a 24/7 on-call rotation.

Software Engineer II, Backend & DevOps

February 2021 – July 2022
Regulated Medical Devices
  • Automated hybrid cloud infrastructure with Terraform and ARM templates, and built auditable CI/CD pipelines with rollbacks in Azure DevOps and Octopus Deploy.
  • Executed Linux OS hardening, network segmentation, and vulnerability remediation across servers.

Solutions & Systems Engineer

March 2019 – February 2021
Software Engineering & Systems Administration
  • Developed Python, Bash, and PowerShell workflows to automate server provisioning and configuration management.
  • Administered Linux/Windows server fleets and performed query and performance tuning on SQL databases.

Certifications & Credentials

☁️
Microsoft Certified
Azure Administrator Associate (AZ-104)

Validated expertise across Azure identity, compute, virtual networking, storage governance, and security.

🏗️
HashiCorp Contributor
terraform-provider-azurerm

5 merged pull requests adding resource schemas, Azure REST API mappings, and Go acceptance tests.

Latest Technical Posts

Practical insights on SRE practices, incident management, cloud architecture, and modern development.

View all posts →

Technical Focus Areas

  • 🤖 Automation & Tooling: Eliminating manual toil with custom Go CLI tools, Python scripts, and CI/CD pipelines.
  • ☁️ Azure & Cloud Governance: Enterprise-grade Azure architectures, Private Endpoints, Key Vault CSI secret delivery, and VNet peering.
  • 🏗️ Terraform (IaC): Standardized, reusable Terraform modules, policy-as-code guardrails, and provider contributions.
  • ☸️ Kubernetes (AKS): Multi-tenant cluster platforms, ingress controllers, RBAC access models, and resource quotas.
  • 📊 Observability & SLIs/SLOs: Building actionable Prometheus metrics, Grafana dashboards, and meaningful error budget alerting.
  • 💻 Software Development: Go backend services and CLI tools, plus Next.js and React for internal tooling and this site.
Career Inquiries

Hiring for SRE or Platform Roles?

Employer names are left off this public site. Email me for the full resume with employer details and references, or connect on LinkedIn.