Trinidad Marroquin
Practice Lead - Site Reliability Engineering & DevOps
SRE / DevOps / Platform Engineer with 25 years of information technology experience across infrastructure operations, automation, CI/CD, cloud platforms, databases, systems administration, and reliability practices, including three years leading a DevOps engineering team.
Professional Summary
Quality-driven and practical information technologist focused on making infrastructure repeatable, observable, secure, and maintainable. Experienced across Kubernetes platforms, Terraform-managed infrastructure, vSphere, image pipelines, CI/CD, Vault, Ansible, NetBox, Linux systems, cloud platforms, database operations, and production support.
I care about clear operating models: versioned changes, useful alerts, readable pipeline output, documented recovery paths, and systems that can be debugged under pressure.
Target Roles
- Site Reliability Engineer.
- DevOps Engineer.
- Platform Engineer.
- Infrastructure Automation Engineer.
- Kubernetes / Cloud / vSphere operations roles.
Core Skills
- Kubernetes platform operations: Rancher, RKE2, Longhorn, Calico, ingress, cert-manager, Velero, cluster autoscaler, node lifecycle, upgrade sequencing, and operational evidence collection.
- Infrastructure as code: Terraform provisioning, module design, state organization, provider behavior, plan review, vSphere automation, and environment promotion.
- Image and node pipelines: Packer image builds, template lifecycle, validation gates, OS template hygiene, and repeatable machine provisioning.
- Secrets and security operations: Vault deployment, policy design, secrets engines, Kubernetes auth, PKI, transit encryption, token lifecycle, lease review, rotation patterns, audit logging, and operational guardrails.
- Automation and inventory: Ansible remediation workflows, per-cluster inventory patterns, Vault-backed variables, NetBox ownership data, Python scripting, Bash scripting, and repeatable operational audits.
- CI/CD and delivery: Concourse, Jenkins, Hudson, Nexus, Maven, validation stages, promotion workflows, deployment safety, and pipeline troubleshooting.
- Observability and incident response: Prometheus, Grafana, SLOs, burn-rate alerting, dashboard design, alert routing, incident review, escalation, and evidence bundles.
- Virtualization and platform operations: vSphere / vCenter, VM lifecycle, resource and template hygiene, CSI/CNS troubleshooting, Kubernetes node support, storage attach/mount triage, and automation integration.
- Linux and middleware operations: Linux systems administration, patching, security maintenance, RPM packaging, Puppet, Tomcat, GlassFish, Apache HTTPD, and Oracle database administration.
- Cloud platforms: AWS, Azure, cloud infrastructure patterns, identity, networking, monitoring, and platform operations.
Core Strengths
- Turning messy infrastructure behavior into clear runbooks, checklists, and operational evidence.
- Building automation that stays inspectable, reversible, and safe for production operators.
- Connecting platform engineering practices across Kubernetes, vSphere, Terraform, Vault, CI/CD, observability, and Linux systems.
- Writing deep technical documentation that helps engineers troubleshoot under pressure.
Certifications
- The University of Texas at San Antonio: graduate-level Data Science Certificate coursework, 9 credit hours, 4.0 GPA. Courses included Programming for Data Science, Statistical Methods in Research, Introduction to Data Science, and Data Organization & Visualization. Transcript issued November 8, 2024.
- Caltech Center for Technology and Management Education / Simplilearn: Post Graduate Program in DevOps, completed June 2, 2022. Certificate ID:
52288585. - The Linux Foundation: KCNA: Kubernetes and Cloud Native Associate. Issued February 2025, expires February 2027. Certificate ID:
LF-b5hanhynvo. - The Linux Foundation: LFS250: Kubernetes and Cloud Native Essentials.
- The Linux Foundation: LFS158: Introduction to Kubernetes.
- The Linux Foundation: LFS151: Introduction to Cloud Infrastructure Technologies.
- The Linux Foundation: LFS101: Introduction to Linux.
- Credly profile: credly.com/users/trinidad-marroquin.
- Previously held Credly-listed certifications: HashiCorp Certified: Terraform Associate (003), AWS Certified Cloud Practitioner.
- Additional technical training and certifications completed over time as part of ongoing retooling across Linux, cloud, DevOps, SRE, and platform engineering. Details available upon request.
Languages
- English
- Spanish
Selected Proof Of Work
- Kubernetes Platform Operations - upgrade sequencing, platform conventions, access patterns, and operational readiness.
- Observability And Incident Response - SLOs, alerting, dashboard design, incident review, and escalation practices.
- Secrets Management With Vault - Vault deployment, policy, auth methods, leases, audit, PKI, transit, and rotation practices.
- Terraform Infrastructure Modules - module boundaries, plan review, and environment ownership.
- vCenter Platform Administration - VM lifecycle, platform hygiene, automation integration, and Kubernetes storage touchpoints.
- Packer Image Pipelines - image versioning, validation gates, base image hardening, and image factory workflows.
- CI/CD Pipeline Design - validation, promotion, credential handling, deployment patterns, Concourse, and Terraform workflows.
Experience
trinidadmarroquin.com - Site Reliability Engineering & DevOps
January 2025 - Present | Texas, United States
- Built a public engineering log focused on SRE, DevOps, Kubernetes, vCenter, Terraform, Vault, Packer, NetBox, observability, CI/CD, security guardrails, and production operations.
- Published field notes, project writeups, and operational references covering failure modes, runbooks, evidence collection, alerting, dashboards, secrets management, Kubernetes storage, vSphere automation, and infrastructure automation.
- Developed companion patterns for safe operational tooling, read-only diagnostics, sanitized public technical documentation, and operator-friendly evidence collection.
Donyati - Practice Lead, Site Reliability Engineering & DevOps
April 2020 - January 2025 | Texas, United States
- Led SRE and DevOps practice direction with hands-on leadership for infrastructure optimization, automation, documentation, and operational process improvement.
- Led a team of DevOps engineers for three years, providing technical direction, mentoring, delivery coordination, and hands-on support across infrastructure automation, CI/CD, cloud, and reliability work.
- Developed training programs, documentation practices, and tool evaluation frameworks for SRE and DevOps processes.
- Standardized best practices across teams using structured documentation for tooling, incident response, infrastructure as code, and operational learning.
- Created reusable templates for tool evaluation and post-incident reviews to improve team efficiency and knowledge sharing.
- Facilitated collaborative learning and team-wide contribution to documentation and best practices for performance optimization and infrastructure scalability.
- Documented deep operational insight into AWS, Azure, Terraform, Kubernetes, Prometheus, Grafana, CI/CD, automation, incident response, and related platform technologies.
Independent Consultant - Independent Consultant
March 2020 - April 2020 | San Antonio, Texas Metropolitan Area
- Helped teams address hands-on technical problems while improving reliability, security, repeatability, and operational resilience.
- Focused on practical guidance for robust systems, adaptable processes, and dependable infrastructure practices.
H-E-B - Lead Technical Specialist
September 2003 - March 2020
- Worked in an infrastructure provisioning role building and configuring virtual machines from templates across development, certification, and production environments.
- Provided input to internal cloud teams for automation and template design improvements.
- Managed middleware platforms including Tomcat, GlassFish, and Apache HTTPD servers.
- Worked with infrastructure teams to build environments for internally developed and commercial applications.
- Supported monitoring, patching, and security maintenance for managed systems.
- Developed RPM packages to maintain consistent deployments across a large number of systems.
- Used Puppet to maintain RPM installations for distributed systems.
- Used Jenkins, Hudson, Nexus, and Maven to move application code from development to certification to production environments.
- Built Python and Bash scripts to automate routine tasks, support administration duties, package repeatable operations, and reduce manual toil.
Valero Energy Corporation - Oracle Database Administrator
June 2001 - September 2003
- Supported Oracle database administration responsibilities.
Education
The University of Texas at San Antonio
Bachelor of Applied Science (BASc), Mathematics and Computer Science, 1994.
Graduate-level Data Science Certificate coursework, 2023-2024. Completed 9 credit hours with 4.0 GPA.