AI Infrastructure & Systems Engineering

Greyson K. Evers

AI Infrastructure & Systems Engineer

Building and operating AI infrastructure and GPU systems across Linux, Ceph and distributed storage, virtualization, high-speed networking, data-center deployment, security, telemetry, and observability—while developing bare-metal automation and platform capabilities.

Salt Lake City, Utahhello@greysonevers.comLinkedInGitHub

Open to remote and strategic relocation.

Last updated:

Experience

Senior Data Center Operations Engineer

DataBank

Feb 2026 – Present

Salt Lake City, Utah

Operates AI and hyperscale data-center infrastructure across GPU compute, high-speed interconnects, power, cooling, telemetry, and incident response.

  • Installed more than 70 B200/B300-class systems while supporting hundreds of GPU systems across mixed NVIDIA environments.
  • Coordinated across operations, facilities, engineering, and vendor teams to remove deployment blockers and sustain service reliability.

Infrastructure Security Lead (GS-12)

Centers for Disease Control and Prevention

Oct 2023 – Jun 2025

Atlanta, Georgia

Led infrastructure security, systems engineering, and operational support across federal cloud, virtualization, distributed storage, monitoring, and legacy environments.

  • Administered an eight-node Ceph environment with approximately 100–200 TB of raw capacity for operational storage and retention needs.
  • Advanced secure telemetry visibility through monitoring initiatives that used risk-controlled, data-diode collection patterns.

Senior Data Center Engineer

BGIS / PayPal data-center environment

Mar 2022 – Oct 2023

Salt Lake City, Utah

Supported high-availability financial data-center infrastructure under strict uptime, change-control, physical-security, and incident-response requirements.

  • Supported large-scale storage-area network deployments and validated end-to-end storage connectivity in a high-availability environment.
  • Contributed to controlled change and recovery exercises designed to protect uptime and operational continuity.

Cybersecurity Analyst

Intermountain Healthcare

Feb 2020 – Jan 2022

Salt Lake City, Utah

Supported security operations for high-availability healthcare systems subject to regulatory, privacy, and incident-response requirements.

  • Strengthened day-to-day access governance and security response across sensitive healthcare environments.
  • Supported vulnerability and endpoint-security workflows while protecting continuity of high-availability systems.

Tech Support

Apple

Atlanta, Georgia

Delivered high-volume technical support across Apple devices, operating systems, identity, networking, and endpoint configuration.

  • Resolved cross-domain endpoint and account issues through structured troubleshooting and clear customer communication.

Expertise

AI systems and GPU platforms

GPU server bring-up · AI infrastructure operations · DGX/HGX-class environment familiarity · GPU telemetry and validation · AI platform support workflows

Linux and bare-metal infrastructure

Linux administration · systemd and host troubleshooting · bare-metal provisioning concepts · BMC/IPMI/Redfish development focus · PXE development focus

Storage and virtualization

Ceph distributed storage · VMware infrastructure · block/object/file storage concepts · capacity and recovery planning · Kubernetes storage learning path

Networking and interconnects

high-speed networking · MPO/MTP fiber familiarity · AOC/DAC familiarity · 25/100 Gb network design learning path · optical/cabling BOM review experience

Security and observability

Nessus lifecycle exposure · PKI/TLS and certificate lifecycle · identity controls and hardening · Prometheus and Grafana · IPMI and SNMP telemetry

Automation and platform operations

Bash and scripting · Python automation development focus · Infrastructure-as-code development focus · Kubernetes platform development focus · repeatable operational runbooks

Education

Western Governors University

Computer Science

Certifications

  • NVIDIA NCA-AIIOIn progress
  • NVIDIA NCP-RIPending results
  • CompTIA Security+
  • CompTIA Network+
  • Lean Six Sigma Yellow Belt
  • Schneider Electric DCCA
  • DCD Generative AI and Data Centers
  • DCD Data Center Design Awareness