Skip to content
View vanlailaptrinh's full-sized avatar

Block or report vanlailaptrinh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
vanlailaptrinh/README.md

Hi 👋, I'm Van Tran (Trần Quốc Văn)

DevOps Engineer | Infrastructure & Automation


About Me

A proactive DevOps Engineer with hands-on experience in operating on-premise, virtualized, and bare-metal infrastructure. Specialized in Kubernetes orchestration, High Availability (HA/DR), Storage Systems (Ceph), and Infrastructure as Code (IaC).

  • Production Experience: Operated multi-node K8s clusters, Ceph storage clusters (upgraded Reef → Squid), and PostgreSQL HA with Patroni/Keepalived at XPERC LTD.
  • Applied AI & Systems: Built an end-to-end AI pipeline (EfficientNet-B0 + ONNX + Spring Boot) for mango leaf disease classification published on Google Play.
  • Core Focus: Bare-metal Automation, High Availability Systems, CI/CD Optimization, and Disaster Recovery.

Languages, Tools & Infrastructure

Virtualization, Storage & Orchestration

Proxmox Ceph Kubernetes Docker Helm Linux

Automation, IaC & CI/CD

Terraform Ansible Bash GitHub Actions GitLab CI Jenkins

High Availability, Networking & DBs

HAProxy Prometheus Grafana PostgreSQL SQL Server MongoDB Cloudflare AWS

Programming & Backend/AI

Java Spring Boot Python C#


Highlighted Projects

  • XCORP Production Infrastructure: Operated Proxmox VE cluster, Ceph storage, multi-node K8s, MetalLB, and PostgreSQL HA cluster (Patroni + HAProxy + Keepalived)[cite: 3].
  • Enterprise HA/DR Solution: Designed and tested SQL Server AlwaysOn AG & Disaster Recovery setups for multi-site Windows Server environments[cite: 3].
  • HPC Slurm Cluster: Configured Slurm job scheduler with Ansys Fluent for engineering batch execution using OpenMPI and MUNGE authentication[cite: 3].
  • LeafGuard AI (Research Project Lead): Built deep learning pipeline (EfficientNet-B0, 97.38% test accuracy) integrated with Spring Boot & React Native app published on Google Play.

Pinned Loading

  1. infrastructure-architectures infrastructure-architectures Public

    Enterprise Infrastructure Architectures: High Availability, Distributed Storage, and Advanced Networking Case Studies

    Shell 1