I build and scale self-hosted AI infrastructure. My current focus is on MLOps and edge computing—specifically optimizing open-weight models (via vLLM and llama.cpp) using AMD ROCm on rootless Linux environments. I am passionate about bridging the gap between raw model inference and full-stack production by building secure, agent-driven workflows for web applications and automated CI/CD pipelines.