Looking for: AI Infrastructure, Platform, DevOps, SRE, Cloud Infrastructure
3+ years in production infrastructure.
At GlueX, I helped reduce AWS spend from ~$25k to $8.5k/month and brought a critical request path from ~2s to under 250ms.
I’m now the first infrastructure person at an AI startup, owning AWS/Nebius infrastructure, deployments, observability, recovery and NVIDIA NIM/VSS workloads.
Platform / DevOps / SRE engineer with 3+ years of experience owning production infrastructure in startup teams.
I am useful when a team needs someone to take messy infrastructure, deployment, cost, observability, or reliability problems and make them boring.
Recent proof: At RestoreAI, I own cloud infrastructure for an AI product across dev, staging, prod, and preview. Recently reduced dev/staging AWS spend from about $700/day to $200/day by rightsizing GPU workloads and removing stale Kubernetes/ECS resources.
At GlueX Protocol, I worked on production blockchain/RPC and data infrastructure across multiple EVM chains. Reduced AWS cost by 45%, improved API latency from around 2s to under 250ms, and built observability with Prometheus/Grafana/Loki.
Best fit: platform engineering, DevOps, SRE, cloud infrastructure, AI infrastructure, data infrastructure, production reliability, technical operations, or infrastructure-heavy support engineering.
Platform / DevOps / SRE engineer with 3+ years of experience owning production infrastructure in startup teams.
I am useful when a team needs someone to take messy infrastructure, deployment, cost, observability, or reliability problems and make them boring.
Recent proof: At RestoreAI, I own cloud infrastructure for an AI product across dev, staging, prod, and preview. Recently reduced dev/staging AWS spend from about $700/day to $200/day by rightsizing GPU workloads and removing stale Kubernetes/ECS resources.
At GlueX Protocol, I worked on production blockchain/RPC and data infrastructure across multiple EVM chains. Reduced AWS cost by 45%, improved API latency from around 2s to under 250ms, and built observability with Prometheus/Grafana/Loki.
Best fit: platform engineering, DevOps, SRE, cloud infrastructure, AI infrastructure, data infrastructure, production reliability, technical operations, or infrastructure-heavy support engineering.
Platform / DevOps / SRE engineer with about 3 years of experience owning production infrastructure in startup teams.
I am useful when a team needs someone to take messy infrastructure, deployment, cost, observability, or reliability problems and make them boring.
Recent proof:
At RestoreAI, I own cloud infrastructure for an AI product across dev, staging, prod, and preview. Recently reduced dev/staging AWS spend from about $700/day to $200/day by rightsizing GPU workloads and removing stale Kubernetes/ECS resources.
At GlueX Protocol, I worked on production blockchain/RPC and data infrastructure across multiple EVM chains. Reduced AWS cost by 45%, improved API latency from around 2s to under 250ms, and built observability with Prometheus/Grafana/Loki.
Best fit: platform engineering, DevOps, SRE, cloud infrastructure, AI infrastructure, data infrastructure, production reliability, technical operations, or infrastructure-heavy support engineering.
Platform / DevOps / SRE engineer with about 3 years of experience building and operating production systems in startup environments.
I work across cloud infrastructure, reliability, automation, observability, data infrastructure, and cost optimization. I have handled production infrastructure, monitoring, incident response, deployment automation, RPC/node infrastructure, benchmarking, and infrastructure migrations.
Most of my recent work has been at the intersection of platform engineering, blockchain infrastructure, data systems, and AI infrastructure. I am not limited to blockchain roles. I am also interested in platform engineering, DevOps, SRE, cloud infrastructure, data infrastructure, AI infrastructure, and technical operations roles.
Best fit: early-stage or fast-moving teams that need someone who can own infrastructure, improve reliability, reduce cost, debug production issues, and build practical internal systems without needing too much hand-holding.
I am a Cloud Platform/DevOps Architect and Member of Technical Staff focused on platform engineering, reliability, and data infrastructure, with about 3 years of experience building and operating production systems in fast-moving startup environments. I work primarily with cloud infrastructure and automation; often using tools like Terraform and GitHub Actions, to make services easier to deploy, monitor, and keep stable, including workloads that support blockchain and AI.
DevOps and data infrastructure engineer with 3+ years of experience running production blockchain and data systems. My work has included node and RPC operations, observability, CI/CD, data pipelines, performance tuning, and infrastructure cost reduction across Web3 and AI environments.
Remote: Yes
Willing to relocate: Yes
Technologies: AWS, Nebius, Linux, Docker, Terraform/Terragrunt, GitHub Actions, PostgreSQL, Redis, Prometheus/Grafana/Loki, NVIDIA NIM/VSS, Python, Bash, SQL
Résumé/CV: https://krissemmy.com/resume
Email: [email protected]
Looking for: AI Infrastructure, Platform, DevOps, SRE, Cloud Infrastructure
3+ years in production infrastructure.
At GlueX, I helped reduce AWS spend from ~$25k to $8.5k/month and brought a critical request path from ~2s to under 250ms.
I’m now the first infrastructure person at an AI startup, owning AWS/Nebius infrastructure, deployments, observability, recovery and NVIDIA NIM/VSS workloads.
Website: https://krissemmy.com
reply