Kustomize
Customize Kubernetes manifests without templating using Kustomize.
Skills for your AI
Customize Kubernetes manifests without templating using Kustomize.
System administration for Linux servers. Manage packages, services, and system configuration. Use when administering Linux systems.
Implement multi-layer LLM caching with exact match, semantic similarity, and provider-side prompt caching.
Reduce LLM API and infrastructure costs through model selection, prompt caching, batching, caching, quantization, and self-hosting strategies.
Set up infrastructure for fine-tuning LLMs with QLoRA, LoRA, and full fine-tuning using Hugging Face TRL, Axolotl, and distributed training with DeepSpeed or FSDP.
Deploy an API gateway for LLM traffic with load balancing, rate limiting, key management, semantic caching, fallback routing, and cost tracking.
Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling.
Build production LLMOps platforms with CI/CD, model promotion workflows, evaluation gates, rollback, and governance across cloud and self-hosted inference.
Configure load balancers and traffic distribution. Implement health checks and SSL termination. Use when distributing traffic across servers.
Configure Grafana Loki for log aggregation and analysis.
Configure a Mac mini as a reliable local LLM server with remote access, observability, and power-safe operation. Use when building an always-on private AI inference server on Apple Silicon.
Manage and secure company devices with MDM solutions
Establish model registry standards, governance controls, metadata schemas, approvals, and lifecycle policies for enterprise AI deployments.
Deploy ML models on Kubernetes with KServe (formerly KFServing) and NVIDIA Triton Inference Server.
Administer MongoDB databases. Configure replica sets, sharding, and backups. Use when managing MongoDB deployments.
Design secure, multi-tenant LLM hosting platforms with tenant isolation, quotas, billing attribution, noisy-neighbor protection, and per-tenant policy controls.
Administer MySQL/MariaDB databases. Configure replication and optimize performance. Use when managing MySQL deployments.
Configure New Relic observability platform for infrastructure and application monitoring.
Configure NFS servers and clients. Implement network file sharing for Linux systems. Use when setting up shared storage.
Configure object storage with S3, GCS, and MinIO. Implement lifecycle policies and access controls. Use when managing object storage.
Run local LLM workloads with Ollama, Open WebUI, and GPU-aware tuning for private development environments.
Set up OpenClaw locally and run it reliably on a Mac mini for private, always-on local agent workflows.
Harden OpenClaw self-hosted environments with baseline host controls, auth tightening, secret handling, network segmentation, and safe update/rollback workflows.
Manage Red Hat OpenShift clusters and deployments.