Trevor Royer
Trevor Royer's contributions
Article
The tokenomics of self-hosted LLMs
Trevor Royer
Learn how to optimize self-hosted LLM cost per token. Cut GPU spending and maximize real-world throughput with autoscaling, right-sizing, and vLLM tuning.
Article
How to check if your model is supported by vLLM in Red Hat AI
Trevor Royer
Confidently deploy LLMs with Red Hat support: Learn how to determine if your model is supported by Red Hat's vLLM community.
Article
Practical strategies for vLLM performance tuning
Trevor Royer
Optimize vLLM performance with practical tuning tips. Learn how to use GuideLLM for benchmarking, adjust GPU ratios, and maximize KV cache to improve throughput.
Article
Autoscaling vLLM with OpenShift AI
Trevor Royer
Implement cost-effective LLM serving on OpenShift AI with this step-by-step guide to configuring KServe's Serverless mode for vLLM autoscaling.
Article
How to navigate LLM model names
Trevor Royer
Learning the naming conventions of large language models (LLMs) helps users select the right model for their needs.
Article
Build and deploy a ModelCar container in OpenShift AI
Trevor Royer
Learn how to build a ModelCar container image and deploy it with OpenShift AI.
Article
How to validate GitOps manifests
Trevor Royer
This article discusses how to validate GitOps manifests to improve the reliability and confidence of changes before merging.
Article
3 patterns for deploying Helm charts with Argo CD
Trevor Royer
Learn about deploying Helm charts with three popular patterns for Argo CD on Red Hat OpenShift. Plus, discover the advantages and disadvantages of each pattern.
The tokenomics of self-hosted LLMs
Learn how to optimize self-hosted LLM cost per token. Cut GPU spending and maximize real-world throughput with autoscaling, right-sizing, and vLLM tuning.
How to check if your model is supported by vLLM in Red Hat AI
Confidently deploy LLMs with Red Hat support: Learn how to determine if your model is supported by Red Hat's vLLM community.
Practical strategies for vLLM performance tuning
Optimize vLLM performance with practical tuning tips. Learn how to use GuideLLM for benchmarking, adjust GPU ratios, and maximize KV cache to improve throughput.
Autoscaling vLLM with OpenShift AI
Implement cost-effective LLM serving on OpenShift AI with this step-by-step guide to configuring KServe's Serverless mode for vLLM autoscaling.
How to navigate LLM model names
Learning the naming conventions of large language models (LLMs) helps users select the right model for their needs.
Build and deploy a ModelCar container in OpenShift AI
Learn how to build a ModelCar container image and deploy it with OpenShift AI.
How to validate GitOps manifests
This article discusses how to validate GitOps manifests to improve the reliability and confidence of changes before merging.
3 patterns for deploying Helm charts with Argo CD
Learn about deploying Helm charts with three popular patterns for Argo CD on Red Hat OpenShift. Plus, discover the advantages and disadvantages of each pattern.