Pinned Loading
Repositories
- llamastack-performance Public
Performance benchmarking and testing framework for LlamaStack on OpenShift
- mlperf-inference-6.1-redhat Public Forked from mlcommons/inference
Reference implementations of MLPerf® inference benchmarks
- composite-dra-driver Public
Composite DRA driver to manage/publish other DRA driver managed resources as a synthetic composite device to the scheduler.
- glm-5.x-playbook Public
Reproducible benchmark playbook for GLM-5.2-FP8 on NVIDIA H200 with vLLM. Covers two topologies: pipeline parallel (PP=2, TP=8, 2 nodes) and single-node replicas (TP=8, MTP enabled).
- fournos Public
Fournos is a Kubernetes operator that schedules benchmark jobs via Kueue and executes them as Tekton PipelineRuns on remote clusters through the FORGE framework.
- performance-dashboard Public
A comprehensive performance analysis dashboard for RHAIIS (Red Hat AI Inference server) benchmarks. This dashboard provides interactive visualizations and analysis of AI model performance across different accelerators, versions, and configurations.
- vllm-parallelism-advisor Public
Interactive decision tree for choosing vLLM multi-GPU parallelism strategies (TP, PP, DP, EP, DCP)
- benchconf Public
Standard benchmark configurations used by Red Hat's Performance and Scale for AI Platforms Team
- psap-public-claude-skills Public
A collection of psap-developed-skills that are publicly available for everyone to build on.
Top languages
Loading…
Most used topics
Loading…