Achieve state of the art inference performance with modern accelerators on Kubernetes
This report presents the forensic synthetic code analysis of llm-d/llm-d, a Shell project with 4,308 GitHub stars. SynthScan v2.0 examined 96,942 lines of code across 1018 source files, recording 500 pattern matches distributed across 13 syntactic categories. The overall adjusted score of 11.0 places this repository in the Low AI signal band.
The scanner applied 160+ deterministic lexical heuristics, multi-line block detectors, abstract syntax tree depth profilers, and a cross-file Jaccard similarity matrix to construct a statistically normalised synthetic code estimate. All matches are individually weighted by severity coefficient and contextual multiplier before summation, and the resulting headline score is temporally discounted to account for the repository's development history relative to the commercial emergence of large language model coding tooling (November 2022 onward).
This chart maps the temporal evolution of the adjusted synthetic code score across successive scan runs. An upward trajectory indicates ongoing incorporation of AI-generated code or expanding LLM-assisted scaffolding; a stable or declining trajectory may reflect active human refactoring, code removal, or the adoption of stricter authorship policies. The dashed secondary line (right axis) independently tracks total raw pattern hit count, which can diverge from the normalised score when codebase size changes significantly between scans.
Classifies detected patterns by their diagnostic confidence and structural impact. CRITICAL patterns (coefficient 10) represent definitive synthetic signatures — hallucinated imports, explicit LLM attribution metadata — virtually never produced by human authors. HIGH (5) indicates strong structural tells such as cross-file repetition or cross-linguistic idioms. MEDIUM (2) covers recognisable conversational padding and AI-specific vocabulary. LOW (1) captures subtle indicators like tautological comments and generic boilerplate that require density to carry independent signal.
This horizontal bar chart decomposes the repository's raw synthetic code score by top-level directory, allowing you to pinpoint precisely which modules or components carry the highest AI authorship density. Directories with disproportionately high scores relative to their size warrant targeted manual review: concentrated AI signatures often trace back to mass-generated configuration layers, auto-ported test suites, LLM-scaffolded boilerplate classes, or entire subsystems authored under heavy copilot assistance. Use this view to prioritise your human code-review effort.
The scanner identified 500 distinct pattern matches across 13 syntactic categories. Each entry below represents a discrete location in the source code where the engine recorded a statistically significant AI authorship indicator. Expand any category row to inspect the individual file paths, line numbers, code snippets, and the lexical context (CODE, COMMENT, or STRING) in which each match was detected.
Reading the findings table: The Severity column indicates the diagnostic confidence level (CRITICAL / HIGH / MEDIUM / LOW). The Context column identifies whether the match occurred inside executable code, an inline comment, or a string literal — comment-context matches receive a ×1.5 weight because LLMs systematically over-annotate. The ⚡ bolt icon marks clustered matches: three or more patterns within a 10-line window, each receiving an additional ×1.5 density multiplier as dense clusters constitute far stronger evidence of synthetic authorship than isolated hits.
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| MEDIUM | docker/scripts/cuda/builder/build-ucx.sh | 5 | # -------------------------------------------- | COMMENT |
| MEDIUM | docker/scripts/cuda/builder/build-ucx.sh | 9 | # -------------------------------------------- | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 174 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 178 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 210 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 213 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 245 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 248 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 287 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/kong.md | 290 | # ───────────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 262 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 265 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 281 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 284 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 300 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | docs/operations/serve-external-apis/litellm.md | 303 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| MEDIUM | guides/flow-control/scripts/tuning_wizard.py | 18 | # ========================================== | COMMENT |
| MEDIUM | guides/flow-control/scripts/tuning_wizard.py | 20 | # ========================================== | COMMENT |
| MEDIUM⚡ | guides/flow-control/scripts/tuning_wizard.py | 45 | # ========================================== | COMMENT |
| MEDIUM⚡ | guides/flow-control/scripts/tuning_wizard.py | 47 | # ========================================== | COMMENT |
| MEDIUM | guides/flow-control/scripts/tuning_wizard.py | 104 | # ========================================== | COMMENT |
| MEDIUM | guides/flow-control/scripts/tuning_wizard.py | 106 | # ========================================== | COMMENT |
| MEDIUM | …s/recipes/observability/alerts/epp-alerting-rules.yaml | 9 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …s/recipes/observability/alerts/epp-alerting-rules.yaml | 11 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …s/recipes/observability/alerts/epp-alerting-rules.yaml | 63 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …s/recipes/observability/alerts/epp-alerting-rules.yaml | 65 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …ed-latency-routing/benchmark-templates/multimodal.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …ed-latency-routing/benchmark-templates/multimodal.yaml | 17 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …edicted-latency-routing/benchmark-templates/guide.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …edicted-latency-routing/benchmark-templates/guide.yaml | 25 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 21 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | guides/agentic-serving/benchmark-templates/guide.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | guides/agentic-serving/benchmark-templates/guide.yaml | 22 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 23 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 25 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 86 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 88 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 96 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 98 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 103 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 105 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 62 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 64 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 119 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 121 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 179 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 181 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 192 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 194 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 213 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 215 | # --------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | scripts/guide.py | 91 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM⚡ | scripts/guide.py | 93 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 164 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 166 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 224 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 226 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 285 | # -------------------------------------------------------------------------- | COMMENT |
| MEDIUM | scripts/guide.py | 287 | # -------------------------------------------------------------------------- | COMMENT |
| 109 more matches not shown… | ||||
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | docs/operations/serve-external-apis/kong.md | 46 | ## Step 1: Create Secret for External Provider Keys | COMMENT |
| LOW | docs/operations/serve-external-apis/kong.md | 62 | ## Step 2: Install Kong Ingress Controller + Gateway via Helm | COMMENT |
| LOW | docs/operations/serve-external-apis/kong.md | 110 | ## Step 3: Create GatewayClass, Gateway, and Placeholder Service | COMMENT |
| LOW | docs/operations/serve-external-apis/kong.md | 159 | ## Step 4: Configure Routes and `ai-proxy` Plugins | COMMENT |
| LOW | docs/operations/serve-external-apis/kong.md | 355 | ## Step 5: Verification | COMMENT |
| LOW | docs/operations/serve-external-apis/litellm.md | 41 | ## Step 1: Create Secrets for LiteLLM | COMMENT |
| LOW | docs/operations/serve-external-apis/litellm.md | 68 | ## Step 2: Set Up PostgreSQL and Redis Backends | COMMENT |
| LOW | docs/operations/serve-external-apis/litellm.md | 208 | ## Step 3: Construct LiteLLM Helm Values and Install | COMMENT |
| LOW | docs/operations/serve-external-apis/litellm.md | 328 | ## Step 4: Verify Both Models | COMMENT |
| LOW | docs/operations/observability/setup.md | 12 | ## Step 1: Install Prometheus and Grafana | COMMENT |
| LOW | docs/operations/observability/setup.md | 145 | ## Step 2: Load Grafana Dashboards | COMMENT |
| LOW | docs/operations/observability/setup.md | 176 | ## Step 3: Install Distributed Tracing (Optional) | COMMENT |
| LOW | docs/operations/observability/alerting.md | 20 | ## Step 1: Apply the Alerting Rules | COMMENT |
| LOW | docs/operations/observability/alerting.md | 31 | ## Step 2: Verify | COMMENT |
| LOW | docs/operations/observability/tracing.md | 24 | ## Step 1: Deploy OTel Collector and Jaeger | COMMENT |
| LOW | docs/operations/observability/tracing.md | 68 | ## Step 2: Enable Tracing on the Model Server and Routing Proxy | COMMENT |
| LOW | docs/operations/observability/tracing.md | 89 | ## Step 3: Enable Tracing on EPP | COMMENT |
| LOW | docs/operations/observability/tracing.md | 104 | ## Step 4: View Traces | COMMENT |
| LOW | docs/operations/observability/metrics.md | 17 | ## Step 1: Enable Model Server Metrics | COMMENT |
| LOW | docs/operations/observability/metrics.md | 76 | ## Step 3: Enable EPP Metrics | COMMENT |
| LOW | docs/operations/observability/metrics.md | 212 | ## Step 4: View Dashboards | COMMENT |
| LOW | docs/operations/observability/metrics.md | 264 | ## Step 5: Query Metrics | COMMENT |
| LOW⚡ | docs/infrastructure/providers/openshift-aws/README.md | 16 | ## Step 1: Create a Red Hat OpenShift Service on AWS (ROSA) Cluster | COMMENT |
| LOW⚡ | docs/infrastructure/providers/openshift-aws/README.md | 25 | ## Step 2: Add a Machineset with NVIDIA GPU Instances | COMMENT |
| LOW⚡ | docs/infrastructure/providers/openshift-aws/README.md | 34 | ## Step 3: Enable GPU support on OpenShift with the NFD and GPU Operators | COMMENT |
| LOW | docs/infrastructure/gateway/gke.md | 26 | ## Step 1: Install Gateway API and Gateway API Inference Extension CRDs | COMMENT |
| LOW | docs/infrastructure/gateway/gke.md | 45 | ## Step 2: Deploy the Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/gke.md | 69 | ## Step 3: Verify the Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/gke.md | 86 | ## Step 4: Send a Request | COMMENT |
| LOW | docs/infrastructure/gateway/agentgateway.md | 18 | ## Step 1: Install Gateway API and Gateway API Inference Extension CRDs | COMMENT |
| LOW | docs/infrastructure/gateway/agentgateway.md | 22 | ## Step 2: Install Agentgateway | COMMENT |
| LOW | docs/infrastructure/gateway/agentgateway.md | 58 | ## Step 3: Deploy the Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/agentgateway.md | 99 | ## Step 4: Send a Request | COMMENT |
| LOW | docs/infrastructure/gateway/envoy-ai-gateway.md | 18 | ## Step 1: Install Gateway API and Gateway API Inference Extension CRDs | COMMENT |
| LOW | docs/infrastructure/gateway/envoy-ai-gateway.md | 22 | ## Step 2: Install Envoy AI Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/envoy-ai-gateway.md | 128 | ## Step 3: Deploy the Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/envoy-ai-gateway.md | 157 | ## Step 4: Send a Request | COMMENT |
| LOW | docs/infrastructure/gateway/istio.md | 15 | ## Step 1: Install Gateway API and Gateway API Inference Extension CRDs | COMMENT |
| LOW | docs/infrastructure/gateway/istio.md | 19 | ## Step 2: Install Istio | COMMENT |
| LOW | docs/infrastructure/gateway/istio.md | 44 | ## Step 3: Deploy the Gateway | COMMENT |
| LOW | docs/infrastructure/gateway/istio.md | 73 | ## Step 4: Send a Request | COMMENT |
| LOW | …prefix-cache/modelserver/tpu/base/vllm/patch-vllm.yaml | 18 | # WARNING: This changes the HOST memory settings, not just the container. | COMMENT |
| LOW | guides/flow-control/tuning.md | 43 | ### Step 1: Gather System Telemetry (For Compute Bound) | COMMENT |
| LOW | guides/flow-control/tuning.md | 55 | ### Step 2: Gather Workload Statistics (For Memory Bound) | COMMENT |
| LOW | guides/flow-control/tuning.md | 65 | ### Step 3: Run the Tuning Wizard | COMMENT |
| LOW | guides/flow-control/tuning.md | 103 | ### Step 4: Apply Configuration | COMMENT |
| LOW | guides/multi-model-routing/README.md | 57 | ## Step 1: Deploy IPP | COMMENT |
| LOW | guides/multi-model-routing/README.md | 82 | ## Step 2: Create Model Mapping ConfigMaps | COMMENT |
| LOW | guides/multi-model-routing/README.md | 95 | ## Step 3: Configure HTTPRoutes | COMMENT |
| LOW | guides/multi-model-routing/README.md | 107 | ## Step 4: Test the Deployment | COMMENT |
| LOW | …s/agentic-serving/modelserver/tpu/vllm/patch-vllm.yaml | 15 | # WARNING: This changes the HOST memory settings, not just the container. | COMMENT |
| LOW | guides/p2p-kv-cache-sharing/benchmarking/README.md | 191 | ## Step 0 - pull-versus-recompute crossover (single request) | COMMENT |
| LOW | guides/batch-serving/asynchronous-processing/README.md | 40 | #### Step 1: Deploy llm-d Router | COMMENT |
| LOW | guides/batch-serving/asynchronous-processing/README.md | 53 | #### Step 2: Configure Values | COMMENT |
| LOW | guides/batch-serving/asynchronous-processing/README.md | 60 | #### Step 3: Deploy Async Processor | COMMENT |
| LOW | guides/batch-serving/batch-gateway/README.md | 45 | ### Step 1: Create the Namespace | COMMENT |
| LOW | guides/batch-serving/batch-gateway/README.md | 52 | ### Step 2: Create the Secrets | COMMENT |
| LOW | guides/batch-serving/batch-gateway/README.md | 64 | ### Step 3: Configure the llm-d Router URL | COMMENT |
| LOW | guides/batch-serving/batch-gateway/README.md | 76 | ### Step 4: Create the File Storage PVC | COMMENT |
| LOW | guides/batch-serving/batch-gateway/README.md | 97 | ### Step 5: Deploy | COMMENT |
| 21 more matches not shown… | ||||
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| MEDIUM | SIGS.md | 54 | | **[SIG Agentic Inference](#sig-agentic-inference)** | Optimizing inference for agentic and multi-step AI workloads | • | CODE |
| MEDIUM | SIGS.md | 286 | - **Slack Channel**: [#sig-agentic-inference](https://llm-d.slack.com/archives/C0ALHNZJCFJ) | CODE |
| MEDIUM | …des/wide-ep-lws/modelserver/gpu/vllm-glm-5.2/README.md | 132 | [blog post](https://llm-d.ai/blog/serving-glm-5-2-agentic-workloads-on-llm-d). | CODE |
| MEDIUM | guides/recipes/router/calibration/calibrate.sh | 12 | # GUIDE_NAME=agentic-serving NAMESPACE=llm-d-agentic-serving \ | COMMENT |
| MEDIUM | guides/predicted-latency-routing/README.md | 344 | Uses the agentic-serving guide's `inference-perf` workload, tuned for the 480B model and very long (up to 256K-token) co | CODE |
| MEDIUM | guides/predicted-latency-routing/README.md | 347 | # Fetch the existing-stack benchmark runner from llm-d-benchmark (the script the agentic-serving guide uses). | COMMENT |
| MEDIUM | …atency-routing/modelserver/tpu/vllm/kustomization.yaml | 4 | # Reuse the agentic-serving guide's TPU model server: Qwen3-Coder-480B-A35B-Instruct-FP8 | COMMENT |
| MEDIUM | …atency-routing/router/predicted-latency-pd.values.yaml | 35 | # the 256K agentic sizing, so a smaller match window keeps EPP prefix-matching | COMMENT |
| MEDIUM | …d-latency-routing/router/predicted-latency.values.yaml | 27 | # Sized for the 256K-context TPU agentic case (Qwen3-Coder-480B): matching longer | COMMENT |
| MEDIUM | …d-latency-routing/router/predicted-latency.values.yaml | 28 | # prefixes raises the cache-hit rate on long, shared agentic prompts. Tradeoff: it | COMMENT |
| MEDIUM | …tency-routing/router/predicted-latency-slo.values.yaml | 30 | # Sized for the 256K-context TPU agentic case (Qwen3-Coder-480B): matching longer | COMMENT |
| MEDIUM | …tency-routing/router/predicted-latency-slo.values.yaml | 31 | # prefixes raises the cache-hit rate on long, shared agentic prompts. Tradeoff: it | COMMENT |
| MEDIUM | guides/agentic-serving/nemotron-3-ultra-550b-h200.md | 165 | # from the guide directory: guides/agentic-serving | COMMENT |
| MEDIUM | guides/agentic-serving/nemotron-3-ultra-550b-h200.md | 173 | # from the guide directory: guides/agentic-serving | COMMENT |
| MEDIUM | guides/agentic-serving/nemotron-3-ultra-550b-h200.md | 203 | curl -LJO "https://raw.githubusercontent.com/llm-d/llm-d/main/guides/${GUIDE_NAME}/benchmark-templates/agentic-serving-n | CODE |
| MEDIUM | guides/agentic-serving/README.md | 45 | - [GLM-5.2-FP8 on H200](glm-5-2-h200.md) — wide expert-parallel P/D-disaggregated serving with MTP and tiered KV-offload | CODE |
| MEDIUM | guides/agentic-serving/glm-5-2-h200.md | 8 | [GLM-5.2 agentic serving blog post](https://llm-d.ai/blog/serving-glm-5-2-agentic-workloads-on-llm-d), | CODE |
| MEDIUM | guides/agentic-serving/glm-5-2-h200.md | 180 | [blog post](https://llm-d.ai/blog/serving-glm-5-2-agentic-workloads-on-llm-d) for the full | CODE |
| MEDIUM | …erving/modelserver/gpu/vllm/glm-5-2/kustomization.yaml | 3 | # Recommended agentic-serving deployment of GLM-5.2-FP8 on H200: the wide-ep-lws | COMMENT |
| MEDIUM | …ng/modelserver/gpu/vllm/nemotron-3-ultra/gke/README.md | 109 | Once the model is stored in your GCS, please refer back to <https://github.com/llm-d/llm-d/blob/main/guides/agentic-serv | CODE |
| MEDIUM⚡ | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 2 | # inference-perf workload template for the agentic-serving guide. | COMMENT |
| MEDIUM⚡ | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 4 | # Render with `envsubst < agentic-serving-nemotron-3-ultra.yaml > config.yaml` AFTER exporting the | COMMENT |
| MEDIUM | guides/agentic-serving/benchmark-templates/guide.yaml | 2 | # inference-perf workload template for the agentic-serving guide. | COMMENT |
| MEDIUM⚡ | …-serving/router/agentic-serving-tpu-disagg.values.yaml | 1 | ## agentic-serving (TPU / P/D-disaggregated) guide overrides for the llm-d router. | COMMENT |
| MEDIUM⚡ | …-serving/router/agentic-serving-tpu-disagg.values.yaml | 4 | ## This deployment serves the agentic code-generation workload on Google TPUs with | COMMENT |
| MEDIUM⚡ | …-serving/router/agentic-serving-tpu-disagg.values.yaml | 8 | ## by the non-disaggregated TPU deployment in agentic-serving.values.yaml. | COMMENT |
| MEDIUM⚡ | …agentic-serving/router/agentic-serving-gpu.values.yaml | 1 | ## agentic-serving (GPU / P/D-disaggregated) guide overrides for the llm-d router. | COMMENT |
| MEDIUM⚡ | …agentic-serving/router/agentic-serving-gpu.values.yaml | 4 | ## This deployment serves the agentic code-generation workload on NVIDIA GPUs with | COMMENT |
| MEDIUM⚡ | …agentic-serving/router/agentic-serving-gpu.values.yaml | 8 | ## by the non-disaggregated TPU deployment in agentic-serving.values.yaml. | COMMENT |
| MEDIUM | …p2p-kv-cache-sharing/benchmark-results/glm-5.2-h200.md | 119 | ## 2P2D agentic C64 policy comparison | COMMENT |
| MEDIUM | …p2p-kv-cache-sharing/benchmark-results/glm-5.2-h200.md | 122 | study](https://llm-d.ai/blog/serving-glm-5-2-agentic-workloads-on-llm-d) | CODE |
| MEDIUM | …sharing/benchmark-results/qwen3-30b-h200-pd-agentic.md | 1 | # Qwen/Qwen3-30B-A3B-Thinking P2P KV Cache Sharing on P/D (H200, agentic) | COMMENT |
| MEDIUM | guides/coord-disaggregation/router/httproute.yaml | 4 | # (conditional-decode, media handling, orchestration) entirely -- there is no portable | COMMENT |
| MEDIUM | guides/coord-disaggregation/router/httproute-3-epp.yaml | 4 | # pipeline (conditional-decode, media handling, orchestration) entirely -- there is | COMMENT |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | docker/scripts/cuda/runtime/install-vllm.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docker/scripts/cuda/builder/build-nixl.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docker/scripts/cuda/builder/build-nvshmem.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docker/scripts/cuda/builder/build-gdrcopy.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docker/scripts/cuda/builder/build-compiled-wheels.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docker/scripts/cuda/builder/build-ucx.sh | 1 | #!/bin/bash | COMMENT |
| LOW | docs/operations/serve-external-apis/litellm.md | 281 | # ───────────────────────────────────────────────────────────────── | COMMENT |
| LOW | …ion/modelserver/gpu/vllm-ds/base/disaggregatedset.yaml | 21 | # domain, zone, ...) and spread slices across domains, so prefill hands KV | COMMENT |
| LOW | …des/no-kubernetes-deployment/router/epp/endpoints.yaml | 1 | # Endpoints file consumed by the file-discovery plugin. | COMMENT |
| LOW | guides/no-kubernetes-deployment/router/envoy/envoy.yaml | 1 | # Envoy config for the no-Kubernetes file-discovery deployment. | COMMENT |
| LOW | …odelserver/gpu/vllm-deepseek-r1-0528/base/prefill.yaml | 81 | --all2all-backend deepep_high_throughput | COMMENT |
| LOW | guides/wide-ep-lws/render/service.yaml | 1 | apiVersion: v1 | COMMENT |
| LOW | guides/wide-ep-lws/monitoring/kustomization.yaml | 1 | apiVersion: kustomize.config.k8s.io/v1alpha1 | COMMENT |
| LOW | guides/wide-ep-lws/router/precise-routing.values.yaml | 1 | # Precise prefix-cache routing overrides for wide-ep-lws.values.yaml: replaces | COMMENT |
| LOW | guides/flow-control/scripts/nightly-deploy-gke.sh | 1 | #!/usr/bin/env bash | COMMENT |
| LOW | …precise-prefix-cache-routing/render/kustomization.yaml | 1 | apiVersion: kustomize.config.k8s.io/v1beta1 | COMMENT |
| LOW | guides/precise-prefix-cache-routing/render/service.yaml | 1 | apiVersion: v1 | COMMENT |
| LOW | …fix-cache-routing/render/standalone/kustomization.yaml | 1 | apiVersion: kustomize.config.k8s.io/v1beta1 | COMMENT |
| LOW | …outing/router/precise-prefix-cache-routing.values.yaml | 1 | ## precise-prefix-cache-routing guide overrides for the router. | COMMENT |
| LOW | guides/recipes/observability/generate-traffic-pd.sh | 1 | #!/bin/bash | COMMENT |
| LOW | …/recipes/autoscaling/metrics-reader/kustomization.yaml | 1 | # Shared OpenShift auth bundle for KEDA -> Thanos Querier: a dedicated | COMMENT |
| LOW | guides/recipes/router/base.values.yaml | 21 | protocol: http # http, grpc | COMMENT |
| LOW | …router/calibration/calibrate-min-cached-token-delta.sh | 1 | #!/bin/bash | COMMENT |
| LOW | …router/calibration/calibrate-min-cached-token-delta.sh | 21 | # ... DP_RANKS=16 ./calibrate-min-cached-token-delta.sh | COMMENT |
| LOW | …router/calibration/calibrate-min-cached-token-delta.sh | 41 | # DP>1 fleet calibrates correctly without this. Set it to | COMMENT |
| LOW | …er/calibration/calibration-min-cached-token-delta.yaml | 41 | # Measures the pull-versus-recompute crossover between two live model | COMMENT |
| LOW | guides/recipes/router/calibration/calibrate.sh | 1 | #!/bin/bash | COMMENT |
| LOW | guides/recipes/router/calibration/calibrate.sh | 21 | # MODEL_NAME — model name vLLM is serving (default: Qwen/Qwen3-32B) | COMMENT |
| LOW | …ed-latency-routing/benchmark-templates/multimodal.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| LOW | …edicted-latency-routing/benchmark-templates/guide.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| LOW | …outing/router/predicted-latency-multimodal.values.yaml | 41 | # The guide's benchmark workload peaks around ~5K tokens per request | COMMENT |
| LOW | …tency-routing/router/predicted-latency-slo.values.yaml | 1 | ## Predicted Latency-Based Scheduling — SLO-aware setup (streamingMode: true). | COMMENT |
| LOW | …-baseline/modelserver/gpu/vllm/gpt-oss/patch-vllm.yaml | 21 | - "--reasoning-parser=openai_gptoss" | COMMENT |
| LOW | …zed-baseline/modelserver/gpu/vllm/base/patch-vllm.yaml | 21 | - name: HF_TOKEN | COMMENT |
| LOW | …baseline/modelserver/gpu/sglang/base/patch-sglang.yaml | 21 | # - "--otlp-traces-endpoint=http://otel-collector:4317" | COMMENT |
| LOW | …ptimized-baseline/modelserver/cpu/vllm/patch-vllm.yaml | 21 | - name: HF_TOKEN | COMMENT |
| LOW | …ptimized-baseline/modelserver/xpu/vllm/patch-vllm.yaml | 21 | - "--disable-access-log-for-endpoints=/health,/metrics,/v1/models" | COMMENT |
| LOW | …ized-baseline/modelserver/amd/sglang/patch-sglang.yaml | 21 | # - "--enable-trace" | COMMENT |
| LOW | …mized-baseline/modelserver/tpu/v6/vllm/patch-vllm.yaml | 21 | # - "--otlp-traces-endpoint=http://otel-collector:4317" | COMMENT |
| LOW | …mized-baseline/modelserver/tpu/v7/vllm/patch-vllm.yaml | 21 | # - "--otlp-traces-endpoint=http://otel-collector:4317" | COMMENT |
| LOW | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| LOW | guides/agentic-serving/benchmark-templates/guide.yaml | 1 | # --------------------------------------------------------------------------- | COMMENT |
| LOW | …e-sharing/modelserver/gpu/vllm/rdma/kustomization.yaml | 1 | apiVersion: kustomize.config.k8s.io/v1beta1 | COMMENT |
| LOW | …e-sharing/modelserver/gpu/vllm/base/kustomization.yaml | 1 | apiVersion: kustomize.config.k8s.io/v1beta1 | COMMENT |
| LOW | …ache-sharing/modelserver/gpu/vllm/base/patch-vllm.yaml | 21 | - --tensor-parallel-size=1 | COMMENT |
| LOW | …-cache-sharing/router/p2p-kv-cache-sharing.values.yaml | 1 | ## p2p-kv-cache-sharing guide overrides for the router. | COMMENT |
| LOW | …-cache-sharing/router/p2p-kv-cache-sharing.values.yaml | 21 | ## arm is the default because affinity is the safer placement in the | COMMENT |
| LOW | …-cache-sharing/router/p2p-kv-cache-sharing.values.yaml | 81 | # hot shared-prefix block has on any fleet above five | COMMENT |
| LOW | …es/coord-disaggregation/coordinator/patch-pd-only.yaml | 1 | # PD-only pipeline: drops replace-media-urls, render, and encode -- see the | COMMENT |
| LOW | guides/coord-disaggregation/router/httproute.yaml | 21 | - backendRefs: | COMMENT |
| LOW | …egation/router/coord-disaggregation-encode.values.yaml | 1 | ## Coordinator Disaggregation guide overrides for the encode-only EPP / InferencePool, | COMMENT |
| LOW | …gation/router/coord-disaggregation-prefill.values.yaml | 1 | ## Coordinator Disaggregation guide overrides for the prefill-only EPP / InferencePool, | COMMENT |
| LOW | …disaggregation/router/coord-disaggregation.values.yaml | 1 | ## Coordinator Disaggregation guide overrides for the single EPP / InferencePool that | COMMENT |
| LOW | …disaggregation/router/coord-disaggregation.values.yaml | 41 | registry: ghcr.io/llm-d | COMMENT |
| LOW | …egation/router/coord-disaggregation-decode.values.yaml | 1 | ## Coordinator Disaggregation guide overrides for the decode-only EPP / InferencePool, | COMMENT |
| LOW | …us-processing/multitenant/values/redis/quota-only.yaml | 1 | # Values for the multi-tenant scenario — team × tier × MODEL (Scenarios A + B), | COMMENT |
| LOW | …ng/multitenant/values/redis/saturation-prometheus.yaml | 1 | # Redis SortedSet backend + self-hosted Prometheus + Grafana — team × tier × MODEL | COMMENT |
| LOW | …s-processing/multitenant/values/pubsub/quota-only.yaml | 1 | # Values for the multi-tenant Pub/Sub demo — team × tier × MODEL (Scenarios A + B). | COMMENT |
| LOW | …g/multitenant/values/pubsub/saturation-prometheus.yaml | 1 | # Pub/Sub backend + self-hosted Prometheus + Grafana — team × tier × MODEL with | COMMENT |
| LOW | …ocessing/multitenant/values/pubsub/saturation-gmp.yaml | 1 | # Scenario C — team × tier × MODEL with per-model saturation (GCP-native, GMP). | COMMENT |
| 36 more matches not shown… | ||||
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | docker/scripts/cpu/install_nixl.py | 45 | def install_system_dependencies(): | CODE |
| LOW | docker/scripts/cpu/install_nixl.py | 74 | def build_and_install_prerequisites(args): | CODE |
| LOW⚡ | guides/flow-control/scripts/tuning_wizard.py | 49 | def calculate_compute_constraint(throughput: float, latency_sec: float) -> int: | CODE |
| LOW⚡ | guides/flow-control/scripts/tuning_wizard.py | 53 | def calculate_memory_constraint( | CODE |
| LOW | guides/flow-control/scripts/tuning_wizard.py | 91 | def calculate_lookahead_buffer(active_batch: int, max_num_batched_tokens: int, isl_mean: Optional[float]) -> int: | CODE |
| LOW | scripts/lint-envvars.py | 36 | def find_locally_defined_vars(script_content: str) -> Set[str]: | CODE |
| LOW | scripts/lint-dockerfile-envvars.py | 12 | def parse_script_requirements(script_path: Path) -> Set[str]: | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 49 | def test_env_default_emitted_verbatim(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 58 | def test_env_override_is_shell_quoted(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 65 | def test_sensitive_without_override_emits_no_placeholder(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 72 | def test_sensitive_with_override_is_exported_quoted(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 78 | def test_unknown_var_override_is_an_error(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 83 | def test_override_outside_declared_values_is_an_error(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 92 | def test_yaml_null_and_bool_env_values_fail_validation(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 105 | def test_non_bool_sensitive_flag_fails_validation(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 114 | def test_provenance_records_var_names_not_values(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 126 | def test_skip_in_filters_by_context(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 163 | def test_fully_filtered_section_is_marked_not_silent(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 170 | def test_steps_join_with_blank_lines_and_no_when_comment(): | CODE |
| LOW⚡ | scripts/tests/test_guide.py | 182 | def test_sections_emitted_in_cli_order(): | CODE |
| LOW | scripts/tests/test_guide.py | 139 | def test_when_filter_uses_default_and_override(): | CODE |
| LOW | scripts/tests/test_guide.py | 152 | def test_when_on_sensitive_var_without_override_is_an_error(): | CODE |
| LOW | scripts/tests/test_guide.py | 195 | def test_unknown_section_is_an_error(): | CODE |
| LOW | scripts/tests/test_guide.py | 200 | def test_emitting_parent_section_concatenates_all_subgroups(): | CODE |
| LOW | scripts/tests/test_guide.py | 216 | def test_cli_refuses_invalid_yaml(tmp_path, capsys): | CODE |
| LOW | scripts/tests/test_guide.py | 243 | def test_cli_rejects_malformed_var(tmp_path, capsys): | CODE |
| LOW | scripts/tests/test_guide.py | 252 | def test_cli_selection_flags_still_guarded(command, tmp_path, capsys): | CODE |
| LOW | scripts/tests/test_guide.py | 284 | def test_nightly_crds_use_env_sh_release_urls(flow_control): | CODE |
| LOW | scripts/tests/test_guide.py | 296 | def test_nightly_deploy_carries_ci_overrides(flow_control): | CODE |
| LOW | scripts/tests/test_guide.py | 310 | def test_nightly_ci_context_drops_clone_and_secrets(flow_control): | CODE |
| LOW | scripts/tests/test_guide.py | 320 | def test_nightly_wrapper_vars_match_this_contract(flow_control): | CODE |
| LOW | scripts/tests/test_guide.py | 331 | def test_nightly_modelserver_mirror_matches_guide(flow_control): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 93 | def test_empty_workspace_returns_none(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 97 | def test_empty_namespace_returns_none(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 101 | def test_no_results_dir_anywhere(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 112 | def test_namespace_mismatch_returns_none(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 117 | def test_results_without_run_metadata_returns_none(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 123 | def test_concurrent_runs_isolated_by_namespace(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 133 | def test_sequential_same_namespace_newest_wins(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 142 | def test_multiple_experiments_in_one_run_returned_together(self): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 196 | def test_stops_at_non_digit_segment(self, mock_kubectl): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 202 | def test_empty_output_returns_none(self, mock_kubectl): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 207 | def test_leading_non_digit_returns_none(self, mock_kubectl): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 401 | def test_nonzero_exit_returns_empty(self, mock_run): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 409 | def test_timeout_returns_empty(self, mock_run): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 416 | def test_generic_exception_returns_empty(self, mock_run): | CODE |
| LOW⚡ | …cripts/nightly-e2e-verification/test_verify_helpers.py | 425 | def test_decode_prefill_match(self, mock_kubectl): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 154 | def test_stray_metadata_outside_results_ignored(self): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 167 | def test_malformed_yaml_skipped(self): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 184 | def test_parses_local_version_tag(self, mock_kubectl): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 231 | def test_missing_aggregated_defaults_to_empty(self): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 238 | def test_non_dict_top_level_values_ignored(self): | CODE |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 342 | def test_name_includes_reducer(self): | CODE |
| LOW | …ly-e2e-verification/tiered-prefix-cache/test_verify.py | 58 | def test_volume_without_pvc_claim(self, mock_kubectl): | CODE |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| MEDIUM | …re/providers/digitalocean/gpu-configs/l40s-values.yaml | 38 | # Enable comprehensive metrics collection | COMMENT |
| MEDIUM | …ders/digitalocean/gpu-configs/rtx-4000-ada-values.yaml | 47 | # Enable comprehensive metrics collection | COMMENT |
| MEDIUM | …ders/digitalocean/gpu-configs/rtx-6000-ada-values.yaml | 36 | # Enable comprehensive metrics collection | COMMENT |
| MEDIUM | guides/pd-disaggregation/benchmark-templates/tpu.yaml | 17 | namespace: *namespace # Namespace where harness is deployed. Typically with stack. | CODE |
| MEDIUM | guides/pd-disaggregation/benchmark-templates/tpu.yaml | 28 | workload: # yaml configuration for harness workload(s) | CODE |
| MEDIUM | guides/flow-control/guide.yaml | 153 | # --harness/--workload stay literal because flow-control has no dedicated | COMMENT |
| MEDIUM | …ed-latency-routing/benchmark-templates/multimodal.yaml | 32 | namespace: *namespace # Namespace where harness is deployed. Typically with stack. | CODE |
| MEDIUM | …edicted-latency-routing/benchmark-templates/guide.yaml | 42 | namespace: *namespace # Namespace where harness is deployed. Typically with stack. | CODE |
| MEDIUM | …edicted-latency-routing/benchmark-templates/guide.yaml | 52 | workload: # yaml configuration for harness workload(s) | CODE |
| MEDIUM | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 38 | namespace: *namespace # Namespace where harness is deployed. Typically with stack. | CODE |
| MEDIUM | …chmark-templates/agentic-serving-nemotron-3-ultra.yaml | 48 | workload: # yaml configuration for harness workload(s) | CODE |
| MEDIUM | guides/agentic-serving/benchmark-templates/guide.yaml | 39 | namespace: *namespace # Namespace where harness is deployed. Typically with stack. | CODE |
| MEDIUM | guides/agentic-serving/benchmark-templates/guide.yaml | 49 | workload: # yaml configuration for harness workload(s) | CODE |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 42 | # Source the shared guide environment so this harness tests the same llm-d Router | COMMENT |
| MEDIUM | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 139 | # This harness is CPU-only by design (no GPU in CI/dev kind). | COMMENT |
| LOW | scripts/warn-vllm-precompiled.sh | 41 | # Append to vllm user's rc if it exists; otherwise just create it | COMMENT |
| MEDIUM | …-workload-autoscaling-keda-epp-ibm-acc-gpu-vllm-x.yaml | 17 | # decode_pods/prefill_pods/harness/workload are accepted only for | COMMENT |
| MEDIUM | …flows/nightly-e2e-flow-control-gke-acc-gpu-vllm-x.yaml | 8 | # Why not the shared benchmark harness? Two reasons, both still true: | COMMENT |
| MEDIUM | …flows/nightly-e2e-flow-control-gke-acc-gpu-vllm-x.yaml | 13 | # with "unable to detect model". (The harness has a flow-control special | COMMENT |
| MEDIUM | .github/workflows/slash-test-nightly.yaml | 13 | # harness=<name> Benchmark harness (default: inference-perf) | COMMENT |
| MEDIUM | .github/scripts/e2e/e2e-validate-flow-control.sh | 274 | # missed by scrape timing, so it is the robust backbone assertion. | COMMENT |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | …ture/providers/digitalocean/verify-do-prerequisites.sh | 60 | # Check if we're on DOKS | COMMENT |
| LOW | …ture/providers/digitalocean/verify-do-prerequisites.sh | 123 | # Check if DO CSI driver is available | COMMENT |
| LOW | …/providers/digitalocean/monitoring/setup-monitoring.sh | 72 | # Check if kubectl can connect to cluster | COMMENT |
| LOW | …/providers/digitalocean/monitoring/setup-monitoring.sh | 177 | # Check if release already exists | COMMENT |
| LOW | …prefix-cache/modelserver/tpu/base/vllm/patch-vllm.yaml | 22 | # Check if the VFIO IOMMU module parameter exists, and if so, increase the | COMMENT |
| LOW | …recipes/observability/generate-prometheus-tls-certs.sh | 185 | # Check if namespace exists | COMMENT |
| LOW | guides/recipes/observability/load-llm-d-dashboards.sh | 31 | # Check if namespace exists | COMMENT |
| LOW | guides/recipes/observability/load-llm-d-dashboards.sh | 37 | # Check if dashboard directory exists | COMMENT |
| LOW | …es/recipes/observability/install-prometheus-grafana.sh | 240 | # Check if certificates already exist | COMMENT |
| LOW | …es/recipes/observability/install-prometheus-grafana.sh | 268 | # Check if user workload monitoring is enabled | COMMENT |
| LOW | …s/agentic-serving/modelserver/tpu/vllm/patch-vllm.yaml | 19 | # Check if the VFIO IOMMU module parameter exists, and if so, increase the | COMMENT |
| LOW | .github/scripts/e2e/e2e-validate.sh | 134 | # Check if we got a specific error | COMMENT |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | guides/wide-ep-lws/monitoring/apply-scrape-configs.sh | 4 | # Usage: | COMMENT |
| LOW | …router/calibration/calibrate-min-cached-token-delta.sh | 16 | # Usage: | COMMENT |
| LOW | guides/recipes/router/calibration/calibrate.sh | 11 | # Usage: | COMMENT |
| LOW | …ynchronous-processing/multitenant/scripts/gcp-setup.sh | 8 | # Usage: | COMMENT |
| LOW | …hronous-processing/multitenant/scripts/gcp-teardown.sh | 4 | # Usage: | COMMENT |
| LOW⚡ | …batch-serving/asynchronous-processing/test/kind-e2e.sh | 16 | # Usage: | COMMENT |
| LOW | …hub/scripts/nightly-e2e-verification/verify-locally.sh | 11 | # Usage: | COMMENT |
| LOW⚡ | …ers/local-llm-d-cuda-builder/build-local-llm-d-cuda.sh | 5 | # Usage: | COMMENT |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| MEDIUM | …recipes/observability/generate-prometheus-tls-certs.sh | 104 | # Define the Prometheus service DNS names | COMMENT |
| MEDIUM | …recipes/observability/generate-prometheus-tls-certs.sh | 197 | # Create the secret | COMMENT |
| MEDIUM | …server/gpu/vllm/nemotron-3-ultra/gke/patch-decode.yaml | 46 | # Define the GCS FUSE volume specifications | COMMENT |
| MEDIUM | …erver/gpu/vllm/nemotron-3-ultra/gke/patch-prefill.yaml | 45 | # Define the GCS FUSE volume specifications | COMMENT |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | …er/calibration/calibration-min-cached-token-delta.yaml | 127 | except Exception: | CODE |
| LOW | …-autoscaling/slo-aware/benchmark-templates/plot_run.py | 97 | except Exception: | CODE |
| LOW | scripts/lint-envvars.py | 94 | except Exception as e: | CODE |
| LOW | scripts/lint-dockerfile-envvars.py | 16 | except Exception: | CODE |
| MEDIUM | scripts/lint-dockerfile-envvars.py | 180 | print(f"Error: Scripts directory not found: {scripts_dir}", file=sys.stderr) | CODE |
| MEDIUM | scripts/lint-dockerfile-envvars.py | 186 | print(f"Error: Dockerfile not found: {dockerfile}", file=sys.stderr) | CODE |
| LOW | .github/workflows/release-matrix.yaml | 66 | except Exception as exc: # noqa: BLE001 - best-effort snapshot | CODE |
| LOW | …hub/scripts/nightly-e2e-verification/verify_helpers.py | 330 | except Exception as e: | CODE |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | …-deepseek-r1-0528/scripts/generate-disaggregatedset.py | 4 | CODE | |
| LOW | scripts/guide.py | 69 | CODE | |
| LOW | …cripts/nightly-e2e-verification/test_verify_helpers.py | 6 | CODE | |
| LOW | …hub/scripts/nightly-e2e-verification/verify_helpers.py | 8 | CODE | |
| LOW | …ly-e2e-verification/tiered-prefix-cache/test_verify.py | 6 | CODE | |
| LOW | …nightly-e2e-verification/tiered-prefix-cache/verify.py | 28 | CODE | |
| LOW | …b/scripts/nightly-e2e-verification/_template/verify.py | 15 | CODE |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW | scripts/guide.py | 296 | CODE | |
| LOW | scripts/guide.py | 359 | CODE | |
| LOW | scripts/guide.py | 640 | CODE | |
| LOW | scripts/guide.py | 1271 | CODE | |
| LOW | scripts/lint-dockerfile-envvars.py | 45 | CODE |
| Severity | File | Line | Snippet | Context |
|---|---|---|---|---|
| LOW⚡ | scripts/guide.py | 81 | __all__ = [ | CODE |