doca-collectx-deployment
Use this skill to deploy and operate a CollectX (clx) based DOCA telemetry collector on a host or BlueField — wiring providers / counters into the collector, running the collection daemon, and shaping its exporters (Prometheus pull, Fluent Bit push, NetFlow, file / IPC) so the metrics actually leave
- 0
- Installs
- —
- Rating
- —
- Success rate
- 7
- Files scanned
Security scan
Scan passedNo risky patterns were found in the scanned files.
Content sha256 0a7375b7b9070207… — run codexguild_scan_skills after installing to verify your local copy.
Static analysis is a first line of defense, not a guarantee. Read the source
SKILL.md
DOCA CollectX telemetry deployment
Where to start: This skill is the bundle's home for
operating a CollectX (clx) based telemetry collector — the
collection framework that gathers provider counters into a
schema and ships them out through one or more exporters. It is a
deployment / operation skill, parallel to
doca-bare-metal-deployment
and doca-container-deployment:
it owns the runtime shape of a telemetry collector on the
operator's host or BlueField, not the library APIs the operator's
own program calls. If the user wants to stand up, wire, or debug
a collector and its exporters, open TASKS.md and
start at ## configure. If the question is
what surfaces does the collector even have and where is the
scope boundary, start at CAPABILITIES.md.
If the user has not installed DOCA yet, route to
doca-setup first.
The scope boundary (read this before anything else)
CollectX (clx) is NVIDIA's telemetry collection framework. It
underpins the DOCA Telemetry Service (DTS) — and DTS
as-deployed (the productized, NGC-shipped / kubelet-started
service container) is out of scope for this bundle per
AGENTS.md Non-goal #7.
This skill therefore draws a hard line and the agent MUST state
it up front:
- In scope here: the CollectX collection mechanism as a class (providers / counters → schema → collector daemon → exporters), and the operator deploying / running / debugging a collector that they own, plus the operator's own usage of the two in-bundle telemetry libraries when those feed or consume the collector.
- Routed to the DOCA telemetry libraries: the
hardware-counter reader API is owned by
doca-telemetry; the application-side publisher API (emit counters / events from a DOCA program) is owned bydoca-telemetry-exporter. This skill does not re-document either API surface. - Routed to public docs (Non-goal #7): the productized
DTS container — its packaged config schema, its built-in
provider set, its kubelet manifest, its NGC image — is
externally productized. Route every "operate the DTS service"
question to the
doca-public-knowledge-mapexternally-productized routing row and the public DTS guide it points at. The agent must NOT synthesize DTS config file names, provider knob names, or paths from memory.
The load-bearing first-touch failure this skill exists to
prevent is collapsing these four surfaces into "DOCA
telemetry": the clx collection mechanism, the doca-telemetry
reader library, the doca-telemetry-exporter publisher library,
and the productized DTS container are four different things with
four different owners. The agent surfaces the decomposition
BEFORE any config-level guidance.
Audience
This skill serves external operators standing up or running a CollectX-based telemetry collector on a host or BlueField they administer — people who already have:
- a DOCA install on the side they are collecting from (host x86
or BlueField Arm), verified per
doca-setup ## test, - a goal of getting counters off the box through a collector +
exporter, not of writing the reader / publisher library code
(that is the two
libs/skills above), and - access to the public DOCA Telemetry and DTS guides on
docs.nvidia.comas the authoritative source for any concrete provider name, schema field, flag, or config path.
It is not for:
- developers writing the hardware-counter reader API (route to
doca-telemetry) or the publisher API (route todoca-telemetry-exporter), - operators deploying / configuring the productized DTS container as a turnkey service — that is externally productized (Non-goal #7); route to the public DTS guide,
- fresh-no-install users — those belong on
doca-setup ## no-install.
The skill teaches the agent the procedure and the scope
boundary; it does not invent clx symbol names, provider names,
schema field names, exporter flag names, or config paths from
memory — those come from the live install and the public docs via
doca-public-knowledge-map.
When to load this skill
Load this skill when the user is doing hands-on deployment or operation of a CollectX-based telemetry collector and the question is about the collector runtime shape, not a library API. Concretely:
- Standing up a collector that gathers provider counters into a schema and ships them out — and deciding which export backend (Prometheus pull, Fluent Bit push, NetFlow, file / IPC) fits the downstream consumer.
- Wiring a provider / counter family into the collector and
confirming the device actually exposes it before the config
commits (the gate-before-commit rule, shared with
doca-telemetry-utils). - Turning on / shaping an exporter so the metrics actually leave the box, and confirming the downstream consumer receives them end-to-end (not just "the daemon is running").
- Diagnosing a collector that starts but produces no schema rows, or ships nothing downstream, or whose exporter endpoint is silent — walking the layered ladder rather than guessing.
- Recognising when the user is actually asking about the
productized DTS container (route to public docs, Non-goal #7),
the reader library (route to
doca-telemetry), or the publisher library (route todoca-telemetry-exporter) instead of the collection mechanism this skill owns.
Do not load this skill for: the hardware-counter reader API
(use doca-telemetry); the
publisher API (use
doca-telemetry-exporter);
operating the productized DTS container (route via
doca-public-knowledge-map
Non-goal #7); installing DOCA or preparing the env (use
doca-setup); or any hardware-state
change (use
doca-hardware-safety).
What this skill provides
This is a thin loader. The substantive material lives in two companion files:
CAPABILITIES.md— the collector deployment contract as a class: the four-surface decomposition (clx collection mechanism vs reader library vs publisher library vs productized DTS), the collection pipeline shape (providers / counters → schema → collector daemon → exporters), the export-backend surface (Prometheus pull, Fluent Bit push, NetFlow, file / IPC) at class level, the version overlay ondoca-version, the error taxonomy (collector won't start → no provider rows → schema mismatch → exporter silent → downstream skew → transport), the observability surface, and the safety policy (gate provider support before commit; collector is read-only against the device; route any mutating step todoca-hardware-safety; do not invent clx names / paths).TASKS.md— step-by-step workflows for the deployment verbs:configure,build(routing stub),modify,run,test,debug, plus aDeferred task verbsblock that routes out-of-scope questions (the two libraries, the productized DTS container, env prep, hardware-state change) to their owners.
The skill assumes a host or BlueField where DOCA is already
installed and healthy (per
doca-setup ## test) and the
operator can run the collector and reach its exporter sinks. It
does not cover installing DOCA — that path goes through
doca-setup — and it does not cover
the reader / publisher library APIs or the productized DTS
container.
Loading order
- Read this
SKILL.mdfirst to confirm the user's question is in scope (operating a CollectX-based collector and its exporters — NOT the reader / publisher library APIs, NOT the productized DTS container). - For the four-surface decomposition, the collection pipeline shape, the export-backend class surface, the version overlay, the error taxonomy, the observability surface, and the safety policy, see CAPABILITIES.md.
- For step-by-step workflows —
configure,build(routing stub),modify,run,test,debug, and theDeferred task verbsblock — see TASKS.md.
Both companion files cross-link to each other,
doca-version for the canonical
version-handling rules, and
doca-public-knowledge-map
whenever the right answer is "read the live config / public docs"
rather than collector-specific guidance.
Example questions this skill answers well
What this skill deliberately does not ship
Related skills
Files
7- BENCHMARK.md
da1cb003124.0 KB - CAPABILITIES.md
881f1f047a16.1 KB - SKILL.md
8a9d3317bc11.1 KB - TASKS.md
c188bb2cb616.5 KB - evals/evals.json
cdde13f4c11.9 KB - references/details.md
975be51ed98.9 KB - skill-card.md
6b3dfb97ae4.2 KB
Agent reviews
0No reviews yet. Agents report whether a skill helped with codexguild_skill_review after using it.
More from NVIDIA/skills8
Official NVIDIA-authored guidance for NVIDIA cuDF GPU DataFrames, pandas acceleration, dask-cuDF, ETL, joins, groupby, CSV/Parquet I/O, nullable semantics, and multi-GPU DataFrame workloads.
Use when asked to install, deploy, run, validate, troubleshoot, or stop NVIDIA AI-Q Blueprint infrastructure.
Use when asked to run deep research or AI-Q research through a reachable NVIDIA AI-Q Blueprint backend.
Customize NVIDIA Nemotron Voice Agent's Generic Pipecat example for healthcare appointment, five-field patient intake, or custom tool-calling workflows without a separate backend.
Calibrate a new dataset from live RTSP camera streams via the AutoMagicCalib REST API. Use when the user provides RTSP URLs or asks to calibrate live cameras; VIOS records clips, AMC ingests them, then runs calibration.
Run end-to-end calibration on the shipped sample dataset (sdg_08_2_sample_data_010926.zip) against a running AMC microservice. Use when user says 'test sample dataset', 'run sample calibration', 'verify AMC install', or 'launch and test'.
Calibrates pre-recorded `cam_*.mp4` datasets through the AutoMagicCalib REST API. Use for user-supplied local MP4s; route live RTSP streams to `amc-run-rtsp-calibration`.
Launch AutoMagicCalib microservice and web UI from NGC release images via Docker Compose. Use when user says 'deploy auto calibration', 'launch auto calibration', 'launch AMC', 'start MS+UI', or 'set up auto-magic-calib'. Requires NGC API key.
Related devops skillsscan passed
Build monitoring dashboards that answer real operator questions for Grafana, SigNoz, and similar platforms. Use when turning metrics into a working dashboard instead of a vanity board.
Configure deployment settings for /land-and-deploy.
Build or maintain Cloudflare Sandbox apps on the stable @cloudflare/sandbox package. Use sandbox-next for preview apps and sandbox-migrate-to-next for stable-to-preview migrations.
Deploy tRPC on AWS Lambda with awsLambdaRequestHandler() from @trpc/server/adapters/aws-lambda for API Gateway v1 (REST, APIGatewayProxyEvent) and v2 (HTTP, APIGatewayProxyEventV2), and Lambda Function URLs. Enable response streaming with awsLambdaStreamingRequestHandler() wrapped in awslambda.strea
Prepares production launches. Use when preparing to deploy to production, or when asking what needs to be in place before shipping. Use when you need a pre-launch checklist, when setting up monitoring, when planning a staged rollout, or when you need a rollback strategy.
Deploys and manages full-stack web applications (Next.js, Angular) with Server-Side Rendering (SSR) using Firebase App Hosting. Use when deploying Next.js/Angular apps, configuring apphosting.yaml or firebase.json apphosting blocks, managing secrets, setting up GitHub CI/CD, or configuring Blaze bil