Agent Skills that guide AI coding agents in operating and troubleshooting self-hosted Arize AX.
Works with Cursor, Claude Code, Codex, GitHub Copilot, Windsurf, and 40+ other agents.
Troubleshoot an existing Arize self-hosted deployment - give your coding agent this prompt:
Install Arize self-hosted skills from https://github.com/Arize-ai/ax-self-hosted-skills and use the arize-alerts-troubleshoot skill to diagnose the alerts currently firing on my cluster.
Install all skills non-interactively — this is what the agent runs for the prompt above:
npx skills add Arize-ai/ax-self-hosted-skills --skill "*" --yesWant to hand-pick skills, agent, and scope yourself? Drop the flags for the interactive wizard:
npx skills add Arize-ai/ax-self-hosted-skillsBoth auto-detect your agent (Cursor, Claude Code, Codex, etc.) and symlink skills into place.
macOS / Linux:
git clone https://github.com/Arize-ai/ax-self-hosted-skills.git
cd ax-self-hosted-skills
./install.sh --project ~/my-projectWindows (PowerShell):
git clone https://github.com/Arize-ai/ax-self-hosted-skills.git
cd ax-self-hosted-skills
.\install.ps1 -Project ~\my-projectThe installer only symlinks (or copies) skills into your agent's skills directory. It does not install Arize AX or the ax CLI. Use --global / -Global instead to install to ~/.<agent>/skills/.
Required on the machine where the agent runs the skill scripts:
| Tool | Why |
|---|---|
kubectl |
Read-only cluster access via safe-kubectl.sh (get, describe, logs, port-forward, …) |
curl |
HTTP GETs against Prometheus / Alertmanager |
jq |
JSON parsing in the shell helpers |
python3 |
catalog-lookup.py, docs-search.py, distribution.py |
tar |
Read Chart.yaml from arize-operator-chart.tgz during version check |
Also required for a useful investigation session:
kubectlcontext pointed at the self-hosted Arize cluster- An unpacked Arize distribution that matches the cluster release (contains
arize.shanddocs/) - Network reachability to Prometheus / Alertmanager (ingress URL or local port-forward)
Set these in the shell before (or while) running the skill:
| Variable | Required | Purpose |
|---|---|---|
ARIZE_DISTRIBUTION_ROOT |
Yes (or pass --distribution-root) |
Path to the unpacked distribution for this cluster. Alias: ARIZE_DIST. Must contain arize.sh and docs/troubleshooting/selfhosted-alerts-table.csv. |
ARIZE_NAMESPACE |
Recommended | Application namespace where Prometheus / Alertmanager run (used by open-ports.sh) |
OPERATOR_NS |
Optional | Operator namespace for version checks (default: arize-operator) |
KUBE_CONTEXT |
Optional | kube context for safe-kubectl.sh / preflight.sh if not current-context |
PROM / PROM_URL |
When querying Prometheus | Prometheus base URL, e.g. http://localhost:9090/prometheus |
AM / AM_URL |
When querying Alertmanager | Alertmanager base URL, e.g. http://localhost:9093/alertmanager |
ARIZE_SKILL_TMP |
Optional | Scratch dir for PID files / JSON dumps (default: /tmp/arize-alerts-troubleshoot) |
CURL_INSECURE |
Optional | Set to 1 to force curl -k for non-localhost HTTPS (or pass --insecure) |
Example:
export ARIZE_DISTRIBUTION_ROOT="/path/to/unpacked/arize-distribution"
export ARIZE_NAMESPACE="arize"
export OPERATOR_NS="arize-operator"
export PROM="http://localhost:9090/prometheus"
export AM="http://localhost:9093/alertmanager"| Skill | Path | Purpose |
|---|---|---|
arize-alerts-troubleshoot |
skills/arize-alerts-troubleshoot/ |
Diagnose firing self-hosted alerts using read-only kubectl / Prometheus / Alertmanager access and the local distribution docs |
Each skill lives under skills/<name>/ with a SKILL.md entrypoint, optional
references/, and helper scripts/.