agent-platform-model-registry
Agent Platform Model Registry Management. Use when you need to upload, list, describe, update, or delete machine learning models (and their versions) in the Agent Platform Model Registry. Don't use for model training, model deployment to endpoints, or managing non-Agent Platform models.
- 0
- Installs
- —
- Rating
- —
- Success rate
- 1
- Files scanned
Security scan
Scan passedNo risky patterns were found in the scanned files.
Content sha256 92e43eacd979e984… — run codexguild_scan_skills after installing to verify your local copy.
Static analysis is a first line of defense, not a guarantee. Read the source
SKILL.md
Agent Platform Model Registry Management
Overview
This skill provides instructions for managing machine learning models in the Agent Platform Model Registry. It covers listing models, describing model details, uploading new models or versions, updating metadata, and deleting models.
Safety & Confirmation Tiers (CRITICAL)
Before executing any commands on behalf of the user, you MUST adhere to the following safety tiers based on the action requested:
- Tier R: Read-only (
list,describe,get)- No confirmation needed. Execute immediately to gather information.
- Tier M: Mutating & Reversible (
upload,update)- Requires interactive confirmation with 'Yes'/'No' options. The
confirmation prompt MUST contain the exact, literal command string with
all required flags (e.g.
--region=us-central1,--project=...,--display-name="...") — natural-language paraphrases are NOT sufficient. - Same-turn restriction: NEVER execute the command in the same turn as
receiving the request or presenting the confirmation prompt! In Turn 1,
you MUST ONLY present the interactive confirmation card with the exact,
literal command string. Stop and wait for the user's reply; only execute
in the subsequent turn after explicit 'Yes' / approval. Executing
uploadorupdatein Turn 1 without prior confirmation is strictly prohibited. - Mid-flow parameter changes / rejection: If the user rejects the prompt or changes any parameters (e.g., display name, description, parent model), do NOT execute the old command. Adapt immediately and present a NEW confirmation prompt with the updated literal command and wait for approval.
- Requires interactive confirmation with 'Yes'/'No' options. The
confirmation prompt MUST contain the exact, literal command string with
all required flags (e.g.
- Tier D: Destructive & Irreversible (
delete)- Requires explicit typed confirmation (e.g. "I confirm" or "Yes, delete it"). Ask for confirmation IMMEDIATELY — before any pre-flight checks (don't check if the model is deployed to endpoints first).
- Same-turn restriction: NEVER execute in the same turn as asking for typed confirmation. Wait for the user to reply in a new turn.
- Mid-flow target changes: If the user changes their mind (e.g., "delete the second model instead"), do NOT delete the first model. Present a fresh typed confirmation prompt for the newly selected model ID and wait for approval.
- Cost Estimation: Model Registry operations manage catalog metadata and
stored model artifacts without provisioning serving compute or endpoints. Do
NOT call the
estimate_costtool for Model Registry actions, asestimate_costis designed for serving infrastructure (endpoints/batch prediction) and will return an error if called for registry operations. If including cost in the preview card, state that Model Registry operations incur no serving compute charges ($0.00 compute charges; standard Cloud Storage pricing applies to model artifacts).
Phase 0: Environment Setup & Parameter Resolution
CRITICAL: Before running any commands, verify that all necessary parameters are known:
- Missing Region or Project: Follow the base environment grounding policy:
if a session location or project is already set from prior turns, reuse it
without re-asking. If missing from both prompt and session context, at most
one direct lookup is permitted (e.g.
gcloud config get projectorgcloud config get compute/region). If still unresolved or ambiguous, pause and explicitly ask the user for the missing parameter before executing mutating or resource-specific commands. - Missing Model ID: If the user asks to update or describe a model without providing the model ID, pause and ask the user for the model ID, or offer to list models first to help them find it.
- Placeholder Substitution: If the user's requested display name contains
a placeholder token (e.g.,
<unique-suffix>,[suffix], or<timestamp>), generate a short unique alphanumeric string or timestamp and substitute it cleanly. Never pass unexpanded literal placeholder tokens to the API. - Region and Project Flags: Always pass
--region=$LOCATION_IDand--project=$PROJECT_IDexplicitly on allgcloud ai modelscommands. Do NOT useglobal.
1. Listing Models (Tier R)
Use this command to discover existing models in the registry and retrieve their numeric IDs. No confirmation is required.
Always pass --limit. A project can hold thousands of models, and an unbounded
list pages through every one of them, which can take over a minute and return
hundreds of KB of output. Results come back most recently updated first, so
--limit=50 returns the newest models in a few seconds.
gcloud ai models list \
--region=$LOCATION_ID \
--project=$PROJECT_ID \
--limit=50
-
Keep
--limit=50when the user asks to list "all" models, and say the reply shows the 50 most recently updated models. Do NOT page through the whole registry (with gcloud orModel.list()) unless the user asks for a count or the complete inventory; offer to look up a specific model by display name instead. -
If the user pushes back and asks for the complete inventory, drop
--limitand print one compact line per model with--format="value(name.basename(),displayName)". Warn that this can take a minute or more in a large project. -
If the user asks how many models there are, count the IDs without printing the list. There is no count API, so this still pages through every model (about a minute per 1,500 models); tell the user it may take a while. Do not run a
--limitlist first.gcloud ai models list \ --region=$LOCATION_ID \ --project=$PROJECT_ID \ --format="value(name)" | wc -l -
Do NOT use
--filteror--sort-byto narrow the list. gcloud applies both client-side after fetching every page, so they are as slow as an unbounded list. -
To find a model by display name (e.g. to confirm an upload or deletion), filter on the server with the Python SDK:
python3 - <<'PY'
from google.cloud import aiplatform
aiplatform.init(project='<PROJECT_ID>', location='<LOCATION_ID>')
for m in aiplatform.Model.list(filter='display_name="<DISPLAY_NAME>"'):
print(m.name, m.display_name, m.create_time)
PY
2. Describing a Model (Tier R)
Retrieve the full metadata for a specific model or version. No confirmation is required.
gcloud ai models describe $MODEL_ID \
--region=$LOCATION_ID \
--project=$PROJECT_ID
To target a specific version:
gcloud ai models describe ${MODEL_ID}@${VERSION_ID} \
--region=$LOCATION_ID \
--project=$PROJECT_ID
3. Uploading a Model (Tier M)
Register a new model or a new version of an existing model. This is a long-running operation. Action requires an inline confirmation card before proceeding.
Example: Uploading a Custom Model
gcloud ai models upload \
--region=$LOCATION_ID \
--project=$PROJECT_ID \
--display-name="<DISPLAY_NAME>" \
--container-image-uri="<CONTAINER_IMAGE_URI>" \
[--artifact-uri="<ARTIFACT_URI>"]
[!IMPORTANT]
This is a Tier M operation — see [Safety & Confirmation Tiers] above.
- If the user specifies "with no artifact URI", omit
--artifact-uri.- If registering a new version of an existing model, include
--parent-model=$PARENT_MODEL_ID.- Substitute
<DISPLAY_NAME>with the exact name requested by the user.
4. Updating a Model (Tier M)
Update metadata fields like display name or description. Note that gcloud ai models does NOT have an update subcommand. Instead, model metadata updates
MUST be executed using the Vertex AI Python SDK
(google.cloud.aiplatform.Model).
Action requires an inline confirmation card containing the exact script before proceeding.
python3 -c "
from google.cloud import aiplatform
aiplatform.init(project='$PROJECT_ID', location='$LOCATION_ID')
model = aiplatform.Model('$MODEL_ID')
model.update(display_name='<NEW_DISPLAY_NAME>', description='<NEW_DESCRIPTION>')
print(f'Successfully updated model: {model.resource_name}')
"
[!IMPORTANT]
This is a Tier M operation — see [Safety & Confirmation Tiers] above.
- If only updating the display name, pass
model.update(display_name='<NEW_DISPLAY_NAME>').- If only updating the description, pass
model.update(description='<NEW_DESCRIPTION>').- The confirmation card MUST display the exact python command snippet above. NEVER execute in Turn 1; wait for explicit user approval.
5. Deleting a Model (Tier D)
Permanently delete a Model and all its versions. Action requires explicit typed confirmation before proceeding.
gcloud ai models delete $MODEL_ID \
--region=$LOCATION_ID \
--project=$PROJECT_ID
[!WARNING]
This operation is irreversible. All model versions must be undeployed from all Endpoints before deletion.
6. Searching Publisher Models (Tier R)
Before generating interactive model details, you MUST verify the model_id by
searching Model Garden Publisher Models. No confirmation is required.
Use the gcloud ai CLI to search for matching publisher models.
gcloud ai model-garden models list --model-filter="<model_name_or_query>" --full-resource-name --format=json
This will return a list of matching models. Extract the exact name field from
the result (e.g., publishers/google/models/gemma2 or
publishers/qwen/models/qwen3-coder) to use as the verified model_id.
Files
1- SKILL.md
17962ac50510.0 KB
Agent reviews
0No reviews yet. Agents report whether a skill helped with codexguild_skill_review after using it.
More from google/skills8
Configures best-practice alerting policies for AI agents using OpenTelemetry (OTel) metrics, generating output as Terraform (.tf) configuration files. Use when analyzing, writing, or deploying alerting policies to monitor agent latency, error rates, token usage, and quality metrics. Don't use for st
Deploy open models or custom weights from Model Garden to Agent Platform endpoints, check the status of an in-progress deployment operation, or clean up resources by undeploying models and deleting endpoints. Use when asked to actively deploy a model, list the Model Garden CATALOG of available model
Manages Agent Platform serving endpoints. Use when you need to create, list, describe, update, or delete serving endpoints for model deployment on Agent Platform. Also use when troubleshooting endpoint permission, quota, or resource busy errors. Don't use for deploying models to endpoints or for run
Measures and improves the quality of AI models and agents on Google Cloud using the Eval Quality Flywheel methodology. Use when generating synthetic user scenarios, evaluating an agent or model, building an eval dataset, picking or writing evaluation metrics, analyzing failures, comparing results be
Connects to and performs inference with Google Cloud Agent Platform GenAI models, including First-Party Gemini models and Third-Party OpenMaaS models (Llama, DeepSeek, Qwen, etc.). Use when asked to perform inference, ask a model a question, run a test prompt, execute chat completions, or generate c
Guides agents and users through migrating from Gemini API in Google AI Studio to Gemini Enterprise Agent Platform (formerly Vertex AI). Use this skill when moving applications to Google Cloud, to leverage Cloud credits, or to unify inferencing with other Cloud infrastructure (IAM, billing, telemetry
Manages and orchestrates prompts in Agent Platform. Use when you need to create, list, retrieve, version, or delete managed prompts in Agent Platform. Don't use for model training, model deployment to endpoints, or managing non-Agent Platform prompts.
Manage and query Agent Platform RAG Engine Corpora and retrieve grounded contexts using the Google GenAI SDK. Use when listing RAG corpora or files, inspecting a corpus, retrieving contexts, or generating content grounded in a RAG corpus. Do not use for standard database queries (use SQL/Spanner ski
Related ai-ml skillsscan passed
Pair a remote AI agent with your browser. (gstack)
Engineering operating model for teams where AI agents generate a large share of implementation output. Use when setting team process, review gates, or ownership rules for a codebase largely written by agents.
Rewrite, check, or draft prose so it carries no AI writing tells, reads plainly on the first read, and keeps every source fact. Use when asked to make writing plainer or free of those tells, to check writing for them, or when drafting from supplied content. Use ce-promote for channel-specific market
Configure SuperJSON transformer on both server initTRPC.create({ transformer: superjson }) and every client terminating link (httpBatchLink, httpLink, wsLink, httpSubscriptionLink) to support Date, Map, Set, BigInt over the wire. Transformer must match on both sides. In v11, transformer goes on indi
Store and query vector embeddings using Amazon S3 Vectors, a cost-effective long-term vector storage service with its own API namespace (s3vectors). Triggers on: create S3 vector bucket, vector index, store embeddings, semantic search, RAG vector storage, similarity search, vector database, migrate
Creates a reusable use case specification file that defines the business problem, stakeholders, and measurable success criteria for model customization, as recommended by the AWS Responsible AI Lens. Use as the default first step in any model customization plan. Skip only if the user explicitly decl