DeepInfra MCP by usefulapi
Run DeepInfra inference, list models and read account rate limits.
- 1.11.1
- Version
- remote
- Transport
- 18
- Tools
Security review
Partly reviewedReviewed 27m ago.
- tools: 18 tools scanned
- metadata: scanned
- mediumReviewRemote tools take credentials as input
Whatever an agent passes to a remote tool leaves the machine. Never send connection strings, tokens or passwords to a third-party MCP server unless it is the service those credentials belong to.
deepinfra_list_api_tokens
Tools (18)
deepinfra_get_account
Get the current DeepInfra account / user details (identity, email, quotas). DeepInfra REST: GET /v1/me.
deepinfra_get_rate_limit
Get the account's current rate limits (per-model / per-endpoint request and token limits). DeepInfra REST: GET /v1/me/rate_limit.
deepinfra_list_api_tokens
List the account's API tokens (metadata only, not secret values). DeepInfra REST: GET /v1/api-tokens.
deepinfra_get_usage
Get spend / usage for a billing period. DeepInfra REST: GET /payment/usage.
deepinfra_get_usage_tokens
Get per-model token usage for a period. DeepInfra REST: GET /payment/usage/tokens.
deepinfra_get_usage_rent
Get GPU rental (dedicated hardware) usage for a time range. DeepInfra REST: GET /payment/usage/rent.
deepinfra_list_invoices
List billing invoices for the account. DeepInfra REST: GET /payment/invoices.
deepinfra_list_deployments
List the account's dedicated deployments. DeepInfra REST: GET /deploy/list.
deepinfra_get_deployment
Get one dedicated deployment by id (config + current status). DeepInfra REST: GET /deploy/{deploy_id}.
deepinfra_get_deployment_stats
Get time-series stats (throughput / latency / replicas) for a dedicated deployment. DeepInfra REST: GET /deploy/{deploy_id}/stats2.
deepinfra_get_gpu_availability
Get GPU availability for LLM deployments (which hardware can currently be provisioned). DeepInfra REST: GET /deploy/llm/gpu_availability.
deepinfra_list_models
List the DeepInfra model catalog (all available models). DeepInfra REST: GET /models/list.
deepinfra_get_model
Get one model's catalog entry (pricing, context length, capabilities). DeepInfra REST: GET /models/{model_name}.
deepinfra_get_hardware
Get the hardware options / GPU configuration available for a given model. DeepInfra REST: GET /v2/hardware.
deepinfra_query_logs
Query inference request logs for a dedicated deployment over a time window. DeepInfra REST: GET /v1/logs/query.
deepinfra_get_live_metrics
Get global live inference metrics across the account (real-time throughput / activity). DeepInfra REST: GET /v1/metrics/live.
deepinfra_start_deployment
Start (resume) a dedicated deployment. WARNING: this resumes a dedicated deployment and may incur GPU charges while it runs. DeepInfra REST: POST /deploy/{deploy_id}/start.
deepinfra_stop_deployment
Stop (pause) a running dedicated deployment. This halts inference and stops accruing GPU charges for it. DeepInfra REST: POST /deploy/{deploy_id}/stop.