io.usefulapi/deepinfra

DeepInfra MCP by usefulapi

Run DeepInfra inference, list models and read account rate limits.

1.11.1
Version
remote
Transport
18
Tools

Security review

Partly reviewed

Reviewed 27m ago.

  • tools: 18 tools scanned
  • metadata: scanned
  • mediumReviewRemote tools take credentials as input

    Whatever an agent passes to a remote tool leaves the machine. Never send connection strings, tokens or passwords to a third-party MCP server unless it is the service those credentials belong to.

    deepinfra_list_api_tokens

Tools (18)

  • deepinfra_get_account

    Get the current DeepInfra account / user details (identity, email, quotas). DeepInfra REST: GET /v1/me.

  • deepinfra_get_rate_limit

    Get the account's current rate limits (per-model / per-endpoint request and token limits). DeepInfra REST: GET /v1/me/rate_limit.

  • deepinfra_list_api_tokens

    List the account's API tokens (metadata only, not secret values). DeepInfra REST: GET /v1/api-tokens.

  • deepinfra_get_usage

    Get spend / usage for a billing period. DeepInfra REST: GET /payment/usage.

  • deepinfra_get_usage_tokens

    Get per-model token usage for a period. DeepInfra REST: GET /payment/usage/tokens.

  • deepinfra_get_usage_rent

    Get GPU rental (dedicated hardware) usage for a time range. DeepInfra REST: GET /payment/usage/rent.

  • deepinfra_list_invoices

    List billing invoices for the account. DeepInfra REST: GET /payment/invoices.

  • deepinfra_list_deployments

    List the account's dedicated deployments. DeepInfra REST: GET /deploy/list.

  • deepinfra_get_deployment

    Get one dedicated deployment by id (config + current status). DeepInfra REST: GET /deploy/{deploy_id}.

  • deepinfra_get_deployment_stats

    Get time-series stats (throughput / latency / replicas) for a dedicated deployment. DeepInfra REST: GET /deploy/{deploy_id}/stats2.

  • deepinfra_get_gpu_availability

    Get GPU availability for LLM deployments (which hardware can currently be provisioned). DeepInfra REST: GET /deploy/llm/gpu_availability.

  • deepinfra_list_models

    List the DeepInfra model catalog (all available models). DeepInfra REST: GET /models/list.

  • deepinfra_get_model

    Get one model's catalog entry (pricing, context length, capabilities). DeepInfra REST: GET /models/{model_name}.

  • deepinfra_get_hardware

    Get the hardware options / GPU configuration available for a given model. DeepInfra REST: GET /v2/hardware.

  • deepinfra_query_logs

    Query inference request logs for a dedicated deployment over a time window. DeepInfra REST: GET /v1/logs/query.

  • deepinfra_get_live_metrics

    Get global live inference metrics across the account (real-time throughput / activity). DeepInfra REST: GET /v1/metrics/live.

  • deepinfra_start_deployment

    Start (resume) a dedicated deployment. WARNING: this resumes a dedicated deployment and may incur GPU charges while it runs. DeepInfra REST: POST /deploy/{deploy_id}/start.

  • deepinfra_stop_deployment

    Stop (pause) a running dedicated deployment. This halts inference and stops accruing GPU charges for it. DeepInfra REST: POST /deploy/{deploy_id}/stop.