﻿---
title: stack es inference put-googlevertexai cli command
description: Create a Google Vertex AI inference endpoint. Behaviour flags: --dry-run — validate all inputs and exit without performing any action 
url: https://www.elastic.co/elastic/docs-builder/docs/4097/reference/elastic-cli/cli/stack/es/inference/put-googlevertexai
applies_to:
  - Elastic Cloud Serverless: Preview
  - Elastic Stack: Preview
---

# stack es inference put-googlevertexai cli command
<cli-modifiers>
</cli-modifiers>

```bash
elastic stack es inference put-googlevertexai \
  --service <service> \
  --service-settings <service-settings> \
  --task-type <task-type> \
  --googlevertexai-inference-id <googlevertexai-inference-id> \
  [options]
```

Create a Google Vertex AI inference endpoint.
**Behaviour flags:**
`--dry-run` — validate all inputs and exit without performing any action

## Options

<definitions>
  <definition term="--service enum required">
    The type of service supported for the specified task type. In this case, `googlevertexai`.
    **Values:** googlevertexai
  </definition>
  <definition term="--service-settings string required">
    Settings used to install the inference model. These settings are specific to the `googlevertexai` service.
  </definition>
  <definition term="--task-type enum required">
    The type of the inference task that the model will perform.
    **Values:** rerank, text_embedding, completion, chat_completion
  </definition>
  <definition term="--googlevertexai-inference-id string required">
    The unique identifier of the inference endpoint.
  </definition>
  <definition term="--timeout string">
    Specifies the amount of time to wait for the inference endpoint to be created.
  </definition>
  <definition term="--chunking-settings string">
    The chunking configuration object.
    Applies only to the `text_embedding` task type.
    Not applicable to the `rerank`, `completion`, or `chat_completion` task types.
  </definition>
  <definition term="--task-settings string">
    Settings to configure the inference task.
    These settings are specific to the task type you specified.
  </definition>
  <definition term="--error-trace">
    When set to `true` Elasticsearch will include the full stack trace of errors
    when they occur.
  </definition>
  <definition term="--filter-path string">
    Comma-separated list of filters in dot notation which reduce the response
    returned by Elasticsearch.
    **Repeatable:** pass `--filter-path` multiple times to supply more than one value
  </definition>
  <definition term="--human">
    When set to `true` will return statistics in a format suitable for humans.
    For example `"exists_time": "1h"` for humans and
    `"exists_time_in_millis": 3600000` for computers. When disabled the human
    readable values will be omitted. This makes sense for responses being consumed
    only by machines.
  </definition>
  <definition term="--pretty">
    If set to `true` the returned JSON will be "pretty-formatted". Only use
    this option for debugging only.
  </definition>
  <definition term="--input-file string">
    path to a JSON file to use as command input
  </definition>
  <definition term="--dry-run">
    validate all inputs and exit without performing any action (preview changes without applying them)
  </definition>
</definitions>


## Global Options

<definitions>
  <definition term="--json">
    output as JSON
  </definition>
</definitions>