﻿---
title: stack es inference put-azureopenai cli command
description: Create an Azure OpenAI inference endpoint. Behaviour flags: --dry-run — validate all inputs and exit without performing any action 
url: https://docs-v3-preview.elastic.dev/elastic/cli/pull/573/cli/stack/es/inference/put-azureopenai
---

# stack es inference put-azureopenai cli command
<cli-modifiers>
</cli-modifiers>

```bash
elastic stack es inference put-azureopenai \
  --service <service> \
  --service-settings <service-settings> \
  --task-type <task-type> \
  --azureopenai-inference-id <azureopenai-inference-id> \
  [options]
```

Create an Azure OpenAI inference endpoint.
**Behaviour flags:**
`--dry-run` — validate all inputs and exit without performing any action

## Options

<definitions>
  <definition term="--service enum required">
    The type of service supported for the specified task type. In this case, `azureopenai`.
    **Values:** azureopenai
  </definition>
  <definition term="--service-settings string required">
    Settings used to install the inference model. These settings are specific to the `azureopenai` service.
  </definition>
  <definition term="--task-type enum required">
    The type of the inference task that the model will perform.
    NOTE: The `chat_completion` task type only supports streaming and only through the _stream API.
    **Values:** completion, chat_completion, text_embedding
  </definition>
  <definition term="--azureopenai-inference-id string required">
    The unique identifier of the inference endpoint.
  </definition>
  <definition term="--timeout string">
    Specifies the amount of time to wait for the inference endpoint to be created.
  </definition>
  <definition term="--chunking-settings string">
    The chunking configuration object.
    Applies only to the `text_embedding` task type.
    Not applicable to the `completion` and `chat_completion` task types.
  </definition>
  <definition term="--task-settings string">
    Settings to configure the inference task.
    These settings are specific to the task type you specified.
  </definition>
  <definition term="--error-trace">
    When set to `true` Elasticsearch will include the full stack trace of errors
    when they occur.
  </definition>
  <definition term="--filter-path string">
    Comma-separated list of filters in dot notation which reduce the response
    returned by Elasticsearch.
    **Repeatable:** pass `--filter-path` multiple times to supply more than one value
  </definition>
  <definition term="--human">
    When set to `true` will return statistics in a format suitable for humans.
    For example `"exists_time": "1h"` for humans and
    `"exists_time_in_millis": 3600000` for computers. When disabled the human
    readable values will be omitted. This makes sense for responses being consumed
    only by machines.
  </definition>
  <definition term="--pretty">
    If set to `true` the returned JSON will be "pretty-formatted". Only use
    this option for debugging only.
  </definition>
  <definition term="--input-file string">
    path to a JSON file to use as command input
  </definition>
  <definition term="--dry-run">
    validate all inputs and exit without performing any action (preview changes without applying them)
  </definition>
</definitions>


## Global Options

<definitions>
  <definition term="--json">
    output as JSON
  </definition>
</definitions>