co.elastic.clients.elasticsearch._types.RequestBase

co.elastic.clients.elasticsearch.inference.ChatCompletionUnifiedRequest

All Implemented Interfaces:: JsonpSerializable

@JsonpDeserializable public class ChatCompletionUnifiedRequest extends RequestBase implements JsonpSerializable

Perform chat completion inference

The chat completion inference API enables real-time responses for chat completion tasks by delivering answers incrementally, reducing response times during computation. It only works with the chat_completion task type for openai and elastic inference services.

IMPORTANT: The inference APIs enable you to use certain services, such as built-in machine learning models (ELSER, E5), models uploaded through Eland, Cohere, OpenAI, Azure, Google AI Studio, Google Vertex AI, Anthropic, Watsonx.ai, or Hugging Face. For built-in models and models uploaded through Eland, the inference APIs offer an alternative way to use and manage trained models. However, if you do not plan to use the inference APIs to use these models or if you want to use non-NLP models, use the machine learning trained model APIs.

NOTE: The chat_completion task type is only available within the _stream API and only supports streaming. The Chat completion inference API and the Stream inference API differ in their response structure and capabilities. The Chat completion inference API provides more comprehensive customization options through more fields and function calling support. If you use the openai service or the elastic service, use the Chat completion inference API.

See Also:

API specification

Nested Class Summary

Nested Classes

Modifier and Type

Class

Description

static class

ChatCompletionUnifiedRequest.Builder

Builder for ChatCompletionUnifiedRequest.

Nested classes/interfaces inherited from class co.elastic.clients.elasticsearch._types.RequestBase
RequestBase.AbstractBuilder<BuilderT extends RequestBase.AbstractBuilder<BuilderT>>
Field Summary

Fields

Modifier and Type

Field

Description

static final JsonpDeserializer<ChatCompletionUnifiedRequest>

_DESERIALIZER

static final Endpoint<ChatCompletionUnifiedRequest,BinaryResponse,ErrorResponse>

_ENDPOINT

Endpoint "inference.chat_completion_unified".
Method Summary

Modifier and Type

Method

Description

final RequestChatCompletion

chatCompletionRequest()

Required - Request body.

protected static JsonpDeserializer<ChatCompletionUnifiedRequest>

createChatCompletionUnifiedRequestDeserializer()

final String

inferenceId()

Required - The inference Id

static ChatCompletionUnifiedRequest

of(Function<ChatCompletionUnifiedRequest.Builder,ObjectBuilder<ChatCompletionUnifiedRequest>> fn)

void

serialize(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper)

Serialize this value to JSON.

final Time

timeout()

Specifies the amount of time to wait for the inference request to complete.

Methods inherited from class co.elastic.clients.elasticsearch._types.RequestBase
toString

Methods inherited from class java.lang.Object
clone, equals, finalize, getClass, hashCode, notify, notifyAll, wait, wait, wait

Field Details
- _DESERIALIZER
  
  public static final JsonpDeserializer<ChatCompletionUnifiedRequest> _DESERIALIZER
- _ENDPOINT
  
  public static final Endpoint<ChatCompletionUnifiedRequest,BinaryResponse,ErrorResponse> _ENDPOINT
  
  Endpoint "inference.chat_completion_unified".
Method Details
- of
  
  public static ChatCompletionUnifiedRequest of(Function<ChatCompletionUnifiedRequest.Builder,ObjectBuilder<ChatCompletionUnifiedRequest>> fn)
- inferenceId
  
  public final String inferenceId()
  
  Required - The inference Id
  API name: inference_id
- timeout
  
  @Nullable public final Time timeout()
  
  Specifies the amount of time to wait for the inference request to complete.
  API name: timeout
- chatCompletionRequest
  
  public final RequestChatCompletion chatCompletionRequest()
  
  Required - Request body.
- serialize
  
  public void serialize(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper)
  
  Serialize this value to JSON.
  
  Specified by:
  
  serialize in interface JsonpSerializable
- createChatCompletionUnifiedRequestDeserializer
  
  protected static JsonpDeserializer<ChatCompletionUnifiedRequest> createChatCompletionUnifiedRequestDeserializer()

Class ChatCompletionUnifiedRequest

Nested Class Summary

Nested classes/interfaces inherited from class co.elastic.clients.elasticsearch._types.RequestBase

Field Summary

Method Summary

Methods inherited from class co.elastic.clients.elasticsearch._types.RequestBase

Methods inherited from class java.lang.Object

Field Details

_DESERIALIZER

_ENDPOINT

Method Details

of

inferenceId

timeout

chatCompletionRequest

serialize

createChatCompletionUnifiedRequestDeserializer