Class PutNvidiaRequest

java.lang.Object
co.elastic.clients.elasticsearch._types.RequestBase
co.elastic.clients.elasticsearch.inference.PutNvidiaRequest
All Implemented Interfaces:
JsonpSerializable

@JsonpDeserializable public class PutNvidiaRequest extends RequestBase implements JsonpSerializable
Create an Nvidia inference endpoint.

Create an inference endpoint to perform an inference task with the nvidia service.

See Also:
  • Field Details

  • Method Details

    • of

    • chunkingSettings

      @Nullable public final InferenceChunkingSettings chunkingSettings()
      The chunking configuration object. Applies only to the text_embedding task type. Not applicable to the rerank, completion, or chat_completion task types.

      API name: chunking_settings

    • nvidiaInferenceId

      public final String nvidiaInferenceId()
      Required - The unique identifier of the inference endpoint.

      API name: nvidia_inference_id

    • service

      public final NvidiaServiceType service()
      Required - The type of service supported for the specified task type. In this case, nvidia.

      API name: service

    • serviceSettings

      public final NvidiaServiceSettings serviceSettings()
      Required - Settings used to install the inference model. These settings are specific to the nvidia service.

      API name: service_settings

    • taskSettings

      @Nullable public final NvidiaTaskSettings taskSettings()
      Settings to configure the inference task. Applies only to the text_embedding task type. Not applicable to the rerank, completion, or chat_completion task types. These settings are specific to the task type you specified.

      API name: task_settings

    • taskType

      public final NvidiaTaskType taskType()
      Required - The type of the inference task that the model will perform. NOTE: The chat_completion task type only supports streaming and only through the _stream API.

      API name: task_type

    • timeout

      @Nullable public final Time timeout()
      Specifies the amount of time to wait for the inference endpoint to be created.

      API name: timeout

    • serialize

      public void serialize(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper)
      Serialize this object to JSON.
      Specified by:
      serialize in interface JsonpSerializable
    • serializeInternal

      protected void serializeInternal(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper)
    • rebuild

      public PutNvidiaRequest.Builder rebuild()
      Returns:
      New PutNvidiaRequest.Builder initialized with field values of this instance
    • setupPutNvidiaRequestDeserializer

      protected static void setupPutNvidiaRequestDeserializer(ObjectDeserializer<PutNvidiaRequest.Builder> op)