Class PutNvidiaRequest
java.lang.Object
co.elastic.clients.elasticsearch._types.RequestBase
co.elastic.clients.elasticsearch.inference.PutNvidiaRequest
- All Implemented Interfaces:
JsonpSerializable
Create an Nvidia inference endpoint.
Create an inference endpoint to perform an inference task with the
nvidia service.
- See Also:
-
Nested Class Summary
Nested ClassesNested classes/interfaces inherited from class co.elastic.clients.elasticsearch._types.RequestBase
RequestBase.AbstractBuilder<BuilderT extends RequestBase.AbstractBuilder<BuilderT>> -
Field Summary
FieldsModifier and TypeFieldDescriptionstatic final JsonpDeserializer<PutNvidiaRequest>Json deserializer forPutNvidiaRequeststatic final Endpoint<PutNvidiaRequest,PutNvidiaResponse, ErrorResponse> Endpoint "inference.put_nvidia". -
Method Summary
Modifier and TypeMethodDescriptionThe chunking configuration object.final StringRequired - The unique identifier of the inference endpoint.static PutNvidiaRequestrebuild()voidserialize(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper) Serialize this object to JSON.protected voidserializeInternal(jakarta.json.stream.JsonGenerator generator, JsonpMapper mapper) final NvidiaServiceTypeservice()Required - The type of service supported for the specified task type.final NvidiaServiceSettingsRequired - Settings used to install the inference model.protected static voidfinal NvidiaTaskSettingsSettings to configure the inference task.final NvidiaTaskTypetaskType()Required - The type of the inference task that the model will perform.final Timetimeout()Specifies the amount of time to wait for the inference endpoint to be created.Methods inherited from class co.elastic.clients.elasticsearch._types.RequestBase
toString
-
Field Details
-
_DESERIALIZER
Json deserializer forPutNvidiaRequest -
_ENDPOINT
Endpoint "inference.put_nvidia".
-
-
Method Details
-
of
public static PutNvidiaRequest of(Function<PutNvidiaRequest.Builder, ObjectBuilder<PutNvidiaRequest>> fn) -
chunkingSettings
The chunking configuration object. Applies only to thetext_embeddingtask type. Not applicable to thererank,completion, orchat_completiontask types.API name:
chunking_settings -
nvidiaInferenceId
Required - The unique identifier of the inference endpoint.API name:
nvidia_inference_id -
service
Required - The type of service supported for the specified task type. In this case,nvidia.API name:
service -
serviceSettings
Required - Settings used to install the inference model. These settings are specific to thenvidiaservice.API name:
service_settings -
taskSettings
Settings to configure the inference task. Applies only to thetext_embeddingtask type. Not applicable to thererank,completion, orchat_completiontask types. These settings are specific to the task type you specified.API name:
task_settings -
taskType
Required - The type of the inference task that the model will perform. NOTE: Thechat_completiontask type only supports streaming and only through the _stream API.API name:
task_type -
timeout
Specifies the amount of time to wait for the inference endpoint to be created.API name:
timeout -
serialize
Serialize this object to JSON.- Specified by:
serializein interfaceJsonpSerializable
-
serializeInternal
-
rebuild
- Returns:
- New
PutNvidiaRequest.Builderinitialized with field values of this instance
-
setupPutNvidiaRequestDeserializer
protected static void setupPutNvidiaRequestDeserializer(ObjectDeserializer<PutNvidiaRequest.Builder> op)
-