hanxiao commited on
Commit
27f7ee4
·
verified ·
1 Parent(s): 8f5391d

add Elastic Inference Service usage

Browse files
Files changed (1) hide show
  1. README.md +19 -0
README.md CHANGED
@@ -53,6 +53,25 @@ GGUF quantizations of [jina-embeddings-v5-text-nano-clustering](https://huggingf
53
 
54
  ## Usage with llama.cpp
55
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
56
  ```bash
57
  # Build llama.cpp (upstream)
58
  git clone https://github.com/ggml-org/llama.cpp
 
53
 
54
  ## Usage with llama.cpp
55
 
56
+ <details open>
57
+ <summary>via <a href="https://www.elastic.co/docs/explore-analyze/elastic-inference/eis">Elastic Inference Service</a></summary>
58
+
59
+ The fastest way to use v5-text in production. Elastic Inference Service (EIS) provides managed embedding inference with built-in scaling, so you can generate embeddings directly within your Elastic deployment.
60
+
61
+ ```bash
62
+ PUT _inference/text_embedding/jina-v5
63
+ {
64
+ "service": "elastic",
65
+ "service_settings": {
66
+ "model_id": "jina-embeddings-v5-text-nano"
67
+ }
68
+ }
69
+ ```
70
+
71
+ See the [Elastic Inference Service documentation](https://www.elastic.co/docs/explore-analyze/elastic-inference/eis) for setup details.
72
+
73
+ </details>
74
+
75
  ```bash
76
  # Build llama.cpp (upstream)
77
  git clone https://github.com/ggml-org/llama.cpp