Google 的向量模型,用于将输入内容转换为数值向量,支持语义检索和相似度计算。
SentencePiece (Gemini)/v1beta/models/gemini-embedding-001:generateContent| Parameter | Type | Default / range | Description |
|---|---|---|---|
inputrequired | string | — | Text or array of texts to embed |
dimensions | integer | >= 1 | Truncate embeddings to this many dimensions |
encoding_format | enum | = float | Wire encoding for the embedding vectors |
user | string | — | End-user identifier for abuse monitoring |
Replace <YOUR_API_KEY> with the API key from your token settings.
All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.
Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.
| Parameter | Type | Default / range | Description |
|---|---|---|---|
inputrequired | string | — | Text or array of texts to embed |
dimensions | integer | >= 1 | Truncate embeddings to this many dimensions |
encoding_format | enum | = float | Wire encoding for the embedding vectors |
user | string | — | End-user identifier for abuse monitoring |
| Supplier | RPM | TPM | RPD |
|---|---|---|---|
| default | 5.4K | 1.1M | 107K |
| Google AI Studio | 3.4K | 685K | 69K |
| Google Vertex | 3.9K | 772K | 77K |
RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply per token group.
Everything you need to decide before creating an account.
