Embeddings
Provider-first embedding architecture — auto-discovered from your model providers.
AgentOven uses a provider-first embedding architecture: when you configure a model provider (OpenAI, Azure OpenAI, Ollama), its embedding capabilities are automatically discovered and made available. No separate embedding configuration required.
How Provider-First Embeddings Work
Instead of requiring separate API keys for embeddings, AgentOven derives embedding capabilities from your existing model providers. If a provider's driver implements the EmbeddingCapableDriver interface, its embedding models become available instantly.
- Register a model provider (e.g., OpenAI with your API key)
- Server auto-discovers embedding capabilities from the driver
- Default embedding model is registered in the embedding registry
- RAG pipeline, vector stores, and ingestion use it automatically
- All available embedding models are discoverable via API
Supported Embedding Models
Each provider exposes different embedding models with varying dimensions and batch sizes:
OpenAI
| Model | Dimensions | Max Batch |
|---|---|---|
text-embedding-3-small | 1536 | 2048 |
text-embedding-3-large | 3072 | 2048 |
text-embedding-ada-002 | 1536 | 2048 |
Azure OpenAI
| Model | Dimensions | Max Batch |
|---|---|---|
text-embedding-3-small | 1536 | 2048 |
text-embedding-3-large | 3072 | 2048 |
text-embedding-ada-002 | 1536 | 2048 |
Ollama
| Model | Dimensions | Max Batch |
|---|---|---|
nomic-embed-text | 768 | 512 |
mxbai-embed-large | 1024 | 512 |
all-minilm | 384 | 512 |
snowflake-arctic-embed | 1024 | 512 |
Configuring via Provider Registration
Register a provider through the API, and embeddings are available immediately:
terminal
$ curl -X POST http://localhost:8080/api/v1/providers \ -H "Content-Type: application/json" \ -H "X-Kitchen: default" \ -d '{ "name": "my-openai", "kind": "openai", "config": { "api_key": "sk-..." }, "models": ["gpt-4o", "gpt-4o-mini"] }' # Embeddings are now auto-available! $ curl http://localhost:8080/api/v1/embeddings/drivers
Fallback: Environment Variables
For quick local development, you can still use environment variables. These are used as a fallback when no providers with embedding capabilities are configured:
terminal
# OpenAI embeddings $ export OPENAI_API_KEY="sk-..." $ export AGENTOVEN_EMBEDDING_MODEL="text-embedding-3-small" # Ollama embeddings $ export OLLAMA_URL="http://localhost:11434" $ export AGENTOVEN_OLLAMA_EMBED_MODEL="nomic-embed-text"
API Endpoints
| Method | Path | Description |
|---|---|---|
GET | /api/v1/embeddings/drivers | List registered embedding drivers |
POST | /api/v1/embeddings/embed | Generate embeddings for text input |
GET | /api/v1/embeddings/health | Health check all embedding drivers |
Enterprise Extensions
AgentOven Pro adds embedding support for enterprise cloud providers:
- AWS Bedrock — Amazon Titan embeddings (v1, v2, Image), Cohere English/Multilingual
- Azure AI Foundry — Azure OpenAI deployment-based embeddings
- Google Vertex AI — text-embedding-004/005, multilingual-002, gecko@003