For e-commerce
Better search, better product content, more languages — on a fixed monthly budget.
Catalogues grow faster than content teams. Use open-weight models to generate and translate product content, power semantic product search and assist customer service — without sending your catalogue and customer conversations to a consumer AI service.
The problem
What gets in the way today.
- Thousands of SKUs without good descriptions
- Keyword search misses what customers mean
- Every new market means another translation backlog
- Support volume spikes around campaigns
Use cases
What teams build with it.
Semantic product search
Product descriptions
Multilingual translation
Recommendations
Customer support assistance
Image search and tagging
How it fits
Where it sits in your architecture.
Your application keeps running where it runs today. It calls a managed endpoint; we operate everything behind it.
- Shop / PIM
- Batch & real-time API calls
- Managed private AI
- Search index & content
Recommended plans
Where most teams start.
Fixed monthly plans. All prices exclude VAT.
Growth
Most popularFor production SaaS applications.
€149/ month
Launch pricing
Higher compute allowance, tailored to the selected model
- Higher compute allocation
- Chat completions
- Embeddings
- Speech-to-text
- Priority queue
- Usage analytics
- Email support
Managed AI Node
RecommendedA dedicated, fully managed AI endpoint. Our flagship.
From€699/ month
Launch pricing
- Dedicated 96 GB Apple Silicon environment
- Model installation and configuration
- OpenAI-compatible endpoint
- Monitoring and alerting
- Model and system updates
- Secure API gateway
- Backup configuration
- Technical support
- Deployment assistance
Recommended models
Models that suit this workload.
Model availability depends on licensing, memory requirements and deployment configuration.
Qwen3 8B
AvailableAlibaba Qwen · 8B (dense)
Compact, fast model for high-volume tasks such as tagging, translation drafts and short replies.
BGE-M3
AvailableBAAI · 568M
Multilingual embedding model for semantic search and RAG across 100+ languages.
BGE Reranker v2 M3
AvailableBAAI · 568M
Cross-encoder reranker that reorders retrieved passages to improve RAG answer quality.
Gemma 3 27B
Private NodeGoogle · 27B (dense)
Multimodal model that accepts text and images — useful for document images and product photos.
Considerations
Worth knowing before you start.
Batch jobs on dedicated capacity
Human in the loop
Customer data
Data & location
Transparent about where your data is processed.
Designed to support GDPR-conscious deployments. Customers remain responsible for determining the appropriate legal basis for their workloads.
- Infrastructure region
- Georgia
- Outside the EEA. All compute and model storage currently run here.
- Commercial operations
- Germany / Europe
- Sales, contracts, onboarding and customer communication.
- International data transfer
- Safeguards available
- For workloads involving EEA personal data, appropriate contractual and technical safeguards may be required.
Tell us what you are building.
An engineer reviews your requirements and comes back with a concrete proposal — model, plan and architecture.