Pricing
Predictable infrastructure pricing without hyperscaler complexity.
Fixed monthly plans for shared inference, dedicated AI nodes and Apple Silicon build machines. You know what private AI costs before the month starts — and so does your finance team.
AI plans per month
Launch pricing- DeveloperManaged AI API€49
- GrowthManaged AI API€149
- ScaleManaged AI API€349
- Dedicated NodePrivate AI NodeFrom €449
- Managed AI NodePrivate AI NodeFrom €699
- High AvailabilityPrivate AI NodeFrom €1,299
All prices exclude VAT.
Plans
Choose a product family.
Most production teams land on the Managed AI Node: a dedicated, fully managed endpoint. Start on the shared API if you are still finding your volume.
Dedicated Apple Silicon nodes, managed end to end.
Dedicated Node
Entry dedicated compute for teams that run their own stack.
From€449/ month
Launch pricing
- Dedicated Apple Silicon node
- Up to 96 GB unified memory
- Private, single-tenant environment
- Private API endpoint
- Monitoring
- Basic managed deployment
Managed AI Node
RecommendedA dedicated, fully managed AI endpoint. Our flagship.
From€699/ month
Launch pricing
- Dedicated 96 GB Apple Silicon environment
- Model installation and configuration
- OpenAI-compatible endpoint
- Monitoring and alerting
- Model and system updates
- Secure API gateway
- Backup configuration
- Technical support
- Deployment assistance
High Availability
Two-node architecture for production workloads that must stay up.
From€1,299/ month
Launch pricing
- Two-node architecture
- Load balancing across nodes
- Redundant capacity
- Monitoring
- Failover design
- Priority support
- Custom deployment
- Private networking
All prices exclude VAT.
Compare features
Everything in each plan, side by side.
On small screens, scroll the tables sideways — the feature column stays in place.
Managed AI API
| Feature | Developer€49 / month | Growth€149 / month | Scale€349 / month |
|---|---|---|---|
| Infrastructure & capacity | |||
| Infrastructure | Shared, managed | Shared, managed | Shared, managed |
| Compute allowance | Fair-use compute allocation on shared infrastructure | Higher compute allowance, tailored to the selected model | High compute allowance, sized with you per model |
| Request priority | Standard | Priority queue | Priority inference |
| Concurrency | Standard | Standard | Higher |
| Endpoints | |||
| Chat completions | Included | Included | Included |
| Embeddings | Included | Included | Included |
| Speech-to-text | Not included | Included | Included |
| Reranking | Not included | Not included | Included |
| OpenAI-compatible request format | Included | Included | Included |
| Account & support | |||
| API keys | Included | Included | Included |
| Usage reporting | Usage dashboard | Usage analytics | Usage analytics |
| Support level | Standard | Standard | Priority |
| Custom model consultation | Not included | Not included | Included |
Private AI Node
| Feature | Dedicated NodeFrom €449 / month | Managed AI NodeFrom €699 / month | High AvailabilityFrom €1,299 / month |
|---|---|---|---|
| Hardware & tenancy | |||
| Dedicated nodes | 1 | 1 | 2 |
| Unified memory | Up to 96 GB | 96 GB | 2 × 96 GB |
| Single-tenant hardware | Included | Included | Included |
| Load balancing across nodes | Not included | Not included | Included |
| Redundant capacity & failover design | Not included | Not included | Included |
| Managed service | |||
| Deployment | Basic managed deployment | Model installation & configuration | Custom deployment |
| Private API endpoint | Included | Included | Included |
| OpenAI-compatible secure API gateway | Not included | Included | Included |
| Monitoring | Monitoring | Monitoring & alerting | Monitoring & alerting |
| Model and system updates | Not included | Included | Included |
| Configuration backups | Not included | Included | Included |
| Deployment assistance | Not included | Included | Included |
| Security & connectivity | |||
| TLS & API key authentication | Included | Included | Included |
| IP allowlisting | Included | Included | Included |
| VPN / private connectivity | On request | Eligible deployments | Private networking |
| Configurable content logging | Included | Included | Included |
| Bring your own compatible model | After compatibility check | After compatibility check | After compatibility check |
| Support & terms | |||
| Support level | Standard | Priority | Enterprise |
| Production SLA options | On request | Available | Available |
Mac Cloud plans are compared on the Mac Cloud page. Production SLA options are available for eligible managed deployments.
Support
Engineers who know your deployment.
Support is run from Germany / Europe, in English and German. Response targets are set in your agreement, not in marketing copy.
Standard
- Included in
- Developer, Growth, Dedicated Node, Developer Mac
- Channels
- Email, Documentation
- Response target
- Defined in your agreement
- Business-hours support (CET/CEST)
- Onboarding documentation
- Status notifications
Priority
- Included in
- Scale, Managed AI Node, Dedicated & CI/CD Mac
- Channels
- Email, Shared chat channel
- Response target
- Defined in your agreement
- Prioritised ticket handling
- Deployment assistance
- Model update planning
Enterprise
- Included in
- High Availability and custom infrastructure
- Channels
- Email, Shared chat channel, Named engineer
- Response target
- Defined in your agreement
- Named technical contact
- Architecture reviews
- Custom escalation path
Service levels. Production SLA options are available for eligible managed deployments. Read the SLA framework.
Beyond the standard plans
Validate first, or design something bigger.
Private AI Pilot
Validate your workload on dedicated infrastructure before committing to a larger deployment.
Contact sales
Scoped and quoted per workload before anything is deployed.
- Model deployment on dedicated capacity
- Private OpenAI-compatible endpoint
- Benchmarking on your real workload
- Architecture recommendation
- Migration plan
Custom Infrastructure
Multi-node clusters, white-label platforms and bespoke deployments.
Custom quote
Designed with an engineer, priced per architecture.
- Multiple dedicated nodes and private clusters
- White-label infrastructure for agencies
- Dedicated model hosting
- Industrial and on-premise-adjacent AI
- Custom networking and VPN
- Large storage and model repositories
- Custom model deployments
Calculator
Shared, dedicated or build it yourself?
Estimate your monthly cost from model, volume and deployment type — and compare it with what assembling and running the same stack in-house would take.
The fine print, up front
What's included, what isn't, and how fair use works.
No hidden line items. If something affects your bill or your architecture, it is on this page.
Prices exclude VAT
All prices exclude VAT. Invoices are issued in EUR; VAT treatment is confirmed during onboarding.
Fair use on the shared API
Managed AI API plans include a compute allowance tailored to the selected model, not a published token quota. Requests above your rate limit are throttled, and we contact you before recommending a larger plan.
One node is one node
A Private AI Node has finite memory. Which models fit depends on parameter count, quantization and context length. Running several large models usually means more than one node.
“From” means configuration-dependent
Dedicated plans are priced from a base configuration. Additional storage, custom networking or unusual model setups are quoted before you commit.
Not included
Application hosting, CUDA-based workloads, model training and third-party model licenses. You are responsible for complying with the license of each model you deploy.
Where it runs
Compute currently runs in Georgia, outside the EEA. An EU region is not available today. Workloads with personal data may require transfer safeguards and your own legal review.
Annual billing: 2 months free, invoiced yearly in advance.
FAQ
Pricing questions.
Can't find your answer? Ask an engineer — we reply with specifics, not a sales sequence.
Is the hardware dedicated?
On Private AI Node and Mac Cloud plans, yes — the physical Apple Mac Studio (Apple M3 Ultra, 96 GB) is assigned to you alone. Managed AI API plans run on shared infrastructure with per-customer authentication, rate limits and isolation at the API layer.
Where is the infrastructure located?
Our current compute region is Georgia. Commercial operations, sales and customer communication are run from Germany / Europe. Georgia is outside the European Economic Area — see our Data Processing page for what that means for personal data.
Can I bring or run my own model?
On Private AI Nodes, yes — including fine-tuned variants of supported architectures — provided the model is technically compatible with Apple Silicon inference runtimes, fits in memory with your required context length, and you have the rights to deploy it. We check compatibility before deployment.
Can I use an OpenAI SDK?
Yes, for supported API patterns. Our endpoints follow the OpenAI request and response format for chat completions, embeddings, audio transcription and model listing. Most applications switch by changing the base URL, the API key and the model name. Not every OpenAI parameter or endpoint is supported — see OpenAI compatibility in the docs.
Is my data used for model training?
No. Lirux does not train models on customer prompts, responses or files. We serve open-weight models; we do not build our own foundation models from customer data.
Can I cancel monthly?
Monthly API plans can be cancelled at the end of any billing period. Dedicated nodes are offered on monthly or longer terms, with notice periods stated in your order. [POLICY: confirm minimum terms before launch]
Can you support custom deployments?
Yes. Multi-node clusters, private networking, custom models, white-label endpoints and specific retention settings are designed with you. Request an architecture recommendation to start.
Can I connect over VPN?
Private connectivity (for example a site-to-site VPN or WireGuard tunnel) is available on eligible Private AI Node and High Availability deployments. Contact us for deployment requirements.
Can European companies use the service?
Yes. European businesses may use infrastructure outside the EEA. Workloads involving personal data may require appropriate international-transfer safeguards and internal legal review. We support contractual and technical measures such as a Data Processing Agreement and Standard Contractual Clauses where applicable. Speak with us about your data classification before deployment.
How does GDPR work?
When you send personal data to Lirux, you are typically the controller and we act as processor under a Data Processing Agreement. Because compute currently runs in Georgia (outside the EEA), a transfer mechanism such as Standard Contractual Clauses and a transfer impact assessment may be required. Our services are designed to support GDPR-conscious deployments, but customers remain responsible for determining the appropriate legal basis for their workloads. This is not legal advice.
Can I get an EU region?
An EU-based region is not available today. We are evaluating one. If EU data residency is a requirement for you, tell us — demand from customers directly informs our roadmap — and we will be transparent about timing.
Can agencies resell your infrastructure?
Yes. Our Partner Program is built for software agencies: partner pricing, client-specific private nodes, custom endpoints and white-label architecture where technically available. Apply on the Partners page.
Know your AI infrastructure cost before you build.
Pick a plan and deploy, or tell us your workload and get a concrete recommendation with a fixed monthly price.