Skip to content

Pricing

Predictable infrastructure pricing without hyperscaler complexity.

Fixed monthly plans for shared inference, dedicated AI nodes and Apple Silicon build machines. You know what private AI costs before the month starts — and so does your finance team.

AI plans per month

Launch pricing
  1. DeveloperManaged AI API€49
  2. GrowthManaged AI API€149
  3. ScaleManaged AI API€349
  4. Dedicated NodePrivate AI NodeFrom €449
  5. Managed AI NodePrivate AI NodeFrom €699
  6. High AvailabilityPrivate AI NodeFrom €1,299

All prices exclude VAT.

Plans

Choose a product family.

Most production teams land on the Managed AI Node: a dedicated, fully managed endpoint. Start on the shared API if you are still finding your volume.

Dedicated Apple Silicon nodes, managed end to end.

Billing period

Dedicated Node

Entry dedicated compute for teams that run their own stack.

From€449/ month

Launch pricing

Deploy a Private Node
  • Dedicated Apple Silicon node
  • Up to 96 GB unified memory
  • Private, single-tenant environment
  • Private API endpoint
  • Monitoring
  • Basic managed deployment

Managed AI Node

Recommended

A dedicated, fully managed AI endpoint. Our flagship.

From€699/ month

Launch pricing

Deploy Managed AI
  • Dedicated 96 GB Apple Silicon environment
  • Model installation and configuration
  • OpenAI-compatible endpoint
  • Monitoring and alerting
  • Model and system updates
  • Secure API gateway
  • Backup configuration
  • Technical support
  • Deployment assistance

High Availability

Two-node architecture for production workloads that must stay up.

From€1,299/ month

Launch pricing

Talk to an Engineer
  • Two-node architecture
  • Load balancing across nodes
  • Redundant capacity
  • Monitoring
  • Failover design
  • Priority support
  • Custom deployment
  • Private networking

All prices exclude VAT.

Compare features

Everything in each plan, side by side.

On small screens, scroll the tables sideways — the feature column stays in place.

Managed AI API

Managed AI API plan comparison
FeatureDeveloper€49 / monthGrowth€149 / monthScale€349 / month
Infrastructure & capacity
InfrastructureShared, managedShared, managedShared, managed
Compute allowanceFair-use compute allocation on shared infrastructureHigher compute allowance, tailored to the selected modelHigh compute allowance, sized with you per model
Request priorityStandardPriority queuePriority inference
ConcurrencyStandardStandardHigher
Endpoints
Chat completionsIncludedIncludedIncluded
EmbeddingsIncludedIncludedIncluded
Speech-to-textNot includedIncludedIncluded
RerankingNot includedNot includedIncluded
OpenAI-compatible request formatIncludedIncludedIncluded
Account & support
API keysIncludedIncludedIncluded
Usage reportingUsage dashboardUsage analyticsUsage analytics
Support levelStandardStandardPriority
Custom model consultationNot includedNot includedIncluded

Private AI Node

Private AI Node plan comparison
FeatureDedicated NodeFrom €449 / monthManaged AI NodeFrom €699 / monthHigh AvailabilityFrom €1,299 / month
Hardware & tenancy
Dedicated nodes112
Unified memoryUp to 96 GB96 GB2 × 96 GB
Single-tenant hardwareIncludedIncludedIncluded
Load balancing across nodesNot includedNot includedIncluded
Redundant capacity & failover designNot includedNot includedIncluded
Managed service
DeploymentBasic managed deploymentModel installation & configurationCustom deployment
Private API endpointIncludedIncludedIncluded
OpenAI-compatible secure API gatewayNot includedIncludedIncluded
MonitoringMonitoringMonitoring & alertingMonitoring & alerting
Model and system updatesNot includedIncludedIncluded
Configuration backupsNot includedIncludedIncluded
Deployment assistanceNot includedIncludedIncluded
Security & connectivity
TLS & API key authenticationIncludedIncludedIncluded
IP allowlistingIncludedIncludedIncluded
VPN / private connectivityOn requestEligible deploymentsPrivate networking
Configurable content loggingIncludedIncludedIncluded
Bring your own compatible modelAfter compatibility checkAfter compatibility checkAfter compatibility check
Support & terms
Support levelStandardPriorityEnterprise
Production SLA optionsOn requestAvailableAvailable

Mac Cloud plans are compared on the Mac Cloud page. Production SLA options are available for eligible managed deployments.

Support

Engineers who know your deployment.

Support is run from Germany / Europe, in English and German. Response targets are set in your agreement, not in marketing copy.

Standard

Included in
Developer, Growth, Dedicated Node, Developer Mac
Channels
Email, Documentation
Response target
Defined in your agreement
  • Business-hours support (CET/CEST)
  • Onboarding documentation
  • Status notifications

Priority

Included in
Scale, Managed AI Node, Dedicated & CI/CD Mac
Channels
Email, Shared chat channel
Response target
Defined in your agreement
  • Prioritised ticket handling
  • Deployment assistance
  • Model update planning

Enterprise

Included in
High Availability and custom infrastructure
Channels
Email, Shared chat channel, Named engineer
Response target
Defined in your agreement
  • Named technical contact
  • Architecture reviews
  • Custom escalation path

Service levels. Production SLA options are available for eligible managed deployments. Read the SLA framework.

Beyond the standard plans

Validate first, or design something bigger.

Private AI Pilot

Validate your workload on dedicated infrastructure before committing to a larger deployment.

Contact sales

Scoped and quoted per workload before anything is deployed.

  • Model deployment on dedicated capacity
  • Private OpenAI-compatible endpoint
  • Benchmarking on your real workload
  • Architecture recommendation
  • Migration plan

Custom Infrastructure

Multi-node clusters, white-label platforms and bespoke deployments.

Custom quote

Designed with an engineer, priced per architecture.

  • Multiple dedicated nodes and private clusters
  • White-label infrastructure for agencies
  • Dedicated model hosting
  • Industrial and on-premise-adjacent AI
  • Custom networking and VPN
  • Large storage and model repositories
  • Custom model deployments

Calculator

Shared, dedicated or build it yourself?

Estimate your monthly cost from model, volume and deployment type — and compare it with what assembling and running the same stack in-house would take.

The fine print, up front

What's included, what isn't, and how fair use works.

No hidden line items. If something affects your bill or your architecture, it is on this page.

Prices exclude VAT

All prices exclude VAT. Invoices are issued in EUR; VAT treatment is confirmed during onboarding.

Fair use on the shared API

Managed AI API plans include a compute allowance tailored to the selected model, not a published token quota. Requests above your rate limit are throttled, and we contact you before recommending a larger plan.

One node is one node

A Private AI Node has finite memory. Which models fit depends on parameter count, quantization and context length. Running several large models usually means more than one node.

“From” means configuration-dependent

Dedicated plans are priced from a base configuration. Additional storage, custom networking or unusual model setups are quoted before you commit.

Not included

Application hosting, CUDA-based workloads, model training and third-party model licenses. You are responsible for complying with the license of each model you deploy.

Where it runs

Compute currently runs in Georgia, outside the EEA. An EU region is not available today. Workloads with personal data may require transfer safeguards and your own legal review.

Annual billing: 2 months free, invoiced yearly in advance.

FAQ

Pricing questions.

Can't find your answer? Ask an engineer — we reply with specifics, not a sales sequence.

Is the hardware dedicated?

On Private AI Node and Mac Cloud plans, yes — the physical Apple Mac Studio (Apple M3 Ultra, 96 GB) is assigned to you alone. Managed AI API plans run on shared infrastructure with per-customer authentication, rate limits and isolation at the API layer.

Where is the infrastructure located?

Our current compute region is Georgia. Commercial operations, sales and customer communication are run from Germany / Europe. Georgia is outside the European Economic Area — see our Data Processing page for what that means for personal data.

Can I bring or run my own model?

On Private AI Nodes, yes — including fine-tuned variants of supported architectures — provided the model is technically compatible with Apple Silicon inference runtimes, fits in memory with your required context length, and you have the rights to deploy it. We check compatibility before deployment.

Can I use an OpenAI SDK?

Yes, for supported API patterns. Our endpoints follow the OpenAI request and response format for chat completions, embeddings, audio transcription and model listing. Most applications switch by changing the base URL, the API key and the model name. Not every OpenAI parameter or endpoint is supported — see OpenAI compatibility in the docs.

Is my data used for model training?

No. Lirux does not train models on customer prompts, responses or files. We serve open-weight models; we do not build our own foundation models from customer data.

Can I cancel monthly?

Monthly API plans can be cancelled at the end of any billing period. Dedicated nodes are offered on monthly or longer terms, with notice periods stated in your order. [POLICY: confirm minimum terms before launch]

Can you support custom deployments?

Yes. Multi-node clusters, private networking, custom models, white-label endpoints and specific retention settings are designed with you. Request an architecture recommendation to start.

Can I connect over VPN?

Private connectivity (for example a site-to-site VPN or WireGuard tunnel) is available on eligible Private AI Node and High Availability deployments. Contact us for deployment requirements.

Can European companies use the service?

Yes. European businesses may use infrastructure outside the EEA. Workloads involving personal data may require appropriate international-transfer safeguards and internal legal review. We support contractual and technical measures such as a Data Processing Agreement and Standard Contractual Clauses where applicable. Speak with us about your data classification before deployment.

How does GDPR work?

When you send personal data to Lirux, you are typically the controller and we act as processor under a Data Processing Agreement. Because compute currently runs in Georgia (outside the EEA), a transfer mechanism such as Standard Contractual Clauses and a transfer impact assessment may be required. Our services are designed to support GDPR-conscious deployments, but customers remain responsible for determining the appropriate legal basis for their workloads. This is not legal advice.

Can I get an EU region?

An EU-based region is not available today. We are evaluating one. If EU data residency is a requirement for you, tell us — demand from customers directly informs our roadmap — and we will be transparent about timing.

Can agencies resell your infrastructure?

Yes. Our Partner Program is built for software agencies: partner pricing, client-specific private nodes, custom endpoints and white-label architecture where technically available. Apply on the Partners page.

Know your AI infrastructure cost before you build.

Pick a plan and deploy, or tell us your workload and get a concrete recommendation with a fixed monthly price.