Managed AI operations
Operate infrastructure, models, retrieval quality and continuous improvement over time.
Deploy production inference, dedicated GPU compute and secure AI applications on infrastructure operated by Arctur in Slovenia. Build with APIs - or let our team deliver and manage the complete solution.
EU operated
Slovenian infrastructure
API → cluster
One deployment continuum
Private by design
Dedicated and on-prem option
Applied AI
RAG, fine-tuning, analytics
Built for real production workloads.
OpenAI-compatible APIs, dedicated endpoints, GPU/HPC capacity, private networking and applied-AI engineering—under one European operating model.
Start with a model endpoint or a single GPU. Expand into a private cluster, a managed knowledge assistant, or an AI system installed in your environment.
Model API
OpenAI-compatible access to curated language, embedding, reranking and multimodal models.
EXPLORE Inference →GPU & HPC capacity
Shared instances, dedicated GPU nodes and high-speed cluster capacity for AI and HPC workloads.
EXPLORE COMPUTE →Managed AI
Private RAG, copilots, data engineering, model adaptation and end-to-end AI application delivery.
EXPLORE SERVICES →AI systems
Buy, rent, host or operate purpose-built AI servers and private AI appliances.
EXPLORE HARDWARE →Use a familiar API surface while retaining a path to reserved throughput, dedicated serving and private connectivity.
Multimodal multilingual reasoning
Efficient coding and language model
Complex enterprise workflows
1M-context high-speed inference
Dedicated or private deployment
Two comparable multimodal 35B-A3B model variants for multilingual chat, image understanding, knowledge assistants, structured extraction and tool-enabled workflows.
| CONTEXT
32K |
MODALITIES
Text · Image · Code · Math |
| REASONING
Strong multilingual |
TOOL CALLING
Supported |
An efficient model for production applications, particularly suitable for coding and language tasks with a long context window.
| CONTEXT
262,144 tokens |
MODALITIES
Text · Image · Code |
| REASONING
Efficient coding |
TOOL CALLING
Supported |
A higher-capability model for complex enterprise reasoning, longer workflows and advanced assistant use cases.
| CONTEXT
64K |
MODALITIES
Text · Image · Code |
| REASONING
Complex enterprise |
TOOL CALLING
Supported |
A high-speed, long-context model for demanding applications that benefit from a 1 million token context window.
| CONTEXT
1m tokens |
MODALITIES
Text |
| REASONING
High-speed long context |
TOOL CALLING
Supported |
Bring an approved open-weight model, deploy one from the Arctur catalogue, or design a dedicated serving environment for your workload.
| CONTEXT
Configurable |
MODALITIES
Model dependant |
| REASONING
Model dependant |
TOOL CALLING
On request |
Every paid plan includes its full monthly commitment as inference credit. Overage discounts apply after included credit is used.
| Plan | Monthly commitment | Included inference credit | Overage discount | API keys | Concurrency | RPM | Support / controls |
|---|---|---|---|---|---|---|---|
| Free | €0 | €0 | 0% | 1 | 5 | 30 | Community / best effort |
| Starter | €59 | €59 | 0% | 3 | 10 | 75 | Usage dashboard, hard spend cap |
| Development | €250 | €250 | 5% | 10 | 25 | 150 | Email support, API key/project budgets |
| Enterprise | €2,000 | €2,000 | 12% | 50 | 150 | 800 | Priority queue, SLA, audit logs, named contact |
| Custom | €5,000+ | 100% of commitment | 15%+ negotiated | Custom | Custom | Custom | Reserved capacity, private deployment/networking, custom SLA |
Concurrency = simultaneous active requests. RPM = requests per minute. Model token rates are quoted separately and vary by model and deployment tier. Custom plans can include dedicated capacity and private network access.
Secure Slovenian infrastructure
InfiniBand and high-speed Ethernet
NVMe-backed performance tiers
24/7 monitoring and support options
A clear progression from standard CPU capacity to dedicated accelerator nodes and private, high-bandwidth AI clusters.
Standard Compute
From €1.80 / hour
EPYC High-Perf
From €3.00 / hour
RTX PRO GPU Node
From €5.00 / hour
MACx Node
From €5.00 / hour
Illustrative configurations and starting rates. Availability, hardware specifications and pricing should be validated before publication.
Our team combines AI infrastructure, data engineering and application delivery so customers do not have to assemble an AI platform from disconnected vendors.
Managed AI operations
Operate infrastructure, models, retrieval quality and continuous improvement over time.
Data engineering & analytics
Data preparation, lakehouse and warehouse delivery, dashboards, forecasting and governed analytics foundations.
Fine-tuning & model adaptation
Dataset readiness, LoRA/QLoRA, evaluation, quantisation and production serving for domain-specific models.
Custom AI products
AI-enabled workflows, document intelligence, agents, vision systems and product integrations built for measurable outcomes.
Architecture sprint
Prioritise the use case, evaluate data and establish deployment/security boundaries.
Proof of value
Build and evaluate a focused working system against customer-approved success criteria.
Production deployment
Integrate identity, access control, data sources, monitoring and operational processes.
Managed AI operations
Operate infrastructure, models, retrieval quality and continuous improvement over time.
Acquire an AI system, rent it in the Arctur data centre, host customer-owned hardware, or deploy it as a managed private AI environment.
Featured Configuration
A high-density AI accelerator server for private inference, RAG, fine-tuning and scalable enterprise workloads. Available as customer-owned hardware, hosted infrastructure or an Arctur-managed service.
| Accelerators | CPU | Connectivity |
|---|---|---|
| 8 × MACx AI | 2 × 128-core AMD Genoa | PCIe 5.0 / 5U |
Commercial Options
| Managed in Arctur DC | from €3,000/mo |
|---|---|
| Purchase hardware | from €79,000 |
| Private/on-prem design | On request |
Hardware examples and commercial figures are indicative. Final configurations, lead times, installation, warranty and sales arrangements are provided in a formal quote.
Tell us whether you need an API, GPU capacity, a private RAG assistant, an AI server, or an end-to-end delivery partner. We will route your request to the appropriate technical team.