Gain access to leading open AI models through a secure, fully managed inference service, provided by Intility from Norwegian data centers under Norwegian jurisdiction.
Gain access to leading open AI models through a secure, fully managed inference service, provided by Intility from Norwegian data centers under Norwegian jurisdiction.
Access leading open models through OpenAI-compatible APIs. Intility develops, secures, and operates the entire service – from underlying infrastructure and model management to monitoring, support, and continuous development.
Choose a model tailored to the specific task.
Qwen3 Embedding 8b
A multilingual embedding model for semantic search, information retrieval, classification, and utilizing the organization’s own data in AI solutions.
Price per 1m tokens
Input
NOK 1,00
Cached Input
NOK 0,30
Output
n/a
GLM-5.2 744B
An advanced reasoning model developed for demanding tasks including coding, tool usage, and complex workflows.
Price per 1m tokens
Input
NOK 20,00
Cached input
NOK 5,00
Output
NOK 75,00
GLM-5.3 744B
Z.AI's next generation open top model, developed for advanced reasoning, software development, agent-based tasks, and cyber defense. The model combines strong code understanding with the ability to identify vulnerabilities, use tools, and solve complex tasks over time.
Connect leading open models to the organization's own applications, data, and workflows. Build new AI features, analyze and structure large amounts of information, make the organization's knowledge more accessible, and automate tasks that previously required manual processing.
Incorporate generative AI into new and existing applications, products, specialized systems, and digital services.
Use language models to analyze, summarize, classify, and structure large amounts of text and other information.
Combine embedding models and language models to find and utilize relevant information from the organization's own data sources.
Use AI to process information, classify inquiries, and automate repetitive tasks in established workflows.
Connect OpenCode and other developer tools to Intility Inference for code generation, debugging, and complex development tasks.
Develop agents that can reason, use tools, and perform tasks across the organization's applications and data sources.
Intility Inference is delivered as an integrated part of an established enterprise platform, developed, secured, and operated by Intility. The platform is already used by hundreds of businesses, including critical and regulated entities with high demands for security, privacy, compliance, availability, and operational control.
From the underlying infrastructure and model management to monitoring, support, and continuous development, the entire service is delivered as a single cohesive offering.
Models and customer data are processed in secure, Norwegian data centers with high operational reliability and control.
The service is provided by a Norwegian supplier and is fully subject to Norwegian law and jurisdiction.
Security is built into the entire delivery, from the underlying infrastructure to ongoing monitoring and incident management.
Transition from experimentation to mission-critical workloads without switching platforms or building a new underlying architecture.
Intility Inference is developed and operated within Intility's established framework for information security, risk management, and operational control. The service follows documented processes for access control, change management, monitoring, incident management, and continuous improvement.
This provides the organization with a clear and documentable security and control environment around the use of open AI models – from underlying infrastructure and data processing to ongoing operation and further development of the service.
ISAE 3402 Type II
Independent attestation of controls related to information security
ISAE 3000 Type I
Independent attestation of established controls and processes
ISAE 3000 Type II
Independent attestation of controls related to privacy and data processing
Trusted Cloud Provider
Registered in the Cloud Security Alliance STAR Registry
NSM Basic Principles
The platform is developed and operated in accordance with NSM's fundamental principles for ICT security
ISO 27001
Certified information security management system
The API is compatible with the OpenAI Chat Completions API. Point to an SDK or framework you are already using against the endpoint at Intility, and keep the rest of the application architecture.
The widest compatibility across SDKs, frameworks, and tools.
For OpenAI SDKs configured against the Responses API.
Receive tokens continuously as server-sent events, so that the responses can be displayed as they are generated.
Define tools and let the model return structured function calls that the application can execute.
Limit the model's responses to structured and machine-readable JSON for use in applications and automated workflows.
Enable extended reasoning for complex tasks and receive a structured summary of the reasoning along with the answer.
Behind Intility Inference is a dedicated professional environment responsible for the entire delivery – from security, platform operations, and model management to consulting and integration with the business's own solutions.
Intility's Security Operations Center monitors the inference platform and handles security incidents around the clock.
Dedicated platform specialists operate, maintain, and further develop the infrastructure, models, and associated services.
People working with these models daily, so a model change or end of life reaches you as advance notice rather than a surprise.
Strategic advisory and hands-on expertise, so the platform turns into working solutions inside your own systems.
Here you will find answers to some of the most common questions about models, data, integration, pricing, and access. Please contact us if you have any further questions.
Tel: +47 93 63 08 87
Mail: [email protected]