LOCAL AI FOR ONE PERSON OR A TEAM

Local LLM, Orchestrated.

The Inomina Gateway inference server runs local models on your machine, and the Inomina Mate agent system carries out multi-step computer work. Inomina brings the model, knowledge, workflow, and tools into one application.

LOCAL + CLOUD MODELS
KimiMoonshot AI DeepSeekDeepSeek GLMZ.ai QwenAlibaba GPTOpenAI ClaudeAnthropic GeminiGoogle LlamaMeta
Built for local LLMs

From interface to inference, one design.

Putting a local model to work takes more than the model: an application to work in, a workspace for each kind of task, a knowledge layer, a workflow engine, an agent that operates a computer, and the inference service beneath them. Inomina implements all six and stacks them as a single design.

All six layers are Inomina's design and implementation

No layer is handed to somebody else's service. Underneath the inference layer we embed open-source execution engines, which Inomina builds and drives; the model lifecycle, the serving API, knowledge, workflows, and the agent above them are our own implementation. Because the boundaries between them belong to the same design, a model loaded once behaves identically in chat, in a workflow, and in the agent system — and changing that behaviour is a change to our own code, not a request to a supplier.

INOMINA / SYSTEM CROSS-SECTION
01
Application Desktop app · browser (Enterprise)
02
Workspaces Chat · translation · speech · documents · research
03
Knowledge Hybrid search · hierarchical index · knowledge graph
04
Workflows Visual graphs · control flow
05
Agents Browser · shell · files · MCP
06
Inference Gateway · model lifecycle · serving APIs
MODULAR
The model The only part you swap
GGUFSafetensorsMLXCloud API
System cross-section
01

Models under your control

Use Gateway to discover, load, serve, and observe local models, while keeping cloud adapters available for workloads that need them.

02

Focused AI workspaces

Move between chat, translation, speech, knowledge search, documents, slides, reports, and research without rebuilding each experience from scratch.

03

An agent system that executes

Use the Mate agent system when a task needs planned browser, shell, file, search, or configured MCP actions—with approvals and execution evidence.

Focused workspaces

Choose the workspace for the job.

Inomina presents different working modes for different outcomes: everyday chat and media, translation and transcription, knowledge retrieval, document and presentation creation, deep research, and executable work through an agent system.

One application, multiple modes

Select Standard, Image, Speech, Translation, Transcribe, Notebook, Global RAG, Agentic RAG, Document, Slide, Report, Deep Research, Fusion, or Mate according to the work in front of you. Availability can depend on configuration.

INOMINA / WORKSPACE MAP WORKSPACE MAP
Workspace selection overview
VISUAL WORKFLOW EXECUTABLE GRAPH
Visual workflow execution
Visual workflows

Design the steps, not just the prompt.

Build an executable graph from LLM, code, generated-tool, RAG, and MCP nodes. Connect inputs and outputs visually instead of hiding the whole process inside one prompt.

Control flow you can inspect

Add conditions, iteration, while loops, parallel branches, Try/Catch, delays, and merge nodes. Configure variables and ports, then inspect execution results from the same workflow.

CAPABILITIES

Run locally. Automate the work.

Manage models, connect repeatable steps, and bring the tools your work already depends on.

01 / LOCAL MODELS

Download, load, and run local models

Gateway brings model discovery, download, compatibility checks, loading, unloading, and service status into one control surface.

02 / AUTOMATION

Turn repeatable work into a workflow

Start with manual input or a webhook, then connect LLM, RAG, and tool nodes with conditions, loops, parallel branches, retries, delays, and merges.

03 / TOOLS

Connect the tools you already use

Call tools from configured MCP servers, use built-in code tools, or create a visual tool and place it directly in the workflow.

Connected platform

One application. Three connected layers.

The application brings model selection, focused AI workspaces, knowledge, and visual workflows together. Gateway supplies inference; the Mate agent system supplies tool execution.

  • 01

    Workspace: chat, media, translation, knowledge, documents, research, and workflows.

  • 02

    Gateway: model routing, local runtimes, model lifecycle, and inference APIs.

  • 03

    Agent System: planned browser, shell, file, search, and MCP execution with approvals and observable output.

inomina.platform CONNECTED
01 / MODEL LAYER
Gateway
local models · cloud adapters · inference

02 / WORKSPACE
Inomina
chat · RAG · documents · visual workflows

03 / EXECUTION
Agent System
browser · shell · files · MCP
INTEGRATED COMPONENTS

Two systems power Inomina.

Inomina is built on two dedicated systems with clearly separated roles: Inomina Gateway runs the models, and Inomina Mate executes the work.

01

Inomina Gateway

The inference server that runs local LLMs. It executes GGUF and Safetensors models on your CPU or GPU, exposes them through app-ready APIs, and manages everything from discovery and download to loading and release.

GGUFSafetensorsMLXCUDA
Learn More →
02

Inomina Mate

The agent system that does computer work on your behalf. It breaks a goal into a plan and executes it with the browser, shell, files, search, and configured MCP tools — progress, approvals, artifacts, previews, and diffs are visible per session.

BrowserShellFilesMCP
Learn More →

A deliberate path from model to result

01

Select and operate a local model through Gateway, with cloud access remaining an explicit option rather than the premise.

02

Use that model in chat, translation, knowledge retrieval, document creation, research, or a visual workflow.

03

When the task requires action, hand it to the Mate agent system with a session, execution policy, approvals, and observable outputs.

One app, any size

Start on one machine. Grow into a team.

The way you work together changes; the application does not. Only what it connects to changes with you.

01

One person

Everything runs on your own computer. The model, the conversations, and the documents stay where you put them.

Personal license · single seat
02

A team

Serve one machine on your network and let colleagues share the same models and knowledge from their own desktops.

Team license · five seats
03

An organization

Seats and concurrent sessions follow your agreement, and members can sign in from a browser without installing the desktop app.

Enterprise license · contract seats

Compare what each license includes on the pricing page.

STARTING POINTS

Start with the layer you need.

Adopt one capability first, then connect the others when the work requires them.

Model layer 01 Inomina Gateway

Begin by serving one local model for a concrete chat, extraction, coding, retrieval, translation, or speech task.

Workspace layer 02 Inomina

Use the model in a focused workspace or connect model, knowledge, tools, and control flow visually.

Execution layer 03 Inomina Mate

Add the Mate agent system when the desired result requires planned work outside a model response.

INOMINA

Start with one model and one real task.

Choose where the model should run, select the workspace for the outcome, and add a workflow or the Mate agent system only when the task needs execution.

Get Started