LLM EdgeBox

Private AI on your terms

A ready-to-use AI appliance for your office: chat assistant, document work and automation — fully offline. Plug it into your network and your team gets a private AI within minutes.

Fully offlineOn your premisesSet up for you

Overview

Ready to go

Models and interface are pre-installed and start automatically.

Maximum privacy

Everything you send and receive stays inside your network.

Familiar interface

A browser-based chat your team already knows how to use.

Features

Work with internal documents

Ask questions about your own files and knowledge base (RAG) — without uploading them anywhere.

Your tone of voice

The assistant uses your terminology and style; the interface can carry your brand.

OpenAI-compatible API

Connect existing tools, agents and scripts to a local endpoint.

Team access

Accounts and access control for your whole team.

Modern open models

Current open-weight models, from fast assistants to large reasoning models.

Updates and support

We keep models and software up to date and help your team get the most out of it.

Configurations & pricing

Choose the option that suits you: build it yourself, have us set it up, or get a turnkey solution.

EdgeBox Basic

A compact appliance for a team of up to 20 people.

$9,900
128 GB unified memory class
  • Models up to ~120B parameters (MoE)
  • Chat interface and API
  • Installation and onboarding
  • 12 months of support
Most popular

EdgeBox Pro

A workstation for demanding teams, tuned to your documents and tasks.

From $24,900
Professional GPU, 96 GB
  • Large models up to ~235B (MoE)
  • Search over your documents (RAG)
  • Customisation for your tasks
  • Priority support for 12 months

Enterprise

Server infrastructure, training on your data and on-site support.

From $60,000
Custom quote
  • Multi-GPU servers
  • Fine-tuning on your data
  • Integration with your systems
  • SLA and on-site support

Prices in USD. Shipping and taxes are calculated individually. Questions: mail@cakescats.com

Technical specifications

DeploymentOn your premises, fully offline; no cloud connection required
ModelsCurrent open-weight LLMs (e.g. Qwen, Llama, Mistral families), updated by us
InterfacesWeb chat, OpenAI-compatible API, document search (RAG)
AccessAccounts, roles, HTTPS on your local network
Basic / Pro hardware128 GB unified-memory appliance / workstation with a 96 GB professional GPU
Setup timeAbout 10 minutes after connecting to your network

Frequently asked questions

Does it work without the internet?
Yes. LLM EdgeBox is fully offline; your data never leaves your network.
What do I need to get started?
Power and a connection to your local network. Everything is pre-installed.
How does it compare to cloud AI?
No per-request fees, no data leaving the company and no dependency on a provider's rules — at the cost of a one-time hardware purchase.
Can it be adapted to our company?
Yes: your terminology, tone of voice, document base and branding. Pro and Enterprise include deeper customisation.
Why are prices higher than last year?
Memory and GPU prices rose sharply in 2026 because of the global memory shortage. In return, today's models are far more capable: Basic runs models that needed a server rack two years ago.

See private AI in action

Book a demo — we'll show EdgeBox on your kind of documents.