BoltProof · How it's set up

See it working. Then see it in your office.

One headless computer in your office runs the AI. Staff connect from their own laptops and phones. A NAS keeps a copy of everything. Here is the full picture, then a live demo.

A BoltProof system is one small computer in your office, running without a screen, that every member of staff can reach from their own laptop or phone. The models, your documents and your chat history stay on that computer and its NAS backup — the only network traffic is encrypted Tailscale access for your authorized devices. This page shows how the parts fit together, what installation involves, and what you will see on the first day.

The whole setup in one picture

BoltProof setup: staff devices connect over Tailscale to a headless server in your office, backed up nightly to a backup server, with an optional off-site copyYour officeStaff laptopsand phonesTailscale app onlyTailscalePrivate, encryptedOffice, home or travelHeadless serverLocal models (Ollama)Open WebUI chatn8n automationsDocument Q&A (RAG)Hardened OSNo screen, keyboard or mouse2-bay NASMirrored drives (RAID 1)Nightly backup of the serverSnapshots (roll back)Optional add-onOptionaloff-site copyencrypted

Solid boxes are things you own. The dashed box on the right is optional. Staff devices can be anywhere; the server and NAS stay in your office.

What each part does

Staff laptops and phones

Staff use whatever they already have: Windows, Mac, iPhone, Android. They install one app (Tailscale) and open a web address. No software licences per seat, no new hardware for staff.

Tailscale

A private network between your devices and the server. Traffic is encrypted end to end. There's no port forwarding and no public-facing IP — the server isn't reachable from the open internet, only from devices your Tailscale network has authorized. You can see who is connected and remove a device in one click.

The headless server

A computer with no monitor or keyboard. It runs the language models (Ollama), the chat interface (Open WebUI), the automation engine (n8n) and the document index. the OS is hardened: disk encryption, firewall, automatic security patches, no sharing services.

2-bay NAS

Two mirrored drives, so one can fail without losing data. Every night it takes a full copy of the server. Snapshots let you roll back to yesterday, last week or last month. Priced in your proposal.

Optional off-site copy

An encrypted copy of the NAS sent to a location you choose: a second office, a home NAS, or a storage provider in the UAE. Protects against fire, flood and theft. You hold the encryption key.

Ongoing support

Quoted per request, no retainer. Ask us to apply updates, swap in newer models or fix a problem remotely, whenever you need it.

How installation works

  1. Short call. We agree the scope, the number of staff and which documents you want the AI to read. You tell us about your existing network and email.
  2. We build it in our workshop. Models, chat, automations and hardening are installed and tested on your server before it reaches you. Your documents are indexed if you have sent them in advance.
  3. Delivery and placement. The computer goes on a shelf, in a cupboard or in your network cabinet. It needs power and one network cable. The NAS sits next to it. About one hour on site; remote guidance elsewhere.
  4. Staff connect. Each person installs Tailscale, signs in, and bookmarks the chat address. Five minutes per person. We do this with your team on a screen-share if you prefer.
  5. Handover. A 60-minute walkthrough for your team: how to chat, how to ask about documents, how to approve an automated draft, who to call. You receive a one-page sheet and a login record.
  6. Thirty-day check. We review what your staff have used, tune the document index and adjust any automation. Included in every package.

What you see on day one

Screen 1

Private chat

A chat window that looks like the tools your staff already know. Drafting letters, summarising a long email thread, rewriting a paragraph, translating to Arabic. Every conversation is stored on your server and nowhere else.

Screen 2

Document Q&A

Ask a question across your own files: contracts, policies, patient leaflets, price lists, tenancy agreements. The answer quotes the source document and page so you can check it. New files added to the shared folder are indexed automatically.

Screen 3

Automation example

An enquiry lands in your shared inbox. n8n reads it, the local model drafts a reply using your price list and standard wording, and the draft appears in a folder or a chat for approval. A person clicks send. Nothing goes out without a human.

Other automations we set up on request: meeting notes to a summary and action list, invoices to a spreadsheet line, a daily digest of new documents, a WhatsApp enquiry to a CRM record.

What happens if the server fails

If the server fails, replacement hardware is restored from the NAS backup and back online in hoursServer failsHardware fault, theft, floodNAS still has everythingBackup + snapshotsReplacement serverFrom stock or Apple retailRestored from NASBack online in hours, not weeks

Hardware fails, offices flood, laptops walk. With the backup add-on, the NAS holds last night's full copy of the server and a chain of snapshots. We put replacement hardware in place, restore from the NAS, reconnect Tailscale, and your staff carry on with the same address, the same chat history and the same documents. Typical downtime is measured in hours. Without the NAS, expect one to three working days to rebuild.

Already live: try the assistant yourself on synthetic documents at demo.boltproof.com — no signup, 10 free messages an hour.

See it working on your own documents

Book a 20-minute live demo on your own documents

We share our screen, you send us two or three of your own non-confidential documents in advance, and you watch the system answer questions about them — then generate a draft. No slides. You can ask anything.

Book a 20-minute live demo on your own documents

Prefer to try it yourself right now?

Try the live sandbox on synthetic documents, no signup required, or compare the three packages / use the configurator to see which one fits your staff count and document volume.

Try the live sandbox demo See packages Open the configurator

For technical buyers

The plain-English explanations above are the whole story for most people. If you want to validate the architecture yourself, here is the detail.

Hardware

Unified memory (RAM and GPU memory in one pool) is what makes running large open-weight models on a desktop machine practical — a 64 to 512 GB unified-memory machine can hold and run models that would otherwise need a rack of GPUs. The machine's GPU and neural-engine cores accelerate inference. We size the hardware to the model sizes and concurrent-user count you need — see the how we work.

Local models: what's actually running

Models run through Ollama, an open-source runtime that loads open-weight models (Qwen, Llama, Gemma, DeepSeek and others) directly on the server. Open WebUI provides the chat interface on top. No model weights or inference requests leave the machine for local models — inference happens entirely on that hardware. We choose specific models based on your memory budget and task mix, and swap them for better releases when you ask us to.

Document retrieval (how Document Q&A actually works)

This is retrieval-augmented generation (RAG): your documents are split into chunks, each chunk is converted to a vector embedding using a local embedding model (nomic-embed), and those vectors are stored in a local index. When you ask a question, the system finds the most relevant chunks by vector similarity, feeds them to the language model as context, and the model answers using that context — citing which file and page the answer came from. New files dropped into a watched folder are indexed automatically. Nothing about this process sends your documents anywhere; the index and the embeddings live on the same machine as the model.

Permissions and access control

Open WebUI gives each staff member their own login. Access to specific document folders can be restricted per user or group — for example, so reception can't see HR files. Every login is tied to a Tailscale-authenticated device, so there's no shared password floating around. We set this up at delivery based on who should see what; ask us if you need more granular controls than the default.

Networking (Tailscale)

Tailscale creates a private mesh network (built on WireGuard) between your staff's devices and the server. There is no port forwarding and no public-facing IP or exposed service — the server is simply unreachable from the open internet. Each device authenticates individually, traffic is encrypted end-to-end, and you can revoke a single device's access instantly from the Tailscale admin console without touching anything else.

Backup

The optional NAS add-on takes a nightly backup of the server (documents, chat history, model configuration and automations) onto two mirrored drives, with versioned snapshots so a file overwritten today can be recovered from yesterday's snapshot. An optional encrypted off-site copy protects against fire, flood or theft; you hold the encryption key. Without the NAS add-on, you're relying on whatever backup discipline you already have — we'll say so plainly if that's a gap worth closing.

Security posture

Every deployment gets OS hardening (full-disk encryption, firewall enabled, automatic security updates, no unnecessary sharing services, a separate admin account) as standard. The fuller checklist — 2FA enforcement, access reviews, ransomware-resistant backup design, incident response planning — is documented on the Security & hardening page.

Remote support

If you ask us for remote support, we connect over the same Tailscale network your staff use — there's no separate remote-access tool or exposed port added for us. We can only reach the server while your Tailscale network authorizes our device, and you can revoke that at any time. Updates are staged with a NAS snapshot taken first, so any update can be rolled back if something breaks.

Model selection

We pick models based on three things: how much unified memory the server has, how many people will be using it at once, and what the work actually is (drafting and Q&A need different strengths than translation or code review). Smaller machines run models in the 8–30B parameter range; larger ones run 70B to 400B-class models, usually quantised to fit memory while keeping quality close to full precision. We test the first real answer against your own documents before handover, not a generic benchmark.

Integrations

n8n handles automations and connects to external systems through whatever export or API they expose — reading an inbox via IMAP, watching a folder for new invoices, calling a practice-management system's API if it has one. If a system only exports CSV or PDF, we build the workflow around that instead. Tell us what you already run (accounting software, CRM, practice management) and we'll say plainly whether it connects before you order.

When external APIs are, and aren't, used

By default, nothing is configured to call an external AI API. If you choose to add a cloud model (GPT, Claude, Gemini) for non-sensitive work — general drafting, marketing copy, research with no client data — that's a separate, clearly labelled option in Open WebUI that you switch on and pay for yourself with your own API key. Staff can see which model, local or cloud, they're using for every conversation. The local vs frontier vs hybrid comparison covers how to decide what should go where.

Frequently asked questions

Is my data sent anywhere?
No. The models that answer your questions run on the server in your office, not in a cloud AI service. Your prompts, documents and chat history stay on that machine and on your NAS. BoltProof does not receive copies. The network traffic that exists is: Tailscale, which carries encrypted traffic between your own authorized devices; software updates; and, if you add it, the encrypted off-site backup copy. If you later choose to switch on a cloud model for non-sensitive tasks, that is a separate, clearly labelled option that you turn on yourself with your own API key.
What if the internet goes down?
The AI keeps working. Staff in the office connect to the server over the local network, so chat, document Q&A and most automations carry on without internet. Anything that depends on outside services, such as reading a mailbox or sending a reply, resumes when the connection returns. Remote staff need internet to reach the office.
Can staff use it from home?
Yes. Each staff member installs the Tailscale app on their laptop or phone and signs in once. From then on the office AI is reachable from home, a client site or abroad, over an encrypted link authenticated to that device. The server stays unreachable from the open internet to anyone outside your Tailscale network, and there's no separate VPN box to manage. You can remove a person's access in one click.
How are updates done?
OS security updates install automatically. Model and application updates are yours to run using the one-page checklist we hand over at delivery, or you can book us for one-off update work, staged with a NAS snapshot first so it can be rolled back.
What models are included?
Each package ships with a set of current open-weight models sized for its memory: a general chat model, a model tuned for documents and summaries, an embedding model for document search, and a coding model if you want one. Smaller machines run mid-size models; larger machines run bigger models with longer context. Models are free to use and we swap them for newer releases when you ask us to.
Can we add ChatGPT for non-sensitive work?
Yes. Open WebUI can show cloud models such as GPT or Claude alongside your local models. We label them clearly and can restrict which staff see them. Staff choose a local model for client files and a cloud model for public tasks such as marketing copy. You supply the API key so billing stays with you.
What happens if the server fails?
If you have the backup add-on, the NAS holds a nightly copy of the whole server plus snapshots. We install replacement hardware, restore from the NAS and reconnect Tailscale. Typical downtime is hours. Without the NAS, we restore from whatever backup you have and rebuild the stack, which takes one to three working days.
Do we need IT staff?
No. The system is built to run without anyone on site touching it. Staff use a browser. If something looks wrong, you message us on WhatsApp. We fix it on a quoted, one-off basis — no retainer to manage. If you already have an IT provider, we work alongside them and give them full documentation.

Prices are in AED. Hardware is new unless marked certified pre-owned. Model names and sizes change as better open-weight releases appear; we tell you what is installed at handover.