On-prem, edge and air-gapped

On-prem AI consulting

For the environments a public API cannot reach

Some data cannot leave the building. We design and deploy AI systems that run entirely inside your infrastructure: self-hosted models, offline installation, and edge deployment for sites with no outbound internet access at all.

The problem with "just use the API"

For a great many organizations, the calculation is not about cost per token. It is that the document set is classified, or export-controlled, or covered by a data-residency clause, or simply too commercially sensitive to hand to a third party that reserves the right to change its terms. Once sending the data out is off the table, most of the vendor landscape goes with it.

The engineering that remains is real but tractable: model selection under fixed hardware, quantization and serving choices that decide whether latency is acceptable, retrieval that works on your actual file formats, and an installation path that a site engineer can run without internet access. That is the work we do.

What an on-prem engagement covers

Self-hosted inference

Model selection sized to the hardware you have or can buy, with serving, quantization, and batching decisions made against your real latency and throughput targets.

Offline installation

An installer that bundles the models, so deployment works with no outbound connectivity. Air-gapped installation is a supported path, not a workaround.

Retrieval over your own files

PDF, Word, PowerPoint, Excel, CSV, Markdown, HTML, and images, including scanned and image-only PDFs read with OCR. Answers link back to the source passage.

Edge deployment

Containerized deployment to constrained hardware at the edge, with the same governance and version control as the datacentre.

Hardware and capacity planning

What GPU, how many, and what it will cost to run, before you buy rather than after.

Governance that travels

Versioning, traceability, and audit evidence that work the same way in a disconnected environment as in the cloud, aligned to ISO/IEC 42001.

Products or consulting?

Where the requirement is knowledge search over your own documents, our SnowShoe.ai platform already does it and ships with an offline installer. Where the requirement is the infrastructure layer (versioned data, automated retraining, deployment pipelines), that is EdgeKube.ai. Both are licensed separately from consulting work.

Where neither fits, we build for your environment and you own the result: clients own project-specific code, models, and documentation as defined in the applicable agreement.

Registered to contract in the United States

U.S. and NATO registration is already in place, so teaming does not wait on paperwork. Every identifier is verifiable in a public registry.

Unique Entity IDJJADR6LWJ743
NCAGE codeL0QB4
U.S. consortiumSOSSEC member
CybersecurityCMMC, scope on request
Controlled goodsCGP registered
PersonnelTop Secret cleared staff

Registry entries establish procurement eligibility. They do not imply endorsement by any government. Full detail is in our capability statement.

Scope an on-prem deployment

Tell us the environment: what the connectivity looks like, what hardware exists, and what the data is. We will tell you what is realistic.

By submitting this form, you agree to be contacted about your enquiry. See our Privacy Policy.

What happens next

  1. An engineer reads it and replies to confirm fit.
  2. A 30-minute discovery call to frame the problem, constraints, and success measures.
  3. A fixed-fee scope with deliverables, milestones, and price, typically within two weeks.

You own the code, models, and documentation we produce. Defense and government enquiries can go direct to [email protected].