Skip to content

Two engineers, a small number of deployments at a time. Currently taking on new work.

SRI

Services

We run the AI. You keep the data.

You buy a machine, or we rent you a dedicated one. We install the model, wire it to your documents, and manage it over an SSH channel you control. Your team gets a search bar and a chat box. Nothing leaves your network.

What we actually do

We run the AI. You keep the data.

You buy a machine, or we rent you a dedicated one. We install the model, wire it to your documents, and manage it over an SSH channel you control. Your team gets a search bar and a chat box. Nothing leaves your network.

FIG.1

01

Unsorted internal documents scanned and returned as a ranked shortlist of use casesWHAT YOU HAVEcontractsticketsemailwikiWORTH BUILDING FIRST1. doc search2. contract review3. ticket triage+ hardware sized, data flow written

We work out what to run

One call, then a spec instead of a slide deck.

We look at the work your team actually does and which parts are blocked today. You get back a shortlist of use cases worth building, the model that fits them, and the exact hardware it needs, priced.

  • 1.1 Use cases ranked by value and by how hard they are to approve
  • 1.2 Model and hardware sized for your volume, not for a benchmark
  • 1.3 A written data flow your security team can sign off on

FIG.2

02

Two setups: hardware in your rack, or a dedicated GPU we rent for you, both managed over SSHSRItwo engineerssshOPTION A / YOUR RACKyour boxyou buy itwe configure itOPTION B / OUR HOSTINGgpu, eusingle tenantlive in daysmodel · retrieval · internal API installed on both

We set the machine up

Your hardware, or a dedicated GPU we rent for you.

Buy a box and keep it in your own rack, or let us put a single-tenant GPU in an EU data centre in your name. Either way we install the model, the document search, and the internal API, and we manage it remotely over SSH.

  • 2.1 Your hardware: you own it, we configure it, access on a channel you control
  • 2.2 Our hosting: dedicated GPU, EU region, live in days instead of a quarter
  • 2.3 Model, retrieval, and internal API wired to SharePoint, ERP, or your own systems

FIG.3

03

Queries circulate inside your perimeter while outbound calls are stopped at the boundaryYOUR PERIMETERmodelon your hostOUTSIDEmodel vendorcloud apiunreachableby design

Nothing leaves the network

No third-party API anywhere in the request path.

Inference happens on your machine. There is no call out to a model provider, so there is no processor agreement, no transfer, and no retention question. Every query is logged inside your own perimeter.

  • 3.1 Zero outbound calls during inference, verifiable on your own firewall
  • 3.2 Role-based access and a full audit trail of every question asked
  • 3.3 Air-gapped option with no route in or out at all

FIG.4

04

A continuous health pulse with maintenance events we handle without your team noticingUPTIMEmonitored by usWHAT WE DO WHILE YOU WORKpatch appliedmodel swappedcapacity raised

We keep it running

Monitoring, patches, and model upgrades on one monthly fee.

AI systems are not set and forget. We watch performance, apply updates, and swap in better open models as they ship. When something breaks, the alert goes to us and not to your service desk.

  • 4.1 Uptime and latency monitoring, with alerts routed to us first
  • 4.2 Model upgrades and capacity planning as usage spreads across teams
  • 4.3 A direct line to the two engineers who built your deployment

The decision on one page

Same capability. Different threat model.

This is the table your security officer will build anyway. Here it is up front.

Cloud AI compared with a private deployment by SRI Systems
CriterionCloud AISRI
Where inference runsVendor infrastructure, region of their choosingHardware you own or rent in your own name
Who holds your documentsA third party, under their retention policyYou. The index never leaves your disk
Processor agreement neededYes, plus transfer assessmentNo third-party processor in the request path
Works with no internetNoYes, on-premise and air-gapped
Audit log locationVendor console, exportable at bestYour own log stack, queryable like any other service
Model choiceWhatever the vendor ships and deprecatesOpen weights you pin, upgrade, or roll back
Cost shapePer token, unbounded, scales with successFixed hardware or fixed monthly, flat under load
If the relationship endsAccess stops, data export on their termsThe system keeps running. You already own it

What it actually gets you

The same tools your competitors use. Without the paperwork.

Search everything you own

Ask a question in plain language and get an answer with citations from your contracts, policies, and archives.

Draft and review faster

First-pass drafting, summarising, and clause comparison on documents that were never allowed near a cloud tool.

Triage the queue

Classify and route tickets, mail, and forms automatically, with a human check where it matters.

Keep the audit trail

Every prompt and answer logged inside your network, attributable to a user, retained on your terms.

What it costs

Three line items. No per-token surprise.

The reason cloud AI budgets explode is that the bill scales with adoption. This does not. You pay to have it built, you pay for the machine, and you pay a flat fee to keep it running.

01

Assessment

Workflow review, model and hardware sizing, integration plan, and the written data flow your security team signs off on. Credited against the deployment if you go ahead.

Fixed fee, one off

02

Deployment

Everything installed, wired to your documents and systems, access control and audit logging in place, runbooks handed to your IT team. Quoted before you commit.

Fixed fee, one off

03

The machine

Buy the hardware and own it outright, or rent the same shape from us as a single-tenant GPU in an EU data centre. We tell you honestly which is cheaper for your volume.

Your capex, or our monthly

04

Support

Monitoring, patching, model upgrades, and capacity planning. Flat under load, because the cost of an extra thousand queries on hardware you already own is zero.

Flat monthly

Both fixed fees are quoted after the discovery call and before you commit to anything. There is no per-seat licence and nothing that expires.

Next step

Tell us what your team is not allowed to do yet.

A 45-minute call. We look at your workflows, your data rules, and whether local AI is worth it for you. If it is not, we will tell you that instead of selling you a project.

Not ready for a call? Send the security brief to whoever has to approve it, or just reply to an email. All three work.