Lume Autonoma Message us

Service · Private LLM

Private LLM fine-tuning for businesses in India

A private LLM is an open-source language model fine-tuned on your own documents and run on your own servers or private cloud, so your data never goes to a third-party AI. Lume Autonoma trains it with LoRA or QLoRA and quantizes it to 8-bit or 4-bit, so it runs fast on modest hardware.

7B model, 16-bit
≈ 14 GB
Same model, 4-bit
≈ 4 GB
Your data
Stays on your servers
Nova, Lume Autonoma's android guide, at a screen

What we build

Who it is for

Businesses whose documents or customer data must stay under their control, such as legal, accounting and consulting firms, and teams that want an assistant that answers the way the business does.

For businesses in India

Your model runs on your own servers or in a private cloud region you choose, including in India, so your data stays where your policies and your clients need it.

How we work

  1. ScanA discovery call. We map where your team's time and money go.
  2. BlueprintA written scope, timeline and fixed quote before any work starts.
  3. BuildYou see working progress every week and test it as we go.
  4. DeployWe launch, train your team and hand over the keys.
  5. EvolveWe monitor, fix and improve, so the system keeps getting better.

What it costs

Every project has a fixed price in rupees, agreed in writing after a discovery call, so there are no hourly surprises. You pay on one of our two payment plans, Elite or Normal, set out in our terms. See the payment plans.

After launch, the monthly plan is ₹24,100 a month and covers security updates, monitoring and fixes. You can cancel it at any time.

Anything that runs in accounts in your name, such as phone lines, AI model usage or a hosting provider, is billed by that provider at its own rates.

Questions

What is a private LLM?

A language model that runs on your own servers or private cloud and is trained on your documents. It answers in your terminology, and your data stays with you instead of going to a public AI service.

Can we have a private ChatGPT for our business?

Yes. Your staff type questions the way they would into ChatGPT, and the answers come from your own files, such as policies, contracts and manuals. It runs on your servers or in a private cloud region you choose, including in India, so no question or document is sent to a third-party AI.

How big a server does it need?

Often a modest one. Quantized to 4-bit, a model with 7 billion parameters needs about 4 GB, against about 14 GB at 16-bit, unquantized, which is why it runs well on modest hardware. 8-bit sits in between.

What is fine-tuning?

It is extra training on your own material, so an open-source model learns your terms, your documents and how your business replies. With LoRA or QLoRA only a small add-on set of weights is trained, not the full model.

How much does a private LLM cost?

One fixed price in rupees covers the project, agreed in writing after a discovery call about your documents, your hardware and what the model must answer. After launch, ₹24,100 a month on the monthly plan pays for security updates, monitoring and fixes.

Who owns the trained model?

Your business does. When every invoice for the project is paid, the weights trained on your data are yours; the open-source base model they build on keeps its own licence terms.

Talk to us

Send us a message and a person from the team will reply, usually the same day.

Other services