Skip to main content

PrivateGPT

PrivateGPT provides an API and browser workbench for private AI applications and document-assisted conversations. Moltern deploys the protected workbench, persists its assigned data and lets it connect to a private Ollama service in the same environment.

Private deployment does not remove the need for access controls, retention rules and careful model selection. Treat imported documents and generated answers according to their business sensitivity.

Upstream production status

The current PrivateGPT workbench identifies itself as Not for Production. Use this catalog profile for controlled evaluation and internal experiments, then complete your own security, recovery, load and release review before placing business-critical data or workflows on it.

Before You Start​

You need:

  • a Moltern workspace and environment;
  • permission to create services and private connections;
  • an API username and strong password;
  • at least 200 mCPU and 1 GiB of available runtime capacity;
  • workspace storage for collections, indexes and imported documents;
  • a private Ollama service with a downloaded compatible model.

Store the PrivateGPT credential in your organisation's password manager. It is used by both the API and browser workbench and is separate from your Moltern identity.

Deploy PrivateGPT​

  1. Open Services and select PrivateGPT.
  2. Enter a unique service name.
  3. Select the project and environment.
  4. Review the initial CPU and memory values.
  5. Enter the API username and password.
  6. Select Preview deploy.
  7. Review the workload and storage impact.
  8. Select Confirm deploy and wait for Running.

PrivateGPT deployment form with protected API credentials

PrivateGPT deployment preview

This profile has no managed model dependency. Deploy Ollama separately, pull a compatible model and then grant PrivateGPT an explicit private connection.

Connect Private Ollama​

Deploy Ollama in the same environment and pull a model before creating the connection.

  1. Open PrivateGPT in Moltern.
  2. Open Access or Private service connections.
  3. Select Connect service.
  4. Choose the Ollama service.
  5. Confirm the connection.
  6. Wait for PrivateGPT to return to Running.

The connection grants only this PrivateGPT workload access to the selected Ollama service. It does not publish Ollama or grant access to other workloads.

Configure The Workbench​

  1. Select Open service from the PrivateGPT page.
  2. Confirm the generated address redirects from / to the /ui workbench.
  3. Enter the PrivateGPT service address in the onboarding form.
  4. Expand authentication and enter the API username and password.
  5. Choose a collection name.
  6. Select Check connection.
  7. Finish onboarding after the workbench confirms API access.

PrivateGPT protected workbench onboarding

The service address alone does not authorize API access. Keep the credential protected and rotate it if it is exposed.

Verify Private Inference​

  1. Confirm the model list contains a model from the attached Ollama service.
  2. Start a new workbench conversation.
  3. Send a small controlled prompt.
  4. Confirm a visible assistant response appears.

PrivateGPT workbench with a visible private Ollama response

Model discovery is not a complete test. A real prompt must return a visible response. The first prompt can be slower while Ollama loads model weights.

Work With Documents​

Before importing business data:

  1. create a collection for one defined purpose;
  2. start with a small representative document;
  3. confirm the parsed content and retrieval result;
  4. verify answers include enough evidence for the intended workflow;
  5. define retention and deletion ownership;
  6. validate export or recovery procedures.

Do not assume a fluent answer is correct. Evaluate retrieval quality with a controlled question whose answer is present in the source. Review how the selected PrivateGPT release handles chunking, embeddings and citations before using it for decisions.

The current production E2E certifies the protected workbench and attached-model inference. Document ingestion, citation quality and restore are not yet part of that release gate.

API Access​

Use the same protected credential for an approved application or agent that needs the PrivateGPT API. Grant the workload an explicit connection and avoid placing credentials in source code, frontend bundles or prompts.

For each integration:

  • use the private service address available to the authorised workload;
  • test model discovery before sending messages;
  • set request timeouts appropriate for local inference;
  • handle non-success responses without logging secrets or source documents;
  • remove the connection when the integration is retired.

Data And Persistence​

PrivateGPT stores its local workbench state, collections and supported document artifacts in its assigned workspace path. The runtime sees only that service path, not the complete team filespace.

Moltern does not request a separate cloud disk for this profile. Storage usage is measured under the PrivateGPT workload.

Production validation performs Stop service and Start service, repeats workbench authentication and requires a second visible response from the attached model.

PrivateGPT workbench after runtime replacement

Stop/Start is a runtime replacement test, not an independent backup restore.

Capacity And Metering​

The validated baseline is:

ResourceValidated value
Instances1
CPU request200 mCPU
Memory request1,024 MiB
Dedicated volume0 GiB

PrivateGPT storage and Ollama model storage are measured separately. Review both workloads when estimating complete cost.

Increase PrivateGPT capacity for document parsing, indexing or API pressure. Increase Ollama capacity for model loading and inference pressure.

Operate PrivateGPT​

Use the Moltern service page for:

  • Overview to open the workbench and review health;
  • Live Logs to diagnose API, indexing and provider failures;
  • Capacity to adjust CPU and memory;
  • Access to add or revoke private service connections;
  • Settings to review protected configuration;
  • Stop service and Start service to replace the runtime without deleting workspace data.

Delete PrivateGPT​

  1. Export documents, indexes or evidence that must be retained.
  2. Remove approved application and agent connections.
  3. Open the PrivateGPT service in Moltern.
  4. Select Delete Service.
  5. Select Delete stored data only when the service path may be removed.
  6. Complete protected confirmation.
  7. Confirm the service, address, assigned path and active allocation are gone.

Deleting PrivateGPT does not delete a separately created Ollama service.

Frequently Asked Questions​

Does PrivateGPT include an AI model?​

No. Attach a private Ollama service containing a compatible model before using the workbench for inference.

Is a document private just because PrivateGPT is self-hosted?​

Not automatically. Limit workspace membership, service attachments and credentials, and define retention and deletion rules for imported content.

Does Stop/Start prove document recovery?​

No. It proves supported state survives runtime replacement. Test a separate export and restore process before relying on PrivateGPT for critical records.

Troubleshooting​

SymptomWhat to check
Workbench connection check failsConfirm the service address, API username and password. Verify the service is Running before retrying.
No models appearConfirm Ollama is Running, contains a model and is connected in the same environment. Wait for rollout completion.
A model is visible but chat failsTest the model directly in Ollama and review PrivateGPT and Ollama logs. Confirm it fits available memory.
First response is slowAllow for cold model loading and compare a second prompt. Try a smaller model when necessary.
A document cannot be indexedCheck file support, parsing errors, workspace storage and embedding configuration for the installed release.
Retrieval answers ignore the sourceVerify ingestion completed, inspect the collection and test a controlled question with an explicit source answer.
Data is missing after deletionDelete stored data is destructive. Restore only from an independently tested export or backup.

Validated Scope​

The production customer E2E covers UI deployment, protected API/workbench access, private Ollama attachment, model discovery, direct completion, browser workbench onboarding, a visible assistant response, runtime replacement, post-restart inference, point-in-time metering and protected cleanup.

Document ingestion, retrieval quality, large collections, concurrent load, GPU behavior, high availability, independent restore and elapsed invoice reconciliation remain outside this validation.

Official Resources​