Uwazi Document Collections - Managed Support

AWS Applications

organise, search and publish structured collections of documents and records

Base
Hardened build
minimal ports, security patches applied at build time
Access
Unique credentials
generated on first boot, readable only by root
Verified
Boots working
services pass a health gate before release
Support
24/7, 365 days
by email and live chat, 24 hour response SLA

Overview

Uwazi is an open source platform for building searchable, structured collections of documents and records. Human rights defenders, investigative journalists, legal teams and researchers use it to turn scattered evidence, case files and archives into a fully searchable library, with custom entity templates, relationships between records, and full text search over every uploaded file.

Why the cloudimg image

cloudimg ships Uwazi production ready: the documented default administrator password is overwritten during the build and rotated to a unique per instance password on first boot, MongoDB, Elasticsearch and Redis are bound to the loopback interface, two dedicated data volumes keep the database and uploaded documents on their own resizable disks, and every deployment is paired with a step by step guide and backed by 24/7 cloudimg support.

Common uses

  • Catalogue case files, testimonies and evidence into structured, related records
  • Build a private working archive or a curated public library from primary source documents
  • Run full text search over large collections of PDFs and scanned documents

Key features

  • Production-ready Uwazi 1.229 built from the official release on native Node.js 20 with dockerised MongoDB 7.0 and Elasticsearch 8.18 (ICU plugin) behind an nginx reverse proxy. Define entity templates, upload documents with OCR-aware full text search, relate records and publish a curated public library. Two independently resizable EBS volumes separate your data from the OS disk for straightforward scaling as collections grow.
  • Secure by default with zero shared credentials: the documented default administrator password is overwritten during the build and rotated to a unique per-instance password on first boot, stored in a root-only file. MongoDB and Elasticsearch are bound to the loopback interface only, never exposed to the network. No two deployments share authentication material, and a fresh session secret is generated on every first boot.
  • 24/7 technical support from cloudimg engineers by email and live chat with a one-hour average response for critical issues. Support covers deployment guidance and sizing, MongoDB and Elasticsearch administration and backup, TLS termination with Let's Encrypt or custom certificates, Uwazi upgrades, and performance tuning - so your team focuses on building the collection rather than operating the infrastructure.

See it running

Real screenshots taken while testing this image against its deployment guide.

Uwazi Document Collections - Managed Support screenshot 1 Uwazi Document Collections - Managed Support screenshot 2 Uwazi Document Collections - Managed Support screenshot 3 Uwazi Document Collections - Managed Support screenshot 4

Description

This is a repackaged open source software product wherein additional charges apply for cloudimg support services.

## Build and Publish Structured Document Collections

Uwazi is an open source platform for organising, searching and publishing collections of documents and records. Human rights defenders, investigative journalists, legal teams and researchers use it to turn scattered evidence, case files and archives into a structured, fully searchable library. This image delivers Uwazi built and configured on a native Node.js application with dockerised MongoDB and Elasticsearch, so your team can define its first template and upload its first document within minutes of launch.

## Who This Is For

Human rights and legal organisations cataloguing case files, testimonies and evidence into structured records that relate to one another, with full text search across every uploaded PDF. Investigative and research teams building a private working archive or a curated public library from primary source documents. Librarians and archivists who need a self-hosted collections platform inside their own VPC without spending ops cycles on installation, search configuration or credential rotation.

## What Makes Uwazi Different

  • Define your own data model with entity templates, properties, thesauri and relationships, so records match your domain rather than a fixed schema
  • Documents become searchable knowledge through OCR-aware full text search over every uploaded file, powered by Elasticsearch
  • Relate records to one another to map connections between people, events, cases and documents
  • Publish a curated public library or keep the whole collection private, with granular control over what is shared
  • Rich document viewer with in-context references, table of contents and text selection

## Application Stack

  • Uwazi 1.229 built from the official production release on Node.js 20
  • nginx reverse proxy on ports 80 and 443 with a self-signed certificate generated per instance
  • MongoDB 7.0 as a single node replica set, published to the loopback interface only
  • Elasticsearch 8.18 with the ICU analysis plugin for multilingual full text search, published to the loopback interface only
  • Redis 6.0 backing the Uwazi job queue, published to the loopback interface only
  • Reach the platform through nginx; the application, database, search and queue ports need no inbound rule and stay closed

## Secure By Default

Unlike images that ship a shared or documented password, this one carries no known credential:

  • Uwazi's default administrator password is overwritten during the build, so the image ships with no usable login
  • On first boot an administrator account is rotated to a per-instance password written to a root-only file
  • A fresh session secret is generated on every first boot
  • MongoDB and Elasticsearch are bound to the loopback interface, never exposed to the network
  • No shared credentials exist anywhere in the image

## Storage Layout

Two dedicated EBS volumes ship with the image, separate from the operating system disk and each independently resizable: the Docker data root holding the MongoDB datastore and the Elasticsearch index, and the application tier holding the Uwazi bundle and your uploaded documents. Both are mounted by filesystem UUID.

## Getting Started

1. Launch the image on an m5.large instance or larger, with ports 80 and 443 open in your security group

2. Connect over SSH and read the generated administrator password from the root-only credentials file

3. Browse to the instance address and sign in as the administrator

4. Create your first entity template in Settings, then upload a document and watch it become searchable

5. Relate records, build thesauri and decide what to publish to a public library

## Recommended Instance Sizing

  • Minimum: m5.large (2 vCPU, 8 GB RAM) for small collections and evaluation
  • Recommended: m5.xlarge (4 vCPU, 16 GB RAM) for active collections with heavy indexing
  • Storage: grow the data volume as your document library grows
  • Networking: ports 80 and 443 for browsers, port 22 for administration

## 24/7 cloudimg Support

Every deployment is backed by cloudimg engineers available around the clock by email and live chat. Support covers:

  • Deployment guidance and instance sizing
  • MongoDB and Elasticsearch administration, backup and restore
  • nginx reverse proxy and TLS termination with Let's Encrypt or your own certificates
  • Uwazi upgrades and migration assistance
  • Performance tuning and troubleshooting

Critical issues receive a one-hour average response time.

## Ready to Evaluate?

Launch an instance to evaluate Uwazi with your team and terminate at any time with no long-term commitment. To discuss sizing for your specific collection or migration planning, email support@cloudimg.co.uk with the subject line "Uwazi Consultation".

Related technologies

document collectionsdocument managementhuman rightsfull text searchdigital archivecase managementknowledge baseself hostedelasticsearchmongodbevidence managementOCR searchresearch archivelegal documents