Production Requirements
Operating System Support
Supported Distributions
- Ubuntu
- Red Hat
- Debian
- Other
- Ubuntu 20.04 LTS (Focal)
- Ubuntu 22.04 LTS (Jammy)
- Ubuntu 24.04 LTS (Noble)
Hardware Requirements
CPU
- 16+ cores (32+ threads)
- Intel Xeon or AMD EPYC
- AVX2 instruction set support
- 3.0 GHz+ base clock
- 32+ cores
- Latest generation processors
- AVX-512 support
Memory (RAM)
- 32 GB - A reasonable starting point for moderate concurrency
- 64 GB+ - For high-volume deployments
Storage
- 500 GB NVMe SSD - Recommended starting point
- 1+ TB NVMe SSD - For high-volume or long-term log retention
- RAID 10 configuration recommended for production
Storage Performance
GPU (Optional)
GPU acceleration significantly improves processing speed: Supported GPUs:- NVIDIA T4 (16 GB)
- NVIDIA A10 (24 GB)
- NVIDIA A100 (40/80 GB)
- NVIDIA L4 (24 GB)
- NVIDIA L40 (48 GB)
- CUDA 11.8 or higher
- NVIDIA driver 520.61.05 or higher
- Compute capability 7.0+
- A GPU can substantially reduce analysis time for model-driven detection and OCR-heavy documents. The size of the improvement depends on the model, document mix, and CPU baseline, so benchmark before sizing around it.
- Recommended for high-volume deployments
Network Requirements
Bandwidth
- 10 Gbps - Recommended for production
- 10+ Gbps with redundancy - For high-volume enterprise deployments
Ports
Required network ports:Firewall Rules
Allow inbound traffic:Software Dependencies
Container Runtime
- Docker
- Kubernetes
- Podman
Required Version: 20.10+Installation:Docker Compose:
Database
PostgreSQL is required for metadata and audit logs: Version: 12+ Recommended: 14+ Storage Requirements:- 10 GB minimum
- 50 GB recommended
- Scales with audit log retention
Messaging
NATS is required for job messaging between services: Version: 2.9+ Memory needs scale with queued job volume; a small allocation is sufficient for typical deployments.High Availability Requirements
For production HA deployment:Compute
- Minimum: 3 nodes
- Recommended: 5+ nodes
- Load balancer: NGINX, HAProxy, or cloud LB
Database
- PostgreSQL: Primary + 2 replicas
- Replication: Streaming replication
- Automatic failover: Patroni or similar
Messaging
- NATS: Clustered (3+ nodes) for redundancy
Storage
- Shared storage: NFS, S3, or similar
- Replication: 3x minimum
- Backup: Daily with off-site replication
Cloud Provider Recommendations
AWS
Instance Types:- Minimum: t3.xlarge (4 vCPU, 16 GB)
- Recommended: c6i.4xlarge (16 vCPU, 32 GB)
- With GPU: g5.2xlarge (1x A10G GPU)
Azure
VM Sizes:- Minimum: Standard_D4s_v5 (4 vCPU, 16 GB)
- Recommended: Standard_D16s_v5 (16 vCPU, 64 GB)
- With GPU: Standard_NC8as_T4_v3 (1x T4 GPU)
Google Cloud
Machine Types:- Minimum: n2-standard-8 (8 vCPU, 32 GB)
- Recommended: n2-standard-16 (16 vCPU, 64 GB)
- With GPU: n1-standard-8 + 1x T4
Capacity Planning
Concurrent Jobs
Estimate required resources based on workload: These pairings are illustrative and unbenchmarked. The “concurrent jobs” column is a planning target, not a measured result — the worker-count guidance in Performance Tuning is a conservative CPU-bound floor and will suggest a lower number for the same hardware.Storage Sizing
Calculate storage needs:Benchmarking
Test your environment:Next Steps
Installation
Install Nvisy on-premises
Getting Started
Configure and start using Nvisy
