The Helicone Helm chart is available for enterprise customers. Schedule a call to discuss access and support.
Overview
The Helicone Helm chart deploys the complete stack on Kubernetes with:- Scalable architecture: Horizontal pod autoscaling for high-traffic workloads
- High availability: Multi-replica deployments with health checks
- Observability: Integrated Grafana and Prometheus monitoring
- GitOps support: Optional ArgoCD for continuous deployment
- Infrastructure as Code: Terraform modules for AWS resources
Architecture
The Helm deployment consists of modular charts:- helicone-core: Application components (web, jawn, workers)
- helicone-infrastructure: Database services (PostgreSQL, ClickHouse, Redis, MinIO)
- helicone-monitoring: Observability stack (Grafana, Prometheus, Beyla)
- helicone-argocd: GitOps continuous deployment
Prerequisites
Required Tools
- kubectl - Kubernetes command-line tool
- Helm 3.0+ - Kubernetes package manager
- AWS CLI (for AWS deployments) - Configured with appropriate permissions
- Terraform (optional) - For infrastructure provisioning
Kubernetes Cluster Requirements
- Kubernetes version: 1.24+
- Node resources:
- Minimum: 3 nodes with 4 CPU / 16GB RAM each
- Recommended: 5+ nodes with 8 CPU / 32GB RAM each for production
- Storage: Dynamic volume provisioning (e.g., AWS EBS CSI driver)
- LoadBalancer: Cloud provider load balancer or ingress controller
Installation
Option 1: Helm Compose (Recommended)
Deploy all components with a single command:- helicone-core (web, jawn, workers)
- helicone-infrastructure (databases, caching, storage)
- helicone-monitoring (Grafana, Prometheus)
- helicone-argocd (GitOps)
Option 2: Manual Helm Installation
Install components individually for granular control:Verify Deployment
Running state with Ready 1/1.
Configuration
values.yaml Structure
The main configuration file controls all deployment aspects:secrets.yaml Structure
AWS Infrastructure (Optional)
Use Terraform to provision AWS resources for production deployments.EKS Cluster
Aurora PostgreSQL
S3 Buckets
Accessing Services
Web Dashboard
values.yaml.
Grafana Monitoring
- Request throughput and latency
- Service health and resource usage
- Database performance metrics
- Error rates and alert triggers
ArgoCD GitOps
Scaling Considerations
Horizontal Pod Autoscaling
The Helm chart includes HPA configurations:Database Scaling
PostgreSQL:- Use AWS Aurora with read replicas
- Configure read/write splitting in application
- Deploy as a cluster with multiple shards
- Configure replication for high availability
- Use Redis Cluster mode or AWS ElastiCache
- Configure sentinel for automatic failover
Storage Scaling
S3/MinIO:- Use lifecycle policies to archive old data
- Consider tiered storage for cost optimization
- Implement time-based partitioning
- Configure TTL policies for old data
Production Best Practices
High Availability
- Deploy core services with
replicaCount: 3+ - Use pod anti-affinity to spread replicas across nodes
- Configure pod disruption budgets
- Enable database replication
Security
- Enable Network Policies to restrict pod-to-pod traffic
- Use Kubernetes Secrets with encryption at rest
- Integrate with AWS Secrets Manager or HashiCorp Vault
- Configure RBAC for least-privilege access
- Enable Pod Security Standards
Monitoring & Alerts
- Configure Prometheus alert rules for critical metrics
- Set up PagerDuty/Opsgenie integration
- Monitor database connection pools
- Track request latency and error rates
Backup & Disaster Recovery
Troubleshooting
Pods not starting
Database connection issues
Ingress/LoadBalancer not working
Performance issues
Upgrading
Helm Chart Updates
Database Migrations
Migrations run automatically as a Kubernetes Job before deployments start. Monitor:Uninstalling
Remove all components:
Getting Enterprise Support
The Helicone Helm chart and Kubernetes support are available for enterprise customers:- Helm chart access: Private repository with production-ready configurations
- Terraform modules: AWS infrastructure as code
- Migration assistance: Help moving from Docker to Kubernetes
- Architecture review: Optimize your deployment for scale and cost
- Priority support: Direct engineering support via Slack/Discord
Next Steps
- Configure application to use Kubernetes endpoints
- Set up monitoring dashboards and alerts
- Implement backup and disaster recovery procedures
- Review API documentation for integration