For non-SmithDB LangSmith services, see Configure LangSmith for scale.
Choose and scale a tier
The tiers are not throughput limits or prescribed configurations. CPU, memory, and cache size are per replica. Cache size is the volume each replica gets: the claim size on a network-attached disk, or theemptyDir limit on local SSD.
Choose the tier whose tested ingestion and query rates most closely match or exceed your expected steady-state load. To estimate your current trace volume, open Settings > Usage in LangSmith. See Granular billable usage.
If you cannot measure your load, start at small and scale up. Under-sizing shows up as slower ingestion and queries rather than data loss.
Apply the per-replica resources and starting replica counts from the selected tier. If the workload later outgrows that starting point, you can scale SmithDB compute horizontally.
SmithDB query, ingestion, and compaction-worker HPAs are enabled by default. They require Kubernetes Metrics Server or another
metrics.k8s.io provider. Verify availability with kubectl get --raw /apis/metrics.k8s.io/v1beta1.Scale with KEDA instead
KEDA scaling for SmithDB compaction workers requires Helm chart
0.17.0 or later.Baseline tiers
Configure resources with Helm
The chart defaults tosmall. Select small, medium, or large:
ephemeral-storage request and limit, which also sizes the generated emptyDir. Replica counts and autoscaling are configured separately.
An explicit component resources block replaces the tier’s CPU and memory. On 0.17 it does not change the cache size; pick a different tier or override deployment.volumes for that. For local SSD values, see Cache storage.
See SmithDB resource tiers for current values and chart details.
Example explicit resource configuration
Example explicit resource configuration
This example overrides the query component at medium-tier sizes.
- LangSmith 0.17
- LangSmith 0.16
The cache claim stays at the tier size. These values request no node storage.
Metastore capacity
The baseline tiers cover SmithDB Kubernetes workloads only. They do not include the PostgreSQL metastore. Use these starting points for a dedicated metastore:
Choose the nearest supported PostgreSQL instance shape from your provider and monitor database resource use and transaction latency during rollout.
Connect these docs to your agent of choice via MCP for real-time answers.

