Product

Investigate failing Kubernetes workloads in one place.

See which namespace and workload needs attention. Open its pod status, restarts, events, logs, requests, limits and routing without piecing them together from separate tools. Fortem runs locally using your existing kubeconfig.

or install yourself

Local install · read-only first · no Helm chart

Want to look first? Open the interactive demo →
Runs
On your workstation
Connects
Through kubeconfig
Starts
Read-only
01

Start with an evidence-backed incident brief

Fortem normalizes the affected workload, recent change, Kubernetes events, safe log-pattern counts and available traffic metrics into a compact local assessment.

  • Observed facts remain clickable
  • Ranked hypotheses cite their evidence
  • One recommended read-only next check
02

See environments together

Fortem treats a namespace as an environment in the first release, then normalizes workload, pod, resource, and event data into one scan-friendly view.

  • Context and namespace scope
  • Ready, progressing, and degraded states
  • Explicit stale and permission states
03

Move from a symptom to the workload

Open an environment, then inspect Deployments and StatefulSets, their pods, container images, rollout details, restart counts, and waiting or termination reasons.

Investigate an OOMKilled container →
  • Workload identity and readiness
  • Pods, restarts, and reasons
  • Images and available change information
04

Put requests beside usage

Requests and limits come from Kubernetes. Actual CPU and memory usage appears only when a compatible metrics source is available; missing telemetry is never rendered as zero.

  • Requests and limits
  • Node capacity, workload placement, and Karpenter NodePool fill
  • Optional metrics-server usage with freshness
05

Keep the evidence in reach

Related Kubernetes events and container logs live beside workload details, with filtering and clear source labels. Readiness is shown as a fact, not a promise that the application is healthy.

  • Namespace and workload events
  • Container log selection and filtering
  • Permission-aware errors
06

Trace the route; add traffic signals when available

Inspect Ingress, Service, and backend relationships from Kubernetes. HTTP request, error, and latency metrics require a separate compatible metrics source and retain their measurement point and time window.

Understand what request counters actually measure →
  • Ingress → Service → backend
  • Optional request and error metrics
  • No invented lost-request percentage
07

Make changes deliberately

Fortem starts read-only. Supported restart and scale actions require an explicit startup flag, an exact object confirmation, and a Kubernetes authorization check before the API accepts them. An assessment never authorizes a mutation.

  • Read-only by default
  • Exact context, namespace, kind, and name
  • SelfSubjectAccessReview before mutation
08

See where GPU capacity is reserved

Nodes show scheduler-visible whole-GPU and MIG capacity beside Pod requests. Fortem Pro can project allocation cost from a rate you provide for each resource type; it does not claim measured GPU utilization or a cloud bill.

  • GPU allocation stays in Free
  • Pro rate scenario requires your USD/unit-hour assumption
  • Missing Node or Pod access stays unknown

Data contracts

Useful without pretending every cluster has every signal.

Optional integrations enrich the same interface. Their absence becomes “unavailable,” with a reason—not a misleading zero.

Fortem rules

Local

The default incident engine is deterministic, runs offline, and records evidence, caveats and provenance. It does not send cluster data to Fortem.

Kubernetes API

Base

The same standard API path supports EKS, GKE, AKS, kind and k3s: contexts, namespaces, workloads, pods, nodes, services, ingresses, events, logs, requests and limits.

metrics-server

Optional

Current CPU and memory usage. The rest of the interface continues to work when it is unavailable.

Prometheus HTTP metrics

Optional

Five-minute request count, RPS, 4xx, 5xx, and p95 latency from a compatible ingress-nginx, Traefik or Istio profile.

Karpenter

Optional

On clusters with a readable Karpenter API, NodePools show provisioned CPU and memory capacity against configured limits. Without it, Fortem keeps the base node view and labels the source unavailable.

Cost estimates

Optional

Fortem has a limited AWS/GCP/Azure node-rate estimate, not provider billing. Pro can model GPU unit-rate cost when you supply the price. The GPU figure is separate and must not be added to the node estimate.

See the investigation flow on your machine.

Install with your coding agent, start in synthetic demo mode, then select a kubeconfig context you control.

or install yourself

Local install · read-only first · no Helm chart