Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

iTechGuides is reader-supported. When you buy through links on our site, we may earn an affiliate commission. As an Amazon Associate I earn from qualifying purchases. Learn more

For a Node.js API, useful health monitoring separates two questions: should this instance receive traffic, and should its process keep running? In Kubernetes, readiness failures remove an instance from service, while liveness failures can trigger a restart. To judge a new release, combine those probe outcomes with API error rate and latency—then compare them with a service-specific baseline rather than applying a universal rollback threshold.

How do I add a health check endpoint to my Node.js API?

Expose lightweight HTTP endpoints that answer distinct operational questions, then configure your orchestrator to probe them. In Kubernetes, the application endpoints you define are separate from the Kubernetes API server’s own /livez and /readyz endpoints; the API server’s older /healthz endpoint is deprecated. See Kubernetes API health endpoints.

Give each endpoint one purpose

  • Liveness: confirm that the application process can respond and make progress. A failed liveness probe can cause Kubernetes to restart the container.
  • Readiness: indicate whether this instance can safely serve requests. A failed readiness probe removes the instance from traffic while it remains unready; it should not restart the process just because the instance is temporarily unable to serve.
  • Startup: use a startup probe when initialization is long or unpredictable. It delays liveness and readiness probing until startup succeeds.

Kubernetes documents these probe behaviors and configuration in Configure Liveness, Readiness and Startup Probes and Liveness, Readiness, and Startup Probes.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Keep liveness independent of external services

A liveness endpoint should normally establish that the process itself is alive, not require a database, network service, or other dependency to be healthy. If a shared dependency goes down and every replica’s liveness check fails, repeated restarts can compound the incident rather than repair the dependency. AWS recommends keeping liveness and readiness distinct and avoiding external dependencies in liveness checks; see AWS guidance on probes and load balancer health checks.

#1 Best Overall
Tecmojo 12U Open Frame Network Rack for IT & AV Gear, AV Rack Floor Standing or Wall Mounted,with 2 PCS 1U Rack Shelves & Mounting Hardware,Network Rack for 19" Networking,Audio and Video Device
  • 【Powerful Load-bearing】12U Network Rack Open Frame is constructed from durable cold rolled steel; Rack shelf supports enhance stability, wall-mounted capacity of 130lbs, the ground-mounted up to 260lbs
  • 【Considerate Designs】Open-frame layout, including a top panel adding space, anti-slip shelf stops fixing devices and compatible racks for stack and expansion to meet requirements of home server rack
  • 【Complete Accessories】A 12U open frame server rack, two ventilated shelves, four shelf stops, four velcro straps and a set of equipment mounting screws
  • 【Versatile Application】Ideal for space-efficient multi-device setups in warehouses, retail, classrooms, offices and more; Excellent choices as AV Rack/IT Rack
  • 【Effortless Setup】 Network Rack includes hardware, a comprehensive manual, mounting hole drilling template and an online assembly video to simplify setup

Use dependency-aware readiness with care

Readiness can reflect whether an instance is able to handle requests, including relevant dependency state. But a check that marks every replica unready when a shared database is unavailable can remove all application capacity from traffic. Decide whether the endpoint reflects instance-specific ability to serve, and consider how the service behaves during a common dependency outage. Amazon EKS discusses application availability and dependency behavior in Running highly-available applications.

Configure probes to tolerate normal behavior

Set probe timing and failure thresholds to allow for expected startup and brief load variation. Overly aggressive settings can misclassify normal behavior as failure. Probe frequency and exec-based checks can also add CPU overhead, particularly at high pod density. Kubernetes explains probe configuration and its operational effects in its probe documentation.

For Node.js services on Kubernetes, Lightship is one package that describes readiness, liveness, startup checks, and graceful shutdown: npm: lightship. It is an implementation option, not a requirement; the key is to preserve the probe semantics above.

What’s actually slowing this PC down?

Pick the symptom - the matching free tool is one click away.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.
Rank #2
Sale
StarTech 42U 4-Post Open Frame Rack, 19in, 22-40in, 1323lb/600kg
  • ADJUSTABLE DEPTH: 4-Post 42U open frame server rack with 4 vertical rails and adjustable mounting depth 22" to 40" (56,0cm to 101,7cm); Compatible with various servers / switches / data / AV and other IT equipment; EIA/ECA-310-E Compliant
  • EASY ASSEMBLY: Mobile network rack with easy-to-follow assembly instructions and online video; Compact flat-pack shipping to avoid damage and facilitate installation; Total product height of 80.3in (204 cm) with casters, 78in (198cm) without casters
  • COLD ROLLED STEEL: Durable 4 Post 19in open frame rack designed for ventilation with 42U mounting height and 1320lb (600kg) weight capacity (stationary); 3 install options included: casters, levelling feet, or base-plate to secure rack to the floor
  • HARDWARE INCLUDED: Rolling computer/data rack includes cage nuts and screws to mount equipment, easy to read Units (U) and depth adjustment markings, cable management hooks for organization, and required assembly tools
  • THE IT PRO'S CHOICE: Designed and built for IT Professionals, this 42U rack is backed for 2-years, including free lifetime 24/5 multi-lingual technical assistance

What is the difference between readiness and liveness?

Signal Question it answers Typical action External dependency check?
Readiness Can this instance accept traffic now? Remove it from traffic while it is unready Possibly, but avoid taking every replica out for one shared outage
Liveness Is the process stuck or unable to make progress? Restart the container after configured failures Normally no; keep it independent of external dependencies
Startup Has this instance completed initialization? Defer the other probes until startup succeeds Only as appropriate to startup behavior

These are Kubernetes probe semantics, not interchangeable names for a generic “healthy” endpoint. AWS also advises using distinct readiness and liveness checks; its guidance is at Configure probes and load balancer health checks.

Which four signals can identify a release that needs attention?

Use four complementary signals during a rollout. The first two are operational recommendations; Kubernetes and AWS define the behavior of the readiness and liveness probes, but do not prescribe universal error-rate, latency, capacity, or restart thresholds.

1. API error rate

Compare the new revision’s error rate with the pre-deployment baseline and the expected traffic profile. A sustained increase that begins with the rollout is a reason to investigate or pause, especially if other signals also worsen. Define which responses count as errors for your API; there is no single threshold established by the cited Kubernetes or AWS guidance.

Rank #3
VEVOR 12U Open Frame Server Rack, 23-40 in Adjustable Depth, Free Standing or Wall Mount Network Server Rack, 4 Post AV Rack with Casters, Holds All Your Networking IT Equipment AV Gear Router Modem
  • Adjustable Depth: 23-40'' adjustable depth is used for servers and network equipment, ensuring enough space for AV equipment, components, and cabling, while allowing you to access ports and equipment from multiple sides.
  • Strong Load Capacity: Ground-Mounted Load Capacity: 500 lbs, Wall-Mounted Load Capacity: 150 lbs. The av rack is made of carbon steel for better weldability performance and can help save space while meeting your need to place multiple devices.
  • User-friendly Design: Ergonomic design makes the open frame av rack easier to use. The additional top panel is able to place other items with more available space. Roller design moves anywhere and anytime, is convenient, and is more energy-saving.
  • Complete Accessories: We provide the accessories you need, including 2 x Pallets, 145 x M5*10 Cross Head Screws, 4 x Casters, 4 x M10*50 Expansion Screws,10 x M6*12 Cage Nuts, 1 x Grounding Wire, 1 x User Manual.
  • Wide Application: The server rack wall mount maximizes the use of available space, suitable for retail venues, classrooms, offices, and other places where space is limited.

2. API latency

Watch for a sustained latency regression against the service’s normal distribution or user-facing objective. Avoid pausing or rolling back because of one slow request. Choose the observation window and acceptable change based on the service’s own baseline; the cited sources do not set a universal latency limit.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

3. Readiness failures or loss of ready capacity

Track the number of instances that remain ready and whether failures cluster in the new revision. Readiness is a traffic-eligibility signal: Kubernetes uses it to decide whether a pod should receive traffic, rather than to restart it solely for being temporarily unable to serve. A falling ready count can therefore reveal reduced capacity even before a liveness failure appears.

4. Liveness failures or rising restarts

Watch for liveness probe failures and a rise in restarts, particularly when they concentrate in the new revision. These can indicate that an application is stuck or cannot make progress. Check whether a shared database, network, or infrastructure issue explains the pattern before attributing it to the release; dependency-driven liveness failures can create restart loops.

Rank #4
AxcessAbles 12U Network Rack with Wheels - 500lb Capacity, 18" Depth | 19-Inch Open Frame AV Rack Case with 3” Caster Wheels | Screws, Spacer, Tool Included
  • Universal 19” Rack Mount Compatibility – Perfect for pro audio, video, IT, and network gear. Compatible with mixers, routers, patch panels, servers, power amps, and more.
  • Heavy-Duty Load Capacity – Built to support up to 550 lbs. Ideal for studio gear, DJ setups, server equipment, and AV components that demand serious stability.
  • Robust Steel Frame & Design – Made with 1.5mm thick steel and weighs 36 lbs for maximum durability, reduced vibration, and long-term reliability in any setting.
  • Mobile & Secure – Preinstalled with 3” industrial-grade caster wheels (lockable), making it easy to move and position your rack exactly where you need it.
  • All-In-One Setup Kit Included – Comes with 34 rack screws (5mm & 6mm), a 1U blank spacer, and an assembly tool—ready for fast installation out of the box.

When should I roll back a deployment?

Pause or roll back when the evidence indicates that the new revision is materially degrading service, not simply because one probe or request failed. A practical rollout controller can evaluate all four signals over an observation window chosen for the API’s traffic and recovery characteristics.

  1. Compare error rate with the pre-deployment baseline and expected traffic profile.
  2. Compare latency with the service’s normal distribution or objective, using a sustained window rather than an isolated slow request.
  3. Check ready capacity and determine whether readiness failures are concentrated in the new revision.
  4. Check liveness failures and restart growth, and look for a shared dependency or infrastructure incident.
  5. Pause the rollout or roll back when multiple signals worsen together after the new revision starts and the impact is consistent with a release regression.

This is an operational decision framework, not a rollback formula prescribed by Kubernetes or AWS. Set thresholds, severity, and observation windows against your service’s baseline and rollout policy. If a shared external dependency is down, a readiness cascade or restart loop may worsen the impact; diagnose that incident before treating every affected pod as a bad release.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.
Independent reader supportYour contribution helps us test, update, and keep practical guides available for everyone.Support on Ko-Fi

How can I distinguish an application regression from a cluster problem?

Correlate probe failures with revision, node conditions, and events. If problems cluster on the new revision across healthy nodes, the release is a stronger suspect. If unrelated workloads on one node are also affected, investigate node health and resource pressure before rolling back the application.

Best Value
VEVOR 9U Open Frame Server Rack, 23''-40'' Adjustable Depth, Free Standing or Wall Mount Network Server Rack, 4 Post AV Rack with Casters, Holds All Your Networking IT Equipment AV Gear Router Modem
  • Adjustable Depth: Depth adjustable from 23" to 40", this open frame server rack accommodates servers and network equipment while providing ample space for A/V gears and cable management. Enjoy easy access to ports and devices from multiple angles.
  • High Weight Capacity: Supports up to 300 lbs on the floor (200 lbs when adjusted to maximum depth) and 200 lbs when wall-mounted (depth cannot be adjusted in wall-mounted mode). Made from carbon steel for superior welding performance and durability, this open frame rack is designed to save space while accommodating multiple devices.
  • User-Friendly Design: Designed with your convenience in mind, this open frame server rack features an top shelf for extra storage and improved space utilization. The rolling casters let you move it effortlessly wherever you need it, making setup and movement a breeze.
  • Widely Applicable: Maximize your space with this adaptable open frame server rack, designed to make the most of every inch. Ideal for retail spots, classrooms, offices, and any area where space is at a premium, it delivers practical solutions for your storage needs.
  • Everything You Need: Our open-frame rack comes with fully equipped accessory kit for easy setup and secure installation: 2 x Trays, 4 x Casters, 1 x set of Screws, 16 x M6*12 Cage Nuts, 1 x Grounding Wire, 1 x Internal & External Hex Wrenches, and 1 x User Manual.

Kubernetes node status includes readiness and resource-pressure conditions; see Node Status. Amazon EKS documents node signals and monitoring-agent events in Detect node health issues with the EKS node monitoring agent. Its automatic node repair responds to specified node conditions, not every resource-pressure condition, as described in Detect node health issues and enable automatic node repair. Node repair addresses infrastructure health; it is not the same action as rolling back an application release.

For visibility beyond in-cluster probes, combine application and infrastructure monitoring with external uptime checks. External checks can help establish whether users can reach the API, while probe alerts and cluster events help explain why capacity changed.

Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.