Kueue
Kueue is a set of APIs and controller for job queueing. It is a job-level manager that decides when a job should be admitted to start (as in pods can be created) and when it should stop (as in active pods should be deleted). Read the overview and watch the Kueue-related talks & presentations to learn more. Job management: Support job queueing based on priorities with different strategies: StrictFIFO and BestEffortFIFO. Advanced Resource management: Comprising: resource flavor fungibility, Fair Sharing, cohorts and preemption with a variety of policies between different tenants. Integrations: Built-in support for popular jobs, e.g. BatchJob, Kubeflow training jobs, RayJob, RayCluster, JobSet, plain Pod and Pod Groups. System insight: Built-in prometheus metrics to help monitor the state of the system, and on-demand visibility endpoint for monitoring of pending workloads. AdmissionChecks: A mechanism for internal or external components to influence whether a workload can be admitted. Advanced autoscaling support: Integration with cluster-autoscaler's provisioningRequest via admissionChecks. All-or-nothing with ready Pods: A timeout-based implementation of All-or-nothing scheduling. Partial admission and dynamic reclaim: mechanisms to run a job with reduced parallelism, based on available quota, and to release the quota the pods complete. Mixing training and inference: Simultaneous management of batch workloads along with serving workloads (such as Deployments or StatefulSets) Multi-cluster job dispatching: called MultiKueue, allows to search for capacity and off-load the main cluster. Topology-Aware Scheduling: Allows to optimize the Pod-to-Pod communication throughput by scheduling aware of the data-center topology. API version: v1beta2, respecting Kubernetes Deprecation Policy. Up-to-date documentation. Test coverage: Unit test testgrid. Integration tests: Baseline suite shard-0 shard-1 shard-2 MultiKueue suite testgrid. E2E tests: Baseline suites for Kubernetes 1.34 1.35 1.36 on Kind. Extended suites for Kubernetes on Kind: 1.34: shard-0 shard-1 shard-2 1.35: shard-0 shard-1 shard-2 1.36: shard-0 shard-1 shard-2 TAS: Baseline suite testgrid. Extended suite: shard-0 shard-1 Sequential tests: Baseline suites: shard-0 shard-1 Extended suites: shard-0 shard-1 E2E Cert Manager test testgrid. DRA test testgrid. MultiKueue: Baseline suite testgrid. Extended suites: shard-0 shard-1 MultiKueue DRA test testgrid. Scheduling performance tests: Baseline suite testgrid. TAS suite testgrid. Large-Scale suite testgrid. Scalability verification via performance tests. Monitoring via metrics. Security: RBAC based accessibility. Stable release cycle (2-3 months). Adopters running on production. Based on community feedback, we continue to simplify and evolve the API to address new use cases.
View Kueue on GitHub
Kueue is a set of APIs and controller for job queueing. It is a job-level manager that decides when a job should be admitted to start (as in pods can be created) and when it should stop (as in active pods should be deleted).
Read the overview and watch the Kueue-related talks & presentations to learn more. Job management: Support job queueing based on priorities with different strategies: StrictFIFO and BestEffortFIFO. Advanced Resource management: Comprising: resource flavor fungibility, Fair Sharing, cohorts and preemption with a variety of policies between different tenants. Integrations: Built-in support for popular jobs, e.g. BatchJob, Kubeflow training jobs, RayJob, RayCluster, JobSet, plain Pod and Pod Groups. System insight: Built-in prometheus metrics to help monitor the state of the system, and on-demand visibility endpoint for monitoring of pending workloads. AdmissionChecks: A mechanism for internal or external components to influence whether a workload can be admitted. Advanced autoscaling support: Integration with cluster-autoscaler's provisioningRequest via admissionChecks. All-or-nothing with ready Pods: A timeout-based implementation of All-or-nothing scheduling. Partial admission and dynamic reclaim: mechanisms to run a job with reduced parallelism, based on available quota, and to release the quota the pods complete. Mixing training and inference: Simultaneous management of batch workloads along with serving workloads (such as Deployments or StatefulSets) Multi-cluster job dispatching: called MultiKueue, allows to search for capacity and off-load the main cluster. Topology-Aware Scheduling: Allows to optimize the Pod-to-Pod communication throughput by scheduling aware of the data-center topology. API version: v1beta2, respecting Kubernetes Deprecation Policy. Up-to-date documentation. Test coverage: Unit test testgrid. Integration tests: Baseline suite shard-0 shard-1 shard-2 MultiKueue suite testgrid. E2E tests: Baseline suites for Kubernetes 1.34 1.35 1.36 on Kind. Extended suites for Kubernetes on Kind: 1.34: shard-0 shard-1 shard-2 1.35: shard-0 shard-1 shard-2 1.36: shard-0 shard-1 shard-2 TAS: Baseline suite testgrid. Extended suite: shard-0 shard-1 Sequential tests: Baseline suites: shard-0 shard-1 Extended suites: shard-0 shard-1 E2E Cert Manager test testgrid. DRA test testgrid. MultiKueue: Baseline suite testgrid. Extended suites: shard-0 shard-1 MultiKueue DRA test testgrid. Scheduling performance tests: Baseline suite testgrid. TAS suite testgrid. Large-Scale suite testgrid. Scalability verification via performance tests. Monitoring via metrics. Security: RBAC based accessibility. Stable release cycle (2-3 months). Adopters running on production.
Based on community feedback, we continue to simplify and evolve the API to address new use cases.
Kueue at a glance
| Stars | 3k |
|---|---|
| Forks | 817 |
| Language | Go |
| License | Apache-2.0 |
| Last update | 2026-09-25 |
| Contributors | 381 |
Where Kueue is listed
Navid.me is reader-supported. When you buy through links on this site, I may earn an affiliate commission. Learn more.
GitHub repos like this
More repo topics
More free tools
Related MCP servers & CLIs
The most actionable AI newsletter for founders
Every week, get proven AI strategies, curated tools, and step-by-step systems to grow your audience, create better content, and build a profitable creator business.
No fluff, no filler, no BS. Just five minutes each week that might level up your online business and life.
P.S. Sign up now to get free access to my ultimate AI tools guide for creators.







































