Nomad forked from v1.6.5 https://developer.hashicorp.com/nomad/docs/v1.6.x

Go to file

Tim Gross 903b5baaa4 keyring: safely handle missing keys and restore GC (#15092 ) When replication of a single key fails, the replication loop breaks early and therefore keys that fall later in the sorting order will never get replicated. This is particularly a problem for clusters impacted by the bug that caused #14981 and that were later upgraded; the keys that were never replicated can now never be replicated, and so we need to handle them safely. Included in the replication fix: * Refactor the replication loop so that each key replicated in a function call that returns an error, to make the workflow more clear and reduce nesting. Log the error and continue. * Improve stability of keyring replication tests. We no longer block leadership on initializing the keyring, so there's a race condition in the keyring tests where we can test for the existence of the root key before the keyring has been initialize. Change this to an "eventually" test. But these fixes aren't enough to fix #14981 because they'll end up seeing an error once a second complaining about the missing key, so we also need to fix keyring GC so the keys can be removed from the state store. Now we'll store the key ID used to sign a workload identity in the Allocation, and we'll index the Allocation table on that so we can track whether any live Allocation was signed with a particular key ID.		2022-11-01 15:00:50 -04:00
.changelog	keyring: safely handle missing keys and restore GC (#15092 )	2022-11-01 15:00:50 -04:00
.circleci	ci: use groups of tests in gha (#15018 )	2022-10-27 09:02:58 -05:00
.github	ci: use groups of tests in gha (#15018 )	2022-10-27 09:02:58 -05:00
.release	Prepare for next release	2022-10-27 13:08:05 -04:00
.semgrep	variables: fix filter on List RPC	2022-10-27 13:08:05 -04:00
.tours	Make number of scheduler workers reloadable (#11593 )	2022-01-06 11:56:13 -05:00
acl	rename SecureVariables to Variables throughout	2022-08-26 16:06:24 -04:00
api	build(deps): bump github.com/stretchr/testify in /api (#15082 )	2022-10-31 08:45:04 -05:00
ci	ci: use groups of tests in gha (#15018 )	2022-10-27 09:02:58 -05:00
client	build: update linters (#15063 )	2022-10-27 15:02:30 -05:00
command	Generate files for 1.4.2 release	2022-10-27 13:08:05 -04:00
contributing	Update architecture-state-store.md (#15049 )	2022-10-27 14:03:43 -04:00
demo	demo/docs: update demo of Kadalu CSI Plugin (#13610 )	2022-07-06 10:24:34 -04:00
dev	docs: swap master for main in Nomad repo	2021-03-08 14:26:31 -05:00
drivers	client: protect user lookups with global lock (#14742 )	2022-09-29 09:30:13 -05:00
e2e	test: use port collision instead of cpu exhaustion (#14994 )	2022-10-21 07:53:26 -07:00
helper	helpers: lockfree lookup of nobody user on unix systems (#14866 )	2022-10-11 08:38:05 -05:00
integrations	spelling: registrations	2018-03-11 18:40:53 +00:00
internal/testing/apitests	cleanup: replace TypeToPtr helper methods with pointer.Of (#14151 )	2022-08-17 18:26:34 +02:00
jobspec	jobspec: allow artifact headers in HCLv1 (#14637 )	2022-09-27 12:18:49 -04:00
jobspec2	hcl2: add strlen function and update docs. (#14463 )	2022-09-06 18:42:40 +02:00
lib	cleanup: rename Equals to Equal for consistency (#14759 )	2022-10-10 09:28:46 -05:00
nomad	keyring: safely handle missing keys and restore GC (#15092 )	2022-11-01 15:00:50 -04:00
plugins	cleanup more helper updates (#14638 )	2022-09-21 14:53:25 -05:00
scheduler	make version checks specific to region (1.4.x) (#14912 )	2022-10-17 16:23:51 -04:00
scripts	build: update go version to go1.19.1 (#14653 )	2022-09-22 09:40:01 -05:00
terraform	terraform: update installed versions of HashiCorp tools. (#13635 )	2022-07-07 16:12:19 +02:00
testutil	Fixing flaky TestOverlap test (#14780 )	2022-10-03 14:35:02 -07:00
tools	ci: use groups of tests in gha (#15018 )	2022-10-27 09:02:58 -05:00
ui	refact: preserve promise.then behavior for acceptance tests (#15003 )	2022-10-24 09:04:39 -04:00
version	Prepare for next release	2022-10-27 13:08:05 -04:00
website	keyring: safely handle missing keys and restore GC (#15092 )	2022-11-01 15:00:50 -04:00
.git-blame-ignore-revs	ignore b0a20b4dc965a38b0c843f47c16685ccad7439da (#13648 )	2022-07-07 15:16:18 -07:00
.gitattributes	Remove invalid gitattributes	2018-02-14 14:47:43 -08:00
.gitignore	ci: use groups of tests in gha (#15018 )	2022-10-27 09:02:58 -05:00
.go-version	build: update go version to go1.19.1 (#14653 )	2022-09-22 09:40:01 -05:00
.golangci.yml	build: update linters (#15063 )	2022-10-27 15:02:30 -05:00
.semgrepignore	build: disable semgrep on structs.go for now	2022-02-01 10:09:49 -06:00
CHANGELOG.md	Merge release 1.4.2 files	2022-10-27 13:31:29 -04:00
CODEOWNERS	add service acct to codeowners for backport merging	2022-05-06 10:06:20 -07:00
GNUmakefile	build: update linters (#15063 )	2022-10-27 15:02:30 -05:00
LICENSE	[COMPLIANCE] Update MPL 2.0 LICENSE (#14884 )	2022-10-13 08:43:12 -04:00
README.md	readme: remove Gitter lobby link. (#14195 )	2022-08-22 10:33:20 +02:00
Vagrantfile	tools: update virtualbox networking configuration (#11561 )	2021-11-24 10:45:58 -05:00
build_linux_arm.go	gofmt all the files	2021-10-01 10:14:28 -04:00
go.mod	build(deps): bump github.com/docker/cli from 20.10.18+incompatible to 20.10.21+incompatible (#15078 )	2022-10-31 08:50:33 -05:00
go.sum	build(deps): bump github.com/docker/cli from 20.10.18+incompatible to 20.10.21+incompatible (#15078 )	2022-10-31 08:50:33 -05:00
main.go	docker_logger: reorder imports to save memory (#14875 )	2022-10-11 13:23:03 -04:00
main_test.go	…

README.md

Nomad

Nomad is a simple and flexible workload orchestrator to deploy and manage containers (docker, podman), non-containerized applications (executable, Java), and virtual machines (qemu) across on-prem and clouds at scale.

Nomad is supported on Linux, Windows, and macOS. A commercial version of Nomad, Nomad Enterprise, is also available.

Website: https://nomadproject.io
Tutorials: HashiCorp Learn
Forum: Discuss

Nomad provides several key features:

Deploy Containers and Legacy Applications: Nomad’s flexibility as an orchestrator enables an organization to run containers, legacy, and batch applications together on the same infrastructure. Nomad brings core orchestration benefits to legacy applications without needing to containerize via pluggable task drivers.
Simple & Reliable: Nomad runs as a single binary and is entirely self contained - combining resource management and scheduling into a single system. Nomad does not require any external services for storage or coordination. Nomad automatically handles application, node, and driver failures. Nomad is distributed and resilient, using leader election and state replication to provide high availability in the event of failures.
Device Plugins & GPU Support: Nomad offers built-in support for GPU workloads such as machine learning (ML) and artificial intelligence (AI). Nomad uses device plugins to automatically detect and utilize resources from hardware devices such as GPU, FPGAs, and TPUs.
Federation for Multi-Region, Multi-Cloud: Nomad was designed to support infrastructure at a global scale. Nomad supports federation out-of-the-box and can deploy applications across multiple regions and clouds.
Proven Scalability: Nomad is optimistically concurrent, which increases throughput and reduces latency for workloads. Nomad has been proven to scale to clusters of 10K+ nodes in real-world production environments.
HashiCorp Ecosystem: Nomad integrates seamlessly with Terraform, Consul, Vault for provisioning, service discovery, and secrets management.

Quick Start

Testing

See Learn: Getting Started for instructions on setting up a local Nomad cluster for non-production use.

Optionally, find Terraform manifests for bringing up a development Nomad cluster on a public cloud in the terraform directory.

Production

See Learn: Nomad Reference Architecture for recommended practices and a reference architecture for production deployments.

Documentation

Full, comprehensive documentation is available on the Nomad website: https://www.nomadproject.io/docs

Guides are available on HashiCorp Learn.

Contributing

See the contributing directory for more developer documentation.

README.md Unescape Escape