consul

Commit Graph

Author	SHA1	Message	Date
Matt Keeler	abed91d069	Generate JSON and Binary Marshalers for Protobuf Types (#6564 ) * Add JSON and Binary Marshaler Generators for Protobuf Types * Generate files with the correct version of gogo/protobuf I have pinned the version in the makefile so when you run make tools you get the right version. This pulls the version out of go.mod so it should remain up to date. The version at the time of this commit we are using is v1.2.1 * Fixup some shell output * Update how we determine the version of gogo This just greps the go.mod file instead of expecting the go mod cache to already be present * Fixup vendoring and remove no longer needed json encoder functions	2019-09-30 15:39:20 -04:00
John Cowen	b3b32dc0f6	ui: UI Release Merge (ui-staging merge) (#6527 ) ## HTTPAdapter (#5637) ## Ember upgrade 2.18 > 3.12 (#6448) ### Proxies can no longer get away with not calling _super This means that we can't use create anymore to define dynamic methods. Therefore we dynamically make 2 extended Proxies on demand, and then create from those. Therefore we can call _super in the init method of the extended Proxies. ### We aren't allowed to reset a service anymore We never actually need to now anyway, this is a remnant of the refactor from browser based confirmations. We fix it as simply as possible here but will revisit and remove the old browser confirm functionality at a later date ### Revert classes to use ES5 style to workaround babel transp. probs Using a mixture of ES6 classes (and hence super) and arrow functions means that when babel transpiles the arrow functions down to ES5, a reference to this is moved before the call to super, hence causing a js error. Furthermore, we the testing environment no longer lets use use apply/call on the constructor. These errors only manifests during testing (only in the testing environment), the application itself runs fine with no problems without this change. Using ES5 style class definitions give us freedom to do all of the above without causing any errors, so we reverted these classes back to ES5 class definitions ### Skip test that seems to have changed due to a change in RSVP timing This test tests a usecase/area of the API that will probably never ever be used, it was more testing out the API. We've skipped the test for now as this doesn't affect the application itself, but left a note to come back here later to investigate further ### Remove enumerableContentDidChange Initial testing looks like we don't need to call this function anymore, the function no longer exists ### Rework Changeset.isSaving to take into account new ember APIs Setting/hanging a computedProperty of an instantiated object no longer works. Move to setting it on the prototype/class definition instead ### Change how we detect whether something requires listening New ember API's have changed how you can detect whether something is a computedProperty or not. It's not immediately clear if its even possible now. Therefore we change how we detect whether something should be listened to or not by just looking for presence of `addEventListener` ### Potentially temporary change of ci test scripts to ensure deps exist All our tooling scripts run through a Makefile (for people familiar with only using those), which then call yarn scripts which can be called independently (for people familar with only using yarn). The Makefile targets always check to make sure all the dependencies are installed before running anything that requires them (building, testing etc). The CI scripts/targets didn't follow this same route and called the yarn scripts directly (usually CI builds a cache of the dependencies first). For some reason this cache isn't doing what it usually does, and it looks as though, in CI, ember isn't installed. This commit makes the CI scripts consistently use the same method as all of the other tooling scripts (Makefile target > Install Deps if required > call yarn script). This should install the dependencies if for some reason the CI cache building doesn't complete/isn't successful. Potentially this commit may be reverted if, the root of the problem is elsewhere, although consistency is always good, so it might be a good idea to leave this commit as is even if we need to debug and fix things elsewhere. ### Make test-parallel consistent with the rest of the tooling scripts As we are here making changes for CI purposes (making test-ci consistent), we spotted that test-parallel is also inconsistent and also the README manual instructions won't work without `ember` installed globally. This commit makes everything consistent and changes the manual instructions to use the local ember instance that gets installed via yarn ### Re-wrangle catchable to fit with new ember 3.12 APIs In the upgrade from ember 3.8 > 3.12 the public interfaces for ComputedProperties have changed slightly. `meta` is no longer a public property of ComputedProperty but of a ComputedDecoratorImpl mixin instead. `7e4ba1096e/packages/%40ember/-internals/metal/lib/computed.ts (L725)` There seems to be no way, by just using publically available methods, to replicate this behaviour so that we can create our own 'ComputedProperty` factory via injecting the ComputedProperty class as we did previously. `3f333bada1/ui-v2/app/utils/computed/factory.js (L1-L18)` Instead we dynamically hang our `Catchable` `catch` method off the instantiated ComputedProperty. In doing it like this `ComputedProperty` has already has its `meta` method mixed in so we don't have to manually mix it in ourselves (which doesn't seem possible) This functionality is only used during our work in trying to ensure our EventSource/BlockingQuery work was as 'ember-like' as possible (i.e. using the traditional Route.model hooks and ember-like Controller properties). Our ongoing/upcoming work on a componentized approach to data a.k.a `<DataSource />` means we will be able to remove the majority of the code involved here now that it seems to be under an amount of flux in ember. ### Build bindata_assetfs.go with new UI changes	2019-09-30 14:47:49 +01:00
Matt Keeler	923d8671a4	Add support for parameterizing the ACL config used with a TestA… (#6559 ) * Add support for parameterizing the ACL config used with a TestAgent Using tokens that are UUIDs will get rid of some warnings * Refactor to allow setting all tokens and change the template to ignore unset values.	2019-09-27 17:06:43 -04:00
R.B. Boyer	c4b92d5534	connect: connect CA Roots in secondary datacenters should use a SigningKeyID derived from their local intermediate (#6513 ) This fixes an issue where leaf certificates issued in secondary datacenters would be reissued very frequently (every ~20 seconds) because the logic meant to detect root rotation was errantly triggering because a hash of the ultimate root (in the primary) was being compared against a hash of the local intermediate root (in the secondary) and always failing.	2019-09-26 11:54:14 -05:00
R.B. Boyer	9566df524e	agent: cache notifications work after error if the underlying RPC returns index=1 (#6547 ) Fixes #6521 Ensure that initial failures to fetch an agent cache entry using the notify API where the underlying RPC returns a synthetic index of 1 correctly recovers when those RPCs resume working. The bug in the Cache.notifyBlockingQuery used to incorrectly "fix" the index for the next query from 0 to 1 for all queries, when it should have not done so for queries that errored. Also fixed some things that made debugging difficult: - config entry read/list endpoints send back QueryMeta headers - xds event loops don't swallow the cache notification errors	2019-09-26 10:42:17 -05:00
Matt Keeler	76cf54068b	Expand the QueryOptions and QueryMeta interfaces (#6545 ) In a previous PR I made it so that we had interfaces that would work enough to allow blockingQueries to work. However to complete this we need all fields to be settable and gettable. Notes: • If Go ever gets contracts/generics then we could get rid of all the Getters/Setters • protoc / protoc-gen-gogo are going to generate all the getters for us. • I copied all the getters/setters from the protobuf funcs into agent/structs/protobuf_compat.go • Also added JSON marshaling funcs that use jsonpb for protobuf types.	2019-09-26 09:55:02 -04:00
Freddy	fdd10dd8b8	Expose HTTP-based paths through Connect proxy (#6446 ) Fixes: #5396 This PR adds a proxy configuration stanza called expose. These flags register listeners in Connect sidecar proxies to allow requests to specific HTTP paths from outside of the node. This allows services to protect themselves by only listening on the loopback interface, while still accepting traffic from non Connect-enabled services. Under expose there is a boolean checks flag that would automatically expose all registered HTTP and gRPC check paths. This stanza also accepts a paths list to expose individual paths. The primary use case for this functionality would be to expose paths for third parties like Prometheus or the kubelet. Listeners for requests to exposed paths are be configured dynamically at run time. Any time a proxy, or check can be registered, a listener can also be created. In this initial implementation requests to these paths are not authenticated/encrypted.	2019-09-25 20:55:52 -06:00
R.B. Boyer	5882e21b2b	agent: tolerate more failure scenarios during service registration with central config enabled (#6472 ) Also: * Finished threading replaceExistingChecks setting (from GH-4905) through service manager. * Respected the original configSource value that was used to register a service or a check when restoring persisted data. * Run several existing tests with and without central config enabled (not exhaustive yet). * Switch to ioutil.ReadFile for all types of agent persistence.	2019-09-24 10:04:48 -05:00
Matt Keeler	100ebd63f9	Allow for enterprise only leader routines (#6533 ) Eventually I am thinking we may need a way to register these at different priority levels but for now sticking this here is fine	2019-09-23 20:09:56 -04:00
R.B. Boyer	af01d397a5	connect: don't colon-hex-encode the AuthorityKeyId and SubjectKeyId fields in connect certs (#6492 ) The fields in the certs are meant to hold the original binary representation of this data, not some ascii-encoded version. The only time we should be colon-hex-encoding fields is for display purposes or marshaling through non-TLS mediums (like RPC).	2019-09-23 12:52:35 -05:00
R.B. Boyer	796de297c8	connect: intermediate CA certs generated with the vault provider lack URI SANs (#6491 ) This only affects vault versions >=1.1.1 because the prior code accidentally relied upon a bug that was fixed in https://github.com/hashicorp/vault/pull/6505 The existing tests should have caught this, but they were using a vendored copy of vault version 0.10.3. This fixes the tests by running an actual copy of vault instead of an in-process copy. This has the added benefit of changing the dependency on vault to just vault/api. Also update VaultProvider to use similar SetIntermediate validation code as the ConsulProvider implementation.	2019-09-23 12:04:40 -05:00
Matt Keeler	51dcd126b7	Add support for implementing new requests with protobufs instea… (#6502 ) * Add build system support for protobuf generation This is done generically so that we don’t have to keep updating the makefile to add another proto generation. Note: anything not in the vendor directory and with a .proto extension will be run through protoc if the corresponding namespace.pb.go file is not up to date. If you want to rebuild just a single proto file you can do so with: make proto-rebuild PROTOFILES=<list of proto files to rebuild> Providing the PROTOFILES var will override the default behavior of finding all the .proto files. * Start adding types to the agent/proto package These will be needed for some other work and are by no means comprehensive. * Add ability to resolve/fixup the agentpb.ACLLinks structure in the state store. * Use protobuf marshalling of raft requests instead of msgpack for protoc generated types. This does not change any encoding of existing types. * Removed structs package automatically encoding with protobuf marshalling Instead the caller of raftApply that wants to opt-in to protobuf encoding will have to call `raftApplyProtobuf` * Run update-vendor to fixup modules.txt Nothing changed as far as dependencies go but the ordering of modules in that file depends on the time they are first seen and its not alphabetical. * Rename some things and implement the structs.RPCInfo interface bits agentpb.QueryOptions and agentpb.WriteRequest implement 3 of the 4 RPCInfo funcs and the new TargetDatacenter message type implements the fourth. * Use the right encoding function. * Renamed agent/proto package to agent/agentpb to prevent package name conflicts * Update modules.txt to fix ordering * Change blockingQuery to take in interfaces for the query options and meta * Add %T to error output. * Add/Update some comments	2019-09-20 14:37:22 -04:00
R.B. Boyer	f9496dc627	sdk: add freelist tracking and ephemeral port range skipping to freeport This should cut down on test flakiness. Problems handled: - If you had enough parallel test cases running, the former circular approach to handling the port block could hand out the same port to multiple cases before they each had a chance to bind them, leading to one of the two tests to fail. - The freeport library would allocate out of the ephemeral port range. This has been corrected for Linux (which should cover CI). - The library now waits until a formerly-in-use port is verified to be free before putting it back into circulation.	2019-09-17 14:30:43 -05:00
R.B. Boyer	7ccaa13514	fix typo of 'unknown' in log messages	2019-09-13 15:59:49 -05:00
R.B. Boyer	20eb0d3e94	cache: remove data race in agent cache In normal operations there is a read/write race related to request QueryOptions fields. An example race: WARNING: DATA RACE Read at 0x00c000836950 by goroutine 30: github.com/hashicorp/consul/agent/structs.(ServiceConfigRequest).CacheInfo() /go/src/github.com/hashicorp/consul/agent/structs/config_entry.go:506 +0x109 github.com/hashicorp/consul/agent/cache.(Cache).getWithIndex() /go/src/github.com/hashicorp/consul/agent/cache/cache.go:262 +0x5c github.com/hashicorp/consul/agent/cache.(Cache).notifyBlockingQuery() /go/src/github.com/hashicorp/consul/agent/cache/watch.go:89 +0xd7 Previous write at 0x00c000836950 by goroutine 147: github.com/hashicorp/consul/agent/cache-types.(ResolvedServiceConfig).Fetch() /go/src/github.com/hashicorp/consul/agent/cache-types/resolved_service_config.go:31 +0x219 github.com/hashicorp/consul/agent/cache.(*Cache).fetch.func1() /go/src/github.com/hashicorp/consul/agent/cache/cache.go:495 +0x112 This patch does a lightweight copy of the request struct so that the embedded QueryOptions fields that are mutated during Fetch() are scoped to just that one RPC.	2019-09-12 16:18:01 -05:00
hashicorp-ci	4f499a8044	update bindata_assetfs.go	2019-09-12 19:39:58 +00:00
Hans Hasselberg	4a20efda9b	agent: handleEnterpriseLeave (#6453 )	2019-09-11 11:01:37 +02:00
R.B. Boyer	a86e63f81e	test: actually wait for the TestAgent to be fully shutdown (#6441 )	2019-09-05 13:36:26 -05:00
Sarah Adams	001137e5e5	test: ensure all TestAgent constructions use a constructor (#6443 ) ensure all TestAgent constructions use a constructor to get start retries + test logs going to the right place Fixes #6435	2019-09-05 10:24:36 -07:00
Sarah Adams	74461406e0	remove funky panic/recover in agent tests (#6442 )	2019-09-04 13:59:11 -07:00
Sarah Adams	4ed5515fca	refactor & add better retry logic to NewTestAgent (#6363 ) Fixes #6361	2019-09-03 15:05:51 -07:00
Pierre Souchay	be50400c62	Distinguish between DC not existing and not being available (#6399 )	2019-09-03 09:46:24 -06:00
Aestek	7c7b7f24fd	Add option to register services and their checks idempotently (#4905 )	2019-09-02 09:38:29 -06:00
Sarah Adams	9f4b329b6d	txn: don't try to decode request bodies > raft.SuggestedMaxDataSize (#6422 ) txn: don't try to decode request bodies > raft.SuggestedMaxDataSize	2019-08-30 10:41:25 -07:00
Matt Keeler	42d608587f	Store primaries root in secondary after intermediate signature (#6333 ) * Store primaries root in secondary after intermediate signature This ensures that the intermediate exists within the CA root stored in raft and not just in the CA provider state. This has the very nice benefit of actually outputting the intermediate cert within the ca roots HTTP/RPC endpoints. This change means that if signing the intermediate fails it will not set the root within raft. So far I have not come up with a reason why that is bad. The secondary CA roots watch will pull the root again and go through all the motions. So as soon as getting an intermediate CA works the root will get set. * Make TestAgentAntiEntropy_Check_DeferSync less flaky I am not sure this is the full fix but it seems to help for me.	2019-08-30 11:38:46 -04:00
R.B. Boyer	7deaba63e1	test: ensure the node name is a valid dns name (#6424 ) The space in the node name was making every test emit a useless warning.	2019-08-29 16:52:13 -05:00
R.B. Boyer	3be88df207	test: explicitly run the pprof tests for 1s instead of the 30s default (#6421 )	2019-08-29 12:06:50 -05:00
R.B. Boyer	d979d8c239	test: add additional http status code assertions in coordinate HTTP API tests (#6410 ) When this test flakes sometimes this happens: --- FAIL: TestCoordinate_Node (1.69s) panic: interface conversion: interface {} is nil, not structs.Coordinates [recovered] FAIL github.com/hashicorp/consul/agent 19.999s Exit code: 1 panic: interface conversion: interface {} is nil, not structs.Coordinates [recovered] panic: interface conversion: interface {} is nil, not structs.Coordinates There is definitely a bug lurking, but the code seems to imply this can only return nil on 404. The tests previously were not checking the status code. The underlying cause of the flake is unknown, but this should turn the failure into a more normal test failure.	2019-08-29 09:55:05 -05:00
Pierre Souchay	58f04815d5	Display IPs of machines when node names conflict to ease troubleshooting When there is an node name conflicts, such messages are displayed within Consul: `consul.fsm: EnsureRegistration failed: failed inserting node: Error while renaming Node ID: "e1d456bc-f72d-98e5-ebb3-26ae80d785cf": Node name node001 is reserved by node 05f10209-1b9c-b90c-e3e2-059e64556d4a with name node001` While it is easy to find the node that has reserved the name, it is hard to find the node trying to aquire the name since it is not registered, because it is not part of `consul members` output This PR will display the IP of the offender and solve far more easily those issues.	2019-08-28 15:57:05 -04:00
Alvin Huang	c516fabfac	revert commits on master (#6413 )	2019-08-27 17:45:58 -04:00
tradel	5a22b77340	update tests to match new method signatures	2019-08-27 14:16:39 -07:00
tradel	1ff46f3f0a	confi\gure providers with DC and domain	2019-08-27 14:16:25 -07:00
tradel	5ba28a6a7b	create a common name for autoTLS agent certs	2019-08-27 14:15:53 -07:00
tradel	9b1ac4e7ef	add subject names to issued certs	2019-08-27 14:15:10 -07:00
tradel	7f36a5b676	construct a common name for each CSR	2019-08-27 14:12:56 -07:00
tradel	672e181399	add serviceID to leaf cert request	2019-08-27 14:12:22 -07:00
tradel	a4312d2e6e	add domain and nodeName to agent cert request	2019-08-27 14:11:40 -07:00
tradel	82ae7caf3e	Added DC and domain args to Configure method	2019-08-27 14:09:01 -07:00
R.B. Boyer	b962fe38cd	test: send testagent logs through testing.Logf (#6411 )	2019-08-27 12:21:30 -05:00
R.B. Boyer	91da908d2f	test: fix TestAgent.Start() to not segfault if the DNSServer cannot ListenAndServe (#6409 ) The embedded `Server` field on a `DNSServer` is only set inside of the `ListenAndServe` method. If that method fails for reasons like the address being in use and is not bindable, then the `Server` field will not be set and the overall `Agent.Start()` will fail. This will trigger the inner loop of `TestAgent.Start()` to invoke `ShutdownEndpoints` which will attempt to pretty print the DNS servers using fields on that inner `Server` field. Because it was never set, this causes a nil pointer dereference and crashes the test.	2019-08-27 10:45:05 -05:00
Alvin Huang	0be1531d80	add nil pointer check for pointer to ACLToken struct (#6407 )	2019-08-27 11:23:28 -04:00
Hans Hasselberg	d051342902	make sure auto_encrypt has private key type and bits (#6392 )	2019-08-27 14:37:56 +02:00
Hans Hasselberg	faa54ab989	auto_encrypt: verify_incoming_rpc is good enough for auto_encrypt.allow_tls (#6376 ) Previously `verify_incoming` was required when turning on `auto_encrypt.allow_tls`, but that doesn't work together with HTTPS UI in some scenarios. Adding `verify_incoming_rpc` to the allowed configurations.	2019-08-27 14:36:36 +02:00
R.B. Boyer	7bc941575c	test: don't leak agent goroutines in TestAgent_sidecarServiceFromNodeService (#6396 ) A goroutine dump using runtime.Stack() before/after shows a drop from 121 => 4.	2019-08-26 15:19:59 -05:00
Hans Hasselberg	f3def8c0d0	make sure auto_encrypt has private key type and bits	2019-08-26 13:09:50 +02:00
hashicorp-ci	3757cb6a96	update bindata_assetfs.go	2019-08-23 22:10:50 +00:00
R.B. Boyer	cc9a6f7993	Merge pull request #6388 from hashicorp/release/1-6 merging release/1-6 into master	2019-08-23 13:44:46 -05:00
Matt Keeler	cbd1857186	Secondary CA `establishLeadership` fix (#6383 ) This prevents ACL issues (or other issues) during intermediate CA cert signing from failing leader establishment.	2019-08-23 11:32:37 -04:00
Hans Hasselberg	3e46352ccb	auto_encrypt: use server-port (#6287 ) AutoEncrypt needs the server-port because it wants to talk via RPC. Information from gossip might not be available at that point and thats why the server-port is being used.	2019-08-23 10:18:46 +02:00
R.B. Boyer	dfcdc41ef8	connect: allow 'envoy_cluster_json' escape hatch to continue to function (#6378 )	2019-08-22 15:11:56 -05:00
R.B. Boyer	7a6faccf2f	docs: document how envoy escape hatches work with the discovery chain (#6350 ) - Bootstrap escape hatches are OK. - Public listener/cluster escape hatches are OK. - Upstream listener/cluster escape hatches are not supported. If an unsupported escape hatch is configured and the discovery chain is activated log a warning and act like it was not configured. Fixes #6160	2019-08-21 15:10:12 -05:00
Matt Keeler	0c3cf47266	Ensure that config entry writes are forwarded to the primary DC (#6339 )	2019-08-20 12:01:13 -04:00
R.B. Boyer	fd1c62ee8b	connect: ensure time.Duration fields retain their human readable forms in the API (#6348 ) This applies for both config entries and the compiled discovery chain. Also omit some other config entries fields when empty.	2019-08-19 15:31:05 -05:00
R.B. Boyer	561b2fe606	connect: generate the full SNI names for discovery targets in the compiler rather than in the xds package (#6340 )	2019-08-19 13:03:03 -05:00
R.B. Boyer	ae79cdab1b	connect: introduce ExternalSNI field on service-defaults (#6324 ) Compiling this will set an optional SNI field on each DiscoveryTarget. When set this value should be used for TLS connections to the instances of the target. If not set the default should be used. Setting ExternalSNI will disable mesh gateway use for that target. It also disables several service-resolver features that do not make sense for an external service.	2019-08-19 12:19:44 -05:00
R.B. Boyer	1a485011d0	connect: updating a service-defaults config entry should leave an unset protocol alone (#6342 ) If the entry is updated for reasons other than protocol it is surprising that the value is explicitly persisted as 'tcp' rather than leaving it empty and letting it fall back dynamically on the proxy-defaults value.	2019-08-19 10:44:06 -05:00
Matt Keeler	318b9ebbe3	Filter out left/leaving serf members when determining if new AC… (#6332 )	2019-08-16 10:34:18 -04:00
R.B. Boyer	72207256b9	xds: improve how envoy metrics are emitted (#6312 ) Since generated envoy clusters all are named using (mostly) SNI syntax we can have envoy read the various fields out of that structure and emit it as stats labels to the various telemetry backends. I changed the delimiter for the 'customization hash' from ':' to '~' because ':' is always reencoded by envoy as '_' when generating metrics keys.	2019-08-16 09:30:17 -05:00
R.B. Boyer	3975cb89bf	agent: blocking central config RPCs iterations should not interfere with each other (#6316 )	2019-08-14 09:08:46 -05:00
hashicorp-ci	d7bb2bbee4	update bindata_assetfs.go	2019-08-13 15:28:06 +00:00
hashicorp-ci	5919c7c184	Merge Consul OSS branch 'master' at commit `8f7586b339`	2019-08-13 02:00:43 +00:00
Sarah Adams	8ff1f481fe	add flag to allow /operator/keyring requests to only hit local servers (#6279 ) Add parameter local-only to operator keyring list requests to force queries to only hit local servers (no WAN traffic). HTTP API: GET /operator/keyring?local-only=true CLI: consul keyring -list --local-only Sending the local-only flag with any non-GET/list request will result in an error.	2019-08-12 11:11:11 -07:00
Mike Morris	61206fdf42	snapshot: add TLS support to HalfCloser interface (#6216 ) Calls net.TCPConn.CloseWrite or mtls.Conn.CloseWrite, which was added in https://go-review.googlesource.com/c/go/+/31318/	2019-08-12 12:47:02 -04:00
Matt Keeler	6a1e0dfed8	Update the v1/agent/service/:service endpoint to output tagged… (#6304 )	2019-08-10 09:15:19 -04:00
R.B. Boyer	913d85ea5b	connect: allow mesh gateways to use central config (#6302 )	2019-08-09 15:07:01 -05:00
Mike Morris	65be58703c	connect: remove managed proxies (#6220 ) * connect: remove managed proxies implementation and all supporting config options and structs * connect: remove deprecated ProxyDestination * command: remove CONNECT_PROXY_TOKEN env var * agent: remove entire proxyprocess proxy manager * test: remove all managed proxy tests * test: remove irrelevant managed proxy note from TestService_ServerTLSConfig * test: update ContentHash to reflect managed proxy removal * test: remove deprecated ProxyDestination test * telemetry: remove managed proxy note * http: remove /v1/agent/connect/proxy endpoint * ci: remove deprecated test exclusion * website: update managed proxies deprecation page to note removal * website: remove managed proxy configuration API docs * website: remove managed proxy note from built-in proxy config * website: add note on removing proxy subdirectory of data_dir	2019-08-09 15:19:30 -04:00
R.B. Boyer	9bbbea1777	connect: ensure intention replication continues to work when the replication ACL token changes (#6288 )	2019-08-07 11:34:09 -05:00
hashicorp-ci	913784e1bf	Merge Consul OSS branch 'master' at commit `d84863799d`	2019-08-06 02:00:30 +00:00
Sarah Adams	d84863799d	fallback to proxy config global protocol when upstream services' protocol is unset (#6277 ) fallback to proxy config global protocol when upstream services' protocol is unset Fixes #5857	2019-08-05 12:52:35 -07:00
R.B. Boyer	8e22d80e35	connect: fix failover through a mesh gateway to a remote datacenter (#6259 ) Failover is pushed entirely down to the data plane by creating envoy clusters and putting each successive destination in a different load assignment priority band. For example this shows that normally requests go to 1.2.3.4:8080 but when that fails they go to 6.7.8.9:8080: - name: foo load_assignment: cluster_name: foo policy: overprovisioning_factor: 100000 endpoints: - priority: 0 lb_endpoints: - endpoint: address: socket_address: address: 1.2.3.4 port_value: 8080 - priority: 1 lb_endpoints: - endpoint: address: socket_address: address: 6.7.8.9 port_value: 8080 Mesh gateways route requests based solely on the SNI header tacked onto the TLS layer. Envoy currently only lets you configure the outbound SNI header at the cluster layer. If you try to failover through a mesh gateway you ideally would configure the SNI value per endpoint, but that's not possible in envoy today. This PR introduces a simpler way around the problem for now: 1. We identify any target of failover that will use mesh gateway mode local or remote and then further isolate any resolver node in the compiled discovery chain that has a failover destination set to one of those targets. 2. For each of these resolvers we will perform a small measurement of comparative healths of the endpoints that come back from the health API for the set of primary target and serial failover targets. We walk the list of targets in order and if any endpoint is healthy we return that target, otherwise we move on to the next target. 3. The CDS and EDS endpoints both perform the measurements in (2) for the affected resolver nodes. 4. For CDS this measurement selects which TLS SNI field to use for the cluster (note the cluster is always going to be named for the primary target) 5. For EDS this measurement selects which set of endpoints will populate the cluster. Priority tiered failover is ignored. One of the big downsides to this approach to failover is that the failover detection and correction is going to be controlled by consul rather than deferring that entirely to the data plane as with the prior version. This also means that we are bound to only failover using official health signals and cannot make use of data plane signals like outlier detection to affect failover. In this specific scenario the lack of data plane signals is ok because the effectiveness is already muted by the fact that the ultimate destination endpoints will have their data plane signals scrambled when they pass through the mesh gateway wrapper anyway so we're not losing much. Another related fix is that we now use the endpoint health from the underlying service, not the health of the gateway (regardless of failover mode).	2019-08-05 13:30:35 -05:00
R.B. Boyer	c395affc93	connect: expose an API endpoint to compile the discovery chain (#6248 ) In addition to exposing compilation over the API cleaned up the structures that would be exchanged to be cleaner and easier to support and understand. Also removed ability to configure the envoy OverprovisioningFactor.	2019-08-02 15:34:54 -05:00
Todd Radel	96be92f3b5	connect: generate intermediate at same time as root (#6272 ) Generate intermediate at same time as root Co-Authored-By: Freddy <freddygv@users.noreply.github.com>	2019-08-02 15:36:03 -04:00
R.B. Boyer	dcb609af83	connect: detect and prevent circular discovery chain references (#6246 )	2019-08-02 09:18:45 -05:00
R.B. Boyer	1d0efdc69e	server: if inserting bootstrap config entries fails don't silence the errors (#6256 )	2019-08-01 23:07:11 -05:00
R.B. Boyer	f02924fafe	connect: simplify the compiled discovery chain data structures (#6242 ) This should make them better for sending over RPC or the API. Instead of a chain implemented explicitly like a linked list (nodes holding pointers to other nodes) instead switch to a flat map of named nodes with nodes linking other other nodes by name. The shipped structure is just a map and a string to indicate which key to start from. Other changes: * inline the compiler option InferDefaults as true * introduce compiled target config to avoid needing to send back additional maps of Resolvers; future target-specific compiled state can go here * move compiled MeshGateway out of the Resolver and into the TargetConfig where it makes more sense.	2019-08-01 22:44:05 -05:00
R.B. Boyer	6393edba53	connect: reconcile how upstream configuration works with discovery chains (#6225 ) * connect: reconcile how upstream configuration works with discovery chains The following upstream config fields for connect sidecars sanely integrate into discovery chain resolution: - Destination Namespace/Datacenter: Compilation occurs locally but using different default values for namespaces and datacenters. The xDS clusters that are created are named as they normally would be. - Mesh Gateway Mode (single upstream): If set this value overrides any value computed for any resolver for the entire discovery chain. The xDS clusters that are created may be named differently (see below). - Mesh Gateway Mode (whole sidecar): If set this value overrides any value computed for any resolver for the entire discovery chain. If this is specifically overridden for a single upstream this value is ignored in that case. The xDS clusters that are created may be named differently (see below). - Protocol (in opaque config): If set this value overrides the value computed when evaluating the entire discovery chain. If the normal chain would be TCP or if this override is set to TCP then the result is that we explicitly disable L7 Routing and Splitting. The xDS clusters that are created may be named differently (see below). - Connect Timeout (in opaque config): If set this value overrides the value for any resolver in the entire discovery chain. The xDS clusters that are created may be named differently (see below). If any of the above overrides affect the actual result of compiling the discovery chain (i.e. "tcp" becomes "grpc" instead of being a no-op override to "tcp") then the relevant parameters are hashed and provided to the xDS layer as a prefix for use in naming the Clusters. This is to ensure that if one Upstream discovery chain has no overrides and tangentially needs a cluster named "api.default.XXX", and another Upstream does have overrides for "api.default.XXX" that they won't cross-pollinate against the operator's wishes. Fixes #6159	2019-08-01 22:03:34 -05:00
R.B. Boyer	8564b6bb38	connect: validate upstreams and prevent duplicates (#6224 ) * connect: validate upstreams and prevent duplicates * Actually run Upstream.Validate() instead of ignoring it as dead code. * Prevent two upstreams from declaring the same bind address and port. It wouldn't work anyway. * Prevent two upstreams from being declared that use the same type+name+namespace+datacenter. Due to how the Upstream.Identity() function worked this ended up mostly being enforced in xDS at use-time, but it should be enforced more clearly at register-time.	2019-08-01 13:26:02 -05:00
Paul Banks	e87cef2bb8	Revert "connect: support AWS PCA as a CA provider" (#6251 ) This reverts commit `3497b7c00d`.	2019-07-31 09:08:10 -04:00
Todd Radel	3497b7c00d	connect: support AWS PCA as a CA provider (#6189 ) Port AWS PCA provider from consul-ent	2019-07-30 22:57:51 -04:00
Todd Radel	2552f4a11a	connect: Support RSA keys in addition to ECDSA (#6055 ) Support RSA keys in addition to ECDSA	2019-07-30 17:47:39 -04:00
freddygv	1a14b94441	Update default gossip encryption key size to 32 bytes	2019-07-30 09:45:41 -06:00
hashicorp-ci	847b90288a	Merge Consul OSS branch 'master' at commit `a1725e6b52`	2019-07-30 02:00:29 +00:00
Matt Keeler	a1725e6b52	Fix flaky tests (#6229 )	2019-07-29 15:07:25 -04:00
Matt Keeler	fcc18c1675	Fix prepared query upstream endpoint generation (#6236 ) Use the correct SNI value for prepared query upstreams	2019-07-29 11:15:55 -04:00
hashicorp-ci	1527616c8c	update bindata_assetfs.go	2019-07-26 23:15:20 +00:00
Matt Keeler	35e67b1d1a	Fix CA Replication when ACLs are enabled (#6201 ) Secondary CA initialization steps are: • Wait until the primary will be capable of signing intermediate certs. We use serf metadata to check the versions of servers in the primary which avoids needing a token like the previous implementation that used RPCs. We require at least one alive server in the primary and the all alive servers meet the version requirement. • Initialize the secondary CA by getting the primary to sign an intermediate When a primary dc is configured, if no existing CA is initialized and for whatever reason we cannot initialize a secondary CA the secondary DC will remain without a CA. As soon as it can it will initialize the secondary CA by pulling the primaries roots and getting the primary to sign an intermediate. This also fixes a segfault that can happen during leadership revocation. There was a spot in the secondaryCARootsWatch that was getting the CA Provider and executing methods on it without nil checking. Under normal circumstances it wont be nil but during leadership revocation it gets nil'ed out. Therefore there is a period of time between closing the stop chan and when the go routine is actually stopped where it could read a nil provider and cause a segfault.	2019-07-26 15:57:57 -04:00
R.B. Boyer	c6c4a2251a	Merge Consul OSS branch master at commit `b3541c4f34`	2019-07-26 10:34:24 -05:00
hashicorp-ci	0c9c5bfa98	update bindata_assetfs.go	2019-07-25 23:41:16 +00:00
Jack Pearkes	4e0a16ab2d	config: correct limit to limits in config example (#6219 ) This isn't yet documented on the website, but wanted to update this to add the missing s.	2019-07-25 12:38:57 -07:00
Matt Keeler	8b54307be2	Allow forwarding of some status RPCs (#6198 ) * Allow forwarding of some status RPCs * Update docs * add comments about not using the regular forward	2019-07-25 14:26:22 -04:00
Jeff Mitchell	1ea7a34756	Make the chunking test multidimensional (#6212 ) This ensures that it's not just a single operation we restores successfully, but many. It's the same foundation, just with multiple going on at once.	2019-07-25 11:40:09 +01:00
Freddy	89158c7a76	auto-encrypt: Fix port resolution and fallback to default port (#6205 ) Auto-encrypt meant to fallback to the default port when it wasn't provided, but it hadn't been because of an issue with the error handling. We were checking against an incomplete error value: "missing port in address" vs "address $HOST: missing port in address" Additionally, all RPCs to AutoEncrypt.Sign were using a.config.ServerPort, so those were updated to use ports resolved by resolveAddrs, if they are available.	2019-07-24 16:49:37 -07:00
Jeff Mitchell	94c73d0c92	Chunking support (#6172 ) * Initial chunk support This uses the go-raft-middleware library to allow for chunked commits to the KV	2019-07-24 17:06:39 -04:00
Matt Keeler	3053342198	Envoy Mesh Gateway integration tests (#6187 ) * Allow setting the mesh gateway mode for an upstream in config files * Add envoy integration test for mesh gateways This necessitated many supporting changes in most of the other test cases. Add remote mode mesh gateways integration test	2019-07-24 17:01:42 -04:00
Freddy	199f6cc41d	Make new config when retrying testServer creation (#6204 )	2019-07-24 08:41:00 -06:00
R.B. Boyer	ad9e7b6ae9	connect: allow L7 routers to match on http methods (#6164 ) Fixes #6158	2019-07-23 20:56:39 -05:00
R.B. Boyer	85cf2706e6	connect: change router syntax for matching query parameters to resemble the syntax for matching paths and headers for consistency. (#6163 ) This is a breaking change, but only in the context of the beta series.	2019-07-23 20:55:26 -05:00
R.B. Boyer	1dbd92e091	connect: validate and test more of the L7 config entries (#6156 )	2019-07-23 20:50:23 -05:00
R.B. Boyer	e039dfd7f8	connect: rework how the service resolver subset OnlyPassing flag works (#6173 ) The main change is that we no longer filter service instances by health, preferring instead to render all results down into EDS endpoints in envoy and merely label the endpoints as HEALTHY or UNHEALTHY. When OnlyPassing is set to true we will force consul checks in a 'warning' state to render as UNHEALTHY in envoy. Fixes #6171	2019-07-23 20:20:24 -05:00
Alvin Huang	ef6b80bab2	resolve circleci config conflicts	2019-07-23 20:18:36 -04:00
Freddy	d86efb83e5	Restore NotifyListen to avoid panic in newServer retry (#6200 )	2019-07-23 14:33:00 -06:00
Pierre Souchay	b4590fb8e8	Display nicely Networks (CIDR) in runtime configuration (#6029 ) * Display nicely Networks (CIDR) in runtime configuration CIDR mask is displayed in binary in configuration. This add support for nicely displaying CIDR in runtime configuration. Currently, if a configuration contains the following lines: "http_config": { "allow_write_http_from": [ "127.0.0.0/8", "::1/128" ] } A call to `/v1/agent/self?pretty` would display "AllowWriteHTTPFrom": [ { "IP": "127.0.0.0", "Mask": "/wAAAA==" }, { "IP": "::1", "Mask": "/////////////////////w==" } ] This PR fixes it and it will now display: "AllowWriteHTTPFrom": [ "127.0.0.0/8", "::1/128" ] * Added test for cidr nice rendering in `TestSanitize()`.	2019-07-23 16:30:16 -04:00
Matt Keeler	d7fe8befa9	Update go-bexpr (#6190 ) * Update go-bexpr to v0.1.1 This brings in: • `in`/`not in` operators to do substring matching • `matches` / `not matches` operators to perform regex string matching. * Add the capability to auto-generate the filtering selector ops tables for our docs	2019-07-23 14:45:20 -04:00
Paul Banks	f38da47c55	Allow raft TrailingLogs to be configured. (#6186 ) This fixes pathological cases where the write throughput and snapshot size are both so large that more than 10k log entries are written in the time it takes to restore the snapshot from disk. In this case followers that restart can never catch up with leader replication again and enter a loop of constantly downloading a full snapshot and restoring it only to find that snapshot is already out of date and the leader has truncated its logs so a new snapshot is sent etc. In general if you need to adjust this, you are probably abusing Consul for purposes outside its design envelope and should reconsider your usage to reduce data size and/or write volume.	2019-07-23 15:19:57 +01:00
Christian Muehlhaeuser	7753b97cc7	Simplified code in various places (#6176 ) All these changes should have no side-effects or change behavior: - Use bytes.Buffer's String() instead of a conversion - Use time.Since and time.Until where fitting - Drop unnecessary returns and assignment	2019-07-20 09:37:19 -04:00
hashicorp-ci	a4431da1cc	Merge Consul OSS branch 'master' at commit `ef257b084d`	2019-07-20 02:00:29 +00:00
javicrespo	b006060d4c	log rotation: limit count of rotated log files (#5831 )	2019-07-19 15:36:34 -06:00
Christian Muehlhaeuser	61ff1d20bf	Avoid unnecessary conversions (#6178 ) Those values already have the right type.	2019-07-19 09:13:18 -04:00
Christian Muehlhaeuser	1366bebf7f	Fixed typos in comments (#6175 ) Just a few nitpicky typo fixes.	2019-07-19 07:54:53 -04:00
Christian Muehlhaeuser	16193665ca	Fixed a few tautological condition mistakes (#6177 ) None of these changes should have any side-effects. They're merely fixing tautological mistakes.	2019-07-19 07:53:42 -04:00
Christian Muehlhaeuser	ed4e64f6b2	Fixed nil check for token (#6179 ) I can only assume we want to check for the retrieved `updatedToken` to not be nil, before accessing it below. `token` can't possibly be nil at this point, as we accessed `token.AccessorID` just before.	2019-07-19 07:48:11 -04:00
Alvin Huang	6f1953d96d	Merge branch 'master' into release/1-6	2019-07-17 15:43:30 -04:00
R.B. Boyer	d7a5158805	xds: allow http match criteria to be applied to routes on services using grpc protocols (#6149 )	2019-07-17 14:07:08 -05:00
R.B. Boyer	b16d7f00bc	agent: avoid reverting any check updates that occur while a service is being added or the config is reloaded (#6144 )	2019-07-17 14:06:50 -05:00
Freddy	5526cfac8b	Reduce number of servers in TestServer_Expect_NonVoters (#6155 )	2019-07-17 11:35:33 -06:00
Freddy	59dbd070d7	More flaky test fixes (#6151 ) * Add retry to TestAPI_ClientTxn * Add retry to TestLeader_RegisterMember * Account for empty watch result in ConnectRootsWatch	2019-07-17 09:33:38 -06:00
hashicorp-ci	fa20c7db97	Merge Consul OSS branch 'master' at commit `95dbb7f2f1`	2019-07-17 02:00:21 +00:00
Sarah Adams	ea2bd5b728	http/tcp checks: fix long timeout behavior to default to user-configured value (#6094 ) Fixes #5834	2019-07-16 15:13:26 -07:00
Freddy	d219e31db8	Update retries that weren't using retry.R (#6146 )	2019-07-16 14:47:45 -06:00
hashicorp-ci	4bb858161e	update bindata_assetfs.go	2019-07-15 18:09:30 +00:00
R.B. Boyer	e8132b61c0	add test for discovery chain agent cache-type (#6130 )	2019-07-15 10:09:52 -05:00
Jack Pearkes	338aed32af	Merge branch 'master' into release/1-6	2019-07-12 14:51:25 -07:00
Matt Keeler	4728329aeb	Various Gateway Fixes (#6093 ) * Ensure the mesh gateway configuration comes back in the api within each upstream * Add a test for the MeshGatewayConfig in the ToAPI functions * Ensure we don’t use gateways for dc local connections * Update the svc kind index for deletions * Replace the proxycfg.state cache with an interface for testing Also start implementing proxycfg state testing. * Update the state tests to verify some gateway watches for upstream-targets of a discovery chain.	2019-07-12 17:19:37 -04:00
Sarah Adams	ce3a6e8660	fix flaky test TestACLEndpoint_SecureIntroEndpoints_OnlyCreateLocalData (#6116 ) * fix test to write only to dc2 (typo) * fix retry behavior in existing test (was being used incorrectly)	2019-07-12 14:14:42 -07:00
R.B. Boyer	bcd2de3a2e	implement some missing service-router features and add more xDS testing (#6065 ) - also implement OnlyPassing filters for non-gateway clusters	2019-07-12 14:16:21 -05:00
R.B. Boyer	9138a97054	Fix bug in service-resolver redirects if the destination uses a default resolver. (#6122 ) Also: - add back an internal http endpoint to dump a compiled discovery chain for debugging purposes Before the CompiledDiscoveryChain.IsDefault() method would test: - is this chain just one resolver step? - is that resolver step just the default? But what I forgot to test: - is that resolver step for the same service that the chain represents? This last point is important because if you configured just one config entry: kind = "service-resolver" name = "web" redirect { service = "other" } and requested the chain for "web" you'd get back a default resolver for "other". In the xDS code the IsDefault() method is used to determine if this chain is "empty". If it is then we use the pre-discovery-chain logic that just uses data embedded in the Upstream object (and still lets the escape hatches function). In the example above that means certain parts of the xDS code were going to try referencing a cluster named "web..." despite the other parts of the xDS code maintaining clusters named "other...".	2019-07-12 12:21:25 -05:00
R.B. Boyer	67a36e3452	handle structs.ConfigEntry decoding similarly to api.ConfigEntry decoding (#6106 ) Both 'consul config write' and server bootstrap config entries take a decoding detour through mapstructure on the way from HCL to an actual struct. They both may take in snake_case or CamelCase (for consistency) so need very similar handling. Unfortunately since they are operating on mirror universes of structs (api.* vs structs.*) the code cannot be identitical, so try to share the kind-configuration and duplicate the rest for now.	2019-07-12 12:20:30 -05:00
Matt Keeler	6e65811db2	Envoy CLI bind addresses (#6107 ) * Ensure we MapWalk the proxy config in the NodeService and ServiceNode structs This gets rid of some json encoder errors in the catalog endpoints * Allow passing explicit bind addresses to envoy * Move map walking to the ConnectProxyConfig struct Any place where this struct gets JSON encoded will benefit as opposed to having to implement it everywhere. * Fail when a non-empty address is provided and not bindable * camel case * Update command/connect/envoy/envoy.go Co-Authored-By: Paul Banks <banks@banksco.de>	2019-07-12 12:57:31 -04:00
Freddy	5873c56a03	Flaky test overhaul (#6100 )	2019-07-12 09:52:26 -06:00
Freddy	4033a4d632	Remove dummy config (#6121 )	2019-07-12 09:50:14 -06:00
Freddy	0b9111714b	Update TestServer creation in sdk/testutil (#6084 ) * Retry the creation of the test server three times. * Reduce the retry timeout for the API wait to 2 seconds, opting to fail faster and start over. * Remove wait for leader from server creation. This wait can be added on a test by test basis now that the function is being exported. * Remove wait for anti-entropy sync. This is built into the existing WaitForSerfCheck func, so that can be used if the anti-entropy wait is needed	2019-07-12 09:37:29 -06:00
Freddy	0f8837824e	Clean up StatsFetcher work when context is exceeded (#6086 )	2019-07-12 08:23:28 -06:00
Matt Keeler	de23af071a	Move ctx and cancel func setup into the Replicator.Start (#6115 ) Previously a sequence of events like: Start Stop Start Stop would segfault on the second stop because the original ctx and cancel func were only initialized during the constructor and not during Start.	2019-07-12 10:10:48 -04:00
hashicorp-ci	7e379bb7f7	update bindata_assetfs.go	2019-07-08 13:20:35 +00:00
Jack Pearkes	e6f1b78efb	Make cluster names SNI always (#6081 ) * Make cluster names SNI always * Update some tests * Ensure we check for prepared query types * Use sni for route cluster names * Proper mesh gateway mode defaulting when the discovery chain is used * Ignore service splits from PatchSliceOfMaps * Update some xds golden files for proper test output * Allow for grpc/http listeners/cluster configs with the disco chain * Update stats expectation	2019-07-08 12:48:48 +01:00
Michael Schurter	b5aab27c21	connect: allow overriding envoy listener bind_address (#6033 ) * connect: allow overriding envoy listener bind_address * Update agent/xds/config.go Co-Authored-By: Kyle Havlovitz <kylehav@gmail.com> * connect: allow overriding envoy listener bind_port * envoy: support unix sockets for grpc in bootstrap Add AgentSocket BootstrapTplArgs which if set overrides the AgentAddress and AgentPort to generate a bootstrap which points Envoy to a unix socket file instead of an ip:port. * Add a test for passing the consul addr as a unix socket * Fix config formatting for envoy bootstrap tests * Fix listeners test cases for bind addr/port * Update website/source/docs/connect/proxies/envoy.md	2019-07-05 16:06:47 +01:00
Matt Keeler	3d562bee5c	Fix Internal.ServiceDump blocking (#6076 ) maxIndexWatchTxn was only watching the IndexEntry of the max index of all the entries. It needed to watch all of them regardless of which was the max. Also plumbed the query source through in the proxy config to help better track requests.	2019-07-04 16:17:49 +01:00
Matt Keeler	62ad0294d4	Don't use WatchedDatacenters in the xds code as thsoe get nil'ed out prior to sending to xds	2019-07-03 09:59:21 -04:00
Matt Keeler	d2b00bd591	xds message ordering (#6061 ) xds message ordering	2019-07-03 09:18:58 -04:00
hashicorp-ci	7a32c5a618	Merge Consul OSS branch 'master' at commit `a58d8e91ac`	2019-07-03 02:00:31 +00:00
Matt Keeler	25f580bcaa	Fix a bunch of xds flaky tests The clusters/endpoints test were still relying on deterministic ordering of clusters/endpoints which cannot be relied upon due to golang purposefully not providing any guarantee about consistent interation ordering of maps. Also fixed a small bug in the connect proxy cluster generation that was causing the clusters slice to be double the size it needed to with the first half being all nil pointers.	2019-07-02 15:53:06 -04:00
Matt Keeler	3eb3ee5a15	Merge pull request #6053 from hashicorp/gateways_and_resolvers Integrate Mesh Gateways with ServiceResolverSubsets	2019-07-02 12:05:08 -04:00
R.B. Boyer	43770b9391	digest the proxy-defaults protocol into the graph (#6050 )	2019-07-02 11:01:17 -05:00
Matt Keeler	a8e2e866e3	Update xds/proxycfg tests to use the same looking trust domain as a normal system This is to prevent confusion about what our SNI fields actually look like.	2019-07-02 10:29:37 -04:00
Matt Keeler	a7421c160f	Implement mesh gateway management of service subsets Fixup some error handling	2019-07-02 10:29:37 -04:00
Matt Keeler	3b6d5e382a	Implement caching for config entry lists Update agent/cache-types/config_entry.go Co-Authored-By: R.B. Boyer <public@richardboyer.net>	2019-07-02 10:11:19 -04:00
R.B. Boyer	4bdb690a25	activate most discovery chain features in xDS for envoy (#6024 )	2019-07-01 22:10:51 -05:00
Matt Keeler	bdebe62fd0	Fix some tests that I broke when refactoring the ConfigSnapshot (#6051 ) * Fix some tests that I broke when refactoring the ConfigSnapshot * Make sure the MeshGateway config is added to all the right api structs * Fix some more tests	2019-07-01 19:47:58 -04:00
Pierre Souchay	fd9237a1ff	Bump timeout in TestManager_BasicLifecycle (#6030 )	2019-07-01 17:02:00 -06:00
Matt Keeler	8d953f5840	Implement Mesh Gateways This includes both ingress and egress functionality.	2019-07-01 16:28:30 -04:00
Matt Keeler	da8db83ddf	Fix secondary dc connect CA roots watch issue The general problem was that a the CA config which contained the trust domain was happening outside of the blocking mechanism so if the client started the blocking query before the primary dcs roots had been set then a state trust domain was being pushed down. This was fixed here but in the future we should probably fixup the CA initialization code to not initialize the CA config twice when it doesn’t need to.	2019-07-01 16:28:30 -04:00
Matt Keeler	4bc1277315	Include a content hash of the intention for use during replication	2019-07-01 16:28:30 -04:00
Matt Keeler	747ae6bdf5	Implement intention replication and secondary CA initialization	2019-07-01 16:28:30 -04:00
Matt Keeler	3943e38133	Implement Kind based ServiceDump and caching of the ServiceDump RPC	2019-07-01 16:28:30 -04:00
R.B. Boyer	2ad516aeaf	do some initial config entry graph validation during writes (#6047 )	2019-07-01 15:23:36 -05:00
hashicorp-ci	43bda6fb76	Merge Consul OSS branch 'master' at commit `e91f73f592`	2019-06-30 02:00:31 +00:00
Sarah Christoff	f09af53894	Remove failed nodes from serfWAN (#6028 ) * Prune Servers from WAN and LAN * cleaned up and fixed LAN to WAN * moving things around * force-leave remove from serfWAN, create pruneSerfWAN * removed serfWAN remove, reduced complexity, fixed comments * add another place to remove from serfWAN * add nil check * Update agent/consul/server.go Co-Authored-By: Paul Banks <banks@banksco.de>	2019-06-28 12:40:07 -05:00
R.B. Boyer	38d76c624e	Allow for both snake_case and CamelCase for config entries written with 'consul config write'. (#6044 ) This also has the added benefit of fixing an issue with passing time.Duration fields through config entries.	2019-06-28 11:35:35 -05:00
Hans Hasselberg	a82e6a7fd3	Release v1.5.2	2019-06-27 22:59:46 +00:00
Hans Hasselberg	33a7df3330	tls: auto_encrypt enables automatic RPC cert provisioning for consul clients (#5597 )	2019-06-27 22:22:07 +02:00
R.B. Boyer	6a52f9f9fb	initial version of L7 config entry compiler (#5994 ) With this you should be able to fetch all of the relevant discovery chain config entries from the state store in one query and then feed them into the compiler outside of a transaction. There are a lot of TODOs scattered through here, but they're mostly around handling fun edge cases and can be deferred until more of the plumbing works completely.	2019-06-27 13:38:21 -05:00
R.B. Boyer	ceef44bbc9	adding new config entries for L7 discovery chain (unused) (#5987 )	2019-06-27 12:37:43 -05:00
Todd Radel	a18b6d5ab9	connect: store signingKeyId instead of authorityKeyId (#6005 )	2019-06-27 16:47:22 +02:00
R.B. Boyer	f7fdf18335	fix test that was failing after #6013 (#6026 )	2019-06-27 09:31:19 -05:00
Aestek	81f8092a42	acl: allow service deregistration with node write permission (#5217 ) With ACLs enabled if an agent is wiped and restarted without a leave it can no longer deregister the services it had previously registered because it no longer has the tokens the services were registered with. To remedy that we allow service deregistration from tokens with node write permission.	2019-06-27 14:24:34 +02:00
Akshay Ganeshen	98a35fbe69	dns: support alt domains for dns resolution (#5940 ) this adds an option for an alt domain to be used with dns while migrating to a new consul domain.	2019-06-27 12:00:37 +02:00
hashicorp-ci	f4304e2e5b	Merge Consul OSS branch 'master' at commit `4eb73973b6`	2019-06-27 02:00:41 +00:00
Pierre Souchay	4eb73973b6	agent: added metadata information about servers into consul service description (#5455 ) This allows have information about servers from HTTP APIs without using the command line.	2019-06-26 23:46:47 +02:00
Sarah Christoff	d3d92d76f3	ui: modify content path (#5950 ) * Add ui-content-path flag * tests complete, regex validator on string, index.html updated * cleaning up debugging stuff * ui: Enable ember environment configuration to be set via the go binary at runtime (#5934) * ui: Only inject {{.ContentPath}} if we are makeing a prod build... ...otherwise we just use the current rootURL This gets injected into a <base /> node which solves the assets path problem but not the ember problem * ui: Pull out the <base href=""> value and inject it into ember env See previous commit: The <base href=""> value is 'sometimes' injected from go at index serve time. We pass this value down to ember by overwriting the ember config that is injected via a <meta> tag. This has to be done before ember bootup. Sometimes (during testing and development, basically not production) this is injected with the already existing value, in which case this essentially changes nothing. The code here is slightly abstracted away from our specific usage to make it easier for anyone else to use, and also make sure we can cope with using this same method to pass variables down from the CLI through to ember in the future. * ui: We can't use <base /> move everything to javascript (#5941) Unfortuantely we can't seem to be able to use <base> and rootURL together as URL paths will get doubled up (`ui/ui/`). This moves all the things that we need to interpolate with .ContentPath to the `startup` javascript so we can conditionally print out `{{.ContentPath}}` in lots of places (now we can't use base) * fixed when we serve index.html * ui: For writing a ContentPath, we also need to cope with testing... (#5945) ...and potentially more environments Testing has more additional things in a separate index.html in `tests/` This make the entire thing a little saner and uses just javascriopt template literals instead of a pseudo handbrake synatx for our templating of these files. Intead of just templating the entire file this way, we still only template `{{content-for 'head'}}` and `{{content-for 'body'}}` in this way to ensure we support other plugins/addons * build: Loosen up the regex for retrieving the CONSUL_VERSION (#5946) * build: Loosen up the regex for retrieving the CONSUL_VERSION 1. Previously the `sed` replacement was searching for the CONSUL_VERSION comment at the start of a line, it no longer does this to allow for indentation. 2. Both `grep` and `sed` where looking for the omment at the end of the line. We've removed this restriction here. We don't need to remove it right now, but if we ever put the comment followed by something here the searching would break. 3. Added `xargs` for trimming the resulting version string. We aren't using this already in the rest of the scripts, but we are pretty sure this is available on most systems. * ui: Fix erroneous variable, and also force an ember cache clean on build 1. We referenced a variable incorrectly here, this fixes that. 2. We also made sure that every `make` target clears ember's `tmp` cache to ensure that its not using any caches that have since been edited everytime we call a `make` target. * added docs, fixed encoding * fixed go fmt * Update agent/config/config.go Co-Authored-By: R.B. Boyer <public@richardboyer.net> * Completed Suggestions * run gofmt on http.go * fix testsanitize * fix fullconfig/hcl by setting correct 'want' * ran gofmt on agent/config/runtime_test.go * Update website/source/docs/agent/options.html.md Co-Authored-By: Hans Hasselberg <me@hans.io> * Update website/source/docs/agent/options.html.md Co-Authored-By: kaitlincarter-hc <43049322+kaitlincarter-hc@users.noreply.github.com> * remove contentpath from redirectFS struct	2019-06-26 11:43:30 -05:00
Pierre Souchay	0e907f5aa8	Support for maximum size for Output of checks (#5233 ) * Support for maximum size for Output of checks This PR allows users to limit the size of output produced by checks at the agent and check level. When set at the agent level, it will limit the output for all checks monitored by the agent. When set at the check level, it can override the agent max for a specific check but only if it is lower than the agent max. Default value is 4k, and input must be at least 1.	2019-06-26 09:43:25 -06:00
hashicorp-ci	4d185baf55	Merge Consul OSS branch 'master' at commit `88b15d84f9` skip-checks: true	2019-06-25 02:00:26 +00:00
Matt Keeler	813e009a2d	Prepare for having different service kinds that are all generic… (#6013 ) Default to internal error when service kind is unknown	2019-06-24 15:05:36 -04:00
Matt Keeler	43c5ba0304	New Cache Types (#5995 ) * Add a cache type for the Catalog.ListServices endpoint * Add a cache type for the Catalog.ListDatacenters endpoint	2019-06-24 14:11:34 -04:00
Matt Keeler	19e70c46bf	Ensure that looking for services by addreses works with Tagged Addresses (#5984 )	2019-06-21 13:16:17 -04:00
Matt Keeler	6cc1451895	Update some tests to fix ContentHash broken by the tagged service addresses (#5996 )	2019-06-20 11:50:18 -04:00
Hans Hasselberg	f13fe4b304	agent: transfer leadership when establishLeadership fails (#5247 )	2019-06-19 14:50:48 +02:00
Aestek	b839f52195	kv: do not trigger watches when setting the same value (#5885 ) If a KVSet is performed but does not update the entry, do not trigger watches for this key. This avoids releasing blocking queries for KV values that did not actually changed.	2019-06-18 15:06:29 +02:00
Aestek	24a0f2bba2	ae: use stale requests when performing full sync (#5873 ) Read requests performed during anti antropy full sync currently target the leader only. This generates a non-negligible load on the leader when the DC is large enough and can be offloaded to the followers following the "eventually consistent" policy for the agent state. We switch the AE read calls to use stale requests with a small (2s) MaxStaleDuration value and make sure we do not read too fast after a write.	2019-06-17 18:05:47 +02:00
Matt Keeler	f3d9b999ee	Add tagged addresses for services (#5965 ) This allows addresses to be tagged at the service level similar to what we allow for nodes already. The address translation that can be enabled with the `translate_wan_addrs` config was updated to take these new addresses into account as well.	2019-06-17 10:51:50 -04:00
Matt Keeler	2557d7a6cc	Fix CAS operations on Services (#5971 ) * Fix CAS operations on services * Update agent/consul/state/catalog_test.go Co-Authored-By: R.B. Boyer <public@richardboyer.net>	2019-06-17 10:41:04 -04:00
Paul Banks	acfcc7daf4	Add rate limiting to RPCs sent within a server instance too (#5927 )	2019-06-13 04:26:27 -05:00
Paul Banks	ffcfdf29fc	Upgrade xDS (go-control-plane) API to support Envoy 1.10. (#5872 ) * Upgrade xDS (go-control-plane) API to support Envoy 1.10. This includes backwards compatibility shim to work around the ext_authz package rename in 1.10. It also adds integration test support in CI for 1.10.0. * Fix go vet complaints * go mod vendor * Update Envoy version info in docs * Update website/source/docs/connect/proxies/envoy.md	2019-06-07 07:10:43 -05:00
Pierre Souchay	4a4c63bda0	Ensure Consul is IPv6 compliant (#5468 )	2019-06-04 10:02:38 -04:00
Matt Keeler	2ba6c3ac00	Update links to envoy docs on xDS protocol (#5871 )	2019-06-03 11:03:05 -05:00
R.B. Boyer	40336fd353	agent: fix several data races and bugs related to node-local alias checks (#5876 ) The observed bug was that a full restart of a consul datacenter (servers and clients) in conjunction with a restart of a connect-flavored application with bring-your-own-service-registration logic would very frequently cause the envoy sidecar service check to never reflect the aliased service. Over the course of investigation several bugs and unfortunate interactions were corrected: (1) local.CheckState objects were only shallow copied, but the key piece of data that gets read and updated is one of the things not copied (the underlying Check with a Status field). When the stock code was run with the race detector enabled this highly-relevant-to-the-test-scenario field was found to be racy. Changes: a) update the existing Clone method to include the Check field b) copy-on-write when those fields need to change rather than incrementally updating them in place. This made the observed behavior occur slightly less often. (2) If anything about how the runLocal method for node-local alias check logic was ever flawed, there was no fallback option. Those checks are purely edge-triggered and failure to properly notice a single edge transition would leave the alias check incorrect until the next flap of the aliased check. The change was to introduce a fallback timer to act as a control loop to double check the alias check matches the aliased check every minute (borrowing the duration from the non-local alias check logic body). This made the observed behavior eventually go away when it did occur. (3) Originally I thought there were two main actions involved in the data race: A. The act of adding the original check (from disk recovery) and its first health evaluation. B. The act of the HTTP API requests coming in and resetting the local state when re-registering the same services and checks. It took awhile for me to realize that there's a third action at work: C. The goroutines associated with the original check and the later checks. The actual sequence of actions that was causing the bad behavior was that the API actions result in the original check to be removed and re-added _without waiting for the original goroutine to terminate_. This means for brief windows of time during check definition edits there are two goroutines that can be sending updates for the alias check status. In extremely unlikely scenarios the original goroutine sees the aliased check start up in `critical` before being removed but does not get the notification about the nearly immediate update of that check to `passing`. This is interlaced wit the new goroutine coming up, initializing its base case to `passing` from the current state and then listening for new notifications of edge triggers. If the original goroutine "finishes" its update, it then commits one more write into the local state of `critical` and exits leaving the alias check no longer reflecting the underlying check. The correction here is to enforce that the old goroutines must terminate before spawning the new one for alias checks.	2019-05-24 13:36:56 -05:00
Freddy	6b31482333	Increase reliability of TestResetSessionTimerLocked_Renew	2019-05-24 13:54:51 -04:00
Pierre Souchay	e892981418	agent: Improve startup message to avoid confusing users when no error occurs (#5896 ) * Improve startup message to avoid confusing users when no error occurs Several times, some users not very familiar with Consul get confused by error message at startup: `[INFO] agent: (LAN) joined: 1 Err: <nil>` Having `Err: <nil>` seems weird to many users, I propose to have the following instead: * Success: `[INFO] agent: (LAN) joined: 1` * Error: `[WARN] agent: (LAN) couldn't join: %d Err: ERROR`	2019-05-24 16:50:18 +02:00
Freddy	17e74985b0	Run TestServer_Expect on its own (#5890 )	2019-05-23 19:52:33 -04:00
Freddy	6c19cacd42	Flaky test: ACLReplication_Tokens (#5891 ) * Exclude non-go workflows while testing * Wait for s2 global-management policy * Revert "Exclude non-go workflows while testing" This reverts commit `47a83cbe9f`.	2019-05-23 19:52:02 -04:00
Freddy	d4ea163b0b	Add retries to StatsFetcherTest (#5892 )	2019-05-23 19:51:31 -04:00
Jack Pearkes	40cec98468	Release v1.5.1	2019-05-22 20:19:12 +00:00
freddygv	40b809bce3	Wait for s2 global-management policy	2019-05-21 17:58:37 -06:00
Freddy	e9259ca97a	Change log line used for verification	2019-05-21 17:07:06 -06:00
Freddy	d1c315fad9	Stop running TestLeader_ChangeServerID in parallel	2019-05-21 15:28:08 -06:00
Sarah Christoff	32b5992d0f	Add retries around `obj`	2019-05-21 13:36:52 -05:00
Sarah Christoff	73d73e0e20	Add retries to all `obj`	2019-05-21 13:31:37 -05:00
Sarah Christoff	2a018e5e0a	Update agent/coordinate_endpoint_test.go Co-Authored-By: Freddy <freddygv@users.noreply.github.com>	2019-05-17 14:32:50 -05:00
Sarah Christoff	b96d9b01bd	Update type assertion logic Logic updated to evaluate with a boolean after the type assertion. This allows us to check if the type assertion succeeded and be more clear with errors.	2019-05-17 13:32:36 -05:00
Kyle Havlovitz	31bb9d67df	Set the dead node reclaim timer at 30s	2019-05-15 11:59:33 -07:00
Kyle Havlovitz	29eb83c9c2	Merge branch 'master' into change-node-id	2019-05-15 10:51:04 -07:00

... 2 3 4 5 6 ...

1813 Commits