Commit graph

8561 commits

Author SHA1 Message Date
pokeeffe-molecula
2e0812ed98 variable fix 2022-01-04 18:09:16 -06:00
pokeeffe-molecula
ca5633c594 add uniqueness to cluster prefix 2022-01-04 16:36:11 -06:00
pokeeffe-molecula
cfacc72c53 smoking or non-smoking? 2022-01-04 15:00:39 -06:00
pokeeffe-molecula
094f70a91a
Merge pull request #1839 from molecula/pipeline-scheduling
Pipeline scheduling
2022-01-04 14:15:29 -06:00
pokeeffe-molecula
13b6f6f133 re-enabling actual test 2022-01-04 13:02:45 -06:00
pokeeffe-molecula
a2be201009 fixed yaml fubar 2022-01-04 12:37:05 -06:00
pokeeffe-molecula
c06d21a189 rules it is.. 2022-01-04 12:33:45 -06:00
pokeeffe-molecula
624975e123 getting scheduling to work 2022-01-04 11:29:00 -06:00
Samir Patel
8690160dd4
Merge pull request #1812 from molecula/54mir/authentication
[FB-1014] Authentication
2022-01-04 09:43:56 -05:00
Samir Patel
ab4e4ac216
Merge branch 'master' into 54mir/authentication 2022-01-04 09:28:58 -05:00
pokeeffe-molecula
c5a75fcd66
Merge pull request #1838 from molecula/get-pipeline-to-run
just run it all the time for now
2022-01-03 23:20:22 -06:00
pokeeffe-molecula
cc9c2761be
Merge branch 'master' into get-pipeline-to-run 2022-01-03 23:19:34 -06:00
pokeeffe-molecula
b4da2be804 just run it all the time for now 2022-01-03 23:19:07 -06:00
pokeeffe-molecula
3e31035e29
Merge pull request #1837 from molecula/get-pipeline-to-run
disabling single node deploy
2022-01-03 22:46:09 -06:00
pokeeffe-molecula
65a2deb842
Merge branch 'master' into get-pipeline-to-run 2022-01-03 22:45:13 -06:00
pokeeffe-molecula
44d635f407 disabling single node deploy 2022-01-03 22:43:25 -06:00
Samir Patel
d537398568
Merge branch 'master' into 54mir/authentication 2022-01-03 23:33:45 -05:00
pokeeffe-molecula
5f8b78ba9a
Merge pull request #1836 from molecula/get-pipeline-to-run
fix gitlab pipeline; disable circleci; remove artifactory
2022-01-03 22:22:06 -06:00
pokeeffe-molecula
56584905c2 fix gitlab pipeline; disable circleci; remove artifactory 2022-01-03 22:11:30 -06:00
pokeeffe-molecula
61c0ef9a1e
Merge pull request #1835 from molecula/fix-failed-deploy-linux-node
Fix failed deploy linux node
2022-01-03 21:50:21 -06:00
Fletcher Haynes
b9e8d3a103 Commented out a failing test as a meta-test 2022-01-03 19:17:11 -08:00
Fletcher Haynes
70a3af97a8 Fixed some variables in the CI file 2022-01-03 19:05:39 -08:00
Fletcher Haynes
c0708e5403 Changed the env vars a CI job was accessing 2022-01-03 16:48:42 -08:00
tgruben
16600a219c
Merge branch 'master' into 54mir/authentication 2022-01-03 17:55:58 -06:00
Fletcher Haynes
58a70863eb
Merge pull request #1823 from molecula/fb901
Fb901
2022-01-03 15:51:24 -08:00
Fletcher Haynes
a40d974377
Merge branch 'master' into fb901 2022-01-03 14:20:08 -08:00
pokeeffe-molecula
51ad67e06a
Update qa/tf/gauntlet/samsung/README.md
Co-authored-by: reese <45641995+reesporte@users.noreply.github.com>
2022-01-03 16:09:26 -06:00
pokeeffe-molecula
804b7b3146
Update qa/tf/README.md
Co-authored-by: reese <45641995+reesporte@users.noreply.github.com>
2022-01-03 16:06:54 -06:00
Ben Johnson
379c5e0bd1
Merge pull request #1832 from molecula/rbf-no-panic
Avoid panics in RBF debug tooling
2022-01-03 13:43:10 -07:00
Fletcher Haynes
e89c04acaf
Merge branch 'master' into fb901 2022-01-03 12:25:26 -08:00
Fletcher Haynes
4016a1d03d Fixed commenting in the gitlab CI file. 2022-01-03 12:21:32 -08:00
garrison.davis@molecula.com
fa9ae13363 Adds gauntlet testing framework for Samsung
This adds the Terraform needed to create a gauntlet testing framework for a cluster that is a mirror of Samsung's. It is meant to be run once a day in CI via the GitLab scheduler.
2022-01-03 12:15:07 -08:00
Ben Johnson
af9795aa1a Avoid panics in RBF debug tooling 2022-01-03 13:13:02 -07:00
tgruben
d5dcdcd40a
Merge branch 'master' into 54mir/authentication 2022-01-03 12:45:15 -06:00
reese
080ea6ad2e
Merge pull request #1806 from molecula/percentile-timestamp-decimal
FB-1095 implement percentiles on timestamp/decimal
2022-01-03 12:26:05 -06:00
tgruben
fb5765676b
Merge branch 'master' into 54mir/authentication 2022-01-03 12:15:10 -06:00
reesporte
b13538e426 update doc comment 2022-01-03 11:16:14 -06:00
reesporte
72adb177ae Merge branch 'master' into percentile-timestamp-decimal 2022-01-03 11:12:23 -06:00
reesporte
6e3ce01ecb explicitly test that getScaledInt works with timestamps 2022-01-03 11:11:54 -06:00
reesporte
fa2391b948 explicitly test untested path of valcountize 2022-01-03 11:11:27 -06:00
reesporte
99f6a1c113 change min to val
bc it could be used for things besides mins
2022-01-03 10:45:23 -06:00
Fletcher Haynes
93b97b9831 Test push to see if pipeline is running on push to master 2021-12-30 17:00:59 -08:00
reesporte
9b77432952 fix bad formatting 2021-12-29 14:22:29 -06:00
reesporte
be66103c45 requirements when auth is enabled
postgres binding is turned off
TLS must be turned on
2021-12-29 11:20:19 -06:00
reesporte
2847c22a4c linter things 2021-12-29 11:06:58 -06:00
reesporte
42f3557c55 fix merge conflicts 2021-12-29 09:05:53 -06:00
Travis Turner
56b9e2aba7
Merge pull request #1831 from molecula/tlt/ignore-down
Stop blocking API called when cluster is DOWN or DEGRADED
2021-12-28 14:14:29 -06:00
Travis
ffd91137e1
Stop blocking API called when cluster is DOWN or DEGRADED
This commit effectively removes the API-level validation that was
blocking certain API methods when the cluster was in a particular state
(namely DOWN and DEGRADED). The thinking is that we shouldn't be
blocking these requests at the API level, but rather should let them
pass through and allow the fact that a node is ACTUALLY down dictate the
behavior.

With this change, two tests were modified. They were previously
expecting the error message from the API validation on DOWN, but now
they check for a "shard unavailable" error, which is what gets returned
for a particular query when the cluster is in an unhealthy state.
2021-12-28 13:52:04 -06:00
Matthew Jaffee
6fd985c8eb
Merge pull request #1828 from molecula/errant-print
fix retry period and change client DialTimeout for commands (e.g. restore/backup)
2021-12-28 13:51:28 -06:00
Matthew Jaffee
1a8c10d5f3 fix backup fail test so it actually fails
A few things were going wrong here.

First, we take a "RetryPeriod" option on backup and restore which is
meant to be roughly the total amount of time we spend retrying any
given request before failing. However we were incorrectly passing that
as the RetryMaxWait which is the maximum amount of time to sleep
between any two attempts. We now do some fuzzy math to figure out
approximately how many attempts we should make given a minimum sleep
of 100ms and the fact that we double the sleep time every attempt.

Second, during the backup test, if a host was totally stopped when we
started the request, it would fail immediately and then retry, but if
the host was stopped during the request (after DNS had resolved), then
the request would wait for the DialTimeout which we default to 30s, so
turning off the cluster for 5 seconds and turning it back on resulted
in the backup completing rather than failing. Because of this, we
change the commandClient to have a default dial timeout of 1 second.

I was tempted to change the global default to 1s which I think would
be fine, but didn't want to break anything too badly.
2021-12-28 13:31:42 -06:00