Skip to content
New issue

Have a question about this project? Sign up for a free GitHub account to open an issue and contact its maintainers and the community.

By clicking “Sign up for GitHub”, you agree to our terms of service and privacy statement. We’ll occasionally send you account related emails.

Already on GitHub? Sign in to your account

roachtest: cdc/cloud-sink-gcs/assume-role failed #86100

Closed
cockroach-teamcity opened this issue Aug 14, 2022 · 11 comments
Closed

roachtest: cdc/cloud-sink-gcs/assume-role failed #86100

cockroach-teamcity opened this issue Aug 14, 2022 · 11 comments
Assignees
Labels
branch-master Failures and bugs on the master branch. C-test-failure Broken test (automatically or manually discovered). O-roachtest O-robot Originated from a bot. T-cdc
Milestone

Comments

@cockroach-teamcity
Copy link
Member

cockroach-teamcity commented Aug 14, 2022

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ d25cb57ccd9bc643ce9058ebd2057cab36b69ad5:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
	cdc.go:1752,cdc.go:289,cdc.go:856,test_runner.go:896: max latency was more than allowed: 1m26.55749219s vs 1m0s

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

/cc @cockroachdb/cdc

This test on roachdash | Improve this report!

Jira issue: CRDB-18571

Epic CRDB-11732

@cockroach-teamcity cockroach-teamcity added branch-master Failures and bugs on the master branch. C-test-failure Broken test (automatically or manually discovered). O-roachtest O-robot Originated from a bot. release-blocker Indicates a release-blocker. Use with branch-release-2x.x label to denote which branch is blocked. labels Aug 14, 2022
@cockroach-teamcity cockroach-teamcity added this to the 22.2 milestone Aug 14, 2022
@blathers-crl blathers-crl bot added the T-cdc label Aug 14, 2022
@HonoreDB HonoreDB removed the release-blocker Indicates a release-blocker. Use with branch-release-2x.x label to denote which branch is blocked. label Aug 29, 2022
@HonoreDB
Copy link
Contributor

Removed the release-blocker tag as this is a test-only issue (I'll leave it up to @samiskin whether to actually close this).

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 389661e823c19f318fa07ec2278336262531692d:

		  |   220.0s        0          747.0          768.3     21.0     31.5     46.1     75.5 withdrawal
		  | _elapsed___errors__ops/sec(inst)___ops/sec(cum)__p50(ms)__p95(ms)__p99(ms)_pMax(ms)
		  |   221.0s        0          782.9          768.5     19.9     29.4     39.8     50.3 deposit
		  |   221.0s        0          781.9          768.4     19.9     29.4     41.9     67.1 withdrawal
		  |   222.0s        0          809.1          768.7     18.9     27.3     35.7     75.5 deposit
		  |   222.0s        0          844.1          768.7     18.9     28.3     37.7     83.9 withdrawal
		  |   223.0s        0          796.9          768.8     19.9     28.3     39.8     75.5 deposit
		  |   223.0s        0          804.9          768.9     19.9     29.4     37.7     48.2 withdrawal
		  |   224.0s        0          807.0          769.0     18.9     28.3     39.8     83.9 deposit
		  |   224.0s        0          810.0          769.1     19.9     28.3     35.7     52.4 withdrawal
		  |   225.0s        0          843.1          769.3     18.9     27.3     37.7     56.6 deposit
		  |   225.0s        0          805.1          769.2     18.9     27.3     35.7     50.3 withdrawal
		  |   226.0s        0          252.8          767.0     19.9     29.4     41.9     52.4 deposit
		  |   226.0s        0          265.8          767.0     19.9     25.2     29.4     39.8 withdrawal
		Wraps: (4) secondary error attachment
		  | UNCLASSIFIED_PROBLEM: context canceled
		  | (1) UNCLASSIFIED_PROBLEM
		  | Wraps: (2) Node 4. Command with error:
		  |   | ``````
		  |   | ./workload run ledger --mix=balance=0,withdrawal=50,deposit=50,reversal=0 {pgurl:1-3} --duration=30m
		  |   | ``````
		  | Wraps: (3) context canceled
		  | Error types: (1) errors.Unclassified (2) *hintdetail.withDetail (3) *errors.errorString
		Wraps: (5) context canceled
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *cluster.WithCommandDetails (4) *secondary.withSecondaryError (5) *errors.errorString

	monitor.go:127,cdc.go:300,cdc.go:847,test_runner.go:917: monitor failure: monitor task failed: read tcp 172.17.0.3:60534 -> 34.139.6.78:26257: read: connection reset by peer
		(1) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).WaitE
		  | 	main/pkg/cmd/roachtest/monitor.go:115
		  | main.(*monitorImpl).Wait
		  | 	main/pkg/cmd/roachtest/monitor.go:123
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.cdcBasicTest
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:300
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.registerCDC.func10
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:847
		  | [...repeated from below...]
		Wraps: (2) monitor failure
		Wraps: (3) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).wait.func2
		  | 	main/pkg/cmd/roachtest/monitor.go:171
		  | runtime.goexit
		  | 	GOROOT/src/runtime/asm_amd64.s:1594
		Wraps: (4) monitor task failed
		Wraps: (5) read tcp 172.17.0.3:60534 -> 34.139.6.78:26257
		Wraps: (6) read
		Wraps: (7) connection reset by peer
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *withstack.withStack (4) *errutil.withPrefix (5) *net.OpError (6) *os.SyscallError (7) syscall.Errno

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ bc2e47da0523b347c28cf024707e80cd35d6c98a:

		  |    81.0s        0          774.0          716.0     19.9     29.4     39.8     46.1 deposit
		  |    81.0s        0          760.0          715.8     19.9     30.4     46.1     88.1 withdrawal
		  |    82.0s        0          813.0          717.2     19.9     28.3     37.7     62.9 deposit
		  |    82.0s        0          809.0          717.0     19.9     27.3     37.7     50.3 withdrawal
		  |    83.0s        0          804.0          718.2     18.9     27.3     39.8     54.5 deposit
		  |    83.0s        0          840.0          718.4     18.9     27.3     35.7     60.8 withdrawal
		  |    84.0s        0          778.9          719.0     19.9     28.3     37.7     46.1 deposit
		  |    84.0s        0          777.9          719.1     19.9     29.4     41.9     60.8 withdrawal
		  |    85.0s        0          799.1          719.9     19.9     26.2     37.7     52.4 deposit
		  |    85.0s        0          813.1          720.2     19.9     28.3     37.7     54.5 withdrawal
		  |    86.0s        0          837.9          721.3     18.9     29.4     41.9     58.7 deposit
		  |    86.0s        0          802.9          721.2     18.9     26.2     35.7     83.9 withdrawal
		  |    87.0s        0          265.7          716.0     19.9     28.3     37.7     50.3 deposit
		  |    87.0s        0          259.7          715.9     19.9     29.4     44.0     83.9 withdrawal
		Wraps: (4) secondary error attachment
		  | UNCLASSIFIED_PROBLEM: context canceled
		  | (1) UNCLASSIFIED_PROBLEM
		  | Wraps: (2) Node 4. Command with error:
		  |   | ``````
		  |   | ./workload run ledger --mix=balance=0,withdrawal=50,deposit=50,reversal=0 {pgurl:1-3} --duration=30m
		  |   | ``````
		  | Wraps: (3) context canceled
		  | Error types: (1) errors.Unclassified (2) *hintdetail.withDetail (3) *errors.errorString
		Wraps: (5) context canceled
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *cluster.WithCommandDetails (4) *secondary.withSecondaryError (5) *errors.errorString

	monitor.go:127,cdc.go:300,cdc.go:847,test_runner.go:917: monitor failure: monitor task failed: read tcp 172.17.0.3:46474 -> 34.148.200.54:26257: read: connection reset by peer
		(1) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).WaitE
		  | 	main/pkg/cmd/roachtest/monitor.go:115
		  | main.(*monitorImpl).Wait
		  | 	main/pkg/cmd/roachtest/monitor.go:123
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.cdcBasicTest
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:300
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.registerCDC.func10
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:847
		  | [...repeated from below...]
		Wraps: (2) monitor failure
		Wraps: (3) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).wait.func2
		  | 	main/pkg/cmd/roachtest/monitor.go:171
		  | runtime.goexit
		  | 	GOROOT/src/runtime/asm_amd64.s:1594
		Wraps: (4) monitor task failed
		Wraps: (5) read tcp 172.17.0.3:46474 -> 34.148.200.54:26257
		Wraps: (6) read
		Wraps: (7) connection reset by peer
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *withstack.withStack (4) *errutil.withPrefix (5) *net.OpError (6) *os.SyscallError (7) syscall.Errno

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 773568fbda06ba9be9fb1bc34a331f21c8891ffa:

		  |   879.0s        0          807.0          795.7     19.9     28.3     35.7     56.6 withdrawal
		  |   880.0s        0          793.0          795.7     19.9     30.4     39.8     71.3 deposit
		  |   880.0s        0          795.0          795.7     19.9     29.4     39.8     62.9 withdrawal
		  | _elapsed___errors__ops/sec(inst)___ops/sec(cum)__p50(ms)__p95(ms)__p99(ms)_pMax(ms)
		  |   881.0s        0          813.0          795.7     18.9     28.3     37.7     65.0 deposit
		  |   881.0s        0          822.0          795.8     18.9     28.3     37.7     56.6 withdrawal
		  |   882.0s        0          814.1          795.8     19.9     28.3     37.7     67.1 deposit
		  |   882.0s        0          798.1          795.8     19.9     28.3     37.7     48.2 withdrawal
		  |   883.0s        0          767.3          795.7     18.9     28.3     35.7     71.3 deposit
		  |   883.0s        0          768.3          795.7     19.9     29.4     44.0     65.0 withdrawal
		  |   884.0s        0            0.0          794.8      0.0      0.0      0.0      0.0 deposit
		  |   884.0s        0            0.0          794.8      0.0      0.0      0.0      0.0 withdrawal
		  |   885.0s        0            0.0          793.9      0.0      0.0      0.0      0.0 deposit
		  |   885.0s        0            0.0          794.0      0.0      0.0      0.0      0.0 withdrawal
		Wraps: (4) secondary error attachment
		  | UNCLASSIFIED_PROBLEM: context canceled
		  | (1) UNCLASSIFIED_PROBLEM
		  | Wraps: (2) Node 4. Command with error:
		  |   | ``````
		  |   | ./workload run ledger --mix=balance=0,withdrawal=50,deposit=50,reversal=0 {pgurl:1-3} --duration=30m
		  |   | ``````
		  | Wraps: (3) context canceled
		  | Error types: (1) errors.Unclassified (2) *hintdetail.withDetail (3) *errors.errorString
		Wraps: (5) context canceled
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *cluster.WithCommandDetails (4) *secondary.withSecondaryError (5) *errors.errorString

	monitor.go:127,cdc.go:300,cdc.go:847,test_runner.go:917: monitor failure: monitor task failed: read tcp 172.17.0.3:50604 -> 35.196.46.116:26257: read: connection reset by peer
		(1) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).WaitE
		  | 	main/pkg/cmd/roachtest/monitor.go:115
		  | main.(*monitorImpl).Wait
		  | 	main/pkg/cmd/roachtest/monitor.go:123
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.cdcBasicTest
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:300
		  | github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests.registerCDC.func10
		  | 	github.com/cockroachdb/cockroach/pkg/cmd/roachtest/tests/cdc.go:847
		  | [...repeated from below...]
		Wraps: (2) monitor failure
		Wraps: (3) attached stack trace
		  -- stack trace:
		  | main.(*monitorImpl).wait.func2
		  | 	main/pkg/cmd/roachtest/monitor.go:171
		  | runtime.goexit
		  | 	GOROOT/src/runtime/asm_amd64.s:1594
		Wraps: (4) monitor task failed
		Wraps: (5) read tcp 172.17.0.3:50604 -> 35.196.46.116:26257
		Wraps: (6) read
		Wraps: (7) connection reset by peer
		Error types: (1) *withstack.withStack (2) *errutil.withPrefix (3) *withstack.withStack (4) *errutil.withPrefix (5) *net.OpError (6) *os.SyscallError (7) syscall.Errno

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 762c1b86fe1c3a70338cee4a91c9e2e4c5e0fcfe:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
(cdc.go:1749).assertValid: max latency was more than allowed: 1m20.633487s vs 1m0s

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_encrypted=false , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

Same failure on other branches

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 0725273ac7f789ba8ed78aacaf73cc953ca47fe8:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
(cdc.go:373).newChangefeed: failed to create changefeed: pq: Use of CHANGEFEED requires an enterprise license. Your evaluation license expired on December 30, 2022. If you're interested in getting a new license, please contact subscriptions@cockroachlabs.com and we can help you out.
(cluster.go:1933).Run: output in run_072409.497885806_n4_workload_run_tpcc: ./workload run tpcc --warehouses=50 --duration=30m  {pgurl:1-3}  returned: context canceled
(cdc.go:283).Close: error shutting down prometheus/grafana: context canceled

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_encrypted=false , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

Same failure on other branches

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 0725273ac7f789ba8ed78aacaf73cc953ca47fe8:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
(cdc.go:373).newChangefeed: failed to create changefeed: pq: Use of CHANGEFEED requires an enterprise license. Your evaluation license expired on December 30, 2022. If you're interested in getting a new license, please contact subscriptions@cockroachlabs.com and we can help you out.
(cluster.go:1933).Run: output in run_072442.777328908_n4_workload_run_tpcc: ./workload run tpcc --warehouses=50 --duration=30m  {pgurl:1-3}  returned: context canceled
(cdc.go:283).Close: error shutting down prometheus/grafana: context canceled

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_encrypted=false , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

Same failure on other branches

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 0725273ac7f789ba8ed78aacaf73cc953ca47fe8:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
(cdc.go:373).newChangefeed: failed to create changefeed: pq: Use of CHANGEFEED requires an enterprise license. Your evaluation license expired on December 30, 2022. If you're interested in getting a new license, please contact subscriptions@cockroachlabs.com and we can help you out.
(cluster.go:1933).Run: output in run_073131.985780557_n4_workload_run_tpcc: ./workload run tpcc --warehouses=50 --duration=30m  {pgurl:1-3}  returned: context canceled
(cdc.go:283).Close: error shutting down prometheus/grafana: context canceled

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_encrypted=false , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

Same failure on other branches

This test on roachdash | Improve this report!

@cockroach-teamcity
Copy link
Member Author

roachtest.cdc/cloud-sink-gcs/assume-role failed with artifacts on master @ 1d7bd69205c2197ccac33df9e2e6d4ff8c0fdbcf:

test artifacts and logs in: /artifacts/cdc/cloud-sink-gcs/assume-role/run_1
(cdc.go:373).newChangefeed: failed to create changefeed: pq: Use of CHANGEFEED requires an enterprise license. Your evaluation license expired on December 30, 2022. If you're interested in getting a new license, please contact subscriptions@cockroachlabs.com and we can help you out.
(cluster.go:1933).Run: output in run_073018.764536711_n4_workload_run_tpcc: ./workload run tpcc --warehouses=50 --duration=30m  {pgurl:1-3}  returned: context canceled
(cdc.go:283).Close: error shutting down prometheus/grafana: context canceled

Parameters: ROACHTEST_cloud=gce , ROACHTEST_cpu=16 , ROACHTEST_encrypted=false , ROACHTEST_ssd=0

Help

See: roachtest README

See: How To Investigate (internal)

Same failure on other branches

This test on roachdash | Improve this report!

@miretskiy
Copy link
Contributor

@samiskin any updates on this?
Going to move this back to triage;

@miretskiy
Copy link
Contributor

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Labels
branch-master Failures and bugs on the master branch. C-test-failure Broken test (automatically or manually discovered). O-roachtest O-robot Originated from a bot. T-cdc
Projects
None yet
Development

No branches or pull requests

4 participants