actions-runner-controller

Commit Graph

Author	SHA1	Message	Date
Yusuke Kuoka	ab1c39de57	feat: HorizontalRunnerAutoscaler Webhook server (#282 ) * feat: HorizontalRunnerAutoscaler Webhook server This introduces a Webhook server that responds GitHub `check_run`, `pull_request`, and `push` events by scaling up matched HorizontalRunnerAutoscaler by 1 replica. This allows you to immediately add "resource slack" for future GitHub Actions job runs, without waiting next sync period to add insufficient runners. This feature is highly inspired by https://github.com/philips-labs/terraform-aws-github-runner. terraform-aws-github-runner can manage one set of runners per deployment, where actions-runner-controller with this feature can manage as many sets of runners as you declare with HorizontalRunnerAutoscaler and RunnerDeployment pairs. On each GitHub event received, the webhook server queries repository-wide and organizational runners from the cluster and searches for the single target to scale up. The webhook server tries to match HorizontalRunnerAutoscaler.Spec.ScaleUpTriggers[].GitHubEvent.[CheckRun\|Push\|PullRequest] against the event and if it finds only one HRA, it is the scale target. If none or two or more targets are found for repository-wide runners, it does the same on organizational runners. Changes: * Fix integration test * Update manifests * chart: Add support for github webhook server * dockerfile: Include github-webhook-server binary * Do not import unversioned go-github * Update README	2021-02-07 17:37:27 +09:00
Jesse Haka	28e80a2d28	Add support for enterprise runners (#290 ) * Add support for enterprise runners * update docs	2021-02-05 09:31:06 +09:00
Jonas Lergell	6c64ae6a01	Actually use 'dockerdContainerResources' to set resources on the dind container (#273 )	2021-01-29 09:18:28 +09:00
Yusuke Kuoka	ace95d72ab	Fix self-update failuers due to /runner/externals mount (#253 ) * Fix self-update failuers due to /runner/externals mount Fixes #252 * Tested Self-update Fixes (#269) Adding fixes to #253 as confirmed and tested in https://github.com/summerwind/actions-runner-controller/issues/264#issuecomment-764549833 by @jolestar, @achedeuzot and @hfuss 🙇 🍻 Co-authored-by: Hayden Fuss <wifu1234@gmail.com>	2021-01-24 10:58:35 +09:00
Johannes Nicolai	94e8c6ffbf	minReplicas <= desiredReplicas <= maxReplicas (#267 ) * ensure that minReplicas <= desiredReplicas <= maxReplicas no matter what * before this change, if the number of runners was much larger than the max number, the applied scale down factor might still result in a desired value > maxReplicas * if for resource constraints in the cluster, runners would be permanently restarted, the number of runners could go up more than the reverse scale down factor until the next reconciliation round, resulting in a situation where the number of runners climbs up even though it should actually go down * by checking whether the desiredReplicas is always <= maxReplicas, infinite scaling up loops can be prevented	2021-01-22 10:11:21 +09:00
ZacharyBenamram	48923fec56	Autoscaling: Percentage runners busy - remove magic number used for round up (#235 ) * remove magic number for autoscaling Co-authored-by: Zachary Benamram <zacharybenamram@blend.com>	2020-12-15 14:38:01 +09:00
ZacharyBenamram	466b30728d	Add "PercentageRunnersBusy" horizontal runner autoscaler metric type (#223 ) * hpa scheme based off busy runners * running make manifests Co-authored-by: Zachary Benamram <zacharybenamram@blend.com>	2020-12-13 08:48:19 +09:00
Yusuke Kuoka	dfffd3fb62	feat: EKS IAM Roles for Service Accounts for Runner Pods (#226 ) One of the pod recreation conditions has been modified to use hash of runner spec, so that the controller does not keep restarting pods mutated by admission webhooks. This naturally allows us, for example, to use IRSA for EKS that requires its admission webhook to mutate the runner pod to have additional, IRSA-related volumes, volume mounts and env. Resolves #200	2020-12-08 17:56:06 +09:00
Juho Saarinen	f710a54110	Don't compare runner connetion token at restart need check (#227 ) Fixes #143	2020-12-08 08:48:35 +09:00
Erik Nobel	a2b335ad6a	Github pkg: Bump github package to version 33 (#222 )	2020-12-06 10:01:47 +09:00
Shinnosuke Sawada	be25715e1e	Use TLS for secure docker connection (#192 )	2020-11-30 08:57:33 +09:00
Reinier Timmer	ee8fb5a388	parametrized working directory (#185 ) * parametrized working directory * manifests v3.0	2020-11-25 08:55:26 +09:00
Erik Nobel	4e93879b8f	[BUG?]: Create mountpoint for /externals/ (#203 ) * runner/controller: Add externals directory mount point * Runner: Create hack for moving content of /runner/externals/ dir * Externals dir Mount: mount examples for '__e/node12/bin/node' not found error	2020-11-25 08:53:47 +09:00
Shinnosuke Sawada	4371de9733	add dockerEnabled option (#191 ) Add dockerEnabled option for users who does not need docker and want not to run privileged container. if `dockerEnabled == false`, dind container not run, and there are no privileged container. Do the same as closed #96	2020-11-16 09:41:12 +09:00
Shinnosuke Sawada	a4061d0625	gofmt ed	2020-11-12 09:20:06 +09:00
Shinnosuke Sawada	83857ba7e0	use tcp DOCKER_HOST instead of sharing docker.sock	2020-11-12 08:07:52 +09:00
Yusuke Kuoka	e613219a89	Fix token registration broken since v0.11.0 (#167 ) Fixes #166	2020-11-11 09:38:05 +09:00
Dan Webb	dcf8524b5c	Adds RUNNER_GROUP argument to the runner registration (#157 ) * Adds RUNNER_GROUP argument to the runner registration Adds the ability to register a runner to a predefined runner_group Resolves #137 * Update README with runner group example - Updates the README with instructions of how to add the runner to a group - Fix code fencing for shell and yaml blocks in the README - Use consistent bullet points (dash not asterisk)	2020-11-10 17:15:54 +09:00
Juho Saarinen	f2a2ab7ede	Check token validity only when creating new pod (#159 ) Fixes #143	2020-11-10 17:02:30 +09:00
Juho Saarinen	40c5050978	Added support for other than public GitHub URL (#146 ) Refactoring a bit	2020-10-28 22:15:53 +09:00
Yusuke Kuoka	faaca10fba	Rename Runner.Spec.dockerWithinRunnerContainer to docker"d"WithinRunnerContainer (#134 ) * Rename Runner.Spec.dockerWithinRunnerContainer to dockerdWithinRunnerContainer Ref https://github.com/summerwind/actions-runner-controller/pull/126#issuecomment-712501790	2020-10-21 21:32:40 +09:00
Juho Saarinen	d16dfac0f8	Restart if pod ends up succeeded (#136 ) Fixes #132	2020-10-21 21:32:26 +09:00
Juho Saarinen	92920926fe	Configurable "runner and DinD in a single container" (#126 )	2020-10-20 08:48:28 +09:00
Brendan Galloway	7d0bfb77e3	Inject Env Vars into Runner defined Container Spec (#127 ) The runner token is now injected into the `runner` container defined within Runner.Spec.Containers[]	2020-10-20 08:43:53 +09:00
Dominic LoBue	a63860029a	Prefer autoscaling based on jobs rather than workflows if available (#114 ) Adds the ability to autoscale on jobs in addition to workflows. We fall back to using workflow metrics if job details are not present. Resolves #89	2020-10-08 09:00:44 +09:00
Yusuke Kuoka	1e466ad3df	Ensure controller-gen is up-to-date and the code and the manifests are in-sync Follow-up for #95 that added /finalizers subresource permission and #103 that upgraded controller-gen from 0.2.4 from 0.3.0	2020-10-06 09:23:03 +09:00
Helder Moreira	7a2fa7fbce	runner-controller: do not delete runner if it is busy (#103 ) Currently, after refreshing the token, the controller re-creates the runner with the new token. This results in jobs being interrupted. This PR makes sure the pod is not restarted if it is busy. Closes #74	2020-10-05 09:06:37 +09:00
Yusuke Kuoka	4733edc20d	Add scaling-down scenario to integration test	2020-08-02 16:10:01 +09:00
Yusuke Kuoka	50487bbb54	Fix the HRA controller name	2020-08-02 10:38:15 +09:00
Yusuke Kuoka	e2164f9946	Fix integration test bugs and do verify scaling out	2020-08-02 10:34:58 +09:00
Yusuke Kuoka	3c3077a11c	Fix crash on startup after the HRDA addition This is a follow-up for #66. The reconciler for the new HorizontalRunnerDeploymentAutoscaler had a terrible flaw that prevented the controller to fail launching due to an error like: ``` indexer conflict: map[field:.metadata.controller:{}] ``` This fixes that, while adding `integration_test.go` to verify its actually fixed and prevent regression in the future.	2020-07-29 21:20:46 +09:00
Moto Ishizawa	e10637ce35	Merge pull request #66 from summerwind/org-runner-autoscale feat: Organizational RunnerDeployment Autoscaling	2020-07-28 19:17:18 +09:00
Yusuke Kuoka	ae30648985	feat: Use HorizontalRunnerAutoscaler for autoscaling	2020-07-27 20:33:44 +09:00
David Liao	c0914743b0	add config to respect image pull policy	2020-07-08 23:53:52 -07:00
Yusuke Kuoka	eca6917c6a	feat: Organizational RunnerDeployment Autoscaling Enhances #57 to add support for organizational runners. As GitHub Actions does not have an appropriate API for this, this is the spec you need: ``` apiVersion: actions.summerwind.dev/v1alpha1 kind: RunnerDeployment metadata: name: myrunners spec: minReplicas: 1 maxReplicas: 3 autoscaling: metrics: - type: TotalNumberOfQueuedAndProgressingWorkflowRuns repositories: # Assumes that you have `github.com/myorg/myrepo1` repo - myrepo1 - myrepo2 template: spec: organization: myorg ``` It works by collecting "in_progress" and "queued" workflow runs for the repositories `myrepo1` and `myrepo2` to autoscale the number of replicas, assuming you have this organizational runner deployment only for those two repositories. For example, if `myrepo1` had 1 `in_progress` and 2 `queued` workflow runs, and `myrepo2` had 4 `in_progress` and 8 `queued` workflow runs at the time of running the reconcilation loop on the runner deployment, it will scale replicas to 1 + 2 + 4 + 8 = 15. Perhaps we might be better add a kind of "ratio" setting so that you can configure the controller to create e.g. 2x runners than demanded. But that's another story. Ref #10	2020-07-03 09:12:47 +09:00
KUOKA Yusuke	5bb2694349	feat: Repository-wide RunnerDeployment Autoscaling (#57 ) * feat: Repository-wide RunnerDeployment Autoscaling This adds `maxReplicas` and `minReplicas` to the RunnerDeploymentSpec. If and only if both fields are set, the controller computes and sets desired `replicas` automatically depending on the demand. The number of demanded runner replicas is computed by `queued workflow runs + in_progress workflow runs` for the repository. The support for organizational runners is not included. Ref https://github.com/summerwind/actions-runner-controller/issues/10	2020-06-27 17:26:46 +09:00
Moto Ishizawa	390f2a62d9	Merge pull request #50 from summerwind/runner-validation-webhook Add validation webhooks	2020-05-08 22:39:13 +09:00
Moto Ishizawa	f80c3c1928	Set volume to pod properly	2020-05-01 08:51:25 +09:00
Moto Ishizawa	e889eaeb04	Add validation webhooks	2020-04-30 22:11:59 +09:00
Reinier Timmer	b96979888c	fix delete pod when runner failed to register	2020-04-29 14:23:58 +09:00
Reinier Timmer	9f57f52e36	organization and repository are now exclusive	2020-04-28 11:14:31 +02:00
Reinier Timmer	8c5b776807	support runner labels	2020-04-28 11:14:31 +02:00
Reinier Timmer	eca3cc7941	add organization info to runner status	2020-04-28 11:14:31 +02:00
Reinier Timmer	fb35dd4131	support for organization runners	2020-04-28 11:14:31 +02:00
Moto Ishizawa	3b8ea2991c	Share runner's working directory with docker sidecar	2020-04-24 22:36:27 +09:00
Moto Ishizawa	3ccc51433f	Use github package to access the GitHub API	2020-04-13 22:28:07 +09:00
Yusuke Kuoka	b411d37f2b	fix: RunnerDeployment should clean up old RunnerReplicaSets ASAP Since the initial implementation of RunnerDeployment and until this change, any update to a runner deployment has been leaving old runner replicasets until the next resync interval. This fixes that, by continusouly retrying the reconcilation 10 seconds later to see if there are any old runner replicasets that can be removed. In addition to that, the cleanup of old runner replicasets has been improved to be deferred until all the runners of the newest replica set to be available. This gives you hopefully zero or at less downtime updates of runner deployments. Fixes #24	2020-04-04 07:55:12 +09:00
Moto Ishizawa	5efdc6efe6	Add permission to create/patch events resource	2020-03-27 23:25:37 +09:00
Aleksandr Stepanov	d4c849ee09	Add variants of PodTemplate spec fields into the Runner spec (#7 ) Resolves #5 Fixes #11 Fixes #12 Changes: * Added podtemplate spec * Rework pod creation logic * Added most using podspecs * Added copy of podspec * Fixed Github List method * Fixed containers * Added ability to override runner's containers * Added ability to override runner's containers * Added ability to override runner's containers * Update controllers/runner_controller.go Co-Authored-By: Moto Ishizawa <summerwind.jp@gmail.com> * Remove optional restartpolicy * Changed naming convention Co-authored-by: Moto Ishizawa <summerwind.jp@gmail.com>	2020-03-20 22:50:50 +09:00
Moto Ishizawa	b1da3092fb	Revert test comment	2020-03-15 21:55:38 +09:00

1 2

72 Commits