ceph-ansible

Commit Graph

Author	SHA1	Message	Date
Guillaume Abrioux	1b49988e4a	do not use dev repo The branch 'master' of ceph/ceph has been renamed to 'main' Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2022-08-03 21:26:22 +02:00
Guillaume Abrioux	fd8aca866d	facts: fix set_radosgw_address.yml use `include_tasks` instead of `import_tasks`. Given that with `import_tasks` statements are preprocessed and the tasks that defines it hasn't been run yet, it will fail and complain like following: ``` The error was: 'ansible.vars.hostvars.HostVarsVars object' has no attribute '_interface' ``` Using `include_tasks` instead fixes this. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `434793e2fe`)	2022-07-06 03:52:54 +02:00
Guillaume Abrioux	07e6762abf	facts: fix deployments with different net interface names Deployments when radosgws don't have the same names for network interface. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2095605 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `f6b49f78a9`)	2022-07-06 03:52:54 +02:00
Guillaume Abrioux	4d2855414e	facts: follow up on `aa0cc93` when these variables are defined in the inventory host file, all tasks are skipped then because the node being played isn't aware about the values from the rgw nodes. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2063029 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `328bd7c975`)	2022-04-21 13:38:23 +02:00
Guillaume Abrioux	1dcc072978	facts: fix mgr/mon collocation `service dump` hangs when no active mgr is available. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2022-04-21 08:48:47 +02:00
Guillaume Abrioux	af5b3f51cc	dashboard: fix regression introduced by ceph/ceph-ansible/pull/7150 when no rgw is present, it fails. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2076192 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `1a56fd6a21`)	2022-04-21 08:48:47 +02:00
Guillaume Abrioux	f224782326	dashboard: support --limit execution with rgw When the following conditions are met: - rgw is deployed, - dashboard is deployed, - playbook is called with --limit, - a node being processed is collocated on either a mon or mgr. The playbook fails because `rgw_instances` is undefined. The idea here is to make sure this variable is always defined. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2063029 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `aa0cc9381d`)	2022-04-14 10:38:34 +02:00
Guillaume Abrioux	86ac9a8c48	dashboard: allow collecting stats from the host This commit makes podman bindmount `/:/rootfs:ro` so the container can collect data from the host. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2028775 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `0f34cd16d8`)	2022-04-14 00:37:38 +02:00
insatomcat	df8674a1c5	do not update Debian cache when package-install is disabled When deploying with --skip-tags=package-install (when there is no access to a repository), the playbook is still trying to update the package cache, which makes the playbook fail. This change prevents the playbook to try to update the cache when the package-install tag is skipped. Signed-off-by: Florent CARLI <florent.carli@rte-france.com> (cherry picked from commit `63f20f5941`)	2022-04-04 13:49:19 +02:00
Guillaume Abrioux	3dd918db23	dashboard: always set `dashboard_server_addr` When running the playbook with `--limit`, if the play targeted doesn't match hosts present in the mgr group the playbook can fail. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2063029 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `72e4654aae`)	2022-03-26 21:57:20 +01:00
Teoman ONAY	de447d168e	Turn off SELinux separation for containers MON and RGW Initially MONs and RGW binded /etc/pki/ca-trust/extracted using the :z flag (introduced to solve an OSP TripleO issue on RHEL - #3638) but using this flag prevents local services (like sssd) running on the host from accessing the certificates/files in that folder. Signed-off-by: Teoman ONAY <tonay@redhat.com> (cherry picked from commit `7e8ce2567e`)	2022-03-10 16:17:35 +01:00
Guillaume Abrioux	5618405b60	adopt: fix node labelling When using group of group, the playbook will apply undesired labels on nodes. This commit fixes it by applying only the expected labels. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2057528 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `266b6e739c`)	2022-03-04 12:51:14 +01:00
Teoman ONAY	10a5e54f8f	Enable user to change the account used for ssh connection By default cephadm uses root account to connect remotely to other nodes in the cluster. This change allows to choose another account. This commit also allows to use a dedicated subnet for cephadm mgmt. Signed-off-by: Teoman ONAY <tonay@redhat.com> (cherry picked from commit `da42f3d139`)	2022-03-04 12:51:14 +01:00
Benoît Knecht	4487e41a1e	ceph-facts: Fix get_def_crush_rule_name.yml in check mode This construct doesn't work as intended since ansible/ansible#74212: ``` item.stdout \| default('{}') \| from_json ``` That PR made the `command` module return `stdout` even in check mode (setting it to the empty string), so `default()` has no effect in that case and `from_json()` fails to parse an empty string. Instead, `default()` needs to be invoked with its second argument set to `True`, so that it replaces any `False` value (such as an empty string) with its first argument: ``` item.stdout \| default('{}', True) \| from_json ``` Signed-off-by: Benoît Knecht <bknecht@protonmail.ch> (cherry picked from commit `7684d892c0`)	2022-02-16 09:50:53 +01:00
Benoît Knecht	3ba0e4bdca	ceph-osd: Fix crush_rules.yml in check mode Set a default value for `item.stdout` before passing it to `from_json()`. The `when` condition doesn't prevent this template from being evaluated in check mode, so it fails if `item.stdout` doesn't contain a valid JSON string. Signed-off-by: Benoît Knecht <bknecht@protonmail.ch> (cherry picked from commit `ef05e9a313`)	2022-02-16 09:50:53 +01:00
Benoît Knecht	9df27fc5c5	ceph-osd: Fix start_osds.yml in check mode This construct doesn't work as intended since ansible/ansible#74212: ``` ceph_osd_ids.stdout \| default('{}') \| from_json ``` That PR made the `command` module return `stdout` even in check mode (setting it to the empty string), so `default()` has no effect in that case and `from_json()` fails to parse an empty string. Instead, `default()` needs to be invoked with its second argument set to `True`, so that it replaces any `False` value (such as an empty string) with its first argument: ``` ceph_osd_ids.stdout \| default('{}', True) \| from_json ``` Signed-off-by: Benoît Knecht <bknecht@protonmail.ch> (cherry picked from commit `0b3a608216`)	2022-02-16 09:50:53 +01:00
John Karasev	2e2d23c79e	ceph-grafana: Add proxy env vars to grafana service template When installing grafana plugins, the container will make http requests. This requires http proxy otherwise installation cannot be performed. Passed the proxy vars from all.yml as env args. Fixes: ceph#6484, ceph#6481 Signed-off-by: John Karasev <john.karasev@intel.com> (cherry picked from commit `79ca442d53`)	2022-02-09 11:35:17 +01:00
Danny Webb	000e93f608	make grafana network a configurable option Signed-off-by: Danny Webb <danny.webb@thehutgroup.com> (cherry picked from commit `189ff93372`)	2022-01-19 10:08:28 +01:00
Guillaume Abrioux	e083d9f62a	container: align systemd units with rpm Update `After=` and `Wants=` parameters in container systemd units and make them be aligned with the systemd units that come from the packaging. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2027440 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `f01536ea19`)	2021-12-15 13:49:49 +01:00
Guillaume Abrioux	53dc75d29c	validate: fix bug when using vault since a variable encrypted with vault is no longer a string but a encrypted object we can't use the filter \| length, we have to convert it to a string before. Fixes: #6991 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `6ad7e52869`)	2021-11-29 13:42:24 +01:00
Guillaume Abrioux	e63df909af	update: support --limit on monitor nodes Change needed in order to support --limit on mon nodes. Otherwise, a call to `hostvars[groups[mon_group_name][0]]['_current_monitor_address']` throws an error: ``` "The error was: 'ansible.vars.hostvars.HostVarsVars object' has no attribute '_current_monitor_address'" ``` Closes: https://bugzilla.redhat.com/show_bug.cgi?id=2014304#c28 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `82eee4303b`)	2021-11-03 08:48:51 +01:00
Seena Fallah	075b1a94d5	ceph-validate: export validate repository vars as a task Signed-off-by: Seena Fallah <seenafallah@gmail.com> (cherry picked from commit `4f6da9d92f`)	2021-10-18 18:38:47 +02:00
Seena Fallah	110b08c290	ceph-common: export repository configuration to a single task Signed-off-by: Seena Fallah <seenafallah@gmail.com> (cherry picked from commit `e79bda9a05`)	2021-10-18 18:38:47 +02:00
Guillaume Abrioux	5e40cb8957	tests: remove all references to ceph_stable_release this is legacy and not needed anymore. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `f277a39dfe`)	2021-10-02 15:50:24 +02:00
Seena Fallah	59c7238741	ceph-defaults: set ceph_stable_release default to the stable branch release ceph_stable_release is a legacy from the time where a single branch of ceph-ansible supported more than one release of ceph Signed-off-by: Seena Fallah <seenafallah@gmail.com> (cherry picked from commit `fb99626987`)	2021-10-02 15:50:24 +02:00
Alex Lambert	de17b232e6	dashboard: allow disabling of unused features Unconfigured dashboard features can lead to empty tabs in the dashboard containing no meaningful content. Allow users to disable dashboard features they know will not be used. A list of features to be disabled allows the user to define a streamlined dashboard as standard across deployments. Defaults to disabling no features, ensuring that users are sure they do not need the dashboard feature before disabling it. Signed-off-by: Alex Lambert <lamberta@microsoft.com> (cherry picked from commit `a9680ab17f`)	2021-09-29 14:28:26 +02:00
Dimitri Savineau	380d25a752	ceph-defaults: set quay.io as the default registry Because the ceph container images are now only pushed to the quay.io registry then this updates the default registry value. The docker.io registry can still be used but doesn't receive updated container images. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `e7b43c1fc6`)	2021-09-09 13:43:02 +02:00
Seena Fallah	688a673c48	ceph-container-engine: allow override container_package_name and container_service_name Only include specific variables when they are undefined Signed-off-by: Seena Fallah <seenafallah@gmail.com> (cherry picked from commit `95bce32270`)	2021-09-08 15:35:19 +02:00
Dimitri Savineau	6baa6e6b84	container: explicitly pull monitoring images We don't pull the monitoring container images (alertmanager, prometheus, node-exporter and grafana) in a dedicated task like we're doing for the ceph container image. This means that the container image pull is done during the start of the systemd service. By doing this, pulling the image behind a proxy isn't working with podman. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1995574 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `5bb7240f87`)	2021-08-23 16:08:16 -04:00
Guillaume Abrioux	6892e02a30	iscsi: don't set default value for trusted_ip_list It restricts access to the iSCSI API. It can be left empty if the API isn't going to be access from outside the gateway node Even though this seems to be a limited use case, it's better to leave it empty by default than having a meaningless default value. We could make this variable mandatory but that would be a breaking change. Let's just add a logic in the template in order to set this variable in the configuration file only if it was specified by users. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1994930 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> Co-authored-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `6802b8dddd`)	2021-08-19 12:06:50 -04:00
Guillaume Abrioux	afe442a18f	containers: introduce target systemd unit This adds ceph-*.target systemd unit files support for containerized deployments. This also fixes a regression introduced by PR #6719 (rgw and nfs systemd units not getting purged) Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1962748 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `09ef465f62`)	2021-08-18 13:42:56 -04:00
Guillaume Abrioux	e7d9d0a7d4	roles: remove leftover from pr #4319 pr #4319 introduced some uesless `become: true` on systemd tasks. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `1db8fa8989`)	2021-08-18 11:08:28 -04:00
Dimitri Savineau	a6b6706fdb	ceph-mon: do not log monitor keyring We don't want to display the keyring in the ansible log. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `e44075abd6`)	2021-08-12 13:31:00 +02:00
Guillaume Abrioux	5b30a72869	common: do not log keyring secret let's not display any keyring secret by default in ansible log. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1980744 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `7511195738`)	2021-08-11 17:01:09 -04:00
Dimitri Savineau	fa8b58fb33	ceph-dashboard: fix TLS cert openssl generation With OpenSSL version prior 1.1.1 (like CentOS 7 with 1.0.2k), the -addext doesn't exist. As a solution, this uses the default openssl.cnf configuration file as a template and add the subjectAltName in the v3_ca section. This temp openssl configuration file is removed after the TLS certificate creation. This patch also move the run_once statement at the block level. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1978869 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `5e0ace7e54`)	2021-08-09 15:14:38 -04:00
Guillaume Abrioux	fa16f6d923	dashboard: subj_alt_names fact refactor the current way the variable is built results in: ``` 2021-08-03 04:18:23,020 - ceph.ceph - INFO - ok: [ceph-sangadi-4x-indpt6-node1-installer] => changed=false ansible_facts: subj_alt_names: \|- subjectAltName=ceph-sangadi-4x-indpt6-node1-installer/subjectAltName=10.0.210.223/subjectAltName=ceph-sangadi-4x-indpt6-node1-installersubjectAltName=ceph-sangadi-4x-indpt6-node2/subjectAltName=10.0.210.252/subjectAltName=ceph-sangadi-4x-indpt6-node2/ ``` which is incorrect. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1978869 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `6f1a0634f7`)	2021-08-09 15:14:38 -04:00
Teoman ONAY	47149a5483	podman pids.max default value is 2048, docker's one is 4096 which are sufficient for the default value (512) of rgw thread pool size. But if its value is increased near to the pids-limit value, it does not leave place for the other processes to spawn and run within the container and the container crashes. pids-limit set to unlimited regardless of the container engine. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1987041 Signed-off-by: Teoman ONAY <tonay@redhat.com> (cherry picked from commit `9b5d97adb9`)	2021-08-05 11:04:24 -04:00
Dimitri Savineau	561a7c02c0	osds: use osd pool ls instead of osd dump command The ceph osd pool ls detail command is a subset of the ceph osd dump command. $ ceph osd dump --format json\|wc -c 10117 $ ceph osd pool ls detail --format json\|wc -c 4740 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `06471a4b82`)	2021-08-03 13:57:20 -04:00
Benoît Knecht	f9478472af	ceph-handler: Fix osd handler in check mode Run the Ceph commands that only gather information (without making any changes to the cluster) when running Ansible in check mode. This allows the tasks that depend on the variables set by those tasks to succeed in check mode. Signed-off-by: Benoît Knecht <bknecht@protonmail.ch> (cherry picked from commit `498acd7527`)	2021-08-02 15:54:04 +02:00
Dimitri Savineau	d7edc71fd5	ceph-defaults: update grafana dashboards source We currently download the grafana dashboars from the ceph@master branch for all ceph releases. We should use the right ceph branch according to the ceph release. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2021-07-27 11:44:50 -04:00
Dimitri Savineau	3e8d9b4a1f	ceph-defaults: add missing grafana dashboards The radosgw-sync-overview and rbd-details grafana dashboars were missing from the list. Closes: #6758 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `f0ccf3ebf0`)	2021-07-27 10:53:47 -04:00
Dimitri Savineau	f5ee8dfb26	alertmanager: allow disable dashboard tls verify When using self-signed/untrusted CA certificates, alertmanager displays an error in logs. With this commit this should make those messages disappear. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1936299 Co-authored-by: Guillaume Abrioux <gabrioux@redhat.com> Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `9f77b929d1`)	2021-07-25 22:02:16 -04:00
Dimitri Savineau	88e07f0bbc	multisite: use node fqdn for endpoints when https When the rgw_multisite_proto variable is set to https then we shoudn't use the IP address in the zone endpoints list but the node FQDN to match the TLS certificate CN. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1965504 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `ad05a08160`)	2021-07-22 22:48:03 +02:00
Dimitri Savineau	f9d60644ad	common: fix py2 pool_list from_json when skipped When using python 2 and the task with a loop is skipped then it generates an error. Unexpected templating type error occurred on ({{ (pool_list.stdout \| from_json)['pools'] }}): expected string or buffer Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `cf6e33346e`)	2021-07-21 08:57:53 -04:00
Guillaume Abrioux	f3a9135241	common: disable/enable pg_autoscaler The PG autoscaler can disrupt the PG checks so the idea here is to disable it and re-enable it back after the restart is done. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `13036115e2`)	2021-07-20 11:48:39 -04:00
Dimitri Savineau	7434157891	ceph-mgr: move mgr module list to common Populating the ceph_mgr_modules list in the mgr_modules doesn't make sense since that file is only executed if the list isn't empty or we're using the dashboard. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `cd06e7c046`)	2021-07-19 15:02:55 -04:00
Dimitri Savineau	925e3efc35	ceph-nfs: allow overriding NFS_CORE_PARAM We already have config override variables for existing block (like ganesha_ceph_export_overrides, ganesha_log_overrides, etc...) or a global one (ganesha_conf_overrides) but redefining the NFS_CORE_PARAM block in that variable will erase all previous values (currently only Bind_Addr). ganesha_core_param_overrides: \| Enable_UDP = false; NFS_Port = 2050; Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1941775 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com> (cherry picked from commit `9817d29543`)	2021-07-19 14:13:02 -04:00
Neelaksh Singh	9c04909d9c	Sensitive key data now hidden in output log Fixes: #6529 Signed-off-by: Neelaksh Singh <neelaksh48@gmail.com> (cherry picked from commit `d18a9860cd`)	2021-07-12 08:49:49 -04:00
Guillaume Abrioux	867376c30b	dashboard: remove "certificate is valid for" error When deploying dashboard with ssl certificates generated by ceph-ansible, we enforce the CN to 'ceph-dashboard' which can makes application such alertmanager complain like following: `err="Post https://mgr0:8443/api/prometheus_receiver: x509: certificate is valid for ceph-dashboard, not mgr0" context_err="context deadline exceeded"` The idea here is to add alternative names matching all mgr/mon instances in the certificate so this error won't appear in logs. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1978869 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `72a0336c71`)	2021-07-07 17:19:11 +02:00
Guillaume Abrioux	d5784c01c0	dashboard: support dedicated network for the dashboard This introduces a new variable `dashboard_network` in order to support deploying the dashboard on a different subnet. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1927574 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `f4f73b6197`)	2021-07-06 14:54:12 +02:00

1 2 3 4 5 ...

2857 Commits (1b49988e4aa311559287eb51dc4040584ad1472a)