ceph-ansible

Commit Graph

Author	SHA1	Message	Date
Neha Ojha	89d950fd3c	infrastructure-playbooks/lv-create.yml: add lvm_volumes to suggested paste Signed-off-by: Neha Ojha <nojha@redhat.com> (cherry picked from commit `65fdad0723`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Neha Ojha	1a0f7baf21	infrastructure-playbooks/lv-create.yml: copy without using a template file Signed-off-by: Neha Ojha <nojha@redhat.com> (cherry picked from commit `50a6d8141c`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Neha Ojha	f1245e6011	infrastructure-playbooks/lv-create.yml: don't use action to copy Signed-off-by: Neha Ojha <nojha@redhat.com> (cherry picked from commit `186c4e11c7`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Neha Ojha	21902f0113	infrastructure-playbooks: standardize variable usage with a space after brackets Signed-off-by: Neha Ojha <nojha@redhat.com> (cherry picked from commit `9d43806df9`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Neha Ojha	fb06c6cb80	vars/lv_vars.yaml: remove journal_device Signed-off-by: Neha Ojha <nojha@redhat.com> (cherry picked from commit `e0293de3e7`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Ali Maredia	10da777634	infrastructure-playbooks: playbooks for creating LVs for bucket indexes and journals These playbooks create and tear down logical volumes for OSD data on HDDs and for a bucket index and journals on 1 NVMe device. Users should follow the guidelines set in var/lv_vars.yaml After the lv-create.yml playbook is run, output is sent to /tmp/logfile.txt for copy and paste into osds.yml Signed-off-by: Ali Maredia <amaredia@redhat.com> (cherry picked from commit `1f018d8612`) Signed-off-by: Sébastien Han <seb@redhat.com>	2018-08-16 17:01:41 +02:00
Sébastien Han	28fc45e346	Revert "osd: generate device list for osd_auto_discovery on rolling_update" This reverts commit `e84f11e99e`. This commit was giving a new failure later during the rolling_update process. Basically, this was modifying the list of devices and started impacting the ceph-osd itself. The modification to accomodate the osd_auto_discovery parameter should happen outside of the ceph-osd. Also we are trying to not play ceph-osd role during the rolling_update process so we can speed up the upgrade. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `3149b2564f`)	2018-08-16 13:35:23 +00:00
Sébastien Han	6f1499800f	rolling_update: register container osd units Before running the upgrade, let's call systemd to collect unit names instead of relaying on the device list. This is more accurate and fix the osd_auto_discovery scenario too. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1613626 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `dad10e8f3f`)	2018-08-16 13:35:23 +00:00
Sébastien Han	51de29046b	contrib: fix generate group_vars samples For ceph-iscsi-gw and ceph-rbd-mirror roles the group_name are named differently (by default) than the role name so we have to change the script to generate the correct name. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1602327 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `315ab08b16`)	2018-08-14 17:51:41 +00:00
Jeffrey Zhang	19c7ca1983	Use /var/lib/ceph/osd folder to filter osd mount point In some case, use may mount a partition to /var/lib/ceph, and umount it will be failure and no need to do so too. Signed-off-by: Jeffrey Zhang <zhang.lei.fly@gmail.com> (cherry picked from commit `85cc61a6d9`)	2018-08-14 14:55:56 +00:00
Mike Christie	c44638ae7e	stable 3.1 igw: add api setting support Port the parts of this upstream commit: commit `91bf53ee93` Author: Sébastien Han <seb@redhat.com> Date: Fri Mar 23 11:24:56 2018 +0800 ceph-iscsi: support for containerize deployment that allows configuration of API settings in roles/ceph-iscsi-gw/templates/iscsi-gateway.cfg.j2 using the iscsi-gws.yml. This fixes Red Hat BZ: https://bugzilla.redhat.com/show_bug.cgi?id=1613963 Signed-off-by: Mike Christie <mchristi@redhat.com>	2018-08-14 10:23:12 +02:00
Mike Christie	2b76e3771d	stable 3.1 igw: enable and start rbd-target-api Backport https://github.com/ceph/ceph-ansible/pull/2984 to stable 3.1. From upstream commit: commit `1164cdc002` Author: Guillaume Abrioux <gabrioux@redhat.com> Date: Thu Aug 2 11:58:47 2018 +0200 iscsigw: install ceph-iscsi-cli package installs the cli package but does not start and enable the rbd-target-api daemon needed for gwcli to communicate with the igw nodes. This just enables and starts it. This fixes Red Hat BZ https://bugzilla.redhat.com/show_bug.cgi?id=1613963. Signed-off-by: Mike Christie <mchristi@redhat.com>	2018-08-14 10:23:12 +02:00
Sébastien Han	e7596d565f	group_vars: resync missing options resync group_vars file with the defaults/main.yml files. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `2dd75a1e6e`)	2018-08-13 18:55:06 +02:00
Guillaume Abrioux	904a0a4017	fail if fqdn deployment attempted fqdn configuration possibility caused a lot of trouble, it's adding a lot of complexity because of multiple cases and the relation between ceph-ansible and ceph-container. Moreover, there is no benefit for such a feature. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1613155 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2018-08-13 18:55:06 +02:00
Guillaume Abrioux	97cf08e897	config: ensure rgw section has the correct name the ceph.conf.j2 always assumes the hostname used to register the radosgw in the servicemap is equivalent to `{{ ansible_hostname }}` which returns the shortname form. We need to detect which form of the hostname was used in case of already deployed cluster and update the ceph.conf accordingly. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1580408 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `f422efb1d6`)	2018-08-13 18:55:06 +02:00
Guillaume Abrioux	95c28e78d1	mgr: backward compatibility for module management Follow up on `3abc253fec` The structure had even changed within `luminous` release. It was first: ``` { "enabled_modules": [ "balancer", "dashboard", "restful", "status" ], "disabled_modules": [ "influx", "localpool", "prometheus", "selftest", "zabbix" ] } ``` Then it changed for: ``` { "enabled_modules": [ "status" ], "disabled_modules": [ "balancer", "dashboard", "influx", "localpool", "prometheus", "restful", "selftest", "zabbix" ] } ``` and finally: ``` { "enabled_modules": [ "status" ], "disabled_modules": [ { "name": "balancer", "can_run": true, "error_string": "" }, { "name": "dashboard", "can_run": true, "error_string": "" } ] } ``` Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `36942af698`)	2018-08-13 16:05:21 +00:00
Guillaume Abrioux	9a013ab333	tests: resync iscsigw group name with master let's align the name of that group in stable-3.1 with master branch. Not having the same group name on different branches is confusing and make some nightlies job failing in the CI. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2018-08-13 12:24:59 +02:00
Guillaume Abrioux	32ef06e80f	tests: fix a typo in testinfra for iscsigws and jewel scenario group name for iscsi-gw nodes in testing is `iscsi-gws`. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2018-08-13 12:24:59 +02:00
Sébastien Han	8ea9d14050	osd: generate device list for osd_auto_discovery on rolling_update rolling_update relies on the list of devices when performing the restart of the OSDs. The task that is builind the devices list out of the ansible_devices dict only runs when there are no partitions on the drives. However during an upgrade the OSD are already configured, they have been prepared and have partitions so this task won't run and thus the devices list will be empty, skipping the restart during rolling_update. We now run the same task under different requirements when rolling_update is true and build a list when: * osd_auto_discovery is true * rolling_update is true * ansible_devices exists * no dm/lv are part of the discovery * the device is not removable * the device has more than 1 sector Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1613626 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `e84f11e99e`)	2018-08-10 16:30:40 +02:00
Sébastien Han	4785799110	rolling_update: add role ceph-iscsi-gw Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1575829 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `e91648a7af`)	2018-08-10 14:38:19 +02:00
Sébastien Han	12083bdab4	mon: fix calamari initialisation If calamari is already installed and ceph has been upgraded to a higher version the initialisation will fail later. So if we detect the calamari-server is too old compare to ceph_rhcs_version we try to update it. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1601755 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `4c9e24a90f`)	2018-08-10 14:15:16 +02:00
Sébastien Han	651058bd1b	rgw: remove useless condition The include does not need a condition on containerized_deployment since we are already in an include than has the same condition. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `5a89479abe`)	2018-08-09 15:38:17 +02:00
Sébastien Han	eba9547a6e	rgw: remove unused file copy_configs.yml was not including and is a leftover so let's remove it. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `3bce117de2`)	2018-08-09 15:38:17 +02:00
Sébastien Han	a16dc0e1de	rgw: ability to use ceph-ansible vars into containers Since the container now simply reads the ceph.conf, we remove all the unnecessary options. Also this PR is the foundation to support multiple backend, such as the new 'beast' from Ceph Mimic. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1582411 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `4d64dd4686`) # Conflicts: # roles/ceph-rgw/tasks/docker/main.yml	2018-08-09 15:38:17 +02:00
Ken Dreyer	1a2c6a3572	common: upgrade/install ceph-test deb first When we deploy a Jewel cluster on Ubuntu with ceph_test: True, we're unable to upgrade that cluster to Luminous. "apt-get install ceph-common" fails to upgrade to luminous if a jewel ceph-test package is installed: Some packages could not be installed. This may mean that you have requested an impossible situation or if you are using the unstable distribution that some required packages have not yet been created or been moved out of Incoming. The following information may help to resolve the situation: The following packages have unmet dependencies: ceph-base : Breaks: ceph-test (< 12.2.2-14) but 10.2.11-1xenial is to be installed ceph-mon : Breaks: ceph-test (< 12.2.2-14) but 10.2.11-1xenial is to be installed In ceph-ansible master, we resolve this whole class of problem by installing all the packages in one operation (see `b338fafd90`). For the stable-3.1 branch, take a less-invasive approach, and upgrade ceph-test prior to any other package. This matches the approach I took for RPMs in `3752cc6f38`, before we had the better solution in `b338fafd90`. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1610997 Signed-off-by: Ken Dreyer <kdreyer@redhat.com>	2018-08-09 14:39:33 +02:00
Graeme Gillies	19958f5c27	Allow mgr bootstrap keyring to be defined In environments where we wish to have manual/greater control over how the bootstrap keyrings are used, we need to able to externally define what the mgr keyring secret will be and have ceph-ansible use it, instead of it being autogenerated Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1610213 Signed-off-by: Graeme Gillies <ggillies@akamai.com> (cherry picked from commit `a46025820d`)	2018-08-09 08:25:27 +00:00
Sébastien Han	b00d2d0439	Resync rhcs_edits.txt We were missing an option so let's add it back. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1519835 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `19518656a7`)	2018-08-08 15:54:32 +02:00
Sébastien Han	a31ce962f7	test: remove osd_crush_location from shrink scenarios This is not needed since this is already covered by docker_cluster and centos_cluster scenarios. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `50be3fd9e8`)	2018-08-07 19:09:58 +00:00
Sébastien Han	b76c7c3afe	test: follow up on osd_crush_location for containers This was fixed by `578aa5c2d5` on non-container, we need to apply the same fix for containers. Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `77d4023fbe`)	2018-08-07 19:09:58 +00:00
Guillaume Abrioux	9403a3df09	iscsigw: install ceph-iscsi-cli package Install ceph-iscsi-cli in order to provide the `gwcli` command tool. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1602785 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `1164cdc002`)	2018-08-07 09:46:25 +02:00
Artur Fijalkowski	290035171f	Fix in regular expression matching OSD ID on non-contenerized deployment. restart_osd_daemon.sh is used to discover and restart all OSDs on a host. To do it the scripts loops the list of ceph-osd@ services in the system. This commit fixes bug in the regular expression responsile for extraction of OSDs - prior version uses `[0-9]{1,2}` expression which is ignoring all OSDS which numbers are greater than 99 (thus longer than 2 digits). Fix removed upper limit of digits in the number. This problem existed in two places in the script. Closes: #2964 Signed-off-by: Artur Fijalkowski <artur.fijalkowski@ing.com> (cherry picked from commit `52d9d406b1`)	2018-08-06 18:50:39 +00:00
Guillaume Abrioux	706d0b8289	defaults: backward compatibility with fqdn deployments This commit ensures we are backward compatible with fqdn deployments. Since ceph-container enforces deployment to be done with shortname, we must keep backward compatibility with clusters already deployed with fqdn configuration Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `0a6ff6bbf8`)	2018-08-06 14:09:35 +00:00
Sébastien Han	31dd4eeecf	rolling_update: set osd sortbitwise upgrade RHCS 2 -> RHCS 3 will fail if cluster has still set sortnibblewise, it stay stuck on "TASK [waiting for clean pgs...]" as RHCS 3 osds will not start if nibblewise is set. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1600943 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `b3266c5be2`)	2018-08-02 14:53:06 +00:00
Sébastien Han	2d5ed5ef8e	config: enforce socket name This was introduced by `59ee2e8d3b` and made our socket checks impossible to run. The PID could be found, but the cctid cannot. This happens during upgrade to mimic and on cluster running on mimic. So let's force the admin socket the way it was so we can properly check for existing instances also the line $cluster-$name.$pid.$cctid.asok is only needed when running multiple instances of the same daemon, thing ceph-ansible cannot do at the time of writing Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1610220 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `ea9e60d48d`)	2018-08-02 12:34:48 +00:00
Guillaume Abrioux	826da2c385	tests: support update scenarios in test_rbd_mirror_is_up() `test_rbd_mirror_is_up()` is failing on update scenarios because it assumes the `ceph_stable_release` is still set to the value of the original ceph release, it means it won't enter in the right part of the condition and fails. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `d8281e50f1`)	2018-08-02 10:06:55 +00:00
Mike Christie	99f84f88af	igw: fix image removal during purge We were not passing in the ceph conf info into the rbd image removal command, so if the clustername was not the default igw purge would fail due to the rbd rm command failing. This just fixes the bug by passing in the ceph conf info which has the clustername to use. This fixes Red Hat bugzilla: https://bugzilla.redhat.com/show_bug.cgi?id=1601949 Signed-off-by: Mike Christie <mchristi@redhat.com> (cherry picked from commit `d572a9a602`)	2018-07-31 10:09:08 +02:00
Mike Christie	f3f734f8f3	igw: do not fail purge on rbd removal errors Instead of failing the entire purge operation when the rbd command fails just log an error. This will allow the higher level target and config cleanup to complete, and the user only has to manually delete the rbd images. Signed-off-by: Mike Christie <mchristi@redhat.com> (cherry picked from commit `6f72f96dad`)	2018-07-31 10:09:08 +02:00
Sébastien Han	47fb07afca	osd: do not remove expose_partition container The container runs with --rm which means it will be deleted by Docker when exiting. Also 'docker rm -f' is not idempotent and returns 1 if the container does not exist. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1609007 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `2ca8c51906`)	2018-07-30 12:09:34 +00:00
Guillaume Abrioux	11d96d0897	ceph-osds: backward compatibility with jewel for osp pools creation If we want to be backward compatible with release prior to luminous, we have to set the rule name accordingly to default values used in jewel. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `053709da97`)	2018-07-30 10:30:19 +02:00
Guillaume Abrioux	75d68040c8	rbd-mirror: bring back compatibility with jewel deployment rbd-mirror can't start when deploying jewel because it needs admin keyring. Getting back this task brings backward compatibility for jewel deployment. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `1ecbbbdcfa`)	2018-07-30 10:30:19 +02:00
Guillaume Abrioux	1ab0f2ce06	iscsigw: do not run common roles when deploying jewel Let's not deploy common roles when iscsigw nodes for jewel deployment. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `a1ca2c8fd3`)	2018-07-30 10:30:19 +02:00
Guillaume Abrioux	40135a5621	tests: leave an OSD node in default crush root jewel used to create a default `rbd` pool in the default crush root `default`, we need to have at least 1 osd to satisfy the PGs for this created pool, otherwise the cluster will be in HEALTH_ERR state because of `pgs stuck unclean`/`pgs stuck inactive` Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `578aa5c2d5`)	2018-07-30 10:30:19 +02:00
Mike Christie	04ec87f31b	ceph ansible 3.1 igw: fix rbd-target-gw startup The problem is rbd-target-gw needs the rbd pool to be created, keyring to be copied over, and the iscsi-gateway.cfg to be setup before starting the rbd-target-gw service. In the master branch this is fixed by this commit: commit `91bf53ee93` Author: Sébastien Han <seb@redhat.com> Date: Fri Mar 23 11:24:56 2018 +0800 ceph-iscsi: support for containerize deployment where the needed setup tasks are done in common.yml which is done before prerequisites.yml. To avoid porting all those changes to 3.1 this patch just moves the rbd-target-gw startup to configure_iscsi.yml after everything has been setup. This fixes red hat bz: https://bugzilla.redhat.com/show_bug.cgi?id=1601325 Signed-off-by: Mike Christie <mchristi@redhat.com>	2018-07-27 11:38:46 +00:00
Sébastien Han	154b1dcc74	rgw: add more config option for civetweb frontend In containerized deployments we now inherite from the radosgw_civetweb_options options when bootstrapping the container. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1582411 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `e2ea5bac51`)	2018-07-25 15:46:03 +00:00
Giulio Fidente	9184b794dd	Run creation of empty rados index object to first monitor When distributing ceph-nfs role, creation of rados index object fails as it assumes availability of client.admin locally. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1607970 Signed-off-by: Giulio Fidente <gfidente@redhat.com> (cherry picked from commit `e85e5ea781`)	2018-07-25 13:01:13 +00:00
Guillaume Abrioux	a4c87e2079	tests: add mimic support in stable-3.1 Add mimic support in stable-3.1 branch so we can test it in nightlies CI jobs. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2018-07-24 18:22:42 +02:00
Guillaume Abrioux	3d19e1bc24	tests: do not deploy all daemons for shrink osds scenarios Let's create a dedicated environment for these scenarios, there is no need to deploy everything. By the way, doing so will save some times. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `b89cc1746f`)	2018-07-23 18:31:17 +02:00
Sébastien Han	36f24a8054	shrink-osd: purge osd on containerized deployment Prior to this commit we were only stopping the container, but now we also purge the devices. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1572933 Signed-off-by: Sébastien Han <seb@redhat.com> (cherry picked from commit `ce1dd8d2b3`)	2018-07-23 14:32:03 +02:00
Guillaume Abrioux	a3fd9c8550	tests: stop hardcoding ansible version In addition to ceph/ceph-build#1082 Let's set the ansible version in each ceph-ansible branch's respective requirements.txt. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2018-07-19 15:38:14 +00:00
Guillaume Abrioux	9db4a23d47	tests: add latest-bis-jewel for jewel tests since no latest-bis-jewel exists, it's using latest-bis which points to ceph mimic. In our testing, using it for idempotency/handlers tests means upgrading from jewel to mimic which is not what we want do. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com> (cherry picked from commit `05852b0301`)	2018-07-17 11:39:49 +00:00

1 2 3 4 5 ...

3856 Commits (f53eca31267710c83545ea5272cc2de22c7b7597) All Branches Search

3856 Commits (f53eca31267710c83545ea5272cc2de22c7b7597)

All Branches