ceph-ansible

Commit Graph

Author	SHA1	Message	Date
Dimitri Savineau	f2d997396e	cephadm-adopt: use custom dashboard images cephadm uses default value for dashboard container images which need to be customized by ansible for upstream or downstream purpose. This feature wasn't present when cephadm-adopt.yml has been designed. Also set the container_image_base variable for upgrade purpose. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-10 16:00:24 +02:00
Dimitri Savineau	b38019e3ca	cephadm-adopt: run orch apply from monitors It looks like we can't run the ceph orch apply commands on nodes other than monitors even if it used to work in the past. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-10 16:00:24 +02:00
Dimitri Savineau	27efcbc0e5	cephadm-adopt: don't fail on systemd reset-failed If the systemd service exists successfully then we don't need to reset the failed state. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-10 16:00:24 +02:00
Dimitri Savineau	fd36433826	cephadm-adopt: copy client.admin keyring The ceph config assimilate-conf command requires the client.admin keyring which isn't present on all nodes most of the time. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-10 16:00:24 +02:00
Dimitri Savineau	14eed63921	tox: add cephadm_adopt scenario This adds an optional cephadm_adopt scenario which is based on all_daemons. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-10 16:00:24 +02:00
Guillaume Abrioux	0f3bae09bd	play: followup on `cc0d969` Remove two other pattern 'iscsigws' in main playbook that have been missed in `cc0d9697c5` Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-09 10:01:05 -04:00
Guillaume Abrioux	86edae724f	rgw: set container memory limit to 4g This commit changes the container memory limit for rgw daemons. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1707488 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-09 15:31:10 +02:00
Guillaume Abrioux	bcc673f66c	facts: refact `ceph_uid` fact There's no need to set this fact with a `set_fact` We can achieve this in `ceph-defaults` Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-09 13:37:29 +02:00
Guillaume Abrioux	f402ab2b87	ceph_volume: fix regression do not skip zapping if osd_fsid is passed Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-08 09:52:53 -04:00
Guillaume Abrioux	40307f810c	tests: add docker hub authentication in jobs This commit makes all jobs authenticating to docker hub in order to avoid the rate limit. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-08 09:52:53 -04:00
Guillaume Abrioux	cc0d9697c5	play: remove backward compatibility group name It's time to remove this old group name. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-08 09:21:19 -04:00
Dimitri Savineau	1438ca0120	ceph-nfs: change ganesha devel source The download.nfs-ganesha.org source for nfs-ganesha on CentOS isn't available anymore. Let's switch back to shaman since we have builds available now. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-06 16:59:25 +02:00
Dimitri Savineau	fc599ed9f5	tests: remove nfs_ganesha_stable_branch variable We don't need to override this variable in the group_vars but use the default value instead. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-06 16:58:59 +02:00
Guillaume Abrioux	5b6f5486f7	tests: update nfs-ganesha to V3.3-stable not really needed in master, commit intended to be backported in octopus branch. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-05 17:10:40 +02:00
Guillaume Abrioux	bbe30bcc69	doc: add a note about deprecated branches This commit adds a note about `stable-3.0` `stable-3.1` branches which are deprecated and not maintained anymore. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-03 09:27:12 +02:00
Guillaume Abrioux	e61488507b	doc: add a note about containerized deployments This commit updates the documentation to add a note about containerized deployments. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-03 09:27:12 +02:00
Guillaume Abrioux	5c254861bd	doc: fix warning treated as an error Typical error: ``` Warning, treated as error: /home/jenkins-build/build/workspace/ceph-ansible-docs-pull-requests/docs/source/day-2/upgrade.rst:2:Title underline too short. ``` Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-03 09:22:42 +02:00
Dimitri Savineau	93754bd70c	ceph-defaults: update nfs-ganesha to 3.3 nfs-ganesha 3.3 is the latest 3.x release available for octopus so we should update to this version. https://download.ceph.com/nfs-ganesha/rpm-V3.3-stable/octopus This will also match the version used in RHCS 5. Ceph container already uses that version too. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-03 06:36:49 +02:00
Dimitri Savineau	c95adc564b	facts: explicitly disable facter and ohai By default, ansible gathers facts from facter and ohai if installed on the remote nodes, given we don't need them, let's exclude these facts from our facts gathering Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-02 17:46:12 +02:00
Dimitri Savineau	1361e84a4e	radosgw: remove INST_PORT environment variable This variable isn't consumed by the container so we can remove it. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-07-02 16:52:29 +02:00
Guillaume Abrioux	7dd68b9ac1	rgw: fix multi instances scaleout When rgw and osd are collocated, the current workflow prevents from scaling out the radosgw_num_instances parameter when rerunning the playbook. The environment file used in the rgw systemd template is rendered when executing the `ceph-rgw` role but during a new run of the playbook (in order to scale out rgw instances), handlers are triggered from `ceph-osd` role which is run before `ceph-rgw`, therefore it tries to start the new rgw daemon whereas its corresponding environment file hasn't been rendered yet and fails like following: ``` ceph-radosgw@rgw.ceph4osd3.rgw1.service failed to run 'start-pre' task: No such file or directory ``` This commit moves the tasks generating this file in `ceph-config` role so it is generated early. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1851906 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-02 10:39:50 -04:00
Guillaume Abrioux	19097026fb	tests: enforce pytest-rerunfailures version This commit enforces the pytest-rerunfailures installed so it's <9.0 This is to avoid the following error: ``` ERROR: pytest-rerunfailures 9.0 has requirement pytest>=5.0, but you'll have pytest 4.6.11 which is incompatible. ``` latest version of pytest-rerunfailures isn't compatible with the version of pytest we are using. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-07-02 15:57:39 +02:00
Dimitri Savineau	72293b6614	vagrant: update centos image to 8.2 CentOS 8.2 (2004) has been relesed so we should switch to this image when using vagrant. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-30 13:56:55 +02:00
Jan Fajerski	d90834b77f	ceph-volume.py: add support for batch refactored code See https://github.com/ceph/ceph/pull/34740 for the batch changes. Signed-off-by: Jan Fajerski <jfajerski@suse.com>	2020-06-30 09:46:27 +02:00
Dimitri Savineau	3592ba1d61	ceph-common: remove copr and sepia repositories All EL8 dependencies are now present on EPEL 8 so we don't need the additional repositories that were only a temporary solution. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-30 08:35:19 +02:00
Guillaume Abrioux	8f9cdf4b10	rolling_update: add any_errors_fatal If a failure occurs in ceph-validate, the upgrade playbook keeps running where we expect it to fail. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-29 12:58:53 -04:00
George Shuklin	3e87f53875	Add container settings for Ubuntu 20 (the same as Ubuntu 18) Signed-off-by: George Shuklin <george.shuklin@gmail.com>	2020-06-29 12:18:58 -04:00
Dimitri Savineau	548ff26256	Add playbook for converting cluster to cephadm The commit adds a new playbook for converting an existing ceph cluster deployed by ceph-ansible to the cephadm orchestrator. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-29 09:21:38 -04:00
Dimitri Savineau	03cd75845f	dashboard: configure mgr backend before restart We need to set the mgr dashboard server ip address before restarting the dashboard module otherwise we can try to bind the dashboard module on an already used address. We already do this configuration for the dashboard port value and ssl setup so we should do the same for server address too. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1851455 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-29 14:59:01 +02:00
Jonathan Rosser	42884e8175	Ansible tests are not filters The use of "\| success" and "\| changed" are not valid syntax for modern ansible releases. Signed-off-by: Jonathan Rosser <jonathan.rosser@rd.bbc.co.uk>	2020-06-26 12:26:25 -04:00
Jonathan Rosser	92288c11c5	Install python routes package as a dependancy rather than directly This is now a dependancy of ceph-mgr so will be installed automatically and does not need a specific task. This change means that ceph-mgr installs correctly on Ubuntu Focal where the python3-routes package is necessary. Signed-off-by: Jonathan Rosser <jonathan.rosser@rd.bbc.co.uk>	2020-06-26 12:26:25 -04:00
Guillaume Abrioux	b7539eb275	dashboard: copy self-signed generated crt to mons This commit makes the playbook copying self-signed generated certificate to monitors. When mons and mgrs are deployed on dedicated nodes the playbook will fail when trying to import certificate and key files since they are generated on mgrs whereas we try to import them from a monitor. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1846995 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-23 09:37:21 -04:00
Dimitri Savineau	d43769dc2a	podman: Add Type and PIDFile value to unit files This changes the way we are running the podman containers via systemd. They are now in dettached mode and Type/PIDFile set. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1834974 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-23 09:37:50 +02:00
Guillaume Abrioux	3f47236470	ceph_volume: make zap function idempotent This commit makes the zap function idempotent, especially when using lvm_volumes variable. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1845668 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-22 22:16:29 -04:00
Dimitri Savineau	bd22f1d1ec	docker: Add Requires on docker service When using docker container engine then the systemd unit scripts only use a dependency on the docker daemon via the After parameter. But if docker is restarted on a live system then the ceph systemd units should wait for the docker daemon to be fully restarted. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1846830 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-22 23:08:50 +02:00
Guillaume Abrioux	37b20b6525	docker2podman: make images pulling optional This commit makes the images pulling skipped if podman isn't installed on the machine. In OSP context, the podman installation is done later in the workflow, it means all `podman pull` commands will fail. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1849559 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-22 12:19:38 -04:00
Dimitri Savineau	296aa13b3c	travis: use tests/requirements.txt Explicitly install ansible-lint pytest pytest-cov via pip results of a specific pytest version (4.3.1) which is not supported for pytest-cov (2.10). Because we are already defining a specific pytest version in the tests requirements then we can install all the python dependencies from that file and remove this from the pip install command. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-19 18:10:55 -04:00
Guillaume Abrioux	1525990f39	requirements: exclude ansible 2.9.10 ansible 2.9.10 seems to have introduced a bug. See https://github.com/ansible/ansible/issues/70168 This commit excludes this version from ceph-ansible requirements. Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-19 17:32:33 -04:00
Dimitri Savineau	e41487dbce	docs: Add upgrade operation. This commit adds a chapter about the ceph upgrade process. Closes: #5393 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-18 17:56:53 +02:00
Dimitri Savineau	829990e60d	ceph-osd: remove ceph-osd-run.sh script Since we only have one scenario since nautilus then we can just move the container start command from ceph-osd-run.sh to the systemd unit service. As a result, the ceph-osd-run.sh.j2 template and the ceph_osd_docker_run_script_path variable are removed. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-18 17:51:13 +02:00
Dimitri Savineau	d67759611e	library/ceph_pool: set name parameter as required The name parameter is required. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-17 16:29:39 +02:00
Dimitri Savineau	0f8a61a3ae	debian/uca: remove the handler notification The "update apt cache" in the ceph-handler role was never called and the handler trigger after adding the uca repository doesn't exist at all. Instead of using a handler for that we can just set the update_cache parameter to true like the other apt_repository tasks. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-17 10:14:03 +02:00
Guillaume Abrioux	b91d60d384	switch_to_containers: don't set noup flag We shouldn't set this flag when running switch_to_containers playbook. Otherwise the playbook fails waiting for pgs to be clean. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1843569 Signed-off-by: Guillaume Abrioux <gabrioux@redhat.com>	2020-06-17 01:32:18 +02:00
Jan Fajerski	1fe8e819f9	lvm_setup: lookup device from inventory, default to /dev/sd* names This fixes a long standing fail in ceph-volumes lvm test suite. Otherwise the default behaviour should not change. Signed-off-by: Jan Fajerski <jfajerski@suse.com>	2020-06-16 18:17:34 +02:00
Dimitri Savineau	cdb30bd125	container: inspect Id field instead of RepoDigests When a container image managed by podman isn't tag anymore then the RepoDigests field when inspecting the image doesn't return any value. This is different from docker workflow and it breaks the ceph-ansible container upgrade when collocated multiple services and using a non fix container tag (like latest or 4). $ podman images REPOSITORY TAG IMAGE ID CREATED SIZE docker.io/ceph/daemon latest 680c9c0d38c3 8 days ago 957 MB <none> <none> 011ee108bfc9 2 months ago 1.01 GB $ podman inspect 680c9c0d38c3 \| jq .[0].RepoDigests[0] "docker.io/ceph/daemon@sha256:20cf789235e23ddaf38e109b391d1496bb88011239d16862c4c106d0e05fea9e" $ podman inspect 011ee108bfc9 \| jq .[0].RepoDigests[0] null Because this field returns "null" then the ansible task trying to determine this value is failing ----------------------------- fatal: [foo]: FAILED! => msg: \|- The task includes an option with an undefined variable. The error was: None has no element 0 The error appears to be in 'roles/ceph-container-common/tasks/fetch_image.yml': line 137, column 3, but may be elsewhere in the file depending on the exact syntax problem. The offending line appears to be: - name: set_fact ceph_osd_image_repodigest_before_pulling ^ here ----------------------------- We don't have this behaviour with docker. $ docker images REPOSITORY TAG IMAGE ID CREATED SIZE docker.io/ceph/daemon latest 680c9c0d38c3 8 days ago 928 MB docker.io/ceph/daemon <none> 011ee108bfc9 2 months ago 986 MB $ docker inspect 680c9c0d38c3 \| jq .[0].RepoDigests[0] "docker.io/ceph/daemon@sha256:45e6f28bb67c81b826acb64fad5c0da1cac3dffb41a88992fe4ca2be79575fa6" $ docker inspect 011ee108bfc9 \| jq .[0].RepoDigests[0] "docker.io/ceph/daemon@sha256:b393a73309d72e43ca7d65cd3519036007947671e373eb59aa75a46185c52231" Instead we should just get the Id field. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1844496 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-16 17:06:25 +02:00
Dimitri Savineau	50140c9b5d	switch_to_container: fix osd systemd regex The systemd LOAD and ACTIVE fileds could have more than one space between both values. This update the systemd regex the same way we're using it in different part of the code. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1843500 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-16 17:04:06 +02:00
Ali Maredia	0175c205fa	rgw multisite: add master zone endpoints to zonegroup We were only adding the endpoints to the master zone but not to the zonegroup. This patch fixes the issue. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1839228 Signed-off-by: Ali Maredia <amaredia@redhat.com>	2020-06-09 09:50:18 -04:00
Dimitri Savineau	2f17f36638	mergify: remove merge on skip ci This rule will probably never be applyied and at the moment this is creating a cancelled job in the CI status. Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-06-09 09:30:10 -04:00
Ansible Deployment User	3f906e0c26	rgwloadbalancer undefined index variable The vrrp_instances variable is using a loop with index but the index_var wasn't defined. As a result, the fact task was failing on this undefined index variable. The task includes an option with an undefined variable. The error was: 'index' is undefined Closes: #5395 Signed-off-by: Florian Faltermeier <florian.faltermeier@uibk.ac.at>	2020-05-26 10:03:25 -04:00
Dimitri Savineau	44e1ebaaff	ceph-nfs: add stable noarch repository When using the stable nfs ganesha repository, we need have both arch and noarch repositories enabled. Currently the noarch repository is missing which cause the non containerized deployment to fail. Closes: #5375 Signed-off-by: Dimitri Savineau <dsavinea@redhat.com>	2020-05-16 07:34:08 +02:00

1 2 3 4 5 ...

5337 Commits (33a544644a671c8b9ffd7c5e761276c1a1ac574d) All Branches Search

5337 Commits (33a544644a671c8b9ffd7c5e761276c1a1ac574d)

All Branches