Earlier  
Posted Nick Remark
#openstack-nova - 2022-11-09
10:24:10 samuelkunkel[m] (hope it is readable now)
10:25:18 frickler seem bauzas was the last one working on it
10:25:51 samuelkunkel[m] currently I will use the quick fix provided https://review.opendev.org/c/openstack/nova/+/838976
10:26:43 bauzas frickler: yup, I need to update my change
10:26:53 bauzas it's a priority I have
10:38:36 samuelkunkel[m] that sounds nice, if you need somebody to test - feel free to reach out to me, have some nodes with mdevs to play on
11:56:41 ygk_12345 HI all
12:04:39 sean-k-mooney samuelkunkel[m]: we not only plan to fix that but backport the fix to wallaby as we require it for our downstream product that far and there is no point doing it downstream only since the fix is backpoartable
12:05:19 sean-k-mooney so given your on yoga that shoudl hopefully also adress your usecase
12:05:33 samuelkunkel[m] yes, that sounds great
12:05:48 samuelkunkel[m] I assume there is currently no estimation possible on a timeframe?
12:06:17 sean-k-mooney well the patch thats propsoed actully works we just need a few comments adressed
12:06:45 auniyal Hi sean-k-mooney
12:07:02 sean-k-mooney downstream we have a dealine of mid decemebr to adress this so i am stongly hoping that we can adress this upstream before then so our product team does not start asking me about it
12:07:11 auniyal how can we run tox functional locally in train branch
12:07:20 auniyal tox -e functional fails
12:07:27 sean-k-mooney use the python3 version
12:07:51 sean-k-mooney or a vm/container based on ubutu 18.04?
12:07:52 samuelkunkel[m] I can second that, it also works on my yoga setup on a non productive cluster. Thanks for the clarification. Until the fix is backported I just use the patch
12:07:56 samuelkunkel[m] thanks for all the information
12:09:00 sean-k-mooney auniyal: so on tain you can use tox -e functional-py36 or tox -e functional-py37
12:10:37 sean-k-mooney auniyal: i would either use ubuntu 18.04/ubuntu-bionic or centos 8 stream to run the tests
12:11:14 sean-k-mooney we use 18.04 in teh ci https://github.com/openstack/nova/blob/stable/train/.zuul.yaml#L72-L119
12:11:19 auniyal got same error, I think its trying to need some package/module
12:11:20 auniyal https://paste.opendev.org/show/bsE4F25vNPl8BaGh7I6I/
12:11:45 sean-k-mooney you are trying to use 3.8
12:12:24 auniyal oh in here - /usr/lib/python3.8/runpy.py
12:12:34 sean-k-mooney do you have 3.6 avaiable
12:12:53 auniyal no right now 3.6
12:12:57 auniyal 3.8
12:13:13 sean-k-mooney ya 3.8 was not released/supported by train
12:13:33 auniyal if I create venv of 3.6 and install test-requirements.txt in it
12:13:36 auniyal will it work
12:13:47 sean-k-mooney so if you want to run these you need to use an operating system that was support hence why i said centos 8 stream or ubuntu 18.04
12:14:14 auniyal ack, will go with ubuntu 18,
12:14:23 auniyal thanks Sean
12:14:42 sean-k-mooney if you host os is too new thing liek sqlight might have issues
12:15:05 sean-k-mooney basically where we have python modules that wrap c libs
12:15:23 sean-k-mooney if your host os lib is too new then the old python bindign might now work
12:15:53 sean-k-mooney so if your currently using say the latest fedora you are likely to have issues with old releases like train
12:16:16 sean-k-mooney i generally use vms or contaienr to work around that if i hit that
12:16:54 auniyal yeah I am using vm , devstack on ubuntu 20
12:17:31 auniyal fo this tests, will go with ubuntu 18
12:17:57 sean-k-mooney ack i used to keep a few vms around for backporting
12:18:31 sean-k-mooney i do that less now just because its rare that i need older then 3.8
12:19:42 auniyal ack
12:19:45 sean-k-mooney i think we added 3.8 in ussuri so train is really the only release that does not supprot 3.8 officall now
12:20:18 sean-k-mooney on it was victoria
12:21:22 auniyal for ussuri also I was dependent on zuul, but as there less conflict so it need less tests
12:23:53 sean-k-mooney frickler: by the way i have been using matrix on and off via the element client pretty seamlessly for irc
12:24:07 sean-k-mooney frickler: i still use weechat as my main irc client
12:24:48 sean-k-mooney but if im not at my work laptop i somethimes use teh eleemnt client form my personal laptop or ipad to chat via teh matrix.org bridge
12:25:44 sean-k-mooney so ya if you keep messanges relitivly short (3-4 lines) it works fine i havent hit the lenght limit personlly
12:26:10 sean-k-mooney of the irc alternivies i have used matrix is really the only one i tollerate
12:27:16 sean-k-mooney if the element desktop clinet ever get the ablity to sign into two matix accounts at once it might even be something i would consider as a replacemnt for weechat
12:27:22 frickler sean-k-mooney: there have also been issues where the bridge disconnects but you do not notice on the matrix side, so my personal suggest is still to not use this, ymmv
12:29:54 sean-k-mooney ya i have not had that issue but i still use irc as my primary interface and matix as what i use when im traveling or not working from my normal location
12:30:12 sean-k-mooney so i porably would not notice if there were tempoiry issues
12:33:36 admin1 i have a vm which is always in a pause state in the hypervisor .. trying to unpause using virsh gives error: Timed out during operation: cannot acquire state change lock (held by monitor=remoteDispatchDomainCreateWithFlags) .. the vm is backed by volume on ceph, but ceph is fine and there are no locks
12:33:45 admin1 what can i do to check/troubleshoot this issue
12:33:56 admin1 i rebooted the hypervisor as well, no luck
12:34:50 sean-k-mooney this might be a lock crated by qemu
12:35:06 sean-k-mooney have you tried stopping the vm and staring it
12:35:14 sean-k-mooney e.g. via a hard reboot
12:35:47 admin1 when i do a vrish destroy, it disappears from virsh list --all
12:36:06 admin1 when i start again (horizon/cli) appears back
12:36:09 admin1 with a paused state
12:36:55 sean-k-mooney ack
12:37:19 sean-k-mooney did you check the qemu instance log for any errors
12:37:35 sean-k-mooney this does not sound like a nova issue by the way
12:37:55 sean-k-mooney this sound like an issue at the qemu/libvirt level and or perhaps the ceph interaction
12:38:19 sean-k-mooney you dont happen to have a kvm error in the instance log do you?
12:38:59 sean-k-mooney we hit an issue with ubutu 22.04 where libvirt incorrectly detected the cpu model
12:39:11 sean-k-mooney it enabled amd cpu flags in the domain on an intel host
12:39:25 sean-k-mooney that left teh vm in a paused state
12:39:59 sean-k-mooney although that would not expaling the lock message but i woudl check the qemu instance log in anycase
12:40:13 admin1 sean-k-mooney thanks . i know what to check for now
12:40:51 admin1 where does qemu/libvirt read the ceph connectioon details like mon addresses ?
12:40:57 admin1 from /etc/ceph/ceph.conf ?
12:41:51 admin1 or is it internally somewhere else
12:43:43 sean-k-mooney we get them form the cinder attachment connection info and then store them in our db and pass it to libvirt
12:43:57 sean-k-mooney so no not from the ceph.conf
12:45:10 sean-k-mooney in recent release of openstack (xena+) we have a nova manage command to refresh the atachment info
12:45:12 sean-k-mooney https://docs.openstack.org/nova/latest/cli/nova-manage.html#volume-attachment-refresh
12:45:29 admin1 this one is not xena yet
12:45:40 admin1 i want to remove 2x mons and use only 1 remaining mon
12:45:46 admin1 how do I update/edit this db ?
12:46:00 sean-k-mooney with great pain and care
12:46:21 sean-k-mooney so we added this command to nova-manage because this is sotred in a json blob in the db
12:46:41 sean-k-mooney while it can be modifed its a pain to do
12:47:15 sean-k-mooney admin1: one option woudl be to grab a xena contaiern or create a xena virtual env and just run nova manage
12:47:58 sean-k-mooney i belive this is implemented such that if you have the new version of nova manage and point it to an old cloud it can work but im not 100% certin of that
12:48:00 admin1 you mean have binaries of xena but connect to existing db to manage/manipulate the entries ?
12:48:08 sean-k-mooney ya
12:48:24 sean-k-mooney so bauzas gibi correct me if im wrong be we have had customer do that right^
12:48:57 sean-k-mooney use the updated contaienr with this command ot repair old dbs when connection infor is out of date
12:49:34 sean-k-mooney admin1: i think we have a downstream backport of this by the way to some release which is why im not 100% sure how we used this downstream with train
12:51:45 sean-k-mooney admin1: ya so we have it backported downstream to train in our 16.2 product
12:52:15 sean-k-mooney and i think we have had custoemr use the 16.2 contaienr to fix this on queens/osp 13
12:52:19 admin1 i am on osa tag 23.1.2

Earlier   Later