Earlier  
Posted Nick Remark
#openstack-nova - 2022-02-17
15:26:17 bauzas sean-k-mooney: nah I was saying about shelved instances
15:26:25 sean-k-mooney yes
15:26:32 sean-k-mooney that fact its shelved does nto change that
15:26:41 bauzas ok, then it was a nit
15:28:27 sean-k-mooney i honestly dont see an impedment to counting server ectra via sharing rps at some point in the futre if we need too or other things like floating ip that are out of scope of the current spec
15:30:25 gibi I agree that offloaded servers if holding floating ips then it should consume quota
15:30:55 gibi and yes we cannot count that with consumer types as no instance allocation exists for offloaded servers in placement
16:01:42 melwitt bauzas, sean-k-mooney: the tl;dr is that unified limits is not meant to have the same behavior as before. shelved offloaded will not consume quota
16:02:08 bauzas melwitt: cool then, I'll switch to +2
16:09:36 melwitt bauzas: I dunno if you've seen this but this is one place where we document the expected differences when counting from placement https://docs.openstack.org/nova/latest/admin/quotas.html#quota-usage-from-placement
16:09:54 melwitt and unified limits builds on top of that
16:10:01 bauzas melwitt: oh thanks, I think I saw it already but forgot
16:10:17 bauzas my fucking brain is so bad...
16:10:18 melwitt and changes a few more things like no more user_id scoped quotas
16:10:41 bauzas and you documented the shelved thing, good
16:10:53 bauzas "Behavior will be different for servers in SHELVED_OFFLOADED state. A server in SHELVED_OFFLOADED state will not have placement allocations, so it will not consume quota usage for cores and ram. Note that because of this, it will be possible for a request to unshelve a server to be rejected if the user does not have enough quota available to support the cores and ram needed by the server to be unshelved."
16:11:04 bauzas perfect, definitely +2 once we're done with the internal meeting
16:11:37 melwitt yeah I aimed to document all differences for operators, it's also in the config option help and was in the release note
16:23:06 bauzas melwitt: my bad, I haven't yet looked at the reno change
16:23:49 opendevreview Felix Huettner proposed openstack/nova stable/ussuri: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829506
16:23:50 melwitt bauzas: sorry I meant the release note for the counting from placement in the past
16:24:02 bauzas ah
16:24:05 bauzas gotchaz
16:24:22 melwitt I was saying I tried to put the info everywhere :)
16:25:09 melwitt reno for the counting from placement is in this list https://docs.openstack.org/releasenotes/nova/train.html#relnotes-20-0-0-stable-train-upgrade-notes
16:35:28 opendevreview Julia Kreger proposed openstack/nova master: WIP Ironic - Handle instance/node host on rebalance https://review.opendev.org/c/openstack/nova/+/813897
16:37:01 opendevreview Julia Kreger proposed openstack/nova master: Ironic - Don't query the API for instance counts https://review.opendev.org/c/openstack/nova/+/829613
16:38:18 opendevreview Jonathan Race proposed openstack/nova master: object/notification for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828369
16:38:18 opendevreview Jonathan Race proposed openstack/nova master: driver/secheduler/docs for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/822053
16:38:19 opendevreview Jonathan Race proposed openstack/nova master: zuul-job for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828372
16:51:21 opendevreview Jonathan Race proposed openstack/nova master: object/notification for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828369
16:51:22 opendevreview Jonathan Race proposed openstack/nova master: driver/secheduler/docs for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/822053
16:51:22 opendevreview Jonathan Race proposed openstack/nova master: zuul-job for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828372
18:07:57 opendevreview Balazs Gibizer proposed openstack/nova master: Record SRIOV PF MAC in the binding profile https://review.opendev.org/c/openstack/nova/+/829248
18:10:47 opendevreview Merged openstack/nova master: Document remote-managed port usage considerations https://review.opendev.org/c/openstack/nova/+/827513
18:43:05 sean-k-mooney gibi: im going to call it a day there ill try and review the rest of your placment patches on monday
19:14:11 chateaulav gibi: are there more verbose logs from grenade, or something that i am missing besides `tempest.lib.exceptions.ServerFault: Got server fault`
19:41:18 opendevreview Jonathan Race proposed openstack/nova master: object/notification for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828369
19:41:19 opendevreview Jonathan Race proposed openstack/nova master: driver/secheduler/docs for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/822053
19:41:19 opendevreview Jonathan Race proposed openstack/nova master: zuul-job for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828372
20:28:05 opendevreview Merged openstack/nova master: Fix to implement 'pack' or 'spread' VM's NUMA cells https://review.opendev.org/c/openstack/nova/+/805649
20:33:50 opendevreview Ilya Popov proposed openstack/nova stable/xena: Fix to implement 'pack' or 'spread' VM's NUMA cells https://review.opendev.org/c/openstack/nova/+/829804
20:35:32 opendevreview Merged openstack/nova master: neutron: Allow to spawn VMs with port without IP address https://review.opendev.org/c/openstack/nova/+/669411
21:18:16 opendevreview Jonathan Race proposed openstack/nova master: object/notification for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828369
21:18:17 opendevreview Jonathan Race proposed openstack/nova master: driver/secheduler/docs for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/822053
21:18:17 opendevreview Jonathan Race proposed openstack/nova master: zuul-job for Adds Pick guest CPU architecture based on host arch in libvirt driver support https://review.opendev.org/c/openstack/nova/+/828372
21:43:34 opendevreview Merged openstack/nova master: [nova/libvirt] Support for checking and enabling SMM when needed https://review.opendev.org/c/openstack/nova/+/825496
#openstack-nova - 2022-02-18
07:51:48 gibi sean-k-mooney: thanks. enjoy PTO today
07:52:51 gibi chateaulav: you have to look into the openstack service logs or 'spread' VM's NUMA cells
07:53:00 gibi sorry
07:53:06 gibi wrong copy paste buffer
07:53:10 gibi https://zuul.opendev.org/t/openstack/build/39dcedc2915b4dce9bacce7d1c21f8fe/logs
07:53:19 gibi so heer in the job result
07:53:51 gibi under controller/logs and comupute1/logs you will find screen-n-cpu.txt with the nova-compute logs
07:56:02 opendevreview Felix Huettner proposed openstack/nova stable/victoria: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829504
07:58:53 brinzhang bauzas, gibi, songwenping: vGPU support in Cyborg may need a slot of PTG, do you have some suggestion?
08:00:06 opendevreview Felix Huettner proposed openstack/nova stable/train: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829507
08:01:10 brinzhang we would like to register 2:00UTC-3:UTC at April 5, is it ok?
08:02:18 opendevreview Ghanshyam proposed openstack/nova master: Separate flavor extra specs policy for server APIs https://review.opendev.org/c/openstack/nova/+/829626
08:04:01 gibi brinzhang: I'm not sure we have a ptg etherpad yet. as far as I know RedHat folks including bauzas is PTO on today.
08:07:08 brinzhang gibi: ack
08:07:22 gibi so let's get back to this on Monday
08:07:25 brinzhang https://etherpad.opendev.org/p/nova-zed-ptg I saw this link, it's nothing else
08:07:52 brinzhang gibi: ok, we can discuss on monday ^^
08:12:03 gibi chateaulav: so for example the logs from your trial with my last suggestion visible here (filtered to error only) https://zuul.opendev.org/t/openstack/build/08789d9f0b0546cb9e6fb2b6f4f0231c/log/compute1/logs/screen-n-cpu.txt?severity=4
08:44:30 opendevreview Felix Huettner proposed openstack/nova stable/stein: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829859
08:45:07 tobias-urdin gibi: friendly request for backport review https://review.opendev.org/c/openstack/nova/+/828407 :)
08:51:40 opendevreview Felix Huettner proposed openstack/nova stable/rocky: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829860
08:55:14 opendevreview Felix Huettner proposed openstack/nova stable/queens: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829861
09:15:33 gibi tobias-urdin: hi! I don't have +2 rights on stable branches :/
09:17:44 tobias-urdin gibi: oh sorry for the noice!
09:17:55 gibi tobias-urdin: no worries
09:29:24 opendevreview Ghanshyam proposed openstack/nova master: Complete phase-1 of RBAC community-wide goal https://review.opendev.org/c/openstack/nova/+/829866
10:51:36 chateaulav gibi: ok,I was trying a few things yesterday, because greens was still failing. I'll put the change back and then investigate that way.
10:51:46 chateaulav Grenade
10:56:29 gibi chateaulav: probably easier to create a unit test that calls obj_make_compatible on a populated ComputeNode object
10:56:43 gibi with that you can troubleshoot locally
10:59:41 chateaulav gibi: Yeah. Biggest thing was trying to have zuul pass which it did. I'll see about doing that, and can test to ensure the new values aren't processed
13:22:34 chateaulav gibi: so then with the current patchset, grenade and everything is happy.
13:23:02 chateaulav it did not like those exceptions or moving the backport for hvspec above super
13:23:52 gibi chateaulav: but it did not like with a different reasons.
13:24:02 chateaulav gotcha
13:24:08 gibi chateaulav: I personally would like to keep rejecting new archs like here https://review.opendev.org/c/openstack/nova/+/828369/14/nova/objects/hv_spec.py
13:24:17 gibi chateaulav: I know that you removed this to please grenade
13:24:27 gibi but I think that just hides the problem
13:25:11 gibi I have not time right now to proposa a patch top of yours showing how to make a unit test to show the same problem in a you local env
13:25:19 chateaulav ok, and thats what im starting to see and understand. getting use to zuul and where to find tthings.
13:25:21 gibi but I think that would be a way forward to troubleshoot
13:25:39 chateaulav makes sense
13:26:09 gibi I try to get to your problem before end of today
13:29:15 chateaulav appreciate it, ill be working all day on it. gonna see about that unit test
14:06:38 opendevreview Felix Huettner proposed openstack/nova stable/train: Gracefull recovery when attaching volume fails https://review.opendev.org/c/openstack/nova/+/829507
14:51:46 jamespage o/ - is there a good reference on how scheduling should behave when hypervisors have partial hugepage memory configuration (say 150GB of 512GB)
14:51:50 jamespage ?
15:05:44 gibi jamespage: I don't think we have. if your instance needs numa topology becasue of cpu pinning or huge pages then the NUMATopologyFilter is responsible to select the proper host
15:11:07 artom jamespage, I don't think the "partial" matters. If the host has enough pages of the correct size to fit the instance, it should pass scheduling
15:13:09 jamespage gibi, artom: interestingly instances with larger page size configuration schedule fine - the problem I'm looking at happens when an instances without large pages gets scheduled to the hypervisor
15:13:52 jamespage deployment is using instances with large amounts of RAM - exceeding the diff between Total RAM - HugePages - Reserved by quite a bit
15:13:56 gibi jamespage: does instances without huge pages are they have any numa related requirements, i.e. cpu pinning?
15:14:59 jamespage gibi: nope no extra specs at all

Earlier   Later