| Posted | Nick | Remark | |
|---|---|---|---|
| #openstack-nova - 2018-07-26 | |||
| 18:55:39 | dansmith | but when I get asked "can we do any of the things that NRP is supposed to enable for us?" I still have to answer no | |
| 18:56:02 | melwitt | mriedem: just trying to get an idea of what's currently going on there ... and taking notes | |
| 18:56:12 | efried | dansmith: Yes, I will agree with that, stipulating s/we/nova/ (as opposed to some other placement consumer). | |
| 18:56:32 | dansmith | efried: yep, been trying to tag my comments with nova when I say stuff like that | |
| 18:56:44 | dansmith | that's my intent at least | |
| 18:57:14 | openstackgerrit | Matt Riedemann proposed openstack/nova master: Fix accumulated nits from port binding for live migration series https://review.openstack.org/583994 | |
| 18:57:16 | mriedem | efried: ^ | |
| 18:57:42 | efried | ack | |
| 18:57:54 | mriedem | dansmith: i got our lazy load thing right here too https://review.openstack.org/#/c/583994/2/nova/compute/manager.py | |
| 18:59:00 | dansmith | mriedem: ack | |
| 18:59:15 | mriedem | i intentionally said that jersey-style | |
| 19:01:45 | mriedem | unrelated, but it seems that stephen's bottom change 564440 is perpetually stuck in the check queue or something weird | |
| 19:01:51 | mriedem | it's in both queues at the same time at least twice today | |
| 19:03:42 | mriedem | and it's already failed in the gate again, | |
| 19:03:47 | mriedem | and queued in check | |
| 19:05:33 | dansmith | cripes | |
| 19:05:48 | dansmith | and some guys keep naking the later patches too | |
| 19:05:49 | dansmith | GAWD | |
| 19:06:26 | mriedem | i nak while nak'ed | |
| 19:06:35 | dansmith | naked nak? | |
| 19:06:42 | mriedem | too far | |
| 19:10:13 | melwitt | mriedem: so the shared storage counting local usage for volume-backed instances saga is finally over? | |
| 19:11:28 | mriedem | not shared storage | |
| 19:11:31 | mriedem | volume-backed root_gb | |
| 19:12:05 | mriedem | https://review.openstack.org/#/q/(status:open+OR+status:merged)+project:openstack/nova+branch:master+topic:bug/1469179 | |
| 19:12:06 | melwitt | oh, I think I got that confused with shared storage reporting. that's still not done | |
| 19:12:26 | melwitt | ok | |
| 19:12:38 | mriedem | we will no longer claim root_gb for volume-backed instances against DISK_GB inventory in placement for new instances | |
| 19:12:41 | mriedem | and heal on moves | |
| 19:13:33 | dansmith | shared ephemeral reporting is also fixed in rocky I think | |
| 19:13:42 | mriedem | via placement? | |
| 19:13:48 | dansmith | we no longer report DISK_GB in inventory if MISC_SHARES_VIA_AGGREGATE is set | |
| 19:14:16 | mriedem | for the libvirt driver* | |
| 19:14:46 | mriedem | i wouldn't talk much about that in rocky though, because we have several places in the code that don't handle that | |
| 19:14:50 | dansmith | not sure, I would have expected that to be applied to each as they converted to update_provider_tree(), unless the others are still using get_inventory() | |
| 19:15:08 | openstackgerrit | Dan Smith proposed openstack/nova master: Assorted cleanups from numa-aware-vswitches series https://review.openstack.org/582651 | |
| 19:15:09 | openstackgerrit | Dan Smith proposed openstack/nova master: Add additional functional tests for NUMA networks https://review.openstack.org/585385 | |
| 19:15:13 | mriedem | yeah that's only libvirt and ironic that implement upt | |
| 19:15:26 | mriedem | and it doesn't apply to ironic | |
| 19:15:58 | mriedem | but like, we don't have any docs on it, we don't have any CI on it, and we have lots of places that aren't aware of it (like move operations) | |
| 19:15:58 | dansmith | I didn't follow that change in, | |
| 19:16:11 | dansmith | but I was hoping that if we're removing that inventory that we handle it the other places we need, | |
| 19:16:24 | dansmith | otherwise we're not going to be able to boot anything on computes that are configured that way :) | |
| 19:16:40 | mriedem | https://review.openstack.org/#/c/560459/ | |
| 19:17:11 | mriedem | well, we should, | |
| 19:17:20 | mriedem | the allocations could go against the provider that has the DISK_GB inventory | |
| 19:17:22 | dansmith | I'm not sure that move ops need to change? | |
| 19:17:31 | mriedem | which is in the MISC_SHARES_VIA_AGGREGATE relationship | |
| 19:17:34 | dansmith | right | |
| 19:17:48 | dansmith | so I'm not sure what you're saying... | |
| 19:17:49 | mriedem | there was something i look up every time this comes up that i know is broken | |
| 19:18:01 | mriedem | i've beaten cdent over the head with a few times already :) | |
| 19:18:32 | dansmith | have we updated the ceph job to do the right thing? | |
| 19:18:41 | dansmith | that'd be a good way to either show it's broken or show that it works | |
| 19:18:53 | dansmith | and you should be able to osc-placement one-shot yourself a sharing provider during setup | |
| 19:19:00 | mriedem | nope, and that's what i suggested in one of the placement update ML threads | |
| 19:19:05 | dansmith | okay | |
| 19:19:17 | dansmith | not sure why we merged that patch without that testing then | |
| 19:19:34 | dansmith | ^ that was rhetorical :D | |
| 19:20:51 | dansmith | mriedem: you're still trying to remember what the missing thing is right? | |
| 19:21:21 | mriedem | well, this is one https://github.com/openstack/nova/blob/master/nova/conductor/tasks/migrate.py#L57 | |
| 19:21:48 | dansmith | ooh right | |
| 19:22:03 | mriedem | and _revert_allocation in compute | |
| 19:22:33 | mriedem | resize to same host probably doesn't handle it either, not sure | |
| 19:22:34 | dansmith | yeah, I remember now.. I was thinking those were only problems we had with the doubling approach, but I remember now | |
| 19:22:39 | mriedem | anywho, | |
| 19:22:52 | mriedem | CI testing with ceph + modeling shared is a todo for the ptg etherpad most likely | |
| 19:22:56 | mriedem | i'll add it | |
| 19:23:10 | dansmith | yeah, that probably needs to be a priority for someone, | |
| 19:23:30 | dansmith | because right now if you associate a sharing provider, things are going to break in unhelpful ways | |
| 19:23:33 | mriedem | i'll assign it to my cat | |
| 19:23:55 | dansmith | melwitt: ^ | |
| 19:24:08 | mriedem | yeah - that's why i've tried to temper enthusiasm / communication that "now it's fixed" | |
| 19:24:17 | mriedem | also because i'm a debby downer | |
| 19:24:58 | dansmith | fair enough, I just saw it had merged when I was triaging a bug related to that stuff | |
| 19:25:05 | dansmith | and figured it had actually been finished | |
| 19:25:29 | melwitt | I've been summoned for ceph stuff eh? | |
| 19:25:39 | dansmith | melwitt: no, for ptl stuff | |
| 19:26:11 | melwitt | oh, thought you were referring to the ceph CI testing add | |
| 19:26:17 | dansmith | I am, | |
| 19:26:41 | dansmith | I'm saying we probably need to be making sure that the test gets updated to validate whether this created a worse hole, | |
| 19:26:50 | dansmith | and maybe decide if we want to either patch that out, | |
| 19:26:54 | dansmith | or reno a known issue or whatever | |
| 19:26:55 | purplerbot | <mriedem> i've beaten cdent over the head with a few times already :) [2018-07-26 19:18:01.246831] [n 1fky] | |
| 19:26:55 | cdent | mriedem: on [t 1fky] | |
| 19:27:07 | melwitt | dansmith: I see, okay | |
| 19:27:35 | cdent | I'm trying to make sure that people are aware, but I'm not sure anybody really reads the weekly pupdates and I'm only willing to take full responsibility for the placement side (where it's ready) | |
| 19:28:03 | cdent | but I agree we have some holes to fill before much more time passes | |
| 19:28:54 | mriedem | L65 https://etherpad.openstack.org/p/nova-ptg-stein | |
| 19:29:21 | mriedem | i will check to see if we have an existing ceph multinode job, but i know the nova-live-migration job is multinode and re-configures itself to use ceph in the 2nd part | |
| 19:29:36 | mriedem | but live migration is less interesting since we don't do claims in the compute or anything | |
| 19:30:19 | dansmith | we really just need a multinode ceph job that boots an instance for base level verification, | |
| 19:30:26 | dansmith | but yeah it should do a cold migration at least | |
| 19:30:40 | dansmith | can we just convert the ceph job to multinode? | |
| 19:31:31 | mriedem | we can do anything we want | |
| 19:31:48 | cdent | anything? | |
| 19:31:49 | mriedem | but i'd rather re-use something than add yet another job | |
| 19:32:04 | mriedem | cdent: with zuul all things are possible (tm) | |
| 19:32:05 | dansmith | mriedem: well just converting the existing one to multinode I mean | |
| 19:32:12 | mriedem | dansmith: i know | |