| Posted | Nick | Remark | |
|---|---|---|---|
| #openstack-nova - 2017-09-26 | |||
| 16:25:21 | bauzas | I mean, super heavy | |
| 16:26:03 | edleafe | bauzas: since this will be sent over RPC, we needed agreement on it so that we don't find ourselves changing it later | |
| 16:26:04 | bauzas | I'd be up concentrating our minds on how we plan to pass that object | |
| 16:26:23 | bauzas | edleafe: we did a couple of RPC changes that didn't require a spec fortunately | |
| 16:26:34 | bauzas | but I leave the mic to mriedem | |
| 16:26:45 | edleafe | bauzas: the idea is to get it close to correct before we make the change | |
| 16:26:46 | dansmith | bauzas: specs are cheap | |
| 16:27:06 | dansmith | if edleafe wants separate specs, I don't think there's a problem | |
| 16:27:11 | edleafe | bauzas: and given the amount of discussion on the Selection object spec, I'd say it was a good thing to do | |
| 16:27:16 | dansmith | we should focus on getting the work done and not the process | |
| 16:27:16 | bauzas | dansmith: well, I'd rather then look at code, but okay :) | |
| 16:27:22 | bauzas | yeah that | |
| 16:27:37 | edleafe | bauzas: the work is being done in parallel | |
| 16:28:00 | bauzas | edleafe: well, okay | |
| 16:28:37 | bauzas | edleafe: the thing is, if you want a spec, fine with me, but then precise the scope | |
| 16:28:48 | bauzas | since it was a work item, I was expecting more | |
| 16:29:14 | dansmith | mriedem: so I was just looking at this for evac and live migration.. this method _moves_ allocations to the destination, not copies AFAICT: https://github.com/openstack/nova/blob/master/nova/scheduler/utils.py#L222-L224 | |
| 16:29:16 | dansmith | mriedem: is that right? | |
| 16:29:18 | bauzas | if the spec isn't targeting to mention how reschedules would be done, fair enough but just make sure you clearly scope that | |
| 16:29:23 | cdent | dansmith: dobne | |
| 16:29:25 | cdent | done! | |
| 16:30:29 | dansmith | cdent: yes the last one is less done | |
| 16:32:40 | mriedem | dansmith: copies | |
| 16:33:10 | dansmith | mriedem: oh does claim_resources() do the doubling thing? | |
| 16:33:17 | mriedem | no | |
| 16:33:30 | mriedem | it takes the allocations for the instance on the source node, and makes those same allocations for the instance on the dest node | |
| 16:33:40 | dansmith | which will erase the source allocation | |
| 16:33:41 | mriedem | it's basically what the scheduler would do, | |
| 16:33:44 | dansmith | because... only one consumer | |
| 16:33:55 | mriedem | oh it calls claim_resources, | |
| 16:33:57 | mriedem | so yeah it doubles | |
| 16:34:02 | mriedem | this is the thing where force=True | |
| 16:34:07 | mriedem | so we don't call the scheduler to double the allocs | |
| 16:34:20 | mriedem | and i said i wanted to move back into the scheduler, but we'd need a skip_filters flag in select_destinations | |
| 16:34:27 | dansmith | okay I didn't think claim_resources was the doubling one, but maybe so, I'll dig a bit | |
| 16:34:39 | mriedem | claim_resources calls the double stuff method | |
| 16:34:59 | dansmith | okay | |
| 16:35:48 | dansmith | cdent: can you look at my comment on the DRY thing and see if you buy what I'm sellin' ? | |
| 16:38:24 | cdent | dansmith: I will buy that with an entire whole dollar, if you comment the plan | |
| 16:38:50 | dansmith | ack | |
| 16:40:23 | dansmith | cdent: you saw the "when we have an atomic operation we should remove this" right? | |
| 16:42:13 | cdent | yes, but (unless I missed it) there’s no “this dupe with that other thing but we don’t care cuz” | |
| 16:42:27 | dansmith | I will add more words | |
| 16:42:31 | cdent | I’ll still buy it for a dollar even if you don’t | |
| 16:45:16 | openstackgerrit | Dan Smith proposed openstack/nova master: Make allocation cleanup honor new by-migration rules https://review.openstack.org/498948 | |
| 16:45:16 | openstackgerrit | Dan Smith proposed openstack/nova master: Pre-create migration object https://review.openstack.org/498950 | |
| 16:45:17 | openstackgerrit | Dan Smith proposed openstack/nova master: Revert allocations by migration uuid https://review.openstack.org/498949 | |
| 16:45:17 | openstackgerrit | Dan Smith proposed openstack/nova master: Refactor resource tracker to account for migration allocations https://review.openstack.org/506419 | |
| 16:45:18 | openstackgerrit | Dan Smith proposed openstack/nova master: Make migration uuid hold allocations for migrating instances https://review.openstack.org/506420 | |
| 16:51:53 | openstackgerrit | Sean Dague proposed openstack/nova master: Move ploop commands to privsep. https://review.openstack.org/492325 | |
| 16:54:07 | openstackgerrit | Sean Dague proposed openstack/nova master: Move ploop commands to privsep. https://review.openstack.org/492325 | |
| 16:58:06 | mriedem | notifications meeting in openstack-meeting-4 in 2 minutes | |
| 17:00:18 | gibi | ... and now it is started | |
| 17:27:00 | openstackgerrit | Moshe Levi proposed openstack/nova master: Don't overwrite binding-profile https://review.openstack.org/505613 | |
| 17:32:29 | openstackgerrit | Eric Berglund proposed openstack/nova master: PowerVM Driver: config drive https://review.openstack.org/409404 | |
| 17:38:24 | mriedem | dansmith: aha, i think i'm hitting issues in devstack where placement isn't getting cleaned up for instances that get 'local' deleted in the api | |
| 17:38:45 | mriedem | not totally sure yet, but failing to burst 500 new instances, hitting NoValidHost | |
| 17:38:53 | mriedem | and i assume it's placement b/c it's not the scheduler filters | |
| 17:39:17 | dansmith | mriedem: and why do you have locally-deleted instances for this test? | |
| 17:39:17 | melwitt | for local deletes, allocations aren't cleaned up till the compute host heals it | |
| 17:39:24 | dansmith | right, what melwitt said | |
| 17:39:27 | mriedem | mysql> select count(id) from consumers; | |
| 17:39:27 | mriedem | | count(id) | | |
| 17:39:27 | mriedem | | 2002 | | |
| 17:39:27 | mriedem | +-----------+ | |
| 17:39:28 | mriedem | 1 row in set (0.01 sec) | |
| 17:40:05 | mriedem | melwitt: there is no compute for these | |
| 17:40:07 | mriedem | they failed during scheduling | |
| 17:40:22 | mriedem | although yeah why would placement have allocations for these... | |
| 17:40:23 | mriedem | wtf | |
| 17:40:37 | mriedem | stack@devstack:~$ nova list | grep -c ERROR | |
| 17:40:37 | mriedem | 1000 | |
| 17:40:39 | melwitt | oh, hm | |
| 17:40:48 | mriedem | so i've got 1000 instances in ERROR state, and 2002 consumers in the api db | |
| 17:41:20 | melwitt | allocations are written at claim time? | |
| 17:41:27 | mriedem | from the scheduler yeah | |
| 17:42:10 | melwitt | so that would explain the ones you do have. but I guess your point is why are there more allocation consumers than non error instances | |
| 17:42:49 | mriedem | that's because i've deleted 1000 over time | |
| 17:43:17 | mriedem | i was hitting messaging timeouts between conductor and the scheduler earlier today, so had 500 in error which i needed to be active, so deleted all of those, restarted conductor and scheduler, and was able to create a single instance | |
| 17:43:22 | mriedem | so tried with 500 more again | |
| 17:43:26 | mriedem | and hit novalidhost on all of those | |
| 17:44:18 | openstackgerrit | Merged openstack/nova master: cleanup test-requirements https://review.openstack.org/507063 | |
| 17:55:32 | openstackgerrit | Dan Smith proposed openstack/nova master: Make live migration hold resources with a migration allocation https://review.openstack.org/507638 | |
| 17:55:45 | dansmith | jaypipes: cdent: ^ quick stab at the live migrate version of this | |
| 17:56:02 | dansmith | it's probably rough at this point, but worth a look I think | |
| 18:32:22 | cdent | dansmith: haven’t had a chance to give it a proper look, but saw a weird when skimming the live migrate thing | |
| 18:32:48 | dansmith | lol | |
| 18:33:04 | dansmith | it's returning True-ish which is what I wanted for the functional tests | |
| 18:33:09 | dansmith | so.. working as designed? :) | |
| 18:33:26 | dansmith | s/returning/being/ | |
| 18:34:36 | openstackgerrit | Merged openstack/nova master: Set the Pike release version for scheduler RPC https://review.openstack.org/507245 | |
| 18:34:51 | cdent | go python! | |
| 18:38:00 | mriedem | wtf, so i can't create multiple instances, i get novalidhost, but i can create one at a time | |
| 18:38:21 | melwitt | are you using multi-create? | |
| 18:38:27 | mriedem | yeah | |
| 18:38:30 | mriedem | wasn't a problem yesterday | |
| 18:38:37 | melwitt | oh | |
| 18:38:41 | mriedem | but i had a bit of a cleaner env yesterday | |
| 18:39:46 | melwitt | multi-create will reject you if any one of min_count can't be accommodated. so one at a time would work if you're in that situation, if some/most of them fit | |
| 18:40:22 | mriedem | yesterday i created 100, then like 200, then 500 more or something | |