Earlier  
Posted Nick Remark
#openstack-nova - 2017-09-26
16:25:02 bauzas edleafe: looks uber too much
16:25:21 bauzas I mean, super heavy
16:26:03 edleafe bauzas: since this will be sent over RPC, we needed agreement on it so that we don't find ourselves changing it later
16:26:04 bauzas I'd be up concentrating our minds on how we plan to pass that object
16:26:23 bauzas edleafe: we did a couple of RPC changes that didn't require a spec fortunately
16:26:34 bauzas but I leave the mic to mriedem
16:26:45 edleafe bauzas: the idea is to get it close to correct before we make the change
16:26:46 dansmith bauzas: specs are cheap
16:27:06 dansmith if edleafe wants separate specs, I don't think there's a problem
16:27:11 edleafe bauzas: and given the amount of discussion on the Selection object spec, I'd say it was a good thing to do
16:27:16 dansmith we should focus on getting the work done and not the process
16:27:16 bauzas dansmith: well, I'd rather then look at code, but okay :)
16:27:22 bauzas yeah that
16:27:37 edleafe bauzas: the work is being done in parallel
16:28:00 bauzas edleafe: well, okay
16:28:37 bauzas edleafe: the thing is, if you want a spec, fine with me, but then precise the scope
16:28:48 bauzas since it was a work item, I was expecting more
16:29:14 dansmith mriedem: so I was just looking at this for evac and live migration.. this method _moves_ allocations to the destination, not copies AFAICT: https://github.com/openstack/nova/blob/master/nova/scheduler/utils.py#L222-L224
16:29:16 dansmith mriedem: is that right?
16:29:18 bauzas if the spec isn't targeting to mention how reschedules would be done, fair enough but just make sure you clearly scope that
16:29:23 cdent dansmith: dobne
16:29:25 cdent done!
16:30:29 dansmith cdent: yes the last one is less done
16:32:40 mriedem dansmith: copies
16:33:10 dansmith mriedem: oh does claim_resources() do the doubling thing?
16:33:17 mriedem no
16:33:30 mriedem it takes the allocations for the instance on the source node, and makes those same allocations for the instance on the dest node
16:33:40 dansmith which will erase the source allocation
16:33:41 mriedem it's basically what the scheduler would do,
16:33:44 dansmith because... only one consumer
16:33:55 mriedem oh it calls claim_resources,
16:33:57 mriedem so yeah it doubles
16:34:02 mriedem this is the thing where force=True
16:34:07 mriedem so we don't call the scheduler to double the allocs
16:34:20 mriedem and i said i wanted to move back into the scheduler, but we'd need a skip_filters flag in select_destinations
16:34:27 dansmith okay I didn't think claim_resources was the doubling one, but maybe so, I'll dig a bit
16:34:39 mriedem claim_resources calls the double stuff method
16:34:59 dansmith okay
16:35:48 dansmith cdent: can you look at my comment on the DRY thing and see if you buy what I'm sellin' ?
16:38:24 cdent dansmith: I will buy that with an entire whole dollar, if you comment the plan
16:38:50 dansmith ack
16:40:23 dansmith cdent: you saw the "when we have an atomic operation we should remove this" right?
16:42:13 cdent yes, but (unless I missed it) there’s no “this dupe with that other thing but we don’t care cuz”
16:42:27 dansmith I will add more words
16:42:31 cdent I’ll still buy it for a dollar even if you don’t
16:45:16 openstackgerrit Dan Smith proposed openstack/nova master: Make allocation cleanup honor new by-migration rules https://review.openstack.org/498948
16:45:16 openstackgerrit Dan Smith proposed openstack/nova master: Pre-create migration object https://review.openstack.org/498950
16:45:17 openstackgerrit Dan Smith proposed openstack/nova master: Revert allocations by migration uuid https://review.openstack.org/498949
16:45:17 openstackgerrit Dan Smith proposed openstack/nova master: Refactor resource tracker to account for migration allocations https://review.openstack.org/506419
16:45:18 openstackgerrit Dan Smith proposed openstack/nova master: Make migration uuid hold allocations for migrating instances https://review.openstack.org/506420
16:51:53 openstackgerrit Sean Dague proposed openstack/nova master: Move ploop commands to privsep. https://review.openstack.org/492325
16:54:07 openstackgerrit Sean Dague proposed openstack/nova master: Move ploop commands to privsep. https://review.openstack.org/492325
16:58:06 mriedem notifications meeting in openstack-meeting-4 in 2 minutes
17:00:18 gibi ... and now it is started
17:27:00 openstackgerrit Moshe Levi proposed openstack/nova master: Don't overwrite binding-profile https://review.openstack.org/505613
17:32:29 openstackgerrit Eric Berglund proposed openstack/nova master: PowerVM Driver: config drive https://review.openstack.org/409404
17:38:24 mriedem dansmith: aha, i think i'm hitting issues in devstack where placement isn't getting cleaned up for instances that get 'local' deleted in the api
17:38:45 mriedem not totally sure yet, but failing to burst 500 new instances, hitting NoValidHost
17:38:53 mriedem and i assume it's placement b/c it's not the scheduler filters
17:39:17 dansmith mriedem: and why do you have locally-deleted instances for this test?
17:39:17 melwitt for local deletes, allocations aren't cleaned up till the compute host heals it
17:39:24 dansmith right, what melwitt said
17:39:27 mriedem mysql> select count(id) from consumers;
17:39:27 mriedem | count(id) |
17:39:27 mriedem | 2002 |
17:39:27 mriedem +-----------+
17:39:28 mriedem 1 row in set (0.01 sec)
17:40:05 mriedem melwitt: there is no compute for these
17:40:07 mriedem they failed during scheduling
17:40:22 mriedem although yeah why would placement have allocations for these...
17:40:23 mriedem wtf
17:40:37 mriedem stack@devstack:~$ nova list | grep -c ERROR
17:40:37 mriedem 1000
17:40:39 melwitt oh, hm
17:40:48 mriedem so i've got 1000 instances in ERROR state, and 2002 consumers in the api db
17:41:20 melwitt allocations are written at claim time?
17:41:27 mriedem from the scheduler yeah
17:42:10 melwitt so that would explain the ones you do have. but I guess your point is why are there more allocation consumers than non error instances
17:42:49 mriedem that's because i've deleted 1000 over time
17:43:17 mriedem i was hitting messaging timeouts between conductor and the scheduler earlier today, so had 500 in error which i needed to be active, so deleted all of those, restarted conductor and scheduler, and was able to create a single instance
17:43:22 mriedem so tried with 500 more again
17:43:26 mriedem and hit novalidhost on all of those
17:44:18 openstackgerrit Merged openstack/nova master: cleanup test-requirements https://review.openstack.org/507063
17:55:32 openstackgerrit Dan Smith proposed openstack/nova master: Make live migration hold resources with a migration allocation https://review.openstack.org/507638
17:55:45 dansmith jaypipes: cdent: ^ quick stab at the live migrate version of this
17:56:02 dansmith it's probably rough at this point, but worth a look I think
18:32:22 cdent dansmith: haven’t had a chance to give it a proper look, but saw a weird when skimming the live migrate thing
18:32:48 dansmith lol
18:33:04 dansmith it's returning True-ish which is what I wanted for the functional tests
18:33:09 dansmith so.. working as designed? :)
18:33:26 dansmith s/returning/being/
18:34:36 openstackgerrit Merged openstack/nova master: Set the Pike release version for scheduler RPC https://review.openstack.org/507245
18:34:51 cdent go python!
18:38:00 mriedem wtf, so i can't create multiple instances, i get novalidhost, but i can create one at a time
18:38:21 melwitt are you using multi-create?
18:38:27 mriedem yeah
18:38:30 mriedem wasn't a problem yesterday
18:38:37 melwitt oh
18:38:41 mriedem but i had a bit of a cleaner env yesterday
18:39:46 melwitt multi-create will reject you if any one of min_count can't be accommodated. so one at a time would work if you're in that situation, if some/most of them fit

Earlier   Later