LUMIERA.clone

Author	SHA1	Message	Date
Ichthyostega	3674d82bdf	Scheduler-test: next goal -- massively parallel work jobs Search processing pattern for massive parallel test. The goal is to get all cores into active processing most of the time, thus we need a graph with low dependency management overhead, which is also consistently wide horizontally to have several jobs in working state all of the time. The investigation aims at finding out about systematic overheads in such a setup.	2024-01-10 20:34:09 +01:00
Ichthyostega	3fb4baefd5	Scheduler-test: optionally allow to propagate immediately This is just another (obvious) degree of freedom, which could be interesting to explore in stress testing, while probably not of much relevance in practice (if a job is expected to become runable earlier, in can as well be just scheduled earlier). Some experimentation shows that the timing measurements exhibit more fluctuations, but also slightly better times when pressure is low, which is pretty much what I'd expect. When raising pressure, the average times converge towards the same time range as observed with time bound propagation. Note that enabling this variation requires to wire a boolean switch over various layers of abstraction; arguably this is an unnecessary complexity and could be retracted once the »experimentation phase« is over. This completes the preparation of a Scheduler Stress-Test setup.	2024-01-09 02:29:35 +01:00
Ichthyostega	0065dabed1	Scheduler-test: investigate volatile in computation-load The `volatile` was used asymmetrically and there was concern that this usage makes the `ComutationalLoad` dependent on concurrency. However, an impact could not be confirmed empirically. Moreover, to simplify this kind of tests, make the `schedDepends` directly configurable in the Stress-Test-Rig.	2024-01-08 03:38:34 +01:00
Ichthyostega	bdc1b089d7	Scheduler-test: binary search working - Result found in typically 6-7 steps; - running 20 instead of 30 samples seems sufficient Breaking point in this example at stress-Factor 0.47 with run-time 39ms	2024-01-04 01:37:43 +01:00
Ichthyostega	e704f4aae0	Scheduler-test: build configurable measurement setup Elaborate the draft to include all the elements used directly in the test case thus far; the goal is to introduce some structuring and leave room for flexible confguration, while implementing the actual binary search as library function over Lambdas. My expectation is to write a series of individual test instances with varying parameters; while it seems possible to add further performance test variations into that scheme later on.	2024-01-03 02:18:15 +01:00
Ichthyostega	6f4bd150fd	Scheduler-test: draft a structure to formalise these investigations - the goal is to run a binary search - the search condition should be factored out - thus some kind of framework or DSL is required, to separate the technicalities of the measurement from the specifics of the actual test case.	2024-01-03 00:45:17 +01:00
Ichthyostega	29699991a0	Scheduler-test: watch statistics with increasing stress - repeated invocations of the same test setup for statistics - the usual nasty 64-node graph with massive fork out - limit concurrency to 4 cores - tabulate data to look for clues regarding a trigger criteria Hypothesis: The Scheduler slips off schedule when all of the following three criteria are met: - more than 55% glitches with Δ > 2ms - σ > 2ms - ∅Δ > 4ms	2024-01-02 18:44:20 +01:00
Ichthyostega	a56afbaf62	Scheduler-test: bugfix in computation-load ...this one was quite silly: obviously we need a separate instance of the memory block ''per invocation'', otherwise concurrent invocations would corrupt each other's allocation. The whole point of this variant of the computation-load is to access a ''private'' memory block...	2024-01-02 16:24:24 +01:00
Ichthyostega	0c9485294e	Scheduler-test: experiment with varying stress levels	2024-01-02 15:45:40 +01:00
Ichthyostega	bb2bbc0e02	Scheduler-test: verify adapted schedule with stress-factor - schedule can now be adapted to concurrency and expected distribution of runtimes - additional stress factor to press the schedule (1.0 is nominal speed) - observed run-time now without Scheduler start-up and pre-roll - document and verify computed numbers	2024-01-02 00:24:28 +01:00
Ichthyostega	813f8721f7	Scheduler-test: build adapted schedule ...based on the adapted time-factor sequence implemented yesterday in TestChainLoad itself - in this case, the TimeBase from the computation load is used as level speed - this »base beat« is then modulated by the timing factor sequence - working in an additional stress factor to press the schedule uniformly - actual start time will be added as offset once the actual test commences	2024-01-01 22:48:27 +01:00
Ichthyostega	f4dd309476	Scheduler-test: change test setup to use a schedule table ...up to now, we've relied on a regular schedule governed solely by the progression of node levels, with a fixed level speed defaulting to 1ms per level. But in preparation of stress tesging, we need a schedule adapted to the expected distribution of computation times, otherwise we'll not be able to factor out the actual computation graph connectivity. The goal is to establish a distinctive breaking point when the scheduler is unable to cope with the provided schedule.	2024-01-01 21:07:16 +01:00
Ichthyostega	55cb028abf	Scheduler-test: document and verify weight adapted timing The helper developed thus far produces a sequence of weight factors per level, which could then be multiplied with an actual delay base time to produce a concrete schedule. These calculations, while simple, are difficult to understand; recommended to use the values tabulated in this test together with a `graphviz` rendering of the node graph (🠲 `printTopologyDOT()`)	2023-12-31 21:59:41 +01:00
Ichthyostega	e9e7d954b1	Scheduler-test: formula to guess expense The intention is to establish a theoretical limit for the expense, given some degree of concurrency. In reality, the expense should always be greater, since the time is not just split by the number of cores; rather we need to chain up existing jobs of various weight on the available cores (which is a special case of the box packing problem). With this formula, an ideal weight factor can be determined for each level, and then summing up the sequence of levels gives us a guess for a sensible timing for the overall scheduler	2023-12-31 03:14:59 +01:00
Ichthyostega	409a60238a	Scheduler-test: extract a generic grouping iterator ...so IterExplorer got yet another processing layer, which uses the grouping mechanics developed yesterday, but is freely configurable through λ-Functions. At actual usage sit in TestChainLoad, now only the actual aggregation computation must be supplied, and follow-up computations can now be chained up easily as further transformation layers.	2023-12-31 00:41:01 +01:00
Ichthyostega	fec117039e	Scheduler-test: need this group aggregation as pipeline rather Yesterday I've written a simple loop-based implementation of a grouping aggregation to count the node weights per level. Unfortunately it turns out we'll use several flavours of this and we'd have to chain up postprocessing -- thus from a usage perspective it would be better to have the same functionality packaged as interator pipeline. This turns out to be surprisingly tricky and there is no suitable library function available, which means I'll have to write one myself. This changeset is the first step into this direction: reformulate the simple for-loop into a demand-driven grouping iterator	2023-12-30 02:15:38 +01:00
Ichthyostega	f04035a030	Scheduler-test: draft calculation of level-weight based schedule ...the idea is to use the sum of node weights per level to create a schedule, which more closely reflects the distribution of actual computation time. Hopefully such a schedule can then be squeezed or stretched by a time factor to find out a ''breaking point'', at which the Scheduler is no longer able to keep up.	2023-12-29 01:07:26 +01:00
Ichthyostega	af680cdfd9	Scheduler-test: adapt tests to changed logic at entrance - now there can not be any direct dispatch anymore when entering events - thus there is no decision logic at entrance anymore - rather the work-function implementation moved down into Layer-2 - so add a unit-test like coverage there (integration in SchedulerService_test)	2023-12-27 00:16:03 +01:00
Ichthyostega	dedfbf4984	Scheduler-test: investigate planning failure - fix mistake in schdule time for planning chunks (must use start, not end of chunk) - allow to configure the heuristics for pre-roll (time reserved for planning a node)	2023-12-23 21:38:51 +01:00
Ichthyostega	90ab20be61	Scheduler-test: press harder with long and massive graph ...observing multiple failures, which seem to be interconnected - the test-setup with the planning chunk pre-roll is insufficient - basically it is not possible to perform further concurrent planning, without getting behind the actual schedule; at least in the setup with DUMP print statements (which slowdown everything) - muliple chained re-entrant calls into the planning function can result - the ASSERTION in the Allocator was triggered again - the log+stacktrace indicate that there is still a Gap in the logic to protect the allocations via Grooming-Token	2023-12-22 00:33:51 +01:00
Ichthyostega	2cd51fa714	Scheduler-test: fix out-of-bound access ...causing the system to freeze due to excess memory allocation. Fortunately it turned out this was not an error in the Scheduler core or memory manager, but rather a sloppiness in the test scaffolding. However, this incident highlights that the memory manager lacks some sanity checks to prevent outright nonsensical allocation requests. Moreover it became clear again that the allocation happens ''already before'' entering the Scheduler — and thus the existing sanity check comes too late. Now I've used the same reasoning also for additional checks in the allocator, limiting the Epoch increment to 3000 and the total memory allocation to 8GiB Talking of Gibitbytes... indeed we could use a shorthand notation for that purpose...	2023-12-21 20:25:43 +01:00
Ichthyostega	c4807abf8a	Scheduler-test: planning for stress-tests	2023-12-19 21:06:23 +01:00
Ichthyostega	9db341bd8b	Scheduler: plan for integration identified three distinct tasks - build the external API - establish component integration - performance testing	2023-10-20 00:59:50 +02:00

23 commits