rl: add the elastic resource benchmark contract, gating, evidence and paired work - #66
Merged
Merged
Conversation
…red work Implements the GPU-free half of the rl-elastic-resource-benchmark change: versioned study manifest with canonical hash and calibration/formal modes, config/edge/pool validation gated by runtime capability attestation, evidence index with tamper-rejecting resume, dry-run plan CLI, layered results and report, and paired splits/schedules/scenarios that reuse the legacy benchmark_rl pairing without changing its CLI. Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
The GPU-free half of the
rl-elastic-resource-benchmarkstudy harness:scripts/benchmark_rl_elastic.pyplans and reports studies comparing fixed GPU partitions (trainer / rollout / standby) against scheduled and automatic resizing inside one RL island.What lands
planCLI, layered results and areportthat exits non-zero while the matrix is incompletebenchmark_rlpairing throughyeto/rl/elastic_benchmark/legacy.pyscripts/benchmark_rl.py, its CLI and its result fields are unchanged.No GPU runner ships here. None of the four commands loads a model, imports torch/ray/miles, or creates cloud resources.
Conflicts
None — every file is new (
yeto/rl/elastic_benchmark/, the script, the test module, the doc). Nothing existing is touched.Spec
The planning contract lives in the miles repository under
openspec/changes/rl-elastic-resource-benchmark/, asdocs/RL_ELASTIC_BENCHMARK.mdstates; there is deliberately no OpenSpec change for it in this repo.Tests
19 pass locally (
tests/test_rl_elastic_benchmark.py).🤖 Generated with Claude Code