Skip to content

馃悰 fix RL training bug#665

Description

@flowerthrower

Edge cases occurring after more than $1000$ training steps could previously cause RL training to fail. These failures remained undetected because the tests used a comparatively small step limit: #573 (comment)

Increase the test step limit and fix the corresponding bugs uncovered by the extended training runs. Mostly implemented by PR #573

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

Labels

bugSomething isn't workingmajorPart of a major release

Type

No type

Projects

No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions