build(deps): Update trl requirement from <0.18,>=0.13 to >=0.13,<0.20 #569
Add this suggestion to a batch that can be applied as a single commit.
This suggestion is invalid because no changes were made to the code.
Suggestions cannot be applied while the pull request is closed.
Suggestions cannot be applied while viewing a subset of changes.
Only one suggestion per line can be applied in a batch.
Add this suggestion to a batch that can be applied as a single commit.
Applying suggestions on deleted lines is not supported.
You must change the existing code in this line in order to create a valid suggestion.
Outdated suggestions cannot be applied.
This suggestion has been applied or marked resolved.
Suggestions cannot be applied from pending reviews.
Suggestions cannot be applied on multi-line comments.
Suggestions cannot be applied while the pull request is queued to merge.
Suggestion cannot be applied right now. Please check back later.
Updates the requirements on trl to permit the latest version.
Release notes
Sourced from trl's releases.
... (truncated)
Commits
5b3ea9d
Release: v0.19 (#3625)c262674
🧰 [SFT] Tool support (#3597)5c3dd3a
🔍 Add test to verify chat template consistency (#3624)4c92de0
⚔️ Fix bf16 fp16 config conflict issue (#3598)67f17f7
📜 Addchat_template_path
parameter toSFTConfig
(#3599)37a71e8
🧬 Addgeneration_kwargs
as a property ofGRPOConfig
to support additional...b0958c6
[GRPO] Fix prompt truncation (max_prompt_length
) with vLLM. (#3601)8bad863
⭐ Addvllm_gpu_memory_utilization
recommendation script (#3554)d004415
🎁 Put the reward computation in a separate function (#3620)9554c2f
🤵♂️ SFT on assistant messages only (#3586)Dependabot will resolve any conflicts with this PR as long as you don't alter it yourself. You can also trigger a rebase manually by commenting
@dependabot rebase
.Dependabot commands and options
You can trigger Dependabot actions by commenting on this PR:
@dependabot rebase
will rebase this PR@dependabot recreate
will recreate this PR, overwriting any edits that have been made to it@dependabot merge
will merge this PR after your CI passes on it@dependabot squash and merge
will squash and merge this PR after your CI passes on it@dependabot cancel merge
will cancel a previously requested merge and block automerging@dependabot reopen
will reopen this PR if it is closed@dependabot close
will close this PR and stop Dependabot recreating it. You can achieve the same result by closing it manually@dependabot show <dependency name> ignore conditions
will show all of the ignore conditions of the specified dependency@dependabot ignore this major version
will close this PR and stop Dependabot creating any more for this major version (unless you reopen the PR or upgrade to it yourself)@dependabot ignore this minor version
will close this PR and stop Dependabot creating any more for this minor version (unless you reopen the PR or upgrade to it yourself)@dependabot ignore this dependency
will close this PR and stop Dependabot creating any more for this dependency (unless you reopen the PR or upgrade to it yourself)