Skip to content

docs: fix Reward Model links in RLHF guides - #9910

Merged
tastelikefeet merged 1 commit into
modelscope:mainfrom
tutao0123:agent/fix-reward-model-doc-links
Aug 24, 2026
Merged

docs: fix Reward Model links in RLHF guides#9910
tastelikefeet merged 1 commit into
modelscope:mainfrom
tutao0123:agent/fix-reward-model-doc-links

Conversation

@tutao0123

Copy link
Copy Markdown
Contributor

PR type

  • Bug Fix
  • New Feature
  • Document Updates
  • More Models or Datasets Support

PR information

The Reward Model examples were reorganized from the removed
examples/train/rlhf/rm.sh script into the
examples/train/rlhf/rm/ directory. PR #9831 updated the corresponding
README links, but the Chinese and English RLHF guides still referenced the
old path.

This updates the four remaining links to the current example directory. No
runtime behavior changes.

Validation

  • Confirmed examples/train/rlhf/rm/ contains train.sh and infer.py.
  • Confirmed no examples/train/rlhf/rm.sh references remain under
    docs/source or docs/source_en.
  • pre-commit run --files docs/source/Instruction/Pre-training-and-Fine-tuning.md docs/source/Instruction/RLHF.md docs/source_en/Instruction/Pre-training-and-Fine-tuning.md docs/source_en/Instruction/RLHF.md
  • git diff --check

Experiment results

Not applicable; this is a documentation-only correction.

@tutao0123
tutao0123 marked this pull request as ready for review August 17, 2026 02:58
@tastelikefeet
tastelikefeet merged commit e86583c into modelscope:main Aug 24, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants