Skip to content

Conversation

lekurile
Copy link
Contributor

@lekurile lekurile commented Aug 11, 2023

This PR adds an explicit LoRA learning rate argument for DS Chat steps 1 through 3.

  • Step 1:
    • lora_learning_rate
  • Step 2:
    • lora_learning_rate
  • Step 3:
    • actor_lora_learning_rate
    • critic_lora_learning_rate

@lekurile lekurile merged commit f98077d into master Aug 14, 2023
LeetJoe pushed a commit to LeetJoe/DeepSpeedExamples that referenced this pull request Sep 15, 2023
This PR adds an explicit LoRA learning rate argument for DS Chat steps 1 through 3.

Step 1:
- lora_learning_rate

Step 2:
- lora_learning_rate

Step 3:
- actor_lora_learning_rate
- critic_lora_learning_rate
hwchen2017 pushed a commit that referenced this pull request Jun 8, 2025
This PR adds an explicit LoRA learning rate argument for DS Chat steps 1 through 3.

Step 1:
- lora_learning_rate

Step 2:
- lora_learning_rate

Step 3:
- actor_lora_learning_rate
- critic_lora_learning_rate
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Labels
None yet
Projects
None yet
Development

Successfully merging this pull request may close these issues.

2 participants