Comments (3)
Hi, thank you for asking, but I doubt if I can fix your problem since the loss curves should be really task-specific, and I am really not an expert.
Anyways, you can send me an email if you want for more discussions or something like that.
from cherry_llm.
Hi, thank you very much for your interest in this work!
Firstly, I would like to declare that this problem does not come from our selected data but probably comes from the Stanford alpaca codebase. You can find our training losses of different models on our hugging face repo: https://huggingface.co/MingLiiii/cherry-alpaca-5-percent-7B/blob/main/trainer_state.json
Then, for this problem, I think directly downgrading the transformers into 4.28.1 will solve this problem: pip install transformers==4.28.1
and probably you need to re-install wandb:
pip install wandb
You can find similar problems here:
tatsu-lab/stanford_alpaca#298
tloen/alpaca-lora#418
tloen/alpaca-lora#170
Hope it works for you.
Please let me know if you have any other questions~
from cherry_llm.
As a beginner, when I see the loss curve becoming very strange, I feel at a loss.
Deeply thank you for your quick and detailed response. The loss curve is normal now.
(I encountered a very oscillatory loss curve while running code for some other projects. I wonder if it would be convenient for you to provide some debugging suggestions.)
from cherry_llm.
Related Issues (20)
- a confusion about Instruction-Following Difficulty (IFD) scores HOT 2
- a confusion about data_by_IFD HOT 3
- Logic behind IFD score HOT 1
- I plan to apply this method on Llama2, which part of this project needs to be changed to adapt to Llama2? HOT 1
- May I ask if this project is suitable for other large models, such as the Baichuan model, to filter high-quality datasets from other fields HOT 4
- about the paper HOT 1
- Multi-round conversation data set HOT 3
- GPT-4/ChatGPT Evaluation Code HOT 1
- How to filter code SFT data? HOT 2
- Questions related to training HOT 5
- Could the Pre-Experienced Model be used in other different dataset? HOT 1
- Any report of time consuming? HOT 1
- Chinese SFT data cannot be displayed. HOT 3
- 'The training of pre-experienced models is discarded for more efficient usage': that means we can only use base model to do cherry analysis and selection? HOT 1
- batch? HOT 1
- 关于Direct Answer Score sθ(A) HOT 2
- Evaluation reproducibility on benchmarks HOT 4
- how many epochs to train on cherry data? HOT 2
- Question about the effect of labels[0, :start_token] = -100 HOT 1
- why is the process so slow HOT 5
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from cherry_llm.