atulkum / atulkum/pointer_summarizer
Training saturates early?
Open
- Dominant language
- Python
- Stars
- 909
- Forks
- 235
- PR merge metrics
- No merged PRs in 30d
Description
I'm using the same hypers but seeing this for my training curve. Why would this happen? Looks like the LR is too high but your curve with the same lr seems fine.
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.