Contributing to the MetaMathQA Benchmark #3521
Replies: 2 comments 8 replies
|
@BenjaminBossan Thanks for setting up this discussion. I’ll continue the VBLoRA MetaMathQA work here. I’ll share the results here once the baseline check is complete. |
|
VBLoRA baseline diagnosis update Completed a full B0 baseline run with the setup discussed earlier:
Diagnostics to isolate the gap:
The base model forward loss matches closely, but the gradient path diverges — likely due to the Unsloth weights or the Transformers version delta. I'm unable to access huggingface.co from this environment to download the original Meta Llama checkpoint for validation. Question: Is the Unsloth baseline acceptable as a reference for tuning experiments (e.g., |
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
A place to discuss community contributions to the MetaMathQA benchmark. Before starting your experiments, read the contribution guideline and discuss your ideas with the maintainers here.
The results can be seen in this Gradio Space. Failed experiments (i.e. ones that didn't lead to an improvement) go to the benchmark graveyard -- check it to see past experiments and upload your failed experiments there.
All reactions