benchmark source-backed
open-jev-laya-bench: fine-tuning vs prompting a 4B model
A Hugging Face dataset with a RESULTS.md comparing a fine-tuned 4B model against prompting the same model on typed decisions, run against Laya.
Notes
Results table in the dataset repo.
source-backed — Public repository, docs or live artifact. About this label