jevland

benchmark source-backed

open-jev-laya-bench: fine-tuning vs prompting a 4B model

A Hugging Face dataset with a RESULTS.md comparing a fine-tuned 4B model against prompting the same model on typed decisions, run against Laya.

Results table in the dataset repo.

source-backed — Public repository, docs or live artifact. About this label

← Back to the directory