view reply Interesting.When I test my models, a 40:1 scale always seems to be the most suitable; if it's higher or lower than that, the model's performance suffers.
Running 14 BananaMindBench Leaderboard 🍌 14 View model rankings for the BananaMindBench text‑completion benchmark