benchmarks

#50

by rfcoder0 - opened Sep 7, 2025

Sep 7, 2025

i was training this to try to make it comparable with an 8b model like my rfcoder0/qwen3-4b-custom-Sile. i wanted to do the same thing with phi but it seems like the stock benchmarks are not correct arc 10 shot was way lower. Is there any others that have this issue.

xujfcn

about 22 hours ago

For those asking about API access — I've been using Crazyrouter as a unified gateway. One API key, OpenAI SDK compatible. Works well for testing different models without managing multiple accounts.

Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.

Tap or paste here to upload images

· Sign up or log in to comment