╔════════════════════════════════╗ ║ Cao - MBs ║ ╚════════════════════════════════╝ New : https://scratch.mit.edu/projects/1219333300/ 5/16/26: New UI, New Model, New Benchmark, New of All Cao Model Benchmarks: Please note: all benchmarks were performed with model parameters rendered as similar as possible. As a reminder, these results do not fully reflect the true power of a model, so they should be interpreted with caution. I conducted these tests to compare which model performed best on Q&A benchmarks (based on my dataset, not on questions/answers written by humans).