I Spent Months Chasing MTEB Scores. Then I Built Something That Actually Works.
I spent months chasing MTEB scores, only to find my embedding models flopped on my own data. Generic benchmarks are misleading β real performance depends on your specific retrieval pipeline. That’s why I built Embench: a free playground to compare embeddings on your own data and taxonomy. Stop relying on leaderboards. Test on what matters.