Embedding models benchmark for code duplication detection
For those who prefer to read the code instead of text, source code is available on GitHub . Relying only on model specification or common benchmarks, we can't predict how it would perform in a specific use case like detecting duplicated code. Focused evaluation revealed, for example, that a general-purpose model can be better than dedicated for code. Or that small model can outperform big providers. submitted by /u/rafal-kochanowski [link] [留言]
本文内容来源于互联网,版权归原作者所有
查看原文