|
You are here |
www.smashcompany.com | ||
| | | | |
www.vellum.ai
|
|
| | | | | Understand the latest benchmarks, their limitations, and how models compare. | |
| | | | |
www.vals.ai
|
|
| | | | | Private, domain-specific benchmarks in legal, tax, and finance. | |
| | | | |
tomasvotruba.com
|
|
| | | | | Last week, I had many interesting discussions about OpenAI and GPT on [Laracon in Porto](https://laracon.eu/). Especially with [Marcel Pociot](https://twitter.com/marcelpociot). I've learned much more in 2 days than on the Internet since December. That feels great, and tips seem basic but effective. But as in any other fresh area, finding out about them takes a lot of work. I want to embrace sharing in the GPT community, so here is cherry-pick list of failures and tricks from people **who were generous to share it with me**. | |
| | | | |
www.confident-ai.com
|
|
| | | In this article, I'm going to go through all the top LLM benchmarks currently used and why they matter. | ||