Viewing mzimmerm's Bookmarks

Large Language Models for Domain-Specific Language Generation: How to Train Your Dragon | by Andreas Mülder | Medium

[https://medium.com/@andreasmuelder/large-language-models-for-domain-specific-language-generation-how-to-train-your-dragon-0b5360e8ed76] - 2024-03-04 09:45:59 - public:mzimmerm

ai, article, code, doc, generate, llm, train - 7 | id:1489780 -

training a model like Llama with 2.7 billion parameters outperformed a larger model like Vicuna with 13 billion parameters. Especially when considering resource consumption, this might be a good alternative to using a 7B Foundation model instead of a full-blown ChatGPT. The best price-to-performance base model for our use case turned out to be Mistral 7b. The model is compact enough to fit into an affordable GPU with 24GB VRAM and outperforms the other models with 7B parameters.

With marked bookmarks

Mark all

| (+) | |

Viewing 1 - 1, 1 links out of 1 links, page: 1

Follow Tags

article - Please Log In To follow this tag

yabs.io

Yet Another Bookmarks Service

Viewing mzimmerm's Bookmarks

Large Language Models for Domain-Specific Language Generation: How to Train Your Dragon | by Andreas Mülder | Medium

Viewing 1 - 1, 1 links out of 1 links, page: 1

Follow Tags

Export: