LocalLLaMA

3 readers

1 users here now

Community to discuss about Llama, the family of large language models created by Meta AI.

founded 1 year ago

MODERATORS

communick@poweruser.forum

We are Higgsfield AI. We have a large GPU cluster and want to finetune your dataset. (alien.top)

submitted 1 year ago by RiskApprehensive9770@alien.top to c/localllama@poweruser.forum

14 comments fedilink hide all child comments

Hey LocalLLaMA. It's Higgsfield AI, and we train huge foundational models.

We have a massive GPU cluster and developed our own infrastructure to manage the cluster and train massive models. We constantly lurked in this subreddit and learned a lot from this passionate community. Right now, we have spare GPUs, and we are excited to give back to this incredible community.

We built a simple web app where you can upload your datasets to finetune it. https://higgsfield.ai/

There's how it works:

You upload the dataset with preconfigured format into HuggingFaсe [1].
Choose your LLM (e.g. LLaMa 70B, Mistral 7B)
Place your submission into the queue
Wait for it to get trained.
Then you get your trained model there on HuggingFace.

[1]: https://github.com/higgsfield-ai/higgsfield/tree/main/tutorials

you are viewing a single comment's thread
view the rest of the comments

[–] herozorro@alien.top 1 points 1 year ago (1 children)

please do something like this, or provide detailed example, on how an open source framework api can be added to a coder LLM.

how do we prepare the data with code sample, docs, so the coder LLM learns it can can do code completions and answer documentation?

[–] RiskApprehensive9770@alien.top 1 points 1 year ago (1 children)

You can train on any dataset as long as it follows our format.

Soon we'll publish a video tutorial.

[–] herozorro@alien.top 1 points 1 year ago

but what would be the proper formatting example for code? just paste in a bunch of files from a repo? or should be more a cheatsheet format?