Spyke

Syndicated from the fediverse. Read and engage on the original instance.

View original on sh.itjust.works

OpenOrca, an open-source dataset and series of instruct-tuned language models

I realized that while Microsoft would probably release their LLaMA-13b based model (as of the time of this writing they still haven't) I concluded that they might not release the dataset. Therefore, I resolved to replicate their efforts, download the data myself, and train the model myself, so that OpenOrca can be released on other sizes of LLaMA as well as other foundational models such as Falcon, OpenLLaMA, RedPajama, MPT, RWKV.

OpenOrca, an open-source dataset and series of instruct-tuned language modelshttps://erichartford.com/openorcaOpen linkView original on sh.itjust.works
17

5 replies

lemmy.sdf.org

I hope this is okay: I made a backup of the blog post and saved it to my website/file hosting site. here is the backup.

I'll remove/blank out this comment when/if I see the page come back online.

EDIT: Okay, so it looks like the OpenOrca project on Eric Hartford's website has been rebranded as Dolphin. My understanding is that someone else is working on an OpenOrca, prompting the rebranding.

3

You reached the end

OpenOrca, an open-source dataset and series of instruct-tuned language models | Spyke