It’s been months upon months since a major announcement like this, but we’ve finally done it: new model releases. Introducing our new models: Pygmalion-2 in 7B, and 13B sizes.
Where We’ve Been Link to this heading The burning question on many peoples’ minds is likely “where have we been?” Why haven’...
I might be a bit late to the party, but for those of you that like ERP and fiction writing:
The people from Pygmalion have released a new model, usable for roleplaying, conversation and storywriting.
It is based on Llama 2 and has been trained on SFW and NSFW roleplay, fictional stories and instruction following conversations. It is available in two sizes, 7b and 13b parameters. They're also releasing a mix with MythoMax-L2 called Mythalion 13B.
Furthermore they're (once again) announcing a website with character sharing and inference (later in october.)
For reference: Pygmalion-6b has been a well known dialogue model for (lewd) roleplay in the times before LLaMA. It had been followed up with an underwhelming successor based on LLaMA (Pygmalion-7b). In their new blogpost they promise to have improved with their new model.
(Personally, I'm curious how it performs compared to MythoMax. There aren't many models around, that excel at roleplay or have been designed specifically for that use case.)
Sorry. I was under the impression that everyone interested in new models has TheBloke's HuggingFace profile on speed-dial. I should have linked them ;-)
Very cool that they have a mix with MythoMax right out of the gate. It'll be interesting to see the differences between MythoMax/Pygmalion-2/Mythalion as everyone kicks the tires.
I did some quick testing yesterday and my initial impressions were that Mythalion and Pyg2 (13B q5_K_M versions btw) were a bit more eloquent and verbose in some situations, but they would often take this too far and start writing novels instead of a dialogue. It also felt like they were more prone to take a sentence and repeat it verbatim as part of all their turns. It's possible that these issues could be toned down by adjusting generation parameters, but MythoMax has been very easy to get good results out of.
It's interesting that you can specify which "mode" pyg2 should operate in as part of system prompts but I didn't test how much difference it actually makes on generation. I told it to be in "instruction following mode" and it seemed good enough at general tasks as well.
If I understand pyg2's model card you're supposed to prefix all turns with <|user|> or <|model|> which I didn't manage to get text-generation-webui to do in chat-instruct mode, so I just used the notepad tab instead.
text-generation-webui "chat" and "chat-instruct" modes are... weird and badly documented when it comes to using a specific prompt template. If you don't want to use the notepad mode, use "instruct" mode and set your turn template with the required tags and include your system prompt in the context (? I forget what it is labeled as) box.
EDIT: Actually I think text-generation-webui might use <|user|> as a special string to mean "substitute the user prefix set in the box directly above the turn template box". Why they have to have a turn template field with "macro" functionality and then separate fields for user and bot prefixes when you could just... put the prefix directly in the turn template I have no idea. It's not as though you would ever want or need to change one without the other anyway. But it's possible that as a result of this you can't actually use <|user|> itself in the turn template...