spow12's picture
Update README.md
901f6c0 verified
|
raw
history blame
No virus
3.61 kB
metadata
language:
  - en
  - ja
license: cc-by-nc-4.0
library_name: transformers
tags:
  - nsfw
  - Visual novel
  - roleplay
  - mergekit
  - merge
base_model:
  - mistralai/Mistral-Small-Instruct-2409
pipeline_tag: text-generation

Model Card for Model ID

image

Merged model using mergekit

This model aimed to act like visual novel character.

Merge Format

models:
  - model: mistralai/Mistral-Small-Instruct-2409_SFT
    layer_range: [0, 56]
  - model: mistralai/Mistral-Small-Instruct-2409
    layer_range: [0, 56]
merge_method: slerp
base_model: mistralai/Mistral-Small-Instruct-2409_SFT
parameters:
  t:
    - filter: self_attn
      value: [0, 0.5, 0.3, 0.7, 1]
    - filter: mlp
      value: [1, 0.5, 0.7, 0.3, 0]
    - value: 0.5 # fallback for rest of tensors
dtype: bfloat16

WaifuModel Collections

Unified demo

WaifuAssistant

Update 2.0

  • 2024.09.23 Update 22B, Ver 2.0

Model Details

Model Description

Currently, chatbot has below personality.

character visual_novel
ムラサメ Senren*Banka
茉子 Senren*Banka
芳乃 Senren*Banka
レナ Senren*Banka
千咲 Senren*Banka
芦花 Senren*Banka
愛衣 Café Stella and the Reaper's Butterflies
栞那 Café Stella and the Reaper's Butterflies
ナツメ Café Stella and the Reaper's Butterflies
Café Stella and the Reaper's Butterflies
涼音 Café Stella and the Reaper's Butterflies
あやせ Riddle Joker
七海 Riddle Joker
羽月 Riddle Joker
茉優 Riddle Joker
小春 Riddle Joker

But you can chat your own Character with persona text.

Feel free to test.

Your feedback will be helpful for improving model.

Feature

  • Fluent Chat performance
  • Reduce repetition problem when generate with many turn(over 20~30)
  • Zero Shot character persona using description of character.
  • 128k context window
  • Memory ability that does not forget even after long-context generation

Demo

You can use Demo in google colab.

Check Here

Bias, Risks, and Limitations

This model trained by japanese dataset included visual novel which contain nsfw content.

So, The model may generate NSFW content.

Use & Credit

This model is currently available for non-commercial & Research purpose only. Also, since I'm not detailed in licensing, I hope you use it responsibly.

By sharing this model, I hope to contribute to the research efforts of our community (the open-source community and anime persons).

This repository can use Visual novel-based RAG, but i will not distribute it yet because i'm not sure if it is permissible to release the data publicly.

Citation

@misc {ChatWaifu_22B_v2.0
    author       = { YoungWoo Nam },
    title        = { ChatWaifu_22B_v2.0 },
    year         = 2024,
    url          = { https://huggingface.co/spow12/ChatWaifu_22B_v2.0 },
    publisher    = { Hugging Face }
}