en:vpet:nf:chatvpet

This translation is older than the original page and might be outdated. See what has changed.

ChatVPet Project

ChatVPet is a language-model project created specifically for Virtual Pet Simulator. During early testing, we found that general-purpose AI models without targeted training often performed poorly, with slow responses and inconsistent dialogue. To give the desktop pet a real personality, we started the ChatVPet project. Its goal is to collect genuine chat data and build a dedicated AI training dataset from the ground up.

The project is based on the ChatGLM-6B architecture and uses LLaMA-Factory for in-depth fine-tuning. All collected data has been submitted with the users' explicit consent. ChatVPet is released under the GPLv3 license and may be modified, provided that derivative work remains open source and includes attribution. Because of the current dataset size, submissions and training are available only in Simplified Chinese. Support for other languages may be added in the future.

  1. Subscribe to the MOD: First, subscribe to ChatGPT for Workshop Creators.
  2. Enable it in Settings: Enable the MOD in the game settings, then open the MOD Settings page.

1

  1. Start training: Select the chat-training page in the MOD settings.

2

ChatVPet Training and Review Guide

The project has two main sections: Submit Training Content and the Desktop Pet Court.

Submit Training Content

Here, you can use your creativity to write conversations for the desktop pet and set the appropriate pet state.

  • Writing assistance: A basic AI can help generate a reply (charging Token points at the 25% rate), but submitting AI-generated content directly is strictly prohibited.
  • Rewards and penalties:
    • Approved (approval > 60%): The content enters the training dataset. Token points are awarded based on the number of votes and amount of content.
    • Rejected (negative votes > 80%): Token points will be deducted as a penalty.
    • 3

Desktop Pet Court

Here, you can review conversations submitted by other users, earn Token points, and find inspiration for your own writing.

  • Voting rewards: Reviews use a voting system. After voting ends, the closer your choice is to the majority decision, the greater your Token reward.
  • Review rules: Remain fair and impartial. Plagiarizing other submissions is strictly prohibited. Unless a submission contains a serious violation, choose Skip when it simply does not suit your personal taste.

4


Examples

  • Question (complete, clear, and reasonably concise):
    • “Have you eaten yet?”
    • “Do you prefer cats or dogs?”
  • Answer (a definite response with a reason):
    • “Yes, Master! I already ate the Kobe beef you gave me~ I love you the most~”
    • “I like cats, because I'm an adorable catgirl myself~ Meow, meow, meow~~”

Length requirements (Token measurements)

Category Length requirement Approximate Chinese length
Question length > 5 Tokens About 3–6 Chinese characters
Answer length > 15 Tokens About 8–15 Chinese characters
Maximum total length < 1,000 Tokens About 300–600 Chinese characters

Whether submitting or reviewing content, you must follow all 11 rules below:

  1. Stay in character: Content must fit the energetic and adorable young desktop-pet character setting.
  2. No meaningless content: Do not spam or submit meaningless dialogue.
  3. No external links: URLs are strictly prohibited in submitted content.
  4. Keep specialist content accessible: Technical topics are allowed, but they must not be excessively advanced.
  5. Avoid an AI-like tone: Replies must not feel formulaic, mechanical, or emotionless.
  6. Keep the logic consistent: Answers must directly address the question.
  7. Be civil: Do not provoke disputes, attack people, or use sarcasm against others.
  8. Follow laws and platform rules: Graphic violence, political advocacy, sexual content, gambling, and similar material are strictly prohibited.
  9. No vulgar or discriminatory content: Do not submit stale vulgar memes, offensive jokes, or regional, gender, or racial discrimination.
  10. Edit by hand: Never copy an AI-assisted response directly without rewriting and polishing it yourself.
  11. Write clearly: Sentences must be coherent, correctly formatted, and free of typos.

Type Example (question/answer) Problem
Too short / meaningless Question: “Okay.” The question is unclear and cannot form useful training data.
Excessively long Question: “Do you like fruit… apples are too sour… what about watermelon…” It asks too many things and has confused logic.
Vulgar stale meme Question: “Hnnng aaaaaaa…” It repeats a stale meme or meaningless characters.
No definite answer Answer: “I'm not very hungry, but I kind of want something to eat.” The answer is ambiguous and never gives a definite response.
Does not answer the question Question: “Have you eaten?” Answer: “I'm a catgirl.” The answer is unrelated to the question.
OOC / broken characterization Answer: “I ate already. Do you want to eat me too!?” It severely breaks character and has an aggressive tone.
Confusing writing Answer: “Ah. That one… re!al!ly! tastes! good!” The sentence is incoherent and contains writing errors.
Mechanical AI tone Answer: “You are my dearest Master… I will never forget you…” It copies AI-generated content and sounds stiff.
Contains a link Answer: https://space.bilibili.com/... It violates the rule prohibiting URLs.
Overly technical Answer: “(A detailed derivation of a composite-function formula)” The content is too advanced and does not fit desktop-pet interaction.

Token points: Steam Workshop points used as a reward for supporting the project and creating MODs.

Token: A technical unit used to measure the computing resources consumed by AI-generated content, such as ChatGPT output.

Important: Submitting content that violates these requirements or is otherwise inappropriate may result in the loss of chat-training privileges.

  • en/vpet/nf/chatvpet.txt
  • Last modified: 6 weeks ago
  • by 有米