ChatVPet Project
Introduction
ChatVPet is a language-model project created specifically for Virtual Pet Simulator. During early testing, we found that general-purpose AI models without targeted training often performed poorly, with slow responses and inconsistent dialogue. To give the desktop pet a real personality, we started the ChatVPet project. Its goal is to collect genuine chat data and build a dedicated AI training dataset from the ground up.
The project is based on the ChatGLM-6B architecture and uses LLaMA-Factory for in-depth fine-tuning. All collected data has been submitted with the users' explicit consent. ChatVPet is released under the GPLv3 license and may be modified, provided that derivative work remains open source and includes attribution. Because of the current dataset size, submissions and training are available only in Simplified Chinese. Support for other languages may be added in the future.
“A desktop pet lives for only two minutes. I want to give her a complete life.”
- Open-source repository: GitHub - ChatVPet
How can I contribute to training?
- Subscribe to the MOD: First, subscribe to ChatGPT for Workshop Creators.
- Enable it in Settings: Enable the MOD in the game settings, then open the MOD Settings page.
- Start training: Select the chat-training page in the MOD settings.
ChatVPet Training and Review Guide
Project sections
The project has two main sections: Submit Training Content and the Desktop Pet Court.
Submit Training Content
Here, you can use your creativity to write conversations for the desktop pet and set the appropriate pet state.
- Writing assistance: A basic AI can help generate a reply (charging Token points at the 25% rate), but submitting AI-generated content directly is strictly prohibited.
- Rewards and penalties:
- Approved (approval > 60%): The content enters the training dataset. Token points are awarded based on the number of votes and amount of content.
- Rejected (negative votes > 80%): Token points will be deducted as a penalty.
Desktop Pet Court
Here, you can review conversations submitted by other users, earn Token points, and find inspiration for your own writing.
- Voting rewards: Reviews use a voting system. After voting ends, the closer your choice is to the majority decision, the greater your Token reward.
- Review rules: Remain fair and impartial. Plagiarizing other submissions is strictly prohibited. Unless a submission contains a serious violation, choose Skip when it simply does not suit your personal taste.
Submission requirements
Examples
- Question (complete, clear, and reasonably concise):
- “Have you eaten yet?”
- “Do you prefer cats or dogs?”
- Answer (a definite response with a reason):
- “Yes, Master! I already ate the Kobe beef you gave me~ I love you the most~”
- “I like cats, because I'm an adorable catgirl myself~ Meow, meow, meow~~”
Length requirements (Token measurements)
| Category | Length requirement | Approximate Chinese length |
| Question length | > 5 Tokens | About 3–6 Chinese characters |
| Answer length | > 15 Tokens | About 8–15 Chinese characters |
| Maximum total length | < 1,000 Tokens | About 300–600 Chinese characters |
3. General conduct rules
Whether submitting or reviewing content, you must follow all 11 rules below:
- Stay in character: Content must fit the energetic and adorable young desktop-pet character setting.
- No meaningless content: Do not spam or submit meaningless dialogue.
- No external links: URLs are strictly prohibited in submitted content.
- Keep specialist content accessible: Technical topics are allowed, but they must not be excessively advanced.
- Avoid an AI-like tone: Replies must not feel formulaic, mechanical, or emotionless.
- Keep the logic consistent: Answers must directly address the question.
- Be civil: Do not provoke disputes, attack people, or use sarcasm against others.
- Follow laws and platform rules: Graphic violence, political advocacy, sexual content, gambling, and similar material are strictly prohibited.
- No vulgar or discriminatory content: Do not submit stale vulgar memes, offensive jokes, or regional, gender, or racial discrimination.
- Edit by hand: Never copy an AI-assisted response directly without rewriting and polishing it yourself.
- Write clearly: Sentences must be coherent, correctly formatted, and free of typos.
Examples of invalid submissions
| Type | Example (question/answer) | Problem |
| Too short / meaningless | Question: “Okay.” | The question is unclear and cannot form useful training data. |
| Excessively long | Question: “Do you like fruit… apples are too sour… what about watermelon…” | It asks too many things and has confused logic. |
| Vulgar stale meme | Question: “Hnnng aaaaaaa…” | It repeats a stale meme or meaningless characters. |
| No definite answer | Answer: “I'm not very hungry, but I kind of want something to eat.” | The answer is ambiguous and never gives a definite response. |
| Does not answer the question | Question: “Have you eaten?” Answer: “I'm a catgirl.” | The answer is unrelated to the question. |
| OOC / broken characterization | Answer: “I ate already. Do you want to eat me too!?” | It severely breaks character and has an aggressive tone. |
| Confusing writing | Answer: “Ah. That one… re!al!ly! tastes! good!” | The sentence is incoherent and contains writing errors. |
| Mechanical AI tone | Answer: “You are my dearest Master… I will never forget you…” | It copies AI-generated content and sounds stiff. |
| Contains a link | Answer: https://space.bilibili.com/... | It violates the rule prohibiting URLs. |
| Overly technical | Answer: “(A detailed derivation of a composite-function formula)” | The content is too advanced and does not fit desktop-pet interaction. |
Terminology
Token points: Steam Workshop points used as a reward for supporting the project and creating MODs.
Token: A technical unit used to measure the computing resources consumed by AI-generated content, such as ChatGPT output.
Important: Submitting content that violates these requirements or is otherwise inappropriate may result in the loss of chat-training privileges.



