Model Information
Description
WAN 2.2 Kiss LoRA v0.2
This is the second version of my WAN 2.2 Kiss LoRA.
After completing my first training experiment, I gained some useful experience and learned a lot from the results. I am now releasing Kiss LoRA v0.2, with a larger dataset, more detailed captions, and additional training steps.
What’s New in v0.2
Compared with the previous version, this release includes several major improvements:
1. Larger Training Dataset
The training dataset has been expanded from 90 videos to 180 videos.
Every training sample is a complete 5-second, 81-frame video clip, rather than a selection of individual frames.
2. More Detailed Captioning
For the first version, I used only one caption for every training video:
kiss
I did this mainly because I was being lazy, and I could not find a reliable method for automatically captioning kissing videos.
For v0.2, I manually reviewed and categorized the training clips more carefully. Different kissing behaviors now use different trigger words, allowing the LoRA to produce a wider range of actions.
3. More Training Steps
The number of training steps has been increased:
v0.1: 1,000 steps
v0.2: 1,500 steps
Trigger Words
The following kissing behaviors were included in the training dataset:
kiss
Standard kissing, with a stronger focus on lip-to-lip contact.
This is the most stable and reliable trigger word.
passionate kiss
Produces more intense and energetic kissing movements.
This prompt generally creates more passionate body language and stronger interaction between the characters.
tongue kiss
Focuses more on visible tongue interaction during kissing.
wet kiss
Intended to produce wetter lips and more visible saliva.
However, this trigger word is currently not very reliable.
lick kiss
Intended for kissing while licking the other person’s lips.
The results are still inconsistent.
tongue lick
Another experimental trigger for tongue-licking behavior.
It does not perform consistently in the current version.
tongue suck
Intended to create tongue-sucking actions.
This trigger word does not appear to work very well yet, possibly because there were not enough matching training samples.
swap spit
Intended to produce visible saliva exchange.
This trigger is also unstable and may require several generations before producing a good result. In some cases, the model may incorrectly interpret saliva as other types of fluid, so expect occasional artifacts or inaccurate generations.
Recommended Base Model
I do not recommend using the official WAN 2.2 model with this LoRA.
In my experience, tongue movements generated by the official model often look unnatural or artificial.
The base model that currently works best with this LoRA is:
WAN 2.2 Smooth Mix I2V V1
This is also the model I personally use for testing.
Smooth Mix already has some native kissing capabilities, but its kissing actions tend to be repetitive and lack variety. However, its motion quality is excellent, and it can make kissing scenes feel much more dynamic and passionate.
For now, I believe it is the most compatible base model for this LoRA.
About Me
I’m a kiss lover.
I created this LoRA purely as a personal passion project because I enjoy kissing scenes and like exploring the many different ways a kiss can be presented.
Kissing is far more varied and interesting than it may first appear. Different movements, emotions, rhythms, and interactions can create completely different results.
When I realized that very few people were creating LoRAs specifically focused on kissing behavior, I decided to make one myself.
I am still a beginner when it comes to training video LoRAs, and I also have a busy full-time job. I can only work on this project during my spare time.
If you are also a kiss lover, feel free to share your feedback, suggestions, test results, or training ideas. You are also welcome to participate in the project so that we can experiment, improve the LoRA, and share our results together.
Future Plans
All current versions were trained using a 512 × 512 bucket resolution.
I am not sure whether training at a higher resolution, such as 768 × 768, would produce better details or more accurate interactions, but it may be worth testing in a future version.
I also plan to experiment with additional behaviors, including:
Face licking
More varied tongue movements
Different kissing positions
More expressive and emotional kissing styles
Improved saliva interaction
More stable close-up kissing
However, collecting and preparing high-quality video datasets is extremely time-consuming. I am also very selective about the quality of the training videos, which makes dataset preparation even more difficult.
Because of this, the next version will probably not be released very soon.
But there will definitely be another version.
Thank you for downloading, testing, and supporting this project. Please share your results and let me know which trigger words work best for you.
My First Wan 2.2 LoRA
This is my first attempt at training a video LoRA.
It is not perfect, but it works—and for me, that already makes it a meaningful first step. The LoRA was trained on Wan 2.2, which I still find to be a reliable foundation for character-focused video generation.
I also experimented with LTX 2.3, but in my experience, it still struggles with character consistency and identity preservation. Facial features can become unstable or distorted, especially during larger movements or longer generations. Because of this, I decided to continue training with Wan 2.2 for now.
Hopefully, more powerful and consistent open video models will arrive in the future. I am looking forward to seeing how the ecosystem develops—and to training better versions as I gain more experience.
Trigger Words
No trigger word is required.
Prompting
The LoRA is designed around scenes featuring two people.
And the prompting just : kiss .
Example prompt:
two people kiss.
They are kissing.
For better results, try adding:
Detailed descriptions of both characters
Clothing, hairstyle, and facial features
Camera movement and shot composition
Lighting and environment
The specific interaction or action
Mood, pacing, and visual style
Since this is my first LoRA, results may vary depending on the base model, resolution, sampler, LoRA strength, and generation workflow.
Thank you for trying it. Feedback, example videos, and workflow suggestions are always welcome.
cxy_kiss
v0.2
Model Details
- Type
- Model Addon
- Subtype
- LoRA
- Created
- Updated
- July 25, 2026