soft whisper ltx23
Karşılaştır
8264°
Otomatik oynat
Yerel olarak çalıştır
a_gentle_whisper.safetensors
349 MB · SafeTensor
İndirmeler Civitai tarafından sunulur. Bazı oluşturucular önce Civitai girişi ister.
Bu Model Hakkında
LTX2 · LoRA'lar
LTX-2.3 can generate dialogue, multi-speaker scenes, and full dynamic range audio including screaming — but it cannot whisper. Use lora strength and choose the correct trigger to control the effect.
Oluşturucu notları
LTX-2.3 Whisper & Soft-Spoken Audio LoRA
Base model: LTX-2.3 · Type: Audio-style LoRA · Rank: 32
---
## What this does
LTX-2.3 can generate dialogue, multi-speaker scenes, and full dynamic range audio including screaming — but it cannot whisper. This LoRA adds two quiet vocal registers to the model:
- Whispering — devoiced, breathy, close-mic delivery
- Soft-spoken — voiced but low-volume, intimate, relaxed
The LoRA targets only the three attention modules that write to the audio branch audio_attn1, audio_attn2, video_to_audio_attn). Video output is provably unchanged — no visual fighting, no style drift.
---
## Usage
Load at strength 1.0. The register is controlled entirely by the manner keyword in your prompt — no special strength tuning needed.
### Trigger words (none, use natural language)
| Whispering | (woman, whispering) | (man, whispering quietly) |
| Soft-spoken | (woman, speaking softly) | (man, speaking softly) |
> Note: Male whisper may requires the extra word quietly to tip the model over. (man, whispering) alone produces soft-spoken, not true whisper.
### Prompt format
Follow the LTX-2.3 dialogue caption style:
```
a [scene description], ([gender], [manner]): "[what they say]", intimate ASMR
```
Examples:
```
a woman sitting close to a microphone in warm dim lighting, (woman, whispering): "close your eyes and listen"
a man at a desk late at night, (man, speaking softly): "I've been thinking about this all day"
a woman doing a skincare routine, (woman, whispering quietly): "this is my favourite step"
```
### Without manner keywords
Using the LoRA without any manner keyword defaults to soft-spoken — a subtle volume-softening effect on whatever the base model would have generated. Useful as a gentle "quieter audio" modifier.
---
## What it can't do
- No intra-clip register mixing. You can't have one character whisper and another speak normally in the same clip. The register applies to the whole generation. For mixed-register dialogue, generate each part separately and cut them together.
- No magic above the vocoder ceiling. The audio chain passes through a mel spectrogram bottleneck. Breathy whisper HF energy gets partially smoothed. Expect intimate and quiet, not studio-crisp ASMR.
- Video is untouched by design. If you want the visuals to also feel ASMR (soft lighting, close-up framing), describe that in the scene prompt — the LoRA won't help or hurt.
---
## Training details
| | |
|---|---|
| Base model | LTX-2.3 dev |
| Steps | 2000 |
| Rank / Alpha | 32 / 32 |
| Target modules | audio_attn1, audio_attn2, video_to_audio_attn |
| Training resolution | 192×192, 97 frames (~4s @ 24fps) |
| Dataset | 74 clips, 8 voices (4F / 4M), 2 registers each |
Clips were 4-second segments sourced from ASMR content across 8 speakers — 4 female (2 soft-spoken, 2 whisper) and 4 male (2 soft-spoken, 2 whisper). Captions used Whisper ASR transcription in (gender, manner): "transcript", intimate ASMR format.
SHA256
DBBDC78B8B6EF0405FE237B21ED6BB61EEE0621665C267EA890EBB7784D5CECA
ModelTrends.ai Model Kimliği
#24136
İçeriği bildir
Örnek çizimler
Bu çizimleri Civitai'de yapan kişilerin paylaştığı prompt'lar ve ayarlar. Nasıl yapıldığını görmek için bir çizim seçin.
Ayarlar
Seed
914554288
Boyut
768x1152
Kullanılan ağırlık
1
Prompt
A video of a nervous blonde woman in a hallway. She is very shy, nervous, and awkward.
"Hey, I'm Linda ... I live downstairs." She says nervously. "Um, like, every night its like, really loud, from your apartment? I'm trying to study and there are a lot of...creaks and, like loud women's moans."
She pauses.
"I don't want to like, stop you from doing whatever, but do you think you could keep it down?
Ayarlar
Boyut
448x640
Prompt
A brunette woman with wavy hair is sitting in a brightly lit office setting at her desk. A computer monitor sits to the side. She is looking at her mobile phone.
She looks up to face the cameras.
(Woman, whispering): “Hey, come here for a second. I have something to tell you.”
Ayarlar
Seed
1445100364
Boyut
768x1152
Kullanılan ağırlık
0.45
Prompt
A video of a naturally pretty, gorgeous Latina Mexican woman in a an apartment filled with moving boxes. She places the box she is holding down on the stack of boxes and turns and faces the viewer.
"I think that's it! That was very nice of you to help me out even though we literally just met."
She looks apologetic.
"I wish I could make you dinner or something to thank you, but ... not really possible at the moment."
She thinks, and then awkwardly giggles.
"Maybe ... do you want a sloppy spitty deepthroat blowjob by any chance?" She asks, flirtatiously.
Ayarlar
Seed
125903531
Boyut
768x1152
Kullanılan ağırlık
0.1
Prompt
Beautiful busty Indian girl in hallway. She is shy, awkward but friendly.
"Oh, hey neighbor." she says, giggling. "Been a while since we hung out."
She smiles suggestively.
"It's so fucking hot out, but I just went to the farmers market and was about to make a salad and take a shower."
She looks to the left and right.
"You should join." She says flirtatiously.
Ayarlar
Boyut
448x640
Prompt
A man, in his 30s, in a t-shirt and shorts, sitting at the far end of a plush living room sofa, with warm, dim lighting. He holds a remote control in his hand. 3/4 shot.
He looks up to face the camera.
(Man, softly spoken): “Hey, let’s watch a movie. It’s your turn to pick.”
Benzer Modeller
Görsel, Video ve Metin Üretimi için Yapay Zeka Modelleri
soft whisper ltx23, açık kaynak yapay zeka görsel, video ve metin modellerinin salt okunur kataloğu ModelTrends.ai'de listelenen Ltx2 ailesine ait bir LoRA modelidir. Diğer Ltx2 modelleriyle karşılaştırın, trendini görmek için popülerlik puanına bakın ve denemek için PirateDiffusion veya BitVector'da açın.
Model Ailesine Göre Göz At
Text / LLM Modelleri
411
Anima Modelleri
509
Chroma Modelleri
14
Flux Modelleri
1,071
Flux 2 / Klein Modelleri
1,237
MiniMax H3 Modelleri
331
Hunyuan Modelleri
197
Ideogram Modelleri
20
Krea2 Modelleri
674
Qwen Modelleri
17
Qwen2 Modelleri
36
Ltx2 Modelleri
150
Stable Diffusion 1.5 Modelleri
6,465
Stable Diffusion XL 1.0 Modelleri
15,432
Zimage Modelleri
804
Wan Modelleri
919

Model
Trends
.ai
© ModelTrends.ai
|
Japonya'da yapıldı
|
© 2026
