Model
Trends
.ai
Model
Trends
.ai
những mô hình AI mã nguồn mở tốt nhất

MiniMax H3 Combat BASE V2: chuyển động chiến đấu, va chạm và kịch tính

Chủ đề: Mô hình
Bởi Captain
Đăng ngày 2026-10-01
Share
Ảnh mẫu
Prompt và thiết lập do những người tạo ra các ảnh này trên Civitai chia sẻ. Chọn một ảnh để xem nó được tạo như thế nào.
Thiết lập
Số bước
12
Guidance
1
Kích thước
1440x832
Prompt
For the target video, at 0.00 seconds into the target video, <Picture 1> (from [Shot 1]) is fully referenced. integrated_multimodal_description: [Shot 1] The scene opens from <Picture 1>, preserving the three young women standing close together outdoors on a city street in winter, with European-style buildings and bare trees in the background under soft daylight. The camera has a natural handheld smartphone selfie shakiness, as if held by one of the girls at arm's length. The girl on the left with long straight black hair, a thick gray scarf, and a dark coat (S1) shivers slightly, pulling her scarf tighter around her neck with one hand. She looks directly into the camera lens with a bright, slightly complaining expression and says: <d>[Chinese] 今天有点冷啊,快看镜头啊!我们三个谁最好看啊?</d> The girl in the middle with chestnut-brown hair, a gray coat with bow buttons, and a white phone with a bear charm (S2) turns her gaze toward the camera, a playful smile forming on her lips. The girl on the right with a single braid, a gray sweater, and a gold necklace (S3) continues looking down at her dark phone, smiling faintly. [Shot 2] At 00:05.000, the handheld camera wobbles slightly as the girls adjust their positions. The middle girl (S2) raises her hand in a cute peace-sign gesture beside her face, tilts her head playfully, and laughs with a cheeky, confident expression, saying: <d>[Chinese] 当然是我啦!</d> She winks one eye playfully after the line. The left girl (S1) reacts with a mock-offended expression, opening her mouth in playful protest and lightly nudging the middle girl's shoulder with her elbow. The right girl (S3) finally looks up from her phone toward the camera, an amused smile spreading across her face as she shakes her head at the other two. [Shot 3] At 00:10.000, the right girl (S3) laughs warmly, her eyes crinkling with genuine amusement, and says in a light, scolding-but-affectionate tone: <d>[Chinese] 两个臭屁精!别拍了,赶快吃火锅去啦!</d> She puts her dark phone away into her coat pocket and starts to turn her body toward the street, gesturing with her hand for the others to follow. The middle girl (S2) giggles and lowers her peace-sign hand, while the left girl (S1) laughs and nods in agreement. The handheld camera begins to lower and tilt as if the person filming is putting the phone down, the three girls starting to move together toward the street, their shoulders brushing as they walk off, still laughing. The frame ends with a slight downward tilt as the selfie video concludes. overall_soundscape: Soft outdoor city street ambience continues throughout — distant traffic hum, occasional car passing, faint wind rustling through the bare tree branches, and the low murmur of pedestrians in the background. The girls' voices are clear and close, captured by the phone's microphone with a slight natural reverb from the open street. Subtle fabric rustling is audible as they adjust their coats and scarves, and the soft click of a phone lock button is heard near the end as the right girl puts her phone away. Light, cheerful laughter from all three girls overlaps naturally throughout the conversation. non_diegetic_music: N/A
Thiết lập
Kích thước
960x544
Prompt
V2BASE
Thiết lập
Kích thước
1376x768
Prompt
integrated_multimodal_description: [Shot 1] A surreal martial-arts action scene. A woman with exactly three legs stands in the center of a large training courtyard. She is wearing a traditional white karate gi with a black belt. She is surrounded by a group of masked ninjas dressed entirely in black, they are bare handed. The woman suddenly raises her body onto her middle leg, using it as a single central pivot. Her two other legs extend horizontally in opposite directions, forming a three-point spinning shape. She begins spinning rapidly around her middle leg like a human whirlwind. The middle leg remains planted firmly on the ground and acts as the central axis, while the two outer legs stay fully extended and rotate around her. Several ninjas attempt to rush toward her from different directions, but each time one approaches, one of her two extended legs swings around and delivers a powerful spinning kick, sending the ninja tumbling backward. Other ninjas try to approach from different angles, only to be knocked away by the rotating legs. The choreography is clear and continuous: one leg remains the stationary pivot while the other two legs remain extended and sweep around her in a wide circular motion. The woman maintains perfect balance and moves with exaggerated martial-arts precision. The scene is spectacular and absurd, played completely straight, with dynamic camera movement emphasizing the circular motion and the increasingly desperate ninjas surrounding her. overall_soundscape: Ambient noise. non_diegetic_music: N/A
Thiết lập
Kích thước
720x960
Prompt
V2BASE
Thiết lập
Kích thước
416x736
Prompt
the girl from example.png finally meets ryu from streetfighter in a dark ally. the girl proves far more superior in hand to hand combat since she is fighting in her own comfyui environment using a modified workflow.
Thiết lập
Số bước
25
Seed
265159317415717
Kích thước
1472x1472
Prompt
prfight2, high-end claymation combat sequence, tactile stop-motion texture. A square golden-brown waffle fighter with short articulated limbs and glossy chocolate-syrup eyes is surrounded in a narrow wooden breakfast arena by a surging horde of round pancakes; only three foreground pancake warriors have crude short stubby articulated limbs and receive individual hit reactions while the background mass closes as one blurred threat. The waffle is on the defensive throughout: every foreground exchange begins with a pancake attack, followed by a visible block, parry, or dodge from the waffle, then one heavier counter that opens space to his right. Sequential single-target exchanges only; the camera pans with him to the right as he moves through the gaps created by his counters. A first pancake lunges from ahead-right with an overhead batter-punch at his head; the waffle ducks under it and turns its momentum aside with a low forearm parry, then drives one heavy upward elbow strike into the pancake’s exposed center; it bends violently upward, cracks along its edge, syrup spurts outward, and collapses flat behind him, opening a narrow gap to his right. A second pancake immediately surges through that gap with a horizontal batter-punch at his chest; the waffle absorbs it against his raised guard in a loud wet snap, recoils half a step right but stays upright, then answers instantly with a low forearm strike under the attacker’s rim; the pancake flips over, batter splashes in an exaggerated arc, lands flat and crumples with visible clay deformation, clearing another lane segment to his right. A third pancake lunges overhead from the closing horde with a broadside edge punch; the waffle ducks under it and counters before it can recover by driving one heavy straight punch into its center; the attacker cracks diagonally, tumbling pieces flatten behind him as he continues rightward. The remaining blurred pancake mass recoils and collapses in broad waves behind him, ending with a low-angle victory pose at the right edge of the cleared arena as flattened pancakes trail behind him. Dynamic tracking shot panning sharply right with his defensive movement throughout, impact-synced camera kicks, exaggerated squash-and-stretch while preserving coherent silhouettes, heavy physical consequences after every attack and counter. Audio: Heavy punch and guard-clash impacts from both sides: wet slaps, batter snaps, floppy pancake squelches on incoming pancake attacks; deep cartoonish bass thuds and crunchy waffle cracks on the waffle’s defensive counters; thick syrup splatter whooshes; low-end impact drops synced to every landed counter hit. Percussive rhythm, no melody.

Tổng quan

MiniMax H3
Combat
BASE (V2) của FourBunny là một video
LoRA
tập trung vào hành động: chiến đấu nhanh hơn, tính liên tục hành động tốt hơn, phản ứng khi bị đánh đáng tin cậy, tổn thương tích lũy, di chuyển không gian và nguyên nhân kết quả vật lý. V2 còn mở rộng hơn các trận đánh, thêm chuyển động cơ thể phong phú hơn, chi tiết âm thanh tinh tế hơn và tính liên tục giữa các cảnh mạnh mẽ hơn trong biểu diễn nhân vật.
Trigger: prfight1 cho chiến đấu bình thường; thêm prfin1 sau đó để có cú kết thúc rõ ràng, hạ gục hoặc quật ngã; thêm prslow1 để nhấn mạnh chậm trong giây lát vào cú đánh quyết định. Tác giả lưu ý slow motion không phải lúc nào cũng đảm bảo.

Tham khảo

Tên
Kiểu
Ý nghĩa
Trigger
prfight1
Cảnh chiến đấu bình thường.
Finisher
prfin1
Sau prfight1 để kết thúc quyết định.
Slow motion
prslow1
Nhấn mạnh chậm trong giây lát vào cú đánh cuối.
LoRA
weight
0.7-1.0
Resolution
1080p / 2 MP
Clip 10 giây mất khoảng 15 phút trên thẻ 12 GB theo tác giả.
AD
Không cần cài đặt – chạy mô hình AI trên đám mây
BitVector Prism là cách dễ nhất để bắt đầu: chọn một mô hình, nhập prompt và tạo ảnh trong một ứng dụng web gọn gàng, không cần cấu hình gì. BitVector cũng có trên Discord và trên web (SpyGlass).

Từng bước

  1. Mô tả các võ sĩ, không gian và ai đang thắng trong phần mô tả tích hợp.
  2. Đặt prfight1 ở đầu; thêm prfin1 nếu clip kết thúc trận đấu.
  3. Mô tả âm thanh (va chạm, thở) trong phần âm thanh tổng thể.
  4. Trọng số 0.9, các bước H3 tiêu chuẩn.
  5. Giữ mỗi cảnh chỉ một lần trao đổi đòn đánh.

Ví dụ

Tập luyện đối kháng

prfight1. integrated_multimodal_description: [Shot 1] Two adult boxers spar in a dim gym. The taller one jabs, the shorter slips and lands a body shot; the taller staggers back into the ropes.
overall_soundscape: Glove impacts, squeaking shoes, heavy breathing.
non_diegetic_music: N/A

Cú kết thúc

prfight1, prfin1, prslow1. integrated_multimodal_description: [Shot 1] A swordswoman parries a strike and sweeps her opponent's legs; he falls and the blade stops at his collar.
overall_soundscape: Steel ring, boots on gravel, wind.

Ý tưởng sáng tạo: đối thoại căng thẳng

prfight1. integrated_multimodal_description: [Shot 1] Two negotiators face off across a table. One slams a hand down and says: <d>[English] Enough.</d> The other leans back slowly.
overall_soundscape: Hand on wood, chair creak.

Mẹo

  • Nói rõ phản ứng khi bị đánh bạn muốn (chao đảo lùi, quỳ một gối).
  • Tổn thương tích lũy: nhắc lại các cú đánh trước trong các cảnh sau.
  • Cũng dùng được cho cảnh đối thoại; cải thiện ngôn ngữ cơ thể.
  • Kết hợp với Better Motion ở 0.4 cho đi bộ và quay người.

Khắc phục sự cố

Trận đấu kết thúc không có cú đánh ấn tượng

Vì sao xảy ra
Không có trigger finisher.

Cách sửa
Thêm prfin1 sau prfight1.

Tay chân bị lỗi trong các pha nhanh

Vì sao xảy ra
Quá nhiều hành động trong một cảnh.

Cách sửa
Mỗi cảnh chỉ một lần trao đổi; thêm Better Motion ở 0.4.

AD
PirateDiffusion
Không cần cài đặt – chạy mô hình AI trên đám mây
PirateDiffusion chỉ có trên Telegram và dành cho người dùng chuyên nghiệp: hàng nghìn mô hình, LoRA và workflow điều khiển bằng lệnh chat, tạo không giới hạn với gói giá cố định.

Câu hỏi

Slow motion có đáng tin không?

Tác giả nói prslow1 không đảm bảo; giữ cảnh ngắn.

Cảnh đối thoại thì sao?

Có, V2 cải thiện tính liên tục biểu diễn trong đó.

Liên kết và nguồn

Mô hình trong hướng dẫn này

Người viết
Captain

Hướng dẫn liên quan

Mô hình
GalaxyAce LoRA cho MiniMax H3: video điện thoại tầm thấp đầu những năm 2010
GalaxyAce LoRA của aiguild tái tạo lại kiểu ảnh của camera điện thoại Samsung Galaxy Ace: màu sắc nhạt, nhiễu cảm biến, hiệu ứng cuộn màn trập và cách khung hình đời thường. Tìm hiểu điểm mạnh theo từng model, prompt và ý tưởng video tìm thấy.
Bởi Captain
2026-10-02
Mô hình
Chuyển động người tốt hơn cho MiniMax H3: chuyển động cơ thể tự nhiên, nhất quán
Better Motion LoRA của AdaptiveVision giúp nhân vật MiniMax H3 (và LTX) di chuyển tự nhiên và nhất quán hơn. Tìm hiểu về trọng số, độ phân giải, số bước và cách viết prompt chuyển động ngắn.
Bởi Captain
2026-10-01
Mô hình
LoRA Hôn Nồng Nàn cho MiniMax H3: cảnh lãng mạn thuyết phục giữa người lớn
LoRA của kermitfrog1202 cải thiện các cảnh hôn trong video MiniMax H3. Tìm hiểu câu kích hoạt, các giá trị cường độ tác giả tìm ra cho các cặp đôi khác nhau và cách dựng cảnh lãng mạn tinh tế.
Bởi Captain
2026-10-01
Mô hình
MiniMax H3 Turbo LoRAs: Video 4-8 bước với gói larryvrh
Gói MiniMax H3 Turbo LoRA (larryvrh 4 bước và bạn bè) giảm số bước tạo video từ hàng chục xuống chỉ còn vài bước. Tìm hiểu số bước, bộ lấy mẫu và độ mạnh từng biến thể, cùng cách viết prompt cho H3 để có clip nhanh và sạch.
Bởi Captain
2026-10-01
Viết prompt
Hướng dẫn MiniMax H3: từ văn bản thành video, từ hình ảnh thành video, video tham khảo và âm thanh
Cách điều khiển MiniMax H3 qua lệnh chat: công thức WHO + WHERE + ACTION + CAMERA + SOUND cho việc tạo video từ văn bản, làm chuyển động cho ảnh tĩnh với hình ảnh thành video, giao việc cho từng ảnh tham khảo trong chế độ tham khảo, và tạo thoại rõ ràng thay vì lí nhí. Có kèm các prompt để sao chép và thử, cùng PDF của tác giả.
Bởi Mark White
2026-09-20

Hướng dẫn khác

Dịch vụ đám mây
Phí cố định so với token: tại sao các gói không giới hạn như Graydient.ai là lựa chọn tốt nhất cho nhà sáng tạo AI năm 2026
So sánh chi phí có xếp hạng, với giá được kiểm tra vào ngày 8 tháng 10 năm 2026, giữa các gói không giới hạn phí cố định như Graydient.ai với giá token và tín dụng từ Midjourney, Leonardo, fal.ai, Replicate, Runway và Kling: chi phí thực sự cho 1.000 hình ảnh và 1.000 video mỗi tháng trên mỗi dịch vụ, và tại sao một khoản phí cố định cho hình ảnh, video, âm thanh, trò chuyện Grok LLM và ứng dụng web không giới hạn là giá trị tốt nhất cho những người sáng tạo hay thử nghiệm.
Bởi Captain
2026-10-08
Mô hình
FLUX.1 Dev: cách tạo prompt, cài đặt tốt nhất và ý tưởng sáng tạo
Hướng dẫn thực tế về FLUX.1 Dev của Black Forest Labs: cách tạo prompt bằng ngôn ngữ tự nhiên, cài đặt hướng dẫn và bước, xếp chồng LoRA, hiển thị chữ và ý tưởng prompt tận dụng điểm mạnh của nó.
Bởi Captain
2026-09-30
Mô hình
Stable Diffusion XL 1.0: lời nhắc, cài đặt và những gì nó vẫn làm tốt nhất
Cách tận dụng tối đa mô hình cơ sở chính thức SDXL 1.0: độ phân giải gốc, cài đặt CFG và sampler, bộ tinh chỉnh, lời nhắc tiêu cực và ý tưởng phong cách mà SDXL vẫn nổi bật.
Bởi Captain
2026-10-01
Mô hình
Stable Diffusion 1.5: mô hình cổ điển, dùng đúng cách
Stable Diffusion 1.5 vẫn đáng để biết: độ phân giải phù hợp, CFG và sampler, cách dùng thư viện lớn LoRA và embedding của nó, cùng ý tưởng prompt sáng tạo phù hợp với mô hình 512 pixel.
Bởi Captain
2026-10-02
Mô hình
FLUX.2 Dev: hướng dẫn sử dụng model mới nhất của Black Forest Labs
FLUX.2 Dev hỗ trợ prompt dài hơn, văn bản tốt hơn, thực tế hơn và chỉnh sửa đa tham chiếu. Hướng dẫn này bao gồm quy trình turbo, cài đặt hướng dẫn và độ phân giải, cấu trúc prompt và ý tưởng tận dụng điểm mạnh mới của nó.
Bởi Captain
2026-10-03
Mô hình
Qwen-Image: prompt dài, chữ chuẩn và poster song ngữ
Qwen-Image là model bạn nên dùng khi chữ trong ảnh quan trọng. Học cách viết prompt mô tả dài, tạo chữ tiếng Anh và Trung chính xác, chỉnh CFG và bước, và khám phá ý tưởng sáng tạo với bố cục phức tạp.
Bởi Captain
2026-09-29