Warning

 

Close
Confirm Action

Are you sure you wish to do this?

Cancel Confirm
AR15.COM
Previous Page
/ 17
Next Page
6/29/2025 1:05:24 PM EDT
[#1]
Nice!

6/29/2025 2:25:07 PM EDT
[#2]
@GySgt_D
Multi-GPU is pretty awesome. It doesn’t speed up video generation, but by loading the clip, vae, and clipvision on the 3090 and the main model on the 5090, I can generate a 293-frame video versus a 173 frame video on the 5090 alone.

6/29/2025 3:01:51 PM EDT
[#3]
Hmmmm

I have a spare 3070 laying around, but I suspect that stuff would over-heat if I cram it into my case.
6/29/2025 3:11:21 PM EDT
[#4]
Quote History
Quoted:
Hmmmm

I have a spare 3070 laying around, but I suspect that stuff would over-heat if I cram it into my case.
View Quote

The helper GPU doesn’t need to run at sustained load any longer than it takes to load the models. Then the 5090 takes over and the 3090 core goes back to idle and it essentially serves as a VRAM pack animal running at 30w.

Also, by plugging the monitor into the helper GPU, you free up about 2gb vram on the 5090.
6/29/2025 3:19:36 PM EDT
[#5]
I've been plugging my monitor into the mobo; using the on-board cpu graphics.  Just to save a little vram on occasion...

As far as a second gpu goes, it isn't so much that it is getting warm, but that it would block the intake of cool air into my precious 5090..


(using a new fp32 T5 text encoder model)




6/29/2025 3:54:21 PM EDT
[#6]
Quote History
Quoted:
I've been plugging my monitor into the mobo; using the on-board cpu graphics.  Just to save a little vram on occasion...

As far as a second gpu goes, it isn't so much that it is getting warm, but that it would block the intake of cool air into my precious 5090..

https://www.youtube.com/watch?v=hrXIH19En2U
(using a new fp32 T5 text encoder model)

https://www.youtube.com/watch?v=ZTR1AtpNH5g

https://www.youtube.com/watch?v=4EzqD5LR6lc
View Quote

Can you put your 5090 on the lower slot?
6/29/2025 4:03:46 PM EDT
[#7]
IIRC, the upper one is the only gen5 pciex16 slot

Which reminds me-  I need to remove the nvme out of the m.2-2 slot.  It is nerfing the gpu bandwidth to x8



This seems to do a good job, btw- Wan2.1 prompt generator


6/29/2025 5:22:28 PM EDT
[#8]
Now I need to do another, being this was my testing phase and seeing this over and over...lol



6/29/2025 10:48:58 PM EDT
[#9]




Friendly reminder that these are meant to be viewed in 1440p.  Smack that gear icon and make it so.

6/29/2025 11:49:06 PM EDT
[#10]
Quote History


Thats pretty damn cool. I like 2nd one best
7/1/2025 12:08:01 AM EDT
[#11]
Testing a new WAN text to video workflow


7/1/2025 12:25:41 AM EDT
[#12]
It really does an amazing job compared to the other models.

7/5/2025 8:43:57 PM EDT
[#13]
7/6/2025 2:10:00 PM EDT
[#14]
Quote History


Is that 100% AI or did it start with a photo and video clip to animate it?

7/6/2025 3:33:02 PM EDT
[#15]
Video-to-video, which is why the movement is so good.

I realize that it is practically cheating, lul.
7/6/2025 8:48:31 PM EDT
[#16]
Quote History

Yeah that looks good. I’ve been traveling since Wednesday and can’t wait to try some new unlimited travel workflows when I get back tomorrow.
7/6/2025 9:20:26 PM EDT
[#17]
Someone help a brother out.

Attached File
7/6/2025 11:03:05 PM EDT
[#18]

7/7/2025 1:55:28 AM EDT
[#19]
As per usual, this was made with Wan2.1 image2video.  Differently this time, a locally-installed LLM (Gemma3 27B) analyzed the picture, and I used its description as the prompt.  



That light reflection at the top of the pic didn't do it any favors.
7/7/2025 5:46:01 AM EDT
[#20]
That's wild. Thanks guys .
7/7/2025 8:55:48 PM EDT
[#21]
Retro anime-style text-to-video workflow. It comes with some naughty NSFW loras that are disabled by default

https://civitai.com/models/1671285/retro-90s-anime-golden-boy-style-lora-wan-14b-t2v-i2v?modelVersionId=1963359



7/7/2025 9:44:35 PM EDT
[#22]
Would someone try to take my car;





...and try to make it realistically doing Ed Roth style stuff?

Something like this Bigfoot pic;


Just the car, but driving and shooting flames like these;


7/8/2025 12:05:52 AM EDT
[#23]
Quote History
Quoted:
Would someone try to take my car;

https://i.postimg.cc/qqRMHVbX/67-Firebird-1.jpg

https://i.postimg.cc/T1CKST7M/AI-67-Firebird-Burnout.gif

...and try to make it realistically doing Ed Roth style stuff?

Something like this Bigfoot pic;
https://i.postimg.cc/qRTDZQ06/Bigfoot_Phone_Cover.jpg

Just the car, but driving and shooting flames like these;
https://i.postimg.cc/5jcMKvy0/Ed-Roth-Mustang.png

View Quote



7/8/2025 2:15:44 AM EDT
[#24]
Thank you.
7/8/2025 2:33:42 PM EDT
[#25]


Made this for a different thread, but........

I then remembered the audience, lul
7/12/2025 1:48:39 PM EDT
[#26]
Playing with SeedVR2 upscaler


@andyBroshankle

Upscaled with SeedVR1, then loaded it into vanilla FramePack F1 with Triton and SageAtttention 2.2
7/12/2025 3:58:45 PM EDT
[#27]
I need to figure out how to make a lora.  Have been using Reactor for years to swap-in faces, but that works poorly when the subject is turning their head....
7/12/2025 7:21:33 PM EDT
[#28]
Original Buffalo Dance video from 1894:


AI Upscaled. I cut the original video into 4 parts, upscaled each, then stitched back together. I also threw in some "fake frames" do double it from 30 to 60fps. Set to 4K in YouTube.


Quote History
Quoted:
I need to figure out how to make a lora.  Have been using Reactor for years to swap-in faces, but that works poorly when the subject is turning their head....
View Quote

That's on my list of things to do.
7/12/2025 10:24:06 PM EDT
[#29]








My brain is too tired to come up with something amusing or creative





7/14/2025 8:10:45 PM EDT
[#31]
I made my first LoRA (low rank adaptation) by training photos of Cristin Milioti. You can train one using photographs of any person, whether its a celebrity, yourself, family member, or deceased relative. It just takes about 20-30 minutes to resize, caption and upload the photos, then let it marinade in the trainer for a few hours. Once the LoRA model is ready, pop it into a text to image or text to video model and type the "trigger word" in the prompt, in this case "Milioti." It is an easier and more accurate way of developing consistent, repeatable characters than image to video.

7/14/2025 11:08:07 PM EDT
[#32]
I may have to try learning that, because I still have trouble getting Framepack to follow prompts. (Even the ones generated on that prompt website). I'm going to try upgrading my PC to Win11 tonight, so if that breaks things I'll have to start over anyway.
7/15/2025 12:30:33 AM EDT
[#33]
I tried using Wan2.1 t2v as a text-to-image model for the first time today.  It does a very good job.  Perhaps even the best job.



7/15/2025 8:16:54 PM EDT
[#34]
I made a new LoRA, "Emma"


Quoted:
I may have to try learning that, because I still have trouble getting Framepack to follow prompts. (Even the ones generated on that prompt website). I'm going to try upgrading my PC to Win11 tonight, so if that breaks things I'll have to start over anyway.
View Quote

If you do exactly what he says in this video and how he says it, you can have a fully functional ComfyUI and LoRA generator. No coding required.
https://www.youtube.com/watch?v=Na5BEQWFmic

Quoted:
I tried using Wan2.1 t2v as a text-to-image model for the first time today.  It does a very good job.  Perhaps even the best job.

https://i.postimg.cc/bJjsn4yY/Wan-T2-I-0208.jpg

https://i.postimg.cc/Bv2vJNxr/Wan-T2-I-Reactor-0211.jpg
View Quote

Wan and Hidream are both great for text to image.
7/15/2025 11:07:57 PM EDT
[#35]
Thank you. I'm bookmarking that video, but I'll wait until after I upgrade my computer. I need to ask my son what motherboard to get, and I'm trying to find a left hand case, as budget friendly as possible.
7/15/2025 11:56:11 PM EDT
[#36]
Quote History
Quoted:
I made a new LoRA, "Emma"
https://www.ar15.com/media/mediaFiles/563643/FusionX_00064-ezgif_com-optimize-3587946.gif


If you do exactly what he says in this video and how he says it, you can have a fully functional ComfyUI and LoRA generator. No coding required.
https://www.youtube.com/watch?v=Na5BEQWFmic


Wan and Hidream are both great for text to image.
View Quote View All Quotes
View All Quotes
Quote History
Quoted:
I made a new LoRA, "Emma"
https://www.ar15.com/media/mediaFiles/563643/FusionX_00064-ezgif_com-optimize-3587946.gif

Quoted:
I may have to try learning that, because I still have trouble getting Framepack to follow prompts. (Even the ones generated on that prompt website). I'm going to try upgrading my PC to Win11 tonight, so if that breaks things I'll have to start over anyway.

If you do exactly what he says in this video and how he says it, you can have a fully functional ComfyUI and LoRA generator. No coding required.
https://www.youtube.com/watch?v=Na5BEQWFmic

Quoted:
I tried using Wan2.1 t2v as a text-to-image model for the first time today.  It does a very good job.  Perhaps even the best job.

https://i.postimg.cc/bJjsn4yY/Wan-T2-I-0208.jpg

https://i.postimg.cc/Bv2vJNxr/Wan-T2-I-Reactor-0211.jpg

Wan and Hidream are both great for text to image.


So is that a completely AI generated female making a sandwich, or one from a photo or video that's re-0animated to be making a sammich?  I guess I'm confused on what the things are, especially a LoRa (which I  think is an artificial human "image" that can be made to do any number of things in a text to video sort of environment.  I know of LoRa as a slow speed RF communication method for micro-controllers.

If she's imaginary, may as well give her nice C cups or so.
7/16/2025 12:02:44 AM EDT
[#37]
Quote History
Quoted:


So is that a completely AI generated female making a sandwich, or one from a photo or video that's re-0animated to be making a sammich?  I guess I'm confused on what the things are, especially a LoRa (which I  think is an artificial human "image" that can be made to do any number of things in a text to video sort of environment.  I know of LoRa as a slow speed RF communication method for micro-controllers.

If she's imaginary, may as well give her nice C cups or so.
View Quote View All Quotes
View All Quotes
Quote History
Quoted:
Quoted:
I made a new LoRA, "Emma"
https://www.ar15.com/media/mediaFiles/563643/FusionX_00064-ezgif_com-optimize-3587946.gif

Quoted:
I may have to try learning that, because I still have trouble getting Framepack to follow prompts. (Even the ones generated on that prompt website). I'm going to try upgrading my PC to Win11 tonight, so if that breaks things I'll have to start over anyway.

If you do exactly what he says in this video and how he says it, you can have a fully functional ComfyUI and LoRA generator. No coding required.
https://www.youtube.com/watch?v=Na5BEQWFmic

Quoted:
I tried using Wan2.1 t2v as a text-to-image model for the first time today.  It does a very good job.  Perhaps even the best job.

https://i.postimg.cc/bJjsn4yY/Wan-T2-I-0208.jpg

https://i.postimg.cc/Bv2vJNxr/Wan-T2-I-Reactor-0211.jpg

Wan and Hidream are both great for text to image.


So is that a completely AI generated female making a sandwich, or one from a photo or video that's re-0animated to be making a sammich?  I guess I'm confused on what the things are, especially a LoRa (which I  think is an artificial human "image" that can be made to do any number of things in a text to video sort of environment.  I know of LoRa as a slow speed RF communication method for micro-controllers.

If she's imaginary, may as well give her nice C cups or so.

She is a character model trained on my GPU using photos. I load the LoRA model into my ComfyUI text to video or text to image workflow and she does whatever I say in the prompt. No further images needed. Give me some prompts and I can generate an image or video. I’m still figuring out this training thing, but this particular LoRA doesn’t seem to shut up

7/16/2025 12:29:58 AM EDT
[#38]
Quote History
Quoted:

She is a character model trained on my GPU using photos. I load the LoRA model into my ComfyUI text to video or text to image workflow and she does whatever I say in the prompt. No further images needed. Give me some prompts and I can generate an image or video. I’m still figuring out this training thing, but this particular LoRA doesn’t seem to shut up

https://www.ar15.com/media/mediaFiles/563643/FusionX_00109-ezgif_com-optimize-3587959.gif
View Quote View All Quotes
View All Quotes
Quote History
Quoted:
Quoted:
Quoted:
I made a new LoRA, "Emma"
https://www.ar15.com/media/mediaFiles/563643/FusionX_00064-ezgif_com-optimize-3587946.gif

Quoted:
I may have to try learning that, because I still have trouble getting Framepack to follow prompts. (Even the ones generated on that prompt website). I'm going to try upgrading my PC to Win11 tonight, so if that breaks things I'll have to start over anyway.

If you do exactly what he says in this video and how he says it, you can have a fully functional ComfyUI and LoRA generator. No coding required.
https://www.youtube.com/watch?v=Na5BEQWFmic

Quoted:
I tried using Wan2.1 t2v as a text-to-image model for the first time today.  It does a very good job.  Perhaps even the best job.

https://i.postimg.cc/bJjsn4yY/Wan-T2-I-0208.jpg

https://i.postimg.cc/Bv2vJNxr/Wan-T2-I-Reactor-0211.jpg

Wan and Hidream are both great for text to image.


So is that a completely AI generated female making a sandwich, or one from a photo or video that's re-0animated to be making a sammich?  I guess I'm confused on what the things are, especially a LoRa (which I  think is an artificial human "image" that can be made to do any number of things in a text to video sort of environment.  I know of LoRa as a slow speed RF communication method for micro-controllers.

If she's imaginary, may as well give her nice C cups or so.

She is a character model trained on my GPU using photos. I load the LoRA model into my ComfyUI text to video or text to image workflow and she does whatever I say in the prompt. No further images needed. Give me some prompts and I can generate an image or video. I’m still figuring out this training thing, but this particular LoRA doesn’t seem to shut up

https://www.ar15.com/media/mediaFiles/563643/FusionX_00109-ezgif_com-optimize-3587959.gif


Ok, so a still photo with her in a pose would work as an accurate female model for a figure drawing?   Those are Normally copyrighted and protected but for people learning art I think this could open up a lot of new material to practice/learn from instead of paying $50/photo just to practice pencil drawing accurately getting the anatomy of muscles correct and things like that.   I suppose chaining from there, it's going to shake the foundations of the pr0n industries as well.


7/17/2025 3:53:31 AM EDT
[#39]
7/17/2025 7:58:58 AM EDT
[#40]
Just bought a rog zephyrus, ryzen 9ai hx, 5070ti 32gb i think. Please I'm sure it's shit or something, but I know nothing and wanted a laptop. My ai sidekick recommended it, for what I need.

But I can access this use with my lt now?

Do you guys have coding knowledge to use this? I'm thinking of learning python. But ai said you're maybe using light coding?
7/17/2025 8:42:00 AM EDT
[#41]
Quote History
Quoted:
Just bought a rog zephyrus, ryzen 9ai hx, 5070ti 32gb i think. Please I'm sure it's shit or something, but I know nothing and wanted a laptop. My ai sidekick recommended it, for what I need.

But I can access this use with my lt now?

Do you guys have coding knowledge to use this? I'm thinking of learning python. But ai said you're maybe using light coding?
View Quote

You have 16 GB VRAM which more than enough to join the party. If you run out or VRAM, you increase the number of blocks to swap to the DRAM, run a lower floating point format (FP8 vs FP16) or lower the resolution. The majority of people on the AI Discords have 4 to 16GB vram and are making impressive stuff, even with as little as 4GB. I’d suggest upgrading the DRAM to 96GB, which is easy on a laptop. You’ll also need more storage than 1TB, but that’s also easy to upgrade.

You don’t need to code to set up ComfyUI or the FramePack standalone. The link to FramePack is in the OP. For ComfyUI, just follow the instructions in this video:
https://youtu.be/Na5BEQWFmic?si=zrZymT9V8p-NYsf-
7/17/2025 8:53:18 AM EDT
[#42]
Quote History
Quoted:

You have 16 GB VRAM which more than enough to join the party. If you run out or VRAM, you increase the number of blocks to swap to the DRAM, run a lower floating point format (FP8 vs FP16) or lower the resolution. The majority of people on the AI Discords have 4 to 16GB vram and are making impressive stuff, even with as little as 4GB. I'd suggest upgrading the DRAM to 96GB, which is easy on a laptop. You'll also need more storage than 1TB, but that's also easy to upgrade.

You don't need to code to set up ComfyUI or the FramePack standalone. The link to FramePack is in the OP. For ComfyUI, just follow the instructions in this video:
https://youtu.be/Na5BEQWFmic?si=zrZymT9V8p-NYsf-
View Quote View All Quotes
View All Quotes
Quote History
Quoted:
Quoted:
Just bought a rog zephyrus, ryzen 9ai hx, 5070ti 32gb i think. Please I'm sure it's shit or something, but I know nothing and wanted a laptop. My ai sidekick recommended it, for what I need.

But I can access this use with my lt now?

Do you guys have coding knowledge to use this? I'm thinking of learning python. But ai said you're maybe using light coding?

You have 16 GB VRAM which more than enough to join the party. If you run out or VRAM, you increase the number of blocks to swap to the DRAM, run a lower floating point format (FP8 vs FP16) or lower the resolution. The majority of people on the AI Discords have 4 to 16GB vram and are making impressive stuff, even with as little as 4GB. I'd suggest upgrading the DRAM to 96GB, which is easy on a laptop. You'll also need more storage than 1TB, but that's also easy to upgrade.

You don't need to code to set up ComfyUI or the FramePack standalone. The link to FramePack is in the OP. For ComfyUI, just follow the instructions in this video:
https://youtu.be/Na5BEQWFmic?si=zrZymT9V8p-NYsf-

Thank you, I'll see if I can figure it out. My ai assistant hallucinated, told me 8gb vram, that liar. Or wait, now it's reversed back to 8gb with pretty strong argument. Either way it's above 6gb from the op. Glad it can handle this, glad the thread popped up to remind me.

ETA now *12gb final answer, believe her this time, but down votes are coming for some of her answers ??.
7/18/2025 2:18:24 AM EDT
[#43]




"The camera glides slowly over a tranquil forest lake at golden hour, capturing vivid reflections of the fiery orange, pink, and purple sunset in the still water. Pine trees on the right glow as sunbeams pierce through them. The camera gently orbits above the lake, dipping low to skim the mirror-like surface, then rising to reveal the endless sky. A warm breeze rustles the grass and ripples the water slightly. Birds fly across the sunset in the distance. Ultra high quality, 4K, cinematic, natural movement, vibrant colors, peaceful yet majestic atmosphere"

Used a video prompt for t2i.  First gen tonight.  I like it.



7/19/2025 4:06:22 PM EDT
[#44]


The sword kept on glitching.

Entirely AI using a character I made.  First step was text-to-image using Hunyuan, then image-to-video using WanFusion.  All locally generated.

The background movement made the first attempts not so good at 60fps, so bumped it up to 180.
7/19/2025 4:15:15 PM EDT
[#45]
Quote History
Quoted:
https://www.youtube.com/watch?v=4iVqXVDO7Lg

The sword kept on glitching.

Entirely AI using a character I made.  First step was text-to-image using Hunyuan, then image-to-video using WanFusion.  All locally generated.
View Quote



That's amazing.  You have ea video clip of somebody that doesn't and hasn't existed before.  Even 10 years ago it was assumed creative3 process would be needed to create characters from some base outline actor, similar to Deep Fakes or in the other direction with Davy Tentacle Face in Pirates of Caribbean and others where motion tracking was used for everything linked to facial muscles because of number and being too hard to create without hitting uncanny valley.

7/19/2025 4:21:33 PM EDT
[#46]
Quote History
Quoted:
https://www.youtube.com/watch?v=4iVqXVDO7Lg

The sword kept on glitching.

Entirely AI using a character I made.  First step was text-to-image using Hunyuan, then image-to-video using WanFusion.  All locally generated.
View Quote

Have you made a LoRA out of her yet? Then you can summon her with text to video.
7/19/2025 4:33:50 PM EDT
[#47]
No.  Every time I attempt to do so, even including the instructions you posted, I get nothing but error messages.

Really want to do so.  The face will glitch badly if turned more than perhaps 45 degrees in any direction.



Even at 150fps ^, the background was a little choppy.
7/19/2025 4:47:46 PM EDT
[#48]
Quote History
Quoted:
No.  Every time I attempt to do so, even including the instructions you posted, I get nothing but error messages.

Really want to do so.  The face will glitch badly if turned more than perhaps 45 degrees in any direction.
View Quote

They're pretty awesome. If you send me the data set, I can make you one. The data set is 20-30 images, each resized 768x768, along with a caption for each image that describes everything in detail (LLMs can help with this).

7/19/2025 11:23:42 PM EDT
[#49]
I hope to do something along these lines soon. My new case will be here in a few days, and then my son will help me with transferring parts and then upgrade to Win11. Then I'll put the better programs on, and hope that I can learn them.

I'm surprised that the one above won't turn a face, because FramePack has been creating and turning them for me.

From this;


To this;


7/19/2025 11:31:45 PM EDT
[#50]
Quote History
Quoted:
I hope to do something along these lines soon. My new case will be here in a few days, and then my son will help me with transferring parts and then upgrade to Win11. Then I'll put the better programs on, and hope that I can learn them.

I'm surprised that the one above won't turn a face, because FramePack has been creating and turning them for me.

From this;
https://i.postimg.cc/DwY6t0C4/2000s-Linda-Church-1.jpg

To this;
https://i.postimg.cc/WpM9h24x/2025-5-22-AI-1st.gif

View Quote

The one GySgt_D and I are using (Wan 2.1) has the same problem as Framepack in that it changes the face. A LoRA solves this problem because it is a digital model of whoever you want it to be, trained with either images, or a combination of images and video. You can then summon them into your scene by entering their name into the prompt. You get the same person every time and it doesn’t change their physical characteristics.


Sign up to continue the discussion

Create a free account to share your thoughts, follow topics, and connect with the AR15.COM community.

Already a member? Sign In

Previous Page
/ 17
Next Page