brokenlcd

joined 2 years ago
[–] [email protected] 0 points 12 hours ago* (last edited 12 hours ago) (1 children)

Is bifurcation necessary because of how CUDA works, or because of bandwidth restraints? Mostly asking because for the secondary card i'll be limited by the x1 link mining risers have (and also because unfortunately both machines lack that capability. :'-) )

Also, if i offload layers to the GPU manually, so that only the context needs to overflow into RAM, will that be less of a slowdown, or will iti comparable to letting model layers into ram? (Sorry for the question bombing, i'm trying understand how mutch i can realistically push the setup before i pull the trigger)

[–] [email protected] 0 points 22 hours ago (3 children)

You need a 15$ electrical relay board that sends power from the motherboard to the second PSU or it won't work.

If you are talking about something like the add2psu boards that jump the PS_ON line of the secondary power supply on when the 12v line of the primary one is ready. Then i'm already on it the diy way. Thanks for the heads up though :-).

expect 1-5token per second (really more like 2-3).

5 tokens per seconds would be wonderful compared to what i'm using right now, since it averages at ~ 1,5 tok/s with 13B models. (Koboldcpp through vulkan on a steam deck) My main concerns for upgrading are bigger context/models + trying to speed up prompt processing. But i feel like the last one will also be handicapped by offloading to RAM.

How much vram is the 3060 youre looking at?

I'm looking for the 12GB version. i'm also giving myself space to add another one (most likely through a 1x mining riser) if i manage to save up enough another card in the future to bump it up to 24 gb with parallel processing, though i doubt i'll manage.

Sorry for the wall of text, and thanks for the help.

 

I have an unused dell optiplex 7010 i wanted to use as a base for an interference rig.

My idea was to get a 3060, a pci riser and 500w power supply just for the gpu. Mechanically speaking i had the idea of making a backpack of sorts on the side panel, to fit both the gpu and the extra power supply since unfortunately it's an sff machine.

What's making me weary of going through is the specs of the 7010 itself: it's a ddr3 system with a 3rd gen i7-3770. I have the feeling that as soon as it ends up offloading some of the model into system ram is going to slow down to a crawl. (Using koboldcpp, if that matters.)

Do you think it's even worth going through?

[–] [email protected] 3 points 1 week ago* (last edited 1 week ago)

I don't have adhd, but femtanyl, kmfdm and justice are wonderful when crunching for exams.

[–] [email protected] 3 points 1 week ago

No no no. You don't get it. The turd is turning into a werewolf mid-shit.

[–] [email protected] 4 points 1 week ago

Tbh, everytime i see soulless corporations trying to look more amicable like this, the only thing that comes to mind is "✨some pretty colors aren't going to erase your sins✨" said in the most cutesy voice imaginable.

[–] [email protected] 4 points 1 week ago (1 children)

Don't give me ideas... I love spicy stuff, and it has been a pretty good deterrent in of itself from having my foodstuffs stolen. So two birds with one stone...

[–] [email protected] 9 points 1 week ago* (last edited 1 week ago) (2 children)

amphotericity is some weird shit, so yes. Water also an acid. (100% butchered the translation)

[–] [email protected] 0 points 1 week ago

It's so hot outside even the birb's melting

[–] [email protected] 3 points 1 week ago

Yeah, but they can't get you to scroll through all the ads if they don't water it down to hell.

[–] [email protected] 21 points 1 week ago (4 children)

I remember solving something similar using an opaque bottle with "GI supplements, don't drink" written in sharpie. Especially since the first time it was actually true and they didn't believe the warning.

[–] [email protected] 3 points 2 weeks ago

Backwards. It will put it back in instead of making it come out.

view more: next ›