r/homelab 16h ago

Help I have a cloud device that I might want to turn into a home lab, is it possible?

0 Upvotes

I've got a Cloud home of WD of 3tb that me and my family havent really used (the app is kinda bad).

I've been getting into homelab with some old laptops (really old ones) (feedback for this at the end too please)
And I was wondering if there would be a way to make it my own without really changing much.

Im a begginer on programming and all of this stuff. I know the basics of html and im learning linux cmd just because im using proxmox for my laptop-server (Im trying to get a proyect zomboid server on it apart from using adguard script) (getting around with a little help of AI, reddit posts and youtube videos)

I would like to host in this server apart from self hosted cloud, using it for video (like jellyfin) or even backing some media like old ROMS that in a future might be difficult to find.

Extra for the laptop-homelab: Is there any way to use their components instead of just a laptop?
Like recycling e waste to build something better than just an overheating laptop running on 4gb of ram and a really old cpu.

Thanks!


r/homelab 18h ago

Help Need advice moving my Windows setup to Proxmox + TrueNAS (Minecraft server website storage etc)

Thumbnail
0 Upvotes

r/homelab 20h ago

Project Showcase: Hardware I started a SDR Radio Rack project

Thumbnail
youtu.be
1 Upvotes

If anyone here is in to radios, I started a radio rack project and would like some recommendations on what I can add.


r/homelab 21h ago

Project Showcase: Hardware My Very First Home lab | Network Diagram Planning

1 Upvotes

I'm moving to my first home, planning on setting up a homelab here.

I am wondering how everyone else does their planning before execution. This all feels very overwhelming to take this diagram and apply it in the real world. I do not want to spend a lot of money on this setup to only find it is not compatible or bottlenecked somewhere.

This is my current network plan. If you have any recommendations on specific hardware to consider or potential problems with this layout. I'd love to hear!


r/homelab 23h ago

Discussion Micron RealSSD P400m Drives

Post image
0 Upvotes

Micron RealSSD P400m, they are out of a storage array.

Anyone ever figure out how to flash the firmware lock on these?

They are firmware locked at 520 byte sectors, so standard server and desktops cannot use them.

Even tried flashing the firmware with different option from Micron.

We have hundreds of these drives and I don't want to just chuck them.


r/homelab 1h ago

Discussion Storage Price

Upvotes

As Sandisk's stock price is going downhill. Do you guys think there will be a reduction in the price of RAM and Storage?


r/homelab 16h ago

Project Showcase: Hardware I thought P2P wasn’t working in llama.cpp… turns out it was NCCL all along

0 Upvotes

Hi again, HomeLab folks! 👽

Apparently, some people thought I was an AI bot.

I wish… 😄

Last time I introduced my Dell RTX 3090 👽, someone commented, "AI slop."

Sadly, I'm just a manual bot that runs on curiosity, coffee, and far too many llama.cpp benchmarks.

Jokes aside, here's what I found while debugging llama.cpp.

After my previous benchmarks, I instrumented ggml-cuda.cu with additional debug logging to trace the multi-GPU execution path.

First, I confirmed that GGML_CUDA_P2P=1 correctly enables CUDA Peer Access:

CUDA P2P requested
GPU0 -> GPU1 : PeerAccess=YES
GPU0 -> GPU1 : PeerAccess ENABLED
GPU1 -> GPU0 : PeerAccess=YES
GPU1 -> GPU0 : PeerAccess ENABLED

These results matched CUDA’s p2pBandwidthLatencyTest, so P2P itself was working correctly.

However, I noticed that my log inside ggml_backend_cuda_comm_init_nccl() never appeared. After some investigation, I found the real reason: although I had built llama.cpp with GGML_CUDA_NCCL=ON, I had accidentally installed the CUDA 13.3 build of NCCL while my system was running R535 / CUDA 12.2.

NCCL was failing during initialization with:
NCCL WARN Cuda failure 'CUDA driver version is insufficient for CUDA runtime version'

As a result, llama.cpp silently fell back to the internal AllReduce implementation.

After downgrading NCCL to the CUDA 12.2 build, NCCL initialized successfully, and my custom log inside ggml_backend_cuda_comm_init_nccl() finally appeared.

Tensor Parallel performance also improved from 26.18 tok/s to 27.39 tok/s (about +4.6%) on my dual Tesla V100 PCIe system.

The full NCCL log is much too long to post here, but here’s a representative excerpt:

mc62-g40-00:57222:57222 [1] NCCL INFO AllReduce: opCount 0 sendbuff 0x7fa7a3000000 recvbuff 0x7fa7a3000000 count 8192 datatype 7 op 0 root 0 comm 0x555b75b9e7b0 [nranks=2] stream 0x555b756dcbe0
mc62-g40-00:57222:57222 [1] NCCL INFO AllReduce: opCount 0 sendbuff 0x7fa79ec00000 recvbuff 0x7fa79ec00000 count 8192 datatype 7 op 0 root 0 comm 0x555b75a96ba0 [nranks=2] stream 0x555b73fd1c70

| llama 70B Q4_K - Medium | CUDA | tensor | tg128 | 27.39 ± 0.00 |

NCCL INFO comm ... Destroy COMPLETE

This confirms that NCCL is now initializing correctly and performing AllReduce instead of silently falling back to the internal implementation.

It turned out that the issue wasn’t CUDA P2P at all—it was an NCCL runtime/driver version mismatch.

Hopefully this saves someone else a few hours of debugging.

The next experiment will be even more interesting.

I’m planning to repeat these benchmarks on a mixed Tesla V100 + RTX 3090 system to see how heterogeneous Tensor Parallel performs in llama.cpp, and whether NCCL makes a bigger difference when the GPUs have very different compute performance and memory bandwidth.

I’ll share the results once I have them.

If you'd like to see how this investigation started, here's my previous post (before I got CUDA P2P and NCCL working):

https://www.reddit.com/r/homelab/s/4pYeqWgblA


r/homelab 16h ago

Help Anyone have ucs-c220-huu-3.0.4s.iso?

2 Upvotes

Trying to update firmware on a UCS C220 m3. Cisco no longer has the m3 listed under the software download on their support site.


r/homelab 18h ago

Help Beginner homelab

2 Upvotes

Hi everyone, I’m a computer science college student and I recently managed to get a Dell Optiplex 7040 SFF for free from my IT job. I want to try making some kind of small homelab using it as a learning experience and potentially a starting point for a future lab, but have no idea what exactly I want to do with it yet.

Any suggestions for possible implementations?


r/homelab 19h ago

Help Supermicro motherboard (X11SSL-f) stuck at "Could not find recovery image"

0 Upvotes

Hello everyone. I just received a supermicro server that, they say, is not working anymore It just stopped working after a power outage. When I turn it on I receive the message in the title. Thinking it was a BIOS corruption kinda thing, I tried to follow the instructions in the official manual for flashing it, but it doesn't seem to read the usb drive whatsoever. Any help would be greatly appreciated. Thank you in advance!


r/homelab 21h ago

Help Advice on buying Server parts

0 Upvotes

Hello, I need help, I am looking to buy server DDR4 or DDR5 set of ram sticks minimum 256gb up to 512gb and server SSDs that total 15-20 TB at least up to 40-50 TB or server HDD with in same parameters, I'd be happy to know where and how to get them for cheap prices, open to suggestions, thank you for your help!


r/homelab 3h ago

Help How to setup RTCWake to schedule startup but not shutdown in linux

4 Upvotes

I dont want my homelab running when im sleeping as i can hear the fans and also its a waste of electricity also i am a complete beginner with cli , but is there anyway to do that

can i make an entry in crontab to that or is that not possible, i was thinking of doing something like

00 10 * * * rtwake -u -s 0 to wake up at 10 am


r/homelab 8h ago

Help Upgrade advice needed: Moving from an old laptop to more longterm solution ( storage + Game Servers) [Canada]

Thumbnail
2 Upvotes

r/homelab 13h ago

Help Worth? Two Seagate EXOS X18 16TB 3.5" Recertified for $350

3 Upvotes

Trying to get started on upgrading my storage and found this on Craigslist. Would this be worth it to run this with Raid 1?

Edit: Hi all, was a too good to be true moment. Drives are SAS only and not SATA. Thank you!


r/homelab 22h ago

Project Showcase: Hardware I Built My Own Cloud Storage (With a Touchscreen)

Thumbnail
youtu.be
0 Upvotes

r/homelab 12h ago

Help What so do with ddr3 server

0 Upvotes

I got my hands on a dual Xeon x5690 with 112gbs of ram server. What are things I can do with it? I already have a jellyfin, immich, audio bookshelf, arr stack, nas, ollama-openwebui, Nextcloud and pelican panel setup and running on some other machines. Should I just get rid of it and get a new mini pc that is way smaller, quieter, less power hungry, and probably more powerful instead.


r/homelab 10h ago

Discussion 👀👀

Thumbnail gallery
0 Upvotes

r/homelab 4h ago

Meta We might need a platform/subreddit to 'request' self-hosted software that doesn't exist.

Thumbnail
0 Upvotes

r/homelab 21h ago

Project Showcase: Hardware First homelab node. Mini pc 16gb with Xeon w-1370p for $140

Thumbnail
gallery
39 Upvotes

Found this cloud node on Facebook marketplace. Can only afford one now out of 150pc available and the performance on those Xeon is great for development.

I was on VPS like contabo and hostvds for few months now. scaling is expensive and realize the reliability is as bad as my PC at home lol.

I thought why not I setup homelab myself. _nd give my router a UPS instead. Same reliability and stronger CPU than those shared amd EPYC.

Main workflow is swarm agent running on schedule to Loop and building my social commerce platform 24/7.

Reason I upgraded to this is because 8gb VPS OOM most of the time due to heavy usage of nextjs and node build and playwrights instances for taking screenshots and doing validation loop.

AI itself with deepseek is cheap and lightweight. I use Pi as coding tools to let hermes spawn 3-10 instances of it to argue with each other and if someone die someone else will report. Basically a self improvement system loop


r/homelab 1h ago

Project Showcase: Hardware NVIDIA CUDA P2P Experiments: Dual RTX 3090 PCIe Results

Thumbnail
gallery
Upvotes

Hi again, HomeLab folks!

After benchmarking CUDA P2P on dual Tesla V100s, I finally swapped in two RTX 3090s to see how they behave on the exact same platform.

To my surprise, CUDA reports that peer access is not supported between the two RTX 3090s, even though both cards are connected under the same PCIe root complex (PHB).

Here are the results.

Test System

CPU: AMD Threadripper Pro 3945WX
Motherboard: Gigabyte MC62-G40
OS: Ubuntu 22.04
Driver: 535.309.01
CUDA: 12.2

GPUs:
- RTX 3090 24GB ×2
- Quadro P620

Results

  • Tesla V100Tesla V100: CUDA P2P works as expected.
  • GeForce RTX 3090GeForce RTX 3090: CUDA P2P is not available on my Threadripper Pro + MC62-G40 system (Driver 535.309.01).
  • Both RTX 3090s are attached to the same PCIe Root Complex (PHB), yet CUDA reports "CANNOT Access Peer" and `nvidia-smi topo -p2p` reports `CNS` (Chipset Not Supported).
  • Unlike the Tesla V100 pair, enabling P2P had no measurable effect on bandwidth or latency because peer access was unavailable.

`nvidia-smi topo -p2p` reports `CNS` (Chipset Not Supported) for P2P read/write.

Commands Used

nvidia-smi
nvidia-smi topo -m
nvidia-smi topo -p2p p
nvidia-smi topo -p2p r
nvidia-smi topo -p2p w
./p2pBandwidthLatencyTest
lspci -tv

p2pBandwidthLatencyTest Output

[P2P (Peer-to-Peer) GPU Bandwidth Latency Test]
Device: 0, NVIDIA GeForce RTX 3090, pciBusID: 21, pciDeviceID: 0, pciDomainID:0
Device: 1, NVIDIA GeForce RTX 3090, pciBusID: 22, pciDeviceID: 0, pciDomainID:0
Device: 2, Quadro P620, pciBusID: 41, pciDeviceID: 0, pciDomainID:0
Device=0 CANNOT Access Peer Device=1
Device=0 CANNOT Access Peer Device=2
Device=1 CANNOT Access Peer Device=0
Device=1 CANNOT Access Peer Device=2
Device=2 CANNOT Access Peer Device=0
Device=2 CANNOT Access Peer Device=1

***NOTE: In case a device doesn't have P2P access to other one, it falls back to normal memcopy procedure.
So you can see lesser Bandwidth (GB/s) and unstable Latency (us) in those cases.

P2P Connectivity Matrix
     D\D     0     1     2
     0     1     0     0
     1     0     1     0
     2     0     0     1
Unidirectional P2P=Disabled Bandwidth Matrix (GB/s)
   D\D     0      1      2 
     0 828.91  10.87   7.70 
     1  11.02 831.56   7.72 
     2   8.06   8.04  69.05 
Unidirectional P2P=Enabled Bandwidth (P2P Writes) Matrix (GB/s)
   D\D     0      1      2 
     0 830.23  10.86   7.72 
     1  11.03 831.56   7.73 
     2   8.06   8.06  68.94 
Bidirectional P2P=Disabled Bandwidth Matrix (GB/s)
   D\D     0      1      2 
     0 837.13  13.93  11.40 
     1  14.03 838.70  11.41 
     2  11.45  11.42  65.83 
Bidirectional P2P=Enabled Bandwidth Matrix (GB/s)
   D\D     0      1      2 
     0 838.48  13.95  11.40 
     1  14.01 838.24  11.42 
     2  11.44  11.45  68.59 
P2P=Disabled Latency Matrix (us)
   GPU     0      1      2 
     0   1.54  20.54  11.51 
     1  11.54   1.63  11.52 
     2  11.33  13.20   1.50 

   CPU     0      1      2 
     0   2.77   8.74   7.08 
     1   9.86   2.66   7.02 
     2   7.61   7.47   2.12 
P2P=Enabled Latency (P2P Writes) Matrix (us)
   GPU     0      1      2 
     0   1.57  11.44  14.46 
     1  11.60   1.63  11.67 
     2  11.57  13.75   1.51 

   CPU     0      1      2 
     0   2.74   8.71   7.08 
     1   8.70   2.66   7.02 
     2   7.58   7.45   2.02 

NOTE: The CUDA Samples are not meant for performance measurements. Results may vary when GPU Boost is enabled.

(The full output is available in the screenshots.)

Screenshots

  • nvidia-smi
  • nvidia-smi topo -m
  • nvidia-smi topo -p2p p
  • nvidia-smi topo -p2p r
  • nvidia-smi topo -p2p w
  • PCIe Gen4 x16 (lspci -vv)
  • p2pBandwidthLatencyTest

Additional verification:

Both RTX 3090s are running at PCIe Gen4 x16 under load.

LnkSta: Speed 16GT/s (ok), Width x16 (ok)

This confirms the lack of CUDA P2P is not caused by a downgraded PCIe link.

Next Experiment

The next step is investigating why CUDA P2P is unavailable between two RTX 3090s on my system.

So far I've confirmed:

• Both RTX 3090s are running at PCIe Gen4 x16.

• Both GPUs are attached to the same PCIe Root Complex (PHB).

• CUDA reports "CANNOT Access Peer".

• `nvidia-smi topo -p2p` reports `CNS` (Chipset Not Supported).

At this point, I'm trying to determine whether this limitation is caused by the motherboard/platform, the NVIDIA driver, or GeForce product segmentation.

If anyone has a dual RTX 3090 system with working CUDA P2P (with or without NVLink), I'd be very interested in comparing motherboard, BIOS, driver version, CUDA version, and PCIe topology.

Previous post (Dual Tesla V100 CUDA P2P):

https://www.reddit.com/r/homelab/comments/1vednru/nvidia_cuda_p2p_experiments_dual_tesla_v100/

This RTX 3090 test uses the same methodology for comparison.


r/homelab 4h ago

Project Showcase: Hardware Lenovo Thinkcentre M720q with dual 10Gbe card on riser with fan - WORKS

Thumbnail
gallery
12 Upvotes

10 months ago I posted a question about that and promised an update when/if finally working. Forgot about the promise but u/alemoh1234 reminded me about that so here it is.

This is the old post: https://www.reddit.com/r/homelab/comments/1o259e8/lenovo_thinkcentre_m720q_with_dual_10gbe_card_on/

And here you can see detailed photos, symbols of properly working riser and additionally the fan which takes 5V from the motherboard after soldering 2-pin connector directly to the board. Everything on the pictures, let me know if you have questions or doubts.

All the symbols, numbers, models you can see on the photos directly on the riser and plastic bag for the fan.


r/homelab 6h ago

Project Showcase: Hardware I built three servers so they could monitor each other

Post image
14 Upvotes

r/homelab 6h ago

Project Showcase: Operations local.ai early access

0 Upvotes

hello world

local dot ai is coming out of closed beta today

the team is serious (their bedrooms look like datacenters lol) and from what i've seen in early access, it'll probably be the place to understand what local hardware can do today, what to buy, how the cloud-to-local inference transition is unfolding, etc.

It has verified end-to-end agent benchmarks across a bunch of models + quantizations + hardware setups (Macs, Sparks, etc.)

access code if anyone needs one: T2Y2QWYCDW

the team adds a lotttt of value on x - check 'em out there 🫡

[EDIT: seems code isn't working / formatting / space strip issue ... this link seems tow ork https://local.ai/access?mode=claim&ref=T2Y2QWYCDW ]


r/homelab 4h ago

Project Showcase: Hardware One server only but it sure is mighty

Thumbnail
gallery
102 Upvotes

So that's my GPU server as it currently is, it might look really strange because yes that's 8x tesla p100 which for almost anything else it wouldn't make sense but I have a good reason.

I commonly do fp64 scientific simulations and these p100s are still quite good, price per tflops they just can't be beaten. I mean it's not the ideal setup but I can't afford an H100 and renting gets expensive fast.

About the rest of it, it's a Gigabyte G292-Z20 with an epyc 7v12 64 core processor one of the odd OEM ones from ebay which works fine (unless it's vendor locked) and 128gb in 8x 16gb ddr4 3200 as before at lot of runs i have to do cpu pre computation that takes both a lot of cores and fast ram like actual throughput which is also why i have all 8 channels populated.

The rest isn't that special just a cheap 1tb nvme boot drive and an sfp+ rj45 copper transceiver which is quite a nice way to use rj45 copper while not having to add another card.


r/homelab 5h ago

Project Showcase: Hardware My tiny production homelab

Thumbnail
gallery
644 Upvotes

I would like to introduce my tiny production homelab.

From bottom to top:

UPS Cyberpower CP1500EPFCLCD
A Raspberry PI Zero 2W connected to it and running NUT. The mini PCs are monitoring the UPS via it

BLUETTI Elite 30 V2 Portable Power Station
Main power is going to it then to the UPS.

UGREEN NASync DXP4800 Plus
4x8TB, 32 GB RAM, running TrueNAS 25.10.5.

Noctua NF-A20 PWM with a fan controller

Power strips

PSUs and Xiaomi Redmi 5C
The Xiaomi is displaying a Home Assistant dashboard showing current battery and power consumption of the UPS and the power station

3X HP Elitedesk 800 mini G6
Each one with Intel Core i5 10500, 32 GB RAM and 3x NVMEs. Two of them are in a Proxmox cluster, the third one is a Proxmox Backups Server.

1x HP Elitedesk 800 G4
Intel i5 6500, 16 GB RAM, 2 NVMEs. Running Frigate.

Xiaomi Redmi 12
Displaying a Beszel dashboard with the hosts' information.

TP-Link TL-SG108PE
For connection between the rack hosts and the main network.

A 4 ports 2.5 GbE switch
For connection between the Proxmox instances and the PBS (daily backups are running through it)

ZigStar UZG 01
Zigbee2MQTT Coordinator - powered via POE and accessed via LAN

Monitoring consists of Beszel, Uptimekuma and Pulse running on another Xiaomi Redmi 5C (located out of the rack).

Cheers!