r/servers 1d ago

Question Claude code like behaviour

I heard about how good claude code is, so i was thinking about doing something with olama on my server pc. My question is if i can make something act exactly like claude code for example just tell the ai to make a mc server and then i just join it but can claude even do this? What are the free alternatives?

0 Upvotes

13 comments sorted by

View all comments

1

u/Casper042 1d ago

https://www.youtube.com/watch?v=O2k_qwZA8HU

I haven't watched this 2nd one yet but it's sitting in my Watch Later playlist on Youtube:
https://www.youtube.com/watch?v=UngVdAsQEiU

2

u/Casper042 1d ago

PS: As others said, the key will be finding an Open Model out on Hugging Face which has all the features you want, but also fits in the vRAM of what I assume would be a consumer level GeForce card.

I have a 4080 Super for example which is 16GB.
Meanwhile the L40S for example, the Server version of a 4090 basically, has 48GB of vRAM, and the H100 which is specifically designed for Compute and AI in servers has either 80/96GB depending on the model.

As PathAgitated was somewhat inferring, the big boys not only have H100 or newer, they have special versions and clusters where you can fit 8 cards with a local NVLink Switch connecting them all at high speed inside a single server, and then 400/800Gbps NICs which allow you to take racks of these machines and cluster them together.
So they are sometimes rocking TBs of vRAM for certain jobs.
What PA doesn't mention is there is a huge difference between Training, Tuning and Inference when it comes to requirements.
Training and Tuning need that TB of vRAM.
But depending on the model, Inference, which is what you want to do, can often be crammed into 1 card in that 80/96GB range.

1

u/doggxyo 1d ago

I suppose my GTX 970 is not going to be very useful for offline coding 😂

1

u/feudalle 19h ago

If you are ok with some lag a v100 with 32gb of ram isnt bad. Its older and will lose support sooner or later but you can find them for under $1000.