Programming

27223 readers

297 users here now

Welcome to the main community in programming.dev! Feel free to post anything relating to programming here!

Cross posting is strongly encouraged in the instance. If you feel your post or another person's post makes sense in another community cross post into it.

Hope you enjoy the instance!

Rules

Follow the programming.dev instance rules
Keep content related to programming in some way
If you're posting long videos try to add in some form of tldr for those who don't want to watch videos

Wormhole

Follow the wormhole through a path of communities !webdev@programming.dev

founded 3 years ago

MODERATORS

snowe@programming.dev

Ategon@programming.dev

UlrikHD@programming.dev

bugsmith@programming.dev

Spyro@programming.dev

Ask Lemmy: What do you currently use for AI coding? (lemmy.world)

submitted 4 days ago by joelthelion@lemmy.world to c/programming@programming.dev

43 comments fedilink hide all child comments

Given how quickly things evolve, it's easy to get lost in the numerous offerings and hard to get the best deal. So, what do you use? Both clients/harnesses and LLM providers or local setups would be interesting.

Personally, I've been using opencode with Github copilot for work. I'm currently looking for cost-effective provider for personal work. Maybe openrouter with one of the cheap models?

you are viewing a single comment's thread
view the rest of the comments

[–] mike_wooskey@lemmy.thewooskeys.com 20 points 4 days ago (1 children)

I use opencode with locally-hosted llama.cpp - usually with qwen3.6-35b-a3b.

I tried opencode go for a couple month, and its definitely nice to have an lln runner with more gram and more GPUs, but I prefer to have all my stuff local whenever it's possible. Also, I'd use up my token allotments fairly quickly on opencode go.

I also tried opentouter and it, too, was great - many more models. But I exhausted by credits even quicker than opencode go, and its also not local.

[–] joelthelion@lemmy.world 6 points 4 days ago (1 children)

What hardware do you use? How fast is it?

[–] mike_wooskey@lemmy.thewooskeys.com 12 points 4 days ago (1 children)

AMD Ryzen 9 9950X CPU and AMD radeon pro w7900 (48GB vram). I get 55tps output pretty consistently, but ingesting context starts around 1500tps and if context size reaches, say, 50K, tps drops to around 200tps. I often have to wait a bit, but it's a price I'm happy to pay for local-only AI

[–] joelthelion@lemmy.world 2 points 4 days ago

Thank you!