I’ve got two setups: a 24GB VRAM GPU at home and a 16GB VRAM GPU at work. Followed a few tutorials and got a basic Ollama setup running Hermes, with Gemma, and Qwen models (very basic). My main use case is game dev tooling for blender or unreal, so mostly Python and C++, and one off small pipeline script.
This is a companion discussion topic for the original entry at https://www.reddit.com/r/LocalLLM/comments/1v8f8d6/local_ai_for_game_dev_tooling_am_i_doing_this/