No issues here. Nemotron is still my main LLM for most simple tasks and it taps Claude and GPT API models for anything more complex. Speed isn’t really an issue. Usually get a response within 20-seconds or so. Not the fastest in the world but not slow enough to be an issue.