Pairing a small local Qwen3 4B model with Immersive Translate for a private, uncapped translation setup — plus a prompt trick that speeds it up.
Running ten locally-hosted models on consumer GPU hardware to see which ones best summarize customer-service calls after speech-to-text.
Installing curl via snap put it ahead of /usr/bin on PATH, which broke the Ollama installer in a confusing way.
A one-line fix for a CUBLAS_STATUS_NOT_INITIALIZED crash: just update Ollama.