Running OpenAI locally hits memory and speed limits
MLOps2 posts from 2 people2d active+00 posts in the last 7 days, 1 the 7 days before (fading)Unverified
Posts per day
The posts behind it
2, newest first| Post | Date | |
|---|---|---|
| Support Prism ternary GGUFs (PQ2_0 type 142 / PTQ1_0 type 143): import fails with unsupported tensor size overflows...and then decodes garbage. Worth knowing for anyone who tries the "portable" pack as a workaround. ## What it would take 1. llama.cpp: kernels for types 142/143 plus the `prism.hadamard.*` activation transform, tracked in ggml-org/llama.cpp#29058 (the metadata contract, block...ollama/ollamaQuentinDanblon | ollama/ollamaQuentinDanblon9 reactions, 2 replies | Sep 187 days ago |
| RFC: image, video and audio generation from diffusion GGUFs (LTX-2)### Prerequisites - [x] I am running the latest code. Mention the version if possible as well. - [x] I carefully followed the [README.md](https://github.com/ggml-org/llama.cpp/blob/master/README.md). - [x] I searched using keywords relevant to my issue to make sure that I am...ggml-org/llama.cppericcurtin | ggml-org/llama.cppericcurtin3 reactions, 6 replies | Sep 72 weeks ago |
Companies and products named
About this problem
Evidence
2 posts from 2 people in 2 places, about 1 a week over 12 days. Mostly on ollama/ollama, ggml-org/llama.cpp. Tools named alongside: llama.cpp, Hugging Face, NVIDIA, Ollama.
Frustration Frustration 3 of 3· Seen on GitHub issues
How it was grouped
Posts that state a pain and match the "local-llm" rule about OpenAI. First post Sep 7, 2026, latest Sep 18, 2026. Unverified: fewer than 3 posts, or one person or place only. Method
The same problem elsewhere
- Running Ollama locally hits memory and speed limits (7)
- Running llama.cpp locally hits memory and speed limits (7)
- Running LLMs locally hits memory and speed limits (5)
- Running Qwen locally hits memory and speed limits (4)
- Running vLLM locally hits memory and speed limits (3)
- Running DeepSeek locally hits memory and speed limits (2)
Other OpenAI problems
History
- 2026-09-25 Added: Running OpenAI locally hits memory and speed limits (2 posts)