Skip to content

Deploying and serving models is harder than training

MLOps7 posts from 7 people7d active+11 post in the last 7 days, 2 the 7 days before (steady)

Posts per day

Posts per day7 posts, Aug 31 to Sep 25

The posts behind it

7, newest first
PostDate
...on. After all it can be just some GPU server spewing text over network. Like there are no way to connect back to it. Nothing except inference running on servers with GPU so there just nothing to "hack".Hacker News commentsSXXSep 214 days ago
[FR] Add STACKIT Model Serving as a native AI Gateway provider...Serving directly discoverable when configuring LLM endpoints. > #### Why is it currently difficult to achieve this use case? STACKIT AI Model Serving is OpenAI-compatible, so users can potentially work around the lack of a dedicated provider by configuring it through an existing...mlflow/mlflowhaya-fatimaSep 187 days ago
Model tries to fit trajectories and infer their relationships simultaneously - convergence is a nightmare, as expected in hindsight. Help is appreciated. [D]Setup: decompose each entity (companies, in my case) into a small set of aspects, treat each as a trajectory over time rather than a static feature, and set up a PDE where the coupling terms are meant to reveal latent relationships between trajectories no fixed correlation...r/MachineLearningu/Frequent-Pen-9898Sep 169 days ago
[BUG] Gateway secret API base updates reuse stored credentials> [!WARNING] > Before submitting a PR, please make sure that: > - A maintainer has triaged this issue and applied the `ready` label > - This issue has no assignee > - No duplicate PR exists > > PRs not meeting these requirements may be automatically closed. ### Issues Policy...mlflow/mlflowyoussefcamaoSep 112 weeks ago
Vulnerability - [BUG]> [!WARNING] > Before submitting a PR, please make sure that: > - A maintainer has triaged this issue and applied the `ready` label > - This issue has no assignee > - No duplicate PR exists > > PRs not meeting these requirements may be automatically closed. ### Issues Policy...mlflow/mlflowkiran71319-beepSep 92 weeks ago
[BUG] Amazon Bedrock Titan/AI21 completions adapters silently drop top_p/top_k> [!WARNING] > Before submitting a PR, please make sure that: > - A maintainer has triaged this issue and applied the `ready` label > - This issue has no assignee > - No duplicate PR exists > > PRs not meeting these requirements may be automatically closed. ### Issues Policy...mlflow/mlflowYash-ChindamSep 33 weeks ago
Repeated "Selected model is at capacity" errors interrupting Codex tasks on ChatGPT Pro### What issue are you seeing? I am a ChatGPT Pro user using the latest Codex app. Starting today, August 31, 2026, Codex has repeatedly shown the following error: "Selected model is at capacity. Please try a different model." This is happening very frequently during normal...openai/codexgucheng2020-svgAug 313 weeks ago

Companies and products named

Company or productPosts naming it
MLflow4
GitHub3
ChatGPT1
Codex1
Llama1
Ollama1
About this problem

Evidence

7 posts from 7 people in 4 places, about 2 a week over 22 days. Mostly on mlflow/mlflow, Hacker News comments, r/MachineLearning. Tools named alongside: MLflow, GitHub, ChatGPT, Codex.

Frustration Frustration 1 of 3· Seen on GitHub issues, Hacker News, Reddit

How it was grouped

Posts that state a pain and match the "model-serving" rule. First post Aug 31, 2026, latest Sep 21, 2026. Corroborated: 3 or more posts from 2 or more people or places. Method

History

  • 2026-09-25 Added: Deploying and serving models is harder than training (3 posts)

Rising problems by email

Mondays: the problems in data, tech and AI that grew fastest that week.

Double opt-in. Unsubscribe any time.