Skip to content

AI agents are unreliable, loop or stall midway

LLM apps9 posts from 9 people8d active+11 post in the last 7 days, 5 the 7 days before (fading)

Posts per day

Posts per day9 posts, Aug 27 to Sep 25

The posts behind it

9, newest first
PostDate
server-filesystem: tools/list schemas declare unsupported $schema draft-07 dialect, rejected by strict 2020-12 validators## Summary `@modelcontextprotocol/server-filesystem` (tested on the current latest, `2026.8.31`, published 2026-09-17) emits `tools/list` responses where every tool's schema declares: Clients that validate `tools/list` strictly against JSON Schema 2020-12 reject every tool...modelcontextprotocol/serverstomekepSep 232 days ago
[RFC]: Stateless Responses API in the Rust frontend## Motivation. Add native Responses API support to the Rust frontend so clients can use its serving pipeline without a Python API-server fallback. The API must preserve structured conversation input and produce consistent Responses objects whether the client requests buffered...vllm-project/vllmai-jzSep 187 days ago
"Agents for Humans Presents Tendril: From My Daughter's Struggling Balcony Garden to a Goal-Driven, Multi-Agent Intelligence on AWS" by Prashant Baj #agentic-ai #strands-agents #amazon-bedrock-agentcore #llm #hackathonBluesky search@communityaws.bsky.socialSep 1411 days ago
pytest exits 4 on a missing path and 5 on zero tests. My verify hook turned both into PASS....in a table. After the command runs, a hook reads the tool payload and writes PASS or FAIL to a ledger. Green ledger means the agent may mark the task done.. Three rows of that table named test files that had been split into siblings. pytest on a path that doesn't exist exits 4;...r/devopsu/coding-osSep 1411 days ago
Ask HN: Multi-agent workflows in production; Where people using 1000s of agents?...agent swarms in production: what is the main use case / need and what is your biggest pain point right now (state sync, token costs, cascading failures, latency)? [As context: I am a founder at Acyclic Labs (YC F26) and we are building infra to scale agents. Looking to map out...Ask HNramstar3000Sep 1312 days ago
bug(langchain): create_agent ends the loop on an invalid structured-output call when an ordinary tool call is in the same batch (no retry feedback)### Submission checklist - [x] This is a bug, not a usage question. - [x] I added a clear and descriptive title that summarizes this issue. - [x] I used the GitHub search to find a similar question and didn't find it. - [x] I am sure that this is a bug in LangChain rather than...langchain-ai/langchainatifsSep 1213 days ago
[Bug]: MCP tool failures are reported to the agent as successes### Bug Description `McpToolSpec` never propagates an MCP tool failure. `CallToolResult.is_error` (serialized as `isError`) does not become `ToolOutput.is_error`, so a tool the server reported as failed is handed to the agent as a success. `grep -rn "isError\|is_error"` over...run-llama/llama_indexConnorMoss02Sep 33 weeks ago
FIXED: "FAILED_PRECONDITION (code 400): User location is not supported" in Google Antigravity / Gemini Subagents...for the API use While standard approaches like Smart DNS or system-wide VPNs are either too slow or prone to sudden stream drops (which instantly crash long-running subagents), there is a highly stable, production-grade architectural workaround to bypass this restriction...google-gemini/gemini-cliMNshrahiliAug 313 weeks ago
...to a fully local model in that, I can have agents looping 24/7 even when my internet is not working. But this is niche enough that if I had to price the advantages they don't seem worth it. If I'm willing to pay the Openrouter tax, I can fire up Openrouter today and just get...Hacker News commentsKarrot_KreamAug 274 weeks ago

Companies and products named

About this problem

Evidence

9 posts from 9 people in 9 places, about 2 a week over 28 days. Mostly on Ask HN, Hacker News comments, langchain-ai/langchain. Tools named alongside: Model Context Protocol, Claude, LangChain, Anthropic.

Frustration Frustration 0 of 3· Seen on GitHub issues, Hacker News, Bluesky, Reddit

How it was grouped

Posts that state a pain and match the "agent-reliability" rule. First post Aug 27, 2026, latest Sep 23, 2026. Corroborated: 3 or more posts from 2 or more people or places. Method

History

  • 2026-09-25 Added: AI agents are unreliable, loop or stall midway (8 posts)

Rising problems by email

Mondays: the problems in data, tech and AI that grew fastest that week.

Double opt-in. Unsubscribe any time.