RequestResponder.__exit__ leaks CancelledError on cancelled request, killing the stdio receive loop
Nessuno ha ancora preso questa issue.
Valutazione
- Difficoltà
- 3/5
- Tempo stimato
- 1-2 giorni
- Idoneità per principianti
- 68/100
Direzione di ricerca
Inizia da mcp.shared.session.RequestResponder.exit e analizza la gestione della cancellazione intorno a mcp/server/lowlevel/server.py, vicino alla riga 766. Riproduci il problema con tests/test_transport_recovery.py, quindi verifica che la cancellazione di una richiesta stdio produca comunque la relativa risposta di errore e che una richiesta successiva riceva una risposta senza che il ciclo di ricezione si interrompa.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Descrizione
Summary
When a client sends notifications/cancelled for an in-flight request handled over stdio, the server's receive loop task group dies. The process stays alive but stops reading stdin, so every subsequent request hangs and the client eventually reports MCP error -32000: Connection closed. The bug is racy — it reproduces in roughly 40–80% of attempts depending on platform timing.
Affects: mcp 1.26.0 and 1.27.1 (latest at time of writing). Confirmed with FastMCP 3.1.x and the in-tree stdio server.
Reproduction
A standalone subprocess test is in mlorentedev/hive tests/test_transport_recovery.py. The relevant flow:
initialize→ acknotifications/initializedtools/call id=2notifications/cancelledforrequestId=2(within a few ms of step 3)- Receive:
{"id": 2, "error": {"code": 0, "message": "Request cancelled"}}← OK tools/call id=3- No response. Server is alive but
proc.stdout.readline()hangs.
Root cause
mcp.shared.session.RequestResponder.__exit__:
def __exit__(self, exc_type, exc_val, exc_tb):
try:
if self._completed:
self._on_complete(self)
finally:
self._entered = False
...
self._cancel_scope.__exit__(exc_type, exc_val, exc_tb) # ← (A)
When notifications/cancelled arrives, RequestResponder.cancel() calls self._cancel_scope.cancel() and sends the error response. The handler task catches the CancelledError in mcp/server/lowlevel/server.py (around line 766) and returns. The with responder: block then exits with exc_type=None while the cancel scope is still in cancelled state — at line (A), anyio's CancelScope.__exit__ re-raises CancelledError.
That exception bubbles up to:
async with anyio.create_task_group() as tg:
async for message in session.incoming_messages:
tg.start_soon(self._handle_message, ...)
…in Server._run. anyio task groups cancel all sibling tasks and propagate the cancellation. The receive loop is one of those sibling tasks, so it dies.
Suggested fix
Swallow the spurious cancellation when the responder has already sent its response:
def __exit__(self, exc_type, exc_val, exc_tb):
try:
if self._completed:
self._on_complete(self)
finally:
self._entered = False
if not self._cancel_scope:
raise RuntimeError("No active cancel scope")
try:
self._cancel_scope.__exit__(exc_type, exc_val, exc_tb)
except BaseException as exc:
if self._completed and isinstance(exc, anyio.get_cancelled_exc_class()):
# cancel() already sent the error response — the scope's
# re-raised cancellation is spurious.
return
raise
Mirrors what we ship in hive src/hive/_compat.py.
Workaround used downstream
We monkey-patch RequestResponder.__exit__ at import time. The patch is self-gated (only fires when _completed=True AND the leaking exception is anyio.get_cancelled_exc_class()), so it stays inert once a fix lands upstream.
Environment
- Reproduced on Windows 11 (
mcp1.26.0, 1.27.1) and via Claude Code as the host. - Tracked downstream in mlorentedev/hive#75.
- Regression test passes 5/5 with the patch applied on Python 3.12; fails 2/5 — 4/5 without it.
- Python 3.13 appears to have additional uncancel semantics that the patch does not yet fully cover — verification in progress.
- Lingua principale
- Python
- Stelle
- 24.3k
- Fork
- 4k
- Merge medio
- 1g 11h
- PR unite (30g)
- 30
Guida per i contributori
Apri la guida per i contributori
Come iniziare
- Leggi tutta la issue e poi la guida ai contributi del progetto.
- Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
- Fai un fork del repository e lavora su un branch.
- Apri una pull request che faccia riferimento al numero della issue.
Altre issue di modelcontextprotocol/python-sdk
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
modelcontextprotocol/python-sdk#3566 ·
-
v1 v2
Difficoltà 2/5 1-3 ore Idoneità per principianti 85/100
modelcontextprotocol/python-sdk#3546 · 5 commenti ·
-
v1 v2
Difficoltà 2/5 1-3 ore Idoneità per principianti 76/100
modelcontextprotocol/python-sdk#3545 · 1 commento ·
-
v1 v2
Difficoltà 1/5 Meno di un'ora Idoneità per principianti 91/100
modelcontextprotocol/python-sdk#3508 · 2 commenti ·
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 64/100
modelcontextprotocol/python-sdk#3504 ·
Tutte le issue di modelcontextprotocol/python-sdk
Issue simili
-
bug confirmed issue
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
open-webui/open-webui#30750 · 1 commento ·
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
-
enhancement
Difficoltà 2/5 1-3 ore Idoneità per principianti 75/100
OpenwaterHealth/openmotion-bloodflow-app#604 · 1 commento ·
-
Difficoltà 2/5 1-3 ore Idoneità per principianti 70/100
-
good first issue
Difficoltà 1/5 Meno di un'ora Idoneità per principianti 90/100