Уроборос в какой-то момент хочет отправить 195226 tokens и больше ничего не может делать (все запросы падают).
Нужен какой-то ручной механизм, как подрезать ему контекст. И желательно диагностику: понимать, что именно у него в контексте, чтобы знать, куда копать.
{
"ts": "2026-04-14T01:57:42.223289+00:00",
"session_id": "ab70d9dec22f49df82d2b7e7e6d1027c",
"direction": "out",
"chat_id": 1,
"user_id": 1,
"text": "⚠️ All models are down. Primary (google/gemma-4-31b-it:free) and fallback (openai/gpt-oss-120b) both returned no response. Stopping. Last provider error: BadRequestError('Error code: 400 - {\'error\': {\'message\': "This endpoint\'s maximum context length is 131072 tokens. However, you requested about 195226 tokens (174258 of text input, 4584 of tool input, 16384 in th... Background consciousness will attempt recovery when the provider is back.",
"format": "markdown",
"source": "",
"sender_label": "",
"sender_session_id": "",
"client_message_id": "",
"telegram_chat_id": 0,
"task_id": "18750a78"
}
Уроборос в какой-то момент хочет отправить 195226 tokens и больше ничего не может делать (все запросы падают).
Нужен какой-то ручной механизм, как подрезать ему контекст. И желательно диагностику: понимать, что именно у него в контексте, чтобы знать, куда копать.
{⚠️ All models are down. Primary (google/gemma-4-31b-it:free) and fallback (openai/gpt-oss-120b) both returned no response. Stopping. Last provider error: BadRequestError('Error code: 400 - {\'error\': {\'message\': "This endpoint\'s maximum context length is 131072 tokens. However, you requested about 195226 tokens (174258 of text input, 4584 of tool input, 16384 in th... Background consciousness will attempt recovery when the provider is back.",
"ts": "2026-04-14T01:57:42.223289+00:00",
"session_id": "ab70d9dec22f49df82d2b7e7e6d1027c",
"direction": "out",
"chat_id": 1,
"user_id": 1,
"text": "
"format": "markdown",
"source": "",
"sender_label": "",
"sender_session_id": "",
"client_message_id": "",
"telegram_chat_id": 0,
"task_id": "18750a78"
}