QuickFix for FinalStreamingStdOutCallbackHandler: Ignore new lines & white spaces (#5497)

# Make FinalStreamingStdOutCallbackHandler more robust by ignoring new lines & white spaces `FinalStreamingStdOutCallbackHandler` doesn't work out of the box with `ChatOpenAI`, as it tokenized slightly differently than `OpenAI`. The response of `OpenAI` contains the tokens `["\nFinal", " Answer", ":"]` while `ChatOpenAI` contains `["Final", " Answer", ":"]`. This PR make `FinalStreamingStdOutCallbackHandler` more robust by ignoring new lines & white spaces when determining if the answer prefix has been reached. Fixes #5433 ## Who can review? Community members can review the PR once tests pass. Tag maintainers/contributors who might be interested: Tracing / Callbacks - @agola11 Twitter: [@UmerHAdil](https://twitter.com/@UmerHAdil) | Discord: RicChilligerDude#7589
2025-09-04 20:46:45 +00:00 · 2023-06-04 00:05:58 +02:00
parent 1f4abb265a
commit 44ad9628c9
2 changed files with 109 additions and 16 deletions
--- a/docs/modules/agents/streaming_stdout_final_only.ipynb
+++ b/docs/modules/agents/streaming_stdout_final_only.ipynb
@@ -1,6 +1,7 @@
 {
 "cells": [
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "23234b50-e6c6-4c87-9f97-259c15f36894",
   "metadata": {
@@ -11,6 +12,7 @@
   ]
  },
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "29dd6333-307c-43df-b848-65001c01733b",
   "metadata": {},
@@ -36,6 +38,7 @@
   ]
  },
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "19a813f7",
   "metadata": {},
@@ -84,6 +87,7 @@
   ]
  },
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "53a743b8",
   "metadata": {},
@@ -92,11 +96,12 @@
   ]
  },
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "23602c62",
   "metadata": {},
   "source": [
-    "By default, we assume that the token sequence ``\"\\nFinal\", \" Answer\", \":\"`` indicates that the agent has reached an answers. We can, however, also pass a custom sequence to use as answer prefix."
+    "By default, we assume that the token sequence ``\"Final\", \"Answer\", \":\"`` indicates that the agent has reached an answers. We can, however, also pass a custom sequence to use as answer prefix."
   ]
  },
  {
@@ -108,26 +113,75 @@
   "source": [
    "llm = OpenAI(\n",
    "    streaming=True,\n",
-    "    callbacks=[FinalStreamingStdOutCallbackHandler(answer_prefix_tokens=[\"\\nThe\", \" answer\", \":\"])],\n",
+    "    callbacks=[FinalStreamingStdOutCallbackHandler(answer_prefix_tokens=[\"The\", \"answer\", \":\"])],\n",
    "    temperature=0\n",
    ")"
   ]
  },
  {
+   "attachments": {},
   "cell_type": "markdown",
   "id": "b1a96cc0",
   "metadata": {},
   "source": [
-    "Be aware you likely need to include whitespaces and new line characters in your token. "
+    "For convenience, the callback automatically strips whitespaces and new line characters when comparing to `answer_prefix_tokens`. I.e., if `answer_prefix_tokens = [\"The\", \" answer\", \":\"]` then both `[\"\\nThe\", \" answer\", \":\"]` and `[\"The\", \" answer\", \":\"]` would be recognized a the answer prefix."
+   ]
+  },
+  {
+   "attachments": {},
+   "cell_type": "markdown",
+   "id": "9278b522",
+   "metadata": {},
+   "source": [
+    "If you don't know the tokenized version of your answer prefix, you can determine it with the following code:"
   ]
  },
  {
   "cell_type": "code",
   "execution_count": null,
-   "id": "9278b522",
+   "id": "2f8f0640",
   "metadata": {},
   "outputs": [],
-   "source": []
+   "source": [
+    "from langchain.callbacks.base import BaseCallbackHandler\n",
+    "\n",
+    "class MyCallbackHandler(BaseCallbackHandler):\n",
+    "    def on_llm_new_token(self, token, **kwargs) -> None:\n",
+    "        # print every token on a new line\n",
+    "        print(f\"#{token}#\")\n",
+    "\n",
+    "llm = OpenAI(streaming=True, callbacks=[MyCallbackHandler()])\n",
+    "tools = load_tools([\"wikipedia\", \"llm-math\"], llm=llm)\n",
+    "agent = initialize_agent(tools, llm, agent=AgentType.ZERO_SHOT_REACT_DESCRIPTION, verbose=False)\n",
+    "agent.run(\"It's 2023 now. How many years ago did Konrad Adenauer become Chancellor of Germany.\")"
+   ]
+  },
+  {
+   "attachments": {},
+   "cell_type": "markdown",
+   "id": "61190e58",
+   "metadata": {},
+   "source": [
+    "### Also streaming the answer prefixes"
+   ]
+  },
+  {
+   "attachments": {},
+   "cell_type": "markdown",
+   "id": "1255776f",
+   "metadata": {},
+   "source": [
+    "When the parameter `stream_prefix = True` is set, the answer prefix itself will also be streamed. This can be useful when the answer prefix itself is part of the answer. For example, when your answer is a JSON like\n",
+    "\n",
+    "`\n",
+    "{\n",
+    "    \"action\": \"Final answer\",\n",
+    "    \"action_input\": \"Konrad Adenauer became Chancellor 74 years ago.\"\n",
+    "}\n",
+    "`\n",
+    "\n",
+    "and you don't only want the action_input to be streamed, but the entire JSON."
+   ]
  }
 ],
 "metadata": {