FRONTDESK
A Support Agent That Does the Work, With a Human on Refunds
実務をこなすサポートエージェント(返金は人が承認)





Overview
概要Agentic AI for customer support: it looks up orders, cancels, redirects and refunds. The money rules are enforced in code, large refunds wait for a human, and every step is on screen.
カスタマーサポート向けのエージェント型AI。注文の照会、キャンセル、配送先変更、返金まで実行します。お金に関するルールはコードで強制し、高額な返金は人の承認を待ち、すべての手順を画面に表示します。
Support is where most companies first put an LLM agent in front of customers, and the hard part isn't the conversation. It's letting software that can be talked into things touch refunds. The design rule: the model decides what to do; code decides whether it's allowed.
多くの企業が最初にLLMエージェントを顧客の前に出すのはサポート窓口です。難しいのは会話ではなく、言いくるめられる可能性のあるソフトウェアに返金を扱わせることです。設計の原則は、「何をするかはモデルが決め、許されるかどうかはコードが決める」です。
How it was built
開発The model never types a refund amount: it picks the items and the tool prices them. Return windows, final-sale items, double refunds and order ownership are checked inside the tool, where no prompt can argue with them. A refund over £100 stops the run; a supervisor approves or declines, and the agent resumes from saved state and reports the decision honestly.
モデルが返金額を入力することはありません。商品を選ぶのはモデル、金額を計算するのはツールです。返品期限、最終セール品、二重返金、注文の所有者はツール内で検証するため、プロンプトで覆すことはできません。100ポンドを超える返金は処理を止め、担当者が承認・却下すると、保存した状態から再開して結果を正直に伝えます。
What it does
機能Refunds over the limit pause with an approval card; nothing is paid until a person approves.
上限を超える返金は承認カードを出して一時停止。人が承認するまで支払いは行われません。
Final-sale items, return windows and double refunds are refused in code, and the customer gets the rule in plain words.
最終セール品、返品期限、二重返金はコードで拒否し、理由をわかりやすく顧客に伝えます。
A fake 'system notice' asking to refund someone else's order goes nowhere: the session only sees the signed-in customer.
他人の注文の返金を求める偽の「システム通知」は通りません。セッションはログイン中の顧客の情報しか扱えません。
An injury report becomes an urgent ticket for a supervisor, without promising an outcome.
けがの報告は担当者向けの緊急チケットにし、結果を約束しません。
Stack
技術構成Python, FastAPI and SQLite, with Claude tool use through the Anthropic SDK. Nine rule tests run in CI. The screenshots come from a demo mode where the model's turns are scripted; the tools, rules and database changes are real.
Python・FastAPI・SQLite、Anthropic SDKによるClaudeのツール利用。9件のルールテストをCIで実行。スクリーンショットはモデルの発話を台本化したデモモードで撮影しており、ツール・ルール・データベースの変更は実際に動作しています。
The model decides what to do. Code decides whether it's allowed.
何をするかはモデルが決め、許されるかどうかはコードが決める。