Um servidor Model Context Protocol somente-leitura que dá ao Claude acesso delimitado à sua conta do Gong: descoberta de calls, definições de trackers, sinais analisados por call, estatísticas de interação por rep e uma ferramenta derivada que reporta quais trackers de risco dispararam em quais calls — separando se foi o cliente que disse ou o seu próprio rep. O scaffold está no bundle de artefatos em apps/web/public/artifacts/mcp-server-gong-revops/, que inclui README.md, pyproject.toml e src/gong_revops_mcp/server.py, instalável com pip install -e ..
Comece pelo que a API não tem, porque isso determina o formato de todo o resto. A API pública do Gong não expõe nenhum endpoint de leitura para os dados do deal board. Os endpoints de CRM (GET /v2/crm/entities) retornam apenas objetos que você subiu antes por uma integração de CRM genérica registrada, e a documentação do Gong marca esse endpoint como verificação em fase de desenvolvimento. Então um servidor que promete “pergunte ao Claude sobre seus deals do Gong” está fazendo uma de três coisas: envolvendo a UI, lendo seu CRM no lugar, ou adivinhando. Este deriva o risco do deal das conversas e diz isso: deal_risk_digest retorna os hits de trackers com uma nota mandando você juntar call_id ao seu CRM para stage, valor e data de fechamento.
Quando usar
Recorra a ele quando uma pergunta recorrente de RevOps custa dez minutos de cliques para uma pessoa: quais contas reclamaram de preço na semana passada, em quais calls um concorrente foi citado, se os reps de um segmento em dificuldade estão monologando. Esses são joins entre os próprios dados do Gong que a UI te obriga a fazer no olho. Dois papéis extraem mais valor. O líder de RevOps que roda um pipeline review semanal pergunta em linguagem natural e cola uma resposta estruturada no deck. O GTM engineer que escreve um script descartável contra /v2/calls/extensive a cada nova pergunta já tem o contentSelector, a paginação por cursor, o rate limiter e a atribuição de falante prontos.
É também o padrão certo se você já roda o servidor MCP do Salesforce para RevOps ou o do Clari e quer a camada de conversa na mesma superfície de chat, para que uma pergunta atravesse de “o que o cliente disse” para “em qual stage está” sem trocar de aba. Essa travessia é o ganho real — nenhum dos dois sistemas responde isso sozinho.
Quando NÃO usar
O Gong já publica um servidor MCP oficial. O Gong anunciou suporte a MCP em 2026 e documenta um servidor MCP hospedado pelo próprio Gong, disponível em qualquer plano do Gong, configurado por um administrador técnico e com acesso governado pelo nível de assento. Ele permite que Claude, ChatGPT e Microsoft Copilot perguntem sobre contas e deals e puxem os insights gerados pela IA do Gong. Teste primeiro. É de primeira parte, não exige hospedar processo nenhum e respeita as permissões por nível de assento do Gong, o que este scaffold não faz. Construa a versão auto-hospedada quando você precisar de uma superfície de ferramentas fixa e auditável — um contentSelector que você controla, um kill-switch de transcrições, um limite de páginas, saída de trackers com atribuição de falante — ou quando seu administrador não habilitar o servidor hospedado.
Você não consegue um administrador para gerar uma API key. As credenciais saem de Company Settings → Ecosystem → API e só um administrador técnico pode criá-las. Não existe key por usuário.
Visibilidade de calls por usuário é determinante para você. Uma única key no nível da conta vê todas as calls dos workspaces que ela cobre, não importa qual pessoa esteja conversando. Se sua instância do Gong restringe quem pode ouvir as calls de quem, este servidor contorna isso. Rode um por analista com keys de escopo estreito, ou não rode.
Você quer transcrições literais no modelo por padrão. Aqui elas estão desligadas, e o design assume que isso é o certo. Se seu workflow é transcrição-primeiro, você vai brigar com o scaffold.
Uma ou duas perguntas por mês. Os filtros da própria UI do Gong são mais rápidos que um setup que você precisa manter.
O que ele expõe
Seis ferramentas de leitura, nenhuma de escrita. A superfície de escrita da API pública é upload de calls e upload de objetos de CRM genérico; nenhuma das duas cabe atrás de um prompt de chat, e somente-leitura elimina toda a classe de falha do tipo “o modelo me entendeu errado e mudou o sistema de registro”.
find_calls — GET /v2/calls. Só metadata: id, título, início, duração, direção, URL do Gong. Delimite a pergunta aqui primeiro.
list_trackers — GET /v2/settings/trackers. Apenas definições de trackers. O Gong não retorna contagens de correspondência por esse endpoint, o que surpreende as pessoas; as contagens de ocorrências vêm do endpoint extensive de calls.
call_signals — POST /v2/calls/extensive. O cavalo de batalha: participantes, correspondências de trackers, ocorrências de trackers, brief do Spotlight, key points, resultado automático da call, tópicos, tempo de fala, estatísticas de interação por pessoa, comentários públicos.
call_transcript — POST /v2/calls/transcript. Desligado a menos que GONG_ALLOW_TRANSCRIPTS=true, limitado a 3 calls, exige uma justificativa.
rep_interaction_stats — POST /v2/stats/interaction. Monólogo mais longo, história de cliente mais longa, interatividade, paciência, taxa de perguntas.
deal_risk_digest — derivada. Junta definições de trackers às ocorrências em um intervalo de datas e classifica cada hit como customer, internal ou unattributed.
Postura de engenharia
A atribuição de falante é o sentido inteiro do digest. “Pricing Pushback” dito pelo seu próprio rep é um sinal de comportamento do rep. Dito pelo cliente, é um sinal do deal. Uma contagem de trackers que soma os dois se move pelos motivos errados e produz um número de risco no qual ninguém consegue agir. deal_risk_digest lê content.trackerOccurrences, procura cada speakerId no array parties da call e separa por afiliação da parte. É por isso que o servidor pede ocorrências e não só contagens, e é a única coisa que um wrapper genérico do Gong não vai fazer para você.
Mídia nunca é solicitada. O contentSelector em server.py é fixo, não controlado por quem chama, e omite media. A key não carrega api:calls:read:media-url. Assim o servidor nunca gera os links assinados de áudio/vídeo de 8 horas do Gong — um link que sobrevive à conversa em que apareceu é um vazamento esperando um print.
Transcrições são um kill-switch, não um prompt.call_transcript checa uma variável de ambiente antes de rodar e limita a três calls. Confiar só numa string de justificativa deixa a fala literal do cliente a uma leitura errada confiante da janela de contexto. A flag transforma “a gente permite isso, afinal?” em uma decisão de deploy em vez de uma decisão por pergunta.
A paginação é limitada e o limite é reportado.GONG_MAX_PAGES vem em 5 por padrão, então uma chamada de ferramenta lê no máximo 500 registros e retorna truncated: true quando parou antes. Um modelo que vê metade dos dados em silêncio responde com confiança a pergunta errada.
As requisições são serializadas, não retentadas. O Gong limita a 3 requisições por segundo e 10.000 requisições por dia por padrão, retornando 429 com um header Retry-After. O scaffold espera 0,34s entre requisições em vez de disparar em paralelo e reagir aos 429, porque uma tempestade de retries reativos gasta cota diária em requisições que iam falhar de qualquer jeito.
Realidade de custos
Três linhas, mais uma que não é linha.
Assinatura do Claude. O que você já paga — Pro a $20/usuário/mês, Max a $100–200/usuário/mês, ou consumo de API. O servidor não muda nada aqui.
Auto-hospedagem. Um processo Python local por usuário do Claude Desktop: custo zero de infraestrutura. Como serviço compartilhado, uma VM pequena a uns $20–50/mês em qualquer cloud.
Cota da API do Gong. Grátis com seu contrato do Gong, não medida em dólares, mas finita: 3 requisições por segundo e 10.000 requisições por dia por empresa por padrão, ampliáveis falando com o suporte do Gong. Faça o orçamento. Um deal_risk_digest sobre 90 dias em um workspace com 4.000 calls são 40 páginas de 100 = 40 requisições. Dez perguntas assim por dia são 400 requisições, confortavelmente dentro do teto. O que estoura o orçamento é um loop de cursor sem limite, que é exatamente o que GONG_MAX_PAGES existe para evitar.
Assentos do Gong. O Gong não publica preço de tabela; é cotado por assento com uma taxa de plataforma. O que você paga não muda com este servidor — ele não adiciona assentos.
O custo em tokens é dominado pelo payload das respostas, e é por isso que server.py enxuga cada resposta antes de devolvê-la. call_signals sobre 20 calls retorna briefs e key points em vez de conteúdo completo e cai nas dezenas baixas de milhares de tokens. Um call_transcript de uma call de 45 minutos é comparável por si só, e esse é o argumento real para deixar as transcrições desligadas.
Frente às alternativas
O servidor MCP oficial do Gong. Coberto acima: teste primeiro. De primeira parte, qualquer plano, permissões por nível de assento, nada para hospedar. Escolha o scaffold auto-hospedado quando você precisar de uma superfície de ferramentas que caiba em um arquivo e possa ser fixada, ou quando o servidor hospedado não estiver habilitado para você.
Um servidor MCP do Gong da comunidade. Existem vários no GitHub e em diretórios de MCP, a maioria envolvendo calls e transcrições. Mais rápido de instalar do que ler este scaffold. O contraponto é que “envolve calls e transcrições” costuma significar transcrições ligadas por padrão, sem limite de páginas e contagens de trackers sem atribuição de falante — as três decisões que este scaffold toma diferente de propósito.
Um script descartável contra /v2/calls/extensive. Controle máximo, e cada time reconstrói na mão a autenticação Basic, a base URL específica da conta, a paginação por cursor, o rate limiter e o join com parties. Este scaffold são umas 450 linhas com tudo isso já conectado.
A própria UI do Gong e o Spotlight. Mais rápido para uma call só e os dados já estão lá. Não consegue juntar dados do Gong ao resto do seu contexto no Claude, que é a única razão para rodar qualquer uma dessas coisas. Se você não tem certeza se um servidor MCP ou um Skill é o formato certo para o seu problema, leia Claude Skill frente a servidor MCP.
Pontos de atenção
O README documenta os sete; os cinco determinantes:
Uma base URL errada retorna 401, não 404. A base URL da API do Gong é específica da conta e https://api.gong.io é um padrão comum, não universal. Times perdem uma tarde depurando credenciais que estavam corretas. Guarda: _raise_for_gong intercepta o 401 e nomeia a base URL que realmente usou, listando o desencontro de base URL como primeira causa antes das credenciais.
Um tracker renomeado se lê como boa notícia.deal_risk_digest compara nomes de trackers exatamente, então um tracker renomeado no Gong para de corresponder e o digest reporta risco zero. Guarda: parcial — rode list_trackers primeiro e cole os nomes reais em GONG_RISK_TRACKERS; os padrões que vêm são placeholders que não correspondem a nada na maioria dos workspaces. Emitir um aviso quando um nome configurado não corresponde a nenhum tracker vivo é o TODO #3 do README.
As estatísticas de interação punem baixo volume de calls. As estatísticas do Gong derivam apenas de calls com Whisper habilitado, então um rep com três calls gravadas é estatisticamente indistinguível de um rep com um problema real. Guarda: rep_interaction_stats devolve esse alerta embutido em cada resposta, para o modelo repetir em vez de dar coaching sobre ruído; junte as contagens de find_calls antes de mostrar os números a um gestor.
Truncamento silencioso. Um loop de cursor interrompido no limite de páginas parece idêntico a uma resposta completa. Guarda: toda ferramenta paginada retorna truncated: true quando parou antes, e find_calls é o jeito barato de checar volume antes de fazer uma pergunta cara.
Deriva de consentimento. Um cliente que consentiu ser gravado não consentiu com isso ser resumido por um modelo de terceiros. Guarda: transcrições ficam desligadas por padrão e URLs de mídia nunca são geradas; revise seu DPA antes de ligar GONG_ALLOW_TRANSCRIPTS.
Stack
Gong — conversation intelligence, trackers, briefs do Spotlight, estatísticas de interação
MCP Python SDK — mcp>=1.2.0; fornece Server, stdio_server e os decoradores do registro de ferramentas
httpx — cliente REST async contra o host da API do Gong da sua conta, autenticação Basic com base64("key:secret")
Claude Desktop ou Claude Code — interface de linguagem natural e chamador de ferramentas
GONG_ALLOW_TRANSCRIPTS — a trava no nível de ambiente que decide se a fala literal do cliente chega ao modelo
GONG_MAX_PAGES — a guarda de cota que também deixa visível ao modelo que a resposta está incompleta
# mcp-server-gong-revops
A read-only MCP server over the Gong public API v2, tuned for RevOps questions that currently require a human to open Gong, filter a call list, read four calls, and write down what they saw. Exposes call discovery, tracker definitions, per-call analyzed signals (tracker matches with speaker attribution, Spotlight brief, key points, call outcome, talk-ratio stats), per-rep interaction stats, and one derived tool — `deal_risk_digest` — that joins tracker definitions to tracker occurrences and splits them by who actually said the thing.
> **STATUS: scaffold — not runtime-tested.** The code follows the official `mcp` Python SDK conventions and the endpoint paths, scopes, and field names track the public Gong API docs (help.gong.io/apidocs) as of 2026-07. It has not been executed against a live Gong account. Response field names in particular vary by account configuration — verify before you rely on it.
## Two things this server refuses to do
**No `get_deals` tool.** Gong's public API has no native read endpoint for deal-board data. The CRM endpoints (`GET /v2/crm/entities`) only read back objects you previously *uploaded* through a registered generic CRM integration, and Gong's own documentation marks that endpoint as development-phase verification only. Any MCP server advertising "ask Claude about your Gong deals" is either wrapping the UI, reading your CRM, or inventing the answer. `deal_risk_digest` is the honest substitute: it derives risk signals from conversations and tells you to join to the CRM for stage and amount.
**No writes.** The public API's write surface is call upload and generic-CRM object upload. Neither belongs behind a chat prompt, and read-only removes the entire class of "the model misread me and mutated the system of record" failure. If you need writes later, add them as separately-named tools with mandatory justification strings — never as a free-text command.
## What it exposes
- `find_calls(fromDateTime, toDateTime?, workspace_id?)` — `GET /v2/calls`. Cheap metadata: id, title, start, duration, direction, Gong URL. Scope a question here first, then pass ids to `call_signals`. Follows at most `GONG_MAX_PAGES` cursor pages of 100.
- `list_trackers(workspace_id?)` — `GET /v2/settings/trackers`. Tracker **definitions only** — ids, names, keywords, affiliation. No match counts; Gong does not return occurrence statistics from this endpoint. Call it to learn what your workspace actually tracks before guessing a tracker name in a question.
- `call_signals(call_ids? | fromDateTime, toDateTime?, workspace_id?)` — `POST /v2/calls/extensive` with a fixed `contentSelector`: parties, tracker matches, tracker occurrences, Spotlight brief, key points, auto call outcome, topics, speaker talk time, per-person interaction stats, public comments. The workhorse tool.
- `call_transcript(call_ids, justification)` — `POST /v2/calls/transcript`. Verbatim monologues with speaker id and millisecond offsets. Disabled unless `GONG_ALLOW_TRANSCRIPTS=true`, capped at `GONG_MAX_TRANSCRIPT_CALLS` (default 3), and requires a justification of at least 10 characters.
- `rep_interaction_stats(fromDate, toDate, user_ids?)` — `POST /v2/stats/interaction`. Longest monologue, longest customer story, interactivity, patience, question rate, per rep.
- `deal_risk_digest(fromDateTime, toDateTime?, tracker_names?, workspace_id?)` — derived. Scans calls in the range, keeps only occurrences of the trackers named in `GONG_RISK_TRACKERS`, and reports each hit split into `customer` / `internal` / `unattributed` by the speaker's party affiliation.
## Setup
### 1. Install
```bash
git clone <wherever you put this>
cd mcp-server-gong-revops
python -m venv .venv
source .venv/bin/activate # or .venv\Scripts\activate on Windows
pip install -e .
```
### 2. Generate Gong API credentials
A **technical administrator** creates these — a standard user seat cannot. In Gong: **Company Settings → Ecosystem → API**, then generate an Access Key and Access Key Secret. Copy the secret immediately; Gong shows it once.
The same page displays **your account's base URL**. Copy it. `https://api.gong.io` is the common value but not a universal one — accounts on regional or dedicated hosts get a different origin, and a wrong base URL returns **401, not 404**, which sends people debugging a credential problem they do not have.
### 3. Grant scopes
Scopes are attached to the key by the administrator who creates it. This server needs five:
| Scope | Used by |
|---|---|
| `api:calls:read:basic` | `find_calls` |
| `api:calls:read:extensive` | `call_signals`, `deal_risk_digest` |
| `api:calls:read:transcript` | `call_transcript` |
| `api:settings:trackers:read` | `list_trackers` |
| `api:stats:interaction` | `rep_interaction_stats` |
Grant only what you intend to use. Omitting `api:calls:read:transcript` is a second, key-level lock on transcripts on top of `GONG_ALLOW_TRANSCRIPTS`. Note what is deliberately **absent**: `api:calls:read:media-url`. The server never requests media URLs, so it never mints the 8-hour signed audio/video links that would otherwise outlive the conversation they appeared in.
### 4. Configure environment
```bash
export GONG_ACCESS_KEY="your-access-key"
export GONG_ACCESS_KEY_SECRET="your-access-key-secret"
export GONG_BASE_URL="https://api.gong.io" # COPY YOURS from Company Settings -> API
export GONG_ALLOW_TRANSCRIPTS="false" # true enables call_transcript
export GONG_MAX_TRANSCRIPT_CALLS="3" # cap per transcript call
export GONG_MAX_PAGES="5" # cursor pages followed per tool call
export GONG_WORKSPACE_ID="" # optional default workspace
export GONG_MIN_REQUEST_INTERVAL="0.34" # seconds between requests (3/s limit)
export GONG_RISK_TRACKERS="Pricing Pushback,Competitor Mention,Budget Freeze,Legal Review,Champion Left"
```
Env var notes:
- **`GONG_ACCESS_KEY` / `GONG_ACCESS_KEY_SECRET`** — from Company Settings → Ecosystem → API. Combined as `base64("key:secret")` and sent as `Authorization: Basic <token>`. If you register this as a Gong OAuth app instead, replace `auth_headers()` with a `Bearer` token.
- **`GONG_BASE_URL`** — account-specific. Copy it rather than trusting the default. This is the single most common setup failure.
- **`GONG_ALLOW_TRANSCRIPTS`** — the PII kill-switch. Transcripts put full verbatim customer speech into model context. Off by default; flip it only after someone has decided that is allowed for this data.
- **`GONG_MAX_TRANSCRIPT_CALLS`** — blast-radius cap. Three transcripts is already a large prompt. Raise it deliberately, never to "just get the analysis done."
- **`GONG_MAX_PAGES`** — the quota guard. Gong pages at 100 records and allows 10,000 requests/day by default; an unbounded cursor loop over a busy workspace can spend a real share of that answering one question. 5 pages = up to 500 records per tool call, and the response reports `truncated: true` so the model knows it did not see everything.
- **`GONG_RISK_TRACKERS`** — which tracker names count as risk. Gong ships no "this tracker means risk" flag, so this is a judgment your team makes. Replace the defaults with your actual tracker names from `list_trackers` — the defaults are placeholders and will match nothing in most workspaces.
- **`GONG_MIN_REQUEST_INTERVAL`** — requests are serialized behind this interval to stay under 3/second. Reactive 429 retries still burn daily quota on requests that were always going to fail.
### 5. Register with Claude
`claude_desktop_config.json` (macOS: `~/Library/Application Support/Claude/`, Windows: `%APPDATA%\Claude\`):
```json
{
"mcpServers": {
"gong-revops": {
"command": "/absolute/path/to/mcp-server-gong-revops/.venv/bin/python",
"args": ["-m", "gong_revops_mcp.server"],
"env": {
"GONG_ACCESS_KEY": "your-access-key",
"GONG_ACCESS_KEY_SECRET": "your-access-key-secret",
"GONG_BASE_URL": "https://api.gong.io",
"GONG_ALLOW_TRANSCRIPTS": "false",
"GONG_MAX_PAGES": "5",
"GONG_RISK_TRACKERS": "Pricing Pushback,Competitor Mention,Legal Review"
}
}
}
}
```
For Claude Code, the same block goes in `.mcp.json` at the project root. Restart the client after editing.
### 6. Sanity check
Run these three in order. Each one isolates a different failure.
1. **"List the Gong trackers in my workspace."** → exercises auth, base URL, and `api:settings:trackers:read` on the cheapest possible request. A 401 here means base URL or credentials; a 403 means scopes. Copy the real tracker names out of the response into `GONG_RISK_TRACKERS`.
2. **"Find Gong calls from the last 7 days."** → exercises `GET /v2/calls` and cursor pagination. If `truncated` comes back `true`, your workspace has more than `GONG_MAX_PAGES × 100` calls in a week; narrow the range in real questions.
3. **"Pull the signals for the three most recent of those calls and tell me which trackers the customer raised."** → exercises `/v2/calls/extensive`, the fixed `contentSelector`, and speaker attribution. If tracker occurrences come back empty while counts are non-zero, your account does not expose `content.trackerOccurrences` and `deal_risk_digest` will report everything as `unattributed`.
## Security model
- **Token scope.** One account-level API key with five read scopes. It is not per-user: the key sees every call in the workspaces it covers, regardless of which human is chatting. Anyone who can talk to this MCP server can read any recorded call. If your Gong instance relies on per-user visibility rules, this server bypasses them — run it per-analyst with narrowly-scoped keys, or do not run it.
- **What leaves Gong.** Call metadata, party names/emails/titles, tracker matches, Spotlight briefs, key points, topics, and interaction stats go into the model context on every `call_signals` call. Verbatim customer speech goes only through `call_transcript`, which is off by default.
- **What never leaves.** Audio and video. The server does not request the `media` field and does not hold `api:calls:read:media-url`, so no signed recording links are minted.
- **Recording consent is upstream.** This server inherits whatever consent posture your Gong instance already has. It does not create a new consent question, but it does widen who can read the result — a recording a customer consented to being *recorded* is not automatically one they consented to being *summarized by a third-party model*. Check your DPA before enabling transcripts.
## Known limits — numbered TODO list before production use
1. **Not runtime-tested.** Every response-slimming function assumes field names from the docs (`metaData.id`, `content.trackers[].occurrences[].speakerId`, `usersAggregateActivity`). Run each tool once against a real account and fix the shapes before trusting output.
2. **No retry with backoff.** A 429 raises with the `Retry-After` value in the message instead of sleeping and retrying. Fine for interactive chat, wrong for unattended use.
3. **`deal_risk_digest` matches tracker names case-insensitively and exactly.** A renamed tracker silently stops matching and the digest reports zero risk — which reads as good news. Add a warning when a configured name matches no tracker returned by `list_trackers`.
4. **No caching.** Asking the same question twice spends the quota twice. A short-lived cache keyed on the filter would cut the common repeat-question cost.
5. **`rep_interaction_stats` has no call-count denominator.** Gong's stats derive only from calls with Whisper enabled, so a rep with three recorded calls looks statistically identical to a rep with a real problem. Join `find_calls` counts before showing these numbers to a manager.
6. **Single workspace assumption in the digest.** `deal_risk_digest` accepts one `workspace_id`; multi-workspace accounts need one call per workspace and a merge step.
7. **Account name comes from party emails, not the CRM.** `external_parties` is a list of names/emails, not a resolved account. Joining on email domain is the usual fix and it is not implemented here.
"""
gong-revops-mcp — read-only MCP server over the Gong public API v2.
Exposes call discovery, tracker definitions, extensive per-call signals (trackers,
tracker occurrences, Spotlight brief, key points, call outcome, per-person interaction
stats), rep interaction stats, and a derived deal-risk digest that joins tracker
definitions to tracker occurrences across a date range.
Two things this server deliberately does NOT do:
1. There is no `get_deals` tool. Gong's public API has no native read endpoint for
deal-board data. `GET /v2/crm/entities` only reads back objects you previously
uploaded through a registered generic CRM integration, and Gong's own docs mark it
as development-phase verification only. Deal risk here is DERIVED from calls plus
trackers (see `deal_risk_digest`); for authoritative deal fields, query the CRM's
own API instead.
2. There are no writes. The public API's write surface is call upload and generic-CRM
object upload, neither of which belongs behind a chat prompt. Read-only removes the
whole class of "the model misread me and changed the system of record" failure.
STATUS: scaffold — not runtime-tested. Endpoint paths, scopes, and field names track
the public Gong API docs (help.gong.io/apidocs) as of 2026-07; verify against your
account before relying on it. Your base URL is account-specific — see README.
Run as: python -m gong_revops_mcp.server
"""
from __future__ import annotations
import asyncio
import base64
import os
import time
from typing import Any
import httpx
from mcp.server import Server
from mcp.server.stdio import stdio_server
from mcp.types import TextContent, Tool
# ----- Configuration (read from env at startup) -----
GONG_ACCESS_KEY = os.environ.get("GONG_ACCESS_KEY")
GONG_ACCESS_KEY_SECRET = os.environ.get("GONG_ACCESS_KEY_SECRET")
# The Gong API base URL is ACCOUNT-SPECIFIC. api.gong.io is the common default, but
# accounts on regional or dedicated hosts get a different origin. Find yours at
# Company Settings -> API (see README) — a wrong base URL presents as 401, not 404,
# which sends people hunting for a credential problem they do not have.
GONG_BASE_URL = os.environ.get("GONG_BASE_URL", "https://api.gong.io").rstrip("/")
# Transcripts are the highest-PII and highest-token surface in the API: full verbatim
# customer speech, thousands of tokens per call. Off unless explicitly opted in.
GONG_ALLOW_TRANSCRIPTS = os.environ.get("GONG_ALLOW_TRANSCRIPTS", "false").lower() == "true"
# Blast-radius cap on transcript pulls. Three calls of transcript is already a large
# prompt; a date-range transcript pull across a team is how you blow a context window
# and a day's API quota in one question.
GONG_MAX_TRANSCRIPT_CALLS = int(os.environ.get("GONG_MAX_TRANSCRIPT_CALLS", "3"))
# Cursor-following guard. Gong pages at 100 records and allows 10,000 calls/day; an
# unbounded cursor loop over a busy workspace can consume a meaningful share of that
# quota answering one question. Five pages = up to 500 records per tool call.
GONG_MAX_PAGES = int(os.environ.get("GONG_MAX_PAGES", "5"))
# Optional default workspace, so callers do not pass workspaceId on every query.
GONG_WORKSPACE_ID = os.environ.get("GONG_WORKSPACE_ID")
# Tracker names that count as risk signals for deal_risk_digest. Gong ships no
# "this tracker means risk" flag — which trackers are risk is a judgment call your
# team makes, so it is configuration, not a hardcoded list.
GONG_RISK_TRACKERS = [
t.strip()
for t in os.environ.get(
"GONG_RISK_TRACKERS",
"Pricing Pushback,Competitor Mention,Budget Freeze,Legal Review,Champion Left",
).split(",")
if t.strip()
]
# Gong throttles at 3 requests/second. We serialize requests behind a minimum
# interval rather than firing concurrently and handling 429s reactively — a reactive
# retry storm still burns daily quota on requests that were always going to fail.
MIN_REQUEST_INTERVAL = float(os.environ.get("GONG_MIN_REQUEST_INTERVAL", "0.34"))
PAGE_SIZE = 100
_rate_lock = asyncio.Lock()
_last_request_at = 0.0
def require_config() -> None:
missing = [
name
for name, value in (
("GONG_ACCESS_KEY", GONG_ACCESS_KEY),
("GONG_ACCESS_KEY_SECRET", GONG_ACCESS_KEY_SECRET),
)
if not value
]
if missing:
raise RuntimeError(f"Required env vars are unset: {', '.join(missing)}")
def auth_headers() -> dict[str, str]:
# Gong's API-key method is HTTP Basic with base64("<access key>:<secret>").
# OAuth apps use "Authorization: Bearer <token>" instead; swap this function if
# you register the server as a Gong app rather than using an account API key.
token = base64.b64encode(
f"{GONG_ACCESS_KEY}:{GONG_ACCESS_KEY_SECRET}".encode()
).decode()
return {
"Authorization": f"Basic {token}",
"Content-Type": "application/json",
}
# ----- Gong REST helpers -----
async def _throttle() -> None:
global _last_request_at
async with _rate_lock:
wait = MIN_REQUEST_INTERVAL - (time.monotonic() - _last_request_at)
if wait > 0:
await asyncio.sleep(wait)
_last_request_at = time.monotonic()
async def gong_request(
method: str, path: str, *, params: dict[str, Any] | None = None, json: dict[str, Any] | None = None
) -> dict[str, Any]:
await _throttle()
async with httpx.AsyncClient(timeout=60.0) as client:
r = await client.request(
method, f"{GONG_BASE_URL}{path}", headers=auth_headers(), params=params, json=json
)
_raise_for_gong(r, path)
return r.json() if r.content else {}
def _raise_for_gong(r: httpx.Response, path: str) -> None:
if r.status_code == 401:
raise PermissionError(
"Gong returned 401. Two causes, in order of likelihood: (1) GONG_BASE_URL "
f"is wrong for this account — {GONG_BASE_URL} is a guess unless you copied it "
"from Company Settings -> API; (2) the access key/secret pair is wrong or "
"revoked. A wrong base URL does NOT return 404."
)
if r.status_code == 403:
raise PermissionError(
f"Gong returned 403 on {path}. The API key is missing a scope. This server "
"needs api:calls:read:basic, api:calls:read:extensive, "
"api:calls:read:transcript, api:settings:trackers:read, and "
"api:stats:interaction. Scopes are set per key by a technical administrator."
)
if r.status_code == 429:
retry_after = r.headers.get("Retry-After", "unknown")
raise RuntimeError(
f"Gong returned 429 (rate limit; Retry-After={retry_after}s). Default limits "
"are 3 requests/second and 10,000 requests/day. Lower GONG_MAX_PAGES, raise "
"GONG_MIN_REQUEST_INTERVAL, or ask Gong support to raise the account limit."
)
r.raise_for_status()
async def paged_post(path: str, body: dict[str, Any], record_key: str) -> tuple[list[Any], bool]:
"""POST through cursor pagination up to GONG_MAX_PAGES. Returns (records, truncated)."""
records: list[Any] = []
cursor: str | None = None
for _ in range(max(1, GONG_MAX_PAGES)):
payload = dict(body)
if cursor:
payload["cursor"] = cursor
data = await gong_request("POST", path, json=payload)
records.extend(data.get(record_key, []) or [])
cursor = (data.get("records") or {}).get("cursor")
if not cursor:
return records, False
return records, True
def _date_filter(arguments: dict[str, Any], *, keys: tuple[str, str]) -> dict[str, Any]:
out: dict[str, Any] = {}
for key in keys:
if v := arguments.get(key):
out[key] = v
workspace = arguments.get("workspace_id") or GONG_WORKSPACE_ID
if workspace:
out["workspaceId"] = workspace
return out
# ----- Server + tool registry -----
server = Server("gong-revops")
# The contentSelector this server sends to /v2/calls/extensive. Fixed, not
# caller-controlled: `media` is deliberately absent so the server never requests
# 8-hour signed audio/video URLs (a separate scope, and a link that outlives the
# conversation it appeared in). `content.trackerOccurrences` is included because
# tracker *counts* without speaker and timestamp cannot tell you whether the customer
# raised pricing or your rep did — which inverts the meaning of the signal.
SIGNALS_CONTENT_SELECTOR: dict[str, Any] = {
"context": "Extended",
"contextTiming": ["Now"],
"exposedFields": {
"parties": True,
"content": {
"trackers": True,
"trackerOccurrences": True,
"brief": True,
"keyPoints": True,
"callOutcome": True,
"topics": True,
},
"interaction": {
"speakers": True,
"personInteractionStats": True,
"questions": True,
},
"collaboration": {"publicComments": True},
},
}
@server.list_tools()
async def list_tools() -> list[Tool]:
return [
Tool(
name="find_calls",
description=(
"List calls in a date range (GET /v2/calls). Cheap metadata only — id, "
"title, start time, duration, participant count. Use this first to scope "
"a question, then pass the ids you care about to call_signals. Pages at "
"100 records; follows at most GONG_MAX_PAGES pages."
),
inputSchema={
"type": "object",
"properties": {
"fromDateTime": {
"type": "string",
"description": "ISO-8601, e.g. 2026-07-01T00:00:00Z. Calls starting at or after.",
},
"toDateTime": {
"type": "string",
"description": "ISO-8601. Calls starting before.",
},
"workspace_id": {"type": "string"},
},
"required": ["fromDateTime"],
},
),
Tool(
name="list_trackers",
description=(
"List keyword/smart tracker DEFINITIONS (GET /v2/settings/trackers). "
"Returns configuration only — names, ids, keywords, affiliation — and no "
"match counts. Occurrence counts come from call_signals. Call this to "
"learn what your workspace actually tracks before assuming a tracker name."
),
inputSchema={
"type": "object",
"properties": {"workspace_id": {"type": "string"}},
},
),
Tool(
name="call_signals",
description=(
"Retrieve analyzed signals for specific calls (POST /v2/calls/extensive): "
"parties, tracker matches with speaker and timestamp, Spotlight brief, key "
"points, auto call outcome, topics, talk ratio and interactivity stats, and "
"public comments. No transcript, no media URLs. This is the workhorse tool."
),
inputSchema={
"type": "object",
"properties": {
"call_ids": {
"type": "array",
"items": {"type": "string"},
"description": "Specific Gong call ids. Preferred over a date range.",
},
"fromDateTime": {"type": "string", "description": "ISO-8601, used when call_ids is omitted."},
"toDateTime": {"type": "string", "description": "ISO-8601."},
"workspace_id": {"type": "string"},
},
},
),
Tool(
name="call_transcript",
description=(
"Retrieve verbatim transcripts for up to GONG_MAX_TRANSCRIPT_CALLS calls "
"(POST /v2/calls/transcript). Disabled unless GONG_ALLOW_TRANSCRIPTS=true. "
"Requires a justification. Prefer call_signals — the brief and key points "
"answer most questions at a fraction of the tokens and the PII exposure."
),
inputSchema={
"type": "object",
"properties": {
"call_ids": {"type": "array", "items": {"type": "string"}},
"justification": {
"type": "string",
"description": "Why the verbatim transcript is needed instead of the brief. Min 10 chars.",
},
},
"required": ["call_ids", "justification"],
},
),
Tool(
name="rep_interaction_stats",
description=(
"Per-rep aggregated interaction stats over a date range "
"(POST /v2/stats/interaction): longest monologue, longest customer story, "
"interactivity, patience, question rate. Covers only calls that had Whisper "
"enabled, so a rep with few recorded calls looks like a rep with bad numbers."
),
inputSchema={
"type": "object",
"properties": {
"fromDate": {"type": "string", "description": "YYYY-MM-DD"},
"toDate": {"type": "string", "description": "YYYY-MM-DD"},
"user_ids": {"type": "array", "items": {"type": "string"}},
},
"required": ["fromDate", "toDate"],
},
),
Tool(
name="deal_risk_digest",
description=(
"Derived signal, not a Gong endpoint. Joins tracker definitions to tracker "
"occurrences across a date range and reports which calls and accounts hit "
"the risk trackers named in GONG_RISK_TRACKERS, split by whether the "
"CUSTOMER or your own rep said it. Gong has no public deals endpoint; this "
"is the closest honest substitute. Attribute nothing to a deal stage from "
"this output — join it to your CRM for that."
),
inputSchema={
"type": "object",
"properties": {
"fromDateTime": {"type": "string", "description": "ISO-8601"},
"toDateTime": {"type": "string", "description": "ISO-8601"},
"tracker_names": {
"type": "array",
"items": {"type": "string"},
"description": "Override GONG_RISK_TRACKERS for this call.",
},
"workspace_id": {"type": "string"},
},
"required": ["fromDateTime"],
},
),
]
@server.call_tool()
async def call_tool(name: str, arguments: dict[str, Any]) -> list[TextContent]:
if name == "find_calls":
params = _date_filter(arguments, keys=("fromDateTime", "toDateTime"))
rows: list[dict[str, Any]] = []
cursor: str | None = None
truncated = False
for page in range(max(1, GONG_MAX_PAGES)):
q = dict(params)
if cursor:
q["cursor"] = cursor
data = await gong_request("GET", "/v2/calls", params=q)
rows.extend(data.get("calls", []) or [])
cursor = (data.get("records") or {}).get("cursor")
if not cursor:
break
truncated = page == max(1, GONG_MAX_PAGES) - 1
return [TextContent(type="text", text=str(_slim_calls(rows, truncated)))]
if name == "list_trackers":
params: dict[str, Any] = {}
workspace = arguments.get("workspace_id") or GONG_WORKSPACE_ID
if workspace:
params["workspaceId"] = workspace
data = await gong_request("GET", "/v2/settings/trackers", params=params)
trackers = [
{
"trackerId": t.get("trackerId"),
"trackerName": t.get("trackerName"),
"affiliation": t.get("affiliation"),
"keywords": [
kw
for lang in (t.get("languageKeywords") or [])
for kw in (lang.get("keywords") or [])
][:20],
}
for t in (data.get("keywordTrackers") or [])
]
return [
TextContent(
type="text",
text=str(
{
"trackers": trackers,
"note": "Definitions only — no match counts. Occurrences come from call_signals.",
}
),
)
]
if name == "call_signals":
body: dict[str, Any] = {"contentSelector": SIGNALS_CONTENT_SELECTOR}
call_ids = arguments.get("call_ids")
if call_ids:
body["filter"] = {"callIds": [str(c) for c in call_ids]}
else:
if not arguments.get("fromDateTime"):
raise ValueError("Pass call_ids, or fromDateTime to scope a date range.")
body["filter"] = _date_filter(arguments, keys=("fromDateTime", "toDateTime"))
calls, truncated = await paged_post("/v2/calls/extensive", body, "calls")
return [TextContent(type="text", text=str(_slim_signals(calls, truncated)))]
if name == "call_transcript":
justification = (arguments.get("justification") or "").strip()
if len(justification) < 10:
raise ValueError("justification is mandatory and must be at least 10 characters.")
if not GONG_ALLOW_TRANSCRIPTS:
raise PermissionError(
"call_transcript is disabled. Verbatim transcripts put full customer "
"speech into the model context. Set GONG_ALLOW_TRANSCRIPTS=true only "
"after confirming that is allowed for this data."
)
call_ids = [str(c) for c in (arguments.get("call_ids") or [])]
if not call_ids:
raise ValueError("call_ids must be a non-empty list.")
if len(call_ids) > GONG_MAX_TRANSCRIPT_CALLS:
raise ValueError(
f"Refusing {len(call_ids)} transcripts in one call; the cap is "
f"{GONG_MAX_TRANSCRIPT_CALLS}. Narrow the question with call_signals "
"first, or raise GONG_MAX_TRANSCRIPT_CALLS deliberately."
)
data = await gong_request(
"POST", "/v2/calls/transcript", json={"filter": {"callIds": call_ids}}
)
return [
TextContent(
type="text",
text=str(
{
"justification": justification,
"callTranscripts": data.get("callTranscripts", []),
}
),
)
]
if name == "rep_interaction_stats":
body: dict[str, Any] = {
"filter": {
"fromDate": arguments["fromDate"],
"toDate": arguments["toDate"],
}
}
if v := arguments.get("user_ids"):
body["filter"]["userIds"] = [str(u) for u in v]
rows, truncated = await paged_post("/v2/stats/interaction", body, "usersAggregateActivity")
return [
TextContent(
type="text",
text=str(
{
"users": rows,
"truncated": truncated,
"caveat": (
"Stats derive only from calls with Whisper enabled. Low call "
"volume reads as poor metrics; check call counts before coaching."
),
}
),
)
]
if name == "deal_risk_digest":
wanted = [t.lower() for t in (arguments.get("tracker_names") or GONG_RISK_TRACKERS)]
body = {
"filter": _date_filter(arguments, keys=("fromDateTime", "toDateTime")),
"contentSelector": SIGNALS_CONTENT_SELECTOR,
}
calls, truncated = await paged_post("/v2/calls/extensive", body, "calls")
return [TextContent(type="text", text=str(_risk_digest(calls, wanted, truncated)))]
raise ValueError(f"Unknown tool: {name}")
# ----- Response slimming (keep model payloads tractable) -----
def _slim_calls(calls: list[dict[str, Any]], truncated: bool) -> dict[str, Any]:
rows = [
{
"id": c.get("id"),
"title": c.get("title"),
"started": c.get("started"),
"duration_s": c.get("duration"),
"direction": c.get("direction"),
"url": c.get("url"),
}
for c in calls
]
return {"count": len(rows), "truncated": truncated, "calls": rows}
def _external_parties(call: dict[str, Any]) -> list[str]:
return [
p.get("name") or p.get("emailAddress") or "?"
for p in (call.get("parties") or [])
if (p.get("affiliation") or "").lower() == "external"
]
def _slim_signals(calls: list[dict[str, Any]], truncated: bool) -> dict[str, Any]:
rows = []
for c in calls:
meta = c.get("metaData") or {}
content = c.get("content") or {}
rows.append(
{
"id": meta.get("id"),
"title": meta.get("title"),
"started": meta.get("started"),
"outcome": (content.get("callOutcome") or {}).get("category"),
"external_parties": _external_parties(c),
"trackers": [
{"name": t.get("name"), "count": t.get("count")}
for t in (content.get("trackers") or [])
if t.get("count")
],
"brief": content.get("brief"),
"key_points": [kp.get("text") for kp in (content.get("keyPoints") or [])],
"topics": [
{"name": t.get("name"), "duration_s": t.get("duration")}
for t in (content.get("topics") or [])
],
}
)
return {"count": len(rows), "truncated": truncated, "calls": rows}
def _risk_digest(
calls: list[dict[str, Any]], wanted: list[str], truncated: bool
) -> dict[str, Any]:
"""Join tracker occurrences to speaker affiliation, per call.
Speaker affiliation is the load-bearing part. "Pricing Pushback" said by your own
rep is a rep-behavior signal; said by the customer it is a deal signal. Counting
them together produces a risk number that moves for the wrong reasons.
"""
hits = []
for c in calls:
meta = c.get("metaData") or {}
parties = {p.get("speakerId"): p for p in (c.get("parties") or []) if p.get("speakerId")}
matched = []
for tracker in (c.get("content") or {}).get("trackers") or []:
if (tracker.get("name") or "").lower() not in wanted:
continue
by_side = {"customer": 0, "internal": 0, "unattributed": 0}
for occ in tracker.get("occurrences") or []:
party = parties.get(occ.get("speakerId"))
affiliation = (party or {}).get("affiliation", "")
if affiliation.lower() == "external":
by_side["customer"] += 1
elif affiliation.lower() == "internal":
by_side["internal"] += 1
else:
by_side["unattributed"] += 1
if not any(by_side.values()):
# Tracker matched but occurrences were not exposed; report the count
# rather than dropping the signal, and mark it unattributed.
by_side["unattributed"] = tracker.get("count") or 0
matched.append({"tracker": tracker.get("name"), "said_by": by_side})
if matched:
hits.append(
{
"call_id": meta.get("id"),
"title": meta.get("title"),
"started": meta.get("started"),
"external_parties": _external_parties(c),
"risk_trackers": matched,
}
)
return {
"calls_scanned": len(calls),
"calls_with_risk_signals": len(hits),
"truncated": truncated,
"risk_trackers_checked": wanted,
"hits": hits,
"note": (
"Derived from tracker occurrences on calls. Gong's public API exposes no "
"deal-board read endpoint — join call_id or account name to your CRM for "
"stage, amount, and close date. Do not treat this as a forecast."
),
}
# ----- Entrypoint -----
async def main() -> None:
require_config()
async with stdio_server() as (read, write):
await server.run(read, write, server.create_initialization_options())
if __name__ == "__main__":
asyncio.run(main())