Developer
News and Updates
Get Support
Sign in
Get Support
Sign in
DOCUMENTATION
Cloud
Data Center
Resources
Sign in
Sign in
DOCUMENTATION
Cloud
Data Center
Resources
Sign in
Last updated Jul 30, 2026

TWG Benchmark

This benchmark answers one question: does connecting your AI agent to Teamwork Graph produce better answers, using fewer tokens, on your own work?

You run a fixed set of real prompts twice — once without Teamwork Graph, once with it — then pick the better answer in a blind review. The tool produces a report comparing the results side by side.

The recommended workflow runs through the TWG AI Context Benchmark skill in Codex or Claude Code and takes about 20–40 minutes end to end. Your data never leaves your machine unless you explicitly choose to share a local report bundle.

This tool is in active development. Results are directional, not final. Read how to interpret your results before quoting any number — token savings and answer quality do not always move together.

How the benchmark works

StageWhat happens
PrepareInstall the Benchmark CLI, sign in to your agent and TWG CLI, and connect at least one code source if needed (GitHub or Bitbucket).
ValidateA readiness check confirms connectors and source reads — no tokens spent.
RunThe same prompts are answered twice: baseline vs Teamwork Graph.
ReviewA browser opens; you pick the better answer for each prompt.
ReportQuality, token, and latency results, broken down by prompt.
ShareOptional: package and send results to Atlassian for analysis.

Baseline arm = your agent on standard Atlassian access, Teamwork Graph off.
Teamwork Graph arm = the same agent, same prompts, with the TWG CLI connected.

The TWG AI Context Benchmark skill drives the supported workflow: readiness, isolated baseline and Teamwork Graph lanes, blind review, and the report handoff.

Get started

  1. Install and configure the Benchmark CLI
  2. Run a benchmark
  3. Review and share results
  4. Troubleshooting and command reference

Rate this page: