Skip to content

Add in LLM Observability Prompt Tracking API - #12161

Merged
gh-worker-dd-mergequeue-cf854d[bot] merged 5 commits into
masterfrom
sabrenner/llmobs-manual-prompt-tracking
Aug 12, 2026
Merged

gh-worker-dd-mergequeue-cf854d[bot] merged 5 commits into
masterfrom
sabrenner/llmobs-manual-prompt-tracking

Conversation

@sabrenner

@sabrenner sabrenner commented Aug 7, 2026 •

Copy link
Copy Markdown
Contributor

What Does This Do

Adds a manual prompt track API for the LLM Observability SDK. This works by exposing a prompt object builder to pass in to a new annotatePrompt method. This method only applies the prompt annotation on LLM spans.

Motivation

Further support parity with Node.js and Python SDKs.

Additional Notes

We'll need to follow up with allowing these to be annotated on auto-instrumented LLM spans (ie OpenAI)

Contributor Checklist

Jira ticket: MLOB-7901

@sabrenner sabrenner added type: feature Enhancements and improvements comp: mlobs ML Observability (LLMObs) labels Aug 7, 2026
@sabrenner

Copy link
Copy Markdown
Contributor Author

@codex review

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 21b2f2e0ba

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "@codex address that feedback".

Comment thread dd-trace-api/src/main/java/datadog/trace/api/llmobs/LLMObs.java
@sabrenner

Copy link
Copy Markdown
Contributor Author

@codex review

@chatgpt-codex-connector

Copy link
Copy Markdown

Codex Review: Didn't find any major issues. More of your lovely PRs please.

Reviewed commit: 866782e2d6

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "@codex review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "@codex address that feedback".

@datadog-datadog-us1-prod

This comment has been minimized.

@dd-octo-sts

dd-octo-sts Bot commented Aug 7, 2026 •

Copy link
Copy Markdown
Contributor

🟢 Java Benchmark SLOs — All performance SLOs passed

Suite Status
Startup 🟢 pass

SLO thresholds are defined here based on automatically generated metrics. A warning is raised when results are within 5% of the threshold.

PR vs. master results
Scenario Candidate master Δ (95% CI of mean)
startup:insecure-bank:iast:Agent 14.04 s 13.97 s [-0.3%; +1.4%] (no difference)
startup:insecure-bank:tracing:Agent 12.93 s 12.97 s [-1.1%; +0.5%] (no difference)
startup:petclinic:appsec:Agent 16.87 s 16.70 s [+0.0%; +2.1%] (maybe worse)
startup:petclinic:iast:Agent 16.42 s 16.98 s [-7.5%; +0.9%] (no difference)
startup:petclinic:profiling:Agent 16.71 s 16.22 s [-1.5%; +7.5%] (no difference)
startup:petclinic:sca:Agent 16.35 s 16.66 s [-6.3%; +2.5%] (no difference)
startup:petclinic:tracing:Agent 16.07 s 16.26 s [-2.4%; +0.0%] (no difference)

Commit: f916f872 · CI Pipeline · Benchmarking Platform UI


Load and DaCapo benchmarks can be triggered manually in the GitLab pipeline. Results will appear in the Benchmarking Platform UI after completion.

@sabrenner
sabrenner marked this pull request as ready for review August 7, 2026 18:33
@sabrenner
sabrenner requested review from a team as code owners August 7, 2026 18:33
@sabrenner
sabrenner requested a review from ygree August 7, 2026 18:33

@datadog-datadog-us1-prod datadog-datadog-us1-prod Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Datadog Autotest: FAIL

A later prompt annotation that only adds version or tags silently resets custom RAG context and query variable keys to defaults, corrupting tracking classification for applications that enrich a prompt across multiple calls.

📊 Validated against 7 scenarios · Open Bits AI session

🤖 Datadog Autotest · Commit b387155 · What is Autotest? · @DataDog review to ask questions · Any feedback? Reach out in #autotest

@ncybul
ncybul added this pull request to the merge queue Aug 12, 2026
@dd-octo-sts

dd-octo-sts Bot commented Aug 12, 2026

Copy link
Copy Markdown
Contributor

/merge

@gh-worker-devflow-routing-ef8351

gh-worker-devflow-routing-ef8351 Bot commented Aug 12, 2026 •

Copy link
Copy Markdown

View all feedbacks in Devflow UI.

2026-08-12 18:46:16 UTC ℹ️ Start processing command /merge


2026-08-12 18:46:21 UTC ℹ️ MergeQueue: pull request added to the queue

The expected merge time in master is approximately 2h (p90).


2026-08-12 19:35:18 UTC ℹ️ MergeQueue: This merge request was merged

@github-merge-queue
github-merge-queue Bot removed this pull request from the merge queue due to failed status checks Aug 12, 2026
@gh-worker-dd-mergequeue-cf854d
gh-worker-dd-mergequeue-cf854d Bot merged commit 0b9aa56 into master Aug 12, 2026
594 checks passed
@gh-worker-dd-mergequeue-cf854d
gh-worker-dd-mergequeue-cf854d Bot deleted the sabrenner/llmobs-manual-prompt-tracking branch August 12, 2026 19:35
@github-actions github-actions Bot added this to the 1.66.0 milestone Aug 12, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

comp: mlobs ML Observability (LLMObs) type: feature Enhancements and improvements

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants