Skip to Main Content
Shape the future of IBM watsonx Orchestrate

Start by searching and reviewing ideas others have posted, and add a comment (private if needed), vote, or subscribe to updates on them if they matter to you.

If you can't find what you are looking for, create a new idea:

  1. stick to one feature enhancement per idea

  2. add as much detail as possible, including use-case, examples & screenshots (put anything confidential in Hidden details field or a private comment)

  3. Explain business impact and timeline of project being affected

[For IBMers] Add customer/project name, details & timeline in Hidden details field or a private comment (only visible to you and the IBM product team).

This all helps to scope and prioritize your idea among many other good ones. Thank you for your feedback!

Specific links you will want to bookmark for future use
Learn more about IBM watsonx Orchestrate - Use this site to find out additional information and details about the product.
Welcome to the IBM Ideas Portal (https://www.ibm.com/ideas) - Use this site to find out additional information and details about the IBM Ideas process and statuses.
IBM Unified Ideas Portal (https://ideas.ibm.com) - Use this site to view all of your ideas, create new ideas for any IBM product, or search for ideas across all of IBM.
ideasibm@us.ibm.com - Use this email to suggest enhancements to the Ideas process or request help from IBM for submitting your Ideas.

Status Submitted
Created by Guest
Created on Sep 19, 2026

Per-sentence source-span grounding flags

Problem: In document-grounded QA, the dominant failure isn't invented facts but interpretive overconfidence — unsupported characterizations of sources laundered into confident general statements. Fluency-optimized models can't supply claim-level sourcing.

Idea: Add a grounded-QA Orchestrate agent that attaches per-sentence source spans and flags any sentence whose support is "interpretive" rather than extractive, routing those sentences to analyst review. Grounded in Hagar, Agustianto & Diakopoulos (2025, Northwestern), who found 30% of outputs on a 300-document reporting task contained hallucinations, mostly interpretive overconfidence.

Value: Analysts and journalists get verifiable, sentence-level provenance — trust moves from the model's fluency to checkable sources.

Idea priority Low