Transfer pricing benchmarking requires an early decision on whether to test a transaction on a standalone basis or aggregate it with other international transactions of the entity. Rule 10A's language on aggregation can support either approach depending on the facts, and taxpayers and tax authorities routinely take opposing positions on which approach the facts justify. A recent ITAT Mumbai ruling in the case of NTT India (formerly Dimension Data India) addressed this question directly, and its reasoning has implications for how AI-assisted benchmarking tools are being designed and used by Indian tax teams.
NTT India had benchmarked its management-fee payment to its Asian regional AE as part of an aggregate, entity-level TNMM analysis, arguing that the overall margin was at arm's length once all international transactions were considered together. The TPO disagreed, extracted the management-fee transaction from that aggregate analysis, applied the CUP method, found no comparable uncontrolled data, and valued the entire service at nil, resulting in an adjustment of nearly ₹93.23 crore. The ITAT deleted the adjustment. In November 2025, a separate ITAT Mumbai bench reached a similar conclusion in an unrelated case, holding that a TPO cannot accept TNMM for a taxpayer's transactions in aggregate and then isolate a single line item for independent nil valuation. The two rulings, from different benches and about ten months apart, apply the same underlying principle.
Rule 10A's aggregation language has not changed. What appears to have shifted is how often tribunals are being asked to police where the aggregation boundary sits, and the consistency with which they are ruling against the department on this point. For a TP practitioner, this strengthens the argument that once an aggregate TNMM position has been accepted, or at least not affirmatively rejected, for an assessee's transactions as a whole, the TPO's room to isolate and independently value a single line item is narrower than it may once have appeared.
This reasoning is also relevant to AI-assisted benchmarking and documentation tools, including agentic platforms and GenAI-based comparable-search products, that are being marketed to Indian tax teams and Big Four practices. Most of these tools make an implicit aggregation-or-segregation choice somewhere in their workflow. Some default to pulling entity-level financials and running margin comparisons across the whole profit and loss account. Others are built to isolate and test each intercompany transaction separately because that is easier to automate and audit.
Neither default is safe on its own. A tool that always aggregates risks reproducing the outcome favourable to the taxpayer in NTT India even in fact patterns where aggregation is not actually justified, inviting a TPO challenge on the opposite theory. A tool that always segregates transactions for cleaner, auditable output risks reproducing the same TPO error that was overturned twice within about a year, testing a management fee, a cost-contribution arrangement, or an IT service fee in isolation when it was never meant to be tested that way. The tribunals' reasoning indicates that this decision has to rest on how closely the transactions are linked on the specific facts, not on a default setting built into a product.
This has a direct implication for how TP teams evaluate any AI benchmarking tool they consider buying or building. Vendors are likely to emphasise comparable-search speed and documentation drafting, but the more relevant question for audit defensibility is narrower: does the tool make its aggregation-or-segregation choice explicit, does it require a person to record the specific factual basis for that choice, and would that basis survive a TPO challenge along the lines the TPO raised in NTT India.
As more Indian captives and GCCs adopt AI-assisted TP documentation workflows, practitioners should treat the aggregation-or-segregation call as a documented, fact-specific judgment that sits with a person on the team, not a default a vendor sets. The NTT India line of rulings gives that judgment more weight than it may have carried before, and it is a reasonable basis for reviewing any AI-generated benchmarking file before it is relied upon in a submission to the TPO or the DRP.