Service Azure permettant d'accéder aux modèles GPT-3 OpenAI avec des capacités d'entreprise.
Hello Patrick AGIN
Welcome to Microsoft Q&A .Thank you for reaching out to us.
Based on the behavior observed, the issue appears to be related to how token usage information is surfaced through the Responses API rather than a problem with model execution itself. Since the deployment successfully generates the expected output and the equivalent Chat Completions requests report non-zero token usage
Managed Compute uses platform-managed runtimes and exposes OpenAI-compatible endpoints for supported models.
Deployment templates may also include runtime-specific configurations and serving optimizations. However, the internal implementation details of request processing, usage accounting, and response construction are not publicly documented.
Because of this, it is not currently possible to determine whether the discrepancy originates within the underlying runtime implementation or another component of the Managed Compute serving stack without further engineering investigation.
Since the deployment is already using the latest deployment template label, there is currently no known upgrade path expected to resolve this behavior.
To help narrow down the source of the discrepancy, the following validation steps are recommended:
- Comparing Responses API usage values with Azure Monitor metrics:
- Input Tokens
- Output Tokens
- Total Tokens
- Performing a comparison using another Managed Compute model and verify usage reporting between:
- /openai/v1/responses
- /chat/completions
As a temporary workaround, Chat Completions can be used when accurate per-request token usage reporting is required, since usage reporting is currently functioning correctly through that API path. For workloads that require the Responses API, Azure Monitor metrics can be used as an alternative source to validate token consumption
The following references might be helpful , please check them out
- Deploy open-source models with managed compute in Microsoft Foundry - Microsoft Foundry | Microsoft Learn
- Managed compute in Microsoft Foundry - Microsoft Foundry | Microsoft Learn
- Deployment options for Microsoft Foundry Models (classic) - Microsoft Foundry (classic) portal | Microsoft Learn
Please let us know if the response was helpful
Thank you