Class AIProcessingInfo
An immutable processing-mode observation for one provider inference attempt.
public sealed class AIProcessingInfo
- Inheritance
-
AIProcessingInfo
- Inherited Members
Remarks
The index counts inference attempts, including server continuations, rather than HTTP requests or LLM rounds. Retries and format repairs can add attempts. AppliedSpeed is null when the server did not report a recognized mode, including failed attempts. This describes a service mode, not a measured token generation rate.
Constructors
AIProcessingInfo(int, InferenceSpeed, InferenceSpeed?, string?, string?)
public AIProcessingInfo(int requestIndex, InferenceSpeed requestedSpeed, InferenceSpeed? appliedSpeed = null, string? rawAppliedMode = null, string? responseId = null)
Parameters
requestIndexintrequestedSpeedInferenceSpeedappliedSpeedInferenceSpeed?rawAppliedModestringresponseIdstring
Properties
AppliedSpeed
public InferenceSpeed? AppliedSpeed { get; }
Property Value
IsDowngraded
True only when a Fast request was explicitly reported as Standard.
public bool IsDowngraded { get; }
Property Value
RawAppliedMode
public string? RawAppliedMode { get; }
Property Value
RequestIndex
public int RequestIndex { get; }
Property Value
RequestedSpeed
public InferenceSpeed RequestedSpeed { get; }
Property Value
ResponseId
public string? ResponseId { get; }