> ## Documentation Index > Fetch the complete documentation index at: https://openrouter.ai/docs/llms.txt > Use this file to discover all available pages before exploring further. # PreferredMaxLatency Preferred maximum latency (in seconds). Can be a number (applies to p50) or an object with percentile-specific cutoffs. Endpoints above the threshold(s) may still be used, but are deprioritized in routing. When using fallback models, this may cause a fallback model to be used instead of the primary model if it meets the threshold. ## Supported Types ### ```go theme={null} preferredMaxLatency := components.CreatePreferredMaxLatencyNumber(float64{/* values here */}) ``` ### PercentileLatencyCutoffs ```go theme={null} preferredMaxLatency := components.CreatePreferredMaxLatencyPercentileLatencyCutoffs(components.PercentileLatencyCutoffs{/* values here */}) ``` ## Union Discrimination Use the `Type` field to determine which variant is active, then access the corresponding field: ```go theme={null} switch preferredMaxLatency.Type { case components.PreferredMaxLatencyTypeNumber: // preferredMaxLatency.Number is populated case components.PreferredMaxLatencyTypePercentileLatencyCutoffs: // preferredMaxLatency.PercentileLatencyCutoffs is populated } ```