vllm.config.aux_output
¶
Configuration for execution auxiliary outputs.
Classes:
-
AuxOutputConfig–Configuration for auxiliary-output delivery.
AuxOutputConfig
¶
Configuration for auxiliary-output delivery.
Methods:
-
compute_hash–Hash AuxOutput settings that alter the model forward graph.
Attributes:
-
enable_return_routed_experts(bool) –Capture and return routed-experts auxiliary outputs.
-
enabled(bool) –Whether any execution auxiliary output is enabled.
-
max_bytes(int | None) –LRU capacity, or
Noneto derive it from the KV cache capacity.
Source code in vllm/config/aux_output.py
enable_return_routed_experts = False
class-attribute
instance-attribute
¶
Capture and return routed-experts auxiliary outputs.
enabled
property
¶
Whether any execution auxiliary output is enabled.
max_bytes = Field(default=None, gt=0)
class-attribute
instance-attribute
¶
LRU capacity, or None to derive it from the KV cache capacity.
compute_hash()
¶
Hash AuxOutput settings that alter the model forward graph.