vllm.entrypoints.generate.label_reads
¶
One-token reads that return the logprobs of chosen label tokens.
Classes:
Functions:
-
next_token_label_reads–Runs one read per input at once, as request
{request_id}-{i}. Each
LabelRead
dataclass
¶
Attributes:
Source code in vllm/entrypoints/generate/label_reads.py
logprobs
instance-attribute
¶
Full-vocabulary logprob of each label token, in logprob_token_ids
order.
next_token_label_reads(engine_client, engine_inputs, sampling_params, request_id, *, lora_request=None, trace_headers=None, priority=0)
async
¶
Runs one read per input at once, as request {request_id}-{i}. Each
read's params set max_tokens=1 and the label tokens as
logprob_token_ids. Raises ValueError naming the item when a read has
no output or lacks a label's logprob.