Try Pixtral 12B in the Workbench
Run this model interactively, tune parameters, and compare outputs.
pixtral-12b
Pixtral 12B is a Multimodal LLM that excels in handling both images and text, supporting tasks like image captioning, visual question answering, and document analysis. It maintains strong performance in text-only tasks as well.
*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.
Example request
- Minimal
- Basic parameters
- All parameters
Fetch model details
The models endpoint returns the full model object, including itsjson_request_schema.