GPT 4o mini - Oxen.ai

Try GPT 4o mini in the Workbench

Run this model interactively, tune parameters, and compare outputs.

Model ID: gpt-4o-mini GPT-4o mini is a Small Language Model (SLM) designed for cost-efficient and fast processing. It excels in handling a wide range of tasks with low latency and reduced computational costs, making it ideal for applications requiring multiple model calls, large context processing, or real-time interactions. Some other noteworthy features of GPT-4o mini include multimodal capabilities, supporting both text and vision inputs, and improved performance in non-English languages compared to its predecessors.

Metric	Value
Parameter Count	Unknown
Mixture of Experts	No
Context Length	128,000 tokens
Multilingual	Yes
Quantized*	Unknown

*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.

Example request

Use the Workbench as a request builder: configure parameters for this model in the UI, then open the API tab to copy the exact cURL or Python call.

Minimal
Basic parameters
All parameters

curl -X POST https://hub.oxen.ai/api/ai/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $OXEN_API_KEY" \
  -d '{
  "model": "gpt-4o-mini",
  "messages": [
    {
      "role": "user",
      "content": "Hello, what can you do?"
    }
  ]
}'

import os
import requests

response = requests.post(
    "https://hub.oxen.ai/api/ai/chat/completions",
    headers={
        "Content-Type": "application/json",
        "Authorization": f"Bearer {os.environ['OXEN_API_KEY']}",
    },
    json={
        "model": "gpt-4o-mini",
        "messages": [
            {
                "role": "user",
                "content": "Hello, what can you do?"
            }
        ]
    },
)
response.raise_for_status()
print(response.json())

curl -X POST https://hub.oxen.ai/api/ai/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $OXEN_API_KEY" \
  -d '{
  "model": "gpt-4o-mini",
  "messages": [
    {
      "role": "user",
      "content": "Hello, what can you do?"
    }
  ],
  "temperature": 0.7,
  "max_tokens": 1024,
  "stream": false
}'

import os
import requests

response = requests.post(
    "https://hub.oxen.ai/api/ai/chat/completions",
    headers={
        "Content-Type": "application/json",
        "Authorization": f"Bearer {os.environ['OXEN_API_KEY']}",
    },
    json={
        "model": "gpt-4o-mini",
        "messages": [
            {
                "role": "user",
                "content": "Hello, what can you do?"
            }
        ],
        "temperature": 0.7,
        "max_tokens": 1024,
        "stream": false
    },
)
response.raise_for_status()
print(response.json())

curl -X POST https://hub.oxen.ai/api/ai/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $OXEN_API_KEY" \
  -d '{
  "model": "gpt-4o-mini",
  "messages": [
    {
      "role": "user",
      "content": "Hello, what can you do?"
    }
  ],
  "temperature": 0.7,
  "max_tokens": 1024,
  "stream": false,
  "top_p": 1.0,
  "frequency_penalty": 0,
  "presence_penalty": 0
}'

import os
import requests

response = requests.post(
    "https://hub.oxen.ai/api/ai/chat/completions",
    headers={
        "Content-Type": "application/json",
        "Authorization": f"Bearer {os.environ['OXEN_API_KEY']}",
    },
    json={
        "model": "gpt-4o-mini",
        "messages": [
            {
                "role": "user",
                "content": "Hello, what can you do?"
            }
        ],
        "temperature": 0.7,
        "max_tokens": 1024,
        "stream": false,
        "top_p": 1.0,
        "frequency_penalty": 0,
        "presence_penalty": 0
    },
)
response.raise_for_status()
print(response.json())

Fetch model details

The models endpoint returns the full model object, including its json_request_schema.

curl -H "Authorization: Bearer $OXEN_API_KEY" https://hub.oxen.ai/api/ai/models/gpt-4o-mini

Request parameters

This model follows the standard OpenAI chat completions request body. See the chat completions reference for the full parameter list.

Try GPT 4o mini in the Workbench

​Example request

​Fetch model details

​Request parameters

Example request

Fetch model details

Request parameters