Models

Grouped by what each one can prove, not by how capable it is. The label on a group is the same label the answer will carry, and if it cannot be established at the moment you ask, the request is refused rather than served one rung down.

Runs inside an attested hardware enclave. Every chat answer carries a receipt you can verify yourself, offline.

  • GPT OSS 120B

    openai-gpt-oss-120b

    OpenAI: GPT OSS 120B

    128K ctx · $0.60/M out

    Verified by us: a live probe returned a receipt for this model.

  • GPT OSS 20B

    openai-gpt-oss-20b

    OpenAI: GPT OSS 20B

    128K ctx · $0.15/M out

    Verified by us: a live probe returned a receipt for this model.

  • Muse Glimmer 30B

    meta-muse-glimmer-30b

    Meta: Muse Glimmer 30B

    128K ctx · $1.1/M out

    Verified by us: a live probe returned a receipt for this model.

  • Nemotron 3.5 Lightning

    nvidia-nemotron-3.5-lightning

    NVIDIA: Nemotron 3.5 Lightning

    256K ctx · $0.20/M out

    Verified by us: a live probe returned a receipt for this model.

  • Qwen2.5 7B Instruct

    qwen-qwen-2.5-7b-instruct

    32K ctx · $0.20/M out

    Verified by us: a live probe returned a receipt for this model.

  • Untold Large

    untold-large

    Phala: Qwen3.6 35B-A3B Uncensored (Aggressive)

    128K ctx · $1.5/M out

    Verified by us: a live probe returned a receipt for this model.

  • Untold Uncensored

    untold-uncensored

    Phala: Gemma-4 26B-A4B Uncensored (Heretic)

    64K ctx · $0.70/M out

    Verified by us: a live probe returned a receipt for this model.

We strip your identity before forwarding, but the model provider still sees the text of your prompt. This is the weakest mode and we say so plainly.

Also at /api/models. Every label here was established by a real request to the model, not by reading what the gateway says about itself.