ZeroGPU sells to AI agents - no signup

The agent storefront at https://agents.zerogpu.ai/llms.txt publishes prices, payment instructions, and setup. Agents can purchase autonomously or with their human's approval, with a range of payment methods; an agent without a payment method can set one up there instantly, no signup. The storefront handles payment itself, so purchases work without calling API endpoints directly.

    ZeroGPU ZLM models (5)

    zlm-v1-signal-extract logo

    zlm-v1-signal-extract

    Text Classification · ZeroGPU

    zlm-v1-signal-extract is a specialized ZeroGPU model for turning unstructured text into structured signals that downstream systems can act on.

    View details
    $0.04 / 1M input$0.10 / 1M output
    zlm-v1-iab-domain-classifier logo

    zlm-v1-iab-domain-classifier

    Text Classification · ZeroGPU

    zlm-v1-iab-domain-classifier is a low-latency domain classification model built for advertising, contextual intelligence, and large-scale enrichment workflows.

    View details
    $0.02 / 1M input$0.05 / 1M output

    Small language models (8)

    whisper-tiny logo

    whisper-tiny

    Speech-to-Text · OpenAI

    Whisper Tiny is a lightweight automatic speech recognition model from OpenAI designed for fast and efficient speech-to-text transcription.

    View details
    $0.02 / minute
    Chatterbox Nano logo

    Chatterbox Nano

    Text-to-Speech · Resemble AI

    Chatterbox Nano is a lightweight text-to-speech model from Resemble AI designed for fast, high-quality speech generation.

    View details
    $0.01 / minute
    GLiNER2.5 Multi v1 logo

    GLiNER2.5 Multi v1

    Text Classification · Fastino

    GLiNER2.5 Multi is a 287M-parameter multilingual information extraction model built on mDeBERTa-v3-base.

    View details
    $0.02 / 1M input$0.05 / 1M output
    deberta-v3-small logo

    deberta-v3-small

    Text Classification · Microsoft

    deberta-v3-small is Microsoft’s lightweight text classification model optimized for fast, low-cost zero-shot classification.

    View details
    $0.02 / 1M input$0.05 / 1M output
    all-minilm-l6-v2 logo

    all-minilm-l6-v2

    Text Embedding · Sentence Transformers

    all-MiniLM-L6-v2 is a lightweight sentence embedding model that converts sentences and short paragraphs into 384-dimensional dense vectors.

    View details
    $0.0040 / 1M input
    bge-small-en-v1.5 logo

    bge-small-en-v1.5

    Text Embedding · BAAI

    bge-small-en-v1.5 is a lightweight English embedding model from BAAI designed for semantic search, retrieval, similarity scoring, ranking, and RAG workflows.

    View details
    $0.0040 / 1M input

    Open-weight and frontier models (12)

    deepseek-v4.1-flash logo

    deepseek-v4.1-flash

    Text Generation · DeepSeek

    DeepSeek V4.1 Flash is an open-weight sparse mixture-of-experts model and the first built on DeepSeek's Causal Encoder-Decoder (CED) architecture, activating 8B parameters on input and 16B on output.

    View details
    $0.30 / 1M input$1.20 / 1M output
    glm-5.3-flash logo

    glm-5.3-flash

    Text Generation · Z.ai

    GLM-5.3-Flash is Z.ai's efficient open-weight model for coding and long-horizon agent tasks.

    View details
    $0.10 / 1M input$0.35 / 1M output
    gpt-5.6-luna logo

    gpt-5.6-luna

    Text Generation · OpenAI

    GPT-5.6 Luna is the cost-optimized model of OpenAI's GPT-5.6 family, designed for cost-sensitive, high-volume workloads.

    View details
    $0.20 / 1M input$1.20 / 1M output
    gpt-5.4-nano logo

    gpt-5.4-nano

    Text Generation · OpenAI

    GPT-5.4 nano is the most cost-efficient model in OpenAI's GPT-5.4 family, built for high-volume and latency-sensitive workloads such as classification, extraction, routing and sub-agent tasks.

    View details
    $0.20 / 1M input$1.25 / 1M output
    gpt-oss-120b logo

    gpt-oss-120b

    Text Generation · OpenAI

    gpt-oss-120b is OpenAI’s open-weight reasoning model built for advanced text generation, coding, research, and agentic workflows.

    View details
    $0.15 / 1M input$0.60 / 1M output
    qwen3-30b-a3b-fp8 logo

    qwen3-30b-a3b-fp8

    Text Generation · Qwen

    qwen3-30b-a3b-fp8 is Qwen’s open-weight mixture-of-experts model designed for efficient reasoning, coding, multilingual text generation, and agentic workflows.

    View details
    $0.10 / 1M input$0.45 / 1M output
    llama-3.1-8b-instruct-fast logo

    llama-3.1-8b-instruct-fast

    Summarization · Meta

    llama-3.1-8b-instruct-fast is Meta’s Llama 3.1 8B Instruct model optimized by ZeroGPU for fast, cost-efficient summarization and text processing at scale.

    View details
    $0.15 / 1M input$0.28 / 1M output
    llama-guard-4-12b logo

    llama-guard-4-12b

    Text Generation · Meta

    Llama Guard 4 12B is Meta’s multimodal safety classification model for moderating text, images, and mixed text-image inputs.

    View details
    $0.18 / 1M input$0.18 / 1M output
    LFM2.5-1.2B-Thinking logo

    LFM2.5-1.2B-Thinking

    Text Generation · Liquid AI

    LFM2.5-1.2B-Thinking is Liquid AI’s compact 1.2B-parameter reasoning model designed for multi-step problem solving, planning, data extraction, and agentic workflows.

    View details
    $0.02 / 1M input$0.05 / 1M output
    LFM2.5-1.2B-Instruct logo

    LFM2.5-1.2B-Instruct

    Text Generation · Liquid AI

    LFM2.5-1.2B-Instruct is Liquid AI’s compact 1.2B-parameter instruction model built for efficient text generation, conversational AI, and lightweight agent workflows.

    View details
    $0.02 / 1M input$0.05 / 1M output
    Qwen3.6 35B A3B logo

    Qwen3.6 35B A3B

    Text Generation · Qwen

    Qwen3.6-35B-A3B is a fast, open-weight multimodal model for applications that need to understand text, images, and video.

    View details
    $0.20 / 1M input$1.50 / 1M output