ZeroGPU sells to AI agents - no signup

The agent storefront at https://agents.zerogpu.ai/llms.txt publishes prices, payment instructions, and setup. Agents can purchase autonomously or with their human's approval, with a range of payment methods; an agent without a payment method can set one up there instantly, no signup. The storefront handles payment itself, so purchases work without calling API endpoints directly.

    ZeroGPU ZLM models (2)

    zlm-v1-signal-extract logo

    zlm-v1-signal-extract

    Text Classification · ZeroGPU

    zlm-v1-signal-extract is a specialized ZeroGPU model for turning unstructured text into structured signals that downstream systems can act on.

    View details
    $0.04 / 1M input$0.10 / 1M output
    zlm-v1-iab-domain-classifier logo

    zlm-v1-iab-domain-classifier

    Text Classification · ZeroGPU

    zlm-v1-iab-domain-classifier is a low-latency domain classification model built for advertising, contextual intelligence, and large-scale enrichment workflows.

    View details
    $0.02 / 1M input$0.05 / 1M output

    Open-weight and frontier models (9)

    deepseek-v4.1-flash logo

    deepseek-v4.1-flash

    Text Generation · DeepSeek

    DeepSeek V4.1 Flash is an open-weight sparse mixture-of-experts model and the first built on DeepSeek's Causal Encoder-Decoder (CED) architecture, activating 8B parameters on input and 16B on output.

    View details
    $0.30 / 1M input$1.20 / 1M output
    gpt-5.6-luna logo

    gpt-5.6-luna

    Text Generation · OpenAI

    GPT-5.6 Luna is the cost-optimized model of OpenAI's GPT-5.6 family, designed for cost-sensitive, high-volume workloads.

    View details
    $0.20 / 1M input$1.20 / 1M output
    gpt-5.4-nano logo

    gpt-5.4-nano

    Text Generation · OpenAI

    GPT-5.4 nano is the most cost-efficient model in OpenAI's GPT-5.4 family, built for high-volume and latency-sensitive workloads such as classification, extraction, routing and sub-agent tasks.

    View details
    $0.20 / 1M input$1.25 / 1M output
    gpt-oss-120b logo

    gpt-oss-120b

    Text Generation · OpenAI

    gpt-oss-120b is OpenAI’s open-weight reasoning model built for advanced text generation, coding, research, and agentic workflows.

    View details
    $0.15 / 1M input$0.60 / 1M output
    qwen3-30b-a3b-fp8 logo

    qwen3-30b-a3b-fp8

    Text Generation · Qwen

    qwen3-30b-a3b-fp8 is Qwen’s open-weight mixture-of-experts model designed for efficient reasoning, coding, multilingual text generation, and agentic workflows.

    View details
    $0.10 / 1M input$0.45 / 1M output
    llama-3.1-8b-instruct-fast logo

    llama-3.1-8b-instruct-fast

    Summarization · Meta

    llama-3.1-8b-instruct-fast is Meta’s Llama 3.1 8B Instruct model optimized by ZeroGPU for fast, cost-efficient summarization and text processing at scale.

    View details
    $0.15 / 1M input$0.28 / 1M output
    LFM2.5-1.2B-Instruct logo

    LFM2.5-1.2B-Instruct

    Text Generation · Liquid AI

    LFM2.5-1.2B-Instruct is Liquid AI’s compact 1.2B-parameter instruction model built for efficient text generation, conversational AI, and lightweight agent workflows.

    View details
    $0.02 / 1M input$0.05 / 1M output
    Qwen3.6 35B A3B logo

    Qwen3.6 35B A3B

    Text Generation · Qwen

    Qwen3.6-35B-A3B is a fast, open-weight multimodal model for applications that need to understand text, images, and video.

    View details
    $0.20 / 1M input$1.50 / 1M output