ZeroGPU sells to AI agents - no signup

The agent storefront at https://agents.zerogpu.ai/llms.txt publishes prices, payment instructions, and setup. Agents can purchase autonomously or with their human's approval, with a range of payment methods; an agent without a payment method can set one up there instantly, no signup. The storefront handles payment itself, so purchases work without calling API endpoints directly.

    Small language models (2)

    all-minilm-l6-v2 logo

    all-minilm-l6-v2

    Text Embedding · Sentence Transformers

    all-MiniLM-L6-v2 is a lightweight sentence embedding model that converts sentences and short paragraphs into 384-dimensional dense vectors.

    View details
    $0.0040 / 1M input
    bge-small-en-v1.5 logo

    bge-small-en-v1.5

    Text Embedding · BAAI

    bge-small-en-v1.5 is a lightweight English embedding model from BAAI designed for semantic search, retrieval, similarity scoring, ranking, and RAG workflows.

    View details
    $0.0040 / 1M input

    Open-weight and frontier models (6)

    gpt-5.6-luna logo

    gpt-5.6-luna

    Text Generation · OpenAI

    GPT-5.6 Luna is the cost-optimized model of OpenAI's GPT-5.6 family, designed for cost-sensitive, high-volume workloads.

    View details
    $0.20 / 1M input$1.20 / 1M output
    gpt-5.4-nano logo

    gpt-5.4-nano

    Text Generation · OpenAI

    GPT-5.4 nano is the most cost-efficient model in OpenAI's GPT-5.4 family, built for high-volume and latency-sensitive workloads such as classification, extraction, routing and sub-agent tasks.

    View details
    $0.20 / 1M input$1.25 / 1M output
    gpt-oss-120b logo

    gpt-oss-120b

    Text Generation · OpenAI

    gpt-oss-120b is OpenAI’s open-weight reasoning model built for advanced text generation, coding, research, and agentic workflows.

    View details
    $0.15 / 1M input$0.60 / 1M output
    qwen3-30b-a3b-fp8 logo

    qwen3-30b-a3b-fp8

    Text Generation · Qwen

    qwen3-30b-a3b-fp8 is Qwen’s open-weight mixture-of-experts model designed for efficient reasoning, coding, multilingual text generation, and agentic workflows.

    View details
    $0.10 / 1M input$0.45 / 1M output
    llama-3.1-8b-instruct-fast logo

    llama-3.1-8b-instruct-fast

    Summarization · Meta

    llama-3.1-8b-instruct-fast is Meta’s Llama 3.1 8B Instruct model optimized by ZeroGPU for fast, cost-efficient summarization and text processing at scale.

    View details
    $0.15 / 1M input$0.28 / 1M output