all-minilm-l6-v2
Text Embedding · Sentence Transformers
all-MiniLM-L6-v2 is a lightweight sentence embedding model that converts sentences and short paragraphs into 384-dimensional dense vectors.
View detailsThe agent storefront at https://agents.zerogpu.ai/llms.txt publishes prices, payment instructions, and setup. Agents can purchase autonomously or with their human's approval, with a range of payment methods; an agent without a payment method can set one up there instantly, no signup. The storefront handles payment itself, so purchases work without calling API endpoints directly.
Text Embedding · Sentence Transformers
all-MiniLM-L6-v2 is a lightweight sentence embedding model that converts sentences and short paragraphs into 384-dimensional dense vectors.
View detailsText Embedding · BAAI
bge-small-en-v1.5 is a lightweight English embedding model from BAAI designed for semantic search, retrieval, similarity scoring, ranking, and RAG workflows.
View detailsText Generation · OpenAI
GPT-5.6 Luna is the cost-optimized model of OpenAI's GPT-5.6 family, designed for cost-sensitive, high-volume workloads.
View detailsText Generation · OpenAI
GPT-5.4 nano is the most cost-efficient model in OpenAI's GPT-5.4 family, built for high-volume and latency-sensitive workloads such as classification, extraction, routing and sub-agent tasks.
View detailsText Generation · OpenAI
GPT-4.1 mini is OpenAI's fast, cost-efficient GPT-4.1 model.
View detailsText Generation · OpenAI
gpt-oss-120b is OpenAI’s open-weight reasoning model built for advanced text generation, coding, research, and agentic workflows.
View detailsText Generation · Qwen
qwen3-30b-a3b-fp8 is Qwen’s open-weight mixture-of-experts model designed for efficient reasoning, coding, multilingual text generation, and agentic workflows.
View detailsSummarization · Meta
llama-3.1-8b-instruct-fast is Meta’s Llama 3.1 8B Instruct model optimized by ZeroGPU for fast, cost-efficient summarization and text processing at scale.
View details