Infron AI Model Marketplace

Browse 445 entries with direct links to detailed pricing, capabilities, and provider information.

  • Moonshot: Kimi K2.6 (free)

    Note: For the free endpoint, all prompts and outputs are logged to help improve the provider’s model, products, and services. This endpoint is provided for t...

  • inclusionAI: Ling-2.6-1T

    Ling-2.6-1T: A Trillion-Parameter Comprehensive Flagship Model for Complex Tasks. Tailored for real–world, complex scenarios, this trillion–parameter model i...

  • Interfaze: Interfaze Beta

    Interfaze is an AI model built on a new architecture that merges specialized DNN/CNN models with LLMs for developer tasks that require deterministic output a...

  • Google: Gemini 3.1 flash lite

    Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio...

  • Qwen: Qwen3.5 27B earica Derestricted Lite

    Lighter earica finetune for responsive creative chat, roleplay, and iterative writing.

  • Qwen: Qwen3.5 27B earica Derestricted

    earica finetune for derestricted creative writing, dialogue, and roleplay.

  • Qwen: Qwen3.5 27B RpRMax V1

    RpRMax v1 finetune for roleplay-focused Qwen3.5 27B conversations and story generation.

  • Qwen: Qwen3.5 27B Queen Derestricted Lite

    Lighter Queen finetune for responsive creative chat, roleplay, and scene iteration.

  • Qwen: Qwen3.5 27B Queen Derestricted

    Queen finetune for derestricted creative writing, roleplay, and expressive character dialogue.

  • Qwen: Qwen3.5 27B NaNovel V2 Derestricted Lite

    Lighter NaNovel finetune for fast creative drafting, dialogue, and roleplay.

  • Qwen: Qwen3.5 27B NaNovel V2 Derestricted

    NaNovel finetune for derestricted novel-style prose, character writing, and long-form scenes.

  • Qwen: Qwen3.5 27B Marvin V2 Derestricted Lite

    Lighter Marvin V2 finetune for responsive creative chat, dialogue, and scene drafting.

  • Qwen: Qwen3.5 27B Marvin V2 Derestricted

    Marvin V2 finetune for derestricted creative writing, character voice, and multimodal chat.

  • Qwen: Qwen3.5 27B Marvin DPO V2 Derestricted Lite

    Lighter Marvin DPO V2 finetune for responsive creative chat and iterative roleplay.

  • Qwen: Qwen3.5 27B Marvin DPO V2 Derestricted

    Marvin DPO V2 finetune for derestricted creative writing, dialogue, and roleplay.

  • Qwen: Qwen3.5 27B Infracelestial

    Qwen3.5 27B Infracelestial finetune for expressive creative chat and long-form roleplay.

  • Qwen: Qwen3.5 27B Anko

    Qwen3.5 27B Anko finetune for creative writing, roleplay, and multimodal chat.

  • Qwen: Qwen3.5 27B BlueStar v3 Derestricted Lite

    Lighter third-generation BlueStar finetune for responsive creative chat, roleplay, and scene drafting.

  • Qwen: Qwen3.5 27B BlueStar v3 Derestricted

    Third-generation BlueStar finetune for creative roleplay, narrative prose, and multimodal Qwen3.5 27B workflows.

  • Rednote-Hilab: Dots.ocr

    dots.ocr is a powerful, multilingual document parser that unifies layout detection and content recognition within a single vision-language model while mainta...

  • xAI: Grok 4.3

    Grok 4.3 is a reasoning model from xAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, ...

  • MiroMind: mirothinker-1-7-deepresearch-mini

    Mid-tier · 30B agent · cost-effective

  • MiroMind: mirothinker-1-7-deepresearch

    Flagship · 235B agent · top-quality research.

  • NVIDIA: Nemotron 3 Nano Omni (free)

    Note: For the free endpoint, all prompts and outputs are logged to help improve the provider’s model, products, and services. This endpoint is provided for t...

  • Qwen: Qwen 3.6 Flash

    The Qwen3.6 native vision-language Flash model series delivers a significant performance boost over the 3.5-Flash version. This model particularly excels in ...

  • OpenAI: GPT-5.5 Pro

    GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context wi...

  • OpenAI: GPT-5.5

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved...

  • DeepSeek: Deepseek V4 Pro

    DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token contex...

  • DeepSeek: Deepseek V4 Flash

    284B total / 13B active params. Your fast, efficient, and economical choice.

  • Qwen: Qwen3.6-27B

    The Qwen3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM re...

  • Xiaomi: Mimo v2.5

    MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni ...

  • Xiaomi: Mimo v2.5 Pro

    MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks...

  • inclusionAI: Ling-2.6-flash

    Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that req...

  • Moonshot: Kimi K2.6

    Kimi K2.6 is an open-source, native multimodal agentic model that significantly advances practical capabilities in long-horizon coding, coding-driven design,...

  • Qwen: Qwen3.6 Max Preview

    The Max model, the largest and most capable variant in the Qwen3.6 series, is now available in a preview version. At present, only its plain-text capabilitie...

  • Qwen: Qwen3.6-35B-A3B

    The Qwen3.6 35B-A3B native vision-language model is built on a hybrid architecture that integrates linear attention mechanisms with a sparse mixture-of-exper...

  • Anthropic: Claude Opus 4.7

    Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus...

  • Z.AI: GLM 5.1

    GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around m...

  • Google: Gemma 4 26B A4B

    Gemma 4 26B A4B is built for developers who need scalable performance without sacrificing core capabilities.Crucially, it retains the massive 256K-token cont...

  • Google: Gemma 4 31B

    Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window...

  • Holo: Holo3-35B-A3B

    Holo3-35B-A3B is our efficient Action Vision-Language Model. With only 3B active parameters, it delivers near-flagship performance at dramatically lower late...

  • Qwen: Qwen 3.6 Plus

    The Qwen 3.6 Native Vision-Language Series Plus models demonstrate exceptional performance comparable to today's cutting-edge models, marking a significant i...

  • xAI: Grok 4.2 Non Reasoning

    Grok 4.20 is xAI's newest flagship model with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the ...

  • xAI: Grok 4.2 Reasoning

    Grok 4.20 is xAI's newest flagship model with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the ...

  • Arcee AI: Trinity Large Thinking

    Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and...

  • Z.AI: GLM 5V Turbo

    GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video...

  • Kwaipilot: KAT-Coder-Pro V2

    KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS inte...

  • ByteDance: Seed 2.0 Pro

    Built for the Agent era, it delivers stable performance in complex reasoning and long-horizon tasks, including multi-step planning, visual-text reasoning, vi...

  • MiniMax: MiniMax M2.7

    MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively partici...

  • OpenAI: GPT-5.4 Nano

    GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text a...

  • OpenAI: GPT-5.4 Mini

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image in...

  • Mistral: Mistral Small 4

    Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It ...

  • OpenAI: Gpt oss safeguard 20b

    gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers l...

  • Z.AI: GLM 5 Turbo

    GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply...

  • Qwen: Qwen3.5-9B

    Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9...

  • ByteDance: Seed 2.0 Lite

    ByteDance-Seed-2.0-lite is a balanced model designed for high-frequency enterprise workloads, optimizing for both capability and cost. Its overall performanc...

  • OpenAI: GPT-5.4

    GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K ou...

  • Google: Gemini 3.1 flash lite preview

    Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality ...

  • Qwen: Tongyi DeepResearch 30B A3B

    Tongyi DeepResearch is an agentic large language model developed by Tongyi Lab, with 30 billion total parameters activating only 3 billion per token. It's op...

  • Qwen: Qwen3 VL 235B A22B Instruct

    Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Ins...

  • Qwen: Qwen3 VL 235B A22B Thinking

    Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model i...

  • Inception: Mercury 2

    Mercury 2 doesn't decode sequentially. It generates responses through parallel refinement, producing multiple tokens simultaneously and converging over a sma...

  • Qwen: Qwen3.5-397B-A17B

    The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixtur...

  • Qwen: Qwen3.5-122B-A10B

    The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-ex...

  • Qwen: Qwen3.5-35B-A3B

    The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mix...

  • Qwen: Qwen3.5-27B

    The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed ...

  • Qwen: Qwen 3.5 Flash

    The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-expe...

  • ByteDance: Seed 2.0 Mini

    ByteDance-Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deploymen...

  • Qwen: Qwen 3.5 Plus

    The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-e...

  • Google: Gemini 3.1 pro preview

    Gemini 3.1 Pro is designed to tackle the most challenging agentic problems with strong coding and state-of-the-art reasoning capabilities as well as complex ...

  • Google: Gemini 3 flash preview

    Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pr...

  • Anthropic: Claude Opus 4.5

    Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offe...

  • Anthropic: Claude Opus 4.6

    Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather th...

  • OpenAI: GPT-4o Mini TTS (Text to Audio)

    GPT-4o Mini TTS is OpenAI's cost-efficient text-to-speech model. It converts text input into natural-sounding audio output, supporting a variety of voices an...

  • OpenAI: GPT-4o Transcribee (Audio to Text)

    The gpt-4o-transcribe model is a cutting-edge speech-to-text solution that leverages the advanced capabilities of GPT-4o to deliver highly accurate audio tra...

  • OpenAI: GPT-4o Mini Transcribe (Audio to Text)

    The gpt-4o-mini-transcribe model is a highly efficient speech-to-text solution designed to deliver accurate audio transcriptions while optimizing for speed a...

  • OpenAI: Gpt 4o mini

    GPT-4o mini is OpenAI's newest model after GPT-4 Omni, supporting both text and image inputs with text outputs. As their most advanced small model, it is man...

  • OpenAI: Gpt 4o

    OpenAI ChatGPT 4o is continually updated by OpenAI to point to the current version of GPT-4o used by ChatGPT. It therefore differs slightly from the API vers...

  • OpenAI: Gpt 4.1 mini

    GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token contex...

  • OpenAI: Gpt 4.1

    GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supp...

  • OpenAI: Gpt 4

    OpenAI's flagship model, GPT-4 is a large-scale multimodal language model capable of solving difficult problems with greater accuracy than previous models du...

  • OpenAI: Gpt 35 turbo

    This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost....

  • OpenAI: Gpt 4.1 nano

    For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size ...

  • OpenAI: Gpt 5 nano

    GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. Wh...

  • OpenAI: gpt 5 mini

    GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning ben...

  • OpenAI: Gpt oss 20b

    gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B ...

  • OpenAI: gpt oss 120b

    gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose ...

  • OpenAI: Gpt 5

    GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that re...

  • OpenAI: Gpt 5 pro

    GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks tha...

  • OpenAI: Gpt 5 codex

    GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessio...

  • OpenAI: Gpt 5.1

    GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natur...

  • OpenAI: Gpt 5.1 codex

    GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development se...

  • OpenAI: Gpt 5.1 codex mini

    GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

  • OpenAI: Gpt 5.1 chat

    GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It u...

  • Qwen: Qwen3 Coder Plus

    Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous pr...

  • Qwen: Qwen3 Coder Flash

    Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in aut...

  • Qwen: Qwen3 0.6B

    Achieves effective integration of reasoning and non-reasoning modes, allowing seamless switching during conversations. The model delivers state-of-the-art (S...

  • Qwen: Qwen3 8B

    Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports s...

  • Qwen: Qwen3 14B

    Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built up...

  • Qwen: Qwen3 30B A3B

    Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, m...

  • Qwen: Qwen3 32B

    Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports se...

  • Qwen: Qwen3 235B A22B

    Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switch...

  • Qwen: Qwen3 30B A3B Instruct 2507

    Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-t...

  • Qwen: Qwen3 30B A3B Thinking 2507

    Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The mod...

  • Qwen: Qwen3 235B A22B Instruct 2507

    Qwen3-235B-A22B-Instruct-2507 has the following features: Type: Causal Language Models Training Stage: Pretraining & Post-training Number of Parameters: 235B...

  • Qwen: Qwen3 235B A22B Thinking 2507

    Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates ...

  • Qwen: Qwen3 Next 80B A3B Instruct

    Qwen3-Next-80B-A3B-Instruct is an 80B parameter model with only 3B activated per token that excels at handling ultra-long contexts up to 256K tokens natively...

  • Qwen: Qwen3 Next 80B A3B Thinking

    Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for ha...

  • Qwen: Qwen2.5 7B Instruct 1M

    Qwen2.5-1M is the long-context version of the Qwen2.5 series models, supporting a context length of up to 1M tokens. Compared to the Qwen2.5 128K version, Qw...

  • Qwen: Qwen2.5 7B Instruct

    Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ra...

  • Tavily: Tavily Search

    Tavily is a search engine optimized for LLMs, aimed at efficient, quick and persistent search results. Unlike other search APIs such as Serp or Google, Tavil...

  • Tavily: Tavily Extract

    Extract web page content from one or more specified URLs using Tavily Extract.

  • Anthropic: Claude Sonnet 4.5

    Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art perfo...

  • OpenAI: GPT 5 chat

    GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that re...

  • Voyage: voyage 4 lite

    Optimized for latency and cost. All embeddings created with the 4 series are compatible with each other

  • Moonshot: Kimi K2.5

    Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on K...

  • Z.AI: GLM 4.7

    GLM-4.7 is Z.AI’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/executio...

  • MiniMax: Minimax M2.1

    MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 1...

  • Meta: Llama 3.1 8B

    Meta's Llama 3.1 is a major upgrade, introducing new model sizes. The flagship 405B model rivals top closed-source AI like GPT-4o, while a new, highly effici...

  • Z.AI: GLM 4.7 Flash

    As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, ...

  • NVIDIA: Nemotron 3 Nano 30B A3B

    NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI syst...

  • MythoMax L2 13B

    One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay.

  • Nous: Hermes 3 Llama 3.1 405B

    Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, mu...

  • Nous: Hermes 3 Llama 3.1 70B

    Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, mu...

  • Anthropic: Claude Opus 4.1

    Claude Opus 4-1 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent ...

  • Anthropic: Claude Haiku 4.5

    Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claud...

  • Google: Gemini 2.5 flash

    Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It in...

  • Google: Gemini 2.5 flash lite preview 06 17

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved thro...

  • Google: Gemini 2.5 flash lite preview 09 2025

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved thro...

  • Google: Gemini 2.5 flash preview 05 20

    Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It in...

  • Google: Gemini 2.5 flash image preview

    Gemini 2.5 Flash Image Preview, AKA Nano Banana is a state of the art image generation model with contextual understanding. It is capable of image generation...

  • Google: Gemini 2.5 flash image

    Gemini 2.5 Flash Image, AKA Nano Banana is a state of the art image generation model with contextual understanding. It is capable of image generation, edits,...

  • Google: Gemini 2.5 flash preview

    Gemini 2.5 Flash is our best model in terms of price and performance, and offers well-rounded capabilities. Gemini 2.5 Flash is our first Flash model model t...

  • Google: Gemini 2.5 flash lite

    Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved thro...

  • Google: Gemini 2.5 pro

    Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabi...

  • Google: Gemini 2.5 pro preview 06 05

    Gemini 2.5 Pro is our most advanced reasoning Gemini model, capable of solving complex problems. Gemini 2.5 Pro can comprehend vast datasets and challenging ...

  • Google: Gemini 2.5 pro preview 05 06

    Gemini 2.5 Pro is our most advanced reasoning Gemini model, capable of solving complex problems. Gemini 2.5 Pro can comprehend vast datasets and challenging ...

  • Google: Gemini 2.5 pro preview 03 25

    Gemini 2.5 Pro is our most advanced reasoning Gemini model, capable of solving complex problems. Gemini 2.5 Pro can comprehend vast datasets and challenging ...

  • Z.AI: GLM 5

    GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it...

  • Qwen: Qwen3 Reranker 0.6B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upo...

  • Jina: Jina Reranker V3

    jina-reranker-v3 is a 0.6B parameter multilingual document reranker introducing a novel last but not late interaction architecture. Unlike ColBERT's separate...

  • OpenAI: O1

    The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-sc...

  • Qwen: Qwen3 Coder 480B A35B

    Qwen3-Coder-480B-A35B-Instruct is a cutting-edge open coding model from Qwen, matching Claude Sonnet’s performance in agentic programming, browser automation...

  • MiniMax: MiniMax M2.5

    MiniMax M2.5 is SOTA in coding, agentic tool use and search, office work, and a range of other economically valuable tasks, boasting scores of 80.2% in SWE-B...

  • MoonshotAI: Kimi K2 0905

    Kimi K2 0905 is the September update of Kimi K2 0711. It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trill...

  • Z.AI: GLM 4.6

    GLM-4.6 achieves comprehensive enhancements across multiple domains, including real-world coding, long-context processing, reasoning, searching, writing, and...

  • Z.AI: GLM 4.6V

    GLM-4.6 achieves comprehensive enhancements across multiple domains, including real-world coding, long-context processing, reasoning, searching, writing, and...

  • Z.AI: GLM 4.6V FlashX

    GLM-4.6 achieves comprehensive enhancements across multiple domains, including real-world coding, long-context processing, reasoning, searching, writing, and...

  • Z.AI: GLM 4.6V Flash

    GLM-4.6 achieves comprehensive enhancements across multiple domains, including real-world coding, long-context processing, reasoning, searching, writing, and...

  • Z.AI: GLM 4.5

    GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Expe...

  • Z.AI: GLM 4.5V

    GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Expe...

  • Z.AI: GLM 4.5X

    GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Expe...

  • Z.AI: GLM 4.5 Air

    GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Expe...

  • Z.AI: GLM 4.5 Flash

    GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Expe...

  • Z.AI: GLM 4 32B 0414 128K

    GLM 4 32B is a cost-effective foundation language model. It can efficiently perform complex tasks and has significantly enhanced capabilities in tool use, on...

  • Moonshot: Kimi K2 Thinking

    Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the tril...

  • Moonshot: Kimi V1 32K

    Kimi v1 32k is a large‑language model in Moonshot AI’s moonshot‑v1 series that offers a 32 k token context window, enabling it to ingest and generate up to r...

  • Moonshot: Kimi V1 128K

    Kimi v1 (also marketed as the moonshot‑v1 family) is a large‑scale Transformer‑based language model developed by Moonshot AI that pushes the frontier of long...

  • ByteDance: Seed 1.8

    A brand-new model optimized specifically for multimodal agent scenarios. It features enhanced agent capabilities, upgraded multimodal comprehension, and more...

  • ByteDance: Seed 1.6

    Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K conte...

  • ByteDance: Seed 1.6 Flash

    Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context w...

  • Voyage: voyage 4 Large

    The best general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other

  • Voyage: voyage 4

    Optimized for general-purpose and multilingual retrieval quality. All embeddings created with the 4 series are compatible with each other

  • Voyage: voyage 3 Large

    The best general-purpose and multilingual retrieval quality.

  • Voyage: voyage 3.5

    Optimized for general-purpose and multilingual retrieval quality.

  • Voyage: voyage 3.5 Lite

    Optimized for latency and cost.

  • Voyage: voyage code 3

    Optimized for code retrieval.

  • Voyage: voyage Finance 2

    Optimized for finance retrieval and RAG.

  • Voyage: voyage Law 2

    Optimized for legal retrieval and RAG. Also improved performance across all domains.

  • Voyage: voyage Code 2

    Optimized for code retrieval (17% better than alternatives) / Previous generation of code embeddings.

  • Voyage: voyage Rerank 2.5

    Our generalist reranker optimized for quality with instruction-following and multilingual support.

  • Voyage: voyage Rerank 2.5 Lite

    Our generalist reranker optimized for both latency and quality with instruction-following and multilingual support.

  • Mistral: Mistral Nemo

    A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, ...

  • DeepSeek: Deepseek OCR

    DeepSeek-OCR is a comprehensive Optical Character Recognition (OCR) model that analyzes and understands complex documents. It excels at challenging OCR tasks...

  • Google: Gemma 3 4B IT

    Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 langu...

  • Google: Gemma 3 12B IT

    Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 langu...

  • Google: Gemma 3 27B IT

    Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 langu...

  • Jina: Jina Deepsearch V1

    DeepSearch is an LLM API that performs iterative search, reading, and reasoning until it finds an accurate answer to a query or reaches its token budget limit.

  • Jina: Jina Embeddings V4

    Jina Embeddings V4 is a 3.8 billion parameter multimodal embedding model that provides unified text and image representation capabilities. Built on the Qwen2...

  • Jina: Jina Clip V2

    Jina CLIP v2 revolutionizes multimodal AI by bridging the gap between visual and textual understanding across 89 languages. This model solves critical challe...

  • Jina: Jina Embeddings V3

    Jina Embeddings v3 is a groundbreaking multilingual text embedding model that transforms how organizations handle text understanding and retrieval across lan...

  • Jina: Jina Colbert V2

    Jina-ColBERT-v2 is a groundbreaking multilingual information retrieval model that solves the critical challenge of efficient, high-quality search across mult...

  • Jina: Jina Clip V1

    Jina CLIP v1 revolutionizes multimodal AI by being the first model to excel equally in both text-to-text and text-to-image retrieval tasks. Unlike traditiona...

  • Jina: Jina Colbert V1 EN

    Jina-ColBERT-v1-en revolutionizes text search by solving a critical challenge in information retrieval: achieving high accuracy without sacrificing computati...

  • Jina: Jina Embeddings V2 Base ES

    Jina Embeddings v2 Base Spanish is a groundbreaking bilingual text embedding model that addresses the critical challenge of cross-lingual information retriev...

  • Jina: Jina Embeddings V2 Base Code

    Jina Embeddings v2 Base Code tackles a critical challenge in modern software development: efficiently navigating and understanding large codebases. For devel...

  • Jina: Jina Embeddings V2 Base DE

    Jina Embeddings v2 Base German addresses a critical challenge in international business: bridging the language gap between German and English markets. For Ge...

  • Jina: Jina Embeddings V2 Base ZH

    Jina Embeddings v2 Base Chinese breaks new ground as the first open-source model to seamlessly handle both Chinese and English text with an unprecedented 8,1...

  • Jina: Jina Embeddings V2 Base EN

    Jina Embeddings v2 Base English is a groundbreaking open-source text embedding model that solves the critical challenge of processing long documents while ma...

  • Jina: Jina Reranker m0

    jina-reranker-m0 is a groundbreaking multimodal multilingual reranker model designed to rank visual documents across multiple languages. What makes this mode...

  • Jina: Jina Reranker V2 Base Multilingual

    Jina Reranker v2 Base Multilingual is a cross-encoder model designed to enhance search accuracy across language barriers and data types. This reranker addres...

  • Jina: Jina Reranker V1 Tiny EN

    Jina Reranker v1 Tiny English represents a breakthrough in efficient search refinement, designed specifically for organizations requiring high-performance re...

  • Jina: Jina Reranker V1 Turbo EN

    Jina Reranker v1 Turbo English addresses a critical challenge in production search systems: the trade-off between result quality and computational efficiency...

  • Jina: Jina Reranker V1 Base EN

    Jina Reranker v1 Base English revolutionizes search result refinement by addressing a critical limitation in traditional vector search systems: the inability...

  • OpenAI: Text Embedding Ada 002

    text-embedding-ada-002 is OpenAI's legacy text embedding model.

  • OpenAI: Text Embedding 3 Small

    text-embedding-3-small is OpenAI's improved, more performant version of the ada embedding model. Embeddings are a numerical representation of text that can b...

  • OpenAI: Text Embedding 3 Large

    text-embedding-3-large is OpenAI's most capable embedding model for both english and non-english tasks. Embeddings are a numerical representation of text tha...

  • Google: Gemini Embedding 001

    gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consi...

  • Google: Text Embedding Large Exp 0307

    text-embedding-large-exp-03-07

  • Google: Text Embedding 005

    text-embedding-005

  • Thenlper: Gte Base

    The gte-base embedding model encodes English sentences and paragraphs into a 768-dimensional dense vector space, delivering efficient and effective semantic ...

  • Thenlper: Gte Large

    The gte-large embedding model converts English sentences, paragraphs and moderate-length documents into a 1024-dimensional dense vector space, delivering hig...

  • Intfloat: E5 Large V2

    The e5-large-v2 embedding model maps English sentences, paragraphs, and documents into a 1024-dimensional dense vector space, delivering high-accuracy semant...

  • Intfloat: E5 Base V2

    The e5-base-v2 embedding model encodes English sentences and paragraphs into a 768-dimensional dense vector space, producing efficient and high-quality seman...

  • Intfloat: Multilingual E5 Large

    The multilingual-e5-large embedding model encodes sentences, paragraphs, and documents across over 90 languages into a 1024-dimensional dense vector space, d...

  • Sentence Transformers: Paraphrase Minilm L6 V2

    The paraphrase-MiniLM-L6-v2 embedding model converts sentences and short paragraphs into a 384-dimensional dense vector space, producing high-quality semanti...

  • Sentence Transformers: All Minilm L12 V2

    The all-MiniLM-L12-v2 embedding model maps sentences and short paragraphs into a 384-dimensional dense vector space, producing efficient and high-quality sem...

  • BAAI: Bge Base En V1.5

    The bge-base-en-v1.5 embedding model converts English sentences and paragraphs into 768-dimensional dense vectors, delivering efficient, high-quality semanti...

  • Sentence Transformers: Multi Qa Mpnet Base Dot V1

    The multi-qa-mpnet-base-dot-v1 embedding model transforms sentences and short paragraphs into a 768-dimensional dense vector space, generating high-quality s...

  • BAAI: Bge Large En V1.5

    The bge-large-en-v1.5 embedding model maps English sentences, paragraphs, and documents into a 1024-dimensional dense vector space, delivering high-fidelity ...

  • BAAI: Bge M3

    The bge-m3 embedding model encodes sentences, paragraphs, and long documents into a 1024-dimensional dense vector space, delivering high-quality semantic emb...

  • Sentence Transformers: All Mpnet Base V2

    The all-mpnet-base-v2 embedding model encodes sentences and short paragraphs into a 768-dimensional dense vector space, providing high-fidelity semantic embe...

  • Sentence Transformers: All Minilm L6 V2

    The all-MiniLM-L6-v2 embedding model maps sentences and short paragraphs into a 384-dimensional dense vector space, enabling high-quality semantic representa...

  • Mistral: Mistral Embed 2312

    Mistral Embed is a specialized embedding model for text data, optimized for semantic search and RAG applications. Developed by Mistral AI in late 2023, it pr...

  • Mistral: Codestral Embed 2505

    Mistral Codestral Embed is specially designed for code, perfect for embedding code databases, repositories, and powering coding assistants with state-of-the-...

  • Qwen: Qwen3 Embedding 4B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series ...

  • Qwen: Qwen3 Embedding 8B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series ...

  • Perplexity: Perplexity Search

    Get ranked search results from Perplexity’s continuously refreshed index with advanced filtering and customization options.

  • firecrawl: Firecrawl Search

    Search the web and get full content from results Firecrawl’s search API allows you to perform web searches and optionally scrape the search results in one op...

  • Exa: Exa Search

    By default, it automatically chooses the best search method using Exa’s embeddings-based model and other techniques to find the most relevant results for you...

  • Cloudsway: Cloudsway Smart Search

    Native semantic AI Search & Data service tailored to AI Agent. It addresses the intelligent retrieval needs of AI Agents across multiple modalities, language...

  • OpenAI: O1 Mini

    The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 models are optimized for math, scienc...

  • OpenAI: O3 Mini

    OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model ...

  • OpenAI: O4 Mini

    OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic cap...

  • Deepseek: Deepseek Chat

    DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on...

  • Deepseek: Deepseek R1

    DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active ...

  • Deepseek: Deepseek R1 Distill Llama 70B

    DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advance...

  • Deepseek: Deepseek Chat V3 0324

    DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the De...

  • Deepseek: Deepseek R1 0528

    May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens. It's 671B parameters in...

  • Deepseek: Deepseek Chat V3.1

    DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It ext...

  • Deepseek: Deepseek V3.1 Terminus

    DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including la...

  • Deepseek: Deepseek V3.1

    DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It ext...

  • Anthropic: Claude Sonnet 4.6

    Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative...

  • Qwen: Qwen3 Reranker 4B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upo...

  • Qwen: Qwen3 Reranker 8B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upo...

  • OpenAI: TTS-1 (Text to Audio)

    TTS is a model that converts text to natural sounding speech. TTS is optimized for realtime or interactive scenarios. For offline scenarios, TTS-HD provides ...

  • OpenAI: TTS-1-HD (Text to Audio)

    TTS-HD is a model that converts text to natural sounding speech. TTS is optimized for realtime or interactive scenarios. For offline scenarios, TTS-HD provid...

  • OpenAI: Whisper-1 (Audio to Text)

    The Whisper models are trained for speech recognition and translation tasks, capable of transcribing speech audio into the text in the language it is spoken ...

  • MiniMax: MiniMax M2

    MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (23...

  • Qwen: Qwen3 Max

    Qwen/qwen3-max, Enhanced with specialized upgrades in agent programming and tool calling. This official release achieves domain SOTA performance, supporting ...

  • Qwen: Qwen3 Coder 30b A3b Instruct

    Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code gen...

  • Qwen: Qwen3 Embeddings 0.6B

    The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. This series ...

  • Meta: Llama 3.3 70B Instruct

    This model delivers enhanced performance for chat, coding, instruction following, mathematics, and reasoning use cases.

  • Mistral: Mistral Small 3 24B Instruct 2501

    Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it fea...

  • Qwen: Qwen3 Omni 30B A3b Instruct

    The Qwen-Omni model accepts combined inputs of text and a single additional modality (image, audio, or video) to generate responses in text or speech. It off...

  • Qwen: Qwen3 Omni 30b A3b Thinking

    The Qwen-Omni model accepts combined inputs of text and a single additional modality (image, audio, or video) to generate responses in text or speech. It off...

  • Qwen: Qwen Mt Plus

    Qwen-MT is a large language model optimized for machine translation, built upon the foundation of the Tongyi Qianwen model. It supports translation across 92...

  • StepFun: Step 3.5 Flash

    Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only...

  • Qwen: Qwen Plus

    The enhanced version of the Qianwen ultra-large-scale language model supports input in different languages, including Chinese and English. Compared to previo...

  • Qwen: Qwen Plus Character

    The Thousand Questions series of role-playing models is a dynamically updated version. Model updates will be announced in advance. It is suitable for anthrop...

  • AionLabs: Aion 2.0

    A variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into ...

  • ByteDance: Skylark Pro SC 250615

    Roleplay model based on Skylark Pro.

  • Qwen: Qwen3 Coder Next

    Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B to...

  • Kwaipilot: KAT-Coder-Pro V1

    KAT-Coder-Pro V1 is KwaiKAT's most advanced agentic coding model in the KAT-Coder series. Designed specifically for agentic coding tasks, it excels in real-w...

  • Meituan: LongCat Flash Chat

    LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activat...

  • Mistral: Ministral 3 14B 2512

    The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B coun...

  • Moonshot: Kimi-K2-Instruct

    State-of-the-art mixture-of-experts agentic intelligence model with 1 T parameters, 128K context, and native tool use

  • Qwen: Qwen3 Coder

    Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.

  • ByteDance: Skylark Pro SC 260215

    Roleplay model based on Skylark Pro.

  • Qwen: Qwen3-VL-Plus

    The Qwen3 series of visual understanding models effectively integrates thinking and non-thinking modes, achieving world-class performance on public benchmark...

  • Qwen: Qwen Mt Lite

    Based on the fully upgraded Qwen3 basic text translation model, it supports mutual translation between 32 languages. The model performance and translation ef...

  • Qwen: Qwen Flash

    The Qwen3 series Flash models effectively blend thinking and non-thinking modes, allowing switching between modes during dialogue. They exhibit excellent per...

  • Qwen: Qwen3-VL-Flash

    The Qwen3 series of small-sized visual understanding models effectively integrates thinking and non-thinking modes, outperforming the open-source version Qwe...

  • MiniMax: MiniMax M2-her

    MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designe...

  • NVIDIA: Llama 3.1 Nemotron 70B Instruct

    NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforce...

  • NVIDIA: Nemotron 3 Super

    NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex mult...

  • NVIDIA: Llama 3.3 Nemotron Super 49B V1.5

    Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context. It...

  • NVIDIA: Nemotron Nano 12B 2 VL

    NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces...

  • NVIDIA: Nemotron Nano 9B V2

    NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoni...

  • Z.AI: GLM 4.7 FlashX

    Lightweight, high-speed GLM-4.7 variant.

  • Z.AI: GLM 4.5 AirX

    Enhanced GLM-4.5 Air variant.

  • Z.AI: GLM 4 32B

    GLM 4 32B is a cost-effective foundation language model. It can efficiently perform complex tasks and has significantly enhanced capabilities in tool use, on...

  • Sao10K: Llama 3.1 Euryale 70B v2.2

    Euryale L3.1 70B v2.2 is a model focused on creative roleplay from Sao10k. It is the successor of Euryale L3 70B v2.1.

  • Sao10K: Llama 3.3 Euryale 70B v2.3

    Euryale L3.3 70B is a model focused on creative roleplay from Sao10k. It is the successor of Euryale L3 70B v2.2.

  • VolcanoEngine: Doubao-Seed 1.8

    Doubao-Seed-1.8 is optimized for multimodal agent scenarios. In terms of agent capabilities, tool use and complex command compliance have been significantly ...

  • VolcanoEngine: Doubao-Seed 2.0 Code

    This coding model is optimized for real-world programming environments and can reliably call tools in common IDEs such as Claude Code. The model is specifica...

  • VolcanoEngine: Doubao-Seed 2.0 Mini

    Designed for low-latency, high-concurrency, and cost-sensitive scenarios, it delivers exceptional model inference speed. Model performance is comparable to D...

  • VolcanoEngine: Doubao-Seed 2.0 Lite

    This model offers a balanced approach to performance and cost for high-frequency enterprise scenarios, surpassing the capabilities of its predecessor, Doubao...

  • VolcanoEngine: Doubao-Seed 2.0 Pro

    A flagship-level, all-around general-purpose model designed for complex reasoning and long-chain task execution scenarios in the Agent era. It emphasizes mul...

  • Meta: Llama 3.1 70B Instruct

    Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue u...

  • Meta: Llama 4 Maverick

    Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 expert...

  • Mistral: Mistral Small 3.2 24B Instruct 2506

    Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved fu...

  • OpenAI: GPT-5.4 Pro

    GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. ...

  • GLM: Embedding 3

    Embedding-3 is the third-generation text embedding model launched by Zhipu AI. Representing a comprehensive upgrade over its predecessors, it offers enhanced...

  • Cohere: Command A 03-2025

    Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding...

  • Inflection: Inflection 3 Pi

    Inflection 3 Pi powers Inflection's [Pi](https://pi.ai/) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to rec...

  • Nous: Hermes 4 70B

    This incarnation of Hermes 4 balances scale and size. It handles complex reasoning tasks, while staying fast and cost effective. A versatile choice for many ...

  • Nous: Hermes 4 405B

    This is the largest model in the Hermes 4 family, and it is the fullest expression of our design, focused on advanced reasoning and creative depth rather tha...

  • Nous: Hermes 4 14B

    Hermes 4 14B is an open-source reasoning model from Nous Research built on Qwen 3 that excels at math, code, logic, and creative tasks

  • Google: Veo 3.1 (Text to Video)

    Google Veo 3.1 converts text prompts into videos with synchronized audio at native 1080p for high-quality outputs. Ready-to-use REST inference API, best perf...

  • Perceptron: Perceptron Mk1

    Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired ...

  • Morph: Morph V3 Fast

    Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the f...

  • Morph: Morph V3 Large

    Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt...

  • Google: Gemini 3.5 Flash

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimi...

  • xAI: Grok Build 0.1

    Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output...

  • Qwen: Qwen3.7 Max

    Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular...

  • DeepSeek: Deepseek V4 Flash (Free)

    Note: For the free endpoint, all prompts and outputs are logged to help improve the provider’s model, products, and services. This endpoint is provided for t...

  • Amazon: Nova 2 Lite

    Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstr...

  • Amazon: Nova Micro 1.0

    Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context ...

  • Amazon: Nova Lite 1.0

    Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output...

  • Amazon: Nova Pro 1.0

    Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. It a...

  • Anthropic: Claude Opus 4.8

    Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with rea...

  • StepFun: Step 3.7 Flash

    Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for...

  • MiniMax: MiniMax M3

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suite...

  • Anthropic: Claude Opus 4.8 (Fast)

    Fast-mode variant of Opus 4.8 - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claud...

  • Qwen: Qwen3.7 Plus

    Among the Qwen3.7 series, the cost-effective Plus model builds on its robust text capabilities while delivering a comprehensive upgrade to its vision‑languag...

  • NVIDIA: Nemotron 3 Ultra

    NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hyb...

  • Anthropic: Claude Fable 5

    Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text out...

  • Moonshot: Kimi K2.7 Code

    MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long conte...

  • Alibaba: Wan 2.7 (Text to Video)

    Generate high-quality images from text prompts using the WAN 2.7 model with advanced prompt understanding and detailed output.

  • Alibaba: Wan 2.7 (Image to Video)

    Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

  • Alibaba: Wan 2.7 (Reference to Video)

    Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

  • Alibaba: Wan 2.7 (Video Edit)

    Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

  • Z.AI: GLM 5.2

    GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering cont...

  • xAI: Grok Imagine Video 1.5 (Image to Video)

    Generate videos from images with audio using xAI's Grok Imagine 1.5 Video model.

  • Alibaba: HappyHorse-1.0 (Text to Video)

    Generate 1080p video with synchronized native audio from a text prompt. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.

  • Alibaba: HappyHorse-1.0 (Image to Video)

    Alibaba's #1-ranked Happy Horse 1.0 — generate 1080p video with synchronized native audio and multilingual lip-sync from text prompts or images.

  • Alibaba: HappyHorse-1.0 (Reference to Video)

    Generate 1080p video with synchronized native audio from a text prompt and references. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.

  • Alibaba: HappyHorse-1.0 (Video Edit)

    HappyHorse video editing supports advanced video editing through natural language instructions. It allows for local or global editing of video elements using...

  • ByteDance: Seed 2.0 Code

    A coding model optimized for real-world development environments, with reliable tool use in common IDEs such as Claude Code. It delivers strong front-end per...

  • Alibaba: HappyHorse-1.1 (Text to Video)

    Happy Horse 1.1 is Alibaba's #1-ranked video model. This text-to-video endpoint generates 1080p video with synchronized native audio and multilingual lip-syn...

  • Alibaba: HappyHorse-1.1 (Image to Video)

    Happy Horse 1.1 is Alibaba's #1-ranked video model. This image-to-video endpoint animates a still image into 1080p video with synchronized native audio and m...

  • Alibaba: HappyHorse-1.1 (Reference to Video)

    Happy Horse 1.1 is Alibaba's #1-ranked video model. This reference-to-video endpoint turns up to 9 reference images into 1080p video with synchronized native...

  • Sakana: Fugu Ultra

    Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration sys...

  • Anthropic: Claude Sonnet 5

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinkin...

  • Meituan: LongCat-2.0

    LongCat-2.0 is a high-performance foundation model designed for agentic workloads. Natively supports tool use, and long-context tasks, with strong performanc...

  • Tencent: HY3

    Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-w...

  • xAI: Grok 4.5

    Grok 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

  • OpenAI: GPT-5.6 Sol

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong a...

  • OpenAI: GPT-5.6 Terra

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for eve...

  • OpenAI: GPT-5.6 Luna

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, ...

  • Kwaipilot: Kat Coder Pro V2.5

    High-end coding agent built for complex software engineering, repository-scale tasks, and autonomous development workflows.

  • Kwaipilot: Kat Coder Air V2.5

    Fast, lightweight coding model designed for interactive development, rapid code iteration, and efficient coding agents.

  • AionLabs: Aion 3.0

    Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in...

  • AionLabs: Aion 3.0 Mini

    Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation ...

  • Meta AI: Muse Spark 1.1

    Muse Spark offers competitive performance in multimodal perception, reasoning, health, and agentic tasks.

  • ByteDance: Seed 2.1 Turbo

    A next-generation model tailored for the coding and agent era.

  • Moonshotai: Kimi K3

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agenti...

  • Thinking Machines: Inkling

    Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for ge...

  • Z.AI: GLM 5.2 Fast

    GLM-5.2-Fast-Preview is the high-speed version of Zhipu AI’s flagship model, GLM-5.2. It supports an ultra-long context window of 1 million tokens and matche...

  • Moonshot: Kimi K2.7 Code Fast

    The K2.7 Code High-Speed ​​version and the standard version share the same underlying model, yet the high-speed version delivers an output speed approximatel...

  • Google: Gemini 3.6 Flash

    Gemini 3.6 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimi...

  • Google: Gemini 3.5 Flash Lite

    Gemini 3.5 Flash Lite is Flash-Lite's first step into the agentic space, prioritizing thinking and tool calling to serve as a quick, efficient, and capable s...

  • Claude Opus 5

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software ...

  • Claude Opus 5 (Fast)

    Fast-mode variant of Opus 5 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https...

  • Qwen: Qwen 3.7 Flash

    Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with st...

  • DeepSeek: Deepseek V4 Flash 0731

    Deepseek v4 Flash, 0731 GA version.

  • Qwen: Qwen3.8 Max

    2.4-trillion-parameter MoE flagship delivering a comprehensive leap in coding and professional work. Autonomously codes and delivers complete projects spanni...

  • Ling 3.0 Flash

    Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token e...

  • Meta AI: Muse Glimmer 30B

    Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on ...

  • Upstage: Solar Pro 4

    Upstage's flagship text model for agentic, coding, document, and business workflows.

  • Upstage: Solar Pro 3

    Upstage's Mixture-of-Experts text model for reasoning and instruction following.

  • Nvidia: Nemotron 3.5 Lightning 30B A3B (Free)

    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput a...

  • DeepSeek: Deepseek V4 Pro 0813

    1.6T total / 49B active params. Performance rivaling the world's top closed-source models.

  • xAI: Grok 4.6

    Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

  • Qwen: Qwen3.8 2.4T A95B

    Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters o...

  • Alibaba: Wan 3.0 (Text to Video)

    Wan 3.0 Text to Video generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinki...

  • Alibaba: Wan 3.0 (Image to Video)

    Wan 3.0 Image to Video animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio co...

  • Alibaba: Wan 3.0 (Reference to Video)

    Wan 3.0 Reference to Video creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second dur...

  • Alibaba: Wan 3.0 (Video Edit)

    Wan 3.0 Reference to Video creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second dur...

  • Google: Gemini 3.7 Flash

    Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that requir...

  • Qwen: Qwen3.8 27B

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and l...

  • Z.AI: GLM 5.3

    GLM-5.3 is an advanced large language model designed to deliver strong performance across a wide range of natural language processing tasks, including text g...

  • Nvidia: Nvidia Nemotron 3.5 Lightning

    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput a...

  • DeepSeek: DeepSeek V4 Flash Vision Experimental

    The deepseek-v4-flash-vision-exp model accepts images alongside text, so you can ask the model to describe pictures, read text from screenshots, analyze char...

  • Google Transcoder: Merge Videos

    Merge ordered video inputs into one MP4 with Google Cloud Transcoder.

  • Qwen: Qwen3.8 27B (Free)

    Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and l...

  • Apodex 1.1

    Apodex's 397B flagship core reasoning model.

  • Apodex 1.1 Mini

    Apodex's 35B low-latency core model for high-volume tasks.

  • Qwen: Qwen3.8 Flash

    Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebas...

  • Z.AI: GLM 5.3 Flash

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention ...

  • Google: Nano Banana (Text to Image)

    Google's famous original image generation and editing model

  • Google: Nano Banana (Image to Image)

    Google's famous original image generation and editing model.

  • Google: Nano Banana 2 (Text to Image)

    Google Nano Banana 2 (Gemini 3.1 Flash Image) delivers Pro-quality image generation at Flash speed with 512px to 4K resolution support. Features include impr...

  • Google: Nano Banana 2 Lite (Text to Image)

    Google Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) delivers Pro-quality image generation at Flash speed with 512px to 4K resolution support. Features in...

  • Google: Nano Banana 2 (Image to Image)

    Nano Banana 2 is Google's new state-of-the-art image generation and editing model.

  • Google: Nano Banana 2 Lite (Image to Image)

    Nano Banana 2 Lite is Google's new state-of-the-art image generation and editing model.

  • Google: Nano Banana Pro (Text to Image)

    Google's Nano Banana pro (Gemini 3.0 Pro Image) is a cutting-edge text-to-image model enabling high-res 4K image generation optimized for phones. Ready-to-us...

  • Google: Nano Banana Pro (Image to Image)

    Google Nano Banana Pro (Gemini 3.0 Pro Image) Edit enables image editing with 4K-capable output. Ready-to-use REST inference API, best performance, no coldst...

  • xAI: Grok Imagine Image (Text to Image)

    Generate highly aesthetic images with xAI's Grok Imagine Image generation model.

  • xAI: Grok Imagine Image (Image to Image)

    Edit images precisely with xAI's Grok Imagine model.

  • xAI: Grok Imagine Image Quality (Text to Image)

    Grok Imagine Pro is an advanced AI model from xAI that creates high-quality visuals from text prompts and allows you to edit or analyze existing images.

  • xAI: Grok Imagine Image Quality (Image to Image)

    Grok Imagine Pro is an advanced AI model from xAI that creates high-quality visuals from text prompts and allows you to edit or analyze existing images.

  • OpenAI: GPT Images 2 (Text to Image)

    GPT Image 2, OpenAI's latest image model, is capable of creating extremely detailed images with fine typography.

  • OpenAI: GPT Images 2 (Image to Image)

    GPT Image 2, OpenAI's latest image model, is capable of creating extremely detailed images with fine typography.

  • Google: Veo 3.1 (Image to Video)

    Google Veo 3.1 is an Image-to-Video model that converts images into high-quality videos with native 1080P output for enhanced detail and creative flexibility...

  • Google: Veo 3.1 (First Last Frame to Video)

    Generate videos from a first and last framed using Google's Veo 3.1.

  • Google: Veo 3.1 (Video Extend)

    Extend and continue Veo 3.1 videos with smooth motion, preserved style, and strong scene coherence. Ready-to-use REST inference API, best performance, no col...

  • Google: Veo 3.1 (Reference to Video)

    Generate Videos from images using Google's Veo 3.1

  • ByteDance: Seedance 2.0 (Text to Video)

    ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.

  • ByteDance: Seedance 2.5 (Text to Video)

    Dreamina Seedance 2.5 generates native 30-second single-shot video at up to 720p from a single text prompt, reasoning about the whole shot at once so motion,...

  • ByteDance: Seedance 2.0 (Image to Video)

    ByteDance's most advanced image-to-video model. Animate still images into cinematic video with synchronized audio, start and end frame control, and motion pr...

  • ByteDance: Seedance 2.5 (Image to Video)

    Dreamina Seedance 2.5 animates a single still into a native 30-second clip at up to 720p, extending one frame into continuous, coherent motion without the dr...

  • ByteDance: Seedance 2.0 (Reference to Video)

    ByteDance's most advanced reference-to-video model. Generate video from up to 9 images, 3 videos, and 3 audio clips with native audio and cinematic camera co...

  • ByteDance: Seedance 2.0 (Real Human Reference to Video)

    ByteDance's most advanced reference-to-video model. Generate video from up to 9 images, 3 videos, and 3 audio clips with native audio and cinematic camera co...

  • ByteDance: Seedance 2.0 Fast (Real Human Reference to Video)

    ByteDance's most advanced reference-to-video model, fast tier. Lower latency and cost with up to 9 images, 3 videos, and 3 audio clips as inputs.

  • ByteDance: Seedance 2.5 (Reference to Video)

    Dreamina Seedance 2.5 generates video from up to 50 multimodal references images, video, audio, and style inputs, locking a character, set, and palette acros...

  • ByteDance: Seedance 2.5 (Real Human Reference to Video)

    Dreamina Seedance 2.5 generates video from up to 50 multimodal references images, video, audio, and style inputs, locking a character, set, and palette acros...

  • ByteDance: Seedance 2.0 Fast (Text to Video)

    ByteDance's most advanced text-to-video model, fast tier. Lower latency and cost with cinematic output, native audio, multi-shot editing, and director-level ...

  • ByteDance: Seedance 2.0 Mini (Text to Video)

    Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

  • ByteDance: Seedance 2.0 Fast (Image to Video)

    ByteDance's most advanced image-to-video model, fast tier. Lower latency and cost with synchronized audio, start and end frame control, and motion prompts.

  • ByteDance: Seedance 2.0 Mini (Image to Video)

    Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

  • ByteDance: Seedance 2.0 Fast (Reference to Video)

    ByteDance's most advanced reference-to-video model, fast tier. Lower latency and cost with up to 9 images, 3 videos, and 3 audio clips as inputs.

  • ByteDance: Seedance 2.0 Mini (Reference to Video)

    Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

  • ByteDance: Seedance 2.0 Mini (Real Human Reference to Video)

    Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

  • xAI: Grok Imagine Video (Text to Video)

    X-AI Grok Imagine Video generates videos from text descriptions using xAI's Grok Imagine Video model. Create high-quality videos with customizable duration, ...

  • xAI: Grok Imagine Video (Image to Video)

    X-AI Grok Imagine Video generates videos from text descriptions using xAI's Grok Imagine Video model. Create high-quality videos with customizable duration, ...

  • xAI: Grok Imagine Video (Reference to Video)

    X-AI Grok Imagine Video generates videos from text descriptions using xAI's Grok Imagine Video model. Create high-quality videos with customizable duration, ...

  • xAI: Grok Imagine Video (Edit Video)

    Edit videos using xAI's Grok Imagine.

  • xAI: Grok Imagine Video (Extend Video)

    Extend videos with xAI's Grok Imagine video model.

  • Wan Spicy: Z-Image Spicy (Text to Image)

    Z Image Spicy text-to-image. Square / portrait / landscape compositions, 256–1536px on each side. Integrate this model via REST with endpoint docs, parameter...

  • Wan Spicy: Qwen Image Edit Spicy (Image to Image)

    Qwen Image Edit Spicy. Add, remove, or modify elements in an existing image with text guidance. Integrate this model via REST with endpoint docs, parameters,...

  • Wan Spicy: Face Swap (Image to Image)

    Face Swap model. Integrate this model via REST with endpoint docs, parameters, and code examples.

  • Wan Spicy: Head Swap (Image to Image)

    Head Swap model. Integrate this model via REST with endpoint docs, parameters, and code examples.

  • Wan Spicy: Wan 2.2 I2V Spicy (Image to Video)

    Image-to-video with WAN 2.2 Spicy. Animate a starting image. 480p or 720p, 5s or 8s clips. Integrate this model via REST with endpoint docs, parameters, and ...

  • Wan Spicy: Wan 2.7 I2V Spicy (Image to Video)

    Image-to-video with WAN 2.7 Spicy. Animate a starting image with optional driving audio. 720p or 1080p, 2–15 second clips. Integrate this model via REST with...

  • Wan Spicy: Wan Animate (Video to Video)

    Wan Animate creates a digital-human animation by combining a character image with a reference motion video. Output duration follows the reference video lengt...

  • Wan Spicy: FlashVSR (Video Super Resolution)

    FlashVSR super-resolution model, supporting upscaling from 480P/720P/1080P to 2K/4K.

  • Alibaba: Wan 2.7 Image (Text to Image)

    Generate high-quality images from text prompts using the WAN 2.7 model with advanced prompt understanding and detailed output.

  • Alibaba: Wan 2.7 Image (Image to Image)

    Generate high-quality images from image prompts using the WAN 2.7 model with advanced prompt understanding and detailed output.

  • Alibaba: Wan 2.7 Image Pro (Text to Image)

    Wan2.7–image-pro,supports text to image, text/image to sequential images, image editing, multi-image reference generation, and interactive editing. Delivers ...

  • Alibaba: Wan 2.7 Image Pro (Image to Image)

    Wan2.7–image-pro,supports text to image, text/image to sequential images, image editing, multi-image reference generation, and interactive editing. Delivers ...

  • ByteDance: Seedream 5.0 Pro (Text to Image)

    ByteDance's Seedream 5.0 Pro is flagship text-to-image model, with deep-thinking prompt understanding, native text in 14 languages, and precise control over ...

  • ByteDance: Seedream 5.0 Pro (Image to Image)

    ByteDance's Seedream 5.0 Pro is flagship text-to-image model, with deep-thinking prompt understanding, native text in 14 languages, and precise control over ...

  • ByteDance: Seedream 5.0 Lite (Text to Image)

    Text to Image endpoint for the fast Lite version of Seedream 5.0, supporting high quality intelligent text-to-image generation.

  • ByteDance: Seedream 5.0 Lite (Image to Image)

    Text to Image endpoint for the fast Lite version of Seedream 5.0, supporting high quality intelligent text-to-image generation.

  • Alibaba: Qwen Image 3.0 (Image to Image)

    Generates images from a text prompt at resolutions up to 2048×2048, with automatic prompt rewriting and prompt-guided resolution selection, building on Qwen'...

  • Alibaba: Qwen Image 3.0 (Text to Image)

    Edits images from one to three reference images and a natural-language instruction, preserving key details such as facial features and identity while applyin...

  • Alibaba: Qwen Image 3.0 Pro (Image to Image)

    Qwen Image 3.0 Pro Edit is a professional-grade image editing model that transforms existing images with natural-language instructions, delivering advanced i...

  • Alibaba: Qwen Image 3.0 Pro (Text to Image)

    Qwen Image 3.0 Pro Text-to-Image is a professional-grade image generation model that creates high-quality images from text prompts, with advanced prompt unde...

  • xAI: Grok Imagine Image 2.0 (Text to Image)

    Generate images from text using xAi's Grok Imagine 2.0 model.

  • xAI: Grok Imagine Image 2.0 (Image to Image)

    Edit images with xAi's Grok Imagine 2.0 model.

  • Deepseek: Deepseek R1 Distill Qwen 14B

    DeepSeek R1 Distill Qwen 14B is a distilled large language model based on Qwen 2.5 14B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini acros...

  • Deepseek: Deepseek R1 Distill Qwen 32B

    DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini acros...

  • Austism: chronos-hermes-13b

    ## Chronos Hermes 13b Chronos Hermes (13b) is an advanced AI language model. This model is particularly adept at producing evocative storywriting and maintai...

  • Teknium: OpenHermes 2.5 Mistral 7B

    ## OpenHermes-2.5-Mistral (7B) Transform your communication strategies with OpenHermes-2.5-Mistral (7B) API, an AI model designed to enhance interaction and ...

  • Nousresearch: nous-hermes-yi-34b

    ## Nous Hermes-2 Yi (34B) Nous-Hermes-2-Yi-34B presents a promising large language model with great abilities in understanding complex questions and reasonin...

  • Migtissera: synthia-13b

    ## Model Overview Meet Synthia-13B, a powerful AI model designed to understand and respond to human input. Synthia-13B is a type of Llama-2-13B model, traine...

  • Orenguteng: llama-3-8b-lexi-uncensored

    ## Llama-3-8B-Lexi-Uncensored This model is based on Llama-3-8b-Instruct, and is governed by [META LLAMA 3 COMMUNITY LICENSE AGREEMENT](https://llama.meta.co...

  • Nitral-AI: Captain BMO 12B

    # Captain\_BMO-12B | Property | Value | | --- | --- | | Parameter Count | 12.2B | | Model Type | Mistral-based Language Model | | Tensor Type | BF16 | | Lice...

  • Sao10k: l3.1-70b-hanami-x1

    ## What is Llama 3.1 70B Hanami x1? Llama 3.1 70B Hanami x1 is a state-of-the-art large language model (LLM) developed by Sao10K, built upon Meta’s Llama 3.1...

  • TheDrummer: Cydonia-22B-v1

    ## Links - Original: https://huggingface.co/TheDrummer/Cydonia-22B-v1 - GGUF: https://huggingface.co/TheDrummer/Cydonia-22B-v1-GGUF - iMatrix: https://huggin...

  • Elinas: Chronos-Gold-12B-1.0

    ## What is Chronos-Gold-12B-1.0? Chronos-Gold-12B-1.0 is a state-of-the-art large language model developed by elinas, featuring 12 billion parameters designe...

  • FallenMerick: MN-Violet-Lotus-12B

    ## MN-Violet-Lotus-12B Overview MN-Violet-Lotus-12B is a 12 billion parameter language model developed by FallenMerick, distinguished by its strong performan...

  • Sao10K: MN-12B-Lyra-v4

    # MN-12B-Lyra-v4 | Property | Value | | --- | --- | | Author | Sao10K | | License | cc-by-nc-4.0 | | Model Size | 12B parameters | | Base Architecture | Mist...