Type

Results for “speech to text” — 17 073

Amazon: Nova Micro 1.0

openrouter.ai

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

Model Models & platforms Self-hosted

Qwen: Qwen3.6 Flash

openrouter.ai

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

Model Models & platforms Self-hosted

Модель summarization · лицензия apache-2.0

Model Models & platforms Self-hosted

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Model Models & platforms Self-hosted

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

Model Models & platforms Self-hosted

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Model Models & platforms Self-hosted

Модель fill-mask · лицензия apache-2.0

Model Models & platforms Self-hosted

Модель feature-extraction · лицензия cc-by-nc-nd-4.0

Model Models & platforms Self-hosted

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...

Model Models & platforms Self-hosted

ALJIACHI/bte-base-ar

huggingface.co

Модель sentence-similarity · лицензия mit

Model Models & platforms Self-hosted

Модель feature-extraction · лицензия apache-2.0

Model Models & platforms Self-hosted

KGESH/nsfw-bge-m3

huggingface.co

Модель feature-extraction

Model Models & platforms Self-hosted

Модель feature-extraction · лицензия apache-2.0

Model Models & platforms Self-hosted

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Model Models & platforms Self-hosted

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Model Models & platforms Self-hosted

Submit a site to the catalog

Just send the link — we will work out the rest.

We will review what you send and add it to the catalog if it fits.

Not sure how to implement it? We can help

Tell us about your task — we will pick the tools and suggest where to start.

0 / 5000
Verification code

Fields marked with an asterisk are required. Your data is used only to reply.