Тип

Тег «audio» — 596

Модель automatic-speech-recognition · лицензия mit

Модель Модели и платформы Своё развёртывание

Модель automatic-speech-recognition · лицензия apache-2.0

Модель Модели и платформы Своё развёртывание

BSC-LT/hubert-base-los-6k

huggingface.co

—

Модель Модели и платформы Своё развёртывание

jhcodec/jhcodec_1.4m

huggingface.co

Модель audio-to-audio · лицензия mit

Модель Модели и платформы Своё развёртывание

cstr/gemma4-e2b-it-GGUF

huggingface.co

Модель automatic-speech-recognition · лицензия apache-2.0

Модель Модели и платформы Своё развёртывание

Модель automatic-speech-recognition · лицензия cc-by-4.0

Модель Модели и платформы Своё развёртывание

zeromodels/whisper_large

huggingface.co

Модель automatic-speech-recognition · лицензия apache-2.0

Модель Модели и платформы Своё развёртывание

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

Модель Модели и платформы Своё развёртывание

Google: Gemini 3.7 Flash

openrouter.ai

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Модель Модели и платформы Своё развёртывание

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Модель Модели и платформы Своё развёртывание

Meta: Muse Spark 1.2

openrouter.ai

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

Модель Модели и платформы Своё развёртывание

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Модель Модели и платформы Своё развёртывание

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

Модель Модели и платформы Своё развёртывание

Google: Gemini 3.6 Flash

openrouter.ai

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Модель Модели и платформы Своё развёртывание

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Модель Модели и платформы Своё развёртывание

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Модель Модели и платформы Своё развёртывание

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Модель Модели и платформы Своё развёртывание

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Модель Модели и платформы Своё развёртывание

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Модель Модели и платформы Своё развёртывание

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Модель Модели и платформы Своё развёртывание

Auto Router (Beta)

openrouter.ai

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

Модель Модели и платформы Своё развёртывание

Meta: Muse Spark 1.1

openrouter.ai

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...

Модель Модели и платформы Своё развёртывание

Google: Gemini 3.5 Flash

openrouter.ai

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Модель Модели и платформы Своё развёртывание

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Модель Модели и платформы Своё развёртывание

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Модель Модели и платформы Своё развёртывание

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Модель Модели и платформы Своё развёртывание

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

Модель Модели и платформы Своё развёртывание

Google Gemini Pro Latest

openrouter.ai

This model always redirects to the latest model in the Google Gemini Pro family.

Модель Модели и платформы Своё развёртывание

This model always redirects to the latest model in the Google Gemini Flash family.

Модель Модели и платформы Своё развёртывание

Xiaomi: MiMo-V2.5

openrouter.ai

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

Модель Модели и платформы Своё развёртывание

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Модель Модели и платформы Своё развёртывание

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

Модель Модели и платформы Своё развёртывание

Предложить сайт в каталог

Пришлите ссылку — остальное мы выясним сами.

Мы рассмотрим, что вы прислали, и добавим в каталог, если подойдёт.

Не знаете, как внедрить? Мы поможем

Расскажите про задачу — подберём инструменты и подскажем, с чего начать.

0 / 5000
Проверочный код

Поля со звёздочкой обязательны. Данные используются только для ответа.