Loading Tavory AI
Loading Tavory AI
Provider directory
Tavory currently lists 19 Nebius models in its live catalog. Explore model capabilities, catalog signals and eligible workspace access. Tavory does not represent a manufacturer relationship unless explicitly stated.
text
Qwen thinking-optimized 80B model designed for sustained multi-step reasoning, structured deliberation, and high-precision problem-solving across math, code, and complex planning tasks.
View modelvideo
Multimodal model featuring a Hybrid Mixture-of-Experts architecture, designed for state-of-the-art performance across chat, retrieval-augmented generation, vision-language understanding, video understanding, and agentic workflows.
View modeltext
Nemotron 3 Super is a 120B hybrid MoE model optimized for efficient multi-agent AI and complex reasoning tasks.
View modeltext
Compact version of Hermes-4 delivering high-quality reasoning and coding with lower inference cost.
View modeltext
Google mid-size model optimized for high-quality instruction following, coding, and multilingual performance.
View modeltext
Nemotron 3 Ultra is a 550B hybrid MoE model from NVIDIA, optimized for the most demanding multi-agent AI and complex reasoning tasks.
View modeltext
Kimi K3 is Moonshot AI's flagship reasoning model with a 1M token context window, strong agentic tool use, and long horizon task execution. Hosted on Nebius.
View modeltext
A lightweight and ultra-fast variant of Llama 3.3 70B, for use when quick response times are needed most.
View modeltext
GLM-5.2 is Zhipu AI's latest flagship model with strong bilingual (Chinese-English) reasoning, long-context understanding, advanced tool use, and agent-oriented capabilities. The latest model in the GLM series, designed to plan, execute, and iterate autonomously on extended, engineering-grade tasks. Hosted on Nebius at fp8 quantization.
View modeltext
Versatile 30B instruct model optimized for high-quality chat, reasoning, and coding.
View modeltext
Hybrid reasoning model trained on verified CoT traces for strong math, coding, and step by step reliability.
View modeltext
Open-source agentic coding model built for polyglot development and precision refactoring, using interleaved-thinking tool calls to reliably execute long, multi-step coding and office workflows.
View modeltext
Balanced Qwen3 flagship tuned for strong general reasoning, chat quality, and tool use.
View modeltext
Open-weight agentic model with configurable reasoning, full CoT visibility, strong tool use, and fine-tuning support.
View modeltext
Zhipu AI flagship multimodal model with strong bilingual Chinese English reasoning, long context understanding, advanced tool use, and agent oriented capabilities.
View modeltext
DeepSeek-V4 is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks.
View modeltext
Generalist model offering strong multilingual reasoning, coding, and long-context performance at mid scale.
View modeltext
Kimi K2.6 is an open-source, native multimodal agentic model built through continual pretraining on approximately 15 trillion mixed visual and text tokens atop Kimi-K2-Base.
View modeltext
The most open, efficient, and accurate omni modal reasoning model for agentic AI.
View modelChoose an eligible model from the catalog or let Smart Mode select a suitable available route. Tavory checks capabilities, plan access and current availability server-side when a request is sent.