Skip to main content

Provider directory

Tensorx models in Tavory

Tavory currently lists 6 Tensorx models in its live catalog. Explore model capabilities, catalog signals and eligible workspace access. Tavory does not represent a manufacturer relationship unless explicitly stated.

Models from Tensorx

text

Deepseek V4 Flash 0731

DeepSeek V4 Flash 0731 is the July 31 update of DeepSeek's efficiency-optimized MoE model (284B total, 13B active parameters) with a 1M-token context window. It is built for fast inference and high-throughput workloads while keeping strong reasoning and coding performance. Reasoning is toggleable and off by default.

View model

text

Glm 5.2

GLM 5.2 is the latest Z.ai flagship model for long horizon coding, reasoning, and agentic workflows. It improves on GLM 5.1 with a 1M token context window, stronger engineering task performance, multiple thinking effort levels, and architectural changes that reduce long context inference cost while improving speculative decoding. Released under the MIT license, it is designed for sustained work over large repositories, complex tool use, and multi step technical tasks.

View model

text

Kimi K3

Kimi K3 is Moonshot AI's flagship reasoning model with a 1M token context window, vision input, strong agentic tool use, and long horizon task execution. Served via TensorX.

View model

text

Kimi K2.7 Code

Kimi K2.7 Code is a coding focused model in Moonshot AI Kimi K2 family, built to complete end to end programming tasks reliably over long contexts. It uses a native multimodal mixture of experts architecture that accepts text and image input, and it always operates in a thinking mode, preserving full reasoning content across multi turn conversations. With a 256K token context window, it targets long horizon coding, agentic task decomposition, and multi turn dialogue. The model activates 32B parameters out of roughly 1T total.

View model

video

Minimax M3

MiniMax M3 is a frontier multimodal model with a 1M token context window built on MiniMax Sparse Attention (MSA). It delivers frontier level performance on coding and agentic tasks, outperforming GPT 5.5 and Gemini 3.1 Pro on SWE Bench Pro and approaching Claude Opus 4.7. It natively handles image and video input and is the first open weight model to combine frontier coding, ultra long context, and native multimodality.

View model

text

Deepseek V4 Pro

DeepSeek V4 Pro is a large scale Mixture of Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M token context window. It is built for advanced reasoning, coding, and long horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks. Built on the same architecture as DeepSeek V4 Flash, it adds a hybrid attention system for efficient long context processing. Reasoning efforts high and xhigh are supported, with xhigh mapping to max reasoning. It is well suited for complex workloads such as full codebase analysis, multi step automation, and large scale information synthesis, where both capability and efficiency are critical.

View model

Using Tensorx models in Tavory

Choose an eligible model from the catalog or let Smart Mode select a suitable available route. Tavory checks capabilities, plan access and current availability server-side when a request is sent.