RunningHub| App Details
Share API call
0
0
Subscribe
DeepSeek Janus Pro text-to-image
Input Parameter
Historical Record
717 runs
Workflows
11
0
App Introduction

Janus Pro is a unified understanding and generation MLLM, which decouples visual encoding for multimodal understanding and generation. Janus Pro is constructed based on the DeepSeek LLM 1.5b base/DeepSeek LLM 7b base.

For multimodal understanding, it uses SigLIP L as the vision encoder, which supports 384 x 384 image input. For image generation, Janus Pro uses the tokenizer from here with a downsample rate of 16.

Janus Pro is a unified understanding and generation MLLM, which decouples visual encoding for multimodal understanding and generation. Janus Pro is constructed based on the DeepSeek LLM 1.5b base/DeepSeek LLM 7b base.

For multimodal understanding, it uses the SigLIP L as the vision encoder, which supports 384 x 384 image input. For image generation, Janus Pro uses the tokenizer from here with a downsample rate of 16.

DeepSeek Janus Pro text-to-image

Text-to-Image
11
0
717 runs
Charged by actual usage
Billing
Run Now
Standard
App Details
My Results
Epsilon
DeepSeek Janus Pro text-to-image
1
Creations
Comments
Publish
History Record
History Record
API Task
No creations available.

No creations available.

Thinking