multimodalart's picture
multimodalart HF Staff
Upload folder using huggingface_hub
7f864cd verified
|
Raw
History Blame
956 Bytes
metadata
title: SenseNova-U1.5-8B-MoT
emoji: 🎨
colorFrom: indigo
colorTo: blue
sdk: gradio
sdk_version: 6.25.0
app_file: app.py
python_version: '3.12'
short_description: Unified text-to-image and image editing model
startup_duration_timeout: 1h
pinned: false

SenseNova-U1.5-8B-MoT

Text-to-image and image editing demo for sensenova/SenseNova-U1.5-8B-MoT, a natively unified multimodal model (18B params, bf16) built on the NEO-unify architecture.

Leave the image upload empty for text-to-image generation, or upload one or more images and write an edit instruction for image editing.

Reference inference configuration (from the model card): cfg_scale=4.0, timestep_shift=3.0, num_steps=50.

The sensenova_u1 package is vendored locally to register the NEO-Unify model architecture with transformers. Running on ZeroGPU.