Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

mifinkelson
/
scena

Text-to-Audio
LTX.io
English
ltx-audio
audio
audio-generation
speech
reference-conditioning
multi-speaker
flow-matching
diffusion
Model card Files Files and versions
xet
Community
1

Instructions to use mifinkelson/scena with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • LTX.io

    How to use mifinkelson/scena with LTX.io:

    # Install the LTX-2 pipelines
    git clone https://github.com/Lightricks/LTX-2.git
    cd LTX-2
    uv sync --frozen
    # Download the weights from this repo, plus the Gemma text encoder
    hf download mifinkelson/scena --local-dir models/scena
    hf download google/gemma-3-12b-it-qat-q4_0-unquantized --local-dir models/gemma-3-12b
    # Fast pipeline (distilled model, no distilled LoRA needed)
    uv run python -m ltx_pipelines.distilled \
        --distilled-checkpoint-path models/scena/<distilled-checkpoint>.safetensors \
        --spatial-upsampler-path models/scena/<spatial-upsampler>.safetensors \
        --gemma-root models/gemma-3-12b \
        --prompt "A beautiful sunset over the ocean" \
        --output-path output.mp4
    # For image-to-video, add: --image path/to/image.jpg 0 0.8
    # HQ pipeline (two-stage, higher quality)
    uv run python -m ltx_pipelines.ti2vid_two_stages_hq \
        --checkpoint-path models/scena/<checkpoint>.safetensors \
        --distilled-lora models/scena/<distilled-lora>.safetensors 0.8 \
        --spatial-upsampler-path models/scena/<spatial-upsampler>.safetensors \
        --gemma-root models/gemma-3-12b \
        --prompt "A beautiful sunset over the ocean" \
        --output-path output.mp4
    # For image-to-video, add: --image path/to/image.jpg 0 0.8
  • Notebooks
  • Google Colab
  • Kaggle
scena
8.52 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 9 commits
mifinkelson's picture
mifinkelson
Update model card examples
604c3d4 verified 25 days ago
  • .gitattributes
    1.52 kB
    initial commit 25 days ago
  • README.md
    4.15 kB
    Update model card examples 25 days ago
  • audio_vae.safetensors
    365 MB
    xet
    Fix: bundle audio_vae+vocoder config metadata (mel_bins etc.) 25 days ago
  • scena.safetensors
    8.15 GB
    xet
    Add ScenA audio-only reference-conditioned checkpoint (arefs-20s-abs-adv-best-sh @02387100) 25 days ago