Video · text-to-image · text-to-video · text-image-to-video · video-refinement

LingBot-Video

robbyant

The inference package is flattened directly under worldfoundry/synthesis/visual_generation/lingbot_video.

IntegratedDedicated environmentlingbot-video

Build your run

Choose a recorded variant, then copy the exact setup or inspection command.

worldfoundry.pipeline
Tasktext-to-image
Environmentlingbot-video
DeviceCUDA 12.8
worldfoundry-eval evaluate \
  --mode model \
  --model-id lingbot-video \
  --model-runner worldfoundry:pipeline \
  --model-manifest-dir worldfoundry/data/models/catalog \
  --requests-path tmp/requests.jsonl \
  --output-dir tmp/model_eval/lingbot-video \
  --metric artifact_count \
  --json
01

Compatibility & versions

Manifest-backed recipe

Integration
integrated
Runner evidence
pending
Environment
lingbot-video
Python
3.10
CUDA
CUDA 12.8
PyTorch
torch==2.12.0.dev20260220+cu130
Source revision
Checkpoint revision
f9789a7d9b4772a47aba62d4eb5282ddefd1da21
Runtime profile
lingbot-video
Pipeline binding
lingbot-video
Runner
worldfoundry.pipeline
Pipeline target
worldfoundry.pipelines.lingbot_video.pipeline_lingbot_video:LingBotVideoGenerationPipeline
Backend stage
in_tree_runtime
Runtime status
dense_quality_and_moe_quality_smoke_verified_a100
Driver status
requires_cuda_13_compatible_driver
Environment kind
Dedicated environment
02

Install environment

The environment resolver reads the recorded profile and chooses the unified or dedicated environment shown here.

bash scripts/setup/model_env_install.sh --model lingbot-video
Package constraints 13
  • torch==2.12.0.dev20260220+cu130
  • torchvision==0.26.0.dev20260220+cu130
  • accelerate>=1.8.0
  • decord>=0.6.0
  • diffusers==0.39.0
  • imageio>=2.35.0
  • imageio-ffmpeg>=0.5.1
  • numpy>=1.26
  • peft==0.19.1
  • pillow>=10.4.0
  • safetensors>=0.4.5
  • scipy>=1.11
  • transformers==5.8.1
Conda packages 3
  • python=3.10
  • pip
  • ffmpeg
03

Checkpoints & assets

Run the local check before allocating compute. Gated, private, and license fields below come directly from the checkpoint manifest.

worldfoundry-eval zoo model-download --model-id lingbot-video --check-local --json
robbyant/lingbot-video-dense-1.3b
Revision
License
apache-2.0
Gated
Private
robbyant/lingbot-video-moe-30b-a3b
Revision
License
apache-2.0
Gated
Private
robbyant/lingbot-video-dense-1.3b
Revision
f9789a7d9b4772a47aba62d4eb5282ddefd1da21
License
Gated
Private
robbyant/lingbot-video-moe-30b-a3b
Revision
f2e538f64afe00cc4ae674db2aeb52e2945edfd5
License
Gated
Private
04

Launch & outputs

The generated command uses the shared evaluation boundary and writes durable result manifests and artifacts.

worldfoundry-eval evaluate \
  --mode model \
  --model-id lingbot-video \
  --model-runner worldfoundry:pipeline \
  --model-manifest-dir worldfoundry/data/models/catalog \
  --requests-path tmp/requests.jsonl \
  --output-dir tmp/model_eval/lingbot-video \
  --metric artifact_count \
  --json

Input contract

FieldRecorded contract
promptRequired
imageRequired
prompt_jsonRequired

Artifact contract

Artifact kindFilename / path
generated_image
generated_video
generated_videolingbot_video.mp4
05

Evidence & sources

Catalog integration, native-demo parity, and runner parity are independent records.

Integration
integrated
Runner evidence
pending
Native demo evidence
Not recorded
Validation imports
torch, torchvision, diffusers, transformers, accelerate, worldfoundry.synthesis.visual_generation.lingbot_video.inference_backend