
Lucy-5B is a model that can create 5-second I2V videos in under 5 seconds, achieving >1x RTF end-to-end

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f9a61%2FT4z71gOSeWv0wALDdi2-b_qVoN8eec.png/tr:w-1920,q-80/T4z71gOSeWv0wALDdi2-b_qVoN8eec.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2FUpscale-5.jpeg/tr:w-1920,q-80/Upscale-5.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

Train Ideogram on your photos, your style, your subject, your look, from a small set of reference images to images that feel consistently yours

Extend videos with audio using LTX-2 Distilled and custom LoRA

Generate video with audio from videos using LTX-2 Distilled

Generate video with audio from videos using LTX-2 Distilled and custom LoRA

Extend video with audio using LTX-2 and custom LoRA

Generate video with audio from videos using LTX-2 and custom LoRA

Cross-eyes for high-quality video using LTX-2.3

Train a LoRA that transforms one audio clip into another, learning a reference→target mapping from paired audio examples.

Train a LoRA that continues an audio clip forward in time, generating the audio that follows a short clean prefix.

Train a LoRA that generates the lead-in to an audio clip, extending audio backward in time from its ending.

Train a LoRA that regenerates masked time spans of an audio clip while keeping the rest unchanged.

Train a LoRA for a joint audio+video transformation, conditioned on a reference clip (its video and audio) to produce a matching target clip.

Train a LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

Train a LoRA that continues a video forward in time — supply an opening clip at inference and the model generates what comes next.

Train a LoRA that generates the lead-in to a video, extending a clip backward in time from its ending.

Train an IC-LoRA that transforms one audio clip into another, conditioned at inference on a reference audio clip.

Train an IC-LoRA for a joint audio+video transformation, conditioned on a reference clip's video and audio to produce a matching target.

Train an IC-LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

Train an IC-LoRA that regenerates only the masked region of a video, guided by the kept pixels and a separate reference/control video.

Train a LoRA that generates the video between keyframes — supply first/last (and optional middle) frames at inference and the model fills the in-between motion.

Train a LoRA that regenerates only the masked region of a video, guided by both the kept pixels and a separate reference/control video.

PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.

Turn photos into mind-blowing, dynamic videos. Your images can can come to life with sharp details, impressive character control and cinematic camera moves.

Turbo is the model to use when you feel the need for speed. Turn your image to stunning video up to 3x faster – all with high quality outputs.