Creative intelligence platform for uncanny AI products

Build and scale creative products with the most direct and expressive video and image generation models available — through the Orrery API.

Start Building
A glass prism casting a vivid rainbow refraction — photo by Design Bits on Pexels

Find a plan that fits you

Build

Access frontier image and video generative features through one straightforward endpoint. Build and iterate generation for your business. Start plans include:


  • Instruction-following system
  • Text to video, image to video
  • Camera control, extend, loop
  • Billed via usage credits
  • Sub-minute generation time
  • You own your inputs and outputs
Start Building
A bright white flower in bloom — photo by venkat krishna on Pexels

Scale

Scale your creative AI products with higher rate limits and hands-on support from the Orrery research team.


  • Every Build tier capability
  • Higher rate limits
  • Sub-minute generation time
  • Onboarding support
  • Ongoing engineering support
  • Billed via monthly invoices
  • You own your inputs and outputs
Talk to us
A close-up of a delicate white flower — photo by Bud Jenkins on Pexels

Orrery Vela 2Video Model Capabilities

gen = create({
  prompt: 'impossible camera flight through a slot canyon',
   model: 'vela-2',
  aspect_ratio: '16:9',
});

Production ready.
Frontier text-to-video model.

Vela 2 opens a new generation of video models: fast coherent motion, dense material detail, and event sequences that follow cause and effect. The result is a much higher share of usable generations and footage that survives an actual edit.

A classic car speeding with heavy motion blur — photo by josh marks on Pexels

Natural Motion

Generate action-heavy shots from a plain sentence. From a high-speed chase to close human movement, Vela 2 raises the bar on motion fidelity — smooth, cinematic, and unnervingly physical.

A solitary ship under dramatic storm clouds at sea — photo by Pavel Khlystunov on Pexels

Storytelling

Tell a story with real cinematic grammar. Vela 2 handles precise camera movement — sweeping panoramas, intimate close-ups, and continuous tracking shots — all specified in text.

A cinematic night street with wet pavement and neon reflections — photo by Shuvo Haque on Pexels

Content Creation

Extend your content toolkit. Whether it is a product promo, a service showcase, or serial storytelling, generate broadcast-usable clips in seconds and keep the pipeline moving.

A television studio with cameras facing a green screen — photo by Shahbaz Zaman on Pexels

Visual Effects

Fold generative work into an existing VFX pipeline. Vela 2 produces previsualization, plate replacement, and first-pass sequences without booking a location.


Orrery PrismImage Model Capabilities

A robotic hand reaching out against a white background — photo by Tara Winstead on Pexels

State of the art creative output, from realism to stylized

Prism delivers category-specific visual quality, crafting images that hold up against professional standards — not generic model output. From campaign design to editorial, every generation is tuned to the requirements you actually ship against. Learn more

State of the art prompt following and text generation

Prism sets a new bar for visual precision. Its handling of your prompt holds every element in place — color, composition, and rendered text — with accuracy that matches what you described.

Resolution / Model Driftwave 3.5 Inkwell Muraldream Halcyon 1.1 Orrery Prism
1080pN/A¢ 6.2¢ 5.4¢ 6.1¢ 1.7
1080p fastN/AN/AN/AN/A¢ 0.4
720p (coming soon)¢ 6.3N/AN/A¢ 5.2¢ 0.9
720p fast (coming soon)¢ 3.8N/AN/A¢ 2.4¢ 0.2

760%
Faster & Cheaper

The Orrery image model runs on the same universal fusion architecture as our video model, which is where the efficiency comes from — a research result, not a quality trade you pay for later.

Character reference

Generate consistent character variations from a single reference photo. Prism holds facial structure and identity steady across poses, expressions, and scenes.

Visual reference

Apply and blend style references with exact control. Prism reads the essence of each reference image and lets you combine distinct visual elements without losing finish quality.


Orrery Vela 1.6Video Model Capabilities

prompt =
"a quiet modern house in slate and oak"

Text to Video

Vela and Prism share an instruction-following system, so users do not need to learn prompt engineering. That means teams can build generative products that reach new markets.

Image
prompt =
"a handbag 360 product shot"

Image to Video

Build workflows that take a static image and return motion. Instruct Vela in ordinary language and get a narrative back, not a slideshow.

1Keyframe
2Keyframe

Keyframe

Control the narrative Vela generates by pinning a start and end image keyframe.

A black-and-white silhouette of a person in profile — photo by Song Wei on Pexels

Extend

…and extend those narratives into stories, all without any pixel-editing complexity inside your app.

Monochrome smoke swirling against a dark backdrop — photo by Hayan on Pexels
loop:
true,

Loop

Create continuous loops for engaging interfaces, product marketing, and backgrounds.

A Weimaraner standing in a sunlit vineyard landscape — photo by Leeloo The First on Pexels
prompt:
"camera push in"

Camera Control

A generative camera that lets even a first-time user get a shot looking right with one plain sentence.

A minimalist ceramic bowl against a teal backdrop — photo by Feng Zou on Pexels
A minimalist empty bowl and spoon in soft light — photo by Steve A Johnson on Pexels
An elegant ceramic vase lit by warm sunlight — photo by Tiarra Sorte on Pexels

Variable Aspect Ratio

Your app can produce content sized for every platform without any video or image editing UI.


Start Building
A light spectrum split through a prism — photo by Nancy Zjaba on Pexels

Cutting edge image and video models at your hand


A new creative intelligence to partner with people and help them make better things

Read why we are building this platform →
A spectrum of rainbow light on a reflective surface — photo by Lars Mai on Pexels

Video carries culture, ideas, and connection around the world. Video is a universal language. Unlike text, making video is a physical act, and editing high-dimensional video data is genuinely hard. So the few who hold the means of production create, and the rest of us consume. They tell the stories and we all listen.

Video can be the most effective medium of visual storytelling. Video — if it were as universally manipulable as text — could be the most effective medium of thought.

To get there we have been training Vela, a family of generative models that produce and manipulate video. Adoption of Vela 1 outran everything we projected, and in directions we did not predict. Which tells us we should go faster.

The Orrery API

We are launching the Orrery API in beta today with the newest family of Vela 1.6 models. It covers high-quality text-to-video, image-to-video, video extension, loop creation, and our camera-control capabilities.

Through our research and engineering we aim to bring about an age of visual literacy in which anyone can make the thing they can picture, and the distance between an idea and an artifact keeps shrinking.


Moderation Controls and Responsible Use

Video is a powerful and persuasive medium. To prevent misuse we built a layered moderation system that pairs AI filters with human oversight. The API lets you tune that system to the norms of your market and your users. With ongoing feedback and learning we keep refining the approach so our models and products stay inside legal standards and get used constructively.

Enterprise Terms of Service
A natural hand alongside a painted arm reaching upward — photo by Darya Grey_Owl on Pexels