Skip to main content
ArtEmotion
Image

Realistic Vision

Credit-based pricing. Supports batch up to 4.

Realistic Vision preview
Output
image
Aspect ratio
1:1, 16:9, 9:16, 4:3, 3:4, Custom
Resolution
1024
Max batch
4

Inputs

  • Batch up to 4 per request

Example prompt

A stunning photorealistic close-up portrait of an elderly Japanese fisherman sitting on a weathered dock at golden hour, deep wrinkles and sun-worn skin reflecting decades at sea, wearing a traditional straw hat, soft ocean mist rising in the background, warm amber and orange sunlight casting dramatic side shadows, ultra detailed skin texture, cinematic depth of field, 8K resolution, award-winning photography
Try this prompt →
LLM-ready

API & LLM schema

Exact request contract for this model. Agents can fetch it from /api/v1/models?id=fal-ai/realistic-vision.

POST/api/v1/generate30 fields · 2 required
FieldTypeRequirementContract
model_idconstantRequiredArtEmotion model identifier.
extraobjectOptionalModel-specific settings may also be nested here.
max_creditsnumberOptionalReject before submission if the estimated list price exceeds this cap. · Range: 1–…
webhook_urlstringOptionalFormat: uri
webhook_secretstringOptionalOptional model input.
folder_idstringOptionalOptional model input.
promptstringRequiredOptional model input.
seedintegerOptionalThe same seed and the same prompt given to the same version of Stable Diffusion will output the same image every time.
negative_promptstringOptionalThe negative prompt to use. Use it to address details that you don't want in the image. · Default: (worst quality, low quality, normal quality, lowres, low details, oversaturated, undersaturated, overexposed, underexposed, grayscale, bw, bad photo, bad photography, bad art:1.4), (watermark, signature, text font, username, error, logo, words, letters, digits, autograph, trademark, name:1.2), (blur, blurry, grainy), morbid, ugly, asymmetrical, mutated malformed, mutilated, poorly lit, bad shadow, draft, cropped, out of frame, cut off, censored, jpeg artifacts, out of focus, glitch, duplicate, (airbrushed, cartoon, anime, semi-realistic, cgi, render, blender, digital art, manga, amateur:1.3), (3D ,3D Game, 3D Game Scene, 3D Character:1.1), (bad hands, bad anatomy, bad body, bad face, bad teeth, bad arms, bad legs, deformities:1.3)
safety_checker_versionstringOptionalThe version of the safety checker to use. v1 is the default CompVis safety checker. v2 uses a custom ViT model. · Allowed: v1, v2 · Default: v1
guidance_scalenumberOptionalThe CFG (Classifier Free Guidance) scale is a measure of how close you want the model to stick to your prompt when looking for a related image to show · Default: 5 · Range: 0–20
num_inference_stepsnumberOptionalThe number of inference steps to perform. · Default: 35 · Range: 1–70
expand_promptbooleanOptionalIf set to true, the prompt will be expanded with additional prompts. · Default: false
formatstringOptionalThe format of the generated image. · Allowed: jpeg, png · Default: jpeg
enable_safety_checkerbooleanOptionalIf set to true, the safety checker will be enabled. · Default: true
lora_pathstringOptionalURL or Hugging Face path to LoRA weights.
lora_scalenumberOptionalStrength of the LoRA effect (0–1). · Default: 1 · Range: 0–1
lora_path_2stringOptionalURL or Hugging Face path to a second LoRA.
lora_scale_2numberOptionalStrength of the second LoRA. · Default: 1 · Range: 0–1
lora_path_3stringOptionalURL or Hugging Face path to a third LoRA.
lora_scale_3numberOptionalStrength of the third LoRA. · Default: 1 · Range: 0–1
embedding_pathstringOptionalURL or Hugging Face path to textual-inversion embedding weights.
embedding_tokensarray<string>OptionalTokens that trigger the embedding when used in your prompt.
embedding_path_2stringOptionalURL or Hugging Face path to a second embedding.
embedding_tokens_2array<string>OptionalTokens for the second embedding.
embedding_path_3stringOptionalURL or Hugging Face path to a third embedding.
embedding_tokens_3array<string>OptionalTokens for the third embedding.
aspect_ratiostringOptionalAllowed: 1:1, 16:9, 9:16, 4:3, 3:4, Custom
resolutionstringOptionalAllowed: 1024
num_imagesintegerOptionalRange: 1–4
Minimal request example
{
  "model_id": "fal-ai/realistic-vision",
  "prompt": "A stunning photorealistic close-up portrait of an elderly Japanese fisherman sitting on a weathered dock at golden hour, deep wrinkles and sun-worn skin reflecting decades at sea, wearing a traditional straw hat, soft ocean mist rising in the background, warm amber and orange sunlight casting dramatic side shadows, ultra detailed skin texture, cinematic depth of field, 8K resolution, award-winning photography"
}
Raw JSON Schema
{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "$id": "https://www.artemotion.ai/api/v1/models?id=fal-ai%2Frealistic-vision",
  "title": "Realistic Vision generation request",
  "description": "Request body accepted by POST /api/v1/generate for fal-ai/realistic-vision.",
  "type": "object",
  "properties": {
    "model_id": {
      "type": "string",
      "const": "fal-ai/realistic-vision",
      "description": "ArtEmotion model identifier."
    },
    "extra": {
      "type": "object",
      "additionalProperties": true,
      "description": "Model-specific settings may also be nested here."
    },
    "max_credits": {
      "type": "number",
      "minimum": 1,
      "description": "Reject before submission if the estimated list price exceeds this cap."
    },
    "webhook_url": {
      "type": "string",
      "format": "uri",
      "maxLength": 2048
    },
    "webhook_secret": {
      "type": "string",
      "maxLength": 512
    },
    "folder_id": {
      "type": "string"
    },
    "prompt": {
      "type": "string"
    },
    "seed": {
      "title": "Seed",
      "description": "The same seed and the same prompt given to the same version of Stable Diffusion will output the same image every time.",
      "type": "integer"
    },
    "negative_prompt": {
      "title": "Negative Prompt",
      "description": "The negative prompt to use. Use it to address details that you don't want in the image.",
      "default": "(worst quality, low quality, normal quality, lowres, low details, oversaturated, undersaturated, overexposed, underexposed, grayscale, bw, bad photo, bad photography, bad art:1.4), (watermark, signature, text font, username, error, logo, words, letters, digits, autograph, trademark, name:1.2), (blur, blurry, grainy), morbid, ugly, asymmetrical, mutated malformed, mutilated, poorly lit, bad shadow, draft, cropped, out of frame, cut off, censored, jpeg artifacts, out of focus, glitch, duplicate, (airbrushed, cartoon, anime, semi-realistic, cgi, render, blender, digital art, manga, amateur:1.3), (3D ,3D Game, 3D Game Scene, 3D Character:1.1), (bad hands, bad anatomy, bad body, bad face, bad teeth, bad arms, bad legs, deformities:1.3)",
      "type": "string"
    },
    "safety_checker_version": {
      "title": "Safety Checker Version",
      "description": "The version of the safety checker to use. v1 is the default CompVis safety checker. v2 uses a custom ViT model.",
      "default": "v1",
      "type": "string",
      "enum": [
        "v1",
        "v2"
      ]
    },
    "guidance_scale": {
      "title": "Guidance Scale",
      "description": "The CFG (Classifier Free Guidance) scale is a measure of how close you want the model to stick to your prompt when looking for a related image to show",
      "default": 5,
      "type": "number",
      "minimum": 0,
      "maximum": 20,
      "multipleOf": 0.1
    },
    "num_inference_steps": {
      "title": "Num Inference Steps",
      "description": "The number of inference steps to perform.",
      "default": 35,
      "type": "number",
      "minimum": 1,
      "maximum": 70,
      "multipleOf": 1
    },
    "expand_prompt": {
      "title": "Expand Prompt",
      "description": "If set to true, the prompt will be expanded with additional prompts.",
      "default": false,
      "type": "boolean"
    },
    "format": {
      "title": "Format",
      "description": "The format of the generated image.",
      "default": "jpeg",
      "type": "string",
      "enum": [
        "jpeg",
        "png"
      ]
    },
    "enable_safety_checker": {
      "title": "Enable Safety Checker",
      "description": "If set to true, the safety checker will be enabled.",
      "default": true,
      "type": "boolean"
    },
    "lora_path": {
      "title": "LoRA Path",
      "description": "URL or Hugging Face path to LoRA weights.",
      "type": "string"
    },
    "lora_scale": {
      "title": "LoRA Scale",
      "description": "Strength of the LoRA effect (0–1).",
      "default": 1,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "lora_path_2": {
      "title": "LoRA Path 2",
      "description": "URL or Hugging Face path to a second LoRA.",
      "type": "string"
    },
    "lora_scale_2": {
      "title": "LoRA Scale 2",
      "description": "Strength of the second LoRA.",
      "default": 1,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "lora_path_3": {
      "title": "LoRA Path 3",
      "description": "URL or Hugging Face path to a third LoRA.",
      "type": "string"
    },
    "lora_scale_3": {
      "title": "LoRA Scale 3",
      "description": "Strength of the third LoRA.",
      "default": 1,
      "type": "number",
      "minimum": 0,
      "maximum": 1
    },
    "embedding_path": {
      "title": "Embedding Path",
      "description": "URL or Hugging Face path to textual-inversion embedding weights.",
      "type": "string"
    },
    "embedding_tokens": {
      "title": "Embedding Tokens",
      "description": "Tokens that trigger the embedding when used in your prompt.",
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "embedding_path_2": {
      "title": "Embedding Path 2",
      "description": "URL or Hugging Face path to a second embedding.",
      "type": "string"
    },
    "embedding_tokens_2": {
      "title": "Embedding Tokens 2",
      "description": "Tokens for the second embedding.",
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "embedding_path_3": {
      "title": "Embedding Path 3",
      "description": "URL or Hugging Face path to a third embedding.",
      "type": "string"
    },
    "embedding_tokens_3": {
      "title": "Embedding Tokens 3",
      "description": "Tokens for the third embedding.",
      "type": "array",
      "items": {
        "type": "string"
      }
    },
    "aspect_ratio": {
      "type": "string",
      "enum": [
        "1:1",
        "16:9",
        "9:16",
        "4:3",
        "3:4",
        "Custom"
      ]
    },
    "resolution": {
      "type": "string",
      "enum": [
        "1024"
      ]
    },
    "num_images": {
      "type": "integer",
      "minimum": 1,
      "maximum": 4
    }
  },
  "required": [
    "model_id",
    "prompt"
  ],
  "additionalProperties": false
}

FAQ

How much does Realistic Vision cost on ArtEmotion?

Credit-based pricing. You pay in ArtEmotion credits — every plan and top-up converts USD to credits at a fixed rate.

Do I get my credits back if Realistic Vision fails?

Yes — failed generations are never charged. The credits are released back to your balance automatically.

Can I call Realistic Vision from the API?

Yes. Use POST /api/v1/generate with model_id: "fal-ai/realistic-vision". See the API reference for the full schema.

Where are my generations stored?

Every output is saved to your personal Library. You can export or delete everything any time from Privacy & deletion.

Ready to generate with Realistic Vision?

Start now →

No commitment. New accounts get 50 free credits on signup. See pricing.