APIXO

모델 카탈로그

최신 및 인기 모델

사용자들이 가장 먼저 사용하는 모델로 시작한 후, 출력 유형이나 워크플로우에 따라 전체 카탈로그를 검색하고 필터링하세요.

워크플로우

87 모델

전체 카탈로그 표시 중

bytedance

Seedance 2.0

Seedance 2.0 is ByteDance's multimodal video model supporting text-to-video, first-and-last-frames, and omni-reference modes. APIXO exclusive: unlimited concurrency, real-person portrait support, and hidden capabilities.

비디오신규인기텍스트 투 비디오이미지 투 비디오
$0.0573/초부터보기

bytedance

Seedance 2.0 Fast

Seedance 2.0 Fast is an APIXO route for ByteDance Seedance 2.0 workflows, exposing text-to-video, first-and-last-frames, and omni-reference modes with optional sound, web search, and 480p/720p output.

비디오신규인기텍스트 투 비디오이미지 투 비디오
$0.044/초부터보기
GPT-Image-2

OpenAI

GPT-Image-2

GPT-Image-2 is OpenAI's next-generation image model for stronger photorealism, cleaner image editing, and sharper in-image text rendering.

이미지신규인기텍스트 투 이미지이미지 투 이미지
$0.03/이미지부터보기

Alibaba

HappyHorse

HappyHorse is Alibaba's video generation and editing model for text-to-video, image-to-video, reference-guided generation, and video-edit workflows with 720p/1080p output.

비디오신규인기텍스트 투 비디오이미지 투 비디오
$0.125/초부터보기
Flux 2

Black Forest Labs

Flux 2

BFL’s latest Pro & Flex pipelines for text-to-image and image-to-image with unified 1K/2K pricing and ~30s generation.

이미지신규인기텍스트 투 이미지이미지 투 이미지
$0.04/이미지부터보기
Nano Banana Pro

Google

Nano Banana Pro

Nano Banana Pro is Google’s Gemini 3 Pro Image route for reasoning-informed image generation, native 4K output, multilingual typography, and spatially precise compositions. APIXO does not expose Search Grounding as an active control on this route.

이미지인기텍스트 투 이미지이미지 투 이미지
$0.08/이미지부터보기
Midjourney

midjourney

Midjourney

Midjourney is an advanced AI image generation model known for artistic, high-quality outputs. It supports text-to-image, image-to-image, image-edit, Vary, and Upscale workflows.

이미지인기텍스트 투 이미지이미지 투 이미지
요청당 $0.1부터보기
Flux Kontext

Black Forest Labs

Flux Kontext

Professional-grade image generation with enhanced prompt understanding and superior quality output.

이미지인기텍스트 투 이미지이미지 투 이미지
$0.04/이미지부터보기
GPT-Image-1

OpenAI

GPT-Image-1

GPT-Image-1 is OpenAI's advanced multimodal model for high-quality image generation with natural language understanding.

이미지인기텍스트 투 이미지이미지 투 이미지
$0.35/이미지부터보기
Nano Banana

Google

Nano Banana

Gemini 2.5 Flash Image Preview (aka Nano Banana) is an advanced AI model excelling in natural language-driven image generation and editing. It produces hyper-realistic, physics-aware visuals with seamless style transformations.

이미지인기텍스트 투 이미지이미지 투 이미지
$0.03/이미지부터보기

OpenAI

GPT-Image-2.5

Create and refine images with GPT Image 2.5. Turn a written brief or reference photo into product visuals, portraits, and illustrations, then fine-tune the details that matter.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.036/이미지부터보기

bytedance

Seedance 2.5

Seedance 2.5 is ByteDance's audiovisual video model for longer, reference-driven storytelling. It combines timeline-based direction, multimodal creative guidance, synchronized sound, multilingual performance, and selective revision to help production teams develop connected scenes instead of isolated short clips.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.085/초부터보기
Qwen Image 3.0

Alibaba

Qwen Image 3.0

Qwen Image 3.0 is Alibaba's text-to-image and image-to-image model with 1K and 2K output, fifteen aspect ratios, up to three reference images, and optional prompt expansion. Both resolution tiers are billed at the same rate.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.03/이미지부터보기
Qwen Image 3.0 Pro

Alibaba

Qwen Image 3.0 Pro

Qwen Image 3.0 Pro is Alibaba's higher-tier route for the same text-to-image and image-to-image contract, pricing 1K and 2K output separately so drafts and final deliverables can be billed differently.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.04/이미지부터보기
MiniMax H3

MiniMax

MiniMax H3

MiniMax H3 is a general-purpose multimodal video model that combines text, image, video, and audio context in one generation workflow. It creates videos up to 2K and 15 seconds with native stereo sound, while supporting first-and-last-frame control, mixed-media references, motion transfer, multi-shot storytelling, and instruction-guided video editing.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.09/초부터보기
Seedream 5.0 Pro

seedream

Seedream 5.0 Pro

Seedream 5.0 Pro is ByteDance's premium single-image model for text-to-image and image-to-image generation, supporting 1K/2K resolution, eight aspect ratios, up to 10 reference images, and strict custom-size control.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.045/이미지부터보기

bytedance

Seedance 1.5 Pro

Seedance 1.5 Pro is ByteDance's per-second video model for fast text-to-video and image-to-video generation with 480p/720p/1080p output, optional sound, aspect ratio control, and fixed-lens camera stability.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.0108/초부터보기

Google

Gemini Omni

Gemini Omni is Google's multimodal video generation model for creating videos from text, image references, source video, reusable audio assets, and character asset IDs.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.1/초부터보기
Wan 2.7 Image

Alibaba

Wan 2.7 Image

Wan 2.7 Image is Alibaba's omni-image API for text-to-image, reference-guided image generation, image editing, sequential images, and high-resolution Omni Image Pro workflows.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.03/이미지부터보기

Alibaba

Wan 2.7

Wan 2.7 is Alibaba's video generation and editing model for text-to-video, image-to-video, reference-guided generation, and video-edit workflows with optional audio input and 720p/1080p output.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.1/초부터보기

Alibaba

Wan 2.6

Wan 2.6 is Alibaba's multi-mode video generation model for text, image, flash image, reference, and flash reference workflows, with optional audio input and 720p/1080p output.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.025/초부터보기
Seedream 5.0

seedream

Seedream 5.0

Seedream 5.0 is ByteDance's next-generation AI image model with real-time web search, controllable editing, and logical reasoning. It supports text-to-image and image-to-image with 2K/3K resolution, multiple aspect ratios, and up to 14 reference images.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.035/이미지부터보기

OpenAI

Sora 2 Pro

Sora 2 Pro is OpenAI’s premium video generation model with higher quality output, supporting text-to-video and image-to-video at 720p and 1080p resolutions with flexible durations of 4, 8, or 12 seconds.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.3/초부터보기
Z-Image LoRA Pro

APIXO

Z-Image LoRA Pro

Z-Image LoRA Pro is an asynchronous text-to-image API tuned for vivid detail, returning one image per request with aspect-ratio presets, exact pixel sizes up to 2560 per side, seed reproducibility, and prompt extension at no extra charge.

이미지신규텍스트 투 이미지
$0.025/이미지부터보기
Z-Image LoRA

APIXO

Z-Image LoRA

Z-Image LoRA is an asynchronous text-to-image API returning one image per request, with aspect-ratio presets, exact pixel size control, seed reproducibility, and prompt extension at no extra charge.

이미지신규텍스트 투 이미지
$0.0156/이미지부터보기

bytedance

Seedance 2.0 Mini

Seedance 2.0 Mini is an APIXO lower-cost route for ByteDance Seedance 2.0 workflows, exposing text-to-video, first-and-last-frames, and omni-reference modes with optional sound, web search, and 480p/720p output.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.012/초부터보기
Wan 2.6 Image

Alibaba

Wan 2.6 Image

Wan 2.6 Image is Alibaba's text-to-image and image-to-image model with prompt, negative prompt, aspect ratio, batch count, and optional seed control.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.04/이미지부터보기
Wan 2.5 Image

Alibaba

Wan 2.5 Image

Wan 2.5 Image is Alibaba's batch-friendly image generation model for text-to-image and image-to-image workflows, defaulting to four generated images per request.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.03/이미지부터보기
Image Upscaler

APIXO

Image Upscaler

Image Upscaler is APIXO's single-image enhancement API for 2K, 4K, and 8K upscaling with JPG, PNG, and WEBP output format control.

이미지신규이미지 투 이미지
$0.01/이미지부터보기
Image Watermark Remover

APIXO

Image Watermark Remover

Image Watermark Remover is APIXO's authorized single-image cleanup API for removing sample marks or overlays from images you own or have permission to process.

이미지신규이미지 투 이미지
$0.015/이미지부터보기

kling

Kling 3.0 Turbo

Kling 3.0 Turbo is Kuaishou's fast video generation model for text-to-video and single-image image-to-video workflows with 720p/1080p output and 3-15 second clips.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.112/초부터보기
MiniMax Image 01

MiniMax

MiniMax Image 01

MiniMax Image 01 is MiniMax’s text-to-image and image-to-image model with prompt optimization, flexible size control, 1K/2K presets, and up to 9 generated images per task.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.0035/이미지부터보기

MiniMax

MiniMax Speech 2.8

MiniMax Speech 2.8 is an async text-to-speech API with HD and turbo quality modes, preset voices, custom MiniMax voice_id support, emotion control, pronunciation dictionaries, and multilingual output settings.

오디오신규텍스트 음성 변환텍스트 투 오디오
요청당 $0.06부터보기

MiniMax

MiniMax Voice

MiniMax Voice creates reusable custom voice IDs from text-described voice design or a single public reference audio clip, then returns preview audio for validation.

오디오신규텍스트 투 오디오
요청당 $0.5부터보기

hailuo

Hailuo 2.3

Hailuo 2.3 is MiniMax's async video model with standard and pro modes for text-to-video and image-to-video generation. Standard mode supports 6s/10s at 768p, while pro mode returns fixed 5s at 1080p.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.056/초부터보기

hailuo

Hailuo 2.3 Fast

Hailuo 2.3 Fast is MiniMax's speed-optimized image-to-video model with standard and pro modes. Standard supports 6s/10s at 768p, while pro returns fixed 6s output at 1080p.

비디오신규이미지 투 비디오
$0.032/초부터보기
Grok Image

xai

Grok Image

Grok Image is xAI's image generation model for text-to-image and image-to-image workflows with simple aspect-ratio control and async task delivery. Text-to-image returns 6 images per request on APIXO.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.02/이미지부터보기

Alibaba

Wan 2.2 Animate

Wan 2.2 Animate API is Alibaba's character animation model that combines one source image and one motion video to generate stylized animated outputs with animate/replace behavior.

비디오신규비디오 효과
$0.04/초부터보기

xai

Grok Video

Grok Video is xAI's async video generation model for text-to-video and image-to-video workflows, with optional continuation via task_id + index and style control.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.02/초부터보기

kling

Kling 3.0 Std

Kling 3.0 Std is Kuaishou's standard-quality video generation model with text-to-video, image-to-video, and motion-control modes. It supports clips up to 15 seconds with optional sound generation and flexible aspect ratios.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.084/초부터보기

MeiGen

InfiniteTalk

InfiniteTalk converts one photo plus audio into audio-driven talking or singing avatar videos with precise lip synchronization. Supports up to 10 minutes at 480p or 720p resolution.

비디오신규이미지 투 비디오
$0.03/초부터보기
Nano Banana 2

Google

Nano Banana 2

Nano Banana 2 is Google’s high-resolution image generation model with 1K/2K/4K output control, 20,000-character prompts, and support for up to 14 reference images. Google Search context is not exposed as an active control in APIXO’s current route.

이미지신규텍스트 투 이미지이미지 투 이미지
$0.05/이미지부터보기

vidu

Vidu Q3

Vidu Q3 is a per-second video generation model that combines standard and Turbo text-to-video plus image-to-video workflows in one API. It supports single-image animation, first-and-last-frame transitions, optional sound and BGM, and output up to 1080p.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.04/초부터보기

kling

Kling 2.5 Turbo Pro

Kling 2.5 Turbo Pro is Kuaishou's high-speed video model for text-to-video and image-to-video creation. It supports 5-10 second clips, optional tail-frame images, aspect ratio control for text-to-video, plus negative prompts and CFG scale guidance.

비디오신규텍스트 투 비디오이미지 투 비디오
비디오당 $0.3부터보기

Lightricks

LTX-2 19B

LTX-2 19B is Lightricks' open-source 19B diffusion transformer for cinematic video generation. It supports text-to-video and image-to-video workflows, LoRA conditioning, and high-fidelity outputs up to 1080p in the API.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.012/초부터보기

kling

Kling 2.1

Kling 2.1 is Kuaishou's multi-tier video model with Standard, Pro, and Master modes for image-to-video and text-to-video creation. It supports 5-10 second clips, optional tail images for Pro, and aspect ratio control for Master text-to-video.

비디오신규텍스트 투 비디오이미지 투 비디오
비디오당 $0.2부터보기

kling

Kling 2.6

Kling 2.6 is Kuaishou's native audio-visual video model that generates video, speech, sound effects, and ambience in one pass. It supports text-to-audio-visual and image-to-audio-visual creation with Chinese and English voice generation and up to 10-second clips.

비디오신규텍스트 투 비디오이미지 투 비디오
비디오당 $0.33부터보기
Suno V5

suno

Suno V5

Latest Suno text-to-music model that returns two polished songs per call with faster queues and richer vocals.

오디오신규텍스트 투 오디오
요청당 $0.12부터보기

Claude

Claude Opus 4.8

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

Claude

Claude Opus 4.7

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

Claude

Claude Opus 4.6

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

Claude

Claude Opus 4.5 20251101

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

Claude

Claude Sonnet 4.6

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.8부터보기

Claude

Claude Sonnet 4.5

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.8부터보기

Claude

Claude Sonnet 4.5 20250929

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.8부터보기

Claude

Claude Haiku 4.5 20251001

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.6부터보기

OpenAI

GPT-5.6 Sol

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

OpenAI

GPT-5.6 Terra

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.5부터보기

OpenAI

GPT-5.6 Luna

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.6부터보기

OpenAI

GPT-5.5

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $3부터보기

OpenAI

GPT-5.4

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.5부터보기

OpenAI

GPT-5.4 Mini

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.45부터보기

OpenAI

GPT-5.3-Codex

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.05부터보기

OpenAI

GPT-5.2

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.05부터보기

OpenAI

GPT-5.2-Codex

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.05부터보기

OpenAI

GPT-5 Mini

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.15부터보기

OpenAI

GPT-6 Astra

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $6부터보기

Gemini

Gemini 3.6 Flash

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.9부터보기

Gemini

Gemini 3.5 Flash

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.9부터보기

Gemini

Gemini 3.1 Pro Preview

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.2부터보기

Gemini

Gemini 3 Flash Preview

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.3부터보기

Gemini

Gemini 2.5 Pro

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.75부터보기

Gemini

Gemini 3.7 Flash

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.9부터보기

Gemini

Gemini 3.8 Flash

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $0.9부터보기

Grok

Grok 4.5

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.2부터보기

Grok

Grok 4.6

채팅, 에이전트, 추론, 구조화된 생성을 위한 OpenAI 호환 게이트웨이입니다.

LLM API
입력 토큰 1M개당 $1.2부터보기

Alibaba

Wan 3.0 Video LoRA

Wan 3.0 Video LoRA supports text-to-video, image-to-video, and reference-to-video workflows with smart duration. Reference mode accepts image, video, and audio inputs.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.0525/초부터보기

Alibaba

Wan 3.0 Video Pro LoRA

Wan 3.0 Video Pro LoRA raises the output ceiling to 4K across text-to-video, image-to-video, and reference-to-video workflows, with smart duration and image, video, and audio references.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.189/초부터보기

Black Forest Labs

FLUX 3

FLUX 3 is Black Forest Labs’ multimodal video model for text-to-video, image-to-video, and video extension with native audio and 720p or 1080p MP4 output.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.17/초부터보기

xai

Grok Imagine Video 1.5

Grok Imagine Video 1.5 is xAI's image-to-video model. It accepts 1–7 reference images and generates 1–15 second videos at 480p, 720p, or 1080p.

비디오신규이미지 투 비디오
$0.02/초부터보기

Alibaba

Wan 3.0 Video

Wan 3.0 Video supports text-to-video, image-to-video, and reference-to-video workflows with smart duration. Reference mode accepts image, video, audio, file, and link inputs.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.07/초부터보기

kling

Kling 3.0 Omni

Kling 3.0 Omni exposes nine modes through one endpoint: standard, pro, and 4K tiers across text-to-video, image-to-video, and reference-to-video workflows.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.084/초부터보기

MiniMax

MiniMax H3 LoRA

MiniMax H3 LoRA is the LoRA-stylized version of MiniMax H3 with 480p and 768p output. Reference images, audio, and video can condition reference-to-video generation.

비디오신규텍스트 투 비디오이미지 투 비디오
$0.04/초부터보기

Google

Veo 3.1

Google DeepMind’s upgraded AI video model with lite, fast, and quality routes, 4/6/8 second duration control, 720p/1080p/4k output, and multi-image reference workflows.

비디오텍스트 투 비디오이미지 투 비디오
비디오당 $0.15부터보기

Alibaba

Wan 2.5

Wan 2.5 is Alibaba's video generation model for text-to-video and image-to-video workflows, with optional audio input, 480p/720p/1080p output, 5 or 10 second clips, and prompt expansion.

비디오텍스트 투 비디오이미지 투 비디오
$0.05/초부터보기

OpenAI

Sora 2

Sora 2 is OpenAI’s synchronized short-video generation model on APIXO, supporting text-to-video and single-image-to-video with realistic motion, generated audio, landscape/portrait framing, and 4-, 8-, or 12-second outputs.

비디오텍스트 투 비디오이미지 투 비디오
$0.1/초부터보기
Seedream 4.5

seedream

Seedream 4.5

Seedream 4.5 is a powerful text-to-image and image-to-image AI model delivering high-quality image generation with support for 2K and 4K resolutions.

이미지텍스트 투 이미지이미지 투 이미지
$0.04/이미지부터보기