Z-Image
Alibaba Z-Image — leichtes, günstiges Text-zu-Bild-Modell mit mehreren Seitenverhältnissen für Poster, Produktbilder und Content-Produktion in großen Mengen.
Input
Output
ab
8 Credits /img
48 Gruppen
Alibaba Z-Image — leichtes, günstiges Text-zu-Bild-Modell mit mehreren Seitenverhältnissen für Poster, Produktbilder und Content-Produktion in großen Mengen.
Input
Output
ab
8 Credits /img
OpenAI GPT-Image 2 — Text-zu-Bild-Generierung und -Bearbeitung mit Inpainting und Referenzmischung bis 4K, tokenbasiert abgerechnet, für Werbedesign.
Input
Output
ab
30 Credits /img
Google Nano Banana — schnelles Bildmodell für Text- und Bild-zu-Bild-Generierung mit bis zu 5 Referenzbildern, ideal für alltägliche Kreativarbeit.
Input
Output
ab
15 Credits /img
Google Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image) — leichtes DeepMind-Modell für schnelle 1K-Generierung und latenzarme Prompt-Bearbeitung.
Input
Output
ab
45 Credits /img
OpenAI GPT Image 2.5 nutzt standardmäßig Flare für schnelle Bildgenerierung und -bearbeitung mit Referenzen, Masken, Transparenz, sechs Qualitätsstufen sowie 1K/2K/4K über APIAny.
Input
Output
ab
60 Credits /img
OpenAI GPT Image 2.5 Flare ist für schnelle Generierung und Bearbeitung mit mehreren Referenzen, Masken, transparentem Hintergrund, 1K bis 4K und Qualität von auto bis max ausgelegt.
Input
Output
ab
60 Credits /img
OpenAI GPT Image 2.5 Sunburst fokussiert präzise Generierung und anspruchsvolle Bearbeitung komplexer Vorgaben, Texte und lokaler Änderungen mit sechs Qualitätsstufen, 1K bis 4K, Referenzen und Masken.
Input
Output
ab
60 Credits /img
Google Nano Banana 2 — hochwertiges Bildmodell mit Text- und Bild-zu-Bild-Generierung, bis 4K Auflösung und 14 Referenzbildern, für E-Commerce und Design.
Input
Output
ab
25 Credits /img
Google Nano Banana Pro — Flaggschiff-Stufe der Nano-Banana-Reihe für Profi-Workflows, liefert 4K-Bilder in höchster Detailtreue mit bis zu 14 Referenzbildern.
Input
Output
ab
60 Credits /img
ByteDance Seedance 5.0, über den offiziellen Volcengine-Ark-Kanal — Text- und Bild-zu-Bild-Generierung bis 3K, ideal für E-Commerce- und Marketingbilder.
Input
Output
ab
70 Credits /img

ByteDance Seedance 2.0 — filmreifes Videomodell mit Soundtrack-Audio und Referenzeingabe, erzeugt 5 bis 15 Sekunden Clips bis 1080p für Werbespots.
Input
Output
ab
720 Credits /video

ByteDance Seedance 1.0 — leichtes, schnelles Bild-zu-Video-Modell, erzeugt 5 bis 10 Sekunden Clips zu geringen Kosten für Kurzvideos in großen Mengen.
Input
Output
ab
70 Credits /video

ByteDance Seedance 1.5 — verbesserte Bewegung und Detailtreue gegenüber 1.0, mit Soundtrack-Audio und Bild-zu-Video-Clips bis 12 Sekunden für Kurzvideos.
Input
Output
ab
150 Credits /video

Google Veo 3.1 Fast — schnelles Text- und Bild-zu-Video-Modell mit nativ synchronisiertem Audio in 16:9 oder 9:16 für Social-Media-Kurzvideos.
Input
Output
ab
330 Credits /video
DeepSeek V4.1 Flash: multimodal reasoning, coding and tool use with a 1M-token context window. Separate from DeepSeek V4 Flash.
Input
Output
Input
≈ $0.375-$0.75
Output
≈ $1.5-$3
Cache Read
≈ $0.0075-$0.015
DeepSeek V4 Flash — latenzarme Variante von DeepSeek V4, behält die Kern-Reasoning-Fähigkeiten bei und senkt Antwortzeit und Kosten für hohe Nebenläufigkeit.
Input
Output
Input
≈ $0.525-$1.05
Output
≈ $1.57-$3.15
Cache Read
≈ $0.0175-$0.035
NSFW Image Detection klassifiziert Risiken erwachsener Inhalte in hochgeladenen oder entfernten Bildern für Moderation, Review-Queues und sichere Veröffentlichung.
Input
Output
ab
0.5 Credits /img

Alibaba Wan 3.0 Prime erzeugt aus Prompts schnelle Videos von 2 bis 30 Sekunden in 480p, 720p oder 1080p, mit wählbarem Format und optionalem Audio.
Input
Output
ab
786 Credits /video

MiniMax H3 (Hailuo 3) Text-zu-Video über EvoLink — 4–15 s in 768p oder 2K, gleiche Parameter wie die offizielle Route, für Ads und Kurzfilme.
Input
Output
ab
760 Credits /video

MiniMax H3 (Hailuo 3) Bild-zu-Video über EvoLink — Steuerung von Erst-/Letztbild, 4–15 s in 768p oder 2K.
Input
Output
ab
760 Credits /video

MiniMax H3 (Hailuo 3) Referenz-zu-Video über EvoLink — bis zu 9 Bilder, 3 Videos und 3 Audios, 4–15 s in 768p/2K.
Input
Output
ab
1,520 Credits /video

ByteDance Seedance 2.5 Text-zu-Video über EvoLink — 4–30 s in 480p/720p/1080p mit Audio, duration -1 für automatische Länge.
Input
Output
ab
1,380 Credits /video

ByteDance Seedance 2.5 Bild-zu-Video über EvoLink — Erst-/Letztbildsteuerung, 4–30 s bis 1080p.
Input
Output
ab
1,380 Credits /video

ByteDance Seedance 2.5 Referenz-zu-Video über EvoLink — bis zu 30 Bilder, 10 Videos und 10 Audios.
Input
Output
ab
1,680 Credits /video
ByteDance Seedance 2.5 Video-Edit über EvoLink — vorhandene Clips per Prompt umschreiben, 4–30 s bis 1080p.
Input
Output
ab
1,680 Credits /video
ByteDance Seedance 2.5 Video-Extend über EvoLink — vorhandene Clips per Prompt fortsetzen, 4–30 s bis 1080p.
Input
Output
ab
1,680 Credits /video

ByteDance Seedance 2.0 Mini Text-zu-Video über EvoLink — günstige 4–15-s-Entwürfe in 480p/720p für Prompt-Tests.
Input
Output
ab
192 Credits /video

ByteDance Seedance 2.0 Mini Bild-zu-Video über EvoLink — Erst-/Letztbild, 4–15 s in 480p/720p für günstige Entwürfe.
Input
Output
ab
192 Credits /video

ByteDance Seedance 2.0 Mini Referenz-zu-Video über EvoLink — Bild-/Video-/Audio-Referenzen, 4–15 s in 480p/720p.
Input
Output
ab
240 Credits /video

Google Gemini Omni Text-zu-Video, über den APIPod-Kanal — erzeugt 6, 8 oder 10 Sekunden Clips in 720p oder 1080p direkt aus Text, ohne Referenzbild.
Input
Output
ab
500 Credits /video
GLM 5.3 Flash: multimodal reasoning, coding and tool use with a 1M-token context window. Thinking is always enabled with low, high or max effort.
Input
Output
Input
≈ $0.14-$0.28
Output
≈ $0.49-$0.98
Cache Read
≈ $0.0405-$0.081
Google Gemini 3.5 Flash — schnelles, kosteneffizientes multimodales Chat-Modell mit Bild- und Texteingabe und einer Million Token Kontextfenster.
Input
Output
Input
≈ $0.938-$1.88
Output
≈ $6.25-$12.5
Cache Read
≈ $0.0938-$0.188
MiniMax M3 — großes Sprachmodell mit einem Kontextfenster von einer Million Token, stark im Verständnis langer Dokumente, Reasoning und Tool-Nutzung.
Input
Output
Input
≈ $1.2-$2.4
Output
≈ $5-$10
Cache Read
≈ $1.5-$3
DeepSeek V4 — auf Reasoning fokussiertes Sprachmodell mit 128K-Token-Kontextfenster, stark bei Codegenerierung und komplexer Logik, sehr günstig.
Input
Output
Input
≈ $0.212-$0.425
Output
≈ $0.475-$0.95
Cache Read
≈ $0.025-$0.05
OpenAI GPT-5.5 — Flaggschiffmodell für komplexes Reasoning, Codegenerierung und mehrstufige Anweisungen, mit 400K Token Kontextfenster für anspruchsvolle Apps.
Input
Output
Input
≈ $4.25-$8.5
Output
≈ $25-$50
Cache Read
≈ $0.425-$0.85
OpenAI GPT-5.6 Sol ist die leistungsstarke GPT-5.6-Stufe für komplexes Reasoning, Codegenerierung, Agenten und anspruchsvolle Produktionsautomatisierung.
Input
Output
Input
≈ $6-$12
Output
≈ $40-$80
Cache Read
≈ $0.85-$1.7
OpenAI GPT-6 — a flagship model for complex reasoning, code generation, and multi-step instructions, with a 400K-token context window for demanding apps
Input
Output
Input
≈ $3.5-$7
Output
≈ $30-$60
Cache Read
≈ $0.531-$1.06
xAI Grok 4.5 ist ein leistungsstarkes Modell für Reasoning, Codeanalyse, lange Kontexte und toolgestützte Assistenten.
Input
Output
Input
≈ $0.6-$1.2
Output
≈ $1.2-$2.4
Cache Read
≈ $0.6-$1.2
Grok 4.6 is an advanced AI model designed to deliver fast, writing, coding, research, data analysis, and creative problem-solving through natural conversation
Input
Output
Input
≈ $0.6-$1.2
Output
≈ $1.2-$2.4
Cache Read
≈ $0.6-$1.2
A premium reasoning route for visual front-end prototyping, repository-scale coding, large evidence sets, long-running agents, and complex knowledge work that benefits from a 1.05M-token working context.
Input
Output
Input
≈ $2-$4
Output
≈ $10-$20
Cache Read
≈ $0.2-$0.4
Google Gemini 2.5 Flash Lite — extrem günstiges, latenzarmes Chat-Modell mit einer Million Token Kontextfenster, ideal für hochfrequente Alltagsaufgaben.
Input
Output
Input
≈ $0.1-$0.2
Output
≈ $0.15-$0.3
Cache Read
≈ $0.01-$0.02
Google Gemini 3.1 Flash Lite — verbessertes Reasoning gegenüber der 2.5-Generation bei günstigem Preis, balanciert Geschwindigkeit und Qualität.
Input
Output
Input
≈ $0.15-$0.3
Output
≈ $1-$2
Cache Read
≈ $0.0125-$0.025
Google Gemini 3.1 Pro — multimodales Flaggschiffmodell von Gemini mit starkem komplexem Reasoning und langer Kontextverarbeitung für Produktivanwendungen.
Input
Output
Input
≈ $1.25-$2.5
Output
≈ $7.5-$15
Cache Read
≈ $0.625-$1.25
OpenAI GPT-4o mini — schnelles, günstiges multimodales Chat-Modell mit zügigen Antworten, ideal für alltägliche Fragen und leichte Programmierhilfe.
Input
Output
Input
≈ $0.275-$0.55
Output
≈ $0.412-$0.825
Cache Read
≈ $0.0138-$0.0275
OpenAI GPT-5.4 — leistungsstarkes Modell für fortgeschrittenes Reasoning, Codegenerierung und Agenten-Workflows, mit 400K Token Kontextfenster.
Input
Output
Input
≈ $2.5-$5
Output
≈ $15-$30
Cache Read
≈ $0.25-$0.5
Anthropic Claude Opus 4.8 — Flaggschiffmodell von Claude für anspruchsvollste Reasoning-, Coding- und Langtextaufgaben, mit exzellenter Kontextverarbeitung.
Input
Output
Input
≈ $2.5-$5
Output
≈ $12-$24
Cache Read
≈ $0.25-$0.5
Anthropic Claude Opus 5 from APIAny — Anthropic’s newest Opus-tier flagship for the hardest coding, long-running agents, and judgment-heavy review.
Input
Output
Input
≈ $3.75-$7.5
Output
≈ $18.75-$37.5
Cache Read
≈ $0.375-$0.75
Anthropic Claude Sonnet 4.6 — ausgewogenes Flaggschiffmodell mit starkem Reasoning bei produktionsreifer Geschwindigkeit und Kosten für stabile Großeinsätze.
Input
Output
Input
≈ $1.5-$3
Output
≈ $7.5-$15
Cache Read
≈ $1-$2
APIAny ist ein Live-Katalog öffentlicher Chat-, Bild-, Video-, Audio- und Sicherheitsmodelle, die Sie mit einem OpenAI-kompatiblen API-Schlüssel aufrufen.
Von APIAny Editorial · Aktualisiert
Filtern Sie nach Typ oder Anbieter, vergleichen Sie Credit-Preise und öffnen Sie eine Modellseite für Playground-Tests und Request-Beispiele.
Quellen
Die Zitate stammen aus offizieller API-Dokumentation, gegen die APIAny implementiert.
The Chat Completions API endpoint will generate a model response from a list of messages comprising a conversation.
The Gemini API provides access to Google's most capable generative AI models.
The Images API provides several endpoints that let you generate images from text prompts or create edits of existing images.
Der APIAny-Modellkatalog ist die Live-Liste öffentlicher Chat-, Bild-, Video-, Audio- und Sicherheitsmodelle, die Sie mit einem API-Key aufrufen.