PlatformModels & Agents

Build apps that generate, reason, and act.

The software a workforce builds now is AI-native: it generates content of every kind, reasons over its own data, and runs agents of its own that act. On Remy, that’s the default. One SDK carries every model in the catalogue, hundreds across 40+ vendors, an agent runtime, and a thousand integrations, with no keys, no per-vendor setup, and every model billed at cost with no markup.


01

Every model, every vendor, one call

Not a curated sample: every model in the catalogue, 305 in all, from 46 vendors, spanning language, image, video, voice, 3D, and retrieval, each reached through one call and one credential.

Language & reasoning

124 models · 18 vendors
OpenAI22 models
GPT-4 Turbo VisionGPT-4o Mini VisionGPT-4o VisionGPT-5GPT-5 miniGPT-5 nanoGPT-5.1GPT‑5.2 ProGPT 5.4GPT 5.4 ProGPT-5.5GPT 5.5 ProGPT 5.6 LunaGPT 5.6 SolGPT 5.6 TerraGPT OSS 120BGPT OSS 20Bo1o1-proo3o3-minio3-pro
xAI17 models
Grok 2 VisionGrok 3 FastGrok 3Grok 3 MiniGrok 3 Mini FastGrok 4Grok 4 FastGrok 4 Fast ReasoningGrok 4.1 FastGrok 4.1 Fast ReasoningGrok 4.20Grok 4.20 ReasoningGrok 4.3Grok 4.3 VisionGrok 4.5Grok 4.6Grok Build 0.1
Google16 models
Gemini 2.5 FlashGemini 2.5 Flash LiteGemini 2.5 Flash VisionGemini 2.5 ProGemini 2.5 Pro VisionGemini 3 FlashGemini 3.1 Flash LiteGemini 3.1 ProGemini 3.5 FlashGemini 3.5 Flash LiteGemini 3.6 FlashGemini 3.7 FlashGemma 3.2Gemma 4 26BGemma 4 31BGoogle Document AI
Alibaba Qwen10 models
Qwen3.5 Omni FlashQwen3.6 FlashQwen3.7 MaxQwen3.7 PlusQwen3 235BQwen3.5 397BQwen3.6-35B-A3BQwen3.8 27BQwen3.8 2.4TQwen3.8 Max
Anthropic10 models
Claude 4.5 HaikuClaude 4.5 OpusClaude 4.6 OpusClaude 4.6 SonnetClaude 4.7 OpusClaude 4.8 OpusClaude 4.5 SonnetClaude 5 OpusClaude 5 SonnetClaude 5 Fable
Mistral AI9 models
Ministral 3 14BMinistral 3 3BMinistral 3 8BMistral CodestralMistral Large 3Mistral Medium 3Mistral NemoMistral OCRMistral Small 3.1 (25.03)
DeepSeek8 models
DeepSeek-V3DeepSeek R1 TurboDeepSeek-R1DeepSeek V3.1DeepSeek V3.2DeepSeek V4 FlashDeepSeek V4 Flash 0731DeepSeek V4 Pro
Z.ai7 models
GLM 4.6GLM 4.6VGLM 4.7GLM 4.7 FlashGLM 5GLM 5.1GLM-5.2
Amazon4 models
Amazon Nova 2 LiteAmazon Nova LiteAmazon Nova MicroAmazon Nova Pro
Meta4 models
Llama 4 MaverickLlama 4 ScoutMuse Glimmer 30BMuse Spark 1.1
Moonshot AI4 models
Kimi K2.5Kimi K2.6Kimi K2.7 CodeKimi K3
NVIDIA4 models
Nemotron 3 Nano 30BNemotron 3 Super 120BNemotron 3 Ultra 550BNemotron 3.5 Lightning
Perplexity4 models
SonarSonar Deep ResearchSonar ProSonar Reasoning Pro
LlamaIndex1 model
LlamaParse
MiniMax1 model
MiniMax M3
Tencent1 model
Hy3
TwelveLabs1 model
Pegasus 1.5
Xiaomi1 model
MiMo V2.5 Pro

Image

67 models · 15 vendors
Black Forest Labs12 models
FLUX.1 [dev] LoRAFLUX.2 [dev] LoRAFLUX.2 [klein] 9BFLUX.2 [turbo]FLUX.1 [dev] Ultra-FastFLUX.1 Kontext [max]FLUX.1 Kontext [pro]FLUX.2 [max]FLUX 1.1 [pro]FLUX 1.1 [pro] UltraFLUX.2 [pro]FLUX.1 [schnell] LoRA
Google8 models
Gemini 2.5 Flash ImageGemini 3 Pro ImageGemini 3.1 Flash ImageGemini 3.1 Flash Lite ImageImagen 3Imagen 3 FastImagen 4 FastImagen 4 Ultra
Alibaba Qwen7 models
Qwen ImageQwen 2 ProQwen Image 3.0Qwen Image 3.0 ProQwen Image Edit PlusZ Image TurboZ Image Turbo Controlnet
Ideogram7 models
Ideogram UpscaleIdeogram V1 RemixIdeogram V2Ideogram V2 RemixIdeogram V3Ideogram V3 RemixIdeogram V4
Alibaba Wan4 models
Wan 2.5Wan 2.6Wan 2.7Wan 2.7 Pro
ByteDance4 models
Seedream 4.5Seedream 4.0Seedream 5.0 LiteSeedream 5.0 Pro
Luma4 models
Photon 1Photon 1 FlashUNI 1.1UNI 1.1 Max
OpenAI4 models
GPT Image LatestGPT Image 1GPT Image 1.5GPT Image 2
Recraft4 models
Recraft 20BRecraft Crisp UpscaleRecraft V4.1Recraft V4.1 Pro
Stability AI3 models
Stable Diffusion 3Stable Image CoreStable Image Ultra
xAI3 models
Grok ImagineGrok Imagine 2.0Grok Imagine Pro
Krea2 models
Krea 2 LargeKrea 2 Turbo
Kuaishou2 models
Kling Image O1Kling Image O3
Microsoft2 models
Microsoft MAI Image 2.5Microsoft MAI Image 2.5 Pro
WaveSpeed1 model
Chroma

Video

51 models · 16 vendors
ByteDance10 models
Omni Human 1.5DreamActor V2Seedance 1.5 ProSeedance 2.0Seedance 2.0 FastSeedance 2.0 MiniSeedance 2.0 Fast TurboSeedance 2.5Seedance 2.5 TurboLatentSync
Kuaishou9 models
Kling 2.6 Pro Motion ControlKling 3.0 Motion ControlKling O1Kling O3Kling 2.6Kling 3.0 ProKling 3.0 Turbo ProKling 3.0AI Avatar Standard
Alibaba Wan5 models
Wan 2.2Wan 2.5Wan 2.6Wan 2.7Wan 3.0
Google4 models
Gemini Omni FlashVeo 3.1Veo 3.1 FastVeo 3.1 Lite
Lightricks4 models
LTX-2 19bLTX-2 19B LipsyncLTX-2.3LTX-2.3 LoRA
Luma3 models
Ray 2Ray Flash 2Ray 3.2
Alibaba2 models
HappyHorse 1.0HappyHorse 1.1
MiniMax2 models
MiniMax H3Hailuo 2.3 Pro
OpenAI2 models
Sora 2Sora 2 Pro
PixVerse2 models
PixVerse V5.6PixVerse C1
Sync2 models
Sync Lipsync 2 ProSync Lipsync 3
xAI2 models
Grok Imagine 1.5Grok Imagine
Black Forest Labs1 model
FLUX 3 Video
HeyGen1 model
HeyGen Video Translate
Runway1 model
Gen-4 Turbo
WaveSpeed1 model
InfiniteTalk

Voice & audio

30 models · 11 vendors
OpenAI9 models
GPT Realtime 2.1GPT Realtime 2.1 MiniGPT-4o-mini TTSGPT TranscribeTTS-1TTS HDWhisper-1Whisper Large v3Whisper Large v3 Turbo
Google5 models
Gemini 2.5 Flash Native AudioGemini 3.1 Flash LiveGemini 3.1 Flash TTSLyria 3Lyria 3 Pro
ElevenLabs4 models
ElevenLabs MusicScribe v1Scribe v2ElevenLabs TTS
Alibaba Qwen3 models
Qwen3 ASR 1.7BQwen3 TTSQwen Audio 3.0 TTS Plus
MiniMax3 models
MiniMax Music 2.5MiniMax Music 3.0Minimax Speech 2.8 HD
Canopy Labs1 model
Orpheus 3B
Cartesia1 model
Cartesia Sonic 3
Deepgram1 model
Deepgram Nova-3
Hexgrad1 model
Kokoro 82M
Mistral AI1 model
Voxtral Mini 3B
xAI1 model
Grok Voice (Think Fast 2.0)

3D

5 models · 4 vendors
Tencent2 models
Hunyuan3D V2 Multi-ViewHunyuan3D v3
Meshy1 model
Meshy 6
Meta1 model
SAM 3D Objects
Tripo3D1 model
Tripo3D v2.5

Embedding & reranking

28 models · 9 vendors
Voyage AI8 models
Voyage 4Voyage 4 LargeVoyage 4 LiteVoyage Code 4Voyage Finance 2Voyage Law 2Voyage Rerank 2.5Voyage Rerank 2.5 Lite
Alibaba Qwen6 models
Qwen3 Embedding 0.6BQwen3 Embedding 4BQwen3 Embedding 8BQwen3 Reranker 0.6BQwen3 Reranker 4BQwen3 Reranker 8B
Cohere4 models
Cohere Embed 4Cohere Rerank 3.5Cohere Rerank 4 FastCohere Rerank 4 Pro
Google3 models
EmbeddingGemma 300MGemini EmbeddingGemini Embedding 2
OpenAI2 models
OpenAI Embedding 3 LargeOpenAI Embedding 3 Small
Perplexity2 models
Perplexity Embed v1 0.6BPerplexity Embed v1 4B
BAAI1 model
BGE-M3
NVIDIA1 model
Llama Nemotron Rerank VL 1B v2
Sentence Transformers1 model
all-MiniLM-L12-v2

The catalogue stays current: models are added as vendors ship them and retired as vendors drop them. An app can list what’s available for a task and choose the right one as it runs, instead of being locked to a single vendor chosen up front.

02

Everything an app can make

One SDK spans every kind of content an app can generate. Each is a single call, and each hands back something the rest of the app can use directly.

Text
Call any chat or reasoning model to get back structured data shaped to your own example, along with the raw text and a success flag.
Image
Generate from a prompt, edit and extend existing images, upscale to print resolution, strip backgrounds, or read an image back as data.
Video
Generate from a prompt or a still image, sync a speaker’s lips to any audio track, or analyze footage frame by frame.
Audio
Transcribe speech to text, synthesize a natural voice from text, or generate original music and sound.
3D & documents
Extract text and structure from PDFs and scans, generate charts and formatted documents, or build a 3D model from a prompt.

Once something exists, an app can reshape it. The editing toolkit you would normally reach for separate software to do is built in, a single call away.

Video
TrimMergeConcatenateResizeCropTranscodeOverlayWatermarkExtract audioExtract framesThumbnailsSubtitlesSpeed changeFace swapBackground removalStabilize
Audio
Extract from videoTranscodeTrimMergeNormalizeDenoiseVoice isolation
Image
ResizeCropRotateConvert formatBackground removalUpscaleCompressWatermark
Documents
Text & OCR extractionSplitMergeConvert to PDFGenerate PDFCharts

Everything an app creates or edits is hosted for you and handed back as a permanent link, with instant resize, crop, and format changes for images. No upload step, no storage to manage.

03

AI, built in, not bolted on

A Remy app isn’t only software. It’s software that generates content of every kind, reasons over the data it already holds, and runs agents of its own that take real action. That’s the kind of app a workforce builds now, and on Remy it’s the default, not a special project.

The usual way to add this is a project in itself. Text is one vendor, images another, transcription another, video another, each with its own SDK, its own keys to store and rotate, its own rate limits, its own line on the bill. Getting a single AI feature into production means wiring up and securing all of that before the feature itself exists.

04

Part of the platform, not an integration

Calling a model isn’t something a Remy app connects out to. It’s a capability of the platform the app already runs on. The app makes one call; the platform picks the model, runs it, meters it, and hands back a result the app can use. Three things follow, and none of them are the developer’s work:

No credentials in the app
The app never holds a provider key. There’s no vendor SDK to install, nothing to rotate, and nothing to leak.
One interface across every vendor
The same call shape reaches OpenAI, Anthropic, Google, and the rest. Switching models means changing an identifier, not rewriting an integration.
One metered edge
Because model calls route through the platform rather than out to a provider directly, every one is measured, attributed, and auditable. That’s the same chokepoint that produces the app’s unit economics and its audit trail.
05

Agents that belong to the app

Beyond single calls, the SDK can run agents of its own: agents that live inside the app and run when it does.

Task agents
A multi-step loop where an agent works toward a goal by calling tools, built right in. Give it a prompt, some starting input, a set of the app’s own actions it can use as tools, and a model. It reasons and calls those tools across as many turns as the work needs (20 by default, up to 100), so it finishes the job without running indefinitely. It shows its thinking and the tools it uses as it goes, remembers where it left off between runs, and reports exactly which tools it called and the tokens and cost it used. Because the tools are the SDK, one agent can search the web, generate an image, post to Slack, and write to the app’s database inside a single loop.
Packaged agents
Full MindStudio agents and workflows can be run from within an app, so a more involved automation is callable as a single step.

That’s the difference between an app that calls AI and an app that is an agent, and both are written in the same SDK.

06

Every system a business already runs on

The same SDK reaches more than 1,000 outside services, on the one credential, and every one is available to a task agent as a tool. An app doesn’t only reason. It reads from and writes to the systems a business already runs on, without assembling and securing a single integration itself.

Communication
SlackGmailDiscordGoogle MeetMicrosoft TeamsIntercomZendeskTwilioTelegram
CRM & Sales
SalesforceHubSpotZohoCalendlyTypeform
Productivity
NotionAirtableAsanaTrelloJiraLinearClickUpConfluence
Google Workspace
SheetsDocsDriveCalendarForms
Marketing & Social
MailchimpInstagramFacebookLinkedInXYouTubeTikTokReddit
Developer & Data
GitHubGitLabStripeShopifySnowflakeMongoDBSentryVercel

A representative sample. The full directory spans 1,000+ integrations and 850+ connector actions, all on the one credential.

07

Bring your own models

Everything above is the platform’s own model layer, and for most teams it’s all they’ll need: every model, one bill, no keys to manage. But an organization with its own model strategy or its own data-boundary requirements isn’t locked into it. Control escalates as far as it needs to:

Bring your own key
Route model calls through the organization’s own provider accounts, under its own contracts and terms. The same call, a different account behind it, and app code that doesn’t change.
Custom and self-hosted models
Register a model the organization hosts or has contracted, a fine-tune, a private or region-specific deployment, and call it the same way as any model in the catalogue.
Local, tunneled models
Reach a model running on the organization’s own hardware or inside its own network over a secure tunnel, so prompts and data never leave that environment at all.

Keys, routing, and usage are governed from the organization dashboard, not scattered across apps. The model layer is a choice, not a lock-in: the managed catalogue for reach, or the organization’s own infrastructure where policy requires it, without changing how apps are written.

08

Easy for the builder, accountable for the organization

Every model call goes through the platform. Nothing routes around it, which is what makes the rest of this true:

Private by default
Every model call runs under enterprise agreements with every model vendor. Data isn’t used to train their models. Where the assurance needs to be absolute, the same escalation from bring-your-own-models puts the data under the organization’s own contract, or keeps it inside its own network entirely.
Metered
The exact cost of every call is captured at the source and attributed by method, user, and model, split between building the app and running it. That’s true unit economics, not a provider bill reverse-engineered later. See the Audit Log deep dive.
Audited
Model use is activity on the platform edge, recorded like any other action: who ran what, when.
No scattered keys
With credentials held by the platform rather than pasted into apps, there’s no sprawl of provider keys across a fleet of tools to inventory, rotate, or lose.

Every app on Remy is AI-native by default: every model, every kind of content, an agent runtime, and a thousand integrations, built into the one SDK it’s written in, with no keys and no setup. And because all of it routes through the same platform, the organization gets the cost, the audit trail, and the visibility that usage produces without doing anything extra for it. The routing that makes an app easy to build is the same routing that makes it easy to govern.

Start building on Remy.

Everything you just read is standard in every app, running from the first deploy.

Start building← Back to the Platform