Unified inference API
Runware exposes one endpoint and one request shape across image, video, audio, 3D, and text tasks, so teams can integrate once and switch models by changing the model string.
Runware is a generative AI inference platform that gives developers a single API for image, video, audio, 3D, and text workloads. Its documentation describes a shared request structure across modalities, with models addressed by identifier and returned through the same API layer.
The product is aimed at teams that want to ship AI features without managing their own GPU infrastructure. Runware combines model access, request routing, managed infrastructure, and usage-based pricing, and it documents both REST and WebSocket flows along with webhook and polling options for asynchronous results.
Runware exposes one endpoint and one request shape across image, video, audio, 3D, and text tasks, so teams can integrate once and switch models by changing the model string.
The platform documents specific task types for image inference, video inference, audio inference, 3D inference, and text inference, each with its own structured payload and response.
Requests can be sent as REST calls for stateless jobs or over WebSockets for persistent sessions, and async tasks can return through webhooks or polling.
The docs describe shared request structure, model schemas, and LLM-readable documentation, which helps developers wire the API into tools and agents more quickly.
Runware says it supports open-source models, partner models, community models, and custom uploads, with standardized addressing across its model catalog.
The Sonic Inference Engine combines Runware-owned hardware and software with preloaded models, region-aware routing, and custom infrastructure design for inference workloads.
Build image generation or editing features with one endpoint, including tasks like text-to-image, image-to-image, inpainting, outpainting, upscaling, and background removal.
Add video, audio, or 3D generation into a product without building separate backends for each modality. The same platform supports structured requests and model-specific parameters.
Use Runware when you need low-latency production inference at scale and want managed infrastructure instead of provisioning and tuning your own GPUs.
Connect the API to coding tools, agent workflows, or app frameworks that already work with the documented integrations and standard request shapes.
Evaluate models in the Playground and then switch to the API once a team has chosen a model and confirmed the output quality and cost.
Runware provides a single API for image, video, audio, 3D, and text generation. The same request shape is used across modalities, with models identified by model ID and tasks sent to the API as JSON.
Pricing is pay-as-you-go. You only pay for successful API requests, and costs vary by model and parameters such as resolution, duration, and quality settings. The pricing page also says new users receive $2 in free credits.
The docs page lists TypeScript, Python, CLI, MCP, ComfyUI, and Vercel AI as supported integration paths, and the site also mentions compatibility with tools such as Claude Code, Cursor, Claude Desktop, ChatGPT, OpenAI-compatible workflows, and several automation or app platforms.
Runware’s documentation covers image generation, image editing, advanced control, video generation, LLMs, media processing, media analysis and safety, audio generation, and 3D asset generation.
The site says Runware offers a REST API for stateless work, WebSockets for persistent low-latency sessions, webhook delivery for async results, and streaming for text inference over SSE. The docs are also structured for LLMs to read end to end.
RLAMA 是一款本地 AI 平台,适用于在 macOS、Linux 和 Windows 上构建 RAG 系统与智能体。支持本地处理、交互式终端工作流和用于文档问答及多智能体自动化的 HTTP API。
Orca 是面向编码代理的 Agent 开发环境,支持在隔离的 git worktree 中并行运行多个 CLI agent,并提供桌面端与移动端协作流程。
Firebase Studio 是一款基于浏览器的全栈应用开发工作区,支持 Gemini 辅助编写代码、应用预览、云模拟器、导入仓库、原型开发、协作与从浏览器部署。
EZsite AI 是一款 AI 网站构建器,可将 URL 转为完整的 React 或 Vue.js 应用。支持托管、自定义域名、代码导出及面向团队的后端功能,帮助快速交付可部署结果。
AI Magicx 是一体化 AI 工作区,集聊天、图片、视频、语音、音乐、邮件和开发任务于一处,帮助创作者、团队和开发者集中使用多种模型,无需在多个工具和订阅间切换。
Paper 是一款设计工具,连接画布、代码和 AI agent,让团队在一个工作流中完成创建、协作与交付。支持桌面应用、基于 MCP 的 agent 访问,以及真实内容和设计 token 工作流。