Beta Roboflow

Roboflow Serverless Video Streaming (WebRTC)

Stream a webcam, RTSP camera or file to Roboflow Cloud over WebRTC and run detection models or Workflows on every frame, getting annotated video and JSON results back.

Est. per minute$0.005 - 0.20
1 high-severity warning

Overview

Best for: Object detection, counting and safety monitoring on live cameras without running your own GPU.

At a glance

Session $/min$0.005
Entry plan $/mo$39
Free tierYes
CPU onlyYes
Concurrency10
WebRTCYes
Self-hostYes
Open weightsYes

Price is CPU streaming; GPU medium about $0.05-$0.10/min, GPU large up to $0.20/min. Free plan 10 credits/month. 10 concurrent streams per workspace by default. GPU large runs about 5 FPS.

Audio in

None

Audio out

None

Languages

n/a

Latency

Quality and FPS can take up to a minute to ramp up at 1080p30 (vendor docs).

Regions

us, eu, ap

Compliance

Not verified in this pass.

Features

  • any Roboflow model or Workflow
  • bidirectional video track with annotations
  • reliable JSON data channel
  • regions us / eu / ap
  • same SDK works against self-hosted Inference server

Pricing

WhatPriceUnitNotes
GPU small / medium / large1 credit per 80 / 60 / 30 minof videoBilled hourly by plan
CPU stream1 credit per 10 hof video
Credit price (Core plan)$2.86 to $3.90per creditMonthly packs; on-demand listed at $6, prepaid from $4
Core plan$39per monthFree tier: 10 credits/month
How the per-minute estimate was worked out

CPU 0.1 credit/h at $2.86 = $0.005/min; GPU large 2 credits/h at $6 = $0.20/min; GPU medium $0.05-$0.10/min

Free tier: Free plan includes 10 credits per month (about 10 hours of GPU-medium streaming).

Source: roboflow.com

Setup

  1. Create a Roboflow workspace, API key and a model or Workflow.
  2. pip install "inference-sdk[webrtc]" (or npm install @roboflow/inference-sdk).
  3. Point the client at https://serverless.roboflow.com and start a WebRTC stream with a StreamConfig (plan, region, outputs).
  4. Handle JSON predictions from the data channel; stop the session to stop billing.

Endpoint

https://serverless.roboflow.com (WebRTC signalling via SDK)

Authentication

Roboflow API key

Quick start python

# pip install "inference-sdk[webrtc]"  (SDK marked experimental)
# Sketch based on documented parameters; check field names
# against the examples/webrtc_sdk scripts in roboflow/inference.
from inference_sdk import InferenceHTTPClient
from inference_sdk.webrtc import StreamConfig, WebcamSource

client = InferenceHTTPClient(
    api_url="https://serverless.roboflow.com",
    api_key="YOUR_API_KEY",
)

session = client.webrtc.stream(
    source=WebcamSource(),
    workflow="your-workflow-id",
    workspace="your-workspace",
    config=StreamConfig(
        data_output=["predictions"],
        requested_plan="webrtc-gpu-small",
        requested_region="us",
    ),
)

@session.on_data("predictions")
def handle(predictions, metadata):
    print(metadata.frame_id, predictions)

session.run()

Written from the current docs. Check the vendor's SDK version before you ship.

Warnings

Streams die at zero credits

A user reported HTTP 402 CreditsExceededError mid-stream even with ~50 credits showing; keep a buffer and enable top-ups for production cameras.

Billed from connection, not from first result

Billing starts when the function spawns and WebRTC connects, including the up to one minute ramp-up.

Credit price varies a lot

The same credit costs $2.86 to $6 depending on pack versus on-demand; budget at the on-demand rate.

Experimental SDK

The WebRTC SDK is labelled experimental; pin versions.

Plus 9 warnings that apply to all avatars and live video APIs. See category warnings.

Limits

  • 10 concurrent streams per workspace by default
  • Billing starts when the serverless function spawns and WebRTC connects
  • Streams stop with HTTP 402 when credits run out

Models and products

NameStatusNotes
webrtc-gpu-mediumGADefault plan, recommended for most Workflows; 60 min per credit.
webrtc-gpu-smallGALower cost; 80 min per credit.
webrtc-gpu-largeGARequired for SAM3 and Rapid Models; about 5 FPS; 30 min per credit.
CPU video streamsGA10 hours per credit.

Docs and sources

Docs

Sources used

Not fully verified

Exact SDK call signature, max stream duration, when billing stops.

Similar avatars + video APIs

Spotted a wrong price or a dead link?