Real-Time Voice & Video

Updated

Build live audio, video, broadcast, and edge-device experiences on Agora's real-time media foundation.

Real-Time Audio & Video is the capability layer for products where live media is the primary user experience. This category covers two-way voice and video interaction, host-audience live experiences, large-scale broadcast delivery, and device-oriented real-time connectivity.

Agora's product map for this layer includes:

  • Voice Calling
  • Video Calling
  • Interactive Live Streaming
  • Fusion CDN broadcast delivery
  • IoT / RTSA edge-device connectivity

What this category is good for

This category is the right entry point when you need:

  • one-to-one or group voice and video calling
  • meetings, classrooms, and telehealth sessions
  • creator-led or host-audience live interaction
  • broadcast-style delivery to larger audiences
  • smart-device, camera, or edge-terminal media connectivity

The shared requirement is stable, low-latency media transport with clear session behavior across networks, devices, and participant roles.

Capability map

Voice Calling

Low-latency voice interaction for calling, social audio, customer communication, and lightweight live sessions.

Video Calling

Real-time face-to-face interaction for meetings, classrooms, telehealth, and collaboration products.

Live Streaming

Host-audience interaction for creator events, live communities, and interactive broadcast-style products.

Broadcast

Fusion CDN-based distribution for larger viewer populations and playback-oriented live delivery.

Edge Devices

RTSA-based device connectivity for smart cameras, displays, embedded endpoints, and IoT-oriented real-time media.

Start here

RTC

Fusion CDN

RTSA

Key characteristics

Shared media foundation

Use the same real-time transport model across calling, live interaction, broadcast expansion, and device connectivity.

Flexible interaction patterns

Support peer-to-peer, group, host-audience, and device-oriented topologies without switching to a different conceptual model.

Scale from interaction to distribution

Start from interactive sessions, then extend into larger audience delivery or downstream cloud media services.

Cross-device coverage

Reach web, mobile, desktop, and edge-device products with consistent session and media semantics.

Works with adjacent layers

Combine media transport with messaging, AI, transcription, recording, and analytics as product requirements grow.

Network resilience

Keep the live experience usable under real-world packet loss, geography, and device constraints.