Built with Agora Conversational AI

The conversational layerfor Physical AI.

Build connected devices that listen, reason, speak, and act with Agora Conversational AI, flexible model providers, tools, and device operations in one platform.

View on GitHub

Agora Conversational AI signal path

A live path from physical input to an intelligent response.

Ready

Physical device

Microphone + speaker + sensors

Vibbo

Agent + device orchestration

Agora Conversational AI

Realtime conversation intelligence

Tools & actions

Device + cloud execution

One continuous voice loop

From microphone input to a spoken response and real-world action.

  1. 01

    Hear clearly

    Wake word, streaming audio, and VAD designed for voice input in real rooms.

  2. 02

    Stay real time

    Agora Conversational AI keeps conversations responsive across changing networks.

  3. 03

    Use your AI stack

    Choose the LLM, ASR, TTS, voice, memory, and persona for every agent.

  4. 04

    Speak and act

    Stream natural speech, handle interruptions, and call device or cloud tools.

A complete real-time voice stack, ready for your product

  • Agora Conversational AI
  • Agora RTC
  • Streaming ASR
  • LLM
  • TTS
  • MCP Tools

Capabilities

Everything the voice loop needs, already wired.

Vibbo connects streaming speech recognition, your model, speech synthesis, tools, and hardware operations around Agora Conversational AI.

Conversational runtime

Real-time voice conversation

Streaming ASR → LLM → TTS with wake word, natural turn taking, and interruption support for responsive conversations.

Live voice session
Agora RTC
Voice detectedInterruptible
ASR

Streaming

LLM

Reasoning

TTS

Speaking

Provider layer

Bring the AI stack you already use.

Choose speech recognition, language models, and voices per agent. Keep the hardware integration unchanged.

ASR

Streaming speech recognition

Configured for this agent

LLM

OpenAI-compatible model

Configured for this agent

TTS

Neural voice

Configured for this agent

Swap providers without rebuilding device firmware.
Device operations

Device and OTA management

Register boards, group them by agent, and roll firmware out over the air.

Workshop agent3 devices

ESP32-S3

Kitchen · firmware 1.6.2

Online

Desk companion

Studio · firmware 1.6.2

Online

Voice speaker

Lab · firmware 1.5.9

Update ready
Session trace

Observable sessions

Inspect every turn from transcript to tool result, then trace what the agent heard, decided, and said.

Voice stream

Realtime

Turn state

Listening

Tool calls

1

Dim the living room lights a little.

set_device_state · brightness: 30%

Done — the living room lights are now at 30%.

System architecture

One live path from microphone to intelligence.

A replaceable hardware-to-cloud path connects the microphone to a live AI agent, while Vibbo keeps every operational control in one place.

01

Hardware

MCU + microphone + speaker + display

02

Firmware

Wake word, audio codec, connectivity

03

Protocol

WebSocket / MQTT / MCP

04

Voice AI cloud

Agora Conversational AI + ASR + LLM + TTS + tools

Built around Agora Conversational AI. Devices and agents meet in Agora real-time channels. Agora Conversational AI orchestrates the live voice session while Vibbo manages agent configuration, providers, tools, devices, and operations.

Built for hardware

Voice experiences that live beyond the screen.

Use one voice platform across products that need to listen, reason, speak, and act in the physical world.

AI companions

Create expressive desktop companions and ambient devices with persistent personalities, voices, memory, and tools.

Smart home devices

Add natural voice control and tool execution to hubs, appliances, panels, and room devices.

Learning hardware

Build interactive learning devices that listen, explain, ask follow-up questions, and adapt their voice experience.

Toys and robots

Give connected toys and robots safe, configurable characters without embedding the whole AI stack in firmware.

Quickstart

From a bare board to a talking device in four steps.

Move from a development board to a managed voice agent without rebuilding the live audio and operations stack.

  1. 01

    Flash the firmware

    Use the prebuilt firmware for ESP32-S3, C3, or P4 — or compile your own from source.

  2. 02

    Pair the device

    The board joins Wi-Fi and shows a six-digit code. Enter it in the console to bind it to your account.

  3. 03

    Configure the agent

    Pick the language model, speech provider, voice, and persona. Changes apply on the next turn.

  4. 04

    Start talking

    Wake the device and speak. Every session, tool call, and transcript stays inspectable.

FAQs

Practical answers for product and hardware teams

Give your hardware a voice.

Create an agent, connect a device, and bring a production voice experience to your hardware.

Voice AI product updates

Follow Vibbo as the platform grows

Get product updates, hardware integration notes, and new voice AI capabilities.