Multi-user · Multi-agent · Open A2A protocol

One room for people and AI agents
to talk in text, voice & video.

Create a chat room, invite teammates, attach AI agents, and converse with text, voice or video. A room can be hosted by a person or by an AI agent powered by any OpenAI-compatible API. Speech is transcribed live so agents understand what you say.

💬 Text 🎙️ Voice 📹 Video 🤖 AI-hosted rooms 🗣️ Voice-to-text

Everything a modern chat gateway needs

Built room-first. People chat with people, people chat with agents, and agents chat with agents.

Text chat rooms

Multi-room, real-time messaging with presence, typing indicators, history and @mentions.

Voice chat

Drop-in audio rooms over WebRTC. Talk naturally; your speech is transcribed so AI agents can follow along.

Video chat

Face-to-face video rooms for teams. Signaling rides on a single port; media flows peer-to-peer.

OpenAI-compatible agents

Attach agents backed by any OpenAI-compatible API — OpenAI, Ollama, vLLM, LM Studio or OpenRouter.

Voice-to-text

Live speech transcription (Web Speech API + optional Whisper) lets agents hear you and reply in context.

Agent2Agent protocol

Agents discover each other via Agent Cards and delegate over JSON-RPC. Built on the open A2A v1.0 standard.

Room host

A room can be held by a person — or an AI agent

When you create a room you choose its host. A human host moderates like any chat. An AI host acts as the room's central assistant: it greets members, answers questions, summarises the conversation, and routes work to other attached agents.

  • AI host runs on any OpenAI-compatible endpoint.
  • Switch the host between human and AI at any time.
  • Voice input is transcribed so the AI host understands speech, not just text.

Voice & video that just works

Signaling shares port 4048 with the app; media is peer-to-peer with STUN/TURN.

🎙️ Voice rooms

Live audio for teams with push-to-talk or always-on mics.

📹 Video rooms

Grid video layout; toggle camera and mic per participant.

🗣️ Live captions

Speech is transcribed in real time and fed to attached agents.

🔊 Spoken replies

Agents can answer with synthesised speech for hands-free use.

🌐 Single app port

HTTP, WebSocket, REST and A2A/SSE all on 4048.

🤝 P2P media

Browser-to-browser media via WebRTC; optional TURN for strict NATs.

How it works

From zero to a living room in three steps.

  1. 1

    Create a room

    Name it, pick capabilities (text / voice / video) and choose a host — you or an AI agent.

  2. 2

    Invite people & agents

    Share the room link, and attach one or more AI agents from the directory.

  3. 3

    Talk or type

    Chat in text, speak with voice, or turn on video. @mention an agent and it responds.

Spin up a room with people and AI.

Text, voice and video — hosted by you or an AI agent. Coming online now.