File size: 5,221 Bytes
534b431
 
 
 
 
 
dfedf76
534b431
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
dfedf76
534b431
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
66
67
68
69
70
71
72
73
74
75
76
77
78
79
80
81
82
83
84
85
86
87
88
89
90
91
92
93
94
95
96
97
98
99
100
101
102
103
104
105
106
107
108
109
110
111
112
113
114
115
116
117
118
119
120
121
122
123
124
125
126
<!doctype html>
<html>

<head>
	<meta charset="utf-8" />
	<meta name="viewport" content="width=device-width, initial-scale=1" />
	<title>Lyon Chatbox</title>
	<link rel="preconnect" href="https://fonts.googleapis.com">
	<link rel="preconnect" href="https://fonts.gstatic.com" crossorigin>
	<link href="https://fonts.googleapis.com/css2?family=Space+Grotesk:wght@400;500;600;700&family=Manrope:wght@400;500;600&display=swap" rel="stylesheet">
	<link rel="stylesheet" href="style.css" />
</head>

<body>
	<header class="hero">
		<div class="topline">
			<div class="brand">
				<span class="logo">πŸ€–</span>
				<span class="brand-name">Reachy Mini</span>
			</div>
			<div class="pill">Cascade pipeline Β· Swappable providers Β· Profiles & reactions</div>
		</div>
		<div class="hero-grid">
			<div class="hero-copy">
				<p class="eyebrow">ChatBox</p>
				<h1>Modular voice conversations for your robot.</h1>
				<p class="lede">
					A cascade pipeline β€” ASR β†’ LLM β†’ TTS β€” where each stage is a swappable provider. Mix cloud APIs and local models, define personality profiles, trigger live reactions, and let the LLM call tools to dance, emote, or look around.
				</p>
				<div class="hero-actions">
					<a class="btn primary" href="#highlights">Explore features</a>
					<a class="btn ghost" href="#story">See how it feels</a>
				</div>
				<div class="hero-badges">
					<span>Cascade ASR β†’ LLM β†’ TTS</span>
					<span>Local + cloud providers</span>
					<span>Live reactions while you speak</span>
					<span>LLM tool calls for movement</span>
				</div>
			</div>
			<div class="hero-visual">
				<div class="glass-card">
					<img src="docs/assets/reachy_mini_dance.gif" alt="Reachy Mini dancing" class="hero-gif">
					<p class="caption">Reachy Mini can move, dance, and emote while holding a natural conversation.</p>
				</div>
			</div>
		</div>
	</header>

	<section id="highlights" class="section features">
		<div class="section-header">
			<p class="eyebrow">What's inside</p>
			<h2>A modular conversational layer for your robot</h2>
			<p class="intro">
				Each piece of the pipeline is independent and swappable β€” pick the providers that fit your setup, budget, and latency needs.
			</p>
		</div>
		<div class="feature-grid">
			<div class="feature-card">
				<span class="icon">πŸ”—</span>
				<h3>Cascade pipeline</h3>
				<p>ASR β†’ LLM β†’ TTS in discrete stages. VAD segments audio, ASR transcribes, LLM reasons, TTS speaks β€” all streaming.</p>
			</div>
			<div class="feature-card">
				<span class="icon">πŸ”€</span>
				<h3>Swappable providers</h3>
				<p>Mix local models (Parakeet, Kokoro) with cloud APIs (Deepgram, Gemini, OpenAI, ElevenLabs). Switch from YAML or CLI flags.</p>
			</div>
			<div class="feature-card">
				<span class="icon">🎭</span>
				<h3>Personality profiles</h3>
				<p>Each profile bundles a system prompt, voice, enabled tools, and reaction triggers. Switch profiles live from the Gradio UI.</p>
			</div>
			<div class="feature-card">
				<span class="icon">⚑</span>
				<h3>Live reactions</h3>
				<p>Keyword and entity triggers fire while the user is still speaking β€” the robot reacts before the LLM even sees the text.</p>
			</div>
			<div class="feature-card">
				<span class="icon">πŸ› οΈ</span>
				<h3>LLM tool calls</h3>
				<p>The LLM can dance, play emotions, move the head, peek through the camera, or toggle head-tracking β€” all via tool calls.</p>
			</div>
		</div>
	</section>

	<section id="story" class="section story">
		<div class="story-grid">
			<div class="story-card">
				<p class="eyebrow">How it feels</p>
				<h3>From config to conversation in one command</h3>
				<ul class="story-list">
					<li><span>βš™οΈ</span> Pick your providers in cascade.yaml β€” local ASR, cloud LLM, local TTS, or any mix.</li>
					<li><span>πŸ—£οΈ</span> Start talking and watch streaming transcripts appear in the Gradio UI.</li>
					<li><span>⚑</span> The robot reacts to keywords instantly β€” before your sentence even finishes.</li>
					<li><span>πŸ’ƒ</span> Ask for a dance or an emotion; the LLM calls tools and the robot moves.</li>
				</ul>
			</div>
			<div class="story-card secondary">
				<p class="eyebrow">Where it shines</p>
				<h3>Demos, teaching, and custom robot personalities</h3>
				<p class="story-text">
					Build a pirate-themed guide, a museum docent, or a classroom assistant β€” each with its own voice, tools, and live reactions. Swap providers to balance cost and latency, run fully local on Apple Silicon, or mix in cloud APIs for maximum quality.
				</p>
				<div class="chips">
					<span class="chip">Cascade pipeline</span>
					<span class="chip">Local + cloud mix</span>
					<span class="chip">Profiles</span>
					<span class="chip">Live reactions</span>
					<span class="chip">Tool calls</span>
				</div>
			</div>
		</div>
	</section>

	<footer class="footer">
		<p>
			Lyon Chatbox by <a href="https://github.com/pollen-robotics" target="_blank" rel="noopener">Pollen Robotics</a>.
			Explore more apps on <a href="https://huggingface.co/spaces/pollen-robotics/Reachy_Mini_Apps" target="_blank" rel="noopener">Hugging Face Spaces</a>.
		</p>
	</footer>

</body>

</html>