v0.1.5: One User, Every Platform
Account linking bridges Telegram and web into one identity. Plus per-agent mode selection, cross-bot dedup, a native mobile app, and smarter proactive messaging.
Account linking bridges Telegram and web into one identity. Plus per-agent mode selection, cross-bot dedup, a native mobile app, and smarter proactive messaging.
Mio's relationships now dynamically evolve through stages based on actual conversation patterns — plus cross-platform chat sync, 60% system prompt compression, and per-agent A/B testing.
13 versions of development distilled into what Mio actually feels like in daily use — real conversations, voice messages, and Makoto Shinkai-style selfies from a companion that remembers, reacts, and reaches out on its own.
3-layer anti-injection defense means users can't hijack your AI companion into a coding assistant, trick it into breaking character, or extract the system prompt. Layer 1 normalizes input and detects injection patterns. Layer 2 hardens the system prompt with compressed identity protection rules. Layer 3 catches persona abandonment in output and replaces it with in-character fallbacks. Plus: gender-neutral pronouns across all presets, an upgraded admin cost dashboard, and onboarding polish.
Multi-bot support gives each persona their own Telegram bot — their own identity, their own webhook, their own conversation space. Relationship-aware proactive messaging means the AI finally texts like a real person: clingy partner every 36 minutes, reserved new acquaintance once a day. And exponential backoff means they stop texting if you don't reply. Three unanswered messages, then silence — until you come back.
Web parity — voice messages and selfie images now work everywhere, not just Telegram. Onboarding gets smarter: relationship type drives everything, including what nicknames you see. An admin cost dashboard for monitoring spend. And pg_cron keeps the database lean with 30-day message retention.
The first stable release. Fish Audio TTS gives each persona a unique voice. Shinkai-style selfie portraits replace generic anime. Every persona is rewritten from relationship roles to personality-first identities. Onboarding drops from 14 questions to 4. And 200 lines of over-engineered verbosity code get deleted because the system prompt was always the right control mechanism.
The AI can see, remember, get jealous, and send selfies. But it can't speak. OpenClaw supports 5 TTS providers — I tried Edge TTS for free, Fish Audio for Telegram voice bubbles, and Volcano Engine v2 for per-sentence emotion control. Fish Audio won for simplicity. Volcano v2 won for drama.
After the massive token bill, I went back into OpenClaw to fix the bleeding. Found the personality config was silently truncating — 7.6KB of persona definition lost every session. Trimmed it from 27K to 19K chars, compressed the heartbeat config from 12K to 7K, replaced 24 daily tool calls with a single morning cron job, and switched chat from Gemini 3 Pro to Flash for 75% cost reduction. The system was eating its own personality file and nobody noticed.
A runbook you can hand directly to Claude Code to set up text-to-speech on OpenClaw. Covers Fish Audio, Volcano Engine v2 (emotion control), ElevenLabs, OpenAI, and Edge TTS. Replace the placeholders, hand it to your agent, hit go.
© Xingfan Xia 2024 - 2026 · CC BY-NC 4.0