v2.1.5 โ€” 5,966 tests passing

NeuralCleave

The intelligence layer for every conversation.

Your personal AI assistant that brings every AI model, conversation, memory, and workflow together in one private, extensible intelligence layer.

๐Ÿง 

3-Tier Memory

Redis, Qdrant & SQLite โ€” your AI actually remembers you.

๐Ÿ”€

Smart Routing

10 task types auto-routed to the optimal model.

๐Ÿ“ก

32 Channels

Telegram, WhatsApp, Discord, Slack and 29 more.

๐Ÿค–

13 Providers

Anthropic, OpenAI, DeepSeek, Ollama and more.

๐ŸŽ™

Voice Ready

Whisper STT, 3-tier TTS, wake word on all platforms.

๐Ÿ”Œ

Plugin SDK

Typed ABCs, entry-points, hot-reload, Hub marketplace.

32Channel Adapters
13LLM Providers
5,076Tests Passing
41REST Endpoints
3-tierMemory System
v2.1Current Release

Why NeuralCleave

Everything you'd want. Nothing you don't.

Built to surpass every personal AI assistant on the market โ€” feature by feature, test by test.

๐Ÿง NeuralCleave leads

3-Tier Memory

Redis (hot session) + Qdrant (vector semantic) + SQLite (long-term). Per-agent namespace isolation. Compaction + archiver included. Works fully offline.

๐Ÿ”€NeuralCleave leads

Task-Aware Model Routing

10 task types automatically routed to the optimal provider โ€” complex reasoning to Claude Opus, code to DeepSeek, cheap inference to Ollama. Privacy mode forces all calls local.

โœจUnique

ReflectionEngine

Every response scored across 4 dimensions โ€” Relevance, Completeness, Accuracy, Tone โ€” on a 0โ€“100 scale. Low-scoring responses are automatically regenerated once.

๐ŸŽ™๏ธNeuralCleave leads

Voice Pipeline

Local Whisper STT (tinyโ†’large-v3), 3-tier TTS (ElevenLabs โ†’ Kokoro โ†’ system), OpenWakeWord always-on detection, voice cloning. Everything runs on-device.

๐ŸงฉNeuralCleave leads

Plugin SDK

Typed Python ABCs with PEP 451 entry-points. Hot-reload without restart. Hub marketplace with PackageScanner safety gate โ€” no supply-chain risk.

๐Ÿค–Parity

Multi-Agent Orchestrator

Named nodes with model overrides, task/keyword/channel routing, priority + round-robin. Per-node LRU memory namespace. Full REST + CLI management.

๐Ÿ“ŠNeuralCleave leads

Prometheus Observability

13 built-in metrics โ€” counters, gauges, histograms โ€” at /api/v1/metrics. Structured JSON logging with ContextLogger. Pre-built Grafana dashboard included.

๐Ÿ–ผ๏ธParity

Live Canvas

The agent can render text, markdown, code, tables, charts, and images to a live visual block stream. Real-time WebSocket broadcast to any connected viewer.

๐Ÿ“ฑParity

Desktop + PWA

Tauri v2 native app for Windows / macOS / Linux. PWA served from the gateway โ€” installable on iOS and Android from any browser. No app store needed.

Pipeline

How NeuralCleave works

Every message you send goes through a four-stage pipeline designed to give you the best possible response, every time.

01

Message arrives on any channel

You send a message from Telegram, WhatsApp, Discord, Slack, voice, or any of the 32 supported adapters. The gateway picks it up instantly.

02

Memory enriches the context

3-tier memory (Redis hot cache โ†’ Qdrant vector search โ†’ SQLite long-term) retrieves everything relevant โ€” past conversations, preferences, and facts you've shared.

03

Smart routing picks the best model

The task router classifies your request across 10 task types โ€” coding, reasoning, summarization, image analysis, and more โ€” then sends it to the optimal LLM from 13 providers.

04

ReflectionEngine scores and replies

Before sending, the 4-dimension ReflectionEngine scores the response for quality, relevance, safety, and tone. If it fails the threshold, it auto-retries with a refined prompt.

Any Channelโ†’Gatewayโ†’3-Tier Memoryโ†’Task Routerโ†’Best LLMโ†’ReflectionEngineโ†’Reply

Use Cases

Built for real workflows

NeuralCleave adapts to how you actually work โ€” not the other way around.

๐ŸŒ…
Daily Briefing

Your morning, automated

NeuralCleave pulls your calendar, checks the weather, skims priority emails, and sends a crisp summary to your Telegram at 8 AM โ€” every day, no setup beyond a TOML file.

Scheduler โ†’ Telegram13 LLM summarizationCalendar + email tools
๐Ÿ’ป
Dev Assistant

Code help, wherever you code

Ask questions in Slack, get PR review comments inline, run terminal commands by voice, and keep a running context of your current project โ€” across every conversation.

Slack + voice inputOllama for local privacyGitHub plugin SDK
๐Ÿ”ฌ
Research Agent

Research with persistent memory

Ask anything across sessions. NeuralCleave remembers what you've already found, cross-references Qdrant vector memory, and saves structured notes to Notion or Obsidian.

3-tier vector memoryDeepSeek reasoning modelNotion plugin
๐ŸŒ
Multi-language

Chat in any language

Switch between languages mid-conversation. NeuralCleave detects locale per channel, adapts tone, and maintains separate memory namespaces for each agent node.

32 channel adaptersPer-node namespace isolationMoonshot + Qwen + ERNIE
๐Ÿ 
Home Automation

Control your environment by voice

Wake word detection triggers NeuralCleave on any platform. Ask it to set timers, play music, dim lights, or brief you on your day โ€” hands-free, locally processed.

OpenWakeWordWhisper STT + 3-tier TTSTool plugin SDK
๐Ÿ›ก๏ธ
Privacy First

Fully local โ€” zero cloud required

Run the entire stack on your machine with Ollama. No API keys sent to third parties, no data leaving your device. Toggle a privacy flag in TOML and you're air-gapped.

Ollama local inferenceprivacy_mode = trueSelf-hosted Redis + Qdrant

Universal Reach

32 Channel Adapters

Every adapter produces a normalised InboundMessage โ€” one interface for every platform. Configure any channel with a single TOML block.

โœˆ๏ธTelegram
๐ŸŽฎDiscord
#๏ธโƒฃSlack
๐Ÿ’ฌWhatsApp
๐Ÿ“งEmail
๐Ÿ“ฑSMS / Twilio
๐ŸŸฆMicrosoft Teams
๐Ÿ”ตGoogle Chat
๐ŸŸฉMatrix
๐Ÿ–ฅ๏ธIRC
๐Ÿ”’Signal
๐ŸŸฉLINE
๐Ÿฆ…Feishu / Lark
๐ŸŽiMessage
๐Ÿ’พSynology Chat
โšกNostr
๐Ÿ“žTwilio Voice
๐Ÿ‡ป๐Ÿ‡ณZalo OA
๐Ÿ’ผWeChat Work
๐ŸงQQ Bot
๐ŸŒTlon / Urbit
๐Ÿ’™Facebook Messenger
๐Ÿš€Rocket.Chat
๐ŸŸฃTwitch
๐Ÿฆ‹Bluesky
๐Ÿ’œViber
๐Ÿ’ฌXMPP / Jabber
โš“Mattermost
๐Ÿ˜Mastodon
โ˜๏ธNextcloud Talk
๐ŸชGeneric Webhook
๐Ÿ”ŒBuilt-in WS / REST

~/.neuralcleave/config.toml

[channels.telegram]
bot_token = "ENV:TELEGRAM_BOT_TOKEN"

[channels.discord]
bot_token = "ENV:DISCORD_BOT_TOKEN"

[channels.slack]
app_token = "ENV:SLACK_APP_TOKEN"

Model Agnostic

13 LLM Providers. One Interface.

Task-aware routing sends each request to the optimal model. Switch providers at runtime with a single API call. All keys use ENV: resolution โ€” no secrets in files.

A
AnthropicPrimary
Claude โ€” complex reasoning, extended thinking
G
Google Gemini
Summarisation, cheap inference, general tasks
O
OpenAI
Code review, broad fallback chain
D
DeepSeek
Code generation & review
M
Mistral AI
Complex reasoning fallback
x
xAI Grok
Complex reasoning
C
Cohere
Summarisation
M
Moonshot / Kimi
General tasks
Z
Zhipu GLM
Intent extraction, cheap inference
A
Alibaba Qwen
Code generation
B
Baidu ERNIE
General tasks
B
ByteDance Doubao
Cheap inference
O
OllamaPrivacy mode
Fully local / offline โ€” any model

Task-Aware Routing

Task TypePrimary Model
complex_reasoningClaude Opus 4.8
code_generationDeepSeek Coder
code_reviewDeepSeek Coder
summarizationGemini 2.5 Flash
intent_extractionGemini 2.5 Flash
task_decompositionClaude Sonnet
reflectionGemini 2.5 Flash
cheap_inferenceOllama (local)
generalGemini 2.5 Flash

Override at runtime: POST /api/v1/settings/model

How We Compare

Built to lead, not follow.

A full-depth comparison against OpenClaw โ€” the leading open-source personal AI assistant. NeuralCleave leads or matches in every tracked dimension.

Feature
Competitors
NeuralCleave
Language
TypeScript / Node.js
Python โ€” better AI/ML ecosystem
Memory system
Flat files / LanceDB
Redis + Qdrant + SQLite (3-tier)
Per-agent memory isolation
โŒ
โœ… LRU namespace per node
LLM providers
4โ€“8
13 (all major + 5 Chinese)
Task-aware routing
โŒ
โœ… 10 task types โ†’ optimal model
Response quality scoring
โŒ
โœ… ReflectionEngine (4D, auto-retry)
Extended thinking
โŒ
โœ… Anthropic budget_tokens
Prometheus metrics
โŒ
โœ… 13 built-in metrics
Local / offline mode
Limited
โœ… Full Ollama + privacy toggle
Voice STT + TTS
Mobile-only
โœ… Whisper + 3-tier TTS (all platforms)
Wake word detection
macOS + iOS only
โœ… OpenWakeWord (all platforms)
Plugin SDK isolation
Markdown files
โœ… Typed ABCs + PEP 451 entry-points
Channel count
~29
32
Test coverage
~200
5,966 tests
Desktop app
โœ…
โœ… Tauri v2 (Win / macOS / Linux)
Mobile companion
โœ… Native apps
โœ… PWA (no app store needed)
Agent orchestration
โœ… Cross-machine
โœ… Multi-node + namespace isolation

Full analysis in docs/COMPETITIVE_ANALYSIS_OPENCLAW.md

Start in 30 seconds.

One command. No account required. Runs entirely on your machine.

Linux / macOS

curl -fsSL https://neuralcleave.com/install.sh | bash
Copy

Windows (PowerShell)

iwr -useb https://neuralcleave.com/install.ps1 | iex
Copy

pip

pip install neuralcleave
Copy

Docker

docker run -p 7432:7432 ghcr.io/theamitchandra/neuralcleave:latest
Copy

Desktop App

NeuralCleave v2.1.5 โ€” Native Tauri App

PyPI

Python 3.12+ required for CLI install. Desktop app ships as a self-contained bundle โ€” no Python needed.