BlackRiver / Experimental systems

Intelligence,on your device.

A collection of experimental language, voice, multimodal and image systems that run locally through WebGPU. Your prompts, reference voices, images and audio stay on your machine.

05Live labs
100%On-device
0Server calls

Choose a system.

Each lab is a working product surface, not a mockup. A WebGPU-capable browser and sufficient device memory are required.

Multilingual TTS / Voice cloning / Automatic ASR
Voice intelligenceLaunch →

OmniVoice

Generate multilingual speech and clone a reference voice entirely in the browser. Parakeet transcribes the reference automatically before private, local WebGPU synthesis.

Core TTS: OmniVoice by k2-fsa · Automatic speech recognition: NVIDIA Parakeet · Browser ONNX packaging and community implementation: VocoLoco.BlackRiver AI implementation: unified loading flow, automatic transcription, long-form generation, interface and deployment.

646 languagesVoice cloningParakeet ASRNo usage quotaWebGPU
27B / 1-bit
Language modelLaunch →

Bonsai 27B

Large-scale private reasoning with a dense 27-billion-parameter model compressed to 1-bit weights.

Model and low-bit research: Prism ML · Base architecture: Qwen3.6-27B · Browser foundation: WebML Community.BlackRiver AI implementation: product interface, integration and deployment.

Text3.8 GBWebGPU
Vision / Audio / Text
Multimodal modelLaunch →

Gemma 4

Understand images, transcribe audio and converse with a compact multimodal model entirely in your browser.

Model: Google · Browser foundation: WebML Community and Hugging Face Transformers.js · WebGPU kernels: Fable 5.BlackRiver AI implementation: multimodal adaptation, interface and site integration.

E2B / E4BVisionAudio
High-speed inference
Compact language modelLaunch →

LFM2.5 230M

A compact Liquid AI language model tuned for extremely fast local generation and structured tasks.

Model: Liquid AI · Browser foundation: WebML Community · Original kernel optimization credits: Fable 5 and Opus 4.8.BlackRiver AI implementation: product interface, navigation and deployment.

230MTextQ4_0
Local diffusion
Image modelLaunch →

Bonsai Image

Generate high-quality images locally with compressed 4B diffusion models and private prompts.

Low-bit models: Prism ML · Base model: FLUX.2 Klein 4B by Black Forest Labs · Browser foundation: WebML Community.BlackRiver AI implementation: product interface, integration and deployment.

4BImageBinary / Ternary

Credits,
clearly.

“BlackRiver AI implementation” describes the unified product interface, BlackRiver navigation, website integration, deployment packaging and selected workflow adaptations. It does not imply that BlackRiver AI trained the underlying models or created the original WebGPU research demos.

OmniVoice

k2-fsa · NVIDIA · VocoLoco

k2-fsa created OmniVoice. Automatic reference transcription uses NVIDIA Parakeet. The browser-ready ONNX work and model packaging are credited to VocoLoco. BlackRiver AI integrated the complete local workflow, interface and deployment surface.

Bonsai 27B

Prism ML · Qwen · WebML Community

Prism ML created the Bonsai 27B low-bit model, derived from Qwen3.6-27B. The original browser work is credited to WebML Community.

Gemma 4 Multimodal

Google · Hugging Face · Fable 5

Google created Gemma 4. The browser foundation comes from WebML Community and Hugging Face tooling; the original optimized WebGPU kernels are credited to Fable 5.

LFM2.5 230M

Liquid AI · WebML Community

Liquid AI created LFM2.5 230M. The original browser demo is credited to WebML Community, with kernel optimization credits retained from that project.

Private by architecture.

Inference happens on your GPU. Content does not need to leave the browser.

Real model weights.

Each experience loads and caches its own quantized model, with no simulated outputs.

Built on WebGPU.

Modern browser compute turns local AI into a directly accessible product surface.