Bonsai 27B WebGPU Kernels
Run a 1-bit 27B LLM locally in your browser on WebGPU
Run a 1-bit 27B LLM locally in your browser on WebGPU
Resize images for visual token budgets while keeping aspect ratio
Unified audio-text intelligence
Identity-preserving instruction image editing on Krea 2
Extend images into larger canvases with Krea 2 outpaint
Stream structured Markdown from document images and PDFs.
Realtime VLM for image and video understanding
Distilled LTX-2.3 identity video from a reference photo
generate a video from an image with a text prompt
Run a 1-bit 27B LLM locally in your browser on WebGPU
Reproduce every ICML 2026 paper with your agent
Demo of the Collection of Qwen Image Edit LoRAs
Voice chat over WebSocket against a HF speech-to-speech
Extract text from images and PDFs instantly
Identity-preserving instruction image editing on Krea 2
Efficient native-resolution image generation and editing
Image edit, text to image, image upscale, remove watermark
WAN2.2 based I2V
generate a video from an image with a text prompt
Extend images into larger canvases with Krea 2 outpaint
Use multiple FLUX.2-Klein LoRAs
Talk to Gemma 4 face to face, with a 3D lip-synced avatar
Open agentic retriever for hard multi-step search
Generate vivid images from text prompts in seconds
Run complete 3.96M and 9.36M text-to-waveform models live.
Animate an image into a video with custom prompts
High-fidelity 3D Generation from images
ltx 2.3 improved image-to-video with 10eros & native audio
Fast Wan 2.2 image-to-video with first/last frames
Stream structured Markdown from document images and PDFs.
Unified AR-LM-based speech enhancement & separation
Generate speech from text using voice design, cloning or presets