Qwen VL
Ask questions about any image
All Spaces of the Week...from all weeks
Ask questions about any image
Chat with an AI model
Dub videos in another language with cloned voice
Chat with an AI assistant using text and images
Create a custom story with characters and plot
Reimagine humans into new scenes! Faces and poses preserved!
Highlight code errors using probability scores
Analyze and visualize dataset characteristics and statistics
Detect small objects in images using SAHI + YOLOX
Process and analyze Arabic text
Duplicate a Hugging Face repository
Analyze text for multiple emotions
Search for images based on text queries
Analyze Hungarian text for grammar, entities, and more
Detect objects in images with YOLO and SAHI
Generate paraphrased questions and translate text for chatbots
Extract and highlight entities from a specific email in a thread
Deblur images to enhance clarity
Create 3D images using your webcam
Generate Dutch text based on a given prompt
Ask questions about text in a PDF
Generate Pokémon images from a numeric seed
Classify video content into categories
Swap faces between two photos with optional anonymization
Segment teeth in panoramic X-rays
Visualize pronunciation differences between two word recordings
Classify document images into predefined categories
Generate line drawing sketches from your photos
Extract text from images in multiple languages
Upload RGB image, get depth map
Play a truck-themed game
Play classic Doom game
Detect and visualize facial landmarks on photos
Generate TensorFlow ops from example input and output
Generate images using latent interpolation
Draw facial landmarks on images
Generate art from text or image prompts
Find relevant passages in documents using semantic search
Generate Magic Eye autostereograms from any photo
Perform NLP tasks like summarization, Q&A, and text generation
Ask questions about Game of Thrones
Generate and control facial images
Generate a poem and illustration from a word
Transform portraits into various artistic styles
Search for medical images using natural language queries
Generate human images from text descriptions
Transform portraits into various artistic styles
Generate multilingual talking-face videos from your text
Generate speech from text in multiple languages
Extract and crop image sections based on text description
Play GPT Golf to guess the target word
Identify flower types and visualize predictions
Play a Unity game where you push blocks into pyramids
Identify dog breeds from images
Remove background from images
Analyze text for sentiment, keywords, POS, emotions, and entities
Generate text-masked images using PIXEL model
Extract information from Indonesian receipts
Remove moire patterns from screen images
Generate text or fill in blanks with a bilingual AI model
Generate a syntactic tree from a sentence
Run and customize deep reinforcement learning simulations
Generate captions for images
Classify images to determine if they are huggable
Fact checking baseline. Dense retrieval + textual entailment
Evaluate gendered pronoun resolution in text
Identify types of food in images
Stitch two images together seamlessly
Detect edges in your photos with various filters
Apply various blur effects to images
Remove background from anime images
Restore and enhance faces in photos with optional upscaling
Generate creative Stable Diffusion prompts
Transfer portrait styles to images and videos
Identify emotion from multi-lingual audio
Play with an AI dog that catches sticks you throw
Generate summaries for long-form text
Generate protein structures using diffusion models
text2text models for document summarization
Convert PDFs and images to structured text and layout data
Score nail psoriasis severity from a hand photo
Generate 3D Mars surface models from images
Generate text answers to various prompts
Enhance image resolution up to 8x
Find images by describing what you're looking for
Clone a voice and generate speech
Generate optical flow or stereo disparity from image pairs
Check if your GitHub repo is in The Stack dataset
Compare text processing speeds with BetterTransformer
Segment images into objects, instances, or scenes
Segment objects in images using text prompts
Generate detailed Stable Diffusion prompts from any image
Generate Simpsons-themed images from text prompts
Train custom Stable Diffusion models
Generate singing voice from lyrics and musical score
Get Music from Generated Spectrogram with Diffusion
Generate optimized prompts for Stable Diffusion
Generate code snippets in Python, Java, JavaScript
Generate images from text prompts
Generate anime character speech in English, Chinese, and Japanese
Generate captions using several AI models
Correct grammar in your text with highlighted edits
Edit images using text instructions
Transform images into pixel art
Train and test custom LoRA DreamBooth models
Generate audio from text descriptions
Generate captions and answer questions from an image
Generate descriptive captions for images using advanced AI technology
Create images from pose-guided prompts
Generate images from text prompts with attention guidance
Run GitHub scripts on Hugging Face Spaces
Generate 3D room layout from an RGB panorama image
Safeguard images against ML-based photo manipulation
Generate a Mario level from text prompts
Generate text with a watermark
Generate images from sketches, edges, poses, and depth maps
Explore and discover diffusion models from the Hugging Face hub
Predict depth map from a single image
Generate text with a powerful RWKV7 language model
Chat with AI models and manage conversation history
Generate captions for images using prompts
Visualize transformer computations with a tuned lens
Generate a talking face video from an image and audio
Generate text based on instructions and input
Ask any questions to the IPCC and IPBES reports
image captioning, VQA
Generate realistic speech and sounds from typed text
Chat with images using MiniGPT-4
Generate and run Python code from natural language queries
text-to-3D & image-to-3D
Retrieve 3D human motion videos from text descriptions
Track, rank and evaluate open LLMs and chatbots
Generate protein sequences and 3‑D structures with diffusion
Evaluate prompt injection challenges and generate submission file
Interact with images using text prompts
Generate spoken audio from text using selectable voices
Chat with an AI assistant powered by Guanaco 33B
Generate music from a text description and optional melody
Interact with Falcon-Chat for personalized conversations
Transcribe audio files to text instantly
QR Code AI Art Generator Blend QR codes with AI Art
Chat with multiple AI models and compare their responses
Edit images with text prompts using diffusion inversion
Segment images using texts, points, or everything mode
Compare LLM hardware performance and find the best model
Generate tiny web apps from descriptions
View the LMArena leaderboard in full‑screen
Translate text between multiple languages
text-to-video
watermark-free Modelscope-based video generation
Create and edit images using DDPM and SEGA techniques
Display a loading screen with a spinner
Manipulate images by dragging points
Generate music from audio tracks
Create 3D mesh from a single image
unlimited Audio generation with a few added features
Compare image generation results from original and compressed AI models
Launch an AI agent to process tasks
Minecraft skin diffusion space for minecraft lovers
Generate images from prompts with feedback guidance
Chat with Llama‑2 13B AI for instant text responses
Chat with the Llama‑2 7B language model
Generate images from text prompts using Stable Diffusion XL
Display leaderboard of language models
Detect water areas in Sentinel‑2 satellite images
Detect burn scars in geotiff images
Generate audio and waveform video from text
Create a story from any uploaded image
Explore fun LoRAs and generate with SDXL
Submit model evaluation results to leaderboard
Compare AI model deployment costs
Experiment with and compare different tokenizers
Detect objects and poses in images with YOLOv8 in your browser
Generate code completions with Code Llama
Describe and highlight entities in images
Track points in a video
Edit videos using text prompts
Generate images from text prompts
Explore speech model benchmarks across languages and datasets
Use AI to translate text between languages
Generate speech from text using a reference voice
Generate stunning high quality illusion artwork
Segment Anything Model on the Browser with Candle/Rust/WASM
Generate unique images by combining random LoRA models
Chat with a private AI model locally
Convert PDFs to markup language using OCR
Replace product image backgrounds with custom scenes
State-of-the-art Zero-shot Object Detection
BLIP2 (cutting edge image captioning) in 🤗transformers
Check if your GPU can run a chosen LLM model
Generate and stream music from text prompts
Create your own AI comic with a single prompt
Explore Hugging Face author rankings and stats
Request evaluation for a new model
Explore code model leaderboard and submit evaluations
Transform and identify speech with MMS
Generate detailed text responses from custom prompts
Display a loading spinner while preparing space
Search for gameplay bugs in GTA V videos using text queries
Generate anime-style images from text descriptions
Generate speaker‑labeled transcripts from video or audio
Generate images from text prompts
Visualize anomaly detection results across different datasets
Generate new images from an uploaded image
Predict and visualize molecular docking poses
Transcribe audio or YouTube video into text
Resize images with and without antialiasing
Convert images of text into readable text
Enhance and restore old photos and AI-generated faces
Generate images from text with Stable Diffusion
Compare classifier performance on datasets
Profile a dataset and publish the report on Hugging Face
Generate code snippets using multiple models
Transform your portrait into Arcane‑style art
Visualize synthetic clustering with image segmentation