Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
fracanz 's Collections
AI Coding
Reasoning
Vision
Audio

Vision

updated Nov 14, 2025
Upvote
-

  • HuggingFaceM4/idefics2-8b-chatty

    Image-Text-to-Text • 8B • Updated Jul 30, 2024 • 410 • 95

  • Running on Zero
    Agents
    Featured
    104

    Phased Consistency Model PCM

    🐠
    104

    Generate images from text prompts


  • Running
    467

    Real-time Whisper WebGPU

    🎤
    467

    Transcribe audio to text instantly using WebGPU


  • stabilityai/stable-diffusion-3.5-medium

    Text-to-Image • 2B • Updated Oct 31, 2024 • 120k • • 1.14k

  • Runtime error
    Agents
    247

    OmniParser demo

    ⚡
    247

    Convert images of screens to structured elements


  • Runtime error
    Agents
    Featured
    249

    TransPixar

    😻
    249

    https://huggingface.co/papers/2501.03006


  • morphic/Wan2.2-frames-to-video

    Image-to-Video • Updated Oct 28, 2025 • • 42

  • peteromallet/Qwen-Image-Edit-InScene

    Image-to-Image • Updated Nov 1, 2025 • 147 • • 92

  • dx8152/Qwen-Image-Edit-2509-Fusion

    Image-to-Image • Updated Nov 12, 2025 • 2.64k • • 253

  • dx8152/Qwen-Edit-2509-Multiple-angles

    Image-to-Image • Updated Apr 21 • 182k • • 981

  • dx8152/Qwen-Image-Edit-2509-Light_restoration

    Image-to-Image • Updated Nov 25, 2025 • 3.65k • • 249
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs