Image
50 items across the graph — tagged with Image.
From the graph · 50
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
We write your reusable computer vision tools. 💜
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdran…
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using th…
Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
一站式原生AI PPT生成应用,几分钟内生成一套幻灯片; 支持上传任意模板图片,上传任意素材&智能解析,一句话/大纲/页面描述自动生成PPT,口头修改指定区域、一键导出可编辑ppt、视频等 - An AI-native slides generator based on nano banana pro🍌
Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming complex documents into clean, structured formats fo…
Hugging Face model with 14399 likes. Tags: diffusers, safetensors, text-to-image, image-generation, flux, en, license:other, endpoints_compatible, diffusers:Flu…
Hugging Face model with 13689 likes. Tags: transformers, safetensors, qwen3_5, image-text-to-text, conversational, license:apache-2.0, eval-results, endpoints_c…
🐍 Geometric Computer Vision Library for Spatial AI
Hugging Face model with 11149 likes. Tags: transformers, safetensors, kimi_k3, feature-extraction, compressed-tensors, conversational, image-text-to-text, custo…
X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data, combining versat…
Techniques for deep learning with satellite & aerial imagery
A Python library for anomaly detection across tabular, time series, graph, text, image, and audio data. 60+ detectors, benchmark-backed ADEngine orchestration,…
A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-e…
Hugging Face model with 8098 likes. Tags: diffusers, onnx, safetensors, text-to-image, stable-diffusion, arxiv:2307.01952, arxiv:2211.01324, arxiv:2108.01073, a…
Hugging Face model with 7057 likes. Tags: diffusers, safetensors, stable-diffusion, stable-diffusion-diffusers, text-to-image, arxiv:2207.12598, arxiv:2112.1075…
OpenCV wrapper for .NET
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and infer…
Hugging Face model with 5680 likes. Tags: diffusers, safetensors, text-to-image, image-generation, flux, en, license:apache-2.0, endpoints_compatible, diffusers…
Framework agnostic sliced/tiled inference + interactive ui + error analysis plots
Hugging Face model with 5191 likes. Tags: diffusers, safetensors, text-to-image, en, arxiv:2511.22699, arxiv:2511.22677, arxiv:2511.13649, license:apache-2.0, d…
Hugging Face model with 5019 likes. Tags: diffusion-single-file, text-to-image, stable-diffusion, en, arxiv:2403.03206, license:other, region:us
Hugging Face model with 4810 likes. Tags: minimax-h3, diffusers, safetensors, text-to-video, image-to-video, image-text-to-video, video-to-video, text-to-audio-…
Hugging Face model with 4738 likes. Tags: transformers, safetensors, qwen4_exp, image-text-to-text, conversational, license:other, eval-results, endpoints_compa…
SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and…
One delightful Ruby framework for every major AI provider. Build AI agents, chatbots, RAG apps, and multimodal workflows in beautiful, expressive code.
Hugging Face model with 4172 likes. Tags: transformers, safetensors, unlimited-ocr, feature-extraction, baidu, vision-language, ocr, custom_code, image-text-to-…
Hugging Face model with 3931 likes. Tags: diffusers, stable-diffusion, text-to-image, dataset:Nerfgun3/bad_prompt, license:creativeml-openrail-m, endpoints_comp…
Hugging Face model with 3746 likes. Tags: diffusers, safetensors, text-to-image, stable-diffusion, en, arxiv:2403.03206, license:other, diffusers:StableDiffusio…
Hugging Face model with 3697 likes. Tags: transformers, safetensors, gemma4, image-text-to-text, conversational, arxiv:2607.02770, base_model:google/gemma-4-31B…
Hugging Face model with 3654 likes. Tags: transformers, pytorch, multi_modality, muiltimodal, text-to-image, unified-model, any-to-any, arxiv:2501.17811, licens…
Hugging Face model with 3571 likes. Tags: gguf, uncensored, qwen3.6, moe, vision, multimodal, image-text-to-text, en, zh, multilingual
✨ Reverse-engineered Python API for Google Gemini web app
Hugging Face model with 3389 likes. Tags: diffusers, safetensors, image-to-video, license:other, diffusers:StableVideoDiffusionPipeline, region:us
Hugging Face model with 3349 likes. Tags: transformers, safetensors, deepseek_vl_v2, feature-extraction, deepseek, vision-language, ocr, custom_code, image-text…
Hugging Face model with 3238 likes. Tags: diffusers, safetensors, stable-diffusion, text-to-image, en, license:creativeml-openrail-m, endpoints_compatible, diff…
Hugging Face model with 2918 likes. Tags: safetensors, qwen3_5, unsloth, qwen, qwen3.5, reasoning, chain-of-thought, Dense, image-text-to-text, conversational
Hugging Face model with 2877 likes. Tags: stable-diffusion, text-to-image, arxiv:2207.12598, arxiv:2112.10752, arxiv:2103.00020, arxiv:2205.11487, arxiv:1910.09…
Hugging Face model with 2855 likes. Tags: transformers, safetensors, kimi_k25, feature-extraction, compressed-tensors, image-text-to-text, conversational, custo…
Unsupervised Learning for Image Registration
Hugging Face model with 2721 likes. Tags: diffusers, safetensors, image-generation, flux, diffusion-single-file, image-to-image, en, arxiv:2506.15742, license:o…
Deep Learning Server and CLI for Torch and TensorRT
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
Medical imaging processing for AI applications.
Emgu CV is a cross platform .Net wrapper to the OpenCV image processing library.
C++ image processing and machine learning library with using of SIMD: SSE, AVX, AVX-512, AMX for x86/x64, NEON, SVE for ARM, HVX for Hexagon
fastdup is a powerful, free tool designed to rapidly generate valuable insights from image and video datasets. It helps enhance the quality of both images and l…
QuPath - Open-source bioimage analysis for research
