Topic

Image

50 items across the graph — tagged with Image.

From the graph · 50

repo
ultralytics/ultralytics

Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking

repo
roboflow/supervision

We write your reusable computer vision tools. 💜

repo
mudler/LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

repo
qdrant/qdrant

Qdrant - High-performance, massive-scale Vector Database and Vector Search Engine for the next generation of AI. Also available in the cloud https://cloud.qdran…

repo
invoke-ai/InvokeAI

Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using th…

repo
lucidrains/vit-pytorch

Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch

repo
Anionex/banana-slides

一站式原生AI PPT生成应用,几分钟内生成一套幻灯片; 支持上传任意模板图片,上传任意素材&智能解析,一句话/大纲/页面描述自动生成PPT,口头修改指定区域、一键导出可编辑ppt、视频等 - An AI-native slides generator based on nano banana pro🍌

repo
Unstructured-IO/unstructured

Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming complex documents into clean, structured formats fo…

model
black-forest-labs/FLUX.1-dev

Hugging Face model with 14399 likes. Tags: diffusers, safetensors, text-to-image, image-generation, flux, en, license:other, endpoints_compatible, diffusers:Flu…

model
Qwen/Qwen3.8-27B

Hugging Face model with 13689 likes. Tags: transformers, safetensors, qwen3_5, image-text-to-text, conversational, license:apache-2.0, eval-results, endpoints_c…

repo
kornia/kornia

🐍 Geometric Computer Vision Library for Spatial AI

model
moonshotai/Kimi-K3

Hugging Face model with 11149 likes. Tags: transformers, safetensors, kimi_k3, feature-extraction, compressed-tensors, conversational, image-text-to-text, custo…

repo
CVHub520/X-AnyLabeling

X-AnyLabeling: A lightweight, efficient, and unified cross-platform desktop application for annotating text, image, video, and multimodal data, combining versat…

repo
satellite-image-deep-learning/techniques

Techniques for deep learning with satellite & aerial imagery

repo
yzhao062/pyod

A Python library for anomaly detection across tabular, time series, graph, text, image, and audio data. 60+ detectors, benchmark-backed ADEngine orchestration,…

repo
roboflow/notebooks

A collection of tutorials on state-of-the-art computer vision models and techniques. Explore everything from foundational architectures like ResNet to cutting-e…

model
stabilityai/stable-diffusion-xl-base-1.0

Hugging Face model with 8098 likes. Tags: diffusers, onnx, safetensors, text-to-image, stable-diffusion, arxiv:2307.01952, arxiv:2211.01324, arxiv:2108.01073, a…

model
CompVis/stable-diffusion-v1-4

Hugging Face model with 7057 likes. Tags: diffusers, safetensors, stable-diffusion, stable-diffusion-diffusers, text-to-image, arxiv:2207.12598, arxiv:2112.1075…

repo
shimat/opencvsharp

OpenCV wrapper for .NET

repo
NVIDIA/DALI

A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and infer…

model
black-forest-labs/FLUX.1-schnell

Hugging Face model with 5680 likes. Tags: diffusers, safetensors, text-to-image, image-generation, flux, en, license:apache-2.0, endpoints_compatible, diffusers…

repo
obss/sahi

Framework agnostic sliced/tiled inference + interactive ui + error analysis plots

model
Tongyi-MAI/Z-Image-Turbo

Hugging Face model with 5191 likes. Tags: diffusers, safetensors, text-to-image, en, arxiv:2511.22699, arxiv:2511.22677, arxiv:2511.13649, license:apache-2.0, d…

model
stabilityai/stable-diffusion-3-medium

Hugging Face model with 5019 likes. Tags: diffusion-single-file, text-to-image, stable-diffusion, en, arxiv:2403.03206, license:other, region:us

model
MiniMaxAI/MiniMax-H3

Hugging Face model with 4810 likes. Tags: minimax-h3, diffusers, safetensors, text-to-video, image-to-video, image-text-to-video, video-to-video, text-to-audio-…

model
Qwen/Qwen3.8-Flash-Next

Hugging Face model with 4738 likes. Tags: transformers, safetensors, qwen4_exp, image-text-to-text, conversational, license:other, eval-results, endpoints_compa…

repo
mcmonkeyprojects/SwarmUI

SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and…

repo
crmne/ruby_llm

One delightful Ruby framework for every major AI provider. Build AI agents, chatbots, RAG apps, and multimodal workflows in beautiful, expressive code.

model
baidu/Unlimited-OCR

Hugging Face model with 4172 likes. Tags: transformers, safetensors, unlimited-ocr, feature-extraction, baidu, vision-language, ocr, custom_code, image-text-to-…

model
WarriorMama777/OrangeMixs

Hugging Face model with 3931 likes. Tags: diffusers, stable-diffusion, text-to-image, dataset:Nerfgun3/bad_prompt, license:creativeml-openrail-m, endpoints_comp…

model
stabilityai/stable-diffusion-3.5-large

Hugging Face model with 3746 likes. Tags: diffusers, safetensors, text-to-image, stable-diffusion, en, arxiv:2403.03206, license:other, diffusers:StableDiffusio…

model
google/gemma-4-31B-it

Hugging Face model with 3697 likes. Tags: transformers, safetensors, gemma4, image-text-to-text, conversational, arxiv:2607.02770, base_model:google/gemma-4-31B…

model
deepseek-ai/Janus-Pro-7B

Hugging Face model with 3654 likes. Tags: transformers, pytorch, multi_modality, muiltimodal, text-to-image, unified-model, any-to-any, arxiv:2501.17811, licens…

model
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

Hugging Face model with 3571 likes. Tags: gguf, uncensored, qwen3.6, moe, vision, multimodal, image-text-to-text, en, zh, multilingual

repo
HanaokaYuzu/Gemini-API

✨ Reverse-engineered Python API for Google Gemini web app

model
stabilityai/stable-video-diffusion-img2vid-xt

Hugging Face model with 3389 likes. Tags: diffusers, safetensors, image-to-video, license:other, diffusers:StableVideoDiffusionPipeline, region:us

model
deepseek-ai/DeepSeek-OCR

Hugging Face model with 3349 likes. Tags: transformers, safetensors, deepseek_vl_v2, feature-extraction, deepseek, vision-language, ocr, custom_code, image-text…

model
prompthero/openjourney

Hugging Face model with 3238 likes. Tags: diffusers, safetensors, stable-diffusion, text-to-image, en, license:creativeml-openrail-m, endpoints_compatible, diff…

model
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled

Hugging Face model with 2918 likes. Tags: safetensors, qwen3_5, unsloth, qwen, qwen3.5, reasoning, chain-of-thought, Dense, image-text-to-text, conversational

model
CompVis/stable-diffusion-v-1-4-original

Hugging Face model with 2877 likes. Tags: stable-diffusion, text-to-image, arxiv:2207.12598, arxiv:2112.10752, arxiv:2103.00020, arxiv:2205.11487, arxiv:1910.09…

model
moonshotai/Kimi-K2.5

Hugging Face model with 2855 likes. Tags: transformers, safetensors, kimi_k25, feature-extraction, compressed-tensors, image-text-to-text, conversational, custo…

repo
voxelmorph/voxelmorph

Unsupervised Learning for Image Registration

model
black-forest-labs/FLUX.1-Kontext-dev

Hugging Face model with 2721 likes. Tags: diffusers, safetensors, image-generation, flux, diffusion-single-file, image-to-image, en, arxiv:2506.15742, license:o…

repo
jolibrain/deepdetect

Deep Learning Server and CLI for Torch and TensorRT

repo
milvus-io/bootcamp

Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.

repo
TorchIO-project/torchio

Medical imaging processing for AI applications.

repo
emgucv/emgucv

Emgu CV is a cross platform .Net wrapper to the OpenCV image processing library.

repo
ermig1979/Simd

C++ image processing and machine learning library with using of SIMD: SSE, AVX, AVX-512, AMX for x86/x64, NEON, SVE for ARM, HVX for Hexagon

repo
visual-layer/fastdup

fastdup is a powerful, free tool designed to rapidly generate valuable insights from image and video datasets. It helps enhance the quality of both images and l…

repo
qupath/qupath

QuPath - Open-source bioimage analysis for research

Related topics