repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
Masoudjafaripour/nanochat-VLM
A minimal, hackable Vision-Language Model built on Karpathy’s nanochat — add image understanding and multimodal chat for under $200 in compute.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 56%ChatImage: Navigating Long-Form LLM Answers through Interactive Images →
- PossiblePossibly related (embedding) · 54%Large Tabular Models Excel Where LLMs Fail →
- PossiblePossibly related (embedding) · 52%Which image program can you talk to like ChatGPT but doesn't have all the stupid rules? →
- PossiblePossibly related (embedding) · 49%Open-Source Software Is Starting to Help Robots Think →
- PossiblePossibly related (embedding) · 49%The Hidden Infrastructure Challenge Behind Every AI-Generated Avatar - SD Times →
- PossiblePossibly related (embedding) · 51%Thinking Machines open sources first multimodal language model, Inkling, focused on low cost and 'resistance to censorship' - VentureBeat →
- PossiblePossibly related (embedding) · 55%Best chat model that fits in 128gb →
Implements
Covers
Covers (incoming)
Related across the graph
paperChatImage: Navigating Long-Form LLM Answers through Interactive ImagesnewsLarge Tabular Models Excel Where LLMs FailnewsThinking Machines open sources first multimodal language model, Inkling, focused on low cost and 'resistance to censorship' - VentureBeatnewsWhich image program can you talk to like ChatGPT but doesn't have all the stupid rules?newsOpen-Source Software Is Starting to Help Robots ThinknewsBest chat model that fits in 128gbnewsThe Hidden Infrastructure Challenge Behind Every AI-Generated Avatar - SD Times
