Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago

Developer-Y/cs-video-courses

List of Computer Science courses with video lectures.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

Implements

paperVideoRAE: Taming Video Foundation Models for Generative Modeling via Representation AutoencoderspaperPeak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic AssessmentpaperRainDancer: RGB-Event Video Deraining with Rain-Oriented Spiking DynamicspaperDo Egocentric Video-Language Models Capture Both Hand- and Object-Centric Cues?paperThe Seriality Gap in Video Diffusion ModelspaperEgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera FootagepaperLIME: Learning Intent-aware Camera Motion from Egocentric VideopaperG2VD: Generalizable AI-Generated Video Detection via Counterfactual Intervention and Causal DisentanglementpaperNo Place to Hide: Benchmarking Video Hallucination with Background-Controlled PairspaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperRayPE: Ray-Space Positional Encoding for 3D-Aware Video GenerationpaperHAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent CollaborationpaperFADRA: Frequency-Aware Diffusion with Residual Adaptation for Video Face RestorationpaperMLVC: Multi-platform Learned Video Codec for Real-World DeploymentpaperMV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-ForcingpaperFrom Draft to Draft-Free: One-Step Video Object Removal via Privileged Distillation and Fast PlantingpaperStitch-Inferencer: Enhance Endoscopic Video Segmentation and Tracking via Panoramic ReconstructionpaperVideoChat3: Fully Open Video MLLM for Efficient and Generalist Video UnderstandingpaperLights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy ConsumptionpaperWorld from Motion: Generative Dynamic Gaussian Reconstruction from Monocular VideopaperEcoVideo: Entropy-Orchestrated Video Generation Paradigm in Cloud-Edge DynamicspaperFlowMark: Mask-Guided Video WatermarkingpaperCycle-World: Mitigating Error Accumulation in Long-term Video World Models via Reverse-Prediction Cycle ConsistencypaperAVSR-Diff: Scale-Agnostic Diffusion Priors for Temporally Consistent Arbitrary-Scale Video Super-ResolutionpaperMotionForesight: Re-purposing Video Models for Future 3D Scene-Flow PredictionpaperFVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video GenerationpaperVLA-ReID: Video-Level Association for Re-Identification in Multi-Object Tracking with Highly Similar ObjectspaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video GenerationpaperFlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality MimicrypaperO-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and ReasoningpaperOmniReasoner: Thinking with Long Audio-Video via Native Tool UsepaperContext-structured Video Anomaly Detection with Large Vision-Language ModelspaperBeyond the Single Camera: Agentic Multi-View Reasoning in Sports Video UnderstandingpaperQCA: Query- and Content-Aware Keyframe Selection for Long Video UnderstandingpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpaperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperVera: Identity-Faithful Human Subject-to-Video GenerationpaperHeadCast: Casting Attention Heads for Efficient Autoregressive Video GenerationpaperMoHallBench: A Benchmark for Motion Hallucination in Video Large Language ModelspaperHarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal ModelspaperAdaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervision in Video Face SwappingpaperElasticTTT: Prior-Preserving Test-Time Tuning for Video EditingpaperPrompt-Adapter Context Routing for Parameter-Efficient Multi-Shot Long Video ExtrapolationpaperGraphVid: Interactive Graph-Controllable Video GenerationpaperEvidence-Backed Video Question AnsweringpaperV-RAE: Rethinking Video Latent Spaces for GenerationpaperTraVEL: Trajectory-Guided Video Embedding Learning for Driving-Video RetrievalpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame InterpolationpaperContext-Matched Distillation: Teacher Causality for Autoregressive Video DistillationpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperGBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack DetectionpaperCaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated?paperBinarized High-Efficiency RAW Video Restoration and BeyondpaperPersonaShot: Benchmarking Person-Centric Narrative Continuity in Multi-Shot Video Generation

Related to (incoming)

contributed_to (incoming)

Related across the graph

paperTraVEL: Trajectory-Guided Video Embedding Learning for Driving-Video RetrievalpaperLIME: Learning Intent-aware Camera Motion from Egocentric VideopaperEgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera FootagepaperOmniReasoner: Thinking with Long Audio-Video via Native Tool UsepersonmundherpaperG2VD: Generalizable AI-Generated Video Detection via Counterfactual Intervention and Causal DisentanglementpaperNo Place to Hide: Benchmarking Video Hallucination with Background-Controlled PairspaperAdaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervision in Video Face SwappingpersonSuraj7879paperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspersonRohitSghpaperRayPE: Ray-Space Positional Encoding for 3D-Aware Video GenerationpaperFVAttn: Adaptive Sparse Attention with Runtime Load Balancing for Video GenerationpaperPeak-End-Net: A Peak-End Rule Inspired Framework for Generalizable Video Aesthetic AssessmentpaperThe Seriality Gap in Video Diffusion ModelspaperHAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent CollaborationpaperVera: Identity-Faithful Human Subject-to-Video Generationpersonpi3t4paperContext-Matched Distillation: Teacher Causality for Autoregressive Video DistillationpaperO-VAD: Industrial Video Anomaly Detection through Object-Centric Tracking and ReasoningpaperFlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality MimicrypaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperV-RAE: Rethinking Video Latent Spaces for GenerationpaperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperWorld from Motion: Generative Dynamic Gaussian Reconstruction from Monocular VideopaperBeyond the Single Camera: Agentic Multi-View Reasoning in Sports Video UnderstandingpersonDeveloper-YpaperLights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy ConsumptionpaperFADRA: Frequency-Aware Diffusion with Residual Adaptation for Video Face RestorationpaperHarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal ModelspaperContext-structured Video Anomaly Detection with Large Vision-Language ModelspersonsolomonbstonerpaperMV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-ForcingpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperGBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack DetectionpaperMoHallBench: A Benchmark for Motion Hallucination in Video Large Language Modelspersonsarendis56personAlaharon123paperEvidence-Backed Video Question AnsweringpaperPrompt-Adapter Context Routing for Parameter-Efficient Multi-Shot Long Video ExtrapolationpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperMLVC: Multi-platform Learned Video Codec for Real-World DeploymentpaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video GenerationpersonPeskyPotatopaperFrom Draft to Draft-Free: One-Step Video Object Removal via Privileged Distillation and Fast PlantingpaperVLA-ReID: Video-Level Association for Re-Identification in Multi-Object Tracking with Highly Similar ObjectspaperElasticTTT: Prior-Preserving Test-Time Tuning for Video EditingpaperMotionForesight: Re-purposing Video Models for Future 3D Scene-Flow PredictionpaperQCA: Query- and Content-Aware Keyframe Selection for Long Video Understandingpersonanishathalyepersonmono635personMasoudKavianipaperVideoRAE: Taming Video Foundation Models for Generative Modeling via Representation AutoencoderspersonspekulatiuspaperAVSR-Diff: Scale-Agnostic Diffusion Priors for Temporally Consistent Arbitrary-Scale Video Super-ResolutionpaperCycle-World: Mitigating Error Accumulation in Long-term Video World Models via Reverse-Prediction Cycle ConsistencypaperFlowMark: Mask-Guided Video WatermarkingpaperEcoVideo: Entropy-Orchestrated Video Generation Paradigm in Cloud-Edge DynamicspaperBinarized High-Efficiency RAW Video Restoration and BeyondpersonbivashpandeypaperPersonaShot: Benchmarking Person-Centric Narrative Continuity in Multi-Shot Video GenerationpaperVideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understandingpersontentenapersonmaxprogrammer007personinf3cti0n95personunfodepaperCaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated?paperDo Egocentric Video-Language Models Capture Both Hand- and Object-Centric Cues?paperRainDancer: RGB-Event Video Deraining with Rain-Oriented Spiking DynamicspersonDateBropersonm229abdpaperStitch-Inferencer: Enhance Endoscopic Video Segmentation and Tracking via Panoramic ReconstructionpersonbhutheshpersonppisapaperHeadCast: Casting Attention Heads for Efficient Autoregressive Video Generationpersonbutter1125modelstabilityai/stable-video-diffusion-img2vid-xtpaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpersonShramkowebpaperGraphVid: Interactive Graph-Controllable Video GenerationpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolationpersontoast1127

Topics