Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /sgl-project/sglang
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 2h ago

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 51%Message Passing Enables Efficient Reasoning →
  • PossiblePossibly related (embedding) · 50%MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments →
  • PossiblePossibly related (embedding) · 49%Little Brains, Big Feats: Exploring Compact Language Models →
  • PossiblePossibly related (embedding) · 49%Transformer →
  • PossiblePossibly related (embedding) · 49%Book Review: Domain-Specific Small Language Models by Guglielmo Iozzia →
  • FuzzyOverlapping authors or contributors · 62%Hindcast: Replaying Prediction Markets to Evaluate LLM Forecasters →

    “Shared author/contributor keys: zhou”

  • FuzzyOverlapping authors or contributors · 62%Self-Evolving Agent Harnesses via Gated Semantic Quality-Diversity →

    “Shared author/contributor keys: luo”

  • FuzzyOverlapping authors or contributors · 62%HoloCount: A Holistic Visual Counting Benchmark for MLLMs →

    “Shared author/contributor keys: wan”

Implements

paperMessage Passing Enables Efficient ReasoningpaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperLittle Brains, Big Feats: Exploring Compact Language ModelspaperHindcast: Replaying Prediction Markets to Evaluate LLM ForecasterspaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperHoloCount: A Holistic Visual Counting Benchmark for MLLMspaperScalable Visual Pretraining for Language IntelligencepaperPoint as Skeleton: Accumulated Point Cloud Enhanced Autoregressive Generation for Closed-Loop Autonomous Driving SimulationpaperText-Driven 3D Indoor Scene Synthesis in Non-Manhattan EnvironmentspaperPanoWorld: Real-World Panoramic GenerationpaperScore Accuracy Along the Forward Diffusion Does Not Certify Numerical Stability in Diffusion SamplingpaperRFMSR: Residual Flow Matching for Image Super-ResolutionpaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperMedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical ConsultationpaperDisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Visual Generation and EditingpaperCoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language ModelspaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperAlignment Is All You Need For X-to-4D GenerationpaperBridging Diffusion Pruning and Step Distillation with Teacher-Aligned RepairpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailpaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and GenerationpaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperCausalMix: Data Mixture as Causal Inference for Language Model TrainingpaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperANet Patu-1: The Value of Connection in the Agent NetworkpaperScaling Behavior Foundation Model for Humanoid RobotspaperSEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningpaperLongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU BudgetpaperQuReC: All-in-One Image Restoration with Query-Specific Guidance and Local-Global Response CalibrationpaperVideo = World + Event StreampaperCan We Trust Item Response Theory for AI Evaluation?paperSciDiagramEdit: Learning to Edit Scientific Diagrams from Paper RevisionspaperMeanFlowNFT: Bringing Forward-Process RL to Average-Velocity GeneratorspaperHierarchical Denoising For Multi-Step Visual ReasoningpaperOnline Neural Space Time Memory for Dynamic Novel View SynthesispaperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperCRISP: Constrained Refinement via Iterative Squeezing Process for Robust Medical Image Segmentation under Domain ShiftpaperUI2App: Benchmarking Visual Interaction Inference in Executable Web Application GenerationpaperTowards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment GuidancepaperNative Video-Action Pretraining for Generalizable Robot ControlpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperToolSciVer: Multimodal Scientific Claim Verification with Visual Tool Augmented Reinforcement LearningpaperAn MLIR-Based Compilation Method for Large Language ModelspaperBayesPO: Bayesian Prompt Optimization via Parallel-Tempered Gradient-Guided Discrete MCMCpaperLoop the Loopies!paperJoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA ModelspaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperVTLoc: Learning-based Tactile Contact Localization in Visual Point CloudspaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperActive rejection enables reliable generalization of universal machine-learning interatomic potentialspaperFourier Geometric Wind Power Forecasting with Numerical Weather PredictionpaperHow Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak ContributionspaperSynH-Rank: Quality-Aware Code Search via Diverse Data Synthesis and Hierarchical Ranking TrainingpaperPersistent Sparse Autoencoders: Learning Feature Timescales in Language ModelspaperLearning from Synthetic Data without Model Collapse in Iterative Instruction TuningpaperVecFontLLM: Anchor-Guided Direct Synthesis of Chinese Vector FontspaperEvoGUI: An Evolution-Aware Benchmark for GUI State-Transition UnderstandingpaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video GenerationpaperLLMs and Agentic AI Systems for Smart Grids: A Tutorial on Architectures and ApplicationspaperSGN: A Similarity-based Generative Network for Data Generation under Distribution ShiftpaperVEHBench: A Stage-Local Diagnostic Benchmark for LLM-Assisted Vibration Energy Harvester DesignpaperPatch Policy: Efficient Embodied Control via Dense Visual RepresentationspaperAn Early Warning of Emerging Biosecurity Risks in Frontier LLMspaperHOMIE: Human-object Centric Video Personalization via Multimodal Intelligent EnchancementpaperSciForma: Structure-Faithful Generation of Scientific DiagramspaperMBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation ModelspaperUR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress ProxiespaperMeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party MeetingspaperIGGT4D: Streaming 4D Instance-Grounded Geometry TransformerpaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperConservative Query and Adaptive Regularization for Offline RL Under Uncertainty EstimationpaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperQCA: Query- and Content-Aware Keyframe Selection for Long Video UnderstandingpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpaperSLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPODpaperPIER: Physics-Informed Environmental Retrieval for Time-Series ModelingpaperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperHow Does Urban Context Relate to Residential Building Health? A Vision-POI Fusion Framework for Building-Level Housing InspectionpaperVera: Identity-Faithful Human Subject-to-Video GenerationpaperPerceptDrive: Perception Prior World-Action Modeling with Adaptive Expert Routing for End-to-End Autonomous DrivingpaperExposure is Optional: Learning Unlike Coordination in Language ModelspaperA Definition and Roadmap for World ModelspaperGenAU: Language-Grounded Industrial Anomaly Understanding with Vision-Language ModelspaperMSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics LearningpaperClimate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobilitypaperCLUIE: Clustering-Aware Recurrent Propagation with Local Structural Compensation for Underwater Image EnhancementpaperSPDCN: Strip-based Deformable Convolutional Network for Steel Surface Defect SegmentationpaperVisual Contrastive Self-DistillationpaperStreaming Multi-Agent Autoregressive Diffusion Model with World State RegisterspaperSANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video GenerationpaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpaperDetecting LLM-Generated Tokens in Human--LLM Coauthored TextpaperAlphaOracle: Oracle bone script decipherment via human-workflow-inspired deep learningpaperScale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic ManipulationpaperEvidence-Backed Video Question AnsweringpaperReduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM InferencepaperBeyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning EvaluationpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperRefusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM SafetypaperWhen Should Multi-Round RAG Stop? Structured Stopping Judgments and Retrieval Reduction in Search-R1paperDARTree: Speculative Diffusion Decoding with Autoregressive Draft TreespaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon ObjectivespaperWhen Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit QuantisationpaperEdit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZpaperIntern-S2-Preview: Scientific Agentic Foundation ModelpaperAutoDesign: Meta-Harness Optimization for Long-Horizon Agentic DesignpaperSimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context ReasoningpaperMathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided RefinementpaperLeading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM ChallengepaperYou Only Pass Once: Answering and Abstaining Together in a Single Forward Pass of a Frozen Language ModelpaperSheetCompass: Hierarchical Relation Graphs for Agentic Spreadsheet ReasoningpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperDesigning Reinforcement Learning for Diffusion Models: A Unified Path-Space ViewpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation ModelspaperHERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geosciencepaperImproving the matrix multiplication exponent with modern optimization and AlphaEvolvepaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpaperHarnessEval-W: Agentifying the Evaluation of Visual Worlds

Related to

glossary_termTransformer

Covers

newsBook Review: Domain-Specific Small Language Models by Guglielmo Iozzia

Related to (incoming)

modeldeepseek-ai/DeepSeek-R1modeldeepseek-ai/DeepSeek-V3modelzai-org/GLM-5.2paperLook Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMsmodelmoonshotai/Kimi-K3modelQwen/Qwen3.8-27B

Implements (incoming)

paperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationpaperEvaluating and Understanding Model Editing for Medical Vision Language ModelspaperToward Real-Time Sentence-Level Sign Language TranslationpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

contributed_to (incoming)

personmerrymercypersonhnyls2002personfzyzcjypersonzhyncspersonslin1237personmickqianpersonBBufpersonFridge003personispobockpersonalisonshaopersonch-wanpersonKangyan-ZhoupersonCatherineSuepersonYing1123personJustinTong0323personShangmingCaipersonb8zhongpersonQiaolin-YupersonByronHsupersonyuan-luopersonyhyang201personyctseng0211personmmangkadpersonhzh0425personalphabetc1

Covers (incoming)

newsSetting Up Your Own Large Language Model - Towards Data SciencenewsSeeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]

Related across the graph

paperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationmodelQwen/Qwen3.8-27BpaperHoloCount: A Holistic Visual Counting Benchmark for MLLMspaperScalable Visual Pretraining for Language IntelligencepaperText-Driven 3D Indoor Scene Synthesis in Non-Manhattan EnvironmentspaperSheetCompass: Hierarchical Relation Graphs for Agentic Spreadsheet ReasoningpaperAlphaOracle: Oracle bone script decipherment via human-workflow-inspired deep learningpaperLittle Brains, Big Feats: Exploring Compact Language ModelspaperHow Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak ContributionspaperPanoWorld: Real-World Panoramic GenerationpersonShangmingCaipaperWhen Should Multi-Round RAG Stop? Structured Stopping Judgments and Retrieval Reduction in Search-R1modelmoonshotai/Kimi-K3paperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperMedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical ConsultationpaperDARTree: Speculative Diffusion Decoding with Autoregressive Draft TreespaperDisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Visual Generation and EditingpaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperScore Accuracy Along the Forward Diffusion Does Not Certify Numerical Stability in Diffusion SamplingpaperRFMSR: Residual Flow Matching for Image Super-ResolutionpaperIntern-S2-Preview: Scientific Agentic Foundation ModelpaperAutoDesign: Meta-Harness Optimization for Long-Horizon Agentic DesignpaperCoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language ModelspaperYou Only Pass Once: Answering and Abstaining Together in a Single Forward Pass of a Frozen Language Modelmodelzai-org/GLM-5.2paperSciForma: Structure-Faithful Generation of Scientific DiagramspaperHERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geosciencepaperPoint as Skeleton: Accumulated Point Cloud Enhanced Autoregressive Generation for Closed-Loop Autonomous Driving SimulationnewsSetting Up Your Own Large Language Model - Towards Data SciencepaperPerceptDrive: Perception Prior World-Action Modeling with Adaptive Expert Routing for End-to-End Autonomous DrivingpaperHindcast: Replaying Prediction Markets to Evaluate LLM ForecasterspaperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperFourier Geometric Wind Power Forecasting with Numerical Weather PredictionpaperExposure is Optional: Learning Unlike Coordination in Language ModelspaperHierarchical Denoising For Multi-Step Visual ReasoningpaperOnline Neural Space Time Memory for Dynamic Novel View Synthesisglossary_termTransformerpaperMSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics Learningmodeldeepseek-ai/DeepSeek-V3paperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperAlignment Is All You Need For X-to-4D GenerationpaperVera: Identity-Faithful Human Subject-to-Video GenerationpaperPersistent Sparse Autoencoders: Learning Feature Timescales in Language ModelspaperLook Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMspaperJoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA ModelspaperBridging Diffusion Pruning and Step Distillation with Teacher-Aligned RepairpaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrailpersonhzh0425paperSPDCN: Strip-based Deformable Convolutional Network for Steel Surface Defect SegmentationpaperLeading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM ChallengepaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperSANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video GenerationpaperELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and GenerationpaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperNative Video-Action Pretraining for Generalizable Robot ControlpersonYing1123paperSGN: A Similarity-based Generative Network for Data Generation under Distribution ShiftpaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperLoop the Loopies!paperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperLLMs and Agentic AI Systems for Smart Grids: A Tutorial on Architectures and ApplicationspaperCausalMix: Data Mixture as Causal Inference for Language Model TrainingpaperViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation ModelspaperRefusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM SafetypaperCan We Trust Item Response Theory for AI Evaluation?paperReduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM InferencepaperUR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress ProxiespaperHOMIE: Human-object Centric Video Personalization via Multimodal Intelligent EnchancementnewsSeeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]paperMBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation ModelspaperConservative Query and Adaptive Regularization for Offline RL Under Uncertainty EstimationpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperSimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context ReasoningpersonJustinTong0323paperEvaluating and Understanding Model Editing for Medical Vision Language ModelspaperUI2App: Benchmarking Visual Interaction Inference in Executable Web Application GenerationpaperMessage Passing Enables Efficient Reasoningpersonyctseng0211paperMeanFlowNFT: Bringing Forward-Process RL to Average-Velocity GeneratorspaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperA Definition and Roadmap for World ModelspaperToolSciVer: Multimodal Scientific Claim Verification with Visual Tool Augmented Reinforcement LearningpaperSEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningpaperToward Real-Time Sentence-Level Sign Language TranslationpaperAn Early Warning of Emerging Biosecurity Risks in Frontier LLMspaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperCLUIE: Clustering-Aware Recurrent Propagation with Local Structural Compensation for Underwater Image EnhancementpersonByronHsupaperEvidence-Backed Video Question AnsweringpaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperVideo = World + Event StreampaperTowards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment GuidancepaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generationpersonyhyang201paperCRISP: Constrained Refinement via Iterative Squeezing Process for Robust Medical Image Segmentation under Domain ShiftpaperVisual Contrastive Self-DistillationpaperScaling Behavior Foundation Model for Humanoid RobotspaperGenAU: Language-Grounded Industrial Anomaly Understanding with Vision-Language ModelspaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperHow Does Urban Context Relate to Residential Building Health? A Vision-POI Fusion Framework for Building-Level Housing InspectionpaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperQCA: Query- and Content-Aware Keyframe Selection for Long Video UnderstandingpaperDetecting LLM-Generated Tokens in Human--LLM Coauthored TextpaperClimate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobilitypaperIGGT4D: Streaming 4D Instance-Grounded Geometry TransformerpaperSLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPODpaperActive rejection enables reliable generalization of universal machine-learning interatomic potentialspersonmmangkadpaperAn MLIR-Based Compilation Method for Large Language ModelspaperBayesPO: Bayesian Prompt Optimization via Parallel-Tempered Gradient-Guided Discrete MCMCpaperVTLoc: Learning-based Tactile Contact Localization in Visual Point CloudspaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperQuReC: All-in-One Image Restoration with Query-Specific Guidance and Local-Global Response CalibrationpersonFridge003modeldeepseek-ai/DeepSeek-R1personmickqianpersonzhyncspaperEvoGUI: An Evolution-Aware Benchmark for GUI State-Transition UnderstandingpaperANet Patu-1: The Value of Connection in the Agent Networkpersonch-wanpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?newsBook Review: Domain-Specific Small Language Models by Guglielmo IozziapaperLearning from Synthetic Data without Model Collapse in Iterative Instruction TuningpaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpaperPIER: Physics-Informed Environmental Retrieval for Time-Series ModelingpaperBeyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning EvaluationpaperImproving the matrix multiplication exponent with modern optimization and AlphaEvolvepaperSynH-Rank: Quality-Aware Code Search via Diverse Data Synthesis and Hierarchical Ranking TrainingpaperMathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided RefinementpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperEdit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZpaperScale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic ManipulationpaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon ObjectivespaperMeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party MeetingspersonKangyan-ZhoupersonCatherineSuepersonb8zhongpersonalphabetc1paperDesigning Reinforcement Learning for Diffusion Models: A Unified Path-Space ViewpersonalisonshaopaperLongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU BudgetpersonfzyzcjypaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpaperVecFontLLM: Anchor-Guided Direct Synthesis of Chinese Vector FontspaperHarnessEval-W: Agentifying the Evaluation of Visual WorldspersonQiaolin-Yupersonslin1237paperVEHBench: A Stage-Local Diagnostic Benchmark for LLM-Assisted Vibration Energy Harvester DesignpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperPatch Policy: Efficient Embodied Control via Dense Visual Representationspersonhnyls2002personispobockpersonyuan-luopersonmerrymercypaperSciDiagramEdit: Learning to Edit Scientific Diagrams from Paper RevisionspaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpersonBBufpaperStreaming Multi-Agent Autoregressive Diffusion Model with World State RegisterspaperWhen Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation
Knowledge path·PHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read Generation→MQwen/Qwen3.8-27B→PHoloCount: A Holistic Visual Counting Benchmark for MLLMs→Rsgl-project/sglang

Topics

attentionblackwellcudadeepseekdiffusionglmgpt-ossinferencellamallm

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Maintenance84
RIS90GitHub verified
Graph trust82Primary
Graph score31978