repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 2h ago
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 51%Message Passing Enables Efficient Reasoning →
- PossiblePossibly related (embedding) · 50%MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments →
- PossiblePossibly related (embedding) · 49%Little Brains, Big Feats: Exploring Compact Language Models →
- PossiblePossibly related (embedding) · 49%Transformer →
- PossiblePossibly related (embedding) · 49%Book Review: Domain-Specific Small Language Models by Guglielmo Iozzia →
- FuzzyOverlapping authors or contributors · 62%Hindcast: Replaying Prediction Markets to Evaluate LLM Forecasters →
“Shared author/contributor keys: zhou”
- FuzzyOverlapping authors or contributors · 62%Self-Evolving Agent Harnesses via Gated Semantic Quality-Diversity →
“Shared author/contributor keys: luo”
- FuzzyOverlapping authors or contributors · 62%HoloCount: A Holistic Visual Counting Benchmark for MLLMs →
“Shared author/contributor keys: wan”
Implements
paperMessage Passing Enables Efficient ReasoningpaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperLittle Brains, Big Feats: Exploring Compact Language ModelspaperHindcast: Replaying Prediction Markets to Evaluate LLM ForecasterspaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperHoloCount: A Holistic Visual Counting Benchmark for MLLMspaperScalable Visual Pretraining for Language IntelligencepaperPoint as Skeleton: Accumulated Point Cloud Enhanced Autoregressive Generation for Closed-Loop Autonomous Driving SimulationpaperText-Driven 3D Indoor Scene Synthesis in Non-Manhattan EnvironmentspaperPanoWorld: Real-World Panoramic GenerationpaperScore Accuracy Along the Forward Diffusion Does Not Certify Numerical Stability in Diffusion SamplingpaperRFMSR: Residual Flow Matching for Image Super-ResolutionpaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperMedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical ConsultationpaperDisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Visual Generation and EditingpaperCoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language ModelspaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperAlignment Is All You Need For X-to-4D GenerationpaperBridging Diffusion Pruning and Step Distillation with Teacher-Aligned RepairpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety GuardrailpaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and GenerationpaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperCausalMix: Data Mixture as Causal Inference for Language Model TrainingpaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperANet Patu-1: The Value of Connection in the Agent NetworkpaperScaling Behavior Foundation Model for Humanoid RobotspaperSEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningpaperLongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU BudgetpaperQuReC: All-in-One Image Restoration with Query-Specific Guidance and Local-Global Response CalibrationpaperVideo = World + Event StreampaperCan We Trust Item Response Theory for AI Evaluation?paperSciDiagramEdit: Learning to Edit Scientific Diagrams from Paper RevisionspaperMeanFlowNFT: Bringing Forward-Process RL to Average-Velocity GeneratorspaperHierarchical Denoising For Multi-Step Visual ReasoningpaperOnline Neural Space Time Memory for Dynamic Novel View SynthesispaperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperCRISP: Constrained Refinement via Iterative Squeezing Process for Robust Medical Image Segmentation under Domain ShiftpaperUI2App: Benchmarking Visual Interaction Inference in Executable Web Application GenerationpaperTowards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment GuidancepaperNative Video-Action Pretraining for Generalizable Robot ControlpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperToolSciVer: Multimodal Scientific Claim Verification with Visual Tool Augmented Reinforcement LearningpaperAn MLIR-Based Compilation Method for Large Language ModelspaperBayesPO: Bayesian Prompt Optimization via Parallel-Tempered Gradient-Guided Discrete MCMCpaperLoop the Loopies!paperJoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA ModelspaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperVTLoc: Learning-based Tactile Contact Localization in Visual Point CloudspaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperActive rejection enables reliable generalization of universal machine-learning interatomic potentialspaperFourier Geometric Wind Power Forecasting with Numerical Weather PredictionpaperHow Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak ContributionspaperSynH-Rank: Quality-Aware Code Search via Diverse Data Synthesis and Hierarchical Ranking TrainingpaperPersistent Sparse Autoencoders: Learning Feature Timescales in Language ModelspaperLearning from Synthetic Data without Model Collapse in Iterative Instruction TuningpaperVecFontLLM: Anchor-Guided Direct Synthesis of Chinese Vector FontspaperEvoGUI: An Evolution-Aware Benchmark for GUI State-Transition UnderstandingpaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video GenerationpaperLLMs and Agentic AI Systems for Smart Grids: A Tutorial on Architectures and ApplicationspaperSGN: A Similarity-based Generative Network for Data Generation under Distribution ShiftpaperVEHBench: A Stage-Local Diagnostic Benchmark for LLM-Assisted Vibration Energy Harvester DesignpaperPatch Policy: Efficient Embodied Control via Dense Visual RepresentationspaperAn Early Warning of Emerging Biosecurity Risks in Frontier LLMspaperHOMIE: Human-object Centric Video Personalization via Multimodal Intelligent EnchancementpaperSciForma: Structure-Faithful Generation of Scientific DiagramspaperMBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation ModelspaperUR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress ProxiespaperMeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party MeetingspaperIGGT4D: Streaming 4D Instance-Grounded Geometry TransformerpaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperConservative Query and Adaptive Regularization for Offline RL Under Uncertainty EstimationpaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperQCA: Query- and Content-Aware Keyframe Selection for Long Video UnderstandingpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpaperSLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPODpaperPIER: Physics-Informed Environmental Retrieval for Time-Series ModelingpaperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperHow Does Urban Context Relate to Residential Building Health? A Vision-POI Fusion Framework for Building-Level Housing InspectionpaperVera: Identity-Faithful Human Subject-to-Video GenerationpaperPerceptDrive: Perception Prior World-Action Modeling with Adaptive Expert Routing for End-to-End Autonomous DrivingpaperExposure is Optional: Learning Unlike Coordination in Language ModelspaperA Definition and Roadmap for World ModelspaperGenAU: Language-Grounded Industrial Anomaly Understanding with Vision-Language ModelspaperMSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics LearningpaperClimate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobilitypaperCLUIE: Clustering-Aware Recurrent Propagation with Local Structural Compensation for Underwater Image EnhancementpaperSPDCN: Strip-based Deformable Convolutional Network for Steel Surface Defect SegmentationpaperVisual Contrastive Self-DistillationpaperStreaming Multi-Agent Autoregressive Diffusion Model with World State RegisterspaperSANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video GenerationpaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpaperDetecting LLM-Generated Tokens in Human--LLM Coauthored TextpaperAlphaOracle: Oracle bone script decipherment via human-workflow-inspired deep learningpaperScale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic ManipulationpaperEvidence-Backed Video Question AnsweringpaperReduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM InferencepaperBeyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning EvaluationpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperRefusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM SafetypaperWhen Should Multi-Round RAG Stop? Structured Stopping Judgments and Retrieval Reduction in Search-R1paperDARTree: Speculative Diffusion Decoding with Autoregressive Draft TreespaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon ObjectivespaperWhen Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit QuantisationpaperEdit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZpaperIntern-S2-Preview: Scientific Agentic Foundation ModelpaperAutoDesign: Meta-Harness Optimization for Long-Horizon Agentic DesignpaperSimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context ReasoningpaperMathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided RefinementpaperLeading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM ChallengepaperYou Only Pass Once: Answering and Abstaining Together in a Single Forward Pass of a Frozen Language ModelpaperSheetCompass: Hierarchical Relation Graphs for Agentic Spreadsheet ReasoningpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperDesigning Reinforcement Learning for Diffusion Models: A Unified Path-Space ViewpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation ModelspaperHERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geosciencepaperImproving the matrix multiplication exponent with modern optimization and AlphaEvolvepaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpaperHarnessEval-W: Agentifying the Evaluation of Visual Worlds
Related to
Covers
Related to (incoming)
Implements (incoming)
paperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationpaperEvaluating and Understanding Model Editing for Medical Vision Language ModelspaperToward Real-Time Sentence-Level Sign Language TranslationpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?
contributed_to (incoming)
personmerrymercypersonhnyls2002personfzyzcjypersonzhyncspersonslin1237personmickqianpersonBBufpersonFridge003personispobockpersonalisonshaopersonch-wanpersonKangyan-ZhoupersonCatherineSuepersonYing1123personJustinTong0323personShangmingCaipersonb8zhongpersonQiaolin-YupersonByronHsupersonyuan-luopersonyhyang201personyctseng0211personmmangkadpersonhzh0425personalphabetc1
Covers (incoming)
Related across the graph
paperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationmodelQwen/Qwen3.8-27BpaperHoloCount: A Holistic Visual Counting Benchmark for MLLMspaperScalable Visual Pretraining for Language IntelligencepaperText-Driven 3D Indoor Scene Synthesis in Non-Manhattan EnvironmentspaperSheetCompass: Hierarchical Relation Graphs for Agentic Spreadsheet ReasoningpaperAlphaOracle: Oracle bone script decipherment via human-workflow-inspired deep learningpaperLittle Brains, Big Feats: Exploring Compact Language ModelspaperHow Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak ContributionspaperPanoWorld: Real-World Panoramic GenerationpersonShangmingCaipaperWhen Should Multi-Round RAG Stop? Structured Stopping Judgments and Retrieval Reduction in Search-R1modelmoonshotai/Kimi-K3paperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperMedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical ConsultationpaperDARTree: Speculative Diffusion Decoding with Autoregressive Draft TreespaperDisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Visual Generation and EditingpaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperScore Accuracy Along the Forward Diffusion Does Not Certify Numerical Stability in Diffusion SamplingpaperRFMSR: Residual Flow Matching for Image Super-ResolutionpaperIntern-S2-Preview: Scientific Agentic Foundation ModelpaperAutoDesign: Meta-Harness Optimization for Long-Horizon Agentic DesignpaperCoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language ModelspaperYou Only Pass Once: Answering and Abstaining Together in a Single Forward Pass of a Frozen Language Modelmodelzai-org/GLM-5.2paperSciForma: Structure-Faithful Generation of Scientific DiagramspaperHERMES: a multi-agent framework for structured knowledge extraction from ultra-long documents in geosciencepaperPoint as Skeleton: Accumulated Point Cloud Enhanced Autoregressive Generation for Closed-Loop Autonomous Driving SimulationnewsSetting Up Your Own Large Language Model - Towards Data SciencepaperPerceptDrive: Perception Prior World-Action Modeling with Adaptive Expert Routing for End-to-End Autonomous DrivingpaperHindcast: Replaying Prediction Markets to Evaluate LLM ForecasterspaperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperFourier Geometric Wind Power Forecasting with Numerical Weather PredictionpaperExposure is Optional: Learning Unlike Coordination in Language ModelspaperHierarchical Denoising For Multi-Step Visual ReasoningpaperOnline Neural Space Time Memory for Dynamic Novel View Synthesisglossary_termTransformerpaperMSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics Learningmodeldeepseek-ai/DeepSeek-V3paperHoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven ReasoningpaperAlignment Is All You Need For X-to-4D GenerationpaperVera: Identity-Faithful Human Subject-to-Video GenerationpaperPersistent Sparse Autoencoders: Learning Feature Timescales in Language ModelspaperLook Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMspaperJoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA ModelspaperBridging Diffusion Pruning and Step Distillation with Teacher-Aligned RepairpaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperIdeas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea GenerationpaperDT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrailpersonhzh0425paperSPDCN: Strip-based Deformable Convolutional Network for Steel Surface Defect SegmentationpaperLeading-Silence Augmentation and Multi-Stage Synthetic Supervision for the Second MLC-SLM ChallengepaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperSANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video GenerationpaperELSA3D: Elastic Semantic Anchoring for Unified 3D Understanding and GenerationpaperAVSCap: Orchestrating Audio-Visual Synergy for Omni-modal Video CaptioningpaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperNative Video-Action Pretraining for Generalizable Robot ControlpersonYing1123paperSGN: A Similarity-based Generative Network for Data Generation under Distribution ShiftpaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperLoop the Loopies!paperPercepCap: Video Captioner with Structured Spatio-Temporal PerceptionpaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperLLMs and Agentic AI Systems for Smart Grids: A Tutorial on Architectures and ApplicationspaperCausalMix: Data Mixture as Causal Inference for Language Model TrainingpaperViCo3D: Empowering LiDAR-based Collaborative 3D Object Detection with Vision Foundation ModelspaperRefusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM SafetypaperCan We Trust Item Response Theory for AI Evaluation?paperReduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM InferencepaperUR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress ProxiespaperHOMIE: Human-object Centric Video Personalization via Multimodal Intelligent EnchancementnewsSeeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]paperMBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation ModelspaperConservative Query and Adaptive Regularization for Offline RL Under Uncertainty EstimationpaperNEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Visual SelectivitypaperSimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context ReasoningpersonJustinTong0323paperEvaluating and Understanding Model Editing for Medical Vision Language ModelspaperUI2App: Benchmarking Visual Interaction Inference in Executable Web Application GenerationpaperMessage Passing Enables Efficient Reasoningpersonyctseng0211paperMeanFlowNFT: Bringing Forward-Process RL to Average-Velocity GeneratorspaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperA Definition and Roadmap for World ModelspaperToolSciVer: Multimodal Scientific Claim Verification with Visual Tool Augmented Reinforcement LearningpaperSEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement LearningpaperToward Real-Time Sentence-Level Sign Language TranslationpaperAn Early Warning of Emerging Biosecurity Risks in Frontier LLMspaperDeep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning ModelspaperCLUIE: Clustering-Aware Recurrent Propagation with Local Structural Compensation for Underwater Image EnhancementpersonByronHsupaperEvidence-Backed Video Question AnsweringpaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperSelf Gradient Forcing: Native Long Video ExtrapolationpaperVideo = World + Event StreampaperTowards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment GuidancepaperWhen Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generationpersonyhyang201paperCRISP: Constrained Refinement via Iterative Squeezing Process for Robust Medical Image Segmentation under Domain ShiftpaperVisual Contrastive Self-DistillationpaperScaling Behavior Foundation Model for Humanoid RobotspaperGenAU: Language-Grounded Industrial Anomaly Understanding with Vision-Language ModelspaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperHow Does Urban Context Relate to Residential Building Health? A Vision-POI Fusion Framework for Building-Level Housing InspectionpaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPUpaperQCA: Query- and Content-Aware Keyframe Selection for Long Video UnderstandingpaperDetecting LLM-Generated Tokens in Human--LLM Coauthored TextpaperClimate-resilient electric vehicle charging infrastructure for sustainable cities: An interpretable causal-ensemble framework for preventive maintenance and low-carbon mobilitypaperIGGT4D: Streaming 4D Instance-Grounded Geometry TransformerpaperSLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPODpaperActive rejection enables reliable generalization of universal machine-learning interatomic potentialspersonmmangkadpaperAn MLIR-Based Compilation Method for Large Language ModelspaperBayesPO: Bayesian Prompt Optimization via Parallel-Tempered Gradient-Guided Discrete MCMCpaperVTLoc: Learning-based Tactile Contact Localization in Visual Point CloudspaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperQuReC: All-in-One Image Restoration with Query-Specific Guidance and Local-Global Response CalibrationpersonFridge003modeldeepseek-ai/DeepSeek-R1personmickqianpersonzhyncspaperEvoGUI: An Evolution-Aware Benchmark for GUI State-Transition UnderstandingpaperANet Patu-1: The Value of Connection in the Agent Networkpersonch-wanpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?newsBook Review: Domain-Specific Small Language Models by Guglielmo IozziapaperLearning from Synthetic Data without Model Collapse in Iterative Instruction TuningpaperDSPrompt: Dynamic Soft Prompt Defense Against M-RAG CorruptionpaperPIER: Physics-Informed Environmental Retrieval for Time-Series ModelingpaperBeyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning EvaluationpaperImproving the matrix multiplication exponent with modern optimization and AlphaEvolvepaperSynH-Rank: Quality-Aware Code Search via Diverse Data Synthesis and Hierarchical Ranking TrainingpaperMathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided RefinementpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperEdit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZpaperScale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic ManipulationpaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon ObjectivespaperMeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party MeetingspersonKangyan-ZhoupersonCatherineSuepersonb8zhongpersonalphabetc1paperDesigning Reinforcement Learning for Diffusion Models: A Unified Path-Space ViewpersonalisonshaopaperLongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU BudgetpersonfzyzcjypaperStreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video GenerationpaperVecFontLLM: Anchor-Guided Direct Synthesis of Chinese Vector FontspaperHarnessEval-W: Agentifying the Evaluation of Visual WorldspersonQiaolin-Yupersonslin1237paperVEHBench: A Stage-Local Diagnostic Benchmark for LLM-Assisted Vibration Energy Harvester DesignpaperCan We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social DisseminationpaperPatch Policy: Efficient Embodied Control via Dense Visual Representationspersonhnyls2002personispobockpersonyuan-luopersonmerrymercypaperSciDiagramEdit: Learning to Edit Scientific Diagrams from Paper RevisionspaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpersonBBufpaperStreaming Multi-Agent Autoregressive Diffusion Model with World State RegisterspaperWhen Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation
