repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 22d ago
chrisliu298/awesome-llm-unlearning
A resource repository for machine unlearning in large language models
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 60%IEEE Rolls Out Large Language Models Virtual Training Course →
- PossiblePossibly related (embedding) · 58%Knowledge Distillation of Black-Box Large Language Models →
- PossiblePossibly related (embedding) · 57%Knowledge Distillation of Black-Box Large Language Models (2024) →
- PossiblePossibly related (embedding) · 57%Understanding Large Language Models →
- PossiblePossibly related (embedding) · 57%AutoTrainess: Teaching Language Models to Improve Language Models Autonomously →
- PossiblePossibly related (embedding) · 46%Probing Chemical Language Models: Effects of Pre-training and Fine-tuning →
- PossiblePossibly related (embedding) · 49%NAVER LABS Europe Submission to the Instruction-following 2026 Short Track →
- PossiblePossibly related (embedding) · 51%Object Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt Optimization →
Covers
Implements
Implements (incoming)
paperProbing Chemical Language Models: Effects of Pre-training and Fine-tuningpaperNAVER LABS Europe Submission to the Instruction-following 2026 Short TrackpaperObject Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt OptimizationpaperBayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty EstimationpaperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction TuningpaperBamiBERT: A New BERT-based Language Model for VietnamesepaperA$^{2}$utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT ConstructionpaperChallenges and Recommendations for LLMs-as-a-Judge in Multilingual Settings and Low-Resource LanguagespaperGrounded autonomous research: a fault-tolerant LLM pipeline from corpus to manuscript in frontier computational physicspaperDecompRL: Solving Harder Problems by Learning Modular Code GenerationpaperThe Future of NLP may not be at NLP Conferences: Scholarly Migration Patterns in Natural Language ProcessingpaperFast Multi-dimensional Refusal Subspaces via RFM-AGOPpaperNeuron-Aware Data Selection for Annotation-Free LLM Self-DistillationpaperDemoPSD: Disagreement-Modulated Policy Self-DistillationpaperLACUNA: A Testbed for Evaluating Localization Precision for LLM UnlearningpaperHeaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language ModelspaperHow Much is Left? LLMs Linearly Encode Their Remaining Output LengthpaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperPluraMath: Extending Mathematical Reasoning Evaluation Beyond High-Resource LanguagespaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperFrom Sinhala to Dhivehi: Cross-Lingual Transfer Learning for Low-Resource Speech RecognitionpaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperTILDE: TILt-based Distributional Erasure for Concept UnlearningpaperOn the feasibility of dependency parsing of non-human sequences without a gold standard. Is evaluation possible in other species?paperData Analysis in the Wild: Benchmarking Large Language Models Against Real-World Data ComplexitiespaperResample or Reroute? Budget-Aware Test-Time Model Selection for Large Language ModelspaperEchoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-NungpaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperPrompt Compression via Activation AggregationpaperIt Takes a MAESTRO To Prune Bad ExpertspaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperThe Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMspaperComplexity-Guided Component-wise Initialization for Language Model PretrainingpaperSuper-Tuning: From Activation-Aware Pruning to Sparse Fine-TuningpaperSelf-Guided Test-Time Training for Long-Context LLMspaperA Sovereign, Open-Source Foundation Model for German and EnglishpaperExtending LLM Context via Associative Recurrent MemorypaperJobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured ResumespaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperLanguage Identification with Succinct Machine-Independent TracespaperExtractable Memorization From First PrinciplespaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language ModelingpaperThe One-Word Census: Answer-Choice Conformity Across 44 Language Models
Covers (incoming)
Related across the graph
news[Paper] How much do language models memorize?paperObject Aligner: A Configurable JSON Schema Similarity Score for Graphs, Applied to LLM Prompt OptimizationnewsKnowledge Distillation of Black-Box Large Language ModelsnewsSetting Up Your Own Large Language Model - Towards Data SciencepaperTILDE: TILt-based Distributional Erasure for Concept UnlearningpaperSuper-Tuning: From Activation-Aware Pruning to Sparse Fine-TuningpaperLanguage Identification with Succinct Machine-Independent TracespaperDemoPSD: Disagreement-Modulated Policy Self-DistillationpaperFast Multi-dimensional Refusal Subspaces via RFM-AGOPpaperEstimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMspaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperEchoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-NungpaperLongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction SynthesispaperTowards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model FinetuningpaperKnowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language ModelingpaperExtending LLM Context via Associative Recurrent MemorypaperDecompRL: Solving Harder Problems by Learning Modular Code GenerationpaperResample or Reroute? Budget-Aware Test-Time Model Selection for Large Language ModelsnewsI developed a 270 million parameter language model entirely from scratch as an independent research projectpaperThe Future of NLP may not be at NLP Conferences: Scholarly Migration Patterns in Natural Language ProcessingpaperIt Takes a MAESTRO To Prune Bad ExpertspaperThe Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMspaperHeaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language ModelspaperJobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured ResumespaperBayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty EstimationpaperUltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic EditingpaperUnlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction TuningpaperA Sovereign, Open-Source Foundation Model for German and EnglishpaperWeak-to-Strong Generalization via Direct On-Policy DistillationpaperOn the feasibility of dependency parsing of non-human sequences without a gold standard. Is evaluation possible in other species?paperPrompt Compression via Activation AggregationnewsDetecting LLM-Generated Texts with "Classical" Machine LearningpaperFrom Sinhala to Dhivehi: Cross-Lingual Transfer Learning for Low-Resource Speech RecognitionpaperLACUNA: A Testbed for Evaluating Localization Precision for LLM UnlearningpaperChallenges and Recommendations for LLMs-as-a-Judge in Multilingual Settings and Low-Resource LanguagespaperPluraMath: Extending Mathematical Reasoning Evaluation Beyond High-Resource LanguagespaperThe One-Word Census: Answer-Choice Conformity Across 44 Language ModelspaperHow Much is Left? LLMs Linearly Encode Their Remaining Output LengthpaperUnderstanding Large Language ModelspaperGrounded autonomous research: a fault-tolerant LLM pipeline from corpus to manuscript in frontier computational physicspaperComplexity-Guided Component-wise Initialization for Language Model PretrainingnewsKnowledge Distillation of Black-Box Large Language Models (2024)paperBamiBERT: A New BERT-based Language Model for VietnamesepaperExtractable Memorization From First PrinciplesnewsIEEE Rolls Out Large Language Models Virtual Training CoursepaperAutoTrainess: Teaching Language Models to Improve Language Models AutonomouslypaperA$^{2}$utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT ConstructionpaperNeuron-Aware Data Selection for Annotation-Free LLM Self-DistillationpaperNAVER LABS Europe Submission to the Instruction-following 2026 Short TrackpaperSelf-Guided Test-Time Training for Long-Context LLMspaperData Analysis in the Wild: Benchmarking Large Language Models Against Real-World Data ComplexitiespaperProbing Chemical Language Models: Effects of Pre-training and Fine-tuning
