GitHub·Code
6 stars·1 linked paper·Updated Jun 2026
GitHub·Code·MIT
This repository introduces MentaLLaMA, the first open-source instruction following large language model for interpretable mental health analysis.
323 stars·1 linked paper·Updated Jul 2026
GitHub·Code
Official Implementation of InstructZero; the first framework to optimize bad prompts of ChatGPT(API LLMs) and finally obtain good prompts!
202 stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
2 stars·1 linked paper·Updated May 2026
GitHub·Code
7 stars·1 linked paper·Updated Jul 2025
GitHub·Code
Code repo for EMNLP 2023 paper "Auto-Instruct: Automatic Instruction Generation and Ranking for Black-Box Language Models"
23 stars·1 linked paper·Updated Nov 2025
GitHub·Code
[NeurIPS 2024] On the Worst Prompt Performance of LLMs
11 stars·1 linked paper·Updated Dec 2025
GitHub·Code·CC-BY-4.0
1 stars·1 linked paper·Updated Jun 2026
GitHub·Code·NOASSERTION
14 stars·2 linked papers·Updated Jul 2026
GitHub·Code·Apache-2.0
This is the official repo for "PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization". PromptAgent is a novel automatic prompt optimization method that autonomously crafts prompts equivalent in quality to those handcrafted by experts,…
355 stars·1 linked paper·Updated Jul 2026
GitHub·Code
Official Code Release for "Beyond Positive Scaling: How Negation Impacts Scaling Trends of Language Models" (ACL 2023 Findings)
10 stars·1 linked paper·Updated Jun 2026
GitHub·Code·Apache-2.0
1 stars·1 linked paper·Updated Nov 2023
GitHub·Code
Code for the paper "True Few-Shot Learning in Language Models" (https://arxiv.org/abs/2105.11447)
143 stars·1 linked paper·Updated Feb 2026
GitHub·Code
[NeurIPS'24] MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution.
11 stars·1 linked paper·Updated Jul 2026
GitHub·Code
Code for "Comments as Natural Logic Pivots: Improve Code Generation via Comment Perspective"
4 stars·1 linked paper·Updated Nov 2025
GitHub·Code
The official repository for the NAACL 2024 paper "UniArk: Improving Generalisation and Consistency for Factual Knowledge Extraction through Debiasing"
2 stars·1 linked paper·Updated Oct 2025
Hugging Face·Dataset
MAIR: A Massive Benchmark for Evaluating Instructed Retrieval MAIR is a heterogeneous IR benchmark that comprises 126 information retrieval tasks across 6 domains, with annotated query-level instructions to clarify each retrieval task and relevance criteria. This repository…
516 downloads·1 linked paper·Updated Oct 2024
GitHub·Code
MAIR: A Massive Benchmark for Evaluating Instructed Retrieval. Evaluate your retrieval models on 126 diverse tasks. [EMNLP 2024]
28 stars·1 linked paper·Updated Jul 2026
GitHub·Code
4 stars·1 linked paper·Updated Dec 2024
GitHub·Code·MIT
[ACL 2025] LIFBench: A comprehensive benchmark and evaluation toolkit for assessing instruction-following capabilities and stability of large language models (LLMs) in long-context scenarios.
7 stars·1 linked paper·Updated Feb 2026
GitHub·Code
JsonTuning: Towards Generalizable, Robust, and Controllable Instruction Tuning
10 stars·1 linked paper·Updated Nov 2024
GitHub·Code·MIT
General technology for enabling AI capabilities w/ LLMs and MLLMs
4.4k stars·3 linked papers·Updated Jul 2026
GitHub·Code·Apache-2.0
0 stars·1 linked paper·Updated Mar 2022
GitHub·Code
In-context Contrastive Learning for Event Causality Identification
15 stars·1 linked paper·Updated Apr 2026
GitHub·Code·MIT
Official implementation of the paper Connecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt Optimizers
250 stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
7 stars·1 linked paper·Updated Mar 2026
GitHub·Code
EMNLP 2024 Survey Paper on LLM Evaluation
3 stars·1 linked paper·Updated Nov 2025
GitHub·Code·MIT
4 stars·1 linked paper·Updated Nov 2025
GitHub·Code
1 stars·1 linked paper·Updated Jul 2023
GitHub·Code
(NAACL 2024) VisLingInstruct: Elevating Zero-Shot Learning in Multi-Modal Language Models with Autonomous Instruction Optimization
9 stars·1 linked paper·Updated Mar 2026
GitHub·Code·MIT
1.4k stars·1 linked paper·Updated Jul 2026
GitHub·Code·NOASSERTION
6 stars·1 linked paper·Updated Apr 2026
GitHub·Code·MIT
[WWW2024 Oral] Harnessing Multi-Role Capabilities of Large Language Models for Open-Domain Question Answering
15 stars·1 linked paper·Updated Apr 2025
GitHub·Code·Apache-2.0
README
4 stars·1 linked paper·Updated Jan 2025
GitHub·Code
6 stars·1 linked paper·Updated Jun 2026
GitHub·Code·MIT
Containing code for surface form impact on math reasoning, accepted to NAACL2024
4 stars·1 linked paper·Updated Dec 2024
GitHub·Code·MIT
[ACL 2024] Source code for InBedder, an instruction-following text embedder
31 stars·1 linked paper·Updated Jun 2026
GitHub·Code
40 stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
[NAACL 2024] Data and code for our paper "Sentiment Analysis in the Era of Large Language Models: A Reality Check"
116 stars·1 linked paper·Updated May 2026
GitHub·Code
Official codebase for "Revisiting the Reliability of Language Models in Instruction-Following"
2 stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
7 stars·1 linked paper·Updated Apr 2026
GitHub·Code·Apache-2.0
Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework
2.6k stars·1 linked paper·Updated Jul 2026
GitHub·Code
[ACL MAIN 2026] Official Implementation of PIAST: Rapid Prompting with In-context Augmentation for Scarce Training data
1 stars·1 linked paper·Updated Apr 2026
GitHub·Code
[ACL 2023 findings] Towards Robust Personalized Dialogue Generation via Order-Insensitive Representation Regularization
17 stars·1 linked paper·Updated Jun 2025
GitHub·Code·MIT
5 stars·1 linked paper·Updated May 2026
GitHub·Code
5 stars·1 linked paper·Updated Dec 2025
GitHub·Code
22 stars·1 linked paper·Updated Dec 2025
GitHub·Code
Multi-level Diagnosis and Evaluation for Robust Tabular Feature Engineering with Large Language Models
1 stars·1 linked paper·Updated Dec 2025
GitHub·Code
8 stars·1 linked paper·Updated Dec 2024
GitHub·Code
131 stars·2 linked papers·Updated Mar 2026
GitHub·Code
The official GitHub page for paper "NegativePrompt: Leveraging Psychology for Large Language Models Enhancement via Negative Emotional Stimuli".
25 stars·1 linked paper·Updated Feb 2026
GitHub·Code
Official implementation of "Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers"
0 stars·1 linked paper·Updated Jan 2026
GitHub·Code·Apache-2.0
The code and files for the paper How are Prompts Different in Terms of Sensitivity
3 stars·1 linked paper·Updated Jan 2026
GitHub·Code·MIT
Optimizing Prompts with Strategy Selection (OPTS)
8 stars·1 linked paper·Updated Jun 2026
GitHub·Code
22 stars·1 linked paper·Updated Jul 2026
GitHub·Code
3 stars·1 linked paper·Updated Aug 2025
GitHub·Code·MIT
Robust machine learning for responsible AI
508 stars·4 linked papers·Updated Jul 2026
GitHub·Code
The MAGMA Benchmark is designed to evaluate the performance of large language models (LLMs) on classical graph algorithms using intermediate steps.
7 stars·1 linked paper·Updated Oct 2025
GitHub·Code
3 stars·1 linked paper·Updated Jan 2026
GitHub·Code·MIT
5 stars·1 linked paper·Updated Oct 2025
GitHub·Code·MIT
Quantifying perturbation impacts for large language models
0 stars·1 linked paper·Updated May 2025
GitHub·Code·NOASSERTION
Official repo for SAO-Instruct: Free-form Audio Editing using Natural Language Instructions presented at NeurIPS 2025
18 stars·1 linked paper·Updated Jul 2026
GitHub·Code·BSD-3-Clause
Code for our work "MSP: Multi-Stage Prompting for Making Pre-trained Language Models Better Translators" in ACL 2022
20 stars·1 linked paper·Updated Feb 2024
GitHub·Code·MIT
Resources for our ICSE'24 poster: Prompt-Enhanced Software Vulnerability Detection Using ChatGPT.
25 stars·1 linked paper·Updated Oct 2025
GitHub·Code·Archived
Label-free Node Classification on Graphs with Large Language Models (LLMS)
94 stars·1 linked paper·Updated Jul 2026
GitHub·Code
0 stars·1 linked paper·Updated Apr 2023
GitHub·Code·MIT
ICML'2022: Black-Box Tuning for Language-Model-as-a-Service & EMNLP'2022: BBTv2: Towards a Gradient-Free Future with Large Language Models
275 stars·2 linked papers·Updated Jul 2026
GitHub·Code·MIT
RUPBench: Benchmarking Reasoning Under Perturbations for Robustness Evaluation in Large Language Models
4 stars·1 linked paper·Updated Sep 2024
GitHub·Code·MIT
Code for our ICLR 2024 paper "PerceptionCLIP: Visual Classification by Inferring and Conditioning on Contexts"
80 stars·1 linked paper·Updated Mar 2026
GitHub·Code·Archived·MIT
Active Example Selection for In-Context Learning (EMNLP'22)
48 stars·1 linked paper·Updated May 2026
GitHub·Code·MIT
[CVPR 2023] OneFormer: One Transformer to Rule Universal Image Segmentation
1.7k stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
We explain why fairness metrics don't correlate and propose CAIRO to make them correlate.
2 stars·1 linked paper·Updated Feb 2025
GitHub·Code
Implementation of paper "Enhancing the Capability and Robustness of Large Language Models through Reinforcement Learning-Driven Query Refinement" (https://arxiv.org/abs/2407.01461).
2 stars·1 linked paper·Updated Jun 2026
GitHub·Code
59 stars·1 linked paper·Updated Mar 2026
GitHub·Code·Apache-2.0
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
27 stars·1 linked paper·Updated Apr 2026
GitHub·Code
Official repository of Single Domain Generalization for Few-Shot Counting via Universal Representation Matching (CVPR 2025)
6 stars·1 linked paper·Updated Jun 2026
GitHub·Code·MIT
Paradigm shift in natural language processing
42 stars·1 linked paper·Updated Jan 2024
GitHub·Code·MIT
288 stars·1 linked paper·Updated Jul 2026
GitHub·Code
Replication Package for KonTest
0 stars·1 linked paper·Updated Sep 2025
GitHub·Code·MIT
Benchmark for Basic Language Abilities of Multimodal Pretrained Transformers
3 stars·1 linked paper·Updated Mar 2025
GitHub·Code·MIT
This is the repository of DEER, a Dynamic Early Exit in Reasoning method for Large Reasoning Language Models.
200 stars·1 linked paper·Updated Jul 2026
GitHub·Code
Source code for template-based NER
212 stars·1 linked paper·Updated May 2026
GitHub·Code
MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games
28 stars·1 linked paper·Updated Jul 2026
GitHub·Code
[CVPR 2024] EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
40 stars·1 linked paper·Updated May 2026
GitHub·Code·Apache-2.0
This is the official implementation of the paper "Paraphrase Types Elicit Prompt Engineering Capabilities"
2 stars·1 linked paper·Updated Oct 2025
GitHub·Code·MIT
Optimization Modeling Using mip Solvers and large language models
284 stars·1 linked paper·Updated Jul 2026
GitHub·Code
53 stars·1 linked paper·Updated May 2026
GitHub·Code·MIT
15 stars·1 linked paper·Updated Jul 2026
GitHub·Code·MIT
Evaluating the Moral Beliefs Encoded in LLMs
38 stars·1 linked paper·Updated Jul 2026
GitHub·Code
2 stars·1 linked paper·Updated Mar 2025
GitHub·Code·MIT
Official code repository for the main conference paper in ACL2023: COLA: Contextualized Commonsense Causality Reasoning from the Causal Inference Perspective
34 stars·1 linked paper·Updated Apr 2026
GitHub·Code
The code and data for our EMNLP 2023 Findings paper: MEAL.
2 stars·1 linked paper·Updated Oct 2025
GitHub·Code
Auxiliary task demands mask the capabilities of smaller language models (COLM 2024)
7 stars·1 linked paper·Updated Nov 2025
GitHub·Code·MIT
Code for ACL paper "Zero-Shot Text Classification via Self-Supervised Tuning"
29 stars·1 linked paper·Updated Jul 2026
GitHub·Code·Apache-2.0
1 stars·1 linked paper·Updated Aug 2025
GitHub·Code·Archived·BSD-3-Clause
Resources for the "CTRLsum: Towards Generic Controllable Text Summarization" paper
149 stars·1 linked paper·Updated Jun 2026
GitHub·Code
[ICRA 2024] Official Implementation of the paper "Parameter-efficient Prompt Learning for 3D Point Cloud Understanding"
30 stars·1 linked paper·Updated Mar 2026
GitHub·Code
Code base for TALC
0 stars·1 linked paper·Updated Oct 2023
GitHub·Code
code, data and model for Paper: AlignMMBench: Evaluating Chinese Multimodal Alignment in Large Vision-Language Models (ACL'25 main)
6 stars·1 linked paper·Updated Jun 2026
GitHub·Code
[COLING 2025 Industry] LoRA Soups
20 stars·1 linked paper·Updated Apr 2026