Co-occurring keywords
Papers
SpikingBERT: Distilling BERT to Train Spiking Language Models Using Implicit Differentiation
AAAI 2024
SemRoDe: Macro Adversarial Training to Learn Representations that are Robust to Word-Level Attacks
NAACL 2024
Unveiling the Magic: Investigating Attention Distillation in Retrieval-Augmented Generation
NAACL 2024
Attention Alignment and Flexible Positional Embeddings Improve Transformer Length Extrapolation
NAACL 2024
SELF-EXPERTISE: Knowledge-based Instruction Dataset Augmentation for a Legal Expert Language Model
NAACL 2024
Deja vu: Contrastive Historical Modeling with Prefix-tuning for Temporal Knowledge Graph Reasoning
NAACL 2024
Instruction Tuning with Human Curriculum
NAACL 2024