The Power of Scale for Parameter-Efficient Prompt Tuning
2021/04/18 by Lester, Brian, Al-Rfou, Rami, Constant, Noah · 298 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
paper · doi:10.48550/arxiv.2104.08691
Abstract
In this work, we explore "prompt tuning", a simple yet effective mechanism for learning "soft prompts" to condition frozen language models to perform specific downstream tasks. Unlike the discrete text prompts used by GPT-3, soft prompts are learned through backpropagation and can be tuned to incorporate signal from any number of labeled examples. Our end-to-end learned approach outperforms GPT-3's "few-shot" learning by a large margin. More remarkably, through ablations on model size using T5, we show that prompt tuning becomes more competitive with scale: as models exceed billions of parameters, our method "closes the gap" and matches the strong performance of model tuning (where all model weights are tuned). This finding is especially relevant in that large models are costly to share and serve, and the ability to reuse one frozen model for multiple downstream tasks can ease this burden. Our method can be seen as a simplification of the recently proposed "prefix tuning" of Li and Liang (2021), and we provide a comparison to this and other similar approaches. Finally, we show that conditioning a frozen model with soft prompts confers benefits in robustness to domain transfer, as compared to full model tuning.
Cited by
- Interpretable Safety Alignment via SAE-Constructed Low-Rank Subspace Adaptation
- Rethinking Fine-Tuning: Unlocking Hidden Capabilities in Vision-Language Models
- FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
- AFA-LoRA: Enabling Non-Linear Adaptations in LoRA with Activation Function Annealing
- Selecting Language Models for Social Science: Start Small, Start Open, and Validate
- Tokenizing Numerical and Embedding Features for LLM RecSys
- Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias
- Influence of Prompt Engineering on Small Language Models for Guarded Query Routing
- A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities
- Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models
- Hierarchy-Aware Fine-Tuning of Vision-Language Models
- MoRAgent: Parameter Efficient Agent Tuning with Mixture-of-Roles
- Latent Implicit Visual Reasoning
- Deadline-Aware Online Scheduling for LLM Fine-Tuning with Spot Market Predictions
- Steering Vision-Language Pre-trained Models for Incremental Face Presentation Attack Detection
- Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
- Video Detective: Seek Critical Clues Recurrently to Answer Question from Long Videos
- Reasoning Palette: Modulating Reasoning via Latent Contextualization for Controllable Exploration for (V)LMs
- Open Ad-hoc Categorization with Contextualized Feature Learning
- Characterizing Mamba's Selective Memory using Auto-Encoders
- FlexAvatar: Learning Complete 3D Head Avatars with Partial Supervision
- PPSEBM: An Energy-Based Model with Progressive Parameter Selection for Continual Learning
- Ladder Up, Memory Down: Low-Cost Fine-Tuning With Side Nets
- Georeferencing complex relative locality descriptions with large language models
- Directional Textual Inversion for Personalized Text-to-Image Generation
- How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control
- Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches
- Rethinking Label Consistency of In-Context Learning: An Implicit Transductive Label Propagation Perspective
- Agile Deliberation: Concept Deliberation for Subjective Visual Classification
- Defect-aware Hybrid Prompt Optimization via Progressive Tuning for Zero-Shot Multi-type Anomaly Detection and Segmentation
- SOP2: Transfer Learning with Scene-Oriented Prompt Pool on 3D Object Detection
- Universal Adversarial Suffixes for Language Models Using Reinforcement Learning with Calibrated Reward
- Universal Adversarial Suffixes Using Calibrated Gumbel-Softmax Relaxation
- PVeRA: Probabilistic Vector-Based Random Matrix Adaptation
- LUNE: Efficient LLM Unlearning via LoRA Fine-Tuning with Negative Examples
- RMAdapter: Reconstruction-based Multi-Modal Adapter for Vision-Language Models
- Personalized Image Descriptions from Attention Sequences
- Group Orthogonal Low-Rank Adaptation for RGB-T Tracking
- David vs. Goliath: Can Small Models Win Big with Agentic AI in Hardware Design?
- STELLA: Guiding Large Language Models for Time Series Forecasting with Semantic Abstractions
- The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems
- MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
- Model Whisper: Steering Vectors Unlock Large Language Models' Potential in Test-time
- Enhancing Instruction-Following Capabilities in Seq2Seq Models: DoLA Adaptations for T5
- Multi-Scale Visual Prompting for Lightweight Small-Image Classification
- NAS-LoRA: Empowering Parameter-Efficient Fine-Tuning for Visual Foundation Models with Searchable Adaptation
- Network Self-Configuration based on Fine-Tuned Small Language Models
- PEFT-Factory: Unified Parameter-Efficient Fine-Tuning of Autoregressive Large Language Models
- promptolution: A Unified, Modular Framework for Prompt Optimization
- AI-Enabled grading with near-domain data for scaling feedback with human-level accuracy
- ZO-ASR: Zeroth-Order Fine-Tuning of Speech Foundation Models without Back-Propagation
- Table as a Modality for Large Language Models
- VFM-ISRefiner: Towards Better Adapting Vision Foundation Models for Interactive Segmentation of Remote Sensing Images
- SocialFusion: Addressing Social Degradation in Pre-trained Vision-Language Models
- Serving Heterogeneous LoRA Adapters in Distributed LLM Inference Systems
- Bridging Modalities via Progressive Re-alignment for Multimodal Test-Time Adaptation
- Behavior-Equivalent Token: Single-Token Replacement for Long Prompts in LLMs
- MoLT: Mixture of Layer-Wise Tokens for Efficient Audio-Visual Learning
- Auxiliary Metrics Help Decoding Skill Neurons in the Wild
- PEFT-Bench: A Parameter-Efficient Fine-Tuning Methods Benchmark
- Profile-LLM: Dynamic Profile Optimization for Realistic Personality Expression in LLMs
- Dual-domain Adaptation Networks for Realistic Image Super-resolution
- Supervised Fine Tuning of Large Language Models for Domain Specific Knowledge Graph Construction:A Case Study on Hunan's Historical Celebrities
- TS-PEFT: Unveiling Token-Level Redundancy in Parameter-Efficient Fine-Tuning
- ILoRA: Federated Learning with Low-Rank Adaptation for Heterogeneous Client Aggregation
- Dynamic Template Selection for Output Token Generation Optimization: MLP-Based and Transformer Approaches
- MGCA-Net: Multi-Grained Category-Aware Network for Open-Vocabulary Temporal Action Localization
- Backdoor Attacks on Open Vocabulary Object Detectors via Multi-Modal Prompt Tuning
- Medical Knowledge Intervention Prompt Tuning for Medical Image Classification
- CoTBox-TTT: Grounding Medical VQA with Visual Chain-of-Thought Boxes During Test-time Training
- GenSIaC: Toward Security-Aware Infrastructure-as-Code Generation with Large Language Models
- Selecting Fine-Tuning Examples by Quizzing VLMs
- GateRA: Token-Aware Modulation for Parameter-Efficient Fine-Tuning
- Text-guided Weakly Supervised Framework for Dynamic Facial Expression Recognition
- Persona-Aware Alignment Framework for Personalized Dialogue Generation
- Advanced Black-Box Tuning of Large Language Models with Limited API Calls
- Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
- Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models
- Remodeling Semantic Relationships in Vision-Language Fine-Tuning
- Adaptation of Foundation Models for Medical Image Analysis: Strategies, Challenges, and Future Directions
- Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs
- Towards Implicit Aggregation: Robust Image Representation for Place Recognition in the Transformer Era
- Reflective Personalization Optimization: A Post-hoc Rewriting Framework for Black-Box Large Language Models
- OvA-LP: A Simple and Efficient Framework for Federated Learning on Non-IID Data
- Order-Level Attention Similarity Across Language Models: A Latent Commonality
- Saliency-Guided Domain Adaptation for Left-Hand Driving in Autonomous Steering
- Plan of Knowledge: Retrieval-Augmented Large Language Models for Temporal Knowledge Graph Question Answering
- Adaptive Phase Shift Information Compression for IRS Systems: A Prompt Conditioned Variable Rate Framework
- GMoPE:A Prompt-Expert Mixture Framework for Graph Foundation Models
- Continual Learning, Not Training: Online Adaptation For Agents
- Keys in the Weights: Transformer Authentication Using Model-Bound Latent Representations
- Reviving Stale Updates: Data-Free Knowledge Distillation for Asynchronous Federated Learning
- Multi-refined Feature Enhanced Sentiment Analysis Using Contextual Instruction
- Efficiency vs. Alignment: Investigating Safety and Fairness Risks in Parameter-Efficient Fine-Tuning of LLMs
- Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
- SteerVLM: Robust Model Control through Lightweight Activation Steering for Vision Language Models
- Don't Let It Fade: Preserving Edits in Diffusion Language Models via Token Timestep Allocation
- Loquetier: A Virtualized Multi-LoRA Framework for Unified LLM Fine-tuning and Serving
- Voice Memory for Agentic Speech Recognition
- Progressive Multimodal Alignment for Continual Instruction Tuning
- ConforNets: Latents-Based Conformational Control in OpenFold3
- Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions
- Teaching Sarcasm: Few-Shot Multimodal Sarcasm Detection via Distillation to a Parameter-Efficient Student
- SCOUT: A Lightweight Framework for Scenario Coverage Assessment in Autonomous Driving
- Zero-Shot Cross-Lingual Transfer using Prefix-Based Adaptation
- LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis
- zFLoRA: Zero-Latency Fused Low-Rank Adapters
- Calibrating and Rotating: A Unified Framework for Weight Conditioning in PEFT
- DualCap: Enhancing Lightweight Image Captioning via Dual Retrieval with Similar Scenes Visual Prompts
- FLoRA: Fused forward-backward adapters for parameter efficient fine-tuning and reducing inference-time latencies of LLMs
- Agent-based Automated Claim Matching with Instruction-following LLMs
- VIPAMIN: Visual Prompt Initialization via Embedding Selection and Subspace Expansion
- Network Intrusion Detection: Evolution from Conventional Approaches to LLM Collaboration and Emerging Risks
- Beyond Higher Rank: Token-wise Input-Output Projections for Efficient Low-Rank Adaptation
- Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
- Group size effects and collective misalignment in LLM multi-agent systems
- PLAN: Proactive Low-Rank Allocation for Continual Learning
- Large Language Models Meet Text-Attributed Graphs: A Survey of Integration Frameworks and Applications
- More Than Memory Savings: Zeroth-Order Optimization Mitigates Forgetting in Continual Learning
- Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
- C-NAV: Towards Self-Evolving Continual Object Navigation in Open World
- Latent Space Factorization in LoRA
- COLA: Continual Learning via Autoencoder Retrieval of Adapters
- Restoring Pruned Large Language Models via Lost Component Compensation
- KORE: Enhancing Knowledge Injection for Large Multimodal Models via Knowledge-Oriented Controls
- FedDEAP: Adaptive Dual-Prompt Tuning for Multi-Domain Federated Learning
- ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
- Contextual Attention Modulation: Towards Efficient Multi-Task Adaptation in Large Language Models
- Prompt-MII: Meta-Learning Instruction Induction for LLMs
- All You Need is One: Capsule Prompt Tuning with a Single Vector
- DEXTER: Diffusion-Guided EXplanations with TExtual Reasoning for Vision Models
- A Guardrail for Safety Preservation: When Safety-Sensitive Subspace Meets Harmful-Resistant Null-Space
- Rewriting History: A Recipe for Interventional Analyses to Study Data Effects on Model Behavior
- Language as a Label: Zero-Shot Multimodal Classification of Everyday Postures under Data Scarcity
- Data-Model Co-Evolution: Growing Test Sets to Refine LLM Behavior
- Evolution of meta's llama models and parameter-efficient fine-tuning of large language models: a survey
- State Space Prompting via Gathering and Spreading Spatio-Temporal Information for Video Understanding
- QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs
- MeTA-LoRA: Data-Efficient Multi-Task Fine-Tuning for Large Language Models
- FedLoRA-Optimizer: Federated LoRA Fine-Tuning with Global and Local Optimization in Heterogeneous Data Scenarios
- CoSPED: Consistent Soft Prompt Targeted Data Extraction and Defense
- In-Context Learning Is Provably Bayesian Inference: A Generalization Theory for Meta-Learning
- Long Exposure: Accelerating Parameter-Efficient Fine-Tuning for LLMs under Shadowy Sparsity
- X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
- StelLA: Subspace Learning in Low-rank Adaptation using Stiefel Manifold
- Parameter-Efficient and Personalized Federated Training of Generative Models at the Edge
- Few-shot multi-token DreamBooth with LoRa for style-consistent character generation
- On the Representations of Entities in Auto-regressive Large Language Models
- Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation
- Analytical Survey of Learning with Low-Resource Data: From Analysis to Investigation
- D-TPT: Dimensional Entropy Maximization for Calibrating Test-Time Prompt Tuning in Vision-Language Models
- Vision Language Models: A Survey of 26K Papers
- AILoRA: Function-Aware Asymmetric Initialization for Low-Rank Adaptation of Large Language Models
- Upfront Chain-of-Thought: A Cooperative Framework for Chain-of-Thought Compression
- LadderMoE: Ladder-Side Mixture of Experts Adapters for Bronze Inscription Recognition
- Prompts Generalize with Low Data: Non-vacuous Generalization Bounds for Optimizing Prompts with More Informative Priors
- FlyLoRA: Boosting Task Decoupling and Parameter Efficiency via Implicit Rank-Wise Mixture-of-Experts
- SliceFine: The Universal Winning-Slice Hypothesis for Pretrained Networks
- Neologism Learning for Controllability and Self-Verbalization
- Learning to Rewrite Prompts for Bootstrapping LLMs on Downstream Tasks
- Prompt reinforcing for long-term planning of large language models
- MASA: Rethinking the Representational Bottleneck in LoRA with Multi-A Shared Adaptation
- WaveSP-Net: Learnable Wavelet-Domain Sparse Prompt Tuning for Speech Deepfake Detection
- Beyond the Seen: Bounded Distribution Estimation for Open-Vocabulary Learning
- FT-MDT: Extracting Decision Trees from Medical Texts via a Novel Low-rank Adaptation Method
- Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
- FedSRD: Sparsify-Reconstruct-Decompose for Communication-Efficient Federated Large Language Models Fine-Tuning
- MHA-RAG: Improving Efficiency, Accuracy, and Consistency by Encoding Exemplars as Soft Prompts
- LLM Based Bayesian Optimization for Prompt Search
- Inoculation Prompting: Eliciting traits from LLMs during training can suppress them at test-time
- DoRAN: Stabilizing Weight-Decomposed Low-Rank Adaptation via Noise Injection and Auxiliary Networks
- HoRA: Cross-Head Low-Rank Adaptation with Joint Hypernetworks
- Large Language Models Hallucination: A Comprehensive Survey
- SPEAR: Soft Prompt Enhanced Anomaly Recognition for Time Series Data
- Decoupling Task-Solving and Output Formatting in LLM Generation
- Deep Generative Continual Learning using Functional LoRA: FunLoRA
- HyperAdaLoRA: Accelerating LoRA Rank Allocation During Training via Hypernetworks without Sacrificing Performance
- Integrating AI and Ensemble Forecasting: Explainable Materials Planning with Scorecards and Trend Insights for a Large-Scale Manufacturer
- TokMem: Tokenized Procedural Memory for Large Language Models
- Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model Alignment
- Retrieval-Augmented Framework for LLM-Based Clinical Decision Support
- Understanding Adversarial Transfer: Why Representation-Space Attacks Fail Where Data-Space Attacks Succeed
- Exploring System 1 and 2 communication for latent reasoning in LLMs
- Efficient Layer-wise LLM Fine-tuning for Revision Intention Prediction
- LoRAFusion: Efficient LoRA Fine-Tuning for LLMs
- TAP: Two-Stage Adaptive Personalization of Multi-task and Multi-Modal Foundation Models in Federated Learning
- Zero-Shot Decentralized Federated Learning
- Vocabulary Customization for Efficient Domain-Specific LLM Deployment
- Revoking Amnesia: RL-based Trajectory Optimization to Resurrect Erased Concepts in Diffusion Models
- Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs
- Towards Structured Knowledge: Advancing Triple Extraction from Regional Trade Agreements using Large Language Models
- FedPOB: Sample-Efficient Federated Prompt Optimization via Bandits
- Prompt and Parameter Co-Optimization for Large Language Models
- Pretraining with hierarchical memories: separating long-tail and common knowledge
- GroupCoOp: Group-robust Fine-tuning via Group Prompt Learning
- Temporal Generalization: A Reality Check
- No Loss, No Gain: Gated Refinement and Adaptive Compression for Prompt Optimization
- Test-Time Policy Adaptation for Enhanced Multi-Turn Interactions with LLMs
- Memory-Efficient Fine-Tuning via Low-Rank Activation Compression
- F-Adapter: Frequency-Adaptive Parameter-Efficient Fine-Tuning in Scientific Machine Learning
- IA2: Alignment with ICL Activations Improves Supervised Fine-Tuning
- Context Parametrization with Compositional Adapters
- Enhancing Low-Rank Adaptation with Structured Nonlinear Transformations
- PSRT: Accelerating LRM-based Guard Models via Prefilled Safe Reasoning Traces
- A Tale of Two Experts: Cooperative Learning for Source-Free Unsupervised Domain Adaptation
- PreLoRA: Hybrid Pre-training of Vision Transformers with Full Training and Low-Rank Adapters
- GALAX: Graph-Augmented Language Model for Explainable Reinforcement-Guided Subgraph Reasoning in Precision Medicine
- Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation
- Crossing the Margin Cliff: Toward Relearn-Robust LLM Unlearning via Margin Calibration
- Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA
- Tight Sample Complexity for Low-Rank Adaptation: Matching Bounds and Rank Selection
- SkillSmith: Learning to Compose Parametric Skills and Textual Knowledge
- Riemannian Optimization for LoRA on the Stiefel Manifold
- Memory in Large Language Models: Mechanisms, Evaluation and Evolution
- TimeMosaic: Temporal Heterogeneity Guided Time Series Forecasting via Adaptive Granularity Patch and Segment-wise Decoding
- HyperAdapt: Simple High-Rank Adaptation
- Data Efficient Adaptation in Large Language Models via Continuous Low-Rank Fine-Tuning
- Advances in Large Language Models for Medicine
- Accurate and Efficient Low-Rank Model Merging in Core Space
- Weights-Rotated Preference Optimization for Large Language Models
- K-DeCore: Facilitating Knowledge Transfer in Continual Structured Knowledge Reasoning via Knowledge Decoupling
- Dynamic Expert Specialization: Towards Catastrophic Forgetting-Free Multi-Domain MoE Adaptation
- Federated Learning with Ad-hoc Adapter Insertions: The Case of Soft-Embeddings for Training Classifier-as-Retriever
- BEFT: Bias-Efficient Fine-Tuning of Language Models
- Distribution-Aligned Decoding for Efficient LLM Task Adaptation
- Adaptive LoRA Experts Allocation and Selection for Federated Fine-Tuning
- PILOT: Steering Synthetic Data Generation with Psychological & Linguistic Output Targeting
- Lost in Translation? Vocabulary Alignment for Source-Free Adaptation in Open-Vocabulary Semantic Segmentation
- Exploring Data and Parameter Efficient Strategies for Arabic Dialect Identifications
- Latent Traits and Cross-Task Transfer: Deconstructing Dataset Interactions in LLM Fine-tuning
- A Systematic Evaluation of Parameter-Efficient Fine-Tuning Methods for the Security of Code LLMs
- Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
- CBP-Tuning: Efficient Local Customization for Black-box Large Language Models
- NeuroStrike: Neuron-Level Attacks on Aligned LLMs
- POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
- Context-Aware Language Models for Forecasting Market Impact from Sequences of Financial News
- MAPGD: Multi-Agent Prompt Gradient Descent for Collaborative Prompt Optimization
- PHLoRA: data-free Post-hoc Low-Rank Adapter extraction from full-rank checkpoint
- CrunchLLM: Multitask LLMs for Structured Business Reasoning and Outcome Prediction
- Sensitivity-LoRA: Low-Load Sensitivity-Based Fine-Tuning for Large Language Models
- CrossPT: Exploring Cross-Task Transferability through Multi-Task Prompt Tuning
- AI-driven Remote Facial Skin Hydration and TEWL Assessment from Selfie Images: A Systematic Solution
- Few-Shot Query Intent Detection via Relation-Aware Prompt Learning
- Pre-Forgettable Models: Prompt Learning as a Native Mechanism for Unlearning
- Enhancing Technical Documents Retrieval for RAG
- Characterizing Fitness Landscape Structures in Prompt Engineering
- Singular Value Few-shot Adaptation of Vision-Language Models
- TeRA: Vector-based Random Tensor Network for High-Rank Adaptation of Large Language Models
- On the Evolution of Federated Post-Training Large Language Models: A Model Accessibility View
- Better by Comparison: Retrieval-Augmented Contrastive Reasoning for Automatic Prompt Optimization
- SSVD: Structured SVD for Parameter-Efficient Fine-Tuning and Benchmarking under Domain Shift in ASR
- Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
- MEPT: Mixture of Expert Prompt Tuning as a Manifold Mapper
- ER-LoRA: Effective-Rank Guided Adaptation for Weather-Generalized Depth Estimation
- BALM-TSF: Balanced Multimodal Alignment for LLM-Based Time Series Forecasting
- Memory Limitations of Prompt Tuning in Transformers
- Integrating Large Language Models with Network Optimization for Interactive and Explainable Supply Chain Planning: A Real-World Case Study
- A Graph Talks, But Who's Listening? Rethinking Evaluations for Graph-Language Models
- FedReFT: Federated Representation Fine-Tuning with All-But-Me Aggregation
- Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning
- ELIXIR: Efficient and LIghtweight model for eXplaIning Recommendations
- Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
- Type-Compliant Adaptation Cascades: Adapting Programmatic LM Workflows to Data
- ChronoLLM: Customizing Language Models for Physics-Based Simulation Code Generation
- J6: Jacobian-Driven Role Attribution for Multi-Objective Prompt Optimization in LLMs
- STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
- SemPT: Semantic Prompt Tuning for Vision-Language Models
- Cross-Prompt Encoder for Low-Performing Languages
- Classifier Language Models: Unifying Sparse Finetuning and Adaptive Tokenization for Specialized Classification Tasks
- Dual Information Speech Language Models for Emotional Conversations
- TAP: Parameter-efficient Task-Aware Prompting for Adverse Weather Removal
- Semantic-Enhanced Time-Series Forecasting via Large Language Models
- AutoAssert 1: A LoRA Fine-Tuned LLM Model for Efficient Automated Assertion Generation
- Score Before You Speak: Improving Persona Consistency in Dialogue Generation using Response Quality Scores
- BoRA: Towards More Expressive Low-Rank Adaptation with Block Diversity
- Contrastive Regularization over LoRA for Multimodal Biomedical Image Incremental Learning
- PRvL: Quantifying the Capabilities and Risks of Large Language Models for PII Redaction
- Optimal Corpus Aware Training for Neural Machine Translation
- Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting
- Aligning LLMs on a Budget: Inference-Time Alignment with Heuristic Reward Models
- Unified modality separation: A vision-language framework for unsupervised domain adaptation
- ETTA: Efficient Test-Time Adaptation for Vision-Language Models through Dynamic Embedding Updates
- A Foundation Model for DAS Signal Recognition and Visual Prompt Tuning of the Pre-trained Model for Downstream Tasks
- DTPA: Dynamic Token-level Prefix Augmentation for Controllable Text Generation
- Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
- Tensorized Clustered LoRA Merging for Multi-Task Interference
- Efficient Morphology-Aware Policy Transfer to New Embodiments
- EmbedGrad: Gradient-Based Prompt Optimization in Embedding Space for Large Language Models
- Variety Is the Spice of Life: Detecting Misinformation with Dynamic Environmental Representations
- MoKA: Mixture of Kronecker Adapters
- Can LLMs Generate High-Quality Task-Specific Conversations?
- Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
- Rein++: Efficient Generalization and Adaptation for Semantic Segmentation with Vision Foundation Models
- Set Pivot Learning: Redefining Generalized Segmentation with Vision Foundation Models
- Adaptive Content Restriction for Large Language Models via Suffix Optimization
- RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
- PersonaTwin: A Multi-Tier Prompt Conditioning Framework for Generating and Evaluating Personalized Digital Twins
- FaceGCD: Generalized Face Discovery via Dynamic Prefix Generation
Related