vix.ing · top · new · best · stats · spec

Exploring How LLMs Capture and Represent Domain-Specific Knowledge

2025/04/23 by Mirian Hipolito Garcia, Garcia, Mirian Hipolito, Camille Couturier +13 · 1 citation
Business, Management and Accounting · Computer Science · #Business Process Modeling and Analysis #FOS: Computer and information sciences #Library Science and Information Systems #Machine Learning (cs.LG) #Semantic Web and Ontologies

paper · pdf · doi:10.48550/arxiv.2504.16871

openalex publication_date 2025/04/23 · openalex created_date 2025/10/18 · openalex updated_date 2026/07/28

Abstract

We study whether Large Language Models (LLMs) inherently capture domain-specific nuances in natural language. Our experiments probe the domain sensitivity of LLMs by examining their ability to distinguish queries from different domains using hidden states generated during the prefill phase. We reveal latent domain-related trajectories that indicate the model's internal recognition of query domains. We also study the robustness of these domain representations to variations in prompt styles and sources. Our approach leverages these representations for model selection, mapping the LLM that best matches the domain trace of the input query (i.e., the model with the highest performance on similar traces). Our findings show that LLMs can differentiate queries for related domains, and that the fine-tuned model is not always the most accurate. Unlike previous work, our interpretations apply to both closed and open-ended generative tasks

Cited by

Related