Efficient Estimation of Word Representations in Vector Space
2013/01/16 by Tomas Mikolov, Tomáš Mikolov, Kai Chen +8 · 2 voices · 484 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL
paper · pdf · doi:10.48550/arxiv.1301.3781
arxiv published 2013/01/16 · arxiv updated 2013/09/07
Abstract
We propose two novel model architectures for computing continuous vector representations of words from very large data sets. The quality of these representations is measured in a word similarity task, and the results are compared to the previously best performing techniques based on different types of neural networks. We observe large improvements in accuracy at much lower computational cost, i.e. it takes less than a day to learn high quality word vectors from a 1.6 billion words data set. Furthermore, we show that these vectors provide state-of-the-art performance on our test set for measuring syntactic and semantic word similarities.
Cited by
- The Gaining Paths to Investment Success: Information-Driven LLM Graph Reasoning for Venture Capital Prediction
- GHaLIB: A Multilingual Framework for Hope Speech Detection in Low-Resource Languages
- Chain-of-thought Reviewing and Correction for Time Series Question Answering
- Transformer Reconstructed with Dynamic Value Attention
- CosineGate: Semantic Dynamic Routing via Cosine Incompatibility in Residual Networks
- AETAS: Analysis of Evolving Temporal Affect and Semantics for Legal History
- Towards High-Level Semantic Intelligence
- Enhancing Code Understanding for Impact Analysis by Combining Transformers and Program Dependence Graphs
- DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing
- Structural Analysis of Journal Columns Using Ordinal Patterns and Information-Theoretic Measures
- Controlling Embedding Spaces with Text-Conditioned Transformations
- CHS-SQL: A Text-to-SQL approach based on Confidence-Guided Heuristic Search Schema Linking process
- Self-attention vector output similarities reveal how machines pay attention
- Better Call Graphs: A New Dataset of Function Call Graphs for Malware Classification
- Detecting cyberbullying in Spanish texts through deep learning techniques
- Event Extraction in Large Language Model
- Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
- Brain-Grounded Axes for Reading and Steering LLM States
- γ(3,4) `Attention' in Cognitive Agents: Ontology-Free Knowledge Representations With Promise Theoretic Semantics
- Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
- Unexpected Knowledge: Auditing Wikipedia and Grokipedia Search Recommendations
- From Essence to Defense: Adaptive Semantic-aware Watermarking for Embedding-as-a-Service Copyright Protection
- Convolutional Lie Operator for Sentence Classification
- Task Matrices: Linear Maps for Cross-Model Finetuning Transfer
- Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
- Citation importance-aware document representation learning for large-scale science mapping
- Semantic Distance Measurement based on Multi-Kernel Gaussian Processes
- PhraseVAE and PhraseLDM: Latent Diffusion for Full-Song Multitrack Symbolic Music Generation
- SUMFORU: An LLM-Based Review Summarization Framework for Personalized Purchase Decision Support
- CLARGA: Multimodal Graph Representation Learning over Arbitrary Sets of Modalities
- ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
- Semantic Geometry for policy-constrained interpretation
- Beyond Detection: A Comprehensive Benchmark and Study on Representation Learning for Fine-Grained Webshell Family Classification
- DMAGT: Unveiling miRNA-Drug Associations by Integrating SMILES and RNA Sequence Structures through Graph Transformer Models
- HOLE: Homological Observation of Latent Embeddings for Neural Network Interpretability
- One Word Is Not Enough: Simple Prompts Improve Word Embeddings
- Small Language Models Can Use Nuanced Reasoning For Health Science Research Classification: A Microbial-Oncogenesis Case Study
- Heard or Halted? Gender, Interruptions, and Emotional Tone in U.S. Supreme Court Oral Arguments
- Retrieving Semantically Similar Decisions under Noisy Institutional Labels: Robust Comparison of Embedding Methods
- Knowing Your Uncertainty -- On the application of LLM in social sciences
- Polynomiogram: An Integrated Framework for Root Visualization and Generative Art
- D-STEER - Preference Alignment Techniques Learn to Behave, not to Believe -- Beneath the Surface, DPO as Steering Vector Perturbation in Activation Space
- Addressing Logical Fallacies In Scientific Reasoning From Large Language Models: Towards a Dual-Inference Training Framework
- In-Context Representation Hijacking
- Investigating the originality of scientific papers across time and domain: A quantitative analysis
- Semantic Nutrition Estimation: Predicting Food Healthfulness from Text Descriptions
- Testing Transformer Learnability on the Arithmetic Sequence of Rooted Trees
- Fiber Bundle Networks: A Geometric Machine Learning Paradigm
- PG-HIVE: Hybrid Incremental Schema Discovery for Property Graphs
- Financial Text Classification Based On rLoRA Finetuning On Qwen3-8B model
- SelfAI: Building a Self-Training AI System with LLM Agents
- Identification of Malicious Posts on the Dark Web Using Supervised Machine Learning
- deepFEPS: Deep Learning-Oriented Feature Extraction for Biological Sequences
- Accumulated Local Effects and Graph Neural Networks for link prediction
- Large Language Models as Search Engines: Societal Challenges
- Fostering Innovation: Streamlining Magnetocaloric Materials Research by Digitalization
- Vector Arithmetic in Concept and Token Subspaces
- GeeSanBhava: Sentiment Tagged Sinhala Music Video Comment Data Set
- A Multiscale Geometric Method for Capturing Relational Topic Alignment
- Learning to Compress: Unlocking the Potential of Large Language Models for Text Representation
- Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition
- R-AVST: Empowering Video-LLMs with Fine-Grained Spatio-Temporal Reasoning in Complex Audio-Visual Scenarios
- Innovation by Displacement
- AssayMatch: Learning to Select Data for Molecular Activity Models
- B+ANN: A Fast Billion-Scale Disk-based Nearest-Neighbor Index
- Standardising the NLP Workflow: A Framework for Reproducible Linguistic Analysis
- Effective Diversification of Multi-Carousel Book Recommendation
- Whistledown: Combining User-Level Privacy with Conversational Coherence in LLMs
- Translation Entropy: A Statistical Framework for Evaluating Translation Systems
- How Good is BLI as an Alignment Measure: A Study in Word Embedding Paradigm
- An Evaluation Framework for Network IDS/IPS Datasets: Leveraging MITRE ATT&CK and Industry Relevance Metrics
- HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
- CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference
- ViConBERT: Context-Gloss Aligned Vietnamese Word Embedding for Polysemous and Sense-Aware Representations
- SCALEX: Scalable Concept and Latent Exploration for Diffusion Models
- Analogical Structure, Minimal Contextual Cues and Contrastive Distractors: Input Design for Sample-Efficient Linguistic Rule Induction
- Data-driven multi-species heat flux closures for two-stream-unstable plasmas with nonlinear sparse regression
- Readability Measures and Automatic Text Simplification: In the Search of a Construct
- Context-Aware Multimodal Representation Learning for Spatio-Temporally Explicit Environmental Modelling
- EPSegFZ: Efficient Point Cloud Semantic Segmentation for Few- and Zero-Shot Scenarios with Language Guidance
- Evaluation of Sentence Representations in Polish
- CL4AC: A Contrastive Loss for Audio Captioning
- TurkEmbed: Turkish Embedding Model on NLI & STS Tasks
- Harmonic Token Projection (HTP): A Vocabulary-Free, Training-Free, Deterministic, and Reversible Embedding Methodology
- TurkEmbed4Retrieval: Turkish Embedding Model for Retrieval Task
- Dual-branch Spatial-Temporal Self-supervised Representation for Enhanced Road Network Learning
- CAE: Character-Level Autoencoder for Non-Semantic Relational Data Grouping
- Large-Scale Representation Learning on Graphs via Bootstrapping
- Forget BIT, It is All about TOKEN: Towards Semantic Information Theory for LLMs
- GRAD: Graph-Retrieved Adaptive Decoding for Hallucination Mitigation
- FusionDP: Foundation Model-Assisted Differentially Private Learning for Partially Sensitive Features
- An Efficient Classification Model for Cyber Text
- The Curved Spacetime of Transformer Architectures
- Smart-Hiring: An Explainable end-to-end Pipeline for CV Information Extraction and Job Matching
- Detecting Vulnerabilities from Issue Reports for Internet-of-Things
- OceanAI: A Conversational Platform for Accurate, Transparent, Near-Real-Time Oceanographic Insights
- Exploring and Mitigating Gender Bias in Encoder-Based Transformer Models
- Embedding based Encoding Scheme for Privacy Preserving Record Linkage
- From the Rock Floor to the Cloud: A Systematic Survey of State-of-the-Art NLP in Battery Life Cycle
- A Survey on Deep Text Hashing: Efficient Semantic Text Retrieval with Binary Representation
- Contrastive Predictive Coding Done Right for Mutual Information Estimation
- Modular Linear Tokenization (MLT)
- Scalable Utility-Aware Multiclass Calibration
- Graph Fusion Network for Text Classification
- Medical Concept Representation Learning from Electronic Health Records and its Application on Heart Failure Prediction
- Characterizing semantic compositions in the brain: A model-driven fMRI re-analysis
- Joint Object and State Recognition using Language Knowledge
- Structure-Invariant Testing for Machine Translation
- From Word Embeddings to Item Recommendation
- Data-driven Rank Breaking for Efficient Rank Aggregation
- Weakly-supervised Domain Adaption for Aspect Extraction via Multi-level Interaction Transfer
- Instances of bias: the gendered semantics of generic masculines in German revealed by instance vectors
- Understanding the Origins of Bias in Word Embeddings
- VICSOM: VIsual Clues from SOcial Media for psychological assessment
- A Survey on Deep Learning for Named Entity Recognition
- Self-Supervised Learning of Context-Aware Pitch Prosody Representations
- Geometric Deep Learning: Going beyond Euclidean data
- Deep Learning Enabled Semantic Communication Systems
- BERTgrid: Contextualized Embedding for 2D Document Representation and Understanding
- Multilingual Stance Detection: The Catalonia Independence Corpus
- Reinforcement Learning-powered Semantic Communication via Semantic Similarity
- You Shall Know a User by the Company It Keeps: Dynamic Representations\n for Social Media Users in NLP
- Uncovering Implicit Gender Bias in Narratives through Commonsense Inference
- Unifying Visual-Semantic Embeddings with Multimodal Neural Language Models
- Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word\n Embeddings
- Dissociating the pre-activation of word meaning and form during sentence comprehension: Evidence from EEG representational similarity analysis
- Enhancing environmental sustainability: the impact of mission-oriented innovation policies on green innovation and patent trends
- Embracing Evolution: A Call for Body-Control Co-Design in Embodied Humanoid Robot
- Short text classification with machine learning in the social sciences: The case of climate change on Twitter
- Large Language Models and the Future of Organization Theory
- The online hostility hypothesis: representations of Muslims in online media
- Tracking Climate and Environmental Attention: A News‐Based Composite Index
- Developing Students’ Statistical Expertise Through Writing in the Age of AI
- Thinking spatially in computational social science
- Natural Language Adversarial Defense through Synonym Encoding
- Zero-Shot Audio Classification with Factored Linear and Nonlinear Acoustic-Semantic Projections
- Uncovering the structure of clinical EEG signals with self-supervised\n learning
- What the Vec? Towards Probabilistically Grounded Embeddings
- Dynamic Network Embeddings for Network Evolution Analysis
- A Survey of Fake News: Fundamental Theories, Detection Methods, and Opportunities
- Natural language processing for innovation search – Reviewing an emerging non-human innovation intermediary
- An Empirical Study and Analysis of Generalized Zero-Shot Learning for Object Recognition in the Wild
- Wikipedia2Vec: An Efficient Toolkit for Learning and Visualizing the Embeddings of Words and Entities from Wikipedia
- #MeTooMaastricht: Building a chatbot to assist survivors of sexual\n harassment
- Continuous-Flow Graph Transportation Distances
- What to Prioritize? Natural Language Processing for the Development of a Modern Bug Tracking Solution in Hardware Development
- Linguists Who Use Probabilistic Models Love Them: Quantification in\n Functional Distributional Semantics
- Skip-Thought Vectors
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators
- Testing of detection tools for AI-generated text
- HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million\n Narrated Video Clips
- Uniting Heterogeneity, Inductiveness, and Efficiency for Graph Representation Learning
- Can LLMs Estimate Cognitive Complexity of Reading Comprehension Items?
- Predicting electric vehicle charging demand using a heterogeneous spatio-temporal graph convolutional network
- Metadata-Driven Retrieval-Augmented Generation for Financial Question Answering
- Text Simplification with Sentence Embeddings
- Out-of-distribution generalization via composition: A lens through induction heads in Transformers
- Mapping the unseen in practice: comparing latent Dirichlet allocation and BERTopic for navigating topic spaces
- Transfer Learning for Multi-lingual Tasks -- a Survey
- SwiftEmbed: Ultra-Fast Text Embeddings via Static Token Lookup for Real-Time Applications
- M3T2IBench: A Large-Scale Multi-Category, Multi-Instance, Multi-Relation Text-to-Image Benchmark
- SALSA: Single-pass Autoregressive LLM Structured Classification
- Frustratingly Easy Task-aware Pruning for Large Language Models
- MMbeddings: Parameter-Efficient, Low-Overfitting Probabilistic Embeddings Inspired by Nonlinear Mixed Models
- Foundation of Intelligence: Review of Math Word Problems from Human Cognition Perspective
- δ-STEAL: LLM Stealing Attack with Local Differential Privacy
- Joint Multimedia Event Extraction from Video and Article
- DeepBalance: Deep-Learning and Fuzzy Oversampling for Vulnerability Detection
- The Need for Standardized Explainability
- DAIL: Beyond Task Ambiguity for Language-Conditioned Reinforcement Learning
- Grounding learning of modifier dynamics: An application to color naming
- No Intelligence Without Statistics: The Invisible Backbone of Artificial Intelligence
- Large-scale User Game Lifecycle Representation Learning
- Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task
- Distributional Semantics: Meaning Through Culture and Interaction
- Artificial intelligence, 21st century competences, and socio‐emotional learning in education: More than high‐risk?
- Approximate Nearest Neighbor Search of Large Scale Vectors on Distributed Storage
- Does Visual Grounding Enhance the Understanding of Embodied Knowledge in Large Language Models?
- Word2vec Skip-gram Dimensionality Selection via Sequential Normalized Maximum Likelihood
- FarsiMCQGen: a Persian Multiple-choice Question Generation Framework
- Beyond "Hallucinations": A Framework for Stable Human-AI Reasoning
- Hierarchical Semantic Retrieval with Cobweb
- Graph Neural Networks: A Review of Methods and Applications
- Unlocking Public Catalogues: Instruction-Tuning LLMs for ICD Coding of German Tumor Diagnoses
- Document Intelligence in the Era of Large Language Models: A Survey
- Do You Get the Hint? Benchmarking LLMs on the Board Game Concept
- Revisiting Query Variants: The Advantage of Retrieval Over Generation of Query Variants for Effective QPP
- Text Anomaly Detection with Simplified Isolation Kernel
- ProtoTopic: Prototypical Network for Few-Shot Medical Topic Modeling
- Attribution Quality in AI-Generated Content:Benchmarking Style Embeddings and LLM Judges
- CARVQ: Corrective Adaptor with Group Residual Vector Quantization for LLM Embedding Compression
- ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification
- SAIL-Embedding Technical Report: Omni-modal Embedding Foundation Model
- WiC: the Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations
- Quest for Orthologs in the era of Data Deluge and AI: Challenges and Innovations in Orthology Prediction and Data Integration
- CNSocialDepress: A Chinese Social Media Dataset for Depression Risk Detection and Structured Analysis
- Connecting Giants: Synergistic Knowledge Transfer of Large Multimodal Models for Few-Shot Learning
- QLENS: Towards A Quantum Perspective of Language Transformers
- Robust ML-based Detection of Conventional, LLM-Generated, and Adversarial Phishing Emails Using Advanced Text Preprocessing
- Glance for Context: Learning When to Leverage LLMs for Node-Aware GNN-LLM Fusion
- Utilizing FastText for Venue Recommendation
- Serialized EHR make for good text representations
- Strong Prediction: Language Model Surprisal Explains Multiple N400 Effects
- Learning Semantic Representations for the Phrase Translation Model
- The Geometry of Reasoning: Flowing Logics in Representation Space
- Steering Embedding Models with Geometric Rotation: Mapping Semantic Relationships Across Languages and Models
- Domain-Adapted Pre-trained Language Models for Implicit Information Extraction in Crash Narratives
- On the Representations of Entities in Auto-regressive Large Language Models
- Contextual Word Representations: A Contextual Introduction
- Web Crawler Restrictions, AI Training Datasets & Political Biases
- Psyzkaller: Learning from Historical and On-the-Fly Execution Data for Smarter Seed Generation in OS kernel Fuzzing
- Understanding the Effects of Domain Finetuning on LLMs
- SMedBERT: A Knowledge-Enhanced Pre-trained Language Model with Structured Semantics for Medical Text Mining
- Learning to Look at the Other Side: A Semantic Probing Study of Word Embeddings in LLMs with Enabled Bidirectional Attention
- A Comparative Study on Structural and Semantic Properties of Sentence Embeddings
- Neologism Learning for Controllability and Self-Verbalization
- Investigating Thematic Patterns and User Preferences in LLM Interactions using BERTopic
- Label Semantics for Robust Hyperspectral Image Classification
- Does Local News Stay Local?: Online Content Shifts in Sinclair-Acquired Stations
- SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora
- End-to-End Test-Time Training for Long Context
- Investigating Industry--Academia Collaboration in Artificial Intelligence: PDF-Based Bibliometric Analysis from Leading Conferences
- ImageNet Large Scale Visual Recognition Challenge
- PTEB: Towards Robust Text Embedding Evaluation via Stochastic Paraphrasing at Evaluation Time with LLMs
- A Framework for Measuring How News Topics Drive Stock Movement
- KEEP: Integrating Medical Ontologies with Clinical Data for Robust Code Embeddings
- Nodes and networks in Diachronic Construction Grammar
- NatGVD: Natural Adversarial Example Attack towards Graph-based Vulnerability Detection
- Detecting Semantic Clones of Unseen Functionality
- Equation Embeddings
- Generating High-Level Test Cases from Requirements using LLM: An Industry Study
- Evaluating Embedding Frameworks for Scientific Domain
- AbsTopK: Rethinking Sparse Autoencoders For Bidirectional Features
- Finding beans in burgers: Deep semantic-visual embedding with\n localization
- Cracking the code of adaptive immunity: The role of computational tools
- Automated Alignment of Math Items to Content Standards in Large-Scale Assessments Using Language Models
- Text-Based Approaches to Item Alignment to Content Standards in Large-Scale Reading & Writing Tests
- Beyond Linear Probes: Dynamic Safety Monitoring for Language Models
- Galton's Law of Mediocrity: Why Large Language Models Regress to the Mean and Fail at Creativity in Advertising
- CustomIR: Unsupervised Fine-Tuning of Dense Embeddings for Known Document Corpora
- The Loss Kernel: A Geometric Probe for Deep Learning Interpretability
- Neural network embeddings recover value dimensions from psychometric survey items on par with human data
- SemShareKV: Efficient KVCache Sharing for Semantically Similar Prompts via Token-Level LSH Matching
- Knowledge Editing with Subspace-Aware Key-Value Mappings
- LogAction: Consistent Cross-system Anomaly Detection through Logs via Active Domain Adaptation
- PEARL: Performance-Enhanced Aggregated Representation Learning
- Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
- Graph Optimization Foundation Model: Tokenizing Graph via A Language-Model Paradigm
- Task Vectors, Learned Not Extracted: Performance Gains and Mechanistic Insight
- Emergent World Representations in OpenVLA
- Advancing mathematics research with generative AI
- Sarcasm Analysis Using Conversation Context
- Using large language models to extract plant functional traits from unstructured text
- The Geometry of Culture: Analyzing the Meanings of Class through Word Embeddings
- Social Media Narratives Across Platforms in Conflict: Evidence from Syria
- An Improved Framework for Scaling Party Positions from Texts with Transformer
- Protocol paper: From Chaos to Order. Augmenting Manual Article Screening with Sentence Transformers in Management Systematic Reviews
- Revisit Systematic Generalization via Meaningful Learning
- Can Local Learning Match Self-Supervised Backpropagation?
- Gender bias in AI-based decision-making systems: a systematic literature review
- Items Outperform Adjectives in a Computational Model of Binary Semantic Classification
- A shared model-based linguistic space for transmitting our thoughts from brain to brain in natural conversations
- Neo-Grounded Theory: A Methodological Innovation Integrating High-Dimensional Vector Clustering and Multi-Agent Collaboration for Qualitative Research
- Representing LLMs in Prompt Semantic Task Space
- One Prompt Fits All: Universal Graph Adaptation for Pretrained Models
- A model of errors in transformers
- SoK: Potentials and Challenges of Large Language Models for Reverse Engineering
- AI Brown and AI Koditex: LLM-Generated Corpora Comparable to Traditional Corpora of English and Czech Texts
- Semantic-Inductive Attribute Selection for Zero-Shot Learning
- MonoCon: A general framework for learning ultra-compact high-fidelity representations using monotonicity constraints
- GRAB: A Risk Taxonomy--Grounded Benchmark for Unsupervised Topic Discovery in Financial Disclosures
- Semantic F1 Scores: Fair Evaluation Under Fuzzy Class Boundaries
- A short survey on almost orthogonal vectors in a few specific large dimensions
- CAD-Tokenizer: Towards Text-based CAD Prototyping via Modality-Specific Tokenization
- Latent Activation Editing: Inference-Time Refinement of Learned Policies for Safer Multirobot Navigation
- Probability Signature: Bridging Data Semantics and Embedding Structure in Language Models
- Estimating affective polarization on a social network
- Generation of focused drug molecule library using recurrent neural network
- PlantBGC: Transformer for Plant BGC Discovery via Label-Free Domain Adaptation and Weak Supervision
- When AI Meets Science: Research Diversity, Interdisciplinarity, Visibility, and Retractions across Disciplines in a Global Surge
- Is Child-Directed Language Optimized for Word Learning? A Computational Study of Verb Meaning Acquisition
- One for All: Neural Joint Modeling of Entities and Events
- Modelling Compositionality and Structure Dependence in Natural Language
- A quantum reservoir computing approach to computer-aided music composition
- The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models
- The Semantic Similarity Effect on Short-Term Memory: Null Effects of Affectively Defined Semantic Similarity
- SimLex-999: Evaluating Semantic Models With (Genuine) Similarity Estimation
- Language model-based B cell receptor sequence embeddings can effectively encode receptor specificity
- Data Sets: Word Embeddings Learned from Tweets and General Data
- A study of word embedding models for measuring topic coherence
- AMELIA: A Family of Multi-task End-to-end Language Models for Argumentation
- Word Embeddings for the Armenian Language: Intrinsic and Extrinsic Evaluation
- Does Diversity of Expertise Drive Citation Impact? Evidence from Computer Science
- TRACE: Early Detection of Chronic Kidney Disease Onset with Transformer-Enhanced Feature Embedding
- Neural Simile Recognition with Cyclic Multitask Learning and Local Attention
- Semantic Search for Information Retrieval
- Global Minimizers of Sigmoid Contrastive Loss
- Heterogeneous co-occurrence embedding for visual information exploration
- Discovering Differential Features: Adversarial Learning for Information Credibility Evaluation
- Learning to Condition: A Neural Heuristic for Scalable MPE Inference
- MAJORScore: A Novel Metric for Evaluating Multimodal Relevance via Joint Representation
- Robustness and Reliability of Gender Bias Assessment in Word Embeddings: The Role of Base Pairs
- Not a Collaborator or a Supervisor, but an Assistant: Striking the Balance Between Efficiency and Ownership in AI-incorporated Qualitative Data Analysis
- Graph Coloring for Multi-Task Learning
- KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
- Can BERT predict fillers for construction elements?
- Predicting the descent into extremism and terrorism
- Query2Prod2Vec Grounded Word Embeddings for eCommerce
- A Comparative Analysis of Transformer Models in Social Bot Detection
- A Weak Supervision Approach for Monitoring Recreational Drug Use Effects in Social Media
- A Systematic Literature Review on Multi-label Data Stream Classification
- Modeling User Redemption Behavior in Complex Incentive Digital Environment: An Empirical Study Using Large-Scale Transactional Data
- Hierarchical Self-Attention: Generalizing Neural Attention Mechanics to Multi-Scale Problems
- SSL-SSAW: Self-Supervised Learning with Sigmoid Self-Attention Weighting for Question-Based Sign Language Translation
- Analyze the Effects of Weighting Functions on Cost Function in the Glove Model
- Graph Neural Network based Service Function Chaining for Automatic Network Control
- Speech-Based Cognitive Screening: A Systematic Evaluation of LLM Adaptation Strategies
- Unsupervised Anomaly Detection in ALS EPICS Event Logs
- Green Recommender Systems: Understanding and Minimizing the Carbon Footprint of AI-Powered Personalization
- Equalizing Gender Biases in Neural Machine Translation with Word Embeddings Techniques
- Textarium: Entangling Annotation, Abstraction and Argument
- Towards the Improvement of Automated Scientific Document Categorization by Deep Learning
- Transformer Query-Target Knowledge Discovery (TEND): Drug Discovery from CORD-19
- Exploring Distributed Vector Databases Performance on HPC Platforms: A Study with Qdrant
- Linear Dimensionality Reduction for Word Embeddings in Tabular Data Classification
- Pun Unintended: LLMs and the Illusion of Humor Understanding
- Cyber Threat Hunting: Non-Parametric Mining of Attack Patterns from Cyber Threat Intelligence for Precise Threats Attribution
- Literaturwissenschaft und Informatik
- Shared-Private Bilingual Word Embeddings for Neural Machine Translation
- Spontaneous eye movements reflect the representational geometries of conceptual spaces
- Compositional Morphology for Word Representations and Language Modelling
- Semantic Fusion with Fuzzy-Membership Features for Controllable Language Modelling
- Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers
- Quantifying Compositionality of Classic and State-of-the-Art Embeddings
- A study of Turkish emotion classification with pretrained language models
- Weakly Supervised Vulnerability Localization via Multiple Instance Learning
- A Survey on Retrieval And Structuring Augmented Generation with Large Language Models
- Dual Encoding for Zero-Example Video Retrieval
- Targeted Test Selection Approach in Continuous Integration
- A Critical Review of Recurrent Neural Networks for Sequence Learning
- Towards Explainable Job Title Matching: Leveraging Semantic Textual Relatedness and Knowledge Graphs
- Let's Simply Count: Quantifying Distributional Similarity Between Activities in Event Data
- Modelling Analogies and Analogical Reasoning: Connecting Cognitive Science Theory and NLP Research
- Topic Modeling with Contextualized Word Representation Clusters
- Multi-Label Image Recognition with Graph Convolutional Networks
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet
- TriagerX: Dual Transformers for Bug Triaging Tasks with Content and Interaction Based Rankings
- Modeling Protein Using Large-scale Pretrain Language Model
- From Detection to Mitigation: Addressing Gender Bias in Chinese Texts via Efficient Tuning and Voting-Based Rebalancing
- Statistical Methods in Generative AI
- Modelling Intertextuality with N-gram Embeddings
- From Noise to Narrative: Tracing the Origins of Hallucinations in Transformers
- Empirical Study of Code Large Language Models for Binary Security Patch Detection
- Iterative Shrinking for Referring Expression Grounding Using Deep Reinforcement Learning
- KnowHow: Automatically Applying High-Level CTI Knowledge for Interpretable and Accurate Provenance Analysis
- Revealing the Numeracy Gap: An Empirical Investigation of Text Embedding Models
- QCSE: A Pretrained Quantum Context-Sensitive Word Embedding for Natural Language Processing
- Self-supervised Learning for Hyperspectral Images of Trees
- Foundational Models and Federated Learning: Survey, Taxonomy, Challenges and Practical Insights
- Between fact and fairy: tracing the hallucination metaphor in AI discourse
- Delta Activations: A Representation for Finetuned Large Language Models
- Contextualized Token Discrimination for Speech Search Query Correction
- A Framework for Generative and Contrastive Learning of Audio Representations
- Explicit and Implicit Data Augmentation for Social Event Detection
- VulRTex: A Reasoning-Guided Approach to Identify Vulnerabilities from Rich-Text Issue Report
- Cooperative Grasping for Collective Object Transport in Constrained Environments
- Explainable Knowledge Graph Retrieval-Augmented Generation (KG-RAG) with KG-SMILE
- Learning Mechanism Underlying NLP Pre-Training and Fine-Tuning
- Using LLMs to create analytical datasets: A case study of reconstructing the historical memory of Colombia
- Multi-modal Transformer for Video Retrieval
- TopoMap: A Feature-based Semantic Discriminator of the Topographical Regions in the Test Input Space
- Training LLMs to be Better Text Embedders through Bidirectional Reconstruction
- Question Answering through Transfer Learning from Large Fine-grained Supervision Data
- An experimental and computational study of an Estonian single-person word naming
- PointAD+: Learning Hierarchical Representations for Zero-shot 3D Anomaly Detection
- Gnowsis: Multimodal multitask learning for oral proficiency assessments
- Representation Learning of Reconstructed Graphs Using Random Walk Graph Convolutional Network
- Jointly Reinforcing Diversity and Quality in Language Model Generations
- chDzDT: Word-level morphology-aware language model for Algerian social media text
- Compiler Bugs Detection in Logic Synthesis Tools via Linear Upper Confidence Bound
- Testing the assumptions about the geometry of sentence embedding spaces: the cosine measure need not apply
- Automatic Product Ontology Extraction from Textual Reviews
- Speaker-Conditioned Phrase Break Prediction for Text-to-Speech with Phoneme-Level Pre-trained Language Model
- Predicting Multi-Type Talented Students in Secondary School Using Semi-Supervised Machine Learning
- Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion
- HSFN: Hierarchical Selection for Fake News Detection building Heterogeneous Ensemble
- Discovering Semantic Subdimensions through Disentangled Conceptual Representations
- Deep Learning Based Concurrency Bug Detection and Localization
- Transparent Semantic Spaces: A Categorical Approach to Explainable Word Embeddings
- From Post To Personality: Harnessing LLMs for MBTI Prediction in Social Media
- Network Representation Learning: From Traditional Feature Learning to Deep Learning
- Are Words Commensurate with Actions? Quantifying Commitment to a Cause from Online Public Messaging
- Character-Aware Neural Language Models
- Tutorial on the Probabilistic Unification of Estimation Theory, Machine Learning, and Generative AI
- SciTopic: Enhancing Topic Discovery in Scientific Literature through Advanced LLM
- Automated Bug Triaging using Instruction-Tuned Large Language Models
- GegenNet: Spectral Convolutional Neural Networks for Link Sign Prediction in Signed Bipartite Graphs
- The Double-edged Sword of LLM-based Data Reconstruction: Understanding and Mitigating Contextual Vulnerability in Word-level Differential Privacy Text Sanitization
- SEA: Sentence Encoder Assembly for Video Retrieval by Textual Queries
- EvoFormer: Learning Dynamic Graph-Level Representations with Structural and Temporal Bias Correction
- Towards Plausible Graph Anonymization
- Deconstructing and reconstructing word embedding algorithms
- Low-Complexity Data-Parallel Earth Mover's Distance Approximations
- A Comparison of Neural Network Training Methods for Text Classification
- Comparative Evaluation of Text and Audio Simplification: A Methodological Replication Study
- Simulating Name-like Vectors for Testing Large-scale Entity Resolution
- ReviewGraph: A Knowledge Graph Embedding Based Framework for Review Rating Prediction with Sentiment Features
- The illusion of a perfect metric: Why evaluating AI's words is harder than it looks
- Can Large Language Models (LLMs) Describe Pictures Like Children? A Comparative Corpus Study
- Prediction is not Explanation: Revisiting the Explanatory Capacity of Mapping Embeddings
- Driving Style Recognition Like an Expert Using Semantic Privileged Information from Large Language Models
- A Risk Manager for Intrusion Tolerant Systems: Enhancing HAL 9000 with New Scoring and Data Sources
- Learning to Steer: Input-dependent Steering for Multimodal LLMs
- SDEC: Semantic Deep Embedded Clustering
- Discrete Word Embedding for Logical Natural Language Understanding
- Feature Request Analysis and Processing: Tasks, Techniques, and Trends
- The Structural Sources of Verb Meaning Revisited: Large Language Models Display Syntactic Bootstrapping
- Investigating Transcription Normalization in the Faetar ASR Benchmark
- Reference Points in LLM Sentiment Analysis: The Role of Structured Context
- CoDiEmb: A Collaborative yet Distinct Framework for Unified Representation Learning in Information Retrieval and Semantic Textual Similarity
- Cross-language Information Retrieval
- Borrowing From the Future: Enhancing Early Risk Assessment through Contrastive Learning
- Towards the Next-generation Bayesian Network Classifiers
- Scalable Geospatial Data Generation Using AlphaEarth Foundations Model
- Copyright Protection for Large Language Models: A Survey of Methods, Challenges, and Trends
- Hybrid-Hierarchical Fashion Graph Attention Network for Compatibility-Oriented and Personalized Outfit Recommendation
- Abundance-Aware Set Transformer for Microbiome Sample Embedding
- Image Captioning with Visual Object Representations Grounded in the Textual Modality
- Semantic Relatedness and Taxonomic Word Embeddings
- A Survey of Cognitive Distortion Detection and Classification in NLP
- Visually Aware Skip-Gram for Image Based Recommendations
- Multimodal Metric Learning for Tag-based Music Retrieval
- DeepBugs: A Learning Approach to Name-based Bug Detection
- Machine Learning for Multimodal Electronic Health Records-based Research: Challenges and Perspectives
- EXSCLAIM! -- An automated pipeline for the construction of labeled materials imaging datasets from literature
- Analyzing HPC Support Tickets: Experience and Recommendations
- MiGrATe: Mixed-Policy GRPO for Adaptation at Test-Time
- DevNous: An LLM-Based Multi-Agent System for Grounding IT Project Management in Unstructured Conversation
- From Source to Target: Leveraging Transfer Learning for Predictive Process Monitoring in Organizations
- Large Language Models for Subjective Language Understanding: A Survey
- GLiClass: Generalist Lightweight Model for Sequence Classification Tasks
- Recovering link-weight structure in complex networks with weight-aware random walks
- Enhancing Rumor Detection Methods with Propagation Structure Infused Language Model
- Towards Real-World Rumor Detection: Anomaly Detection Framework with Graph Supervised Contrastive Learning
- SEVADE: Self-Evolving Multi-Agent Analysis with Decoupled Evaluation for Hallucination-Resistant Irony Detection
- Adversarial Video Promotion Against Text-to-Video Retrieval
- Vec2Summ: Text Summarization via Probabilistic Sentence Embeddings
- Making Global Norms
- Vocabulary Manipulation for Neural Machine Translation
- Identifier Namespaces in Mathematical Notation
- Medical knowledge embedding based on recursive neural network for multi-disease diagnosis
- AI with Symbolic Empathy: Shannon-Neumann Insight Guided Logic
- Benchmarking Pretrained Molecular Embedding Models For Molecular Representation Learning
- Hypergraph Neural Network with State Space Models for Node Classification
- Zero-Shot Learning by Convex Combination of Semantic Embeddings
- Quantifying the Generalization Gap: A New Benchmark for Out-of-Distribution Graph-Based Android Malware Classification
- A Study of the Framework and Real-World Applications of Language Embedding for 3D Scene Understanding
- An Effective Approach for Node Classification in Textual Graphs
- Compressing Large Language Models with PCA Without Performance Loss
- Graph Representation Learning with Massive Unlabeled Data for Rumor Detection
- Factor Augmented Supervised Learning with Text Embeddings
- MalFlows: Context-aware Fusion of Heterogeneous Flow Semantics for Android Malware Detection
- LLM-based IR-system for Bank Supervisors
- Semantic Structure in Large Language Model Embeddings
- Improving Hospital Risk Prediction with Knowledge-Augmented Multimodal EHR Modeling
- UEChecker: Detecting Unchecked External Call Vulnerabilities in DApps via Graph Analysis
- NATLM: Detecting Defects in NFT Smart Contracts Leveraging LLM
- Compositionality in the semantic network: a model-driven representational similarity analysis
- Unlocking New York City Crime Insights using Relational Database Embeddings
- Investigating the Invertibility of Multimodal Latent Spaces: Limitations of Optimization-Based Methods
- Real-time News Story Identification
- Modality-Aware Feature Matching: A Comprehensive Review of Single- and Cross-Modality Techniques
- Resource-Efficient Adaptation of Large Language Models for Text Embeddings via Prompt Engineering and Contrastive Fine-tuning
- Explaining Natural Language Processing Classifiers with Occlusion and Language Modeling
- BanditSum: Extractive Summarization as a Contextual Bandit
- A Scalable Pipeline for Estimating Verb Frame Frequencies Using Large Language Models
- Self-Supervised Contextual Language Representation of Radiology Reports to Improve the Identification of Communication Urgency
- Argument Invention from First Principles
- Adversarial Defence without Adversarial Defence: Enhancing Language Model Robustness via Instance-level Principal Component Removal
- Comparison of Information Retrieval Techniques Applied to IT Support Tickets
- Priority-Aware Clinical Pathology Hierarchy Training for Multiple Instance Learning
- An empirical comparison of deep-neural-network architectures for next activity prediction using context-enriched process event logs
- Wafer Defect Root Cause Analysis with Partial Trajectory Regression
- Balancing the composition of word embeddings across heterogenous data\n sets
Discussions
Related