vix.ing · top · new · best · stats · spec

Moon, Seungwhan

  1. AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model
    2023/09/27 by Seungwhan Moon, Andrea Madotto, Moon, Seungwhan +24 · 1 voice · 6 citations
    Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Natural Language Processing Techniques
  2. Large Language Models as Zero-shot Dialogue State Tracker through Function Calling
    2024/02/16 by Zekun Li, Li, Zekun, Zhiyu Zoey Chen +17 · 9 citations
    Computer Science · #Topic Modeling #Speech and dialogue systems
  3. Adding Chit-Chat to Enhance Task-Oriented Dialogues
    2020/10/24 by Sun, Kai, Moon, Seungwhan, Crook, Paul +7 · 4 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  4. SIMMC 2.0: A Task-oriented Dialog Dataset for Immersive Multimodal Conversations
    2021/04/18 by Kottur, Satwik, Moon, Seungwhan, Geramifard, Alborz +1 · 4 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  5. Active Federated Learning
    2019/09/27 by Jack Goetz, Kshitiz Malik, Goetz, Jack +8 · 3 citations
    Computer Science · #Distributed Sensor Networks and Detection Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Privacy-Preserving Technologies in Data
  6. Leveraging Slot Descriptions for Zero-Shot Cross-Domain Dialogue State Tracking
    2021/05/10 by Lin, Zhaojiang, Liu, Bing, Moon, Seungwhan +7 · 3 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  7. Continual Learning in Task-Oriented Dialogue Systems
    2020/12/31 by Andrea Madotto, Madotto, Andrea, Zhaojiang Lin +15 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Topic Modeling
  8. IMU2CLIP: Multimodal Contrastive Learning for IMU Motion Sensors from Egocentric Videos and Text
    2022/10/26 by Seungwhan Moon, Moon, Seungwhan, Andrea Madotto +11 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Video Analysis and Summarization
  9. SnapNTell: Enhancing Entity-Centric Visual Question Answering with Retrieval Augmented Multimodal LLM
    2024/03/07 by Qiu, Jielin, Madotto, Andrea, Lin, Zhaojiang +7 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  10. Zero-Shot Dialogue State Tracking via Cross-Task Transfer
    2021/09/10 by Lin, Zhaojiang, Liu, Bing, Madotto, Andrea +8 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  11. Fighting FIRe with FIRE: Assessing the Validity of Text-to-Video Retrieval Benchmarks
    2022/10/10 by Rodriguez, Pedro, Azab, Mahmoud, Silvert, Becka +4 · 1 citation
    #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  12. Embodied Executable Policy Learning with Language-based Scene Summarization
    2023/06/09 by Jielin Qiu, Mengdi Xu, Qiu, Jielin +7 · 1 citation
    Computer Science · #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning #Human Pose and Action Recognition
  13. Proactive Assistant Dialogue Generation from Streaming Egocentric Videos
    2025/06/06 by Yichi Zhang, Zhang, Yichi, Dong Xin +12 · 2 citations
    Computer Science · Psychology · #Multimodal Machine Learning Applications #Topic Modeling #Social Robot Interaction and HRI