Arrow Research search

Author name cluster

Yang Hao

Possible papers associated with this exact author name in Arrow. This page groups case-insensitive exact name matches and is not a full identity disambiguation profile.

3 papers
1 author row

Possible papers

3

NeurIPS Conference 2025 Conference Paper

Distributional LLM-as-a-Judge

  • Luyu Chen
  • Zeyu Zhang
  • Haoran Tan
  • Quanyu Dai
  • Yang Hao
  • Zhenhua Dong
  • Xu Chen

LLMs have emerged as powerful evaluators in the LLM-as-a-Judge paradigm, offering significant efficiency and flexibility compared to human judgments. However, previous methods primarily rely on single-point evaluations, overlooking the inherent diversity and uncertainty in human evaluations. This approach leads to information loss and decreases the reliability of evaluations. To address this limitation, we propose a novel training framework that explicitly aligns the LLM-generated judgment distribution with human evaluation distributions. Specifically, we propose a distributional alignment objective based on KL divergence, combined with an auxiliary cross-entropy regularization to stabilize the training process. Furthermore, due to limited human annotations, empirical human distributions are merely noisy estimates of the true underlying distribution. We therefore incorporate adversarial training to ensure a robust alignment with this true distribution, rather than overfitting to its imperfect approximation. Extensive experiments across various LLM backbones and evaluation tasks demonstrate that our framework significantly outperforms existing closed-source LLMs and conventional single-point alignment methods, with superior alignment quality, strong robustness, and competitive evaluation accuracy.

AAAI Conference 2022 Conference Paper

Self-Supervised Audio-and-Text Pre-training with Extremely Low-Resource Parallel Data

  • Yu Kang
  • Tianqiao Liu
  • Hang Li
  • Yang Hao
  • Wenbiao Ding

Multimodal pre-training for audio-and-text has recently been proved to be effective and has significantly improved the performance of many downstream speech understanding tasks. However, these state-of-the-art pre-training audio-text models work well only when provided with large amount of parallel audio-and-text data, which brings challenges on many languages that are rich in unimodal corpora but scarce of parallel cross-modal corpus. In this paper, we investigate whether it is possible to pre-train an audio-text multimodal model with extremely low-resource parallel data and extra non-parallel unimodal data. Our pre-training framework consists of the following components: (1) Intra-modal Denoising Auto-Encoding (IDAE), which is able to reconstruct input text (audio) representations from a noisy version of itself. (2) Cross-modal Denoising Auto-Encoding (CDAE), which is pre-trained to reconstruct the input text (audio), given both a noisy version of the input text (audio) and the corresponding translated noisy audio features (text embeddings). (3) Iterative Denoising Process (IDP), which iteratively translates raw audio (text) and the corresponding text embeddings (audio features) translated from previous iteration into the new less-noisy text embeddings (audio features). We adapt a dual cross-modal Transformer as our backbone model which consists of two unimodal encoders for IDAE and two cross-modal encoders for CDAE and IDP. Our method achieves comparable performance on multiple downstream speech understanding tasks compared with the model pre-trained on fully parallel data, demonstrating the great potential of the proposed method.

JBHI Journal 2016 Journal Article

Guest Editorial: MobiHealth 2014, IEEE HealthCom 2014, and IEEE BHI 2014

  • Metin Akay
  • Gouenou Coatrieux
  • Yang Hao
  • Dimitrios I. Fotiadis
  • Andrew Laine
  • Benny Lo
  • Konstantina S. Nikita
  • Norbert Noury

The papers in this special section were presented at three well-known conferences organized in 2014: EAI Mobihealth, IEEE HealthCom, and IEEE Biomedical and Health Informatics. EAI Mobihealth is an annually organized conference, which started in 2010, to address the demands of the rapidly evolving disciplines of wireless communications, mobile computing, and sensing technologies in healthcare. The IEEE-Healthcom is held every year since 1999 in different countries in Asia, Europe, and in America. It aims at bringing together interested parties working in the field of healthcare to exchange ideas, discuss innovative and emerging solutions, and develop collaborations. The IEEE Biomedical Health Informatics Conference started in 2013 and is organized every year providing the forum to showcase enabling technologies of computing, devices, imaging, sensors, and systems that optimize the acquisition, transmission, processing, storage, retrieval, visualization, and analysis of medical data. The aim of this special section is to present an overview of recent advances in sensing technologies, monitoring of patients, security and privacy of data transfer, provision of collaborative environments, data gathering and analysis from various sources, and predictive models, which all finally target the best strategy for patient monitoring and treatment.

v2026.09.13