MLLGJul 1

Deep Multitask Learning for Mixed-Type Outcomes with Shared Sparsity

arXiv:2607.009953.3
Predicted impact top 78% in ML · last 90 daysOriginality Incremental advance
AI Analysis

This work addresses the challenge of multitask learning with heterogeneous outcome types, which is common in high-dimensional biological applications, by enabling information sharing across tasks with different scales.

The paper proposes a multitask deep learning framework for mixed-type outcomes (e.g., continuous, binary) that uses unknown monotone transformations to align task-specific losses, enabling shared sparsity across tasks. The method achieves competitive prediction and variable selection in simulations and real gene-expression studies.

Most existing multitask learning approaches are limited by their reliance on task-specific loss functions tailored to the scale and type of each outcome. When outcomes differ across tasks, these losses are generally not directly comparable, which makes it difficult to formulate a unified objective and may limit information sharing across tasks. We propose a multitask transformation framework in which task-specific responses may differ through unknown monotone transformations. Motivated by high-dimensional biological applications in which the predictor dimension may diverge with the sample size while only a common subset of predictors is informative, we consider shared sparsity across tasks. Under this framework, we estimate the target functions and identify important predictors by optimizing a smoothed rank-based criterion with a group-Lasso penalty, implemented through a multitask deep neural network with a shared first layer. We establish the nonasymptotic excess-risk bounds, and variable-selection consistency for the proposed estimator. Simulation studies show that the proposed method achieves competitive prediction and variable-selection performance compared with competing approaches. Analyses of gene-expression studies with continuous, binary, and mixed outcomes further illustrate that the proposed method improves prediction and identifies biologically meaningful shared predictors.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes