Large Sign Language Models: Toward 3D American Sign Language Translation

arXiv — cs.CV•Wednesday, November 12, 2025 at 5:00:00 AM

The introduction of Large Sign Language Models (LSLM) marks a pivotal advancement in the translation of 3D American Sign Language (ASL), aiming to enhance digital communication for the hearing-impaired community. Unlike traditional methods that rely on 2D video, LSLM leverages 3D sign language data to capture rich spatial, gestural, and depth information, resulting in more accurate translations. This innovative approach not only improves accessibility but also explores the integration of complex multimodal languages into Large Language Models (LLMs), expanding their capabilities beyond text-based inputs. By investigating direct translation from 3D gesture features to text and incorporating instruction-guided settings, LSLM offers greater flexibility in communication. This work lays a foundational step toward creating inclusive, multimodal intelligent systems that can understand diverse forms of human communication, ultimately benefiting the hearing-impaired community.

— via World Pulse Now AI Editorial System

Read Original

Was this article worth reading? Share it

Recommended Readings

arXiv — stat.MLa day ago

Silenced Biases: The Dark Side LLMs Learned to Refuse

NegativeArtificial Intelligence

Safety-aligned large language models (LLMs) are increasingly used in sensitive applications where fairness is crucial. Evaluating their fairness is complex, often relying on standard question-answer methods that misinterpret refusal responses as indicators of fairness. This paper introduces the concept of silenced biases, which are unfair preferences hidden within the models' latent space, masked by safety-alignment. Previous methods have limitations, prompting the need for new approaches to uncover these biases effectively.

Read full article

via arXiv — stat.ML

arXiv — cs.LGa day ago

Fair In-Context Learning via Latent Concept Variables

PositiveArtificial Intelligence

The paper titled 'Fair In-Context Learning via Latent Concept Variables' explores the in-context learning (ICL) capabilities of large language models (LLMs) in handling tabular data. It highlights the potential for LLMs to inherit biases from pre-training data, which can lead to discrimination in high-stakes applications. The authors propose an optimal demonstration selection method using latent concept variables to enhance task adaptation and fairness, alongside data augmentation strategies to minimize correlations between sensitive variables and predictive outcomes.

Read full article

via arXiv — cs.LG

arXiv — cs.CL2 days ago

Modeling and Predicting Multi-Turn Answer Instability in Large Language Models

NeutralArtificial Intelligence

The paper titled 'Modeling and Predicting Multi-Turn Answer Instability in Large Language Models' discusses the evaluation of large language models (LLMs) in terms of their robustness during user interactions. The study employs multi-turn follow-up prompts to assess changes in model answers and accuracy dynamics using Markov chains. Results indicate vulnerabilities in LLMs, with a 10% accuracy drop for Gemini 1.5 Flash after a 'Think again' prompt over nine turns, and a 7.5% drop for Claude 3.5 Haiku with a reworded question. The findings suggest that accuracy can be modeled over time.

Read full article

via arXiv — cs.CL

arXiv — cs.CL2 days ago

A Multifaceted Analysis of Negative Bias in Large Language Models through the Lens of Parametric Knowledge

NeutralArtificial Intelligence

A recent study published on arXiv examines the phenomenon of negative bias in large language models (LLMs), which refers to their tendency to generate negative responses in binary decision tasks. The research highlights that previous studies have primarily focused on identifying negative attention heads that contribute to this bias. The authors introduce a new evaluation pipeline that categorizes responses based on the model's parametric knowledge, revealing that the format of prompts significantly influences the responses more than the semantics of the content itself.

Read full article

via arXiv — cs.CL

arXiv — cs.CL2 days ago

HI-TransPA: Hearing Impairments Translation Personal Assistant

PositiveArtificial Intelligence

HI-TransPA is an innovative personal assistant designed to aid hearing-impaired individuals in communication. Utilizing the Omni-Model paradigm, it integrates indistinct speech with lip dynamics to facilitate translation and dialogue. The system employs a multimodal preprocessing pipeline that enhances speech clarity by detecting facial landmarks and stabilizing lip movements, ultimately improving communication effectiveness for users.

Read full article

via arXiv — cs.CL

arXiv — cs.CL2 days ago

Identifying and Analyzing Performance-Critical Tokens in Large Language Models

NeutralArtificial Intelligence

The paper titled 'Identifying and Analyzing Performance-Critical Tokens in Large Language Models' explores how large language models (LLMs) utilize in-context learning (ICL) for few-shot learning. It categorizes tokens in ICL prompts into content, stopword, and template tokens, aiming to identify those that significantly impact LLM performance. The study reveals that template and stopword tokens have a greater influence on performance than informative content tokens, challenging existing assumptions about human attention to informative words.

Read full article

via arXiv — cs.CL

arXiv — cs.CL2 days ago

LDC: Learning to Generate Research Idea with Dynamic Control

PositiveArtificial Intelligence

Recent advancements in large language models (LLMs) highlight their potential in automating scientific research ideation. Current methods often produce ideas that do not meet expert standards of novelty, feasibility, and effectiveness. To address these issues, a new framework is proposed that combines Supervised Fine-Tuning (SFT) and controllable Reinforcement Learning (RL) to enhance the quality of generated research ideas through a two-stage approach.

Read full article

via arXiv — cs.CL

arXiv — cs.CL2 days ago

Pre-Attention Expert Prediction and Prefetching for Mixture-of-Experts Large Language Models

PositiveArtificial Intelligence

The paper titled 'Pre-Attention Expert Prediction and Prefetching for Mixture-of-Experts Large Language Models' introduces a method to enhance the efficiency of Mixture-of-Experts (MoE) Large Language Models (LLMs). The authors propose a pre-attention expert prediction technique that improves accuracy and reduces computational overhead by utilizing activations before the attention block. This approach aims to optimize expert prefetching, achieving about a 15% improvement in accuracy over existing methods.

Read full article

via arXiv — cs.CL