Source archive

arXiv — cs.LG

AIarxiv.org100 reports
Open source

Recent reports

Showing 1-12 of 100 reports

Jul 9, 2026arxiv.orgPositive

FPTQuant: Function-Preserving Transforms for LLM Quantization

The introduction of FPTQuant presents a significant advancement in the quantization of large language models (LLMs) by implementing three innovative function-preserving transforms. These transforms aim to enhance the efficiency of LLMs during inference without compromising performance, addressing the challenges posed by large magnitude outliers in naive quantization methods.

Jul 9, 2026arxiv.orgPositive

ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies

Researchers have introduced ORCAID, a new method designed to extract interpretable rule-based policies from reinforcement learning (RL) agents operating in environments with continuous action spaces. This method employs an efficient oblique decision tree training algorithm that enhances the interpretability of complex RL policies through a three-stage split search process.

Jul 9, 2026arxiv.orgPositive

TF-Engram: A Train-Free Engram with SSD-Backed Memory for Large Language Models

The introduction of TF-Engram marks a significant advancement in the field of large language models (LLMs), providing a train-free Engram system that enhances memory storage and retrieval efficiency. This system utilizes SSD-backed memory and Early-Exit Guided Predictive Prefetching to improve performance during autoregressive decoding, resulting in a notable increase in downstream task scores for the Qwen3-0.6B model.

Jul 9, 2026arxiv.orgPositive

Momentum Based Reward Design for Low Emission Traffic Signal Control

A recent study introduces a Momentum-Based Reward Function (MBRF) for low emission traffic signal control, addressing urban traffic congestion and environmental pollution. This innovative approach utilizes Deep Reinforcement Learning (DRL) to enhance adaptive traffic signal systems, evaluated through SUMO simulations, demonstrating improved throughput-emission trade-offs compared to traditional delay and queue-based rewards.

Jul 9, 2026arxiv.orgPositive

Co-LMLM: Continuous-Query Limited Memory Language Models

The introduction of Continuous-Query Limited Memory Language Models (CO-LMLM) marks a significant advancement in the field of artificial intelligence, allowing models to externalize factual knowledge to a knowledge base instead of relying solely on internal weights. This innovative approach enables the generation of flexible vector queries, integrating human-readable knowledge into outputs while overcoming limitations of traditional language models.

Jul 9, 2026arxiv.orgPositive

Dual Attention Heads for Personalized Federated Learning in ECG Classification

A new approach to personalized federated learning, named FedDualAtt, has been proposed for electrocardiogram (ECG) classification, addressing the challenges posed by heterogeneous data across healthcare providers. This method utilizes dual attention heads, with global heads capturing cross-site patterns and local heads adapting to specific institution characteristics. Experimental results indicate that FedDualAtt significantly outperforms existing methods in ECG classification tasks.

Jul 9, 2026arxiv.orgNeutral

Prior-matched evaluation of operational Earth-observation classifiers: a three-number reporting method demonstrated on Sentinel-1 internal-wave detection

The Internal Waves Service has implemented a prior-matched evaluation method for operational classifiers used in detecting internal solitary waves from Sentinel-1 satellite data. This approach addresses the discrepancies between balanced-test precision and real operational performance, revealing a significant gap in reported metrics.

Jul 9, 2026arxiv.orgNeutral

Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering

A comparative study has been conducted on eight open-source pretrained Vision-Language Models (VLMs) for Document Visual Question Answering (DocVQA), evaluating their performance across three document domains: industrial documents, infographics, and presentation slides. The study assesses model capabilities through zero-shot evaluations, supervised finetuning, and few-shot learning. Findings indicate a decline in performance for complex visual layouts despite strong zero-shot baselines.

Jul 9, 2026arxiv.orgPositive

Object Search in Partially-Known Environments via LLM-informed Model-based Planning and Prompt Selection

A novel framework for object search in partially-known environments has been introduced, utilizing LLM-informed model-based planning and prompt selection. This approach leverages a large language model (LLM) to estimate the likelihood of locating target objects across various locations, enhancing search efficiency through informed planning and travel cost analysis.

Jul 9, 2026arxiv.orgPositive

From Jumps to Signatures: a Generative Method for Temporal Point Processes

A new paper introduces sigTPP, a generative model for Temporal Point Processes (TPPs) that utilizes rough path signatures to extend signature methods to discrete event sequences. This approach addresses limitations in existing neural TPP models, which optimize per-event objectives without a global sequence-level loss. The proposed interarrival embedding offers a stable transition from jump paths to continuous paths of bounded variation.

Jul 9, 2026arxiv.orgNegative

POPS: Recovering Unlearned Multi-Modality Knowledge in MLLMs with Prompt-Optimized Parameter Shaking

A novel adversarial strategy named Prompt-Optimized Parameter Shaking (POPS) has been proposed to recover unlearned multi-modality knowledge in Multimodal Large Language Models (MLLMs). This method aims to address concerns regarding privacy-sensitive information that may be unintentionally encoded in these models, which can be exploited by malicious users.