Research Feed Page 27

Search source-linked summaries of recent AI and machine-learning papers by topic, by date, or by whether they include code or a diagram.

Filter papers All papers

Browse by date

Resource filters

Sort options

Research results

Benchmarks & Evals / Efficiency & Inference By Shiqi Huang 2026-08-11
DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains

The paper introduces a decision-aware ranking method for supply chains that selects interventions by balancing recovered net value against operational costs.

Safety & Alignment / Benchmarks & Evals By Rahul Gupta 2026-08-11
From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop

The paper analyzes six years of TrustNLP workshop proceedings to map how research has shifted from post-hoc model interpretability to proactive control of generative systems.

Training & Fine-Tuning / Benchmarks & Evals By Artyom Sabitov 2026-08-11
Batch Size or Negatives? A Selection Rule for Memory-Constrained Recommender Training

The paper provides a selection rule to optimally allocate a fixed memory budget between batch size and negative samples when training large scale recommender systems.

Computer Vision By Rui Xu 2026-08-11
MaskFlow: Precise, Consistent and Seamless Regional Image Editing

MaskFlow provides a method for accurate and seamless regional image editing by incorporating user masks directly into the generation process and using a gradient domain refinement module.

Multimodal / Benchmarks & Evals By Yejin Jeon 2026-08-11
VoxSumm: A Multilingual Corpus of Long-Form Spoken News for Joint Summarization and Translation

VoxSumm introduces a new corpus and framework for simultaneously summarizing and translating long-form spoken news content.

Efficiency & Inference / Multimodal By Muxin Fu 2026-08-11
StreamFlow: Dynamic Memory Flows for Streaming Video Understanding

StreamFlow reduces video redundancy and improves memory efficiency by using a dynamic caching system that selectively retrieves relevant visual data during inference.

Benchmarks & Evals By Thanh-Dan Bui 2026-08-11
REAP: Relation-Aware Elicitation and Parsing for Closed-Book Knowledge Base Construction from LLMs

The REAP system improves closed-book knowledge base construction by using a two-stage process of relation-aware prompting and hybrid parsing to extract structured data from large language models.

Benchmarks & Evals / Efficiency & Inference By Zeinab Ghamlouch 2026-08-11
ReLTEx: Reliable LLM-based Taxonomy Expansion

ReLTEx improves automated taxonomy expansion by using LLMs for candidate generation combined with a structure-aware classifier to ensure hierarchical consistency.

Multimodal / Safety & Alignment By Tao Lin 2026-08-11
Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMs

The paper introduces a flexible backdoor paradigm that enables dynamic, post-training control over a Vision Language Model output by injecting trigger patterns into training data.

Reinforcement Learning / Benchmarks & Evals By Taojie Zhu 2026-08-11
ConRub-Med: Reinforcement Learning with Consensus Rubrics for Open-Ended Medical Question Answering

ConRub-Med enhances medical question answering by using automated consensus rubrics to improve reinforcement learning feedback for model responses.

Benchmarks & Evals / Training & Fine-Tuning By Tsofia Cohen 2026-08-11
MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems, Solutions, and Rationales

The paper introduces MUSE, a large-scale knowledge base of 36,960 structured problem-solution-rationale triplets extracted directly from full-text scientific papers.

Multimodal / Benchmarks & Evals By Mouxiao Huang 2026-08-11
CapProbe: Evaluating Detailed Image Captions via Full-Scene Dense Question Answering

CapProbe is a benchmark for assessing detailed image captions by using region-aligned factual questions to verify specific visual content.

Reasoning / Reinforcement Learning By Yuetian Du 2026-08-11
CARE: Confidence-Aware Reasoning for Reliable Medical VQA

The paper introduces a confidence-aware training framework that aligns medical diagnostic predictions with actual accuracy to reduce clinical decision errors.

Multimodal By Ke Ma 2026-08-11
R4DSG: Relative 4D Scene Graph Memory for Object-Centric Question Answering in Long Egocentric Video

The paper introduces a relative 4D scene graph memory system to help AI assistants answer object-centric questions in long egocentric videos.

Computer Vision / Efficiency & Inference 2026-08-11
Every Packet Counts: Dispersing Information for Loss-Resilient Learned Image Compression

The paper introduces a new image compression architecture that disperses information across packets to maintain stable visual quality even when network connections drop data.

Efficiency & Inference By Minwoo Kim 2026-08-11
ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover

ImpactHO improves LLM performance during user handovers by intelligently prioritizing and transferring essential parts of the KV cache over constrained network links.

Computer Vision / Efficiency & Inference By Vladimir Iglovikov 2026-08-11
AlbumentationsX: One Augmentation Pipeline for Images and Related Annotations

AlbumentationsX provides a centralized framework to apply synchronized image transformations across images and associated annotations.

Reasoning By Davide Rinaldi 2026-08-11
sLTN: Structural Logic Tensor Networks

The paper introduces Structural Logic Tensor Networks (sLTN) to enable logic-based reasoning over sequential or connected data structures by treating positional axes as primary components.

Multimodal / Benchmarks & Evals By Huafeng Chen 2026-08-11
PRMU: A Corpus-Free Benchmark for Person-Centric Knowledge Unlearning in Multimodal Large Language Models

The paper introduces a corpus-free framework and benchmark for deleting specific person-related knowledge from multimodal large language models without needing the original training data.

Benchmarks & Evals By Nikolai Bolik 2026-08-11
Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders

The paper investigates whether latent activations from sparse autoencoders function as meaningful, additive components for representing conceptual similarity.