15 papers across AI, ML, NLP, and CV from the last 24 hours.
Three currents run through today's batch. First, AI governance is moving from abstract principles to concrete mechanisms: one paper formalizes participatory governance of deployed agents through compute-budget allocation, while another audits digital sovereignty in Nigerian mobile commerce apps. Second, medical imaging is shifting from model performance chasing toward real-world accessibility — smartphone-based tooth detection, lightweight on-device sign language recognition, and synthetic lesion generation via optimal transport all target settings where clinical infrastructure is unavailable. Third, a quiet but deep statistical undercurrent runs through several papers: scalable VARMA estimation, learning under monotone adversaries, uncertainty-aware LiDAR localization, and stochastic dynamics on persistence diagram spaces all treat uncertainty and identifiability as first-class concerns rather than afterthoughts.
The standout is AV-AIVAT, which achieves 74x cheaper agent evaluation by applying anytime-valid statistical stopping to imperfect-information game benchmarks. The framing is deceptively simple — stop evaluating as soon as evidence suffices, without invalidating confidence guarantees — but it cuts directly at a practical bottleneck in agent research: every benchmark run costs inference, and most of that spend happens after the answer is already clear.
Collectively, this batch suggests a field growing more mature — less interested in raw capability claims, more focused on how to evaluate, govern, and deploy AI systems under real constraints.
George Grispos, Sajda Qureshi · 2026-08-06
The use of e-commerce mobile applications is expanding in Nigeria, creating both opportunities and risks, including fraud and reduced user control over digital technologies, raising concerns about digital sovereignty. This research examines how Artificial Intelligence (AI) in Nigerian mobile applications affects digital sovereignty, examined through platform transparency as a key indicator of user
Boning Li, Yu Chen, Longbo Huang · 2026-08-06
Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time. Since the number of games needed is unknown, fixed-budget evaluations either keep paying after the result is settled or stop before the agents can be told apart, while naive optional stopping with an ordinary confidence interval invalidates the state
Praphul Chandra, Sujit Gujar, Ganesh Ghalme · 2026-08-06
We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle that governance should control an AI agent through resource allocation so as to make authorization self enforcing via compute budgets. The mechanism seeks to establish the Safe AI paradigm that compute is an effective governance lever. We situate our w
Donna Hooshmand, Shubham Shahi, Cameron Barrie · 2026-08-06
From natural-language query interfaces to automated report generation, data analysis tools need a description of the data: the real-world entities it contains, which columns function as measures or identifiers, and how tables connect into units of analysis. Today, this semantic layer is usually written by hand. This is a knowledge-acquisition bottleneck that limits the scalability of analytic syst
Jonas Gann, Michael Gertz · 2026-08-06
Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. However, its reasoning process remains largely opaque: intermediate reasoning steps are difficult to verify and cannot be reliably attributed to specific evidence. Moreover, missing user-specific context is rarely detected systematically, often lead
Zahra Khodakarami, Yue Li, Pulkit Khandelwal · 2026-08-06
White matter hyperintensities (WMH), bright regions on Fluid-attenuated Inversion Recovery (FLAIR) scans are associated with cerebrovascular pathology and neurodegeneration. FLAIR is usually acquired with thick slices in clinical settings, giving it poor through-plane resolution. Super-resolution (SR) is a widely used method for recovering an isotropic volume from an anisotropic scan. Yet whether
Jacek Komorowski · 2026-08-06
LiDAR-based Scene Coordinate Regression (SCR) maps point clouds directly to 3D scene coordinates, enabling precise 6-DoF localisation without explicit map retrieval. However, existing methods produce deterministic predictions, discarding aleatoric uncertainty that could improve robustness and downstream decision-making. We present UQ-Loc, which extends the LightLoc architecture with an anisotropic
Arash Nedaei, Henna Tiensuu, Elina V\u00e4yrynen · 2026-08-06
Oral health issues affect billions globally, but the cost and limited access to professional dental care hinder preventive oral healthcare. Research relies on clinical-grade radiographs or intraoral camera images, unavailable for public self-screening. This study introduces a tooth localisation and numbering model for smartphone photographs. We developed a customised Mask Region-based Convolutiona
Robin Trombetta, Carole Lartizien · 2026-08-06
The development of deep learning over the past decade has revolutionized medical imaging segmentation, allowing the extraction of precise descriptors from large volumes to characterize pathologies. Data augmentation is a technique widely regarded as a way to improve model training. It includes simple transformations like spatial operations or intensity modifications, but also more advanced synthes
Ziqi Cai, Siqi Yang, Yimu Wang · 2026-08-06
Current video world models struggle in multiplayer environments because they entangle world state with view-dependent visual latents, leading to redundant compute, view inconsistencies, and poor scalability. We propose MAS (Multiplayer world models with Authoritative Shared State) to resolve this limitation. Inspired by multiplayer game architectures, MAS disentangles world dynamics and view rende
Saad Ahmed, Md Khalid Syfullaha · 2026-08-06
Deaf and hard-of-hearing people in Bangladesh communicate mainly through Bangla Sign Language (BdSL). Automatic BdSL recognition on personal devices could widen access to education and services. Existing systems use controlled-setting datasets without expert verification and heavyweight pretrained backbones unsuited to on-device use. We introduce RSBdSL38, 10,874 expert-validated images spanning a
Daniel Paulin, Victor Elvira · 2026-08-06
Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, the parametrization is identified only up to equivalence, and every evaluation costs a pass over the entire series. Yet their moving-average term captures with a few parameters what a pure autoregression matches only with many lags. We introduce an est
Anay Mehrotra · 2026-08-06
A monotone adversary observes an i.i.d. labeled sample and appends a finite number of further examples of its choice, every one of them labeled correctly by the target hypothesis. The learner sees a uniform shuffle of the combined sample and is scored on the original distribution. Every example is correctly labeled, but the insertions depend on the clean sample, so the combined sample is not excha
Dae-Jin Lee · 2026-08-06
We propose a new unit of analysis for longitudinal data: the Latent Memory Table. The scientific contribution is not the encoder. It is that table, treated as a reusable statistical object on the same footing as a matrix of principal-component scores, a table of estimated random effects, or a table of predicted probabilities. We estimate a statistical table that summarizes recent longitudinal hist
Farzana Nasrin · 2026-08-06
Persistence diagrams (PDs) provide stable and interpretable summaries of multiscale topological structure. While substantial progress has been made in the statistical analysis of PDs, existing literature often treats diagrams as static objects and provide limited frameworks for probabilistic modeling and stochastic evolution on PD space. We introduce a reinforcement learning framework for stochast
This digest is generated automatically from arXiv submissions. Not affiliated with arXiv or Cornell University.