Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 2,604 results for author: Rahul

  1. arXiv:2609.39099  [pdf, ps, other] 

    cs.LG cs.AI

    A Generalisation Signal Need Not Be a Model-Selection Signal

    Authors: Aditya Nagarsekar, M P Ashish Bhat, Aadi Nesarkar, Vrishti Godhwani, Rahul Yedida, Aditya Challa, Danda Sravan, Snehanshu Saha

    Abstract: Model selection in computational biology often relies on validation data drawn from the training regime, even when deployment lies outside it. When validation no longer preserves which model is best, a natural alternative is to rank candidates using properties of the trained network itself. We test this idea using a novel, forward-only proxy motivated by the norm of the Hessian, alongside common H… ▽ More

    Submitted 30 September, 2026; originally announced September 2026.

    Comments: Accepted (poster) at the NeurIPS 2026 Workshop "I Can't Believe It's Not Better: Failure Modes of AI in Biology" (ICBINB-BIO). 24 pages, 4 figures, 22 tables. Code: https://github.com/AdityaNagarsekar/A-Generalisation-Signal-Need-Not-Be-a-Model-Selection-Signal

  2. arXiv:2609.38276  [pdf, ps, other] 

    cs.LG eess.SP stat.AP

    Kinematic signatures of impairment: Detecting alcohol intoxication in e-scooter riders using sensor data and machine learning

    Authors: Rahul Rajendra Pai, Marco Dozza, Alexander Rasch, Ali Mohammadi, Marco Capuccini

    Abstract: Alcohol intoxication is a leading contributor to fatal and severe-injured e-scooterist crashes. Current countermeasures, such as temporal restrictions or pre-ride cognitive screening, cannot continuously assess an e-scooterist's physical motor control or impairment in real time. We conducted a controlled experiment in which 25 participants rode an instrumented e-scooter through a test track while… ▽ More

    Submitted 29 September, 2026; originally announced September 2026.

  3. arXiv:2609.38176  [pdf, ps, other] 

    cs.LG cond-mat.dis-nn cond-mat.stat-mech

    Breakdown of Local Denoising as Semantic Speciation

    Authors: Guangkuo Liu, Mert Okyay, Yifan F. Zhang, Fangjun Hu, Rahul Nandkishore, Xun Gao

    Abstract: The dynamics of generative models exhibit two apparently distinct temporal windows: a speciation window, in which a sample commits to a semantic class, and a nonlocality window, in which local context windows become insufficient for generation. Motivated by evidence of their near-concurrence in a variety of frontier models, we investigate their relationship through the spatial distribution of sema… ▽ More

    Submitted 29 September, 2026; originally announced September 2026.

    Comments: 9 pages main, 13 pages appendix, 4 figures. Comments very welcome

  4. arXiv:2609.37995  [pdf, ps, other] 

    cs.NI cs.CE

    Pricing IoT Data Delivered via LEO Satellites

    Authors: Rahul Ramachandran, Suman Banerjee

    Abstract: IoT terminals served by LEO satellite constellations transmit data to passing satellites in discrete uplink windows. That data is delivered to buyers only when the satellite reaches a ground station. This store-and-forward structure makes the achievable price a discontinuous function of the delivery completeness threshold (a phenomenon we call a pricing cliff) and creates a geographic pricing asym… ▽ More

    Submitted 29 September, 2026; originally announced September 2026.

  5. arXiv:2609.37800  [pdf, ps, other] 

    cs.LG cs.AI

    Challenges and Solutions for Bandits in the Wild: Warm-Started Mixture Bandits for Cross-Cohort Slate Recommendation

    Authors: Serafima Lebedeva, Sumantrak Mukherjee, Ali Arshad Sadal, Ilias Ekşi, Rahul Sharma, Julia Mueller, Theresa Dombrowski, Jakob Karolus, Viktor Bengs, Eyke Hüllermeier, Sebastian Vollmer

    Abstract: Many recommender services repeatedly encounter cold-start cohorts, where new users arrive with little or no interaction history. This creates two challenges: learning user preferences quickly from limited feedback and sustaining useful recommendations when each user has a finite catalog that can become repetitive or depleted over time. We propose CohortMix-TS, a warm-started mixture bandit that le… ▽ More

    Submitted 29 September, 2026; originally announced September 2026.

    Comments: 11 pages, 3 figures, preprint

  6. arXiv:2609.36323  [pdf, ps, other] 

    cs.AI cs.DB cs.SE

    Towards an AI Software Factory for Data Systems

    Authors: Anna Pavlenko, Bogdan Crivat, Brandon Haynes, Carlo Curino, Fotis Psallidas, Jaro Slawinski, Johannes Freischuetz, Laura Pereira Sanchez, Markus Weimer, Mathieu Demarne, Matthias Jasny, Mauktik Gandhi, Max Bovykin, Mirco Milletari, Purbasha Ghosh, Qiushi Bai, Raghu Ramakrishnan, Rahul Pandita, Sergiy Matusevich, Shivaram Venkataraman, Subru Krishnan, Md. Tareq Mahmood, Tiemo Bang, Venkatesh Emani, Xuan Zhao , et al. (1 additional authors not shown)

    Abstract: AI-assisted coding tools deliver significant acceleration of coding, but only limited impact across the end-to-end software development lifecycle (SDLC)--an Amdahl's law effect! In this paper, we discuss our progress towards building an AI SW Factory that accelerates all the stages of SDLC-Targeting, Coding, Reviewing, and Ops. The AI SW Factory produces a metadata exhaust that enables self-impr… ▽ More

    Submitted 28 September, 2026; originally announced September 2026.

    Comments: 6 pages, 5 figures, 1 table

    ACM Class: H.2.4; I.2.11; D.2.9

  7. arXiv:2609.36289  [pdf, ps, other] 

    cs.SE cs.AI cs.HC

    How Much Prompt Is Enough? A Blackbox Minimization of Few-Shots in LLMs

    Authors: Ali Alfageeh, Rahul Gopinath, Amin Alipour

    Abstract: Prompts are the primary mechanism for directing the behavior of large language models (LLMs). Yet the internal structure and causal hierarchy of prompts remain poorly understood: which parts are causally necessary and which are redundant is an open question. This opacity can have severe consequences. Subtle prompt variations can silently shift model outputs in critical software systems, and engine… ▽ More

    Submitted 28 September, 2026; originally announced September 2026.

  8. arXiv:2609.36120  [pdf, ps, other] 

    cs.LG cs.AI

    ThinQuant: Scalable Rotation Learning for Weight and Activation Quantization of LLMs

    Authors: Mehdi Makni, Ryan Lucas, Rahul Mazumder

    Abstract: Learned rotations play an important role in enabling low-bit weight and activation quantization of large language models by smoothing outliers in the activation distribution. State-of-the-art approaches include gradient-based procedures such as SpinQuant and computationally friendlier gradient-free approaches such as DartQuant, but both remain hard to scale to the largest architectures. To address… ▽ More

    Submitted 28 September, 2026; originally announced September 2026.

  9. arXiv:2609.34791  [pdf, ps, other] 

    math.OC cs.LG

    Finite-Time Concentration and Convergence Rates for Projected Two-Time-Scale Stochastic Approximation with Markov Noise

    Authors: Rahul Singh, Vivek S. Borkar, Eric Moulines

    Abstract: We study finite-time concentration and convergence rates for projected two-time-scale stochastic approximation driven by a controlled Markov chain. The averaged fast map is contractive, while the slow iterate is projected onto a compact convex polyhedron. The associated projected ordinary differential equation may have a discontinuous vector field at the boundary, preventing a direct application o… ▽ More

    Submitted 28 September, 2026; originally announced September 2026.

  10. arXiv:2609.33895  [pdf, ps, other] 

    cs.CV cs.LG

    Residual-Stream Burden Shapes Representation Learning in Diffusion Transformers

    Authors: Tongtong Liang, Siqi Kou, Ziqiao Xi, Esha Singh, Kun Zhou, Zhijie Deng, Alexander Cloninger, Yu-Xiang Wang, Rahul Parhi

    Abstract: In diffusion-based generation, a neural network can be trained to predict the clean data, the noise, or the velocity from a noisy input. These prediction targets are interconvertible and describe the same generative process, yet plain Diffusion Transformers operating on large pixel patches succeed with clean prediction and fail with noise or velocity prediction. We argue that this asymmetry arises… ▽ More

    Submitted 27 September, 2026; originally announced September 2026.

    Comments: Under review

  11. arXiv:2609.33499  [pdf, ps, other] 

    cs.LG cs.IT math.OC

    The cost of useful natural gradient updates

    Authors: Subhransu S. Bhattacharjee, Dylan Campbell, Rahul Shome

    Abstract: What information is needed to turn a natural-gradient direction into a useful finite update? Under a population Kullback-Leibler (KL) budget, we call a step useful if it is feasible and loses at most a fraction $\varepsilon$ of the best feasible gain along the direction. We construct a four-state exponential family whose laws share their initial gradient, scalar Fisher information and natural grad… ▽ More

    Submitted 27 September, 2026; originally announced September 2026.

    Comments: 43 pages, 8 figures, 10 tables

    MSC Class: 68Q32; 68Q17 ACM Class: I.2.6; F.1.3; G.3

  12. arXiv:2609.31606  [pdf, ps, other] 

    cs.RO

    Learning Robot Policies from Sparse Success Signals via STL-Guided Stein Variational Policy Gradient

    Authors: Hongrui Zheng, Cristian Ioan Vasile, Antonio Loquercio, Rahul Mangharam

    Abstract: Learning robot policies for tasks with sparse success signals is challenging when completion depends on coordinated actions, precise contact outcomes, or satisfying several conditions together. Intricate physical interactions with the world further complicate these requirements. Prior work using conventional reward shaping mechanisms provides dense feedback but local progress might not translate i… ▽ More

    Submitted 25 September, 2026; originally announced September 2026.

  13. arXiv:2609.31245  [pdf, ps, other] 

    cs.CL

    RupeeBias: Auditing Demographic Bias in Indian Economic Guidance from Large Language Models

    Authors: Pavithra P M Nair, Bhavik Talaviya, Shourya Bhushan, Rahul Pankajakshan, Seema Guruvadoo, Avinash Agarwal, Gilad Gressel, Krishnashree Achuthan

    Abstract: Individuals turn to large language models (LLMs) for guidance across a wide range of economic tasks, from comparing loan options and planning savings to deciding what raise to ask for or how much to charge for their services. LLMs are known to reproduce social biases, and biased economic guidance may influence what users believe they are worth, what they ask for, and what they ultimately accept. T… ▽ More

    Submitted 28 September, 2026; v1 submitted 25 September, 2026; originally announced September 2026.

    Comments: Code: https://github.com/lab105/RupeeBias Dataset: https://huggingface.co/datasets/lab-105/RupeeBias

  14. arXiv:2609.30817  [pdf, ps, other] 

    cs.FL math.CO

    Quadratic bounds for uncompletable words and matrix mortality

    Authors: Rahul Chandelkar, Samrath Singh Chadha

    Abstract: Every finite nonempty incomplete uniquely decipherable code with maximum word length $k$ has an uncompletable word of length at most $4k^2-3k$. The bound is independent of the number of codewords and their total length. Deleting a complete codeword cycle gives a finite path-counting identity; Kraft equality then supplies a short word of deficient compressed mass. Cyclic averaging and padding turn… ▽ More

    Submitted 28 September, 2026; v1 submitted 25 September, 2026; originally announced September 2026.

    Comments: Lean formalization and implementation pilot included as ancillary material

  15. arXiv:2609.30571  [pdf, ps, other] 

    cs.AI

    HARDEN: Constrained Evolutionary Search for Harder, Answer-Preserving Evaluation Cases

    Authors: Aditya Kumaran, Rahul Singhal, Karime Maamari, Amine Mhedhbi, Pradyumna Tambwekar

    Abstract: Language models are often evaluated on curated benchmarks that underrepresent the complexity of enterprise deployments. We introduce HARDEN, a constrained evolutionary search method to adapt the input of existing evaluation cases into more challenging variants while keeping their expected outputs fixed. HARDEN searches along generated domain-specific complexity axes while enforcing feasibility con… ▽ More

    Submitted 24 September, 2026; originally announced September 2026.

  16. arXiv:2609.30279  [pdf, ps, other] 

    cs.LG math.AC

    Neural Ideals and Neural Codes: An Algebraic Framework for Neural Network Classification and Feature Interpretation

    Authors: Venkata Subbaiah Yerrapati, Rahul Dixit, Ajay Kumar Shukla

    Abstract: Understanding the features captured by the hidden layers of neural networks is a fundamental challenge in machine learning, despite their widespread success across various classification problems. In this work, we propose an algebraic framework for examining neural networks that model classification problems. Certain results, such as the correspondence between the neural network and neural ideals,… ▽ More

    Submitted 19 August, 2026; originally announced September 2026.

    MSC Class: 13P25; 68T07

  17. arXiv:2609.29952  [pdf, ps, other] 

    cs.AI cs.CL cs.MA

    Augur: A Synthetic Decision Lab for Rehearsing Reactions to Product and Policy Changes

    Authors: Rahul Khedar, Mayank Malhotra, Avinash Karn

    Abstract: Before a product or policy change ships, the question that matters is how people will react to it. Augur rehearses that reaction offline: it builds a typed knowledge graph from the change documents, populates a grounded persona market, simulates the interaction, and returns an auditable decision memo recommending one of five actions. We assemble Gold-50, fifty real product and policy episodes whos… ▽ More

    Submitted 24 September, 2026; originally announced September 2026.

    Comments: 19 pages, 15 figures, 11 tables

  18. arXiv:2609.28854  [pdf, ps, other] 

    cs.CL cs.AI

    Persuaded, Not Informed: Incentive-Misaligned Witnesses Defeat In-Context Grounding

    Authors: Rahul Balakavi

    Abstract: Language-model agents increasingly answer questions over customer-relationship management (CRM) records, such as whether to qualify a sales lead. We identify a failure mode not addressed by a stronger model: when the context contains an assertion by a party with an incentive toward optimism - here the sales representative, a witness recorded in the CRM - the model treats the assertion as evidence… ▽ More

    Submitted 23 September, 2026; originally announced September 2026.

    Comments: 9 pages, 4 figures, IEEE conference format. Ancillary files contain the evaluation harness, pre-specifications, and per-run result files

    ACM Class: I.2.7; H.3.3

  19. arXiv:2609.28029  [pdf, ps, other] 

    math.NA cs.CL cs.LG

    Tensor Decomposition of Transformer Key-Value Caches: Spectral Structure and Format Comparison

    Authors: Rahul Krishnan, Volker Schulz

    Abstract: The key-value (KV) cache of autoregressive transformers can be viewed as a fourth-order tensor spanning attention heads, tokens, features, and grouped layers. We measure the singular-value spectra of all four mode unfoldings on Mistral-7B-v0.3 and LLaMA-2-13B and compare four standard tensor decompositions: Tucker, CP, tensor train, and t-SVD, at matched storage. The spectra partition the four axe… ▽ More

    Submitted 23 September, 2026; originally announced September 2026.

    Comments: 18 pages, 3 figures, 8 tables. Submitted to SIAM Journal on Matrix Analysis and Applications (SIMAX)

    MSC Class: 15A18; 15A69; 65F55; 68T07

  20. arXiv:2609.22132  [pdf, ps, other] 

    cs.NI cs.LG

    WiNeRF: Measurement Constrained Radiance Fields for Actionable Wireless Channel Modeling

    Authors: Saif Ur Rahman, Rafid Umayer Murshed, Anton Dmitriev, Cagri Tanriover, Rahul C. Shah, Elahé Soltanaghai

    Abstract: Wireless embedded systems increasingly rely on wireless channel information for decision making, yet practical platforms operate under severe constraints, including few antennas, narrow bandwidth, and sparse, noisy measurements. While neural field based approaches inspired by Neural Radiance Fields (NeRFs) have recently been explored for continuous wireless channel modeling, existing approaches de… ▽ More

    Submitted 24 August, 2026; originally announced September 2026.

    Comments: Accepted to EWSN 2026

  21. arXiv:2609.19914  [pdf, ps, other] 

    cs.DS cs.DB

    ZigZag Trie: A Novel Index for Contextual Queries

    Authors: Ling Li, Daniel Gibney, Sharma V. Thankachan, Rahul Shah, Grigorios Loukides, Solon P. Pissis

    Abstract: There is increasing interest in queries about the context of a string $P$ in a longer text $T$, i.e., the set of all string pairs $(L,R)$, with $|L|=|R|=q$, for a given $q$, such that the string $LPR$ occurs in $T$. Such contextual queries are important in several domains but are challenging to answer efficiently. This is because the length of $T$ in applications is massive and existing indexes do… ▽ More

    Submitted 17 September, 2026; originally announced September 2026.

    Comments: 29 pages

  22. arXiv:2609.19418  [pdf, ps, other] 

    econ.EM cs.LG math.ST stat.ML

    Stable Policy Learning

    Authors: Harvey Barnhard, Giacomo Opocher, Rahul Singh

    Abstract: In evidence-based policymaking, typically one experimental sample is observed, then a learned policy recommendation is implemented at scale. Policies learned from the experimental data can perform well in expected welfare, yet random sampling in the experiment can produce recommendations with poor welfare outcomes. In this paper, we ask: how should policy learning algorithms balance expected welfa… ▽ More

    Submitted 16 September, 2026; originally announced September 2026.

  23. arXiv:2609.18813  [pdf, ps, other] 

    cs.RO

    Asymptotically Optimal Multi-Robot Task and Motion Planning

    Authors: Thi Thuy Ngan Duong, Cheuk Tung Shadow Yiu, Rahul Shome, Yoonchang Sung

    Abstract: Multi-robot task and motion planning (MR-TAMP) requires jointly reasoning about discrete task decisions and continuous collision-free motions of multiple interacting robots. Although asymptotically optimal algorithms have been developed for task and motion planning, extending these guarantees to the multi-robot setting introduces an important challenge: different task transitions may involve diffe… ▽ More

    Submitted 16 September, 2026; originally announced September 2026.

  24. arXiv:2609.17572  [pdf, ps, other] 

    cs.LG

    Disentangling Algorithmic Bias from Archival Artifacts: A Controlled Audit of Vision-Language Model Valuation in Metropolitan Museum Archives

    Authors: Manpreet Singh, Rhythm Bhatia, Rahul Joshi

    Abstract: Auditing vision-language models (VLMs) for societal bias requires distinguishing direct algorithmic valuation disparities from confounders embedded within archival metadata. In this study, we audit Contrastive Language-Image Pretraining (CLIP) models using historical artwork metadata from the Metropolitan Museum of Art Open Access collection (N = 1,500 total objects; N = 743 attributed works: Male… ▽ More

    Submitted 29 July, 2026; originally announced September 2026.

  25. arXiv:2609.16694  [pdf, ps, other] 

    cs.CR

    Toward Secure AI-Powered Penetration Testing Agents: Security Threats, Guardrails, and Architectural Perspectives

    Authors: Rahul Dev T Y, Hiran V Nath

    Abstract: LLM-powered autonomous agents are transforming the penetration testing space with dynamic, multi-step offensive security workflows that require minimal supervision by humans. These agents leverage sophisticated reasoning abilities and external security tools to independently carry out reconnaissance, identify vulnerabilities, devise exploitation plans, and perform post-exploitation operations. But… ▽ More

    Submitted 15 September, 2026; originally announced September 2026.

  26. arXiv:2609.14450  [pdf, ps, other] 

    cs.RO

    EcoBoat: Design and Experimental Validation of an Autonomous Body-Board Boat For Cleaning Water Bodies

    Authors: M. Aman Ansari, Saifullah Khan, Rahul Kulkarni, PB Sujit

    Abstract: Cleaning water bodies such as swimming pools and lakes typically demands significant manual effort or reliance on costly, sensor-intensive robotic systems. This paper presents EcoBoat, a low-cost autonomous surface vehicle built on a modified hull, designed to collect floating debris in both indoor and outdoor water bodies. For indoor environments, EcoBoat uses ultrasonic sensors to detect boundar… ▽ More

    Submitted 13 September, 2026; originally announced September 2026.

  27. arXiv:2609.13283  [pdf, ps, other] 

    cs.CV cs.AI cs.LG

    Multimodal-Multiresolution Foundation Model for Lunar Remote Sensing

    Authors: Paolo Fraccaro, Gabby Nyirjesy, Daniela Szwarcman, Himanshu Patil, Vishal Gaur, Rohit Lal, Rachel A. Slank, Geoffrey Dawson, Hiyam Debary, Michael K. Barker, Andrew Annex, Vishnu Viswanathan, Zachary Morse, Ethan I. Schaefer, Nikolaos Dionelis, Ankur Kumar, Campbell D. Watson, Manil Maskey, Rebekah I. Dawson-Rigas, Juan Bernabé-Moreno, Rahul Ramachandran, Sujit Roy

    Abstract: We present a multimodal foundation model for lunar remote sensing, pretrained from scratch on SomBench, a geographically partitioned corpus of nearly two million co-registered tile bundles spanning 11 modalities at two spatial scales (1 m/pixel and 100 m/pixel). The model adapts the TerraMind masked-token architecture with two lunar-specific extensions: acquisition geometry is provided as explicit… ▽ More

    Submitted 8 September, 2026; originally announced September 2026.

  28. arXiv:2609.13277  [pdf, ps, other] 

    cs.CV cs.LG

    SomBench: Benchmark Dataset for Advancing Machine Learning in Lunar Science

    Authors: Himanshu Patil, Gabby Nyirjesy, Rachel A. Slank, Vishal Gaur, Daniela Szwarcman, Paolo Fraccaro, Nikolaos Dionelis, Michael K. Barker, Andrew Annex, Vishnu Viswanathan, Zachary Morse, Ethan I. Schaefer, Hiyam Debary, Ankur Kumar, Rohit Lal, Geoffrey Dawson, Campbell Watson, Rebekah I. Dawson-Rigas, Manil Maskey, Juan Bernabé-Moreno, Rahul Ramachandran, Sujit Roy

    Abstract: Lunar orbital missions, such as Lunar Reconnaissance Orbiter, Kaguya/SELENE, Gravity Recovery and Interior Laboratory, and Lunar Prospector, among others, provide rich multi-instrument observations, but their heterogeneity in sampling, projection, and conventions limits reproducible machine learning (ML). We introduce SomBench, a unified, spatially-aligned, ML-ready lunar dataset aggregating 30+ c… ▽ More

    Submitted 7 September, 2026; originally announced September 2026.

  29. arXiv:2609.12239  [pdf, ps, other] 

    cs.DC

    Specifying Paxos for System Builders: Pseudocode Made Executable

    Authors: Yanhong A. Liu, Rahul Sihag

    Abstract: This paper presents a precise executable specification---as a faithful mapping from the pseudocode---of Paxos for System Builders, a practical protocol for replication and consensus in distributed systems. Paxos for System Builders has both a robust implementation in C and a clean pseudocode for critical protocol details. This paper shows how the protocol pseudocode can be expressed easily, esse… ▽ More

    Submitted 10 September, 2026; originally announced September 2026.

    Comments: 41 pages. Extended version of invited paper at the 28th International Symposium on Stabilization, Safety, and Security of Distributed Systems. 2026

  30. arXiv:2609.11207  [pdf, ps, other] 

    cs.LG cs.DS math.OC

    Convex Optimization with Nested Evolving Feasible Sets (CONES) under Time-Varying Loss Functions

    Authors: Rahul Vaze

    Abstract: Convex Optimization with Nested Evolving Feasible Sets (CONES)} was introduced in \cite{CONESVaze} where the objective function \(f\) remains fixed but the feasible region evolves over time as a nested sequence \(S_1 \supseteq S_2 \supseteq \cdots \supseteq S_T\). The goal of an online algorithm is to simultaneously minimize the regret with respect to hindsight static optimal benchmark and the tot… ▽ More

    Submitted 10 September, 2026; originally announced September 2026.

  31. arXiv:2609.09697  [pdf, ps, other] 

    cs.CR

    PrivAudit: A Dual-Lens Auditing Framework for Website Privacy Practices under the CCPA

    Authors: Mohamed Moustafa Dawoud, Riya Aggarwal, Likith Rahul Krishnamurthy, Ram Sundara Raman

    Abstract: Five years after the enforcement of the California Consumer Privacy Act (CCPA), understanding how website privacy practices evolve at scale in response to regulation remains a key challenge for both researchers and regulators. Prior work and regulatory efforts have focused on manual and case-specific enforcement, but there remain no scalable approaches to systematically audit two key user-facing f… ▽ More

    Submitted 9 September, 2026; originally announced September 2026.

    Comments: Extended version of a paper accepted to the 2026 ACM SIGSAC Conference on Computer and Communications Security (CCS '26)

  32. arXiv:2609.09188  [pdf, ps, other] 

    cs.CV

    Lensless Gaze Is Not Private by Default: Auditing Identity Leakage Across Disclosure Surfaces

    Authors: Rahul Vimalkanth, Kaushik Mitra

    Abstract: Lensless near-eye sensing is often described as privacy-friendly because its coded measurements are visually unintelligible. Yet visual unintelligibility reflects human interpretation, not what a learned adversary can recover. We therefore treat identity privacy as a systems property of disclosure surfaces: representations crossing sensing, storage, computation, and output boundaries. We audit a s… ▽ More

    Submitted 30 September, 2026; v1 submitted 31 August, 2026; originally announced September 2026.

    Comments: 16 pages, 5 figures. Code available at https://github.com/xoxo121/Lensless-Gaze-Is-Not-Private-by-Default

  33. arXiv:2609.08592  [pdf, ps, other] 

    cs.AI cs.CL

    A Three-Tier Persona Vector for Controllable User Simulation in Agentic Evaluation

    Authors: Rahul Khedar, Eshita, Sneha Teja Sree Reddy Thondapu, Mayank Malhotra, Arup Kumar Das, Jitesh Chandra Mishra, Arun Menon, Avinash Karn, Mouli V

    Abstract: Evaluating tool-augmented LLM agents requires diverse, realistic user inputs yet most evaluation frameworks use flat role descriptions ("you are an angry customer") that produce near-identical conversations regardless of the underlying scenario. In this paper, we propose a three-tier persona vector with 23 operationalized dimensions: 6 categorical demographics (jurisdiction, age, channel, device,… ▽ More

    Submitted 8 September, 2026; originally announced September 2026.

    Comments: 7 pages, 3 figures, 6 tables. Extended treatment of the persona component of StateGen (arXiv:2606.16307)

  34. arXiv:2609.06849  [pdf, ps, other] 

    cs.AI

    Learning transferable human physiology from two million hours of sleep with SleepFM-2

    Authors: Rahul Thapa, Christopher Sun, William Theodor Lehn-Schioler, Sophia Claire Kivelson, Umaer Hanif, Hyatt Moore IV, Harrison G. Zhang, Hafsa Ahmed, Marcus Dige, Niels R. Lorenzen, Elisabeth Roxane M. Heremans, Adrien Specht, Ulysse Gimenez, Robin Guillard, Andreas Brink-Kjaer, James Zou, Emmanuel Mignot

    Abstract: Sleep provides a nightly window into health by capturing coordinated activity across the brain, heart, muscles and respiratory system. We introduce SleepFM-2, a sleep foundation model developed and evaluated on 282,511 polysomnography recordings from 26 cohorts, including 235,865 used for pretraining. These data span more than two million hours of multimodal physiology. Compared with SleepFM, Slee… ▽ More

    Submitted 6 September, 2026; originally announced September 2026.

  35. arXiv:2609.05334  [pdf, ps, other] 

    cs.CV cs.AI cs.LG

    Lightweight Vision Transformer Compression for On-Device Plant Disease Detection in Resource-Constrained Agricultural Field Conditions

    Authors: Mahadev Sunil Kumar, Bhavika Gondi, Desaisetty Venkata Satya Sai Swapnith, Gangireddy Rahul Jogi, Sudheesh Manalil, Arnab Raha, Amitava Mukherjee, Parthasarathy Seethapathy, G. Gopakumar

    Abstract: Chilli (Capsicum annuum) is one of India's most economically significant crops, yet its productivity is persistently threatened by diseases that are difficult to identify without expert intervention. While Vision Transformers (ViTs) have achieved high classification accuracy, their large computational footprint makes deployment on resource constrained devices challenging. Existing compression appr… ▽ More

    Submitted 4 September, 2026; originally announced September 2026.

  36. arXiv:2609.03413  [pdf, ps, other] 

    cs.CY

    The 5P Reflection Model for Education in the Generative Artificial Intelligence (GenAI) Era

    Authors: Rajan Kadel, Samar Shailendra, Islam Mohammad Tahidul, Urvashi Rahul Saxena, Aakanksha Sharma, Sabitra Kaphle

    Abstract: Contributions: A reflection model suitable for the era of Generative Artificial Intelligence (GenAI) is introduced. The proposed model is an integrated model that extracts features from various existing models and also incorporates technological aspects of GenAI. Background: Universities worldwide are facing challenges in adopting GenAI into their curricula, as it has impacted academic integrity… ▽ More

    Submitted 3 September, 2026; originally announced September 2026.

  37. arXiv:2609.03221  [pdf, ps, other] 

    cs.CL cs.AI cs.CY cs.LG stat.AP

    Instability Floors: Separating Bias from Noise in Fairness Audits of Clinical LLM Agents with FairMedAgent

    Authors: Rohith Reddy Bellibatlu, Manpreet Singh, Deepak Parashar, Rahul Joshi

    Abstract: Counterfactual fairness audits of clinical language-model agents report a flip rate: how often an action changes when only the patient's demographic descriptor changes. Part of that rate is not demographic. A stochastic agent also changes its own action when nothing changes, and a flip rate cannot be interpreted without knowing how often. We measured it. Re-running one condition ten times over six… ▽ More

    Submitted 28 September, 2026; v1 submitted 2 September, 2026; originally announced September 2026.

    Comments: 27 pages (13 main plus 14 supplementary), 4 figures, 3 tables. Code: https://github.com/rohithreddybc/FairMedAgent (v0.1.5, commit 3982974; concept DOI 10.5281/zenodo.22165979). Trajectories: https://huggingface.co/datasets/Rohithreddybc/FairMedAgent

  38. arXiv:2609.02736  [pdf, ps, other] 

    cond-mat.mes-hall cs.DC quant-ph

    QArray+: A physics-informed GPU-accelerated simulator for quantum dot arrays

    Authors: Pranav Vaidhyanathan, Barnaby van Straaten, Alice Petrillo, Rahul Marchand, Edwin De Nicolo, Menno Veldhorst, Brucek Khailany, Taylor L. Patti, Natalia Ares

    Abstract: Semiconductor quantum-dot arrays are a compelling platform for scalable quantum technologies, yet their practical operation is hindered by the complexity of tuning large-scale devices. Existing automation tools rely on simplified physical models---such as constant-capacitance approximations and equilibrium Hubbard models---which assume instantaneous relaxation to a steady state. These frameworks f… ▽ More

    Submitted 2 September, 2026; originally announced September 2026.

    Comments: P.V and B.v.S contributed equally to this work. 19 pages, 13 figures

  39. arXiv:2609.02730  [pdf, ps, other] 

    cs.CL

    CORAL: An LLM-Native Harness for Production Recommender Systems

    Authors: Muhammad Rafay Azhar, Yuhang Zhou, Gilbert Jiang, Yuchen Wang, Rahul Sharma, Matthew DeSousa, Jiayi Liu, Xin Guo, Lizhu Zhang, Xiangjun Fan

    Abstract: Production recommender systems shape what billions of people see, and sustaining their performance requires continual optimization: as content, user behavior, and upstream models shift, the choices governing retrieval, ranking, and serving must be revisited. Traditionally, human engineers test such changes through online experiments--a slow, reactive process limited by engineering effort, leaving… ▽ More

    Submitted 2 September, 2026; originally announced September 2026.

    Comments: Accepted by RecSys '26 OARS Workshop

  40. arXiv:2609.01861  [pdf, ps, other] 

    cs.AI

    Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization

    Authors: Yuhan Chen, Zhihua Tian, Mahavir Dabas, Charith Peris, Rahul Gupta, Ming Jin, Feiyang Kang, Siyuan Zhang, Nan Wang, Ruoxi Jia

    Abstract: The performance of an LLM agent depends on the scaffold around a frozen model. A common way to improve that scaffold is to use a coding agent as an optimizer: it reads current scores and traces and iteratively edits the source, producing a new candidate each round. Each edit is chosen according to a belief about how the environment will respond: what went wrong, and which change should help. That… ▽ More

    Submitted 1 September, 2026; originally announced September 2026.

  41. arXiv:2609.01792  [pdf, ps, other] 

    cs.SD cs.CV

    Efficient Passive Acoustic Monitoring of Killer Whales Using a Two-Stage Detection and Ecotype Classification Cascade

    Authors: Daniela Ruiz, Manuel Castellote, Zhongqi Miao, Carl Chalmers, Bruno Demuro, Rahul Dodhia, Pablo Arbelaez, Juan M. Lavista

    Abstract: Passive acoustic monitoring of killer whales is particularly important for conservation of the endangered Southern Resident killer whale population, but requires accurate models that can operate in real time under severe class imbalance and deployment shift. We propose a lightweight ResNet-based two-stage cascade that first detects killer whale vocalizations and then classifies confident detection… ▽ More

    Submitted 1 September, 2026; originally announced September 2026.

  42. arXiv:2609.01616  [pdf, ps, other] 

    cs.IR

    Incident Memory: Training-Free Operational Memory through Sequential Pattern Mining and Velocity-Stratified Retrieval

    Authors: Adarsh Agrawal, Rahul Suresh Babu

    Abstract: Incident response is a memory problem: teams accumulate tickets, traces, postmortems, and wiki pages, but the knowledge needed for the next incident is rarely stored with its order, freshness, and provenance intact. We present Incident Memory, a deterministic system that accumulates operational knowledge without model training. It combines (i) velocity-stratified retrieval, which ages structural,… ▽ More

    Submitted 29 June, 2026; originally announced September 2026.

    Comments: 14 pages, 5 figures, 8 tables (main text); includes appendix

    ACM Class: I.2.7; H.3.3; D.2.5

  43. arXiv:2609.00898  [pdf, ps, other] 

    cs.CV cs.AI cs.LG

    Vision-Language-Guided Pseudo-Labels for Unsupervised Domain Adaptation in Semantic Segmentation for Waste Sorting

    Authors: Udo Schlegel, Shubhangi, Gabriel Dax, Sai Rahul Kaminwar, Florian Karl, Thomas Seidl

    Abstract: Obtaining labeled data for semantic segmentation in applied settings (e.g., autonomous driving, industrial waste sorting) is expensive and often infeasible at scale. We present a cross-modal pseudo-labeling pipeline that enables unsupervised domain adaptation without any target-domain annotations. The pipeline is built on two core foundation models: SAM generates class-agnostic region proposals, a… ▽ More

    Submitted 1 September, 2026; originally announced September 2026.

    Comments: 18 pages, 2 figures, 2 tables, accepted at ECML-PKDD 2026

  44. arXiv:2609.00441  [pdf, ps, other] 

    cs.AI

    Conversation Coach: A Voice-enabled AI System that Helps Practice Difficult Workplace Conversations

    Authors: Fanyou Wu, Suraj Maharjan, Ainur Yessenalina, Dennis Xu Chen, Rahul Srivastava, Srinivasan H. Sengamedu

    Abstract: Effective manager-employee communication is critical for retaining high performers and developing underperformers, yet training managers in these skills remains costly. Text-based chatbots offer a scalable approach but cannot provide realistic rehearsal: managers need to practice speaking aloud to build confidence before high-stakes conversations. In this paper, we propose Conversation Coach, a vo… ▽ More

    Submitted 31 August, 2026; originally announced September 2026.

    Journal ref: EMNLP 2026 Industry Track

  45. arXiv:2608.29567  [pdf, ps, other] 

    cs.CV

    MotionSync: Non-Causal Refinement of Causal Tracker for Label-Efficient 3D Perception

    Authors: Rahul Ahuja, Bala Murali Manoghar Sai Sudhakar, Shashwata Gupta, Venkatraman Narayanan, Varun Ravi Kumar, Senthil Yogamani

    Abstract: Three-dimensional box-and-track annotation is the cost bottleneck in autonomous-driving data engines, and the offline systems built to relieve it replace the online perception stack outright, so a team needing both regimes maintains and reconciles two. MotionSync makes the causal/non-causal boundary an explicit architectural seam instead. A strictly causal tracker, built on a strong published base… ▽ More

    Submitted 30 August, 2026; originally announced August 2026.

  46. arXiv:2608.28802  [pdf, ps, other] 

    cs.CV cs.AI

    A Large-scale Evaluation of Text-guided Models for Facial Editing

    Authors: Rahul Nair, Saurav Pandit, Hannah Kerner

    Abstract: Facial appearance editing powers popular applications like FaceApp and Photoshop. Generative Adversarial Networks (GANs) and 3D Morphable Models (3DMMs) have been widely used for facial editing. GANs can perform varied facial edits (e.g., changing hair color, hairstyle), but often produce unstable edits. 3DMMs produce stable edits, but can only alter pose and facial expression. Recently, text-guid… ▽ More

    Submitted 28 August, 2026; originally announced August 2026.

    Comments: ACM Multimedia (ACMMM) 2026 Oral

  47. arXiv:2608.28656  [pdf, ps, other] 

    cs.RO cs.CV

    RedLight-VLA: Models for traffic-rule grounding and behavioral emphasis in driving policies

    Authors: Bala Murali Manoghar Sai Sudhakar, Sourab Bapu Sridhar, Sandipan Das, Rahul Ahuja, Meda Lazar, Ashish Garg, Pratik Likhar, Senthil Yogamani

    Abstract: Behavior-cloned Vision-Language-Action (VLA) driving policies struggle with rare rule-governed maneuvers at signalized intersections. Braking and launching examples contribute little to averaged trajectory loss, while fused representations lack explicit supervision for the governing traffic-light and stop-line state. We present RedLight-VLA, a training objective that uses expert futures and automa… ▽ More

    Submitted 20 August, 2026; originally announced August 2026.

  48. arXiv:2608.28594  [pdf, ps, other] 

    cs.AI

    From Question-First to Analyst-First: Domain-Expert Skills and Verified Knowledge Compilation for Proactive Enterprise Analytics

    Authors: Harmohit Singh, Rahul Sharma

    Abstract: Conversational analytics systems assume the user already has a well-formed question, leaving a non-expert facing a blank query box on an unfamiliar enterprise schema. Commercial 'proactive' tools narrow this gap only by detecting statistical anomalies over analyst-curated metric layers, and academic next-question recommenders depend on query logs that a fresh dataset lacks. We describe a productio… ▽ More

    Submitted 15 June, 2026; originally announced August 2026.

    Comments: 21 pages, 4 figures, 2 tables

  49. arXiv:2608.25275  [pdf, ps, other] 

    cs.AI cs.RO

    PhaseShift: Topology-Aware Data Harmonization and Model Consolidation Across Signalized Intersections

    Authors: Yash Ranjan, Artur Kumik, Rahul Sengupta, Anand Rangarajan, Sanjay Ranka

    Abstract: Learned traffic-behavior models are commonly trained separately for each intersection, creating model portfolios that cannot share evidence across sites. We present PhaseShift, a topology-aware framework that harmonizes heterogeneous roadside trajectories into a shared actor-centric representation and trains one reusable backbone. Ego-relative coordinates, trajectory-induced movement paths, normal… ▽ More

    Submitted 25 August, 2026; originally announced August 2026.

  50. arXiv:2608.24974  [pdf, ps, other] 

    cs.LG cs.AI eess.SP

    Clearing the Underbrush: AI-Enhanced RF Interference Suppression

    Authors: Rahul Jain, Pierre Trepagnier, Rick Gentile, Joey Botero, Alexia Schulz

    Abstract: AI-based structured interference rejection has grown more popular because deep learning approaches can outperform traditional methods by jointly considering the signal of interest (SOI) and the signal mixture (SOI plus interference). This work builds on a previous AI-enabled approach utilizing autoregressive transformer-based models by adding a Finite Scalar Quantization (FSQ) tokenizer layer whic… ▽ More

    Submitted 25 August, 2026; originally announced August 2026.

    Comments: 7 pages, 10 figures, Accepted to the 2026 IEEE Military Communications Conference (MILCOM)