Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 59 results for author: Shubhangi

  1. arXiv:2609.00898  [pdf, ps, other] 

    cs.CV cs.AI cs.LG

    Vision-Language-Guided Pseudo-Labels for Unsupervised Domain Adaptation in Semantic Segmentation for Waste Sorting

    Authors: Udo Schlegel, Shubhangi, Gabriel Dax, Sai Rahul Kaminwar, Florian Karl, Thomas Seidl

    Abstract: Obtaining labeled data for semantic segmentation in applied settings (e.g., autonomous driving, industrial waste sorting) is expensive and often infeasible at scale. We present a cross-modal pseudo-labeling pipeline that enables unsupervised domain adaptation without any target-domain annotations. The pipeline is built on two core foundation models: SAM generates class-agnostic region proposals, a… ▽ More

    Submitted 1 September, 2026; originally announced September 2026.

    Comments: 18 pages, 2 figures, 2 tables, accepted at ECML-PKDD 2026

  2. arXiv:2607.15554  [pdf, ps, other] 

    cs.CC cs.DS math.CO

    On the CGGRT Criterion for Detecting Bipartite Perfect Matchings in NC

    Authors: Swastik Kopparty, Shubhangi Saraf

    Abstract: The recent breakthrough work of Chatterjee, Ghosh, Gurjar, Raj and Thierauf [CGGRT26] gives the first deterministic NC algorithm for the bipartite matching problem. They show how to detect as well as find perfect matchings in bipartite graphs in NC. In this note we present an arguably simpler-to-state variation of the NC detection criterion of [CGGRT26], with improved parameters.

    Submitted 16 July, 2026; originally announced July 2026.

    Comments: 8 pages

  3. arXiv:2606.27293  [pdf, ps, other] 

    cs.CC

    Deterministic Algorithms for Low Individual Degree Factors of Sparse Polynomials

    Authors: Somnath Bhattacharjee, Rishabh Kothary, Shanthanu S. Rai, Shubhangi Saraf

    Abstract: We study factoring algorithms for general sparse polynomials and sparse polynomials of bounded individual degree and prove the following results. 1. We give a deterministic polynomial-time algorithm which takes as input an $n$-variate $s$-sparse polynomial $f$ of bounded individual degree $d$ and outputs a list of circuits which contains all factors of $f$, although there might be additional spu… ▽ More

    Submitted 25 June, 2026; originally announced June 2026.

  4. arXiv:2603.05829  [pdf, ps, other] 

    cs.LG cs.CL

    Test-Time Adaptation via Many-Shot Prompting: Benefits, Limits, and Pitfalls

    Authors: Shubhangi Upasani, Chen Wu, Jay Rainton, Bo Li, Urmish Thakker, Changran Hu, Qizheng Zhang

    Abstract: Test-time adaptation enables large language models (LLMs) to modify their behavior at inference without updating model parameters. A common approach is many-shot prompting, where large numbers of in-context learning (ICL) examples are injected as an input-space test-time update. Although performance can improve as more demonstrations are added, the reliability and limits of this update mechanism r… ▽ More

    Submitted 17 March, 2026; v1 submitted 5 March, 2026; originally announced March 2026.

  5. arXiv:2603.02631  [pdf, ps, other] 

    cs.CL

    Cross-Family Speculative Prefill: Training-Free Long-Context Compression with Small Draft Models

    Authors: Shubhangi Upasani, Ravi Shanker Raju, Bo Li, Mengmeng Ji, John Long, Chen Wu, Urmish Thakker, Guangtao Wang

    Abstract: Prompt length is a major bottleneck in agentic large language model (LLM) workloads, where repeated inference steps and multi-call loops incur substantial prefill cost. Recent work on speculative prefill demonstrates that attention-based token importance estimation can enable training-free prompt compression, but this assumes the existence of a draft model that shares the same tokenizer as the tar… ▽ More

    Submitted 12 March, 2026; v1 submitted 3 March, 2026; originally announced March 2026.

    Journal ref: ICLR 2026 (WS)

  6. arXiv:2602.16069  [pdf, ps, other] 

    cs.SE cs.LG

    The Limits of Long-Context Reasoning in Automated Bug Fixing

    Authors: Ravi Raju, Mengmeng Ji, Shubhangi Upasani, Bo Li, Urmish Thakker

    Abstract: Rapidly increasing context lengths have led to the assumption that large language models (LLMs) can directly reason over entire codebases. Concurrently, recent advances in LLMs have enabled strong performance on software engineering benchmarks, particularly when paired with agentic workflows. In this work, we systematically evaluate whether current LLMs can reliably perform long-context code debug… ▽ More

    Submitted 6 March, 2026; v1 submitted 17 February, 2026; originally announced February 2026.

    Comments: Accepted to ICLR 2026 ICBINB workshop

  7. arXiv:2602.01312  [pdf, ps, other] 

    cs.LG

    Imperfect Influence, Preserved Rankings: A Theory of TRAK for Data Attribution

    Authors: Han Tong, Shubhangi Ghosh, Haolin Zou, Arian Maleki

    Abstract: Data attribution, tracing a model's prediction back to specific training data, is an important tool for interpreting sophisticated AI models. The widely used TRAK algorithm addresses this challenge by first approximating the underlying model with a kernel machine and then leveraging techniques developed for approximating the leave-one-out (ALO) risk. Despite its strong empirical performance, the t… ▽ More

    Submitted 1 February, 2026; originally announced February 2026.

  8. arXiv:2511.08650  [pdf, ps, other] 

    cs.LG

    A Lightweight CNN-Attention-BiLSTM Architecture for Multi-Class Arrhythmia Classification on Standard and Wearable ECGs

    Authors: Vamsikrishna Thota, Hardik Prajapati, Yuvraj Joshi, Shubhangi Rathi

    Abstract: Early and accurate detection of cardiac arrhythmias is vital for timely diagnosis and intervention. We propose a lightweight deep learning model combining 1D Convolutional Neural Networks (CNN), attention mechanisms, and Bidirectional Long Short-Term Memory (BiLSTM) for classifying arrhythmias from both 12-lead and single-lead ECGs. Evaluated on the CPSC 2018 dataset, the model addresses class imb… ▽ More

    Submitted 11 November, 2025; originally announced November 2025.

  9. arXiv:2511.03092  [pdf, ps, other] 

    cs.AI cs.AR cs.DC

    SnapStream: Efficient Long Sequence Decoding on Dataflow Accelerators

    Authors: Jonathan Li, Nasim Farahini, Evgenii Iuliugin, Magnus Vesterlund, Christian Häggström, Guangtao Wang, Shubhangi Upasani, Ayush Sachdeva, Rui Li, Faline Fu, Chen Wu, Ayesha Siddiqua, John Long, Tuowen Zhao, Matheen Musaddiq, Håkan Zeffer, Yun Du, Mingran Wang, Qinghua Li, Bo Li, Urmish Thakker, Raghu Prabhakar

    Abstract: The proliferation of 100B+ parameter Large Language Models (LLMs) with 100k+ context length support have resulted in increasing demands for on-chip memory to support large KV caches. Techniques such as StreamingLLM and SnapKV demonstrate how to control KV cache size while maintaining model accuracy. Yet, these techniques are not commonly used within industrial deployments using frameworks like vLL… ▽ More

    Submitted 8 April, 2026; v1 submitted 4 November, 2025; originally announced November 2025.

  10. arXiv:2510.27535  [pdf] 

    cs.CL

    Patient-Centered Summarization Framework for AI Clinical Summarization: A Mixed-Methods Design

    Authors: Maria Lizarazo Jimenez, Ana Gabriela Claros, Kieran Green, David Toro-Tobon, Felipe Larios, Sheena Asthana, Camila Wenczenovicz, Kerly Guevara Maldonado, Luis Vilatuna-Andrango, Cristina Proano-Velez, Satya Sai Sri Bandi, Shubhangi Bagewadi, Megan E. Branda, Misk Al Zahidy, Saturnino Luz, Mirella Lapata, Juan P. Brito, Oscar J. Ponce-Ponte

    Abstract: Large Language Models (LLMs) are increasingly demonstrating the potential to reach human-level performance in generating clinical summaries from patient-clinician conversations. However, these summaries often focus on patients' biology rather than their preferences, values, wishes, and concerns. To achieve patient-centered care, we propose a new standard for Artificial Intelligence (AI) clinical s… ▽ More

    Submitted 31 October, 2025; originally announced October 2025.

    Comments: The first two listed authors contributed equally Pages: 21; Figures:2; Tables:3

  11. arXiv:2510.16481  [pdf, ps, other] 

    math.CO cs.CG

    Integer points in dilates of polytopes

    Authors: Shubhangi Saraf, Narmada Varadarajan

    Abstract: In this paper we study how the number of integer points in a polytope grows as we dilate the polytope. We prove new and essentially tight bounds on this quantity by specifically studying dilates of the Hadamard polytope. Our motivation for studying this quantity comes from the problem of understanding the maximal number of monomials in a factor of a multivariate polynomial with $s$ monomials. A… ▽ More

    Submitted 18 October, 2025; originally announced October 2025.

  12. arXiv:2510.04618  [pdf, ps, other] 

    cs.LG cs.AI cs.CL

    Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models

    Authors: Qizheng Zhang, Changran Hu, Shubhangi Upasani, Boyuan Ma, Fenglu Hong, Vamsidhar Kamanuru, Jay Rainton, Chen Wu, Mengmeng Ji, Hanchen Li, Urmish Thakker, James Zou, Kunle Olukotun

    Abstract: Large language model (LLM) applications such as agents and domain-specific reasoning increasingly rely on context adaptation: modifying inputs with instructions, strategies, or evidence, rather than weight updates. Prior approaches improve usability but often suffer from brevity bias, which drops domain insights for concise summaries, and from context collapse, where iterative rewriting erodes det… ▽ More

    Submitted 29 March, 2026; v1 submitted 6 October, 2025; originally announced October 2025.

    Comments: ICLR 2026; 32 pages

  13. arXiv:2506.23220  [pdf, ps, other] 

    cs.CC

    Constant-depth circuits for polynomial GCD over any characteristic

    Authors: Somnath Bhattacharjee, Mrinal Kumar, Shanthanu Rai, Varun Ramanathan, Ramprasad Saptharishi, Shubhangi Saraf

    Abstract: We show that the GCD of two univariate polynomials can be computed by (piece-wise) algebraic circuits of constant depth and polynomial size over any sufficiently large field, regardless of the characteristic. This extends a recent result of Andrews & Wigderson who showed such an upper bound over fields of zero or large characteristic. Our proofs are based on a recent work of Bhattacharjee, Kumar… ▽ More

    Submitted 29 June, 2025; originally announced June 2025.

  14. arXiv:2506.23214  [pdf, ps, other] 

    cs.CC

    Closure under factorization from a result of Furstenberg

    Authors: Somnath Bhattacharjee, Mrinal Kumar, Shanthanu S. Rai, Varun Ramanathan, Ramprasad Saptharishi, Shubhangi Saraf

    Abstract: We show that algebraic formulas and constant-depth circuits are closed under taking factors. In other words, we show that if a multivariate polynomial over a field of characteristic zero has a small constant-depth circuit or formula, then all its factors can be computed by small constant-depth circuits or formulas respectively. Our result turns out to be an elementary consequence of a fundamenta… ▽ More

    Submitted 29 June, 2025; originally announced June 2025.

  15. arXiv:2505.21495  [pdf, ps, other] 

    cs.RO

    CLAMP: Crowdsourcing a LArge-scale in-the-wild haptic dataset with an open-source device for Multimodal robot Perception

    Authors: Pranav N. Thakkar, Shubhangi Sinha, Karan Baijal, Yuhan, Bian, Leah Lackey, Ben Dodson, Heisen Kong, Jueun Kwon, Amber Li, Yifei Hu, Alexios Rekoutis, Tom Silver, Tapomayukh Bhattacharjee

    Abstract: Robust robot manipulation in unstructured environments often requires understanding object properties that extend beyond geometry, such as material or compliance-properties that can be challenging to infer using vision alone. Multimodal haptic sensing provides a promising avenue for inferring such properties, yet progress has been constrained by the lack of large, diverse, and realistic haptic dat… ▽ More

    Submitted 12 January, 2026; v1 submitted 27 May, 2025; originally announced May 2025.

  16. arXiv:2505.13173  [pdf, ps, other] 

    cs.CL

    A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs

    Authors: V. S. D. S. Mahesh Akavarapu, Hrishikesh Terdalkar, Pramit Bhattacharyya, Shubhangi Agarwal, Vishakha Deulgaonkar, Pralay Manna, Chaitali Dangarikar, Arnab Bhattacharya

    Abstract: Large Language Models (LLMs) have demonstrated remarkable generalization capabilities across diverse tasks and languages. In this study, we focus on natural language understanding in three classical languages -- Sanskrit, Ancient Greek and Latin -- to investigate the factors affecting cross-lingual zero-shot generalization. First, we explore named entity recognition and machine translation into En… ▽ More

    Submitted 31 May, 2025; v1 submitted 19 May, 2025; originally announced May 2025.

    Comments: Accepted to ACL 2025 Findings

    ACM Class: I.2.7

  17. arXiv:2504.08063  [pdf, ps, other] 

    cs.CC cs.DS

    Deterministic factorization of constant-depth algebraic circuits in subexponential time

    Authors: Somnath Bhattacharjee, Mrinal Kumar, Varun Ramanathan, Ramprasad Saptharishi, Shubhangi Saraf

    Abstract: While efficient randomized algorithms for factorization of polynomials given by algebraic circuits have been known for decades, obtaining an even slightly non-trivial deterministic algorithm for this problem has remained an open question of great interest. This is true even when the input algebraic circuit has additional structure, for instance, when it is a constant-depth circuit. Indeed, no effi… ▽ More

    Submitted 14 June, 2025; v1 submitted 10 April, 2025; originally announced April 2025.

    Comments: Some changes in Section 8.1 to reflect a subtle issue regarding the underlying field in the reduction from irreducibility to divisibility

  18. arXiv:2503.08879  [pdf, other] 

    cs.CL cs.AI cs.LG

    LLMs Know What to Drop: Self-Attention Guided KV Cache Eviction for Efficient Long-Context Inference

    Authors: Guangtao Wang, Shubhangi Upasani, Chen Wu, Darshan Gandhi, Jonathan Li, Changran Hu, Bo Li, Urmish Thakker

    Abstract: Efficient long-context inference is critical as large language models (LLMs) adopt context windows of ranging from 128K to 1M tokens. However, the growing key-value (KV) cache and the high computational complexity of attention create significant bottlenecks in memory usage and latency. In this paper, we find that attention in diverse long-context tasks exhibits sparsity, and LLMs implicitly "know"… ▽ More

    Submitted 11 March, 2025; originally announced March 2025.

  19. arXiv:2410.10017  [pdf, other] 

    cs.RO cs.CV cs.GR

    REPeat: A Real2Sim2Real Approach for Pre-acquisition of Soft Food Items in Robot-assisted Feeding

    Authors: Nayoung Ha, Ruolin Ye, Ziang Liu, Shubhangi Sinha, Tapomayukh Bhattacharjee

    Abstract: The paper presents REPeat, a Real2Sim2Real framework designed to enhance bite acquisition in robot-assisted feeding for soft foods. It uses `pre-acquisition actions' such as pushing, cutting, and flipping to improve the success rate of bite acquisition actions such as skewering, scooping, and twirling. If the data-driven model predicts low success for direct bite acquisition, the system initiates… ▽ More

    Submitted 13 October, 2024; originally announced October 2024.

  20. arXiv:2401.09472  [pdf, other] 

    cs.CV eess.IV eess.SY

    Plug-in for visualizing 3D tool tracking from videos of Minimally Invasive Surgeries

    Authors: Shubhangi Nema, Abhishek Mathur, Leena Vachhani

    Abstract: This paper tackles instrument tracking and 3D visualization challenges in minimally invasive surgery (MIS), crucial for computer-assisted interventions. Conventional and robot-assisted MIS encounter issues with limited 2D camera projections and minimal hardware integration. The objective is to track and visualize the entire surgical instrument, including shaft and metallic clasper, enabling safe n… ▽ More

    Submitted 12 January, 2024; originally announced January 2024.

  21. arXiv:2312.15874  [pdf, ps, other] 

    cs.CC

    Lower Bounds for Set-Multilinear Branching Programs

    Authors: Prerona Chatterjee, Deepanshu Kush, Shubhangi Saraf, Amir Shpilka

    Abstract: In this paper, we prove super-polynomial lower bounds for the model of \emph{sum of ordered set-multilinear algebraic branching programs}, each with a possibly different ordering ($\sum \mathsf{smABP}$). Specifically, we give an explicit $nd$-variate polynomial of degree $d$ such that any $\sum \mathsf{smABP}$ computing it must have size $n^{ω(1)}$ for $d$ as low as $ω(\log n)$. Notably, this cons… ▽ More

    Submitted 19 February, 2024; v1 submitted 25 December, 2023; originally announced December 2023.

  22. arXiv:2312.13438  [pdf, ps, other] 

    stat.ML cs.LG

    Independent Mechanism Analysis and the Manifold Hypothesis

    Authors: Shubhangi Ghosh, Luigi Gresele, Julius von Kügelgen, Michel Besserve, Bernhard Schölkopf

    Abstract: Independent Mechanism Analysis (IMA) seeks to address non-identifiability in nonlinear Independent Component Analysis (ICA) by assuming that the Jacobian of the mixing function has orthogonal columns. As typical in ICA, previous work focused on the case with an equal number of latent components and observed mixtures. Here, we extend IMA to settings with a larger number of mixtures that reside on a… ▽ More

    Submitted 20 December, 2023; originally announced December 2023.

    Comments: 6 pages, Accepted at Neurips Causal Representation Learning 2023

  23. arXiv:2310.19149  [pdf, ps, other] 

    cs.CC cs.DM math.CO

    Simple Constructions of Unique Neighbor Expanders from Error-correcting Codes

    Authors: Swastik Kopparty, Noga Ron-Zewi, Shubhangi Saraf

    Abstract: In this note, we give very simple constructions of unique neighbor expander graphs starting from spectral or combinatorial expander graphs of mild expansion. These constructions and their analysis are simple variants of the constructions of LDPC error-correcting codes from expanders, given by Sipser-Spielman [SS96] (and Tanner [Tan81]), and their analysis. We also show how to obtain expanders with… ▽ More

    Submitted 25 January, 2024; v1 submitted 29 October, 2023; originally announced October 2023.

    Comments: Updated introduction, corrected minor typos

  24. arXiv:2309.02949  [pdf, other] 

    cs.NI cs.ET

    Evaluation of NR-Sidelink for Cooperative Industrial AGVs

    Authors: Shubhangi Bhadauria, Klea Plaku, Yash Deshpande, Wolfgang Kellerer

    Abstract: Industry 4.0 has brought to attention the need for a connected, flexible, and autonomous production environment. The New Radio (NR)-sidelink, which was introduced by the third-generation partnership project (3GPP) in Release 16, can be particularly helpful for factories that need to facilitate cooperative and close-range communication. Automated Guided Vehicles (AGVs) are important for material ha… ▽ More

    Submitted 6 September, 2023; originally announced September 2023.

  25. arXiv:2307.07453  [pdf, other] 

    physics.flu-dyn cs.CE

    Investigation of Deep Learning-Based Filtered Density Function for Large Eddy Simulation of Turbulent Scalar Mixing

    Authors: Shubhangi Bansude, Reza Sheikhi

    Abstract: A filtered density function (FDF) model based on deep neural network (DNN), termed DNN-FDF, is introduced for large eddy simulation (LES) of turbulent flows involving conserved scalar transport. The primary objectives of this study are to develop the DNN-FDF models and evaluate their predictive capability in accounting for various filtered moments, including that of non-linear source terms. A syst… ▽ More

    Submitted 29 September, 2023; v1 submitted 14 July, 2023; originally announced July 2023.

  26. arXiv:2302.00332  [pdf, other] 

    cs.AI cs.CV

    iPAL: A Machine Learning Based Smart Healthcare Framework For Automatic Diagnosis Of Attention Deficit/Hyperactivity Disorder (ADHD)

    Authors: Abhishek Sharma, Arpit Jain, Shubhangi Sharma, Ashutosh Gupta, Prateek Jain, Saraju P. Mohanty

    Abstract: ADHD is a prevalent disorder among the younger population. Standard evaluation techniques currently use evaluation forms, interviews with the patient, and more. However, its symptoms are similar to those of many other disorders like depression, conduct disorder, and oppositional defiant disorder, and these current diagnosis techniques are not very effective. Thus, a sophisticated computing model h… ▽ More

    Submitted 1 February, 2023; originally announced February 2023.

  27. arXiv:2207.06137  [pdf, other] 

    stat.ML cs.AI cs.LG

    Probing the Robustness of Independent Mechanism Analysis for Representation Learning

    Authors: Joanna Sliwa, Shubhangi Ghosh, Vincent Stimper, Luigi Gresele, Bernhard Schölkopf

    Abstract: One aim of representation learning is to recover the original latent code that generated the data, a task which requires additional information or inductive biases. A recently proposed approach termed Independent Mechanism Analysis (IMA) postulates that each latent source should influence the observed mixtures independently, complementing standard nonlinear independent component analysis, and taki… ▽ More

    Submitted 13 July, 2022; originally announced July 2022.

    Comments: 10 pages, 14 figures, UAI CRL 2022 final camera-ready version

  28. arXiv:2205.00611  [pdf, ps, other] 

    cs.CC

    Improved Low-Depth Set-Multilinear Circuit Lower Bounds

    Authors: Deepanshu Kush, Shubhangi Saraf

    Abstract: We prove strengthened lower bounds for constant-depth set-multilinear formulas. More precisely, we show that over any field, there is an explicit polynomial $f$ in VNP defined over $n^2$ variables, and of degree $n$, such that any product-depth $Δ$ set-multilinear formula computing $f$ has size at least $n^{Ω\left( n^{1/Δ}/Δ\right)}$. The hard polynomial $f$ comes from the class of Nisan-Wigderson… ▽ More

    Submitted 1 May, 2022; originally announced May 2022.

    Comments: 14 pages, To appear in Computational Complexity Conference (CCC) 2022

  29. arXiv:2202.06844  [pdf, other] 

    stat.ML cs.AI cs.LG

    On Pitfalls of Identifiability in Unsupervised Learning. A Note on: "Desiderata for Representation Learning: A Causal Perspective"

    Authors: Shubhangi Ghosh, Luigi Gresele, Julius von Kügelgen, Michel Besserve, Bernhard Schölkopf

    Abstract: Model identifiability is a desirable property in the context of unsupervised representation learning. In absence thereof, different models may be observationally indistinguishable while yielding representations that are nontrivially related to one another, thus making the recovery of a ground truth generative model fundamentally impossible, as often shown through suitably constructed counterexampl… ▽ More

    Submitted 14 February, 2022; originally announced February 2022.

    Comments: 5 pages, 1 figure

  30. VerSaChI: Finding Statistically Significant Subgraph Matches using Chebyshev's Inequality

    Authors: Shubhangi Agarwal, Sourav Dutta, Arnab Bhattacharya

    Abstract: Approximate subgraph matching, which is an important primitive for many applications like question answering, community detection, and motif discovery, often involves large labeled graphs such as knowledge graphs, social networks, and protein sequences. Effective methods for extracting matching subgraphs, in terms of label and structural similarities to a query, should depict accuracy, computation… ▽ More

    Submitted 18 August, 2021; originally announced August 2021.

  31. arXiv:2105.01751  [pdf, ps, other] 

    cs.CC cs.LG

    Reconstruction Algorithms for Low-Rank Tensors and Depth-3 Multilinear Circuits

    Authors: Vishwas Bhargava, Shubhangi Saraf, Ilya Volkovich

    Abstract: We give new and efficient black-box reconstruction algorithms for some classes of depth-$3$ arithmetic circuits. As a consequence, we obtain the first efficient algorithm for computing the tensor rank and for finding the optimal tensor decomposition as a sum of rank-one tensors when then input is a constant-rank tensor. More specifically, we provide efficient learning algorithms that run in random… ▽ More

    Submitted 4 May, 2021; originally announced May 2021.

    MSC Class: 68Q25; 68Q32; 68Q06; 68W30; 68W40; 14N07 ACM Class: F.2.0; I.2.6

  32. V2X in 3GPP Standardization: NR Sidelink in Rel-16 and Beyond

    Authors: Mehdi Harounabadi, Dariush Mohammad Soleymani, Shubhangi Bhadauria, Martin Leyh, Elke Roth-Mandutz

    Abstract: The 5G mobile network brings several new features that can be applied to existing and new applications. High reliability, low latency, and high data rate are some of the features which fulfill the requirements of vehicular networks. Vehicular networks aim to provide safety for road users and several additional advantages such as enhanced traffic efficiency and in-vehicle infotainment services. Thi… ▽ More

    Submitted 22 April, 2021; originally announced April 2021.

  33. arXiv:2010.14570  [pdf, other] 

    cs.IR cs.LG

    Addressing Purchase-Impression Gap through a Sequential Re-ranker

    Authors: Shubhangi Tandon, Saratchandra Indrakanti, Amit Jaiswal, Svetlana Strunjas, Manojkumar Rangasamy Kannadasan

    Abstract: Large scale eCommerce platforms such as eBay carry a wide variety of inventory and provide several buying choices to online shoppers. It is critical for eCommerce search engines to showcase in the top results the variety and selection of inventory available, specifically in the context of the various buying intents that may be associated with a search query. Search rankers are most commonly powere… ▽ More

    Submitted 27 October, 2020; originally announced October 2020.

  34. arXiv:2008.09657  [pdf, other] 

    cs.SI cs.LG stat.ML

    GraphReach: Position-Aware Graph Neural Network using Reachability Estimations

    Authors: Sunil Nishad, Shubhangi Agarwal, Arnab Bhattacharya, Sayan Ranu

    Abstract: Majority of the existing graph neural networks (GNN) learn node embeddings that encode their local neighborhoods but not their positions. Consequently, two nodes that are vastly distant but located in similar local neighborhoods map to similar embeddings in those networks. This limitation prevents accurate performance in predictive tasks that rely on position information. In this paper, we develop… ▽ More

    Submitted 20 August, 2021; v1 submitted 19 August, 2020; originally announced August 2020.

    Journal ref: IJCAI 2021

  35. arXiv:2004.13875  [pdf, other] 

    cs.IT eess.SP

    6G White Paper on Machine Learning in Wireless Communication Networks

    Authors: Samad Ali, Walid Saad, Nandana Rajatheva, Kapseok Chang, Daniel Steinbach, Benjamin Sliwa, Christian Wietfeld, Kai Mei, Hamid Shiri, Hans-Jürgen Zepernick, Thi My Chinh Chu, Ijaz Ahmad, Jyrki Huusko, Jaakko Suutala, Shubhangi Bhadauria, Vimal Bhatia, Rangeet Mitra, Saidhiraj Amuru, Robert Abbas, Baohua Shao, Michele Capobianco, Guanghui Yu, Maelick Claes, Teemu Karvonen, Mingzhe Chen , et al. (2 additional authors not shown)

    Abstract: The focus of this white paper is on machine learning (ML) in wireless communications. 6G wireless communication networks will be the backbone of the digital transformation of societies by providing ubiquitous, reliable, and near-instant wireless connectivity for humans and machines. Recent advances in ML research has led enable a wide range of novel technologies such as self-driving vehicles and v… ▽ More

    Submitted 28 April, 2020; originally announced April 2020.

  36. arXiv:1908.03825  [pdf, other] 

    cs.IR cs.LG

    Influence of Neighborhood on the Preference of an Item in eCommerce Search

    Authors: Saratchandra Indrakanti, Svetlana Strunjas, Shubhangi Tandon, Manojkumar Rangasamy Kannadasan

    Abstract: Surfacing a ranked list of items for a search query to help buyers discover inventory and make purchase decisions is a critical problem in eCommerce search. Typically, items are independently predicted with a probability of sale with respect to a given search query. But in a dynamic marketplace like eBay, even for a single product, there are various different factors distinguishing one item from a… ▽ More

    Submitted 17 October, 2019; v1 submitted 10 August, 2019; originally announced August 2019.

  37. arXiv:1903.12243  [pdf, ps, other] 

    cs.CC cs.CR cs.IT

    DEEP-FRI: Sampling outside the box improves soundness

    Authors: Eli Ben-Sasson, Lior Goldberg, Swastik Kopparty, Shubhangi Saraf

    Abstract: Motivated by the quest for scalable and succinct zero knowledge arguments, we revisit worst-case-to-average-case reductions for linear spaces, raised by [Rothblum, Vadhan, Wigderson, STOC 2013]. We first show a sharp quantitative form of a theorem which says that if an affine space $U$ is $δ$-far in relative Hamming distance from a linear code $V$ - this is the worst-case assumption - then most el… ▽ More

    Submitted 28 March, 2019; originally announced March 2019.

    Comments: 36 pages

  38. arXiv:1809.01331  [pdf, other] 

    cs.CL

    Neural MultiVoice Models for Expressing Novel Personalities in Dialog

    Authors: Shereen Oraby, Lena Reed, Sharath TS, Shubhangi Tandon, Marilyn Walker

    Abstract: Natural language generators for task-oriented dialog should be able to vary the style of the output utterance while still effectively realizing the system dialog actions and their associated semantics. While the use of neural generation for training the response generation component of conversational agents promises to simplify the process of producing high quality responses in new domains, to our… ▽ More

    Submitted 5 September, 2018; originally announced September 2018.

    Comments: Interspeech 2018

  39. arXiv:1808.06655  [pdf, ps, other] 

    math.AC cs.CC

    Deterministic Factorization of Sparse Polynomials with Bounded Individual Degree

    Authors: Vishwas Bhargava, Shubhangi Saraf, Ilya Volkovich

    Abstract: In this paper we study the problem of deterministic factorization of sparse polynomials. We show that if $f \in \mathbb{F}[x_{1},x_{2},\ldots ,x_{n}]$ is a polynomial with $s$ monomials, with individual degrees of its variables bounded by $d$, then $f$ can be deterministically factored in time $s^{\mathrm{poly}(d) \log n}$. Prior to our work, the only efficient factoring algorithms known for this… ▽ More

    Submitted 20 August, 2018; originally announced August 2018.

    MSC Class: 13P05; 12Y05 ACM Class: F.2.1

  40. Arithmetic Circuits with Locally Low Algebraic Rank

    Authors: Mrinal Kumar, Shubhangi Saraf

    Abstract: In recent years, there has been a flurry of activity towards proving lower bounds for homogeneous depth-4 arithmetic circuits, which has brought us very close to statements that are known to imply $\textsf{VP} \neq \textsf{VNP}$. It is open if these techniques can go beyond homogeneity, and in this paper we make some progress in this direction by considering depth-4 circuits of low algebraic rank,… ▽ More

    Submitted 15 June, 2018; originally announced June 2018.

    MSC Class: 68Q15; 68Q17 ACM Class: F.1.3

    Journal ref: Theory of Computing 13(6):1-33, 2017

  41. arXiv:1805.08352  [pdf, other] 

    cs.CL

    Controlling Personality-Based Stylistic Variation with Neural Natural Language Generators

    Authors: Shereen Oraby, Lena Reed, Shubhangi Tandon, T. S. Sharath, Stephanie Lukin, Marilyn Walker

    Abstract: Natural language generators for task-oriented dialogue must effectively realize system dialogue actions and their associated semantics. In many applications, it is also desirable for generators to control the style of an utterance. To date, work on task-oriented neural generation has primarily focused on semantic fidelity rather than achieving stylistic goals, while work on style has been done in… ▽ More

    Submitted 21 May, 2018; originally announced May 2018.

    Comments: To appear at SIGDIAL 2018

  42. arXiv:1805.01498  [pdf, ps, other] 

    cs.IT

    Improved decoding of Folded Reed-Solomon and Multiplicity Codes

    Authors: Swastik Kopparty, Noga Ron-Zewi, Shubhangi Saraf, Mary Wootters

    Abstract: In this work, we show new and improved error-correcting properties of folded Reed-Solomon codes and multiplicity codes. Both of these families of codes are based on polynomials over finite fields, and both have been the sources of recent advances in coding theory. Folded Reed-Solomon codes were the first explicit constructions of codes known to achieve list-decoding capacity; multivariate multipli… ▽ More

    Submitted 3 May, 2018; originally announced May 2018.

  43. arXiv:1711.00092  [pdf, ps, other] 

    cs.CL

    Summarizing Dialogic Arguments from Social Media

    Authors: Amita Misra, Shereen Oraby, Shubhangi Tandon, Sharath TS, Pranav Anand, Marilyn Walker

    Abstract: Online argumentative dialog is a rich source of information on popular beliefs and opinions that could be useful to companies as well as governmental or public policy agencies. Compact, easy to read, summaries of these dialogues would thus be highly valuable. A priori, it is not even clear what form such a summary should take. Previous work on summarization has primarily focused on summarizing wri… ▽ More

    Submitted 31 October, 2017; originally announced November 2017.

    Comments: Proceedings of the 21th Workshop on the Semantics and Pragmatics of Dialogue (SemDial 2017)

  44. arXiv:1710.10520  [pdf, other] 

    cs.CL

    A Dual Encoder Sequence to Sequence Model for Open-Domain Dialogue Modeling

    Authors: Sharath T. S., Shubhangi Tandon, Ryan Bauer

    Abstract: Ever since the successful application of sequence to sequence learning for neural machine translation systems, interest has surged in its applicability towards language generation in other problem domains. Recent work has investigated the use of these neural architectures towards modeling open-domain conversational dialogue, where it has been found that although these models are capable of learnin… ▽ More

    Submitted 28 October, 2017; originally announced October 2017.

  45. arXiv:1710.10498  [pdf, other] 

    cs.CL cs.IR

    Topic Based Sentiment Analysis Using Deep Learning

    Authors: Sharath T. S., Shubhangi Tandon

    Abstract: In this paper , we tackle Sentiment Analysis conditioned on a Topic in Twitter data using Deep Learning . We propose a 2-tier approach : In the first phase we create our own Word Embeddings and see that they do perform better than state-of-the-art embeddings when used with standard classifiers. We then perform inference on these embeddings to learn more about a word with respect to all the topics… ▽ More

    Submitted 28 October, 2017; originally announced October 2017.

  46. arXiv:1701.01717  [pdf, ps, other] 

    cs.CC math.AG

    Towards an algebraic natural proofs barrier via polynomial identity testing

    Authors: Joshua A. Grochow, Mrinal Kumar, Michael Saks, Shubhangi Saraf

    Abstract: We observe that a certain kind of algebraic proof - which covers essentially all known algebraic circuit lower bounds to date - cannot be used to prove lower bounds against VP if and only if what we call succinct hitting sets exist for VP. This is analogous to the Razborov-Rudich natural proofs barrier in Boolean circuit complexity, in that we rule out a large class of lower bound techniques under… ▽ More

    Submitted 6 January, 2017; originally announced January 2017.

    MSC Class: 68Q15; 68Q17; 68W30; 14Q20 ACM Class: F.1.3; F.2.2

  47. On the number of ordinary lines determined by sets in complex space

    Authors: Abdul Basit, Zeev Dvir, Shubhangi Saraf, Charles Wolf

    Abstract: Kelly's theorem states that a set of $n$ points affinely spanning $\mathbb{C}^3$ must determine at least one ordinary complex line (a line passing through exactly two of the points). Our main theorem shows that such sets determine at least $3n/2$ ordinary lines, unless the configuration has $n-1$ points in a plane and one point outside the plane (in which case there are at least $n-1$ ordinary lin… ▽ More

    Submitted 10 November, 2021; v1 submitted 26 November, 2016; originally announced November 2016.

    Comments: Appeared in Discrete Comput. Geom. This version corrects some errors from the previous version, and clarifies the analysis

    Journal ref: Discrete Comput Geom 61, 778-808 (2019)

  48. arXiv:1605.05412  [pdf, other] 

    cs.IT

    Maximally Recoverable Codes for Grid-like Topologies

    Authors: Parikshit Gopalan, Guangda Hu, Swastik Kopparty, Shubhangi Saraf, Carol Wang, Sergey Yekhanin

    Abstract: The explosion in the volumes of data being stored online has resulted in distributed storage systems transitioning to erasure coding based schemes. Yet, the codes being deployed in practice are fairly short. In this work, we address what we view as the main coding theoretic barrier to deploying longer codes in storage: at large lengths, failures are not independent and correlated failures are inev… ▽ More

    Submitted 20 September, 2016; v1 submitted 17 May, 2016; originally announced May 2016.

  49. arXiv:1504.06213  [pdf, ps, other] 

    cs.CC

    Sums of products of polynomials in few variables : lower bounds and polynomial identity testing

    Authors: Mrinal Kumar, Shubhangi Saraf

    Abstract: We study the complexity of representing polynomials as a sum of products of polynomials in few variables. More precisely, we study representations of the form $$P = \sum_{i = 1}^T \prod_{j = 1}^d Q_{ij}$$ such that each $Q_{ij}$ is an arbitrary polynomial that depends on at most $s$ variables. We prove the following results. 1. Over fields of characteristic zero, for every constant $μ$ such that… ▽ More

    Submitted 23 April, 2015; originally announced April 2015.

  50. arXiv:1504.05653  [pdf, ps, other] 

    cs.CC

    High rate locally-correctable and locally-testable codes with sub-polynomial query complexity

    Authors: Swastik Kopparty, Or Meir, Noga Ron-Zewi, Shubhangi Saraf

    Abstract: In this work, we construct the first locally-correctable codes (LCCs), and locally-testable codes (LTCs) with constant rate, constant relative distance, and sub-polynomial query complexity. Specifically, we show that there exist binary LCCs and LTCs with block length $n$, constant rate (which can even be taken arbitrarily close to 1), constant relative distance, and query complexity… ▽ More

    Submitted 22 April, 2015; originally announced April 2015.