Skip to main content
arXiv is now an independent nonprofit! Learn more

Showing 1–50 of 60 results for author: Saifullah

  1. arXiv:2609.14450  [pdf, ps, other] 

    cs.RO

    EcoBoat: Design and Experimental Validation of an Autonomous Body-Board Boat For Cleaning Water Bodies

    Authors: M. Aman Ansari, Saifullah Khan, Rahul Kulkarni, PB Sujit

    Abstract: Cleaning water bodies such as swimming pools and lakes typically demands significant manual effort or reliance on costly, sensor-intensive robotic systems. This paper presents EcoBoat, a low-cost autonomous surface vehicle built on a modified hull, designed to collect floating debris in both indoor and outdoor water bodies. For indoor environments, EcoBoat uses ultrasonic sensors to detect boundar… ▽ More

    Submitted 13 September, 2026; originally announced September 2026.

  2. arXiv:2609.00002  [pdf, ps, other] 

    cs.AI

    HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models

    Authors: Yun-Jian Zhang, Chen-Wei Liang, Tian-Yi Zhang, Jian Ding, Yi-Lun Wu, Ao-Bo Li, Wei-Cong Su, Saifullah, Hong-Yu An, Mu-Jiang-Shan Wang

    Abstract: World models enable language-model agents to predict environment dynamics and plan before acting. In text environments, the model must learn symbolic action effects from serialized state descriptions, but the role of serialization structure remains underexplored. We present HyperWorld, a controlled study of state serialization for learned textual world models. We compare raw observations with thre… ▽ More

    Submitted 12 June, 2026; originally announced September 2026.

    Comments: 10 pages, 4 figures, 3 tables

  3. arXiv:2608.28633  [pdf, ps, other] 

    cs.CL cs.AI cs.CY

    PAUSE: Editable Strategy Artifacts for Long-Form Cultural Story Adaptation

    Authors: Taaha Kazi, Vasu Sharma, Mohammad Saifullah, Abdur Rahman

    Abstract: Generative AI systems increasingly mediate cultural adaptation, but their cultural decisions are often hidden inside prompts, transient model plans, or final prose. We study PAUSE (Pause-And-Update Strategy Editing), an intervention that exposes an editable adaptation strategy as a human control surface for cultural decisions in long-form story adaptation. The strategy is a structured artifact tha… ▽ More

    Submitted 7 August, 2026; originally announced August 2026.

    Comments: 6 pages, 1 figure, 3 tables. Accepted at the 1st Workshop on Culture x AI: Evaluating AI as a Cultural Technology, ICML 2026. Project page: https://abdur75648.github.io/pause/

  4. arXiv:2608.26648   

    cs.CV cs.LG

    Hierarchical Channel Stacking: A Structured Decision Framework for AI-Generated Image Detection

    Authors: Saifullah Shoaib, Akash Borigi, Rupendra Lekkala, Amaury Lendasse, Edward Ratner, Sai Sowjanya Bhamidipati, Alexander Schlager, Peggy Lindner

    Abstract: Many synthetic-image detectors produce accurate predictions but offer limited insight into how those decisions are formed. This paper introduces Hierarchical Channel Stacking (HCS), a compact framework for AI-generated image detection that converts intermediate CNN activations into a structured 60-dimensional representation organized across three progressively deeper backbone stages. HCS uses per-… ▽ More

    Submitted 2 September, 2026; v1 submitted 27 August, 2026; originally announced August 2026.

    Comments: Withdrawn due to insufficient consent approval

  5. arXiv:2607.22597  [pdf, ps, other] 

    cs.AI

    HyCE-RAG: Hypergraph Chain-of-Evidence Retrieval-Augmented Generation for Explainable Multi-hop Question Answering

    Authors: Hong-Yu An, Yun-Jian Zhang, Chen-Wei Liang, Tian-Yi Zhang, Jian Ding, Yi-Lun Wu, Ao-Bo Li, Wei-Cong Su, Saifullah, Mujiangshan Wang

    Abstract: Multi-hop question answering requires systems to retrieve evidence from multiple documents and connect scattered facts into a coherent reasoning process. Standard retrieval-augmented generation (RAG) mainly relies on semantic similarity between a query and text chunks, and therefore often fails to model structural relations among entities, facts, and evidence units. Graph-based RAG improves this b… ▽ More

    Submitted 12 June, 2026; originally announced July 2026.

    Comments: 15 pages, 3 figures, 4 tables

    MSC Class: 68T50; 68T30; 68P20 ACM Class: I.2.7; H.3.3; I.2.4

  6. arXiv:2607.05577  [pdf, ps, other] 

    cs.AI cs.CL cs.IR

    Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction

    Authors: Mohammad Saifullah, Thomas Kornmaier, Taaha Kazi, Vasu Sharma, Aditya Sanjiv Kanade, Aanand Kumar Yadav

    Abstract: Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether an event preceded the narration that revealed it, whether a setup paid off, and how a relationship shifted. General-purpose retrieval and agent-memory systems represent entities and facts but not the narratological structure these questions turn on… ▽ More

    Submitted 6 July, 2026; originally announced July 2026.

    Comments: 23 pages, 4 figures; 9-page main text plus appendix. Preprint

  7. arXiv:2606.17391  [pdf, ps, other] 

    cs.CL cs.AI cs.LG

    NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama

    Authors: Logan Mann, Abdur Rahman, Mohammad Saifullah, Taaha Kazi, Vasu Sharma

    Abstract: Long-form serialized audio drama, with arcs that run for 200 to 800 episodes, is a major creative medium and a setting where frontier large language models (LLMs) fail. We benchmark 21 models, spanning classical, fine-tuned, open-frontier, closed-frontier, and reasoning tiers, on a uniform set of structural narrative metrics. All closed-frontier systems saturate at a plot-beat F1 in the band [0.78… ▽ More

    Submitted 15 June, 2026; originally announced June 2026.

    Comments: 10 pages. Accepted to the ICML 2026 Workshops on High-dimensional Learning Dynamics (HiLD) and Culture x AI

  8. arXiv:2605.01323  [pdf, ps, other] 

    cs.CL cs.AI

    SiNFluD: Creating and Evaluating Figurative Language Dataset for Sindhi

    Authors: Wazir Ali, Adeeb Noor, Saifullah Tumrani

    Abstract: In this article, we introduce SiNFluD, a novel benchmark dataset for Sindhi figurative language classification. We first collect raw text from various blogs, social media platforms, and literary sources, and subsequently prepare the corpus for annotation. Two native annotators label the data using the Doccano text annotation tool, achieving an inter-annotator agreement of 0.81. We then establish b… ▽ More

    Submitted 9 May, 2026; v1 submitted 2 May, 2026; originally announced May 2026.

  9. arXiv:2602.21824  [pdf, ps, other] 

    cs.LG

    DocDjinn: Controllable Synthetic Document Generation with VLMs and Handwriting Diffusion

    Authors: Marcel Lamott, Saifullah Saifullah, Nauman Riaz, Yves-Noel Weweler, Tobias Alt-Veit, Ahmad Sarmad Ali, Muhammad Armaghan Shakir, Adrian Kalwa, Momina Moetesum, Andreas Dengel, Sheraz Ahmed, Faisal Shafait, Ulrich Schwanecke, Adrian Ulges

    Abstract: Effective document intelligence models rely on large amounts of annotated training data. However, procuring sufficient and high-quality data poses significant challenges due to the labor-intensive and costly nature of data acquisition. Additionally, leveraging language models to annotate real documents raises concerns about data privacy. Synthetic document generation has emerged as a promising, pr… ▽ More

    Submitted 25 February, 2026; originally announced February 2026.

  10. arXiv:2601.08166  [pdf, ps, other] 

    cs.AI

    ZeroDVFS: Zero-Shot LLM-Guided Core and Frequency Allocation for Embedded Platforms

    Authors: Mohammad Pivezhandi, Mahdi Banisharif, Abusayeed Saifullah, Ali Jannesari

    Abstract: Dynamic voltage and frequency scaling (DVFS) and task-to-core allocation are critical for thermal management and balancing energy and performance in embedded systems. Existing approaches either rely on utilization-based heuristics that overlook stall times, or require extensive offline profiling for table generation, preventing runtime adaptation. Building upon hierarchical multi-agent scheduling,… ▽ More

    Submitted 28 February, 2026; v1 submitted 12 January, 2026; originally announced January 2026.

    Comments: 56 pages, 14 figures, 18 tables (including appendix)

  11. arXiv:2601.06425  [pdf, ps, other] 

    cs.DC cs.AI

    HiDVFS: Hierarchical Multi-Agent DVFS for Real-Time OpenMP DAG Workloads

    Authors: Mohammad Pivezhandi, Abusayeed Saifullah, Ali Jannesari

    Abstract: Leakage power in multicore embedded systems now rivals dynamic power, so DVFS schedulers must respect deadlines and thermal limits, not just average makespan. Existing heuristics lack per-core, temperature-aware control and overlook the irregular execution of OpenMP DAGs. We propose HiDVFS, a general, extensible hierarchical multi-agent DVFS scheduler: a profiler agent selects cores and frequencie… ▽ More

    Submitted 8 July, 2026; v1 submitted 9 January, 2026; originally announced January 2026.

    Comments: 52 pages, 24 figures, 32 tables (supplement included as appendices). Under review at IEEE TPDS. v2: fairness-corrected GearDVFS baseline with fair-port study, multi-seed same-window re-measurement and full number audit, 12-benchmark BOTS evaluation, real-time evaluation (feasibility gate, conformal shield, mixed-criticality). Data: https://doi.org/10.5281/zenodo.21212162

    ACM Class: I.2.6; C.3; D.4.1

  12. arXiv:2512.12091  [pdf, ps, other] 

    cs.LG

    GraphPerf-RT: A Graph-Driven Performance Model for Hardware-Aware Scheduling of OpenMP Codes

    Authors: Mohammad Pivezhandi, Mahdi Banisharif, Saeed Bakhshan, Abusayeed Saifullah, Ali Jannesari

    Abstract: Autonomous AI agents on embedded platforms require real-time, risk-aware scheduling under resource and thermal constraints. Classical heuristics struggle with workload irregularity, tabular regressors discard structural information, and model-free reinforcement learning (RL) risks overheating. We introduce GraphPerf-RT, a graph neural network surrogate achieving deep learning accuracy at heuristic… ▽ More

    Submitted 21 January, 2026; v1 submitted 12 December, 2025; originally announced December 2025.

    Comments: 49 pages, 4 figures, 19 tables

    ACM Class: C.4; D.1.3; I.2.6

  13. arXiv:2511.17656  [pdf, ps, other] 

    cs.MA cs.LG cs.RO

    Multi-Agent Coordination in Autonomous Vehicle Routing: A Simulation-Based Study of Communication, Memory, and Routing Loops

    Authors: KM Khalid Saifullah, Daniel Palmer

    Abstract: Multi-agent coordination is critical for next-generation autonomous vehicle (AV) systems, yet naive implementations of communication-based rerouting can lead to catastrophic performance degradation. This study investigates a fundamental problem in decentralized multi-agent navigation: routing loops, where vehicles without persistent obstacle memory become trapped in cycles of inefficient path reca… ▽ More

    Submitted 20 November, 2025; originally announced November 2025.

  14. arXiv:2509.02563  [pdf, ps, other] 

    cs.LG cs.CL

    DynaGuard: A Dynamic Guardian Model With User-Defined Policies

    Authors: Monte Hoover, Vatsal Baherwani, Neel Jain, Khalid Saifullah, Joseph Vincent, Chirag Jain, Melissa Kazemi Rad, C. Bayan Bruss, Ashwinee Panda, Tom Goldstein

    Abstract: Guardian models play a crucial role in ensuring the safety and ethical behavior of user-facing AI applications by enforcing guardrails and detecting harmful content. While standard guardian models are limited to predefined, static harm categories, we introduce DynaGuard, a suite of dynamic guardian models offering novel flexibility by evaluating text based on user-defined policies, and DynaBench,… ▽ More

    Submitted 6 October, 2025; v1 submitted 2 September, 2025; originally announced September 2025.

    Comments: 22 Pages

  15. arXiv:2508.04233  [pdf, ps, other] 

    cs.CV

    DocVCE: Diffusion-based Visual Counterfactual Explanations for Document Image Classification

    Authors: Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: As black-box AI-driven decision-making systems become increasingly widespread in modern document processing workflows, improving their transparency and reliability has become critical, especially in high-stakes applications where biases or spurious correlations in decision-making could lead to serious consequences. One vital component often found in such document processing workflows is document i… ▽ More

    Submitted 6 August, 2025; originally announced August 2025.

  16. arXiv:2508.04208  [pdf, ps, other] 

    cs.CR

    DP-DocLDM: Differentially Private Document Image Generation using Latent Diffusion Models

    Authors: Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: As deep learning-based, data-driven information extraction systems become increasingly integrated into modern document processing workflows, one primary concern is the risk of malicious leakage of sensitive private data from these systems. While some recent works have explored Differential Privacy (DP) to mitigate these privacy risks, DP-based training is known to cause significant performance deg… ▽ More

    Submitted 6 August, 2025; originally announced August 2025.

    Comments: Accepted in ICDAR 2025

  17. arXiv:2506.16409  [pdf, ps, other] 

    cs.NI

    LoRaIN: A Constructive Interference-Assisted Reliable and Energy-Efficient LoRa Indoor Network

    Authors: Mahbubur Rahman, Abusayeed Saifullah

    Abstract: LoRa is a promising communication technology for enabling the next-generation indoor Internet of Things applications. Very few studies, however, have analyzed its performance indoors. Besides, these indoor studies investigate mostly the RSSI and SNR of the received packets at the gateway, which, as we show, may not unfold the poor performance of LoRa and its MAC protocol, LoRaWAN, indoors in terms… ▽ More

    Submitted 19 June, 2025; originally announced June 2025.

  18. arXiv:2505.14692  [pdf, ps, other] 

    cs.SE cs.CL cs.LG

    Sentiment Analysis in Software Engineering: Evaluating Generative Pre-trained Transformers

    Authors: KM Khalid Saifullah, Faiaz Azmain, Habiba Hye

    Abstract: Sentiment analysis plays a crucial role in understanding developer interactions, issue resolutions, and project dynamics within software engineering (SE). While traditional SE-specific sentiment analysis tools have made significant strides, they often fail to account for the nuanced and context-dependent language inherent to the domain. This study systematically evaluates the performance of bidire… ▽ More

    Submitted 22 April, 2025; originally announced May 2025.

  19. arXiv:2504.01261  [pdf, other] 

    cs.RO cs.CV

    ForestVO: Enhancing Visual Odometry in Forest Environments through ForestGlue

    Authors: Thomas Pritchard, Saifullah Ijaz, Ronald Clark, Basaran Bahadir Kocer

    Abstract: Recent advancements in visual odometry systems have improved autonomous navigation; however, challenges persist in complex environments like forests, where dense foliage, variable lighting, and repetitive textures compromise feature correspondence accuracy. To address these challenges, we introduce ForestGlue, enhancing the SuperPoint feature detector through four configurations - grayscale, RGB,… ▽ More

    Submitted 1 April, 2025; originally announced April 2025.

    Comments: Accepted to the IEEE Robotics and Automation Letters

  20. arXiv:2503.19152  [pdf, other] 

    eess.IV cs.AI cs.CV

    PSO-UNet: Particle Swarm-Optimized U-Net Framework for Precise Multimodal Brain Tumor Segmentation

    Authors: Shoffan Saifullah, Rafał Dreżewski

    Abstract: Medical image segmentation, particularly for brain tumor analysis, demands precise and computationally efficient models due to the complexity of multimodal MRI datasets and diverse tumor morphologies. This study introduces PSO-UNet, which integrates Particle Swarm Optimization (PSO) with the U-Net architecture for dynamic hyperparameter optimization. Unlike traditional manual tuning or alternative… ▽ More

    Submitted 24 March, 2025; originally announced March 2025.

    Comments: 9 pages, 6 figures, 4 tables, Gecco 2025 Conference

    MSC Class: 68Q07 ACM Class: I.4.6; I.2

    Journal ref: GECCO '25 Companion: Proceedings of the Genetic and Evolutionary Computation Conference Companion, 2025

  21. arXiv:2502.15716  [pdf, ps, other] 

    cs.DC cs.LG

    Feature-Aware Task-to-Core Allocation in Embedded Multi-core Platforms via Statistical Learning

    Authors: Mohammad Pivezhandi, Abusayeed Saifullah, Prashant Modekurthy

    Abstract: Optimizing task-to-core allocation can substantially reduce power consumption in multi-core platforms without degrading user experience. However, existing approaches overlook critical factors such as parallelism, compute intensity, and heterogeneous core types. In this paper, we introduce a statistical learning approach for feature selection that identifies the most influential features-such as co… ▽ More

    Submitted 10 January, 2026; v1 submitted 26 January, 2025; originally announced February 2025.

    Comments: 15 pages, 9 figures. Published in IEEE RTCSA 2025

    ACM Class: C.3; D.4.1; I.2.6

  22. arXiv:2501.13357  [pdf, other] 

    cs.CV eess.IV

    A light-weight model to generate NDWI from Sentinel-1

    Authors: Saleh Sakib Ahmed, Saifur Rahman Jony, Md. Toufikuzzaman, Saifullah Sayed, Rashed Uz Zzaman, Sara Nowreen, M. Sohel Rahman

    Abstract: The use of Sentinel-2 images to compute Normalized Difference Water Index (NDWI) has many applications, including water body area detection. However, cloud cover poses significant challenges in this regard, which hampers the effectiveness of Sentinel-2 images in this context. In this paper, we present a deep learning model that can generate NDWI given Sentinel-1 images, thereby overcoming this clo… ▽ More

    Submitted 22 January, 2025; originally announced January 2025.

  23. arXiv:2412.10155  [pdf, other] 

    cs.CV cs.AI

    WordVIS: A Color Worth A Thousand Words

    Authors: Umar Khan, Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: Document classification is considered a critical element in automated document processing systems. In recent years multi-modal approaches have become increasingly popular for document classification. Despite their improvements, these approaches are underutilized in the industry due to their requirement for a tremendous volume of training data and extensive computational power. In this paper, we at… ▽ More

    Submitted 13 December, 2024; originally announced December 2024.

  24. arXiv:2409.14178  [pdf, ps, other] 

    cs.LG

    FlowRL: Flow-Augmented Few-Shot Reinforcement Learning for Semi-Structured Sensor Data

    Authors: Mohammad Pivezhandi, Abusayeed Saifullah

    Abstract: Reinforcement learning (RL) in few-shot scenarios with limited sensor data is challenging due to insufficient training samples, particularly in applications like Dynamic Voltage and Frequency Scaling (DVFS) where sensor readings are semi-structured with inherent correlations. We propose Flow-Augmented Reinforcement Learning (FlowRL), a novel method that leverages continuous normalizing flows to ge… ▽ More

    Submitted 10 January, 2026; v1 submitted 21 September, 2024; originally announced September 2024.

    Comments: 13 pages, 5 figures, 2 tables

    ACM Class: I.2.6; I.2.8

  25. arXiv:2408.15720  [pdf, other] 

    cs.CL

    An Evaluation of Sindhi Word Embedding in Semantic Analogies and Downstream Tasks

    Authors: Wazir Ali, Saifullah Tumrani, Jay Kumar, Tariq Rahim Soomro

    Abstract: In this paper, we propose a new word embedding based corpus consisting of more than 61 million words crawled from multiple web resources. We design a preprocessing pipeline for the filtration of unwanted text from crawled data. Afterwards, the cleaned vocabulary is fed to state-of-the-art continuous-bag-of-words, skip-gram, and GloVe word embedding algorithms. For the evaluation of pretrained embe… ▽ More

    Submitted 28 August, 2024; originally announced August 2024.

    Comments: arXiv admin note: substantial text overlap with arXiv:1911.12579

  26. arXiv:2408.09800  [pdf, other] 

    cs.CV

    Latent Diffusion for Guided Document Table Generation

    Authors: Syed Jawwad Haider Hamdani, Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: Obtaining annotated table structure data for complex tables is a challenging task due to the inherent diversity and complexity of real-world document layouts. The scarcity of publicly available datasets with comprehensive annotations for intricate table structures hinders the development and evaluation of models designed for such scenarios. This research paper introduces a novel approach for gener… ▽ More

    Submitted 19 August, 2024; originally announced August 2024.

    Comments: Accepted in ICDAR 2024

  27. arXiv:2407.15608  [pdf, other] 

    cs.CL

    StylusAI: Stylistic Adaptation for Robust German Handwritten Text Generation

    Authors: Nauman Riaz, Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: In this study, we introduce StylusAI, a novel architecture leveraging diffusion models in the domain of handwriting style generation. StylusAI is specifically designed to adapt and integrate the stylistic nuances of one language's handwriting into another, particularly focusing on blending English handwriting styles into the context of the German writing system. This approach enables the generatio… ▽ More

    Submitted 22 July, 2024; originally announced July 2024.

    Comments: Accepted in ICDAR 2024

  28. arXiv:2407.03830  [pdf, other] 

    cs.CV

    DocXplain: A Novel Model-Agnostic Explainability Method for Document Image Classification

    Authors: Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: Deep learning (DL) has revolutionized the field of document image analysis, showcasing superhuman performance across a diverse set of tasks. However, the inherent black-box nature of deep learning models still presents a significant challenge to their safe and robust deployment in industry. Regrettably, while a plethora of research has been dedicated in recent years to the development of DL-powere… ▽ More

    Submitted 4 July, 2024; originally announced July 2024.

    Comments: Accepted in ICDAR 2024

  29. arXiv:2406.19314  [pdf, other] 

    cs.CL cs.AI cs.LG

    LiveBench: A Challenging, Contamination-Limited LLM Benchmark

    Authors: Colin White, Samuel Dooley, Manley Roberts, Arka Pal, Ben Feuer, Siddhartha Jain, Ravid Shwartz-Ziv, Neel Jain, Khalid Saifullah, Sreemanti Dey, Shubh-Agrawal, Sandeep Singh Sandha, Siddartha Naidu, Chinmay Hegde, Yann LeCun, Tom Goldstein, Willie Neiswanger, Micah Goldblum

    Abstract: Test set contamination, wherein test data from a benchmark ends up in a newer model's training set, is a well-documented obstacle for fair LLM evaluation and can quickly render benchmarks obsolete. To mitigate this, many recent benchmarks crowdsource new prompts and evaluations from human or LLM judges; however, these can introduce significant biases, and break down when scoring hard questions. In… ▽ More

    Submitted 18 April, 2025; v1 submitted 27 June, 2024; originally announced June 2024.

    Comments: ICLR 2025 Spotlight

  30. arXiv:2405.08813  [pdf, other] 

    cs.CV cs.LG cs.MM

    CinePile: A Long Video Question Answering Dataset and Benchmark

    Authors: Ruchit Rawal, Khalid Saifullah, Miquel Farré, Ronen Basri, David Jacobs, Gowthami Somepalli, Tom Goldstein

    Abstract: Current datasets for long-form video understanding often fall short of providing genuine long-form comprehension challenges, as many tasks derived from these datasets can be successfully tackled by analyzing just one or a few random frames from a video. To address this issue, we present a novel dataset and benchmark, CinePile, specifically designed for authentic long-form video understanding. This… ▽ More

    Submitted 20 October, 2024; v1 submitted 14 May, 2024; originally announced May 2024.

    Comments: Project page with all the artifacts - https://ruchitrawal.github.io/cinepile/. Updated version with adversarial refinement pipeline and more model evaluations

  31. Comparative Analysis of Image Enhancement Techniques for Brain Tumor Segmentation: Contrast, Histogram, and Hybrid Approaches

    Authors: Shoffan Saifullah, Andri Pranolo, Rafał Dreżewski

    Abstract: This study systematically investigates the impact of image enhancement techniques on Convolutional Neural Network (CNN)-based Brain Tumor Segmentation, focusing on Histogram Equalization (HE), Contrast Limited Adaptive Histogram Equalization (CLAHE), and their hybrid variations. Employing the U-Net architecture on a dataset of 3064 Brain MRI images, the research delves into preprocessing steps, in… ▽ More

    Submitted 8 April, 2024; originally announced April 2024.

    Comments: 9 Pages, & Figures, 2 Tables, International Conference on Computer Science Electronics and Information (ICCSEI 2023)

    ACM Class: I.4.3; I.4.6

    Journal ref: E3S Web Conf. E3S Web Conf., Volume 501, 2024

  32. arXiv:2402.14020  [pdf, other] 

    cs.LG cs.CL cs.CR

    Coercing LLMs to do and reveal (almost) anything

    Authors: Jonas Geiping, Alex Stein, Manli Shu, Khalid Saifullah, Yuxin Wen, Tom Goldstein

    Abstract: It has recently been shown that adversarial attacks on large language models (LLMs) can "jailbreak" the model into making harmful statements. In this work, we argue that the spectrum of adversarial attacks on LLMs is much larger than merely jailbreaking. We provide a broad overview of possible attack surfaces and attack goals. Based on a series of concrete examples, we discuss, categorize and syst… ▽ More

    Submitted 21 February, 2024; originally announced February 2024.

    Comments: 32 pages. Implementation available at https://github.com/JonasGeiping/carving

  33. arXiv:2402.10005  [pdf, other] 

    cs.SD cs.AI cs.LG eess.AS

    ML-ASPA: A Contemplation of Machine Learning-based Acoustic Signal Processing Analysis for Sounds, & Strains Emerging Technology

    Authors: Ratul Ali, Aktarul Islam, Md. Shohel Rana, Saila Nasrin, Sohel Afzal Shajol, A. H. M. Saifullah Sadi

    Abstract: Acoustic data serves as a fundamental cornerstone in advancing scientific and engineering understanding across diverse disciplines, spanning biology, communications, and ocean and Earth science. This inquiry meticulously explores recent advancements and transformative potential within the domain of acoustics, specifically focusing on machine learning (ML) and deep learning. ML, comprising an exten… ▽ More

    Submitted 17 December, 2023; originally announced February 2024.

    Comments: 7 pages, 5 figures, Article

    MSC Class: 68Qxx; 68Uxx; 68Vxx; 68Wxx; 68Txx; 68-XX ACM Class: J.7; D.2; G.4

  34. arXiv:2401.12455  [pdf, ps, other] 

    cs.MA cs.AI cs.LG eess.SY

    Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure management

    Authors: M. Saifullah, K. G. Papakonstantinou, A. Bhattacharya, S. M. Stoffels, C. P. Andriotis

    Abstract: Life-cycle management of large-scale transportation systems requires determining a sequence of inspection and maintenance decisions to minimize long-term risks and costs while dealing with multiple uncertainties and constraints that lie in high-dimensional spaces. Traditional approaches have been widely applied but often suffer from limitations related to optimality, scalability, and the ability t… ▽ More

    Submitted 24 February, 2026; v1 submitted 22 January, 2024; originally announced January 2024.

  35. arXiv:2310.03777  [pdf, other] 

    cs.CL

    PrIeD-KIE: Towards Privacy Preserved Document Key Information Extraction

    Authors: Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

    Abstract: In this paper, we introduce strategies for developing private Key Information Extraction (KIE) systems by leveraging large pretrained document foundation models in conjunction with differential privacy (DP), federated learning (FL), and Differentially Private Federated Learning (DP-FL). Through extensive experimentation on six benchmark datasets (FUNSD, CORD, SROIE, WildReceipts, XFUND, and DOCILE… ▽ More

    Submitted 5 October, 2023; originally announced October 2023.

  36. arXiv:2309.16257  [pdf] 

    cs.CV cs.AI eess.IV

    Nondestructive chicken egg fertility detection using CNN-transfer learning algorithms

    Authors: Shoffan Saifullah, Rafal Drezewski, Anton Yudhana, Andri Pranolo, Wilis Kaswijanti, Andiko Putro Suryotomo, Seno Aji Putra, Alin Khaliduzzaman, Anton Satria Prabuwono, Nathalie Japkowicz

    Abstract: This study explored the application of CNN-Transfer Learning for nondestructive chicken egg fertility detection for precision poultry hatchery practices. Four models, VGG16, ResNet50, InceptionNet, and MobileNet, were trained and evaluated on a dataset (200 single egg images) using augmented images (rotation, flip, scale, translation, and reflection). Although the training results demonstrated tha… ▽ More

    Submitted 28 September, 2023; originally announced September 2023.

    Comments: 18 pages, 9 figures, 1 table, journal article published

    MSC Class: CS-Class: 68T07; 68T45; 68U10; ICT-Class: 94A08 ACM Class: I.2; I.4; I.5

    Journal ref: Jurnal Ilmiah Teknik Elektro Komputer dan Informatika (JITEKI), Vol 9, No 3 (2023)

  37. arXiv:2307.03996  [pdf, other] 

    cs.SE

    ReviewRanker: A Semi-Supervised Learning Based Approach for Code Review Quality Estimation

    Authors: Saifullah Mahbub, Md. Easin Arafat, Chowdhury Rafeed Rahman, Zannatul Ferdows, Masum Hasan

    Abstract: Code review is considered a key process in the software industry for minimizing bugs and improving code quality. Inspection of review process effectiveness and continuous improvement can boost development productivity. Such inspection is a time-consuming and human-bias-prone task. We propose a semi-supervised learning based system ReviewRanker which is aimed at assigning each code review a confide… ▽ More

    Submitted 8 July, 2023; originally announced July 2023.

  38. arXiv:2307.00028  [pdf, other] 

    cs.CV cs.AI cs.CL cs.LG

    Seeing in Words: Learning to Classify through Language Bottlenecks

    Authors: Khalid Saifullah, Yuxin Wen, Jonas Geiping, Micah Goldblum, Tom Goldstein

    Abstract: Neural networks for computer vision extract uninterpretable features despite achieving high accuracy on benchmarks. In contrast, humans can explain their predictions using succinct and intuitive descriptions. To incorporate explainability into neural networks, we train a vision model whose feature representations are text. We show that such a model can effectively classify ImageNet images, and we… ▽ More

    Submitted 28 June, 2023; originally announced July 2023.

    Comments: 5 pages, 2 figures, Published as a Tiny Paper at ICLR 2023

  39. arXiv:2306.13651  [pdf, other] 

    cs.CL cs.LG

    Bring Your Own Data! Self-Supervised Evaluation for Large Language Models

    Authors: Neel Jain, Khalid Saifullah, Yuxin Wen, John Kirchenbauer, Manli Shu, Aniruddha Saha, Micah Goldblum, Jonas Geiping, Tom Goldstein

    Abstract: With the rise of Large Language Models (LLMs) and their ubiquitous deployment in diverse domains, measuring language model behavior on realistic data is imperative. For example, a company deploying a client-facing chatbot must ensure that the model will not respond to client requests with profanity. Current evaluations approach this problem using small, domain-specific datasets with human-curated… ▽ More

    Submitted 29 June, 2023; v1 submitted 23 June, 2023; originally announced June 2023.

    Comments: Code is available at https://github.com/neelsjain/BYOD. First two authors contributed equally. 21 pages, 22 figures

  40. arXiv:2306.04634  [pdf, other] 

    cs.LG cs.CL cs.CR

    On the Reliability of Watermarks for Large Language Models

    Authors: John Kirchenbauer, Jonas Geiping, Yuxin Wen, Manli Shu, Khalid Saifullah, Kezhi Kong, Kasun Fernando, Aniruddha Saha, Micah Goldblum, Tom Goldstein

    Abstract: As LLMs become commonplace, machine-generated text has the potential to flood the internet with spam, social media bots, and valueless content. Watermarking is a simple and effective strategy for mitigating such harms by enabling the detection and documentation of LLM-generated text. Yet a crucial question remains: How reliable is watermarking in realistic settings in the wild? There, watermarked… ▽ More

    Submitted 1 May, 2024; v1 submitted 7 June, 2023; originally announced June 2023.

    Comments: 9 pages in the main body. Published at ICLR 2024. Code is available at https://github.com/jwkirchenbauer/lm-watermarking

  41. arXiv:2305.14637  [pdf, other] 

    cs.CV cs.LG

    Learning UI-to-Code Reverse Generator Using Visual Critic Without Rendering

    Authors: Davit Soselia, Khalid Saifullah, Tianyi Zhou

    Abstract: Automated reverse engineering of HTML/CSS code from UI screenshots is an important yet challenging problem with broad applications in website development and design. In this paper, we propose a novel vision-code transformer (ViCT) composed of a vision encoder processing the screenshots and a language decoder to generate the code. They are initialized by pre-trained models such as ViT/DiT and GPT-2… ▽ More

    Submitted 3 November, 2023; v1 submitted 23 May, 2023; originally announced May 2023.

  42. arXiv:2303.00725  [pdf, other] 

    cs.CV

    OSRE: Object-to-Spot Rotation Estimation for Bike Parking Assessment

    Authors: Saghir Alfasly, Zaid Al-huda, Saifullah Bello, Ahmed Elazab, Jian Lu, Chen Xu

    Abstract: Current deep models provide remarkable object detection in terms of object classification and localization. However, estimating object rotation with respect to other visual objects in the visual context of an input image still lacks deep studies due to the unavailability of object datasets with rotation annotations. This paper tackles these two challenges to solve the rotation estimation of a pa… ▽ More

    Submitted 1 March, 2023; originally announced March 2023.

  43. Privacy Meets Explainability: A Comprehensive Impact Benchmark

    Authors: Saifullah Saifullah, Dominique Mercier, Adriano Lucieri, Andreas Dengel, Sheraz Ahmed

    Abstract: Since the mid-10s, the era of Deep Learning (DL) has continued to this day, bringing forth new superlatives and innovations each year. Nevertheless, the speed with which these innovations translate into real applications lags behind this fast pace. Safety-critical applications, in particular, underlie strict regulatory and ethical requirements which need to be taken care of and are still active ar… ▽ More

    Submitted 8 November, 2022; originally announced November 2022.

    Comments: Under Submission

    Journal ref: Frontiers in Artificial Intelligence, 2024 (7), 1236947

  44. Tourism's trend Ranking on Social Media Data Using Fuzzy-AHP vs. AHP

    Authors: Shoffan Saifullah

    Abstract: Tourism is an exciting thing to be visited by people in the world. Search for attractive and popular places can be done through social media. Data from social media or websites can be used as a reference to find current travel trends and get information about reviews, stories, likes, forums, blogs, and feedback from a place. However, if the search is done manually one by one, it takes a long time,… ▽ More

    Submitted 28 September, 2022; originally announced September 2022.

    Comments: 7 pages, 5 Tables, 3 figures

    Journal ref: JIKO (Jurnal Informatika dan Komputer), 6(2), 2022, 153-159

  45. arXiv:2208.05865  [pdf, ps, other] 

    cs.CR cs.NI

    Transparent and Tamper-Proof Event Ordering in the Internet of Things Platforms

    Authors: Mahbubur Rahman, Abusayeed Saifullah

    Abstract: Today, the audit and diagnosis of the causal relationships between the events in a trigger-action-based event chain (e.g., why is a light turned on in a smart home?) in the Internet of Things (IoT) platforms are untrustworthy and unreliable. The current IoT platforms lack techniques for transparent and tamper-proof ordering of events due to their device-centric logging mechanism. In this paper, we… ▽ More

    Submitted 11 August, 2022; originally announced August 2022.

    Comments: 12 pages, 13 eps figures

    ACM Class: C.2.0; C.2.1; C.2.2; C.2.3; C.2.4

  46. Identification of chicken egg fertility using SVM classifier based on first-order statistical feature extraction

    Authors: Shoffan Saifullah, Andiko Putro Suryotomo

    Abstract: This study aims to identify chicken eggs fertility using the support vector machine (SVM) classifier method. The classification basis used the first-order statistical (FOS) parameters as feature extraction in the identification process. This research was developed based on the process's identification process, which is still manual (conventional). Although currently there are many technologies in… ▽ More

    Submitted 9 January, 2022; originally announced January 2022.

    Comments: 9 Pages, 5 Figures, 2 Tables

    MSC Class: 94A08 ACM Class: I.5.1; I.4.m

    Journal ref: ILKOM Jurnal Ilmiah, 13(3), (2021), 285-293

  47. K-means segmentation based-on lab color space for embryo detection in incubated egg

    Authors: Shoffan Saifullah, Rafal Drezewski, Alin Khaliduzzaman, Lean Karlo Tolentino, Rabbimov Ilyos

    Abstract: The quality of the hatching process influences the success of the hatch rate besides the inherent egg factors. Eliminating infertile or dead eggs and monitoring embryonic growth are very important factors in efficient hatchery practices. This process aims to sort eggs that only have embryos to remain in the incubator until the end of the hatching process. This process aims to sort eggs with embryo… ▽ More

    Submitted 1 August, 2022; v1 submitted 3 March, 2021; originally announced March 2021.

    Comments: 11 pages, 6 figures, ICoSiET Conference 2020, Jurnal Ilmiah Teknik Elektro Komputer dan Informatika (JITEKI)

    Journal ref: J. Ilm. Tek. Elektro Komput. dan Inform., 2022, Vol. 7, No. 2, p. 175-185

  48. Fuzzy-AHP approach using Normalized Decision Matrix on Tourism Trend Ranking based-on Social Media

    Authors: Shoffan Saifullah

    Abstract: This research discusses multi-criteria decision making (MCDM) using Fuzzy-AHP methods of tourism. The fuzzy-AHP process will rank tourism trends based on data from social media. Social media is one of the channels with the largest source of data input in determining tourism development. The development uses social media interactions based on the facilities visited, including reviews, stories, like… ▽ More

    Submitted 8 February, 2021; originally announced February 2021.

    Comments: 8 pages, 4 figures

    Journal ref: Jurnal Informatika, 13(2), (2019), 16-23

  49. Segmentasi Citra Menggunakan Metode Watershed Transform Berdasarkan Image Enhancement Dalam Mendeteksi Embrio Telur

    Authors: Shoffan Saifullah

    Abstract: Image processing can be applied in the detection of egg embryos. The egg embryos detection is processed using a segmentation process. The segmentation divides the image according to the area that is divided. This process requires improvement of the image that is processed to obtain optimal results. This study will analyze the detection of egg embryos based on image processing with image enhancemen… ▽ More

    Submitted 8 February, 2021; originally announced February 2021.

    Comments: 8 pages, in Indonesian language, 6 figures

    ACM Class: I.4.6

    Journal ref: Systemic: Information System and Informatics Journal, 5(2), (2019), 53-60

  50. arXiv:2102.00302  [pdf, ps, other] 

    cs.NI

    LPWAN in the TV White Spaces: A Practical Implementation and Deployment Experiences

    Authors: Mahbubur Rahman, Dali Ismail, Venkata P Modekurthy, Abusayeed Saifullah

    Abstract: Low-Power Wide-Area Network (LPWAN) is an enabling Internet-of-Things (IoT) technology that supports long-range, low-power, and low-cost connectivity to numerous devices. To avoid the crowd in the limited ISM band (where most LPWANs operate) and cost of licensed band, the recently proposed SNOW (Sensor Network over White Spaces) is a promising LPWAN platform that operates over the TV white spaces.… ▽ More

    Submitted 30 January, 2021; originally announced February 2021.

    Comments: ACM Transactions on Embedded Computing Systems (Accepted), 2021