Paper Digest: Recent Papers on Transformer
Paper Digest Team extracted all recent Transformer (NLP) related papers on our radar, and generated highlight sentences for them. The results are then sorted by relevance & date. In addition to this ‘static’ page, we also provide a real-time version of this article, which has more coverage and is updated in real time to include the most recent updates on this topic.
Since 2018, Paper Digest has built a foundation of data spanning decades of conferences, journals, and research topics. The platform features a daily digest service that sifts through tens of thousands of new papers, clinical trials, news articles, and community posts, filtering the noise to highlight what matters most to specific interests. Beyond daily updates, dozens of built-in research tools streamline the academic workflow, supporting efficient reading and writing, comprehensive literature reviews, and automated research report generation.
Paper Digest Team
New York City, New York, 10017
team@paperdigest.org
TABLE 1: Paper Digest: Recent Papers on Transformer
| Paper | Author(s) | Source | Date | |
|---|---|---|---|---|
| 1 | Deep Learning-Based Approach to Detect Depressive Tendencies from Social Media Posts By Leveraging Advanced NLP Techniques and Multimodal Data Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a deep learning-based approach to detect depressive tendencies from social media posts by leveraging advanced Natural Language Processing (NLP) techniques and multimodal data analysis. |
Suganthi D.; A. Geetha; | International Journal of Innovative Science and Research … | 2026-09-04 |
| 2 | A Deterministic Procedure-Aware Bilingual Retrieval-Augmented Generation Framework for Trustworthy High-Stakes AI Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This has made these models unsuitable for contexts in which incorrect responses can have serious implications. Therefore, in this context, this paper proposes a Deterministic Procedure-Aware bilingual Retrieval-Augmented Generation (DPAM-RAG) model, which can be highly beneficial in designing religious advisory systems. |
Abdullah Bin Sawad; Muhammad Binsawad; | Applied Sciences | 2026-09-03 |
| 3 | Integrating Opcode N-Grams and Word Embeddings for Enhanced Malware Classification: A Comparative Study with Transformer-Based Representations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work proposes a comparative framework for malware classification that evaluates the synergy between traditional feature engineering and modern deep learning architectures. |
Siddhita Joshi; Sonya Hu; Fabio Di Troia; | Electronics | 2026-09-03 |
| 4 | A Knowledge-Enhanced Iterative Reasoning Framework for Accurate and Traceable Fault Diagnosis in Distributed Service Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a knowledge-enhanced iterative reasoning framework that integrates large language models (LLMs) with a numerical domain knowledge graph (KG). |
Yuze Zhang; Jian Zhang; Junyuan Wang; Shan Zhang; | Sensors | 2026-09-02 |
| 5 | A Comparative Assessment of ChatGPT, DeepSeek and Human Translations of Cultural References in The Moroccan Novel For Bread Alone Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study evaluates the effectiveness of AI vis-à-vis human translations of CRs. |
Said Faiq; Mai Zaki; | Journal of Language Teaching and Research | 2026-09-02 |
| 6 | A Low-Resource Arabic Dataset and Transformer-Based Benchmark for Dark Pattern Detection in E-Commerce Mobile Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To the best of our knowledge, this paper presents the first ML-based benchmark for Arabic dark pattern detection. |
Reham Alabduljabbar; | Electronics | 2026-09-02 |
| 7 | How Perturbations Propagate: A Multi-Level Analysis of Robustness in Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study how six naturalistic and synthetic input perturbations propagate through decoder-only language models at three levels: output behavior, hidden-state geometry, and attention-head function. |
Dun Li Chan; Emily Liu; Niyathi Allu; Christian Hoang; | arxiv-cs.CL | 2026-09-02 |
| 8 | Polish ModernBERT: The Long and Short of Polish Language Understanding Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce \textbf{Polish ModernBERT}, a family of four Polish encoders available at Base and Large scales, each with 512-token and 8K context variants. |
Michał Perełkiewicz; Sławomir Dadas; Rafał Poświata; Małgorzata Grębowiec; | arxiv-cs.CL | 2026-09-01 |
| 9 | Neural Turing Machines for Efficient Natural Language Summarization: Architecture, Optimization, and Performance Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods This study presents a Neural Turing Machine (NTM)-based framework for abstractive text summarization. |
Kartik Reddy Katti; Kartikeya Reddy Katti; Amanul Islam; | Frontiers in Artificial Intelligence | 2026-08-31 |
| 10 | Utilising Speech-derived Biomarkers to Detect Alzheimer’s Disease with BERT-based Language Models: A Machine Learning Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We processed speech transcriptions from DementiaBank to surface discriminatory speech biomarkers—verbal pauses, disfluencies, and unintelligible words. We then fine-tuned and evaluated lightweight LMs (BERT, AlBERT, and DistilBERT) on these biomarker-conditioned transcripts for automatic AD classification. |
ZARA KHANNA et. al. | Frontiers in Artificial Intelligence | 2026-08-31 |
| 11 | Detoxifying Toxic Communication: A Design Science Approach to Responsible AI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study adopts a Design Science Research approach to create a responsible AI artifact that detects and detoxifies toxic communication. |
Hossein Arshadi Soufiani; Henry M. Kim; Hjalmar Turesson; Syed Mohammad Arham Noman; Anav Setia; | arxiv-cs.CY | 2026-08-31 |
| 12 | Tracing Distinguishability Through Transformer Processing with Stochastic LayerNorm Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We instead give representations volume, turning similarity into statistical distinguishability. |
Kieran Murphy; | arxiv-cs.LG | 2026-08-31 |
| 13 | A Lightweight DistilBERT-Attention Model for Aspect-Based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces an efficient DistilBERT-Attention model for aspect-based sentiment analysis (ABSA), designed to balance classification accuracy against computational cost. |
Mohammad Abu Kausar; Mohammad Nasar; Sallam O. F. Khairy; S. M. Emdad Hossain; Arockiasamy Soosaimanickam; | Journal of Computers, Mechanical and Management | 2026-08-31 |
| 14 | Sentiment Analysis of Imbalanced Dataset Through Data Augmentation and Generative Annotation Using DistilBERT and Low‐Rank Fine‐Tuning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose a framework that leverages large language models and lightweight transformer fine tuning to improve sentiment classification on imbalanced datasets. |
Hossein Nekkouei Nasrabadi; Mohammad Hossein Moattar; | Applied AI Letters | 2026-08-31 |
| 15 | Automated Classification of SAP Literature: Predicting Impact and Trends Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This contribution provides an automated framework for literature analysis in the SAP context, enabling researchers and practitioners to identify high-impact studies and support the identification of emerging research directions. |
Kawkab Bouressace; Tamás Orosz; | Acta Technica Jaurinensis | 2026-08-28 |
| 16 | GPT-only Vs. GPT with RAG: A Study on Accuracy in Handling University-Specific Queries Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: While Large Language Models (LLMs) have demonstrated impressive capabilities in general natural language processing, their accuracy often diminishes in domain-specific contexts where precise, factual responses are crucial. This study addresses this limitation within the higher education sector by comparing two approaches to handling university-specific queries. |
Meltem Cakar; | Athens Journal of Τechnology & Engineering | 2026-08-28 |
| 17 | SCIT: Testing Causal Cache Carriers in Latent Chain-of-Thought Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce SCIT, the Suffix Cache Interchange Test, a causal protocol that constructs exact source-recipient counterfactuals, patches declared cache segments, and identifies which transformer object carries the counterfactual computation. |
Yi Ding; Lijun Huang; Menglin Yang; | arxiv-cs.CL | 2026-08-27 |
| 18 | TSRB: Transformer-based Semantic Refinement Block for Sentiment Analysis Using Scene Text Images Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work aims to use scene text images for sentiment analysis to assist in understanding the intentions of captured scenes. |
SOUTIK MUKHERJEE et. al. | International Journal of Pattern Recognition and Artificial … | 2026-08-27 |
| 19 | Deep Learning-based BDS-3 Ephemeris And positioning Performance Prediction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study employs Gated Recurrent Unit (GRU), Temporal Convolutional Network (TCN), and Transformer for ephemeris prediction, and proposes two Transformer-based positioning performance prediction methods. |
Xinyu Hao; Bo Yan; Qianqian He; Dingfan Xing; | Journal of Applied Geodesy | 2026-08-26 |
| 20 | AFAR: Automated Feedback for Arabic Responses of Short-Answer Questions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: There is still limited research applying LLMs to provide feedback for short-answer tasks in low-resource languages, such as Arabic. To bridge this gap, this study introduces AFAR (Automated Feedback for Arabic Responses), a novel framework designed to generate personalized feedback on students’ short-answer responses in Arabic using GPT-4. |
Sara Alqaidi; Omaima Almatrafi; Arwa Wali; | TEM Journal | 2026-08-26 |
| 21 | Tradespace Analysis Using GPT: A Comparative Study with Humans Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract With large language models (LLMs) becoming ever more popular and their usage expanding into various domains, this study explores how effective an LLM like GPT-4 would be in analyzing requirement specification documents and using relevant information from those to create a tradespace matrix. |
Mozhdeh Rahmanpour; Aathira Anil Kumar; David Joy; Mathew Baby; Beshoy Morkos; | Journal of Computing and Information Science in Engineering | 2026-08-26 |
| 22 | BanglaMamba: Exploring State Space Models for Bangla Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose BanglaMamba and compare it with pre-trained BanglaBERT and a similarly configured BERT model trained from scratch. |
M. K. Khalidi Siam; | arxiv-cs.CL | 2026-08-25 |
| 23 | The Changing Geometry of Grammar: Dimensionality and Neighborhood Reorganization Across Transformer Layers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Such data tend to concentrate on lower-dimensional sub-manifolds, a form of compression quantified by the Intrinsic Dimensionality (ID), the minimum number of independent variables needed to represent them without significant information loss. In this work, we ask whether the grammatical role of tokens, as marked by their part-of-speech (PoS) tag, shapes the local geometry of this manifold. |
SAMUELE VALLISA et. al. | arxiv-cs.CL | 2026-08-25 |
| 24 | Hallucination Rate of Peer-Reviewed Citations Generated By Large Language Models in Neurocritical Care Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: IMPORTANCE: Large language models (LLMs) are increasingly used for scientific literature retrieval, yet their citation accuracy in specialized clinical domains remains poorly … |
Ali Seifi; Ali Seyfi; | Critical Care Explorations | 2026-08-25 |
| 25 | TianoForge: An Automated Bug Triage Approach for The TianoCore UEFI Firmware Development Community Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a novel approach to bug triage in the TianoCore open-source UEFI firmware development ecosystem. |
Nazanin Siavash; Terrance E. Boult; Armin Moin; | arxiv-cs.SE | 2026-08-24 |
| 26 | Research on Intelligent Diagnosis of DC Magnetic Bias of Power Transformers Based on Vibration Signals and Improved 2DWT-CNN-Transformer Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Aiming at the problem that the time–frequency characteristics of transformer vibration signals under DC bias are complex and the adjacent bias levels are difficult to distinguish, this paper proposes a 2DWT-CNN-Transformer diagnostic method that combines two-dimensional discrete wavelet transform, a convolutional neural network, and Transformer Encoder. |
HUIDA DUAN et. al. | Electronics | 2026-08-24 |
| 27 | The Communication Map of A Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the communication map, which charts every potential communication channel in a language model from weights alone, generalizing the composition score of Elhage et al. (2021) into a single coupling coefficient covering all 18 connection classes, from entire attention head circuits to single neurons. |
Richard Zhe Wang; | arxiv-cs.LG | 2026-08-22 |
| 28 | A Distilbert Case-Based Deep Learning Model for Part-of-Speech Tagging for Under-Resourced Kenyan Language: A Case of Dholuo Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed method combines contextual transformer representations with complementary word representations and a training strategy that is optimized for the linguistic features of Dholuo, while previous studies primarily used multilingual transformer models or traditional sequence-labeling methods. |
Maureen Otieno; Lilian Wanzare; Calvins Otieno; | International Journal of Computer Trends and Technology | 2026-08-22 |
| 29 | AraCTI-NER: A Dataset and Benchmark for Arabic Cyber Threat Intelligence Named Entity Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce AraCTI-NER, a dataset of 10,312 token-level annotated samples (275,530 tokens; 42,360 entity spans) over eight STIX-inspired entity types, built by an LLM-assisted pipeline seeded with authentic Arabic cybersecurity articles, structurally validated and rebalanced through targeted generation. |
Joud Alghamdi; Souham Meshoul; | Electronics | 2026-08-21 |
| 30 | LEXF-ATT-XLM: A Hybrid Lexicon-enhanced Attention Model for Hate Speech Detection in Low-resource Language Roman Urdu Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose LEXF-ATT-XLM, a novel hybrid deep learning architecture that synergizes contextual language modeling with explicit domain knowledge. |
Jaweria Jalil Awan; Muhammad Hamid; Tagrid Abdullah N. Alshalali; | PLOS One | 2026-08-20 |
| 31 | A Comparative Review of Modern Large Language Model Paradigms: GPT-4, BERT, Gemini, and DeepSeek Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review provides comparative analysis of GPT-4, BERT (bidirectional encoder representations from transformers), Gemini, and DeepSeek large language models (LLM), focusing architectures, training methodologies, and real-world applications. |
Kavish Sanghvi; Aparna S. Sharma; Surbhi Hooda; | Computer Science and Information Technologies | 2026-08-19 |
| 32 | Malformer: A Multi-Modal Malware Detector Using Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we present Malformer, a quadrimodal malware detection model that incorporates text, image, graph, and audio representations of Windows executables. |
SAMUEL HOWARD et. al. | arxiv-cs.CR | 2026-08-19 |
| 33 | Lightweight Multimodal-Guided Diffusion Transformer for Agricultural Product Packaging Visual Concept Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To address the problems of high costs during the conceptual design stage of agricultural product packaging and weak correlation between visual styles and market feedback, this study proposes a Lightweight Multimodal-Guided Diffusion Transformer (LMG-DiT). |
Yanan Jiang; Yongxiao Liu; | International Journal of Pattern Recognition and Artificial … | 2026-08-18 |
| 34 | Algorithmic Amplification of Geopolitical Narratives and User Engagement in The 2024 U. S. Presidential Election Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines how geopolitical narratives were amplified through presidential campaign communication on YouTube during the 2024 U. S. presidential election, and the extent to which such narratives were associated with differential user engagement (operationalized as a platform-based engagement proxy) compared to content related to domestic issues. |
Ahmed Hassan; | Frontiers in Political Science | 2026-08-18 |
| 35 | Efficient INT8 Inference of Small NLP Models on Server CPUs with PyTorch Native Stack Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We integrate SmoothQuant into TorchAO and optimize the resulting inference path for Intel Xeon CPUs through graph-level fusion in TorchInductor and efficient INT8 GEMM kernel selection across oneDNN-, AVX512_VNNI-, and AMX-based implementations. |
Weiwen Xia; Yuxin Cui; E Cao; | arxiv-cs.CL | 2026-08-18 |
| 36 | Vocabulary Generation and Automatic Evaluation in Chinese Reading Comprehension Using Deep Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study develops a Transformer-based Contrastive Generation and Evaluation (TCGE) framework combining deep learning and generative evaluation for vocabulary generation and automatic assessment of Chinese reading comprehension oriented to internet social neologisms, and constructs cognitive processing modeling aligned with human reading logic. |
Mengyang Li; Shufang Lu; | International Journal of Pattern Recognition and Artificial … | 2026-08-18 |
| 37 | Q-Interference: Memory-Efficient Phase-Aware Quantum-Inspired Attention Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Experiments on public benchmark datasets and baseline models show that the proposed reformulation trains stably in a controlled GPT-style setting and provides a consistent memory advantage over naive phase-aware interference attention. These results support the specific contribution of this work: an exact memory-efficient reformulation that makes phase-aware interference attention practical within a standard GPT pipeline. |
EMAMA NAHID et. al. | arxiv-cs.CL | 2026-08-17 |
| 38 | Comprehensive Overview of A Transformer-Based Model for Detecting Fake News and Sentiment Trends on The Social Media Platform Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The spread of information, including fake content and deceptive narratives, has been greatly expedited by the quick development of social media platforms, especially Twitter. It … |
Gurpreet Kaur; Dr. Jagdev Singh Rana; | International Journal for Research in Applied Science and … | 2026-08-17 |
| 39 | RecurrentGPT: Expressive Depth Through Recurrent Modulation in Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce RecurrentGPT, a recurrent depth transformer where fixed-depth prelude and coda blocks bracket a single shared core iterated R times. |
Amr Hegazy; Amr Alanwar; Mostafa Elhoushi; | arxiv-cs.CL | 2026-08-15 |
| 40 | SAPE: Sandwich Adapters for Parameter Efficiency in Large Language Model Fine-Tuning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The potential of sharing patterns inspired by the inherent hierarchical structure of Transformer architectures remains unexplored in PEFT. To address this gap, we introduce SAPE (Sandwich Adapters for Parameter Efficiency), a PEFT framework based on a sandwich-style hard weight-sharing topology. |
Mohammad Aref Jafari-Raddani; Morteza Mohajjel Kafshdooz; | arxiv-cs.LG | 2026-08-15 |
| 41 | StateM: Reaching 95.3% Raw Accuracy, or A \$15 Frontier Run, on Terminal-Bench 2.1 Via Harness Scaling Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce StateM, an agent-native runtime that organizes execution around durable states, phase-local context, checked transitions, recoverable runbooks, and versioned procedural practices that agents and users can inspect together. |
Ziheng Qin; Yaxin Lu; Zhangyang Atlas Wang; Kai Wang; | arxiv-cs.AI | 2026-08-15 |
| 42 | Perbandingan Akurasi Model AI Dalam Menilai Kepribadian Pengguna Menggunakan Pendekatan NLP (Natural Language Processing) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study compares the effectiveness of deep learning approaches (BERT and DistilBERT) against a rule-based method in predicting Big Five personality traits from text data. |
Ryan Hadi Ardhiansyah –; | Jurnal Informatika dan Teknik Elektro Terapan | 2026-08-13 |
| 43 | Extracting Fine-Grained Sentiment Features About Library Services from Reader Feedback Text Using RoBERTa Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Robustly Optimized Bidirectional Encoder Representations from Transformers Approach (RoBERTa) model, fine-tuned for domain adaptation with library domain data, and introducing the Aspect-Based Sentiment Analysis (ABSA) framework to extract and classify fine-grained sentiment related to collection services in reader feedback. |
C. H. Song; | Advanced Electromagnetics | 2026-08-13 |
| 44 | Adversarial Robustness in Smishing Detection: A Comparative Analysis of Adversarial Fragility in Classical Vs. Transformer-Based Detection Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study evaluates adversarial robustness for five model architectures: three classical lexical models (Random Forest, XGBoost, CNN+BiLSTM) and two multilingual transformers (mBERT, XLM-RoBERTa), using a dataset of 27,037 messages. |
Denzel Chiuseni; Athanase Bahizire; Silva Hama; Jema David Ndibwile; | arxiv-cs.CR | 2026-08-13 |
| 45 | From BERT to Frontier Agents: Eight Years of Language-Model Progress, The Collapse of The Capability-Cost Curve, and The Rise of Task-Targeted Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In a grade school math test using the Qwen 2 point 5 model basic methods solved 58 of 100 problems while advanced sampling solved up to 79. |
Pranav Kumar Kaliaperumal; | arxiv-cs.LG | 2026-08-13 |
| 46 | LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study whether financial time series are useful as an additional input on the task of classifying sentences from Federal Reserve communication as hawkish, dovish, or neutral. |
Michael Schlee; Fabian Lukassen; Christoph Weisser; | arxiv-cs.CL | 2026-08-12 |
| 47 | An Exploration of Novel Identification Techniques for Mitigating Multi-label Online Toxicity on Social Media Platforms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a hybrid architecture, which is called Identification of Multi-Label Toxicity Classification (IMLTC), to fine-grained multi-label toxicity detection in texts. |
Abarna Sundaramurthy; Sheeba Jayaraj Immanuel; | Journal of Intelligent & Fuzzy Systems: Applications in … | 2026-08-12 |
| 48 | Assessing Reliability of BERT-Based Models on Question Answering Tasks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Recent advancements in natural language processing (NLP), particularly those based on transformer architectures, have significantly accelerated progress across various NLP tasks. This study focuses on the reliability of transformer-based question answering (QA) models, specifically BERT models and its variants (RoBERTa, ALBERT, DistilBERT). |
Pooja Yadav; Priyanka Harjule; Basant Agarwal; Marko Robnik Šikonja; | arxiv-cs.CL | 2026-08-11 |
| 49 | Assessing Transformer Models for Abstractive Summarization of Scientific Articles Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we investigate the performance of three pre-trained Transformer-based models—T5, BART, and GPT-2—on the task of abstractive summarization using the CL-SciSumm 2019 dataset. |
Emad Nabil; | Islamic University Journal of Applied Sciences | 2026-08-11 |
| 50 | Resilient Semantic Threat Detection at The Edge: A Knowledge Distillation Framework for SMS Spam Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a high-efficiency detection framework utilizing DistilBERT, a distilled knowledge representation of the BERT transformer. |
Mrinal Mrinal; Neeraj Kumar; | International Journal of Creative and Open Research in … | 2026-08-11 |
| 51 | Does News Tone from Large Language Models (LLMs) Predict Stock Returns? Evidence from Korea Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates whether textual tone derived from large language models (LLMs) can predict future stock returns. |
Cheol-Won Yang; | Korean Journal of Financial Studies | 2026-08-11 |
| 52 | A Deep Learning Model for Prediction of Unknown Gene Functionality: Gene Bio‐BERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: ABSTRACT The Human Genome Project (HGP) was a large international research effort that timelined between 1990 and 2003, marking the successful mapping of the entire human genome. |
Srinivas Kudipudi; Vyshnavi Durga Chirumamilla; Pavani Ippili; | PROTEOMICS | 2026-08-10 |
| 53 | UNMASK: Discovering and Causally Verifying Spurious Shortcuts in Text Classifiers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present U N M ASK, a fully automated pipeline that discovers, causally verifies, and mitigates spurious correlations in text classifiers without additional human annotation. |
Chidaksh Ravuru; Shashank Srivastava; | arxiv-cs.CL | 2026-08-10 |
| 54 | LegoLM: Structured Weight Sharing for Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present \LegoLM{}, a structured weight-sharing compression framework for large language models grounded in a systematic study of why global weight sharing fails and how to fix it. |
Joseph Bingham; | arxiv-cs.LG | 2026-08-09 |
| 55 | ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing systems predominantly focus on classification, while current report-generation methods often produce outputs that remain inadequate for practical clinical use. To address these challenges, we propose ECG-LENS, an end-to-end ECG report-generation framework that jointly integrates multi-lead signal modeling, diagnosis-aware representations, and clinically grounded text generation. |
AKANTA DAS et. al. | arxiv-cs.AI | 2026-08-06 |
| 56 | Neuro-semantic Graph Fusion for Explainable Depression Risk Trajectory Mapping Using A Graph-enhanced RoBERTa Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes Graph-RoBERTa-CL, a hybrid neuro-semantic framework that explicitly integrates three synergistic components: (i) RoBERTa-based Transformer semantic encoding for deep contextual representation, (ii) Graph Attention Networks (GAT) for relational contextual aggregation across semantically related posts, and (iii) Supervised Contrastive Learning (SCL) for latent space regularization under class imbalance, collectively enabling explainable depression risk trajectory mapping from social media corpora. |
M. Karthiga; Emerson Raja Joseph; Subhash Patil; K. Saranya; | Frontiers in Artificial Intelligence | 2026-08-05 |
| 57 | BnBERT-iPET: Sparse Few-Shot Language Modeling for Bengali Via Lottery Ticket Pruning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce BnBERT-iPET, a sparse few-shot language modeling approach for Bengali, and experimentally show that a lightweight few-shot-learned language model retaining only 10% of the edges of an initial model such as BERT can perform neck and neck with much larger models on challenging tasks for a resource-constrained language such as Bengali. |
Sajib Hossain; Md Kamrus Samad; Anan Ghosh; Labib Imam Chowdhury; Nabeel Mohammed; | arxiv-cs.LG | 2026-08-05 |
| 58 | Counterfactual Analysis Via Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Counterfactual analysis aims to predict potential outcomes under hypothetical scenarios, offering valuable insights for decision-making. |
Zonghao Yang; | arxiv-cs.AI | 2026-08-05 |
| 59 | Benchmarking Classical and Transformer-Based Models for Document Sensitivity Classification Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Automatic sensitivity classification of organizational documents is a critical yet underserved problem, where the consequences of misclassification range from regulatory … |
Aleesha Zainab; Muhammad Ahmed Khalid; Faheem Ullah Khan; Asifullah Khan; | arxiv-cs.LG | 2026-08-05 |
| 60 | When Do PEFT Adaptations Leak Structure? Measuring Black-Box Structural Bounds in Public-Base Model Services Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present VectorHijack-SR, a measurement methodology that converts paired victim/base residuals into calibrated structural bounds over PEFT family, layer locality, and coarse rank, while separating metadata visibility from open-world validity and operational exploitability. |
Zhongjiang Yao; Shuangshuang Liang; Chun Yang; LiWei Chen; Gang Shi; | arxiv-cs.CR | 2026-08-05 |
| 61 | ANNOTARES: A Dataset for Extracting Logical Structures from German Statutory Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we introduce the task of identifying and segmenting legal conditions (Tatbestand) and legal consequences (Rechtsfolge) within German statutory texts. |
Ronja Schwarz; Jannik Strötgen; | arxiv-cs.CL | 2026-08-04 |
| 62 | Ordinal Sentiment Classification in Cancer Support Forums: A Controlled Benchmark of Machine Learning and Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods: We benchmarked representative classical, recurrent, and transformer-based models for four-class ordinal sentiment classification using the Mental Health Insights—Vulnerable Cancer Patients dataset ( N = 10,392). |
ZHONGYAN WANG et. al. | Frontiers in Digital Health | 2026-08-04 |
| 63 | Differential Accuracy and Concordance of ChatGPT and DeepSeek for Image-Based Questions in Undergraduate Medical Education Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Large language models (LLMs) such as ChatGPT (GPT) and DeepSeek (DS) are increasingly explored in medical education and practice. |
Lorraine Silva Requena; Joyce Santana Rizzi; Zilda Maria Tosta Ribeiro; Pedro Tadao Hamamoto Filho; Renato Ferretti; | Medical Science Educator | 2026-08-04 |
| 64 | ChaosProbe: A Neurochaotic Lens on Frozen Transformer Input-Embedding Spaces Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Yet frozen transformer input-embedding spaces may also be examined through their responses to a controlled deterministic probe before contextual computation or task-specific adaptation. Guided by this response-based view, we introduce \emph{ChaosProbe}, a deterministic neurochaos-inspired method for constructing response-based fingerprints of frozen transformer input-embedding spaces. |
Kunal Kumar Pant; Nithin Nagaraj; | arxiv-cs.LG | 2026-08-03 |
| 65 | One QK Channel, Many Sources: Guarding Low-Precision Attention Collapse Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We isolate a reproduced GPT-2-class collapse to the streaming-softmax accumulator, where fp32 accumulation repairs it, and use the fault as an assay for moving controlled errors across sources. |
SHUXIAO XIE et. al. | arxiv-cs.LG | 2026-08-03 |
| 66 | Feed-Forward Steering in Transformer Residual Dynamics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We extend this framework by incorporating the feed-forward network (FFN) term as a local steering field acting on each token state. |
Timur Mudarisov; Mikhail Burtsev; Radu State; | arxiv-cs.LG | 2026-08-03 |
| 67 | DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work introduces DeBERTa-Sentinel, a responsible AI-generated text detection framework leveraging DeBERTa-v3’s disentangled attention to capture subtle structural irregularities in synthetic content. |
Muhammad Yousaf Rehman; Muhammad Islam; | arxiv-cs.CL | 2026-08-02 |
| 68 | AI-Driven Detection and Mitigation of Disinformation Threats to U.S. National Security: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This systematic review synthesizes the peer-reviewed and governmental literature published between 2019 and 2024 on AI-driven disinformation detection and mitigation. The review examines four detection technology classes: natural language processing (NLP) classifiers, deepfake forensics, network propagation analysis and content provenance systems, alongside the U.S. policy and institutional frameworks designed to counter these threats. |
Mohammed Hafiz Nabila; | American Journal of Applied Research and AI | 2026-08-01 |
| 69 | Transfer Learning in Neural Networks: Leveraging Pre-trained Models for Improved Performance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The quantitative comparative analysis of the pre-trained models includes ResNet50, VGG16, BERT, GPT, and the baseline CNN and LSTM models. |
Abdul Sttar Ismail Wdaa; Iraq Ali Hussein; Ali Azeez Ahmed; | Future Technology | 2026-07-31 |
| 70 | Analyzing Public Opinion on International Conflict Through YouTube Comments Using NLP Techniques Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines public opinion formation and polarization surrounding international armed conflicts by analyzing a corpus of 2.4 million YouTube comments harvested from 847 conflict-related video uploads spanning the 2022–2024 period. |
Acep Muchtarom; Alim Jaizul Wahid; Nurfadhlina Abdul Halim; | International Journal of Mathematics, Statistics, and … | 2026-07-31 |
| 71 | Public Sentiment Analysis of The Iran–United States Conflict on BBC News YouTube Comments Using DistilBERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines public sentiment expressed in YouTube comments on BBC News coverage of the Iran–United States conflict, employing DistilBERT—a distilled variant of BERT—as the primary classification model. |
Charis Maulana; Setyo Luthfi Okta Yohandoko; Rifki Saefullah; | International Journal of Mathematics, Statistics, and … | 2026-07-31 |
| 72 | Uncovering Bias: Leveraging Large Language Models for News Analysis Inthe Context of Palestine Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we demonstrate the potential of lightweight transformer-based language models to enable bias detection in digital journalism with a scalable and consistent approach. |
Saba Saddique; Usman Ahmad; | Journal of Intelligent Systems and Computer Applications | 2026-07-31 |
| 73 | PERBANDINGAN KINERJA PRE-TRAINED LANGUAGE MODEL BAHASA INDONESIA BERBASIS TRANSFORMER UNTUK ANALISIS SENTIMEN ULASAN TOKOPEDIA GOOGLE PLAY Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Analisis sentimen terhadap ulasan aplikasi e-commerce pada Google Play Store merupakan salah satu pendekatan penting untuk memahami persepsi pengguna. Namun, karakteristik data … |
Muhammad Kevin; Hanafi Hanafi; | INTECOMS: Journal of Information Technology and Computer … | 2026-07-30 |
| 74 | A Comparative Analysis of Automated Techniques for Security Bug Report Identification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: As a result, it is difficult to compare their results and draw reliable conclusions about the effectiveness of existing approaches, leaving researchers and practitioners without clear guidance on which techniques are most suitable for the task. To address this gap, we conducted a comparative analysis of several promising automated techniques to identify security-related bug reports using benchmark datasets. |
Muhammad Laiq; | arxiv-cs.SE | 2026-07-30 |
| 75 | Overview of RAG-based and LLM-based Approaches to Personalization in Healthcare AI Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To address these gaps, Retrieval-Augmented Generation (RAG) can couple LLMs with external knowledge sources at inference time, producing responses that are more grounded, up-to-date, and tailored to individual patient needs. This systematic review examines how personalization is operationalized in LLM-only and RAG-enhanced healthcare AI systems. |
Manal Althobaiti; Minhee Jun; | Frontiers in Artificial Intelligence | 2026-07-29 |
| 76 | GPT Outperforms BERT and LIWC for Stance and Anger Detection in German News Articles and User Comments Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Natural language processing (NLP) has become essential for analyzing this data, with recent advances in Large Language Models (LLMs) driving rapid progress. These models show strong potential for psychological research, though their validity in detecting attitudes varies across domains. |
LUKAS MAYRHOFER et. al. | Frontiers in Political Science | 2026-07-29 |
| 77 | The Ray Tracing Sampler: Bayesian Sampling of Neural Networks for Everyone Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using the simplest ray tracing method, we sample the posterior distributions of neural network outputs for a variety of different architectures, including a preliminary exploration of the 1.5 billion-parameter GPT-2 (Generative Pre-trained Transformer 2) architecture, all on a single consumer-level GPU. |
Peter Behroozi; | The Open Journal of Astrophysics | 2026-07-29 |
| 78 | Automated Multilabel Mpox Research Classification with Explainable Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study focuses on using multilabel classification to categorize 14590 Mpox research articles into key topics such as outbreaks, vaccination, and epidemiology. |
Tanjim Taharat Aurpa; | arxiv-cs.CL | 2026-07-29 |
| 79 | Long Live Fine-tuning: Task-specific Transformers Outperform Zero-shot LLMs for Misinformation Response Classification on Reddit Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract As large language models (LLMs) become widely used tools for online information access and verification, it is important to determine when their zero-shot flexibility is sufficient and when task-specific supervision remains valuable. We examine this question in a controlled misinformation-response classification setting comprising 900 Reddit comments associated with three PolitiFact-verified claims in environment, health, and immigration. |
JooYoung Lee; Lin Tian; Angela Brillantes; Adriana-Simona Mihăiță; Marian-Andrei Rizoiu; | Social Network Analysis and Mining | 2026-07-29 |
| 80 | MDTransformer: A Hardware-Software Co-Design of Mode-Division Photonic Transformer Accelerator with Inverse-Designed Coherent Crossbar Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, state-of-the-art rely on expensive multi-wavelength light generation and large dot-product units due to active phase-shifter components, thus making their approach inefficient and impractical. To address this, we propose MDTransformer, a novel hardware-software co-design of PTA based on mode-division optical dataflow and operations. |
Solomon Micheal Serunjogi; Rachmad Vidya Wicaksana Putra; Ayat Taha; Muhammad Shafique; Mahmoud Rasras; | arxiv-cs.AR | 2026-07-28 |
| 81 | GPT-Red: Automated Red Teaming Via Self-Play at Scale Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce \textbf{GPT-Red}, an automated red-teaming agent that is trained to discover novel prompt injection attacks against frontier LLMs. |
ERIC WALLACE et. al. | arxiv-cs.CR | 2026-07-28 |
| 82 | Recursive Transformers for Semiconductor Thermo-mechanical Reliability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a hardware-aware evaluation of three recursive transformer paradigms for surrogate thermo-mechanical analysis of advanced packages: a)Tiny Recursive Model, b) our proposed Depth Recursive transformer, c) and a simple recursive transformer. |
Kart-leong Lim; | arxiv-cs.LG | 2026-07-28 |
| 83 | BioSentinel at EXIST 2026: Soft-Label Optimization with XLM-RoBERTa for Sexism Intent Classification in Memes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a text-centric approach built on xlm-roberta-base (270M parameters) trained with a composite loss function combining KL divergence on soft annotator distributions and weighted cross-entropy on hard labels. |
Chandru Munisamy; Karthikeya Raguveer; Alapan Kuila; | arxiv-cs.CL | 2026-07-27 |
| 84 | PAP_NER: A Large-scale Vietnamese Administrative Named Entity Recognition Corpus and Hybrid Deep Learning Architecture Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present PAP_NER, the first large-scale, gold-standard Vietnamese administrative NER corpus comprising 162,801 sentences with 205,807 entity annotations across five entity types critical for e-Government workflows: Agency (CQ), Legal Document (VBPL), Object (ĐT), Datetime (NG), and Quantity (SL). |
DINH-DIEN LA et. al. | PLOS One | 2026-07-27 |
| 85 | An Explainable Multimodal Framework for Breast Ultrasound Report Generation Using Vision-Language Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a trustworthy and explainable framework for automated breast ultrasound report generation that combines Vision-Language Modelling (VLM) with multi-level Explainable Artificial Intelligence (XAI). |
Prashanth Gowda Attahalli Shivakumar; Azhar Mahmood; Shaheen Khatoon; | Journal of Imaging | 2026-07-27 |
| 86 | Physics Transformer: Tailoring Transformer for General PDE Prediction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Consequently, effectively applying Transformers to PDEs requires a tokenizer that respects the functional nature of physical fields and constructs physically expressive tokens from arbitrary discretizations.To this end, we propose \methodname{Physics Transformer}, a function-projection-based Transformer architecture for physical field prediction. |
GUOZE SUN et. al. | arxiv-cs.LG | 2026-07-27 |
| 87 | Detecting Phishing Websites Using A Hybrid Approach with DistilBERT, GNN and LightGBM Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing detection systems frequently rely on singular modelling approaches and thus fall short in addressing the multidimensional and continuously evolving nature of these attacks. To overcome this challenge, the present work proposes a hybrid phishing detection framework that integrates three complementary techniques: DistilBERT (Distilled Bidirectional Encoder Representations from Transformers) for semantic analysis of URL text, Graph Neural Networks (GNN) for modelling structural relationships among URL components, and LightGBM (Light Gradient Boosting Machine) for efficient metadata-based feature classification. |
Ms. I. Shalini; Ms. G. Sujini; | International Journal for Research in Applied Science and … | 2026-07-27 |
| 88 | Enhanced Sentiment Analysis Using RoBERTa and BiLSTM: A Context-Aware Hybrid Deep Learning Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In contrast, Transformer-based models offer improved efficiency through parallel processing. To address these challenges, this paper presents a context-aware hybrid deep learning approach by integrating the Robustly Optimized BERT Pretraining Approach (RoBERTa) with Bidirectional Long Short-Term Memory (BiLSTM) networks. |
Dr. Veguru Gayatri; Dr. Rajani Rajalingam; | International Journal for Research in Applied Science and … | 2026-07-27 |
| 89 | Enhancing Text-Based Emotion Detection in Turkish and English: A Sentence-Level Enrichment Approach Using BERT and DistilBERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a novel approach to emotion detection in text by focusing on sentence-level emotion enrichment using transformer-based models, specifically BERT and DistilBERT. |
Senem Kumova Metin; Hande Aka Uymaz; | ACM Transactions on Asian and Low-Resource Language … | 2026-07-27 |
| 90 | Human–AI Co-Regulation in Adaptive Learning: Developing GPT-Supported Self-Regulated Learning Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods A Design-Based Research approach integrated with mixed methods and learning analytics was employed. |
MUH. HUSEIN BAYSHA et. al. | F1000Research | 2026-07-27 |
| 91 | Bigger Is Not Always Better: Computational Efficiency in Lexical Prehospital Triage Modeling Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Materials and Methods We conducted a retrospective study via a civilian critical care air transport service from 2012 to 2021 (approved by the University of Pittsburgh IRB). |
AARON C WEIDMAN et. al. | Military Medicine | 2026-07-26 |
| 92 | Efektivitas Model Pembelajaran LOK-R Berbasis AI (GPT) Terhadap Literasi Digital Dan Pemahaman Ekonomi Digital Siswa SMP Negeri 6 Surabaya Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Perkembangan teknologi digital menuntut peserta didik memiliki literasi digital dan pemahaman ekonomi digital yang memadai, namun pembelajaran IPS di sekolah masih didominasi … |
Lutfiah Khasanah; Nuansa Bayu Segara; Hendri Prastiyono; Asnimawati Asnimawati; | Journal of Educational Research and Humaniora (JERH) | 2026-07-26 |
| 93 | Augmenting Legal Reasoning with BERT: The Second Iteration of Bekenbey AI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we build upon the Bekenbey AI model, which is the first study in the literature to establish a connection between the legal domain and Generative AI. |
Ali Deveci; Mehmet Ali Erkan; İhsan Tolga Medeni; Tunç Durmuş Medeni; | Düzce Üniversitesi Bilim ve Teknoloji Dergisi | 2026-07-24 |
| 94 | A Hybrid RoBERTa–BiLSTM Framework for Aspect-based Sentiment Analysis of Monkeypox Tweets Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
TDN Pavani; Kuppusamy P.; | Egyptian Informatics Journal | 2026-07-24 |
| 95 | SAGE: Safety-First Defense-in-Depth Guardrails for Verified Lifecycle Control of High-Impact Generative AI Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: High-impact generative AI makes catastrophic misuse a lifecycle-control problem, not merely a prompt-filtering problem. SAGE is a safety-first, authorization-separated … |
Mahdi Eslamimehr; | arxiv-cs.AI | 2026-07-24 |
| 96 | A Hybrid Transformer–Ontology Framework for Halal Food Classification Using Multilingual Ingredient Label Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A hybrid neural–symbolic classification framework is proposed that integrates a multilingual transformer-based semantic encoder (XLM-R) with ontology-driven lexical reasoning for automated halal ingredient assessment. |
MOHD AZMI AL BETAR et. al. | Journal of Computational and Cognitive Engineering | 2026-07-24 |
| 97 | Automated Generation and Human Evaluation of Neurosurgical Board Examination Self-Assessment Questions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The goal of this study was to develop an automated system to generate board-style neurosurgical multiple-choice questions using state-of-the-art vision-language models and compare their quality with authentic self-assessment questions. |
ANTON ALYAKIN et. al. | Neurosurgery Practice | 2026-07-23 |
| 98 | Activity Classification in E-Commerce Product Reviews Using Deep Learning and Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Therefore, they do little to enhance the e-commerce experience by helping consumers make more informed purchasing decisions based on products’ intended uses without requiring them to read numerous reviews during the decision-making process. To address this problem, this paper investigates the feasibility of automatically identifying and classifying product usage activities from e-commerce reviews. |
Tinashe Wamambo; Arooj Fatima; Bethwel Kiplagat; Mahdi Maktab Dar Oghaz; Cristina Luca; | Informatics | 2026-07-23 |
| 99 | A Four-module Neural Architecture for The Automatic Extraction and Classification of Causal Relations in Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents a four-module system for the automatic extraction and classification of causal relationships from texts in the Kazakh language, based on the fine-tuning of the KazBERT transformer language model. |
ROMAN TABERKHAN et. al. | Frontiers in Artificial Intelligence | 2026-07-23 |
| 100 | Faster IndexTTS-2: Accelerating and Streaming Autoregressive Zero-Shot Text-to-Speech Synthesis on GPUs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present Faster IndexTTS-2, which accelerates all neural network components of IndexTTS-2 for production deployment on GPUs using NVIDIA TensorRT and TensorRT-LLM. |
Muyang Du; Shuang Yu; Junjie Lai; | arxiv-cs.AI | 2026-07-23 |
| 101 | Parameter-free Adaptive Sparse Attention Via Compression-Based Content Selection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We show that classical data compression provides an effective masking signal with \textbf{no additional parameters}. |
Debarshi Kundu; Swaroop Ghosh; Vasant Honavar; | arxiv-cs.LG | 2026-07-23 |
| 102 | TriAgent: Divergence-Aware Multi-Agent Committees for Cost-Efficient Financial Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present TriAgent, a multi-agent committee stratified by contextual granularity — a word-level lexicon (VADER), a sentence-level domain transformer (FinBERT), and a cross-sentence reasoner (Qwen2.5, 0.5B-14B-4bit, with Mistral-7B and Phi-3.5-mini cross-family checks). |
Isabel Xu; Cynthia Xu; Rachel Ren; Cong Guo; Jiacheng Ding; | arxiv-cs.CL | 2026-07-22 |
| 103 | Transformer-Based Language Models for Clinical Decision Support Using Clinical Notes: A Scoping Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We extracted data on clinical tasks, model architectures, enhancement strategies, and evaluation metrics; mapped each study by primary purpose, care setting, and primary model approach; and charted reported validation design, direct human comparison, fairness assessment, workflow evaluation, and clinical deployment. |
Saahoon Hong; Hunhui Na; | Information | 2026-07-22 |
| 104 | Evaluating Large Language Models for Symbolic Security Protocol Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study evaluates whether Large Language Models (LLMs) can perform comparable analysis. |
Paolo Modesti; Syed Ahmed; Ioannis Sfyrakis; Derek Enodolomwanyi; | arxiv-cs.CR | 2026-07-22 |
| 105 | Research on Emotional Persuasion Effect Identification Method of Legal Large Language Models for Human AI Collaborative Justice Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Focusing on emotional persuasion risks in collaborative justice, this study constructs a legal emotional persuasion corpus, integrates case facts, statutory references, emotion intensity, persuasion strategies, and human judgment changes, designs a multi-feature fusion identification model, and conducts a simulated judicial judgment experiment. |
Suran Chen; Jingyu Yang; | Theoretical and Natural Science | 2026-07-21 |
| 106 | A Multi-Model Text Mining Approach to Tourism Image Analysis of The Historic Centre of Macao Based on User-Generated Content Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Focusing on the Historic Centre of Macao (HCM), a World Cultural Heritage site, this study collected 3781 tourist reviews from Rednote, Ctrip, Dianping, and TripAdvisor and developed a multi-model text-mining framework integrating TF-IDF, BERTopic, and RoBERTa. |
Xiao Xu; Qiaoyun Zhang; | Buildings | 2026-07-21 |
| 107 | Cross-Domain Faithfulness Evaluation of SHAP and Attention-Based Explanations in Transformer NLP Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates whether explainability methods remain faithful and stable under domain shift in transformer-based text classification. |
Dony Bahtera Firmawan; Brian Rizqi Paradisiaca Darnoto; | Journal of Computing Theories and Applications | 2026-07-21 |
| 108 | The Role of Large Language Models As Screening Assistants in The Diagnosis of Placenta Accreta Spectrum Pathologies Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Objective Large language models (LLMs) are increasingly utilized in modern medicine. |
IASON PSILOPATIS et. al. | International Journal of Gynecology & Obstetrics | 2026-07-20 |
| 109 | Multilingual AI-Generated Text Detection in Arabic, English, and Turkish Using A Hybrid Transformer–Graph Convolutional Network Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This problem is especially challenging in Turkish, Arabic, and English due to their distinct linguistic structures, including agglutinative morphology in Turkish, root-based morphology in Arabic, and semantic ambiguity in English. To address these challenges, this study proposes a hybrid architecture that combines a Transformer-based DistilBERT model with a Graph Convolutional Network (GCN). |
Ayca Bostancioglu; Bihter Das; Muzeyyen Bulut Bulut Ozek; | Applied Sciences | 2026-07-20 |
| 110 | Beyond The Surface: Characterizing Adversarial Boundaries in Synthetic Text Attribution Across Heterogeneous Domains Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To overcome all three limitations, this paper proposes a hybrid detection framework which combines semantically deep embeddings from the RoBERTa transformer with a set of carefully designed language statistics (vocabulary richness, burstiness, and information entropy) and linguistic statistics (part-of-speech distributions, Flesch Reading Ease scores). |
Anita Rani; Ms. Suman; | International Journal of Scientific Research in Computer … | 2026-07-20 |
| 111 | Adversarial Robustness of Phishing Email Detection: A Comparative Study of TF-IDF + Logistic Regression and Fine-Tuned DistilBERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Pairwise error analysis shows the models agreed on 54.9% of adversarial samples but each made a similar number of exclusive errors (24 and 25 respectively), indicating partly complementary rather than identical failure modes. |
Tanveer Ahmed; Seyedali Pourmoafil; | arxiv-cs.CR | 2026-07-20 |
| 112 | Mobius Learning: Cyclic Depth Folding in Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Mobius Learning, a training architecture based on cyclic depth folding, in which different data streams follow cyclically shifted block orders. |
Tongtian Zhu; | arxiv-cs.LG | 2026-07-20 |
| 113 | BanClickThumb: A Multimodal Dataset and Transformer Fusion Benchmarks for Clickbait Detection in Bengali YouTube Videos Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Detecting Bengali clickbait remains challenging because publicly available multimodal datasets are limited. To address this gap, we introduce BanClickThumb, a curated dataset of 7,147 Bengali YouTube thumbnail-title pairs from five content domains, annotated by ten annotators with high agreement (Cohen’s Kappa: 0.83-0.93). |
Md. Ariful Islam; Md Tanvirul Islam; Md. Maruf Hossain Miru; Md Khalid Syfullah; | arxiv-cs.CV | 2026-07-19 |
| 114 | NERBench-Chhattisgarh: A Multi-Family NER Dataset for Low-Resource Indic Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present NERBench-Chhattisgarh, a gold-standard Named Entity Recognition (NER) dataset covering seven under-resourced languages spoken in Central India: Baigani, Chhattisgarhi, Surgujia, Sadri, Kudukh, Halbi, and Gondi. |
Rajesh Kumar Mundotiya; | sigir | 2026-07-17 |
| 115 | Cost-efficient Generative AI Summarization for Scalable Automated Essay Scoring in Educational Assessment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a generative AI-assisted summarization framework to improve long-form essay representation while maintaining scoring reliability. |
Haowei Hua; | arxiv-cs.CL | 2026-07-17 |
| 116 | Evaluating The Informational Accuracy of Large Language Models in Patient‑directed Orthodontic Retainer Guidance: A Cross‑sectional Comparison of Chat GPT‑4.1, Gemini 2.5, Microsoft Copilot GPT-4.1 and DeepSeek‑V3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Muhammad Mughni; Muhammad Raif Ilyas; Syed Muhammad Ali Gilani; Mubassar Fida; Rashna Hoshang Sukhia; | BMC Oral Health | 2026-07-15 |
| 117 | DeepLoop: Depth Scaling for Looped Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This reuse changes the residual-scaling problem: in an untied Transformer, each residual branch receives and applies its own parameter update, whereas in a looped Transformer one shared update aggregates gradients from repeated visits and is read back by those same visits in the next linearized forward pass. We formalize this tied-depth effect through a first-order perturbation bound controlled by a visit-alignment coefficient $κ_R$. |
Shuzhen Li; Yifan Zhang; Jiacheng Guo; Quanquan Gu; Mengdi Wang; | arxiv-cs.LG | 2026-07-15 |
| 118 | Application-Oriented Intelligent Sentiment Analysis: Classical Vs Deep Models with Multi-embedding Strategies Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a comprehensive comparative analysis of sentiment classification models, systematically evaluating both conventional machine learning techniques (e.g., SVM, Naive Bayes, Decision Tree, Random Forest) and advanced deep learning architectures (CNN, RNN, LSTM, BERT-LSTM) across various datasets (IMDB, Yelp, Amazon). |
Mohd. Danish; Saifullah Khalid; | Journal of Intelligent & Fuzzy Systems: Applications in … | 2026-07-15 |
| 119 | Sentiment Analysis of The Indonesian Megathrust Earthquake and Tsunami Issue Using BERT and Roberta Methods Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to analyze public opinion and identify the main topics related to the megathrust earthquake issue using Bidirectional Encoder Representations from Transformers (BERT) and Robustly Optimized BERT Pretraining Approach (RoBERTa) models. |
Hendra Rahman; Taswanda Taryo; Sudarno Wiharjo; | Jurnal Teknologi Informatika dan Komputer | 2026-07-15 |
| 120 | Privacy Leakage in Federated Learning in Radiology Reports: A Comparative Evaluation of Tokenizer-Driven Privacy Risks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We quantify gradient-based text reconstruction in FL and compare privacy risk across three tokenizers with the model architecture held fixed. |
SANTHOSH PARAMPOTTUPADAM et. al. | arxiv-cs.LG | 2026-07-15 |
| 121 | General-purpose Named Entity Recognition Using Transformer-based Fine-tuned Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The research article proposes a transformer-based fully fine-tuned XLNet model with 117M parameters. |
PARTH GOEL et. al. | PeerJ Computer Science | 2026-07-14 |
| 122 | Recent Advances in Transformer and Large Language Models for UAV Applications IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Unlike previous surveys, this work presents a unified taxonomy of Transformer-based UAV models, highlights emerging applications such as precision agriculture and autonomous navigation, and provides comparative analyses through structured tables and performance benchmarks. |
Hamza Kheddar; Yassine Habchi; Mohamed Chahine Ghanem; Mustapha Hemis; Dusit (Tao) Niyato; | ACM Computing Surveys | 2026-07-14 |
| 123 | Evolving Optimal Text Clusters: A Novel GA-driven Framework for Dynamic Ensemble Fusion of Multi-model Contextual Embeddings Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a new Genetic Algorithm (GA)-based ensemble model that dynamically optimizes the best fusion weights of SBERT, RoBERTa, and DistilBERT embeddings without ground-truth labels. |
Ali Sabah; Zaid Alaa; | PLOS One | 2026-07-13 |
| 124 | Traditional Machine Learning AndDistilBERTfor Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study compares three approaches for binary sentiment classification on the Internet Movie Database (IMDb) movie review dataset, namely Term Frequency-Inverse Document Frequency (TF-IDF)+Naive Bayes, TF-IDF+Logistic Regression, and DistilBERT, a distilled version of Bidirectional Encoder Representations from Transformers (BERT). |
Lekang Sun; | Applied and Computational Engineering | 2026-07-13 |
| 125 | Tinjauan Arsitektur Transformer Dan Penerapannya Pada Pemrosesan Bahasa Alami Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Perkembangan pesat dalam bidang pemrosesan bahasa alami (PBA) tidak terlepas dari kemunculan arsitektur Transformer yang pertama kali diperkenalkan pada tahun 2017. Arsitektur ini … |
Marcel Alezandro Sihombing; Irene Lestaria Sinaga; Octav Kornelius Hutagaol; Judea Tirta Jordan Simamora; Alex Septama Sihite; | IKRA-ITH Informatika : Jurnal Komputer dan Informatika | 2026-07-13 |
| 126 | Transformer Based End to End Web Application Firewall Pipeline for Intelligent Threat Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes an intelligent, transformer-based end-to-end Web Application Firewall (WAF) pipeline that leverages the self-attention mechanism of transformer architectures to capture long-range dependencies within HTTP request sequences. |
Muralidharan V; Gokulkrishnan S; Agesta Jenifer A; | International Scientific Journal of Engineering and … | 2026-07-12 |
| 127 | Complexity-Guided Component-wise Initialization for Language Model Pretraining Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Pretrained language models often exhibit structured weight spectra, suggesting that training may repeatedly produce similar layerwise and component-wise organization. We ask … |
Konstantin Garbers; Nicholas Oh; | arxiv-cs.CL | 2026-07-10 |
| 128 | A Security Monitoring and Warning Method for Economic Growth and Unemployment in Financial Social Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional security monitoring and warning methods often rely on lagging official statistical data, making it difficult to capture nonlinear correlations and sudden signals in economic dynamics in real time. Therefore, this paper proposes a Boruta-SHAP and Transformer-XL (BST-XL) security monitoring and warning model based on Transformer-XL. |
Lei Liu; Yuanyuan Wen; | Frontiers in Physics | 2026-07-10 |
| 129 | Physics-Informed Semantic Prompt Learning for Few-Shot Low-Altitude Radar Target Recognition in Remote Sensing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Reliable recognition remains difficult because birds, balloons, and UAVs often produce weak radar responses, share similar trajectory-level signatures, and are difficult to annotate at scale. To address these challenges, this paper proposes a physics-informed semantic prompt learning framework for few-shot low-altitude radar target recognition. |
Junrong Tu; Jihui Tu; Wenqing Feng; Zhaoyang Liu; | Remote Sensing | 2026-07-10 |
| 130 | Low Resource Word Sense Disambiguation in Oromo with Fine Tuned Small Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study uses a manually created dataset from the Oromo–English Dictionary to examine the efficacy of transformer-based models for lexical-sample WSD in Oromo. |
Liyachew Edeti; Million Meshesha; Feda Negesse; | Scientific Reports | 2026-07-10 |
| 131 | From Idea to Implementation: Evaluating The Influence of Large Language Models in Software Development—An Opinion Paper Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, opinions from 11 experts regarding their experience with LLMs for software development have been gathered and analyzed to draw insights that can guide successful and responsible integration. |
SARGAM YADAV et. al. | Applied AI Letters | 2026-07-10 |
| 132 | Prediction of Bearing Remaining Life Based on STFT-SWT and Transformer-BiLSTM Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study enhances bearing RUL prediction through improved data preprocessing and deep learning models.An enhanced wavelet threshold denoising method processes vibration signals, followed by a novel STFT-SWT time-frequency analysis for feature extraction. |
deyi wang; Xia Yang; shuangshuang Liu; Ruixiang Hou; | Engineering Research Express | 2026-07-10 |
| 133 | The Silent Freeze: Predicting When Low-Precision Training Stops Learning Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Training in reduced floating-point precision can silently halt learning: when a gradient-descent weight update falls below half the unit in the last place (ULP) of the weight, it … |
Zekai Shang; | arxiv-cs.LG | 2026-07-09 |
| 134 | ZkComposer: Decomposing Proof Construction to Scale ZkML Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce zkComposer, a modular proof-construction framework that unlocks an additional dimension of parallelism, in addition to the parallelism in existing proof kernels. |
PAWAN KUMAR SANJAYA et. al. | arxiv-cs.CR | 2026-07-09 |
| 135 | Attention Degradation, Function Token Anchoring, and The Limits of Attention-Based Intervention in Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present six coordinated experiments across GPT-2, LLaMA-3.2-1B/3B, OPT-1.3B, and distilgpt2. |
Sagar Dangal; Manoj Shakya; | arxiv-cs.AI | 2026-07-09 |
| 136 | Bedside Triage By Large Language Models in Acute Pancreatitis: A Scenario‐Based Comparative Evaluation of GPT‐4, GPT‐5, and Gemini Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Conclusion No model should be used as a stand‐alone bedside decision‐maker for AP. |
Yahya Kemal Çalışkan; Fatih Başak; Olgun Erdem; | World Journal of Surgery | 2026-07-09 |
| 137 | Smart Complaint Detection and Severity Analysis in Financial Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed model integrates a pre-trained roberta-base network as the shared encoder with task-specific attention mechanisms and classifier heads for each task. |
Shravya Koulammagari; B. Kranthi Kiran; | International Journal for Research in Applied Science and … | 2026-07-08 |
| 138 | Hierarchical Classification of Arabic Legal Cases Using Transformer Architectures and Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Automated classification of Arabic legal texts presents unique challenges stemming from the formal register of judicial language, domain-specific Sharī‘a terminology, and the severe class imbalance inherent in hierarchical legal taxonomies. This paper addresses these challenges through a systematic investigation of hierarchical multi-class classification applied to a dataset of 1146 Arabic judicial cases sourced from the Saudi Ministry of Justice open data portal. |
Nourah Alangari; Nouf Alshenaifi; Huda Almuzaini; | Electronics | 2026-07-08 |
| 139 | Comparative Evaluation of Domain-specific and General-purpose Transformer Models for Arabic Poet Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A comparative evaluation framework is introduced to contrast domain-specific and general-purpose language models across prolific authorship and cross-era stylistic variation. |
Sarah Alnefaie; | Scientific Reports | 2026-07-08 |
| 140 | PaSTO-GNN: Prompt-aware Spatio-temporal Graph Neural Networks for Automatic Essay Scoring Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing deep learning approaches—including recurrent, convolutional, and transformer-based models—primarily focus on textual semantics, yet they often overlook the spatio-temporal nature of essay composition, where meaning evolves across sentences and paragraphs through discourse progression. To address this gap, this study presents a prompt-aware Spatio-Temporal Graph Neural Network (PaSTO-GNN) for AES. |
Areej Alhothali; | Frontiers in Artificial Intelligence | 2026-07-08 |
| 141 | Cross-Dataset Generalization in Urdu Fake News Detection: An Empirical Study with XLM-RoBERTa and A Length Confound Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents the first cross-dataset generalisation study for Urdu fake news detection, using two publicly available balanced datasets: the Ax-to-Grind Urdu corpus (10,083 articles, 15 domains) and the Notri-Fact Urdu dataset (13,388 articles). |
Muhammad Abdullah Haroon; | arxiv-cs.CL | 2026-07-07 |
| 142 | Artificial Intelligence in Healthcare: Historical Development, Terminology, Concepts, and Classifications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Results . The article examines the development of AI in clinical medicine and healthcare, tracing its progression from the expert systems of the 1970s to contemporary large language models and multimodal systems. |
D. I. Korabelnikov; A. I. Lamotkin; I. A. Lamotkin; | FARMAKOEKONOMIKA. Modern Pharmacoeconomics and … | 2026-07-06 |
| 143 | Classification of Songs in Spanish with LLMs: An Analysis of The Construction of A Dataset Through Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The accuracy metric reached an impressive 96.34%, highlighting the effectiveness of this approach in song sentiment analysis. This study underscores the importance of understanding emotions in songs and offers practical solutions to enhance the capabilities of language models in this task. |
Tania Alcántara; Omar García-Vázquez; Mayte H Laureano; Grigori Sidorov; | Journal of Intelligent & Fuzzy Systems: Applications in … | 2026-07-06 |
| 144 | Comparative Evaluation of Transformer-Based Models for Plain Language Classification in Hungarian Legal–Administrative Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents the first systematic evaluation of transformer-based models for sentence-level Plain Language classification in Hungarian tax administrative texts. |
István Üveges; | Electronics | 2026-07-06 |
| 145 | Temporal and Linguistic Enrichment for Abstractive Text Summarization Using Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we present a lightweight hybrid abstractive summarization model that enhances temporal awareness and linguistic flexibility. |
KHALED ABDALGADER et. al. | Discover Artificial Intelligence | 2026-07-05 |
| 146 | Transformer-Based NLP for Construction Contract Clause Classification: Implications for Sustainable Construction Project Governance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a text classification framework integrating transformer-based contextual embeddings (BERT, ALBERT, RoBERTa, and DistilBERT) with machine learning and deep learning models (RNN, GRU, and LSTM) to analyze FIDIC and JCT contract provisions. |
Anıl Demircan; Latif Onur Uğur; | Sustainability | 2026-07-03 |
| 147 | A Lightweight Infrared Thermal Image Recognition Network with Thermal Noise Suppression and Dynamic Feature Fusion for Transformer Oil Leakage Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Detecting transformer oil leakage from infrared thermal images is challenging due to complex thermal backgrounds, low contrast, and sensor noise. To address these issues, this study proposes IR-OilNet, a lightweight CNN–Transformer fusion network designed for real-time infrared leakage detection in inspection robots. |
WENBI TAN et. al. | Processes | 2026-07-03 |
| 148 | Conditional Co-Ablation: Recovering Self-Repair Backups in Transformer Circuits Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Such first-order scoring is natural when component importance is additive, but becomes misleading when a transformer self-repairs: after a primary component is removed, a dormant backup can take over, muting the primary’s measured effect while the backup itself appears irrelevant on the intact model. We recast this failure as a recovery task, conditional circuit completion, and introduce Conditional Co-Ablation (CoAx), a label-free, output-grounded score that asks how much each remaining unit’s ablation effect grows once a primary set has been removed. |
Zhiren Gong; Zihao Zeng; Chau Yuen; Wei Yang Bryan Lim; | arxiv-cs.LG | 2026-07-02 |
| 149 | Echoes of Unrest: A Multimodal NLP Framework for Early Warning of Fake News and Violence-Driven Mob Activity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Incidents in South Asia and elsewhere demonstrate how false information disseminated via platforms such as Facebook and WhatsApp can trigger real-world harm, often spreading faster than fact-checking efforts can respond. To address this challenge, this chapter presents a multilingual, multimodal Natural Language Processing (NLP) framework for early detection of misinformation and violence-prone dynamics. |
MD. MARUF BANGABASHI et. al. | arxiv-cs.CL | 2026-07-02 |
| 150 | A Scoping Review of Applications of Natural Language Processing for Chronic Pain Research Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Earlier studies focused on rule-based methods, while recent work adopted transformer-based models (e.g., BERT, RoBERTa, BioBERT) and LLMs (e.g., GPT-3.5 for zero- or few-shot learning frameworks). |
Swati Rajwal; Selen Bozkurt; Jeanmarie Perrone; Anne Marie McKenzie-Brown; Abeed Sarker; | Discover Artificial Intelligence | 2026-07-02 |
| 151 | A Comparative Evaluation of Deep Learning and Rule-Based Models for Sentiment Analysis of 5G/6G Public Discourse on Social Media Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to provide a domain-specific empirical evaluation of sentiment analysis models by examining classification performance, deployment-oriented inference efficiency, and lightweight domain adaptation. |
Hangliang Ding; Jinfeng Li; | Big Data and Cognitive Computing | 2026-07-02 |
| 152 | Developing A Natural Language Processing System Using Transformer-based Models for Adverse Drug Event Detection in Electronic Health Records Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Two data processing methods, window-based and split-based approaches, were compared to identify the optimal processing method. |
Jingyuan Wu; Xiaodi Ruan; Elizabeth McNeer; Katelyn M. Rossow; Leena Choi; | PLOS One | 2026-07-01 |
| 153 | On Reversibility As Language Model Behavioral Property in Parametric Knowledge Editing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce an operational definition of reversibility as model’s property, and empirically investigate it by comparing two prominent editing methods, ROME and MEMIT, across models of different scales commonly used in knowledge editing tests (GPT-2 XL and GPT-J 6B) and using a benchmark derived from CounterFact. |
Emanuele Caddeo; Manuela Sanguinetti; Maurizio Atzori; | Applied Sciences | 2026-07-01 |
| 154 | Understanding The Inner Workings of Large Language Models in Medicine Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Background Large language models (LLMs) are increasingly influencing medical practice, education, and research. Their responsible integration into healthcare requires expertise in … |
Georg Fuellen; Hans Jarchow; Johann-Christian Põder; | F1000Research | 2026-07-01 |
| 155 | Cross-lingual Relation Extraction with Large Language Models: Zero-Shot, Few-Shot, and Fine-Tuned Evaluation on Romanian Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We investigate the feasibility of cross-lingual RE for Romanian by combining automatic dataset translation with large language model (LLM) inference. |
Dragos-Mitrut Vasile; Elena-Simona Apostol; Stefan-Adrian Toma; Adrian Paschke; Ciprian-Octavian Truica; | arxiv-cs.CL | 2026-06-30 |
| 156 | Detecting and Correcting Factual Error in LLM Text Series Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The evaluation of the generated summary is the key focus in this thesis. |
Anisha Soni Anisha Soni; Miss Pratibha Tiwari Miss Pratibha Tiwari; | International Scientific Journal of Engineering and … | 2026-06-30 |
| 157 | Memory-Efficient Probabilistic Neuro-Symbolic Integration for Explainable Natural Language Inference Using Transformer-Based Foundation Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective: This study aims to present a memory-optimized probabilistic neuro-symbolic hybrid architecture that unifies transformer-based neural networks with logic-based symbolic reasoning systems. |
Zahraa Sameer Ibrahim; Haedar Ahmed Mukhef; Hayder Hasan Ali; | Al-Mustansiriyah Journal of Science | 2026-06-30 |
| 158 | Applying Large Language Models to Spam Detection in The Kazakh Low-resource Language Setting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract In this paper, we test the efficacy of multilingual transformer-based large language models for spam classification tasks on the low-resource Kazakh language. |
Kumisbek Mukhammed-Ali; Shormakova Assem; | Scientific Reports | 2026-06-30 |
| 159 | Analyzing Medium and Long Text Indonesian Tourism Feedback Using Topic Modeling and Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Background Problems: This study examines the challenge of identifying the most suitable topic modeling and sentiment analysis techniques for analyzing medium- and long-text feedback in the Indonesian tourism context. |
Sulisetyo Puji Widodo; Isnaeni Noviyanti; | Jurnal Aplikasi Statistika & Komputasi Statistik | 2026-06-30 |
| 160 | Representation As A Bottleneck for Mechanistic Interpretability: The Manifestation Unit Protocol Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The schema absorbs attention-head primitives without modification, set-recovers known IOI circuit members under retrieval-budget-matched controls, and reveals an irreducible two-field core (S+R) with remaining fields either redundant or actively interfering. We present this as schema infrastructure for mechanistic interpretability rather than frontier-scale validation. |
Hussein Chouman; Wataru Sasaki; Tomokazu Matsui; Hirohiko Suwa; Keiichi Yasumoto; | arxiv-cs.LG | 2026-06-30 |
| 161 | Integrating Sentiment Analysis, Text Mining and Predictive Clustering for Resilient Supply Chain Management in Industry 5.0: A Systematic Literature Review and Future Research Agenda Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: An integrated approach combining web scraping, natural language processing and unsupervised learning can transform weak digital signals into operational early warnings without relying solely on internal transactional data. This study constructs a systematic literature review of work integrating these methods in supply chain management for 2019-2026, following the PRISMA 2020 protocol. |
Singgih Saptadi; Wiwik Budiawan; I Gede Indra Aryasa; | World Journal of Advanced Research and Reviews | 2026-06-29 |
| 162 | Intelligent Arabic News Classification Systems Using AraBERT Transformer for Digital Media Engineering Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a transformer-based framework for Arabic news classification centered on fine-tuned AraBERT, a bidirectional encoder pre-trained exclusively on large-scale Arabic corpora. |
Rehab Ahmed; Omar A. Alkhudaydi; Hussain A. Almasabi; | Engineering Systems and Intelligent Technologies (ESIT) | 2026-06-29 |
| 163 | Linguistic Signatures for Enhanced Emotion Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines whether linguistic features can serve as reliable interpretable signals for emotion recognition in text. |
Florian Lecourt; Madalina Croitoru; Konstantin Todorov; | www | 2026-06-29 |
| 164 | When Transformers Learn impossible Languages, What Do They Learn? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Recent work suggests that transformer language models show a bias towards human languages over unnatural (impossible) languages argued to be unacquirable by humans. |
Ram Janarthan; Coleman Haley; Sharon Goldwater; | arxiv-cs.CL | 2026-06-29 |
| 165 | Graph-Enhanced Transformer for Cross-Domain Sentiment Analysis: Integrating RoBERTa with Graph Attention Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a hybrid RoBERTa–graph attention network (GAT) framework that integrates transformer-based contextual embeddings with graph-based relational learning. |
Moteechand Patel; Abhinav Shukla; Pritendra Kumar Malakar; R. Kanesaraj Ramasamy; Parul Dubey; | Future Internet | 2026-06-29 |
| 166 | HomonymSenseNet: A Context-Aware Transformer Framework with Dynamic Sense Memory and Adaptive Negative Sampling for Semantic Representation Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The approach learns highly discriminative contextual embeddings while maintaining low-frequency semantic senses that are under-represented during training. |
S Subi; B. Shanthini; | VFAST Transactions on Software Engineering | 2026-06-28 |
| 167 | Exploring The Cryptographic Limits of Transformer Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In recent work it has been shown that colluding AI agents can use steganographic methods to exchange malicious information. |
Stefan Domunco; Andis Draguns; Philip Torr; Isaac Robinson; Christian Schroeder de Witt; | arxiv-cs.CR | 2026-06-28 |
| 168 | Model Internal Sleuthing: Finding Lexical Identity and Inflectional Features in Modern Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We find a consistent pattern: inflectional features are linearly decodable throughout the model, while lexical identity is prominent early but increasingly weakens with depth. |
Michael Li; Nishant Subramani; | acl | 2026-06-27 |
| 169 | GiLT: Augmenting Transformer Language Models with Dependency Graphs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose Graph-Infused Layers Transformer Language Model (GiLT) which leverages dependency graphs for augmenting Transformer language models. |
Tianyu Huang; Yida Zhao; Chuyan Zhou; Kewei Tu; | acl | 2026-06-27 |
| 170 | Labeling Training Data for Entity Matching Using Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We investigate knowledge distillation for entity matching along the following dimensions: pair-selection strategy, teacher model, label post-processing method, and student model. |
Aaron Steiner; Christian Bizer; | arxiv-cs.CL | 2026-06-27 |
| 171 | On The (In-)Security of The Shuffling Defense in The Transformer Secure Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we show that the shuffling defense is not as robust as previously claimed. |
ZHENGYI LI et. al. | acl | 2026-06-27 |
| 172 | Learning from Textual Radiology Reports: A Benchmark Dataset for Coronary CT Angiography Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our analysis reveals that direct approaches, including state-of-the-art LLMs (GPT-4o, GPT-o3) and fine-tuned BERT models underperform on diverse real-world clinical data. To address these limitations, we propose a two-stage pipeline that decouples structuring from classification: an LLM-based parser normalizes heterogeneous reports into structured format, followed by fine-tuned BERT classification. |
SUDHARSHAN BALAJI et. al. | acl | 2026-06-27 |
| 173 | DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using Multi-VALUE’s linguistically-grounded transformations, we introduce D-CUBE (Dialectal Disinformation Detection Corpus), a core corpus component of comprising 195K samples derived from established disinformation benchmarks. |
JASON S LUCAS et. al. | acl | 2026-06-27 |
| 174 | OPTIMIZING SYNTACTIC-SEMANTIC RELATION EXTRACTION FOR THE KAZAKH LANGUAGE WITH TRANSFORMER ARCHITECTURES AND SYNTHETIC CORPORA Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In our work, we investigate how the use of synthetic data can partially compensate for the lack of linguistic resources. |
G. Bektemyssova; A. Sabdenov; R. Satybaldiyeva; A. Bykov; Binti Ali Nor’ashikin; | Herald of the Kazakh-British Technical University | 2026-06-27 |
| 175 | SenseRel: A Sense-Level Benchmark for Denotational and Connotational Meaning Relations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce the first sense-level benchmark, SenseRel, for modeling semantic relations between word senses, uniting denotational and connotational aspects of meaning. |
PIERLUIGI CASSOTTI et. al. | acl | 2026-06-27 |
| 176 | Enhancing Job Evaluation with Data Augmentation and Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this research, we propose to improve job evaluation by semi-automating a manual, time-consuming, and inconsistent process with text-based classification models. |
Samaneh Jalilian; Niels van Weeren; Mohammad Shokri; Thijmen Bijl; Suzan Verberne; | acl | 2026-06-27 |
| 177 | Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs’ Hallucinations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Building on these, we propose EAACD, an expert-aware adaptive contrast decoding that uses expert differences in MoE’s higher layers to mitigate hallucinations on QA tasks. |
XINYUE FANG et. al. | acl | 2026-06-27 |
| 178 | A Mechanistic Account of Attention Sinks in GPT-2: One Circuit, Broader Implications for Mitigation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Transformers commonly exhibit an attention sink: disproportionately high attention to the first position. We study this behavior in GPT-2–style models with learned query biases and absolute positional embeddings. |
Yuval Ran-Milo; Hila Ofek; Shahar Mendel; | acl | 2026-06-27 |
| 179 | Story Generation in Multi-Agent Systems: A DynamicCollaboration Framework Based on Reinforcement Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a Multi-Agent Dynamic Collaboration Framework for story generation, tackling narrative coherence, diversity, and adaptability. |
Yaolin Li; | Highlights in Science, Engineering and Technology | 2026-06-26 |
| 180 | Detection and Classification of Cyberbullying in Social Media Using Transformer-Based NLP Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a transformer-based Natural Language Processing (NLP) framework for the detection and classification of cyberbullying in social media text. |
Aniket Sawane Prof. Sagar Dhanake; Shradha Suse Romaan Shaikh; | International Journal of Advanced Research in Science … | 2026-06-26 |
| 181 | VASAE: Naming SAE Dictionary Directions with Vocabulary-Aligned Anchoring Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Vocabulary-Aligned Sparse Autoencoder (VASAE), a method that trains SAE features under vocabulary-aligned anchoring and assigns each feature an intrinsic token name: the token string whose embedding is nearest to that feature. |
Kairui Zhang; Ziwen Yu; Zahraa S. Abdallah; Martha Lewis; | arxiv-cs.CL | 2026-06-26 |
| 182 | Blinded By The Bot: Benchmarking GPT and Gemini Against Human Authors in Otolaryngology Reviews Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: For each topic, four AI‐generated reviews (GPT‐4.0 and Gemini 2.0; narrative and PRISMA‐style) and one human‐authored peer‐reviewed review were included, yielding a total of 10 manuscripts (8 AI‐generated, 2 human‐authored). |
SHOLEM HACK et. al. | World Journal of Otorhinolaryngology – Head and Neck Surgery | 2026-06-26 |
| 183 | Decoding The Predator’s Lexicon: A Deep Neural Framework for Disrupting Online Child Exploitation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research presents a unified, modular deep learning architecture engineered explicitly for the automated, low-latency recognition of predatory grooming trajectories within unstructured text streams. |
Dr. Sanjay Nag Dr. Sanjay Nag; Anjan Bera Anjan Bera; | International Journal of Creative and Open Research in … | 2026-06-25 |
| 184 | Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper describes team HSA_CORAL’s submission to the FinCausal 2026 shared task on extracting cause-effect relations from financial narratives via extractive question answering in English and Spanish. |
Akash Kumar Gautam; Serhii Hamotskyi; Christian Hänig; | arxiv-cs.CL | 2026-06-25 |
| 185 | Cross-Platform Generalisation Failure in Mental Health Natural Language Processing: A Five-Axis Fairness Audit of Transformer Models on Social Media Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce the Cross-Platform Fairness Evaluation (CPFE) framework — a five-axis audit protocol covering discriminative performance, calibration, statistical significance, prediction equity, and attribution stability — and apply it to four transformer models (BERT, RoBERTa, Emotion-DistilRoBERTa, GoEmotions-RoBERTa) trained on a Kaggle mental health corpus (n=35,556) and evaluated on Reddit (n=6,257) and Twitter (n=2,883) test sets with emotion labels mapped to clinical proxies. |
Rajveer Singh Pall; Sameer Yadav; | arxiv-cs.CL | 2026-06-25 |
| 186 | Aspect-Aware Sentiment Analysis of Code-Mixed Amazon Indian Reviews Using Roberta Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The paper proposes an aspect-based sentiment analysis tool that uses an XLM-RoBERTa model in a Django web application for code-mixed Amazon product reviews in India. |
Prof. D. V. Mehta Prof. D. V. Mehta; Shriraj Ranaware Shriraj Ranaware; Rutuja Kharat Rutuja Kharat; Karishma Bargaje Karishma Bargaje; Nikita Khude Nikita Khude; | International Journal of Creative and Open Research in … | 2026-06-25 |
| 187 | Transformer-Based Classification of Bacterial Raman Spectra with LOOCV Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, a transformer-based approach was systematically evaluated using a nested leave-one-replicate-out cross-validation framework and compared with conventional machine-learning pipelines combining PCA or ICA with LDA, SVM, and Random Forest classifiers. |
Jamile Mohammad Jafari; Thomas Bocklitz; | arxiv-cs.LG | 2026-06-25 |
| 188 | TH-GNN: Heterogeneous Temporal Graph Neural Networks for LLM-Agent Shilling Attack Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose TH-GNN, a heterogeneous temporal graph neural network with a two-layer Heterogeneous Graph Transformer backbone that applies per-type and per-relation attention augmented with learnable sinusoidal temporal encodings on every edge. |
Shivam Swarup; Divya Prakash Shrivastava; Rakesh Thakur; | arxiv-cs.CL | 2026-06-24 |
| 189 | Analyzing The Impact of KV Representation Compression on Explainability in Lightweight Transformer-based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed approach achieves compression ratios of 2.46–3.05 while maintaining high representational similarity (cosine similarity > 0.994). |
Misun Lee; Yeonghyeon Gu; | Scientific Reports | 2026-06-24 |
| 190 | An AI-Based Hybrid Model to Identify People Who May Be Depressed Based on Their Anonymous Posts on Social Media Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research proposes a Hybrid Explainable Multimodal Deep Learning Framework (HEMDL) for early detection of depressive tendencies in anonymous social media text. |
Dr. Santosh Gaikwad Dr. Santosh Gaikwad; Shrijay Ramdas Kale Shrijay Ramdas Kale; | International Scientific Journal of Engineering and … | 2026-06-24 |
| 191 | Human Rights Case Analysis Using Ai and Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This project leverages AI techniques—specifically Transformer-based Legal-BERT and Doc2Vec combined with Support Vector Machine (SVM)—to automate the analysis of human rights case documents. |
Dipali S. Jadhav; Dr. N. R. Wankhade; Santhosh R. Agrawal; Ashwini Gaikwad; Priyanka U. Mandlik; | International Research Journal on Advanced Engineering Hub … | 2026-06-24 |
| 192 | Optimizing Abstractive Summarization With Fine-Tuned PEGASUS Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Abstractive text summarization is the technique of generating a short and concise summary comprising the salient ideas of a source text without making a subset of the salient … |
S. Rafi; Naimur Rahman; Kazi Nazibul Islam; Hamreen Ahmad; Farig Sadeque; | ArXiv | 2026-06-24 |
| 193 | Lightweight Transformer Models for On-Device Fault Detection: A Benchmark Study on Resource-Constrained Deployment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a benchmark comparing traditional ML methods (Random Forest, XGBoost, SVM, Logistic Regression) against lightweight transformer architectures (DistilBERT, TinyBERT-6L, TinyBERT-4L, MobileBERT) for binary fault detection across three public datasets: NASA C-MAPSS turbofan degradation, SECOM semiconductor manufacturing, and UCI AI4I 2020 predictive maintenance. |
Disha Patel; | arxiv-cs.LG | 2026-06-23 |
| 194 | P30 Using Large Language Models for The Annotation of Skin Single-cell RNA Sequencing Datasets Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Introduction and aims Large language models (LLMs) have recently been shown to accurately annotate cell types in single-cell RNA sequencing (scRNAseq) datasets, using differentially expressed genes (DEGs). |
Tanzil Rujeedawa; Joseph Inns; Richard Gallon; Neil Rajan; | British Journal of Dermatology | 2026-06-23 |
| 195 | Can Scale Save Us From Plasticity Loss in Large Language Models? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To determine whether loss of plasticity remains a problem in the modern transformer-based LLM paradigm, we study plasticity loss in GPT-style Transformer models trained on a multilingual continual learning problem. |
J. Fernando Hernandez-Garcia; Tomás Figliolia; Beren Millidge; | arxiv-cs.AI | 2026-06-23 |
| 196 | An LLM-based Two-Stage Transformer Framework for Cross-Domain Bearing Fault Diagnosis with Limited Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a knowledge-guided two-stage transfer learning framework that employs a lightweight GPT-2-style Transformer with causal self-attention for hierarchical feature extraction from vibration signals, establishing explicit pathways where pre-trained encoder weights and fault prototype embeddings serve as knowledge carriers from multi-source pre-training to target adaptation. |
JINGHAN WANG et. al. | arxiv-cs.LG | 2026-06-23 |
| 197 | News Classification at Scale: A Web Scraping and BERT-Based Approach Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The quick growth of digital news media resulted in the accumulation of a tremendous amount of unstructured text data. The manual collection and classification of such data is not … |
Ms. Khushali Domadiya; | International Journal of Advanced Research in Science … | 2026-06-22 |
| 198 | The Energy Consumption of Transformer Fine-Tuning: A Roofline-Inspired Scaling Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a framework for modeling the energy consumption of Transformer training on multiple GPUs. |
Mansour Zoubeirou a Mayaki; | arxiv-cs.LG | 2026-06-22 |
| 199 | Writing Creativity, Cohesion, and Formal Linguistic Competence in LLMs: A Comparative Evaluation Based on English and Chinese Continuation Writing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Identifying such difference would be beneficial to teachers to adopt LLMs in writing practice. This study addresses this issue by comparing ChatGPTs and C-dominant LLMs in story continuation tasks, with a focus on cohesion, creativity, and formal linguistic competence. |
YAO ZHANG et. al. | PLOS One | 2026-06-22 |
| 200 | Analisis Sentimen Komentar YouTube Terhadap Korporasi MBG: Studi Perbandingan Metode BERT Dan Support Vector Machine (SVM) Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Program Makan Bergizi Gratis (MBG) merupakan program pemerintah yang memperoleh berbagai tanggapan dari masyarakat melalui media sosial, khususnya YouTube. Komentar yang … |
Najma Alfisyahrina; Riza Ibnu Adam; Aries Suharso; | IKRA-ITH Informatika : Jurnal Komputer dan Informatika | 2026-06-22 |
| 201 | Advancing Low-Resource African Language Technologies: Morphological Feature Integration for Kiswahili Question Answering Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study extended XLM-RoBERTa with explicit representations of 17 Kiswahili morphemes to build a morphologically-enhanced architecture, encoded as multi-hot vectors and introduced through learned projection layers and the pre-trained encoder being frozen to maintain multilingual knowledge. |
Collins S. Wanjala; Lilian Wanzare; Calvins Otieno; | International Journal of Computer Trends and Technology | 2026-06-22 |
| 202 | Sub-Billion, Super-Frontier: Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Large language models (LLMs) achieve strong relation extraction (RE), but their computational demands and reliance on proprietary APIs limit deployment in resource-constrained or privacy-sensitive settings. We investigate how far small language models (SLMs) can close this gap across general-domain and literary text. |
Despina Christou; Grigorios Tsoumakas; | arxiv-cs.CL | 2026-06-21 |
| 203 | Learning-based Orchestration for Low-latency AI Deployment in Hybrid Cloud–edge Platforms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce and analyze a resource-conscious deep learning-based scheduling system for managing the deployment of AI models on distributed cloud edges. |
Ahmed Albugmi; | Scientific Reports | 2026-06-20 |
| 204 | Rancang Bangun Sistem Penerjemah Multibahasa Daerah Maluku Utara Dengan Integrasi GPT Sebagai Pelestarian Budaya Lokal Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research discusses the development of a web-based regional multilingual translation system that integrates the Rule-Based Machine Translation (RBMT) method and GPT to translate Indonesian into Ternate, Makian Dalam, and Galela languages using 5,572 vocabulary entries based on a digital dictionary, where RBMT is used for rule-based word matching and GPT is used to improve sentence structure to make the translations more natural and contextually appropriate. |
Anggiah Salim; Abdul Haris Muhammad; Sakina Sudin; | Jurnal Komputer Antartika | 2026-06-20 |
| 205 | When Do LLMs Replace Fine-Tuned NLU? A Decision Framework for Intent Detection in Production Conversational Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: On full ATIS and CLINC150 we compare a fine-tuned RoBERTa, a TF-IDF+logistic-regression baseline, sentence-embedding kNN, and Claude Haiku zero-shot, reporting bootstrap 95% confidence intervals and paired significance tests. |
Carson Rodrigues; Oysturn Vas; | arxiv-cs.CL | 2026-06-19 |
| 206 | GPT-4.1 and Llama 3.3 70 Fail to Detect Clinically Relevant Errors in Radiology Reports in Zero-shot Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Applied to GPT-4.1 and Llama 3.3 70B in zero-shot settings, the framework reveals a performance gap between pattern-based and reasoning-dependent error detection that warrants investigation across additional models and optimization strategies. |
TUGBA AKINCI D’ANTONOLI et. al. | European Radiology | 2026-06-19 |
| 207 | Large Language Models As Data-driven Engines for Benchmarking Preventive and Clinical Knowledge in Chinese Dental Examinations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods This study evaluated GPT–4o, GPT–4.5, and DeepSeek-R1 using 300 standardized dental competency items from institutional examinations (2023–2025). |
YONG ZENG et. al. | Frontiers in Oral Health | 2026-06-19 |
| 208 | Keyless Attention: Value-Space Routing and Value-Only Caching for Efficient Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose Keyless Attention, an attention mechanism that eliminates the key projection entirely, operating over queries and values only. |
Xin Gao; | arxiv-cs.CL | 2026-06-19 |
| 209 | Hybrid Text Summarizer Using SBERT Extractive Filtering and Fine-Tuned BART Abstractive Generation on A Custom Dataset Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a hybrid deep learning framework that integrates the complementary strengths of both paradigms. |
Aadarsha Chaulagain; Aaditya Bhandari; Bishwa Karna; Jagadish Pokharel; Binod Wosti; | International Journal on Engineering Technology | 2026-06-18 |
| 210 | OpenEvidence Performs at Similar Levels Compared to Current and Previous GPT Models on Orthopedic Training and Education Questions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: METHODS We conducted an analysis of orthopedic board-style questions obtained from Orthobullets, a widely used educational platform for orthopedic resident education and board preparation. |
KASHIF JAVID et. al. | World Journal of Orthopedics | 2026-06-17 |
| 211 | Contextual Semantic Classification of Trafficking-Related Advertisements Using DistilBERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a lightweight transformer-based framework for identifying potential trafficking-related risk indicators in publicly accessible online advertisements using contextual semantic classification and leakage-aware evaluation. |
Bakhita Salman; Muneeb Yassin; Jose Leonidez; | Information | 2026-06-17 |
| 212 | EMTF: An Explainable Transformer-Based Framework for Drug Review Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The main aim of this paper, to improve healthcare sentiment analysis, was to integrate transformer-based deep learning with explainable AI techniques. |
Venkataramana Battula; Rohith Kollu; Srichandana Abbineni; V. Srinadh; | International Journal of Drug Delivery Technology | 2026-06-17 |
| 213 | The Utility of Large Language Models to Assist With Emergency Triage Decisions Within Otolaryngology Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: DeepSeek and the otolaryngology resident demonstrated intermediate performance, while Grok and the emergency clinicians performed lowest. Group‐level analyses showed no significant difference between the large language model and otolaryngology cohorts; both were rated higher than emergency clinicians in this sample. |
SHOLEM HACK et. al. | Otolaryngology–Head and Neck Surgery | 2026-06-17 |
| 214 | Explaining Attention with Program Synthesis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A longstanding goal of research on interpretable deep learning is to replace opaque neural computations with human-meaningful symbolic descriptions. |
Amiri Hayes; Belinda Li; Jacob Andreas; | arxiv-cs.LG | 2026-06-17 |
| 215 | HLS-GPT: A Generative Pretrained Transformer (GPT) for Continental-Scale NASA Harmonized Landsat and Sentinel-2 (HLS) Reflectance Reconstruction Across All Bands on Arbitrary Dates Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present HLS-GPT, a large-scale generative pretrained Transformer model for reconstructing NASA Harmonized Landsat Sentinel-2 30 m surface reflectance for all bands, any date, and any pixel location. |
Junjie Li; Hankui K. Zhang; David P. Roy; | arxiv-cs.CV | 2026-06-16 |
| 216 | Bifrost: Hybrid TEE-FHE Inference for Privacy-Preserving Transformer and LLM Serving Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present Bifrost, a hybrid TEE-FHE serving architecture in which secrets are provisioned only to an attested CPU TEE, while the accelerator, device memory, driver/runtime stack, and host software remain outside the trusted computing base. |
Chenghao Chen; Kailun Qin; Xiaolin Zhang; Chi Zhang; Dawu Gu; | arxiv-cs.CR | 2026-06-15 |
| 217 | Data-Driven Decoding of Russell’s Circumplex Model of Affect Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Furthermore, in a zero-shot setting using generic text embeddings, projected fine-grained emotion terms fall close to their established human-mapped coordinates. Our contribution is a novel, data-driven framework for validating emotion models, demonstrating that Russell’s circumplex structure is intrinsically encoded in the embeddings of these modalities rather than being solely an artifact of human labeling, thereby bridging the gap between psychological theory and representation learning. |
Amdjed Belaref; Samir Sadok; Zineb Noumir; Renaud Seguier; | arxiv-cs.CL | 2026-06-15 |
| 218 | Privacy from Symmetry: Orthogonally Equivariant Transformers for LLM Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose an orthogonal obfuscation procedure in which the client multiplies embeddings by a secret orthogonal matrix before transmission. |
Alexander Yukhimchuk; Andrey Shulga; Mladen Kolar; Martin Takáč; | arxiv-cs.LG | 2026-06-15 |
| 219 | Comparative Evaluation of Machine Learning Methods for Protecting LLMs from Prompt Injection Attacks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our research contributes to improving the robustness of AI-driven security systems against Prompt Injection by presenting a comparative study that can help in selecting and deploying ML-based defense against Prompt Injection attacks. |
NAZARII DZHALIUK et. al. | International Journal of Information Security | 2026-06-15 |
| 220 | LiFT: Local Search Via Linear Programming for Overfitting-Controlled Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a Linear Programming (LP)-based local search framework for fine-tuning pretrained transformer models with explicit control against overfitting. |
Abhishek Shukla; Anikeit Khanna; Ankur Sinha; Faiz Hamid; | arxiv-cs.LG | 2026-06-15 |
| 221 | Transformer-Based Brain MRI Classification for Early Alzheimer’s and Parkinson’s Disease Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a transformer-based deep learning system for automated three-class classification of AD, PD, and cognitively normal subjects from structural brain MRI scans. |
Gladiss Merlin N R; Donthamsetty Sai Satwika; A.S. Gayathri Prasanna; | International Research Journal on Advanced Engineering Hub … | 2026-06-15 |
| 222 | DinoFlow : Self‐supervised Pretraining in Flow Cytometry Enables Accurate Detection of Common Hematopathological Disorders Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To maximize clinical utility, we develop a method that enables identification of multiple common disorders and quality indicators. |
Brendan O’Fallon; Muir Morrison; Mattia Medina Grespan; Nicholas C. Spies; David P. Ng; | Cytometry Part B: Clinical Cytometry | 2026-06-15 |
| 223 | Compositional Reasoning Depth Predicts Clinical AI Failure: Empirical Evidence Consistent with Transformer Compositionality Limits in Electronic Health Record Question Answering Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Motivated by theoretical results on transformer compositionality limits, we introduce a pre-specified hop-count taxonomy — the number of distinct reasoning steps required to answer a clinical question from an EHR — as a principled predictor of model failure. |
Sanjay Basu; | arxiv-cs.CL | 2026-06-15 |
| 224 | Attention Alignment Between Humans and Vision-Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Vision-language models implement both, allowing us to treat each component as a separable hypothesis about what drives where we look. |
ISAAC R. CHRISTIAN et. al. | arxiv-cs.CV | 2026-06-15 |
| 225 | The Reservoir Attention Network: Cross-Pass State in Pretrained Transformers Via Content-Addressable Reservoir Injection Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: A feasibility and dynamics study of the Reservoir Attention Network (RAN), an architecture that injects a fixed, randomly-initialized reservoir into the mid-layer attention of a … |
Emma Leonhart; | arxiv-cs.LG | 2026-06-14 |
| 226 | Performance of GPT-based Large Language Models in Hepatocellular Carcinoma Stratification: Liver Function Assessment, BCLC Staging, and Treatment Recommendations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Large language models (LLMs) like GPT have been proposed to support complex clinical decision-making. |
MAX MASTHOFF et. al. | Scientific Reports | 2026-06-12 |
| 227 | Cross-Lingual Sentiment Classification in Sustainable Mobility: A Zero-Shot Domain Transfer Evaluation Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: By releasing the annotated multilingual dataset and code publicly, this work provides a reproducible exploratory evaluation framework for annotation-scarce, domain-specific multilingual NLP. |
Ainhoa Serna; Jon Kepa Gerrikagoitia; Juan de Oña; | AI | 2026-06-12 |
| 228 | Multi-Label Toxic Comment Detection Using BERT-Based NLP and Machine Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a BERT-based toxic comment detection system using Natural Language Processing (NLP) and deep learning techniques for multi-label classification of online comments. |
Binita Adhikari; Pratistha Sapkota; Rajad Shakya; | International Journal on Engineering Technology and … | 2026-06-12 |
| 229 | How Linear Is A Transformer Feed-Forward Block? Per-Block Linear Recoverability Is Learned, Not Architectural Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We treat each FFN as a position-wise input-to-output map and split it into the exact least-squares linear approximation plus a residual. |
Stuart Whipp; | arxiv-cs.LG | 2026-06-12 |
| 230 | Different Layers, Different Manifolds: Module-Wise Weight-Space Geometry in Transformer Optimization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we ask whether different transformer modules prefer different manifold geometries. |
Kirato Yoshihara; | arxiv-cs.LG | 2026-06-11 |
| 231 | GIMeT: A Multimodal Transformer Framework for Automated Gastrointestinal Disease Diagnosis Using Endoscopic Images and Physiological Signals Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: These limitations often lead to variability in diagnosis and increased clinician workload. To address these challenges, we propose GIMeT (GastroIntestinal Multimodal Transformer), a novel deep learning framework that integrates endoscopic imaging, physiological time-series data, and optional clinical text information for automated GI disease classification and lesion segmentation. |
Qianyun Lin; Annie Cheung; | Journal of Computer Science and Frontier Technologies | 2026-06-10 |
| 232 | 3-Key-Input: Exploring The Theoretical Minimum Keys for Text Entry Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: How far can we reduce the number of physical keys if we endow an ambiguous keyboard with modern language models? |
Naoki Kimura; | arxiv-cs.HC | 2026-06-10 |
| 233 | A Hybrid Transformer–XGBoost Framework with SHAP Explainability for Multimodal Student Performance Prediction in Higher Education Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Accurate student performance prediction is critical for early interventions and data-driven educational decisions; thus, this study develops a hybrid framework combining Transformer-based deep learning and XGBoost to predict academic outcomes using textual data, LMS interaction logs, assessment scores, and demographics. |
Ms. Amal Mohammed Hassan Nouri; Assoc. Prof. Dr. Fakhreldin Saeed; | Journal of Arabian Peninsula Centre for Medical and Applied … | 2026-06-10 |
| 234 | Large-Scale Evaluation of Five Large Language Models in Anesthesia Decision-Making for Hip Fracture Surgery Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: METHODS: We evaluated five general-purpose LLMs (DeepSeek 3.2, Gemini 2.5 Flash, GPT-5, GPT-5 mini, GPT-5 nano) using 216 standardized hip fracture surgery vignettes crossing six surgery types, two sexes, and 18 patient variables. |
ROBERT CHEN et. al. | Anesthesia & Analgesia | 2026-06-10 |
| 235 | SpikeDecoder: Realizing The GPT Architecture with Spiking Neural Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose SpikeDecoder, a fully SNN-based implementation of the Transformer decoder block, for applications in natural language processing. |
Claas Beger; Florian Walter; Alois Knoll; | arxiv-cs.NE | 2026-06-10 |
| 236 | Understanding Transformer-Based Classifications of Medical Text Using A Large Language Model for The Attribution of Feature Importance: Proof-of-Concept Algorithm Development and Validation Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Generative large language models may offer a novel approach to generating interpretable, context-aware explanations as autonomous agents. |
FANGWEN ZHOU et. al. | JMIR Medical Informatics | 2026-06-10 |
| 237 | I Understand How You Feel: Enhancing Deeper Emotional Support Through Multilingual Emotional Validation in Dialogue System Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: For timing detection, we propose MEGUMI, a Multilingual Emotion-aware Gated Unit for Mutual Integration, that fuses frozen XLM-RoBERTa semantics with language-specific emotion encoders via cross-modal attention and gated fusion. |
Zi Haur Pang; Yahui Fu; Koji Inoue; Tatsuya Kawahara; | arxiv-cs.CL | 2026-06-10 |
| 238 | Enhancing Fake News Detection in Sport: A Deep Learning Approach to Reliable Sport Information Dissemination Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article proposes a multi-layer embedding fusion framework for detecting fake news in sports by integrating complementary representations from multiple transformer-based language models, including Robustly Optimized Bidirectional Encoder Representations from Transformers (RoBERTa), Decoding-Enhanced Bidirectional Encoder Representations from Transformers version 3 (DeBERTa-v3), and Large Language Model Meta AI 2 (LLaMA2). |
Jianxiong Gao; Jianwei Gao; Huiling Zou; | PeerJ Computer Science | 2026-06-10 |
| 239 | Recoverable But Not Stationary:Local Linear Structures in Weights and Activations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We develop random search theory with a Gaussian local-linear theorem that justifies the effectiveness of random parameter search even in very high dimensions. |
Irina Piontkovskaia; Sergey Nikolenko; | arxiv-cs.LG | 2026-06-09 |
| 240 | Early Comparative Evaluation of Transformer Models for Multilingual Software Vulnerability Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents an early comparative evaluation of BERT, RoBERTa, and CodeBERT for binary vulnerability detection across HTML, Python, JavaScript, and PHP using the CVEFixes dataset and language-wise three-fold stratified cross-validation. |
Fiza Naseer; Javad Khan; Muhammad Yaqoob; Alexios Mylonas; | arxiv-cs.SE | 2026-06-09 |
| 241 | Trajectory Geometry of Transformer Representations Across Layers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Rather than probing for pre-specified features, we characterize trajectory geometry using five metrics computed directly in the ambient space: trajectory length, curvature, a semantic convergence index, layerwise cosine similarity, and representational stability. |
Vishal Pandey; Gopal Singh; | arxiv-cs.LG | 2026-06-08 |
| 242 | A Multi-Modal Transformer Model with GatedLSTM for Sarcasm Detection in Tweets UsingCross-Attention and Emoji Integration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed model utilizes RoBERTa for the contextual processing of textual content to generate contextualized text embeddings, whereas emojis are encoded using Emoji-BERT to capture emoji-specific semantic and emotional cuing. |
Shaikh Ambreen Mohd Ibrahim; Manoj M. Deshpande; Vijaykumar N. Pawar; | International Journal of Information Engineering and … | 2026-06-08 |
| 243 | An Efficient Algorithm for Streaming BPE Tokenization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce the concept of delay for a list of BPE merge rules, which corresponds to the amount of lookahead needed before tokens can be finalized. |
Konstantinos Mamouras; Angela W. Li; Yudi Yang; | Proceedings of the ACM on Programming Languages | 2026-06-08 |
| 244 | FrenchNews-7: Benchmarking Cross-Publisher French News Editorial Desk Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present FrenchNews-7, a cross-publisher France-based French-language news editorial desk classification benchmark combining a large multi-outlet corpus, a URL-derived seven-class taxonomy, and a fine-tuned CamemBERT classifier. |
Amr Sobhy; | arxiv-cs.CL | 2026-06-08 |
| 245 | Small Language Models Efficacy Prototyped for Oromo Word Sense Disambiguation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: DistilBERT provides the optimum trade-off between efficiency and performance, whereas XLM-RoBERTa excels with combined POS categorization. These findings demonstrate that by allowing semantic processing in low-resource languages, lightweight multilingual Transformer topologies offer a consistent, efficient, and context-sensitive approach for WSD tasks. |
Liyachew Edeti; Million Meshesha; Feda Negesse; | Discover Computing | 2026-06-07 |
| 246 | Beyond Self-Attention: Sub-Quadratic Vision Transformers for Fast Image Captioning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed model focuses on improving computational efficiency by restructuring the vision transformer architecture. |
Chiradeep Ghosh; Dakshina Ranjan Kisku; | arxiv-cs.CV | 2026-06-07 |
| 247 | Transformer Models in Digital Image Processing: A Systematic Review of Architectures and Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we conduct a comprehensive systematic review of Transformer architectures for digital image processing from 2020 to 2026, and we cover the key foundational models, such as Vision Transformer (ViT), Swin Transformer, DeiT and BEiT, and their numerous variants. |
Manar Abdulkareem Al-Abaji; Meaad Salih; Maher Khalaf Hussein; | Protek : Jurnal Ilmiah Teknik Elektro | 2026-06-06 |
| 248 | A Hybrid CNN-Transformer Architecture for Robust Multi-Feature Heart Disease Prediction Using Clinical and Electrophysiological Biomarkers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces CardioViT, a novel hybrid architecture that synergistically integrates Convolutional Neural Networks (CNNs) with Vision Transformer (ViT) modules to perform robust, multi-feature cardiac risk stratification. |
T. Sumathi; P. Jessie; Vijayalakshmi B.; Divya Vahini Suresh; M. Sakthivadivel; | International Journal of Drug Delivery Technology | 2026-06-06 |
| 249 | Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a pre-intervention screening framework for forecasting SAE steering side effects from feature statistics computed before steering. |
Evan Duan; | arxiv-cs.LG | 2026-06-06 |
| 250 | Research on Transformer Fault Diagnosis and Early Warning Technology Based on Digital Twins Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Conventional transformer operation and maintenance methods are often limited by fixed alarm rules and insufficient integration of monitoring information, making it difficult to achieve intelligent fault diagnosis and timely warning support. To address these issues, this paper proposes a digital twin-assisted framework for transformer fault diagnosis and early warning by integrating transformer digital modeling, online monitoring information, and dissolved gas analysis (DGA)-based intelligent diagnosis. |
Kaibo Hu; Lifeng Yu; Wuyin Gu; Kefeng Qian; Jiahe Sun; | Frontiers in Smart Grids | 2026-06-05 |
| 251 | DOC GPT: A Retrieval-Augmented Generation Based Intelligent Document Chatbot Using Local Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research presents DOC GPT, an intelligent Retrieval-Augmented Generation (RAG) based chatbot that allows users to interact with uploaded documents using local Large Language Models (LLMs). |
Abhishek R; Raviprakash DK; | International Journal for Research in Applied Science and … | 2026-06-05 |
| 252 | Accelerating Reproducible Research in Synthetic EHR Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, head-to-head comparison of existing generative models is hindered by disjointed codebases, incompatible data loaders, conflicting library dependencies, and inconsistent evaluation protocols. To address these gaps, we introduce a lightweight, end-to-end benchmarking framework for reproducible synthetic EHR evaluation, organized as a unified pipeline spanning data ingestion, standardized model training, and architecture-agnostic evaluation. |
Jalen Jiang; Chufan Gao; Ethan Rasmussen; Stephen Z. Xie; Jimeng Sun; | arxiv-cs.LG | 2026-06-05 |
| 253 | Enhancing Phishing Detection Using BERT and Graph Neural Network Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a novel hybrid deep learning architecture for phishing detection that integrates BERT and Graph Neural Networks through cross-modal attention fusion. |
Zainab Jibril Amedu; Prema Kirubakaran; Ridwan Kolapo; | International Journal of Innovative Science and Research … | 2026-06-05 |
| 254 | Event-Based Sentiment Analysis of Financial News Using Large Language Models: A Comprehensive Framework Integrating RAG, GNNs, and Multi-Agent Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a framework for event-based sentiment analysis of financial news that leverages Large Language Models (LLMs). |
Amit Kulkarni; Varun Dogra; | Information | 2026-06-05 |
| 255 | An Expanded Synthetic Conversation Dataset for Multi-Turn Smishing Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present COVA-X, an expanded dataset of 10,985 conversations spanning eight elder-targeted scam categories, produced by an improved generation pipeline addressing contamination, label mismatch, stage-direction bleed, and prompt-design failures from the first iteration. |
Carl Lochstampfor; Ayan Roy; | arxiv-cs.CL | 2026-06-04 |
| 256 | From Hand-Crafted Features to Large Language Models: A Comparative Evaluation of Android Malware Detection Paradigms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a rigorous, unified comparative evaluation of three methodological paradigms-classical machine learning, Transformer-based architectures, and generative Large Language Models (LLMs)-for static Android malware detection. |
Egemen Taşkın; İbrahim Alper Doğru; | Applied Sciences | 2026-06-03 |
| 257 | LG-Transformer: Learned-graph Transformer Framework Enabling Diverse Physicochemical Properties Prediction Toward Fuel Design Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Although artificial intelligence-based models demonstrate significant potential to accelerate fuel design, most existing methods cannot utilize the internal and external information within and between fuel molecules with interpretability, limiting their generalizability for diverse properties prediction. To address these challenges, a deep learning framework, the learned graph feature fusion Transformer (LG-Transformer), is proposed. |
JIABO ZHANG et. al. | Nature Communications | 2026-06-03 |
| 258 | Trajectory Dynamics in Language Model Hidden States Predict Human Processing Costs Beyond Surprisal Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce trajectory extrapolation error: at each word, we fit a linear trajectory to the preceding hidden states of a transformer language model and measure deviation from the extrapolated path. |
Elan Barenholtz; | arxiv-cs.CL | 2026-06-03 |
| 259 | Using Large Language Models to Support High Volume Application Review for An Undergraduate Research Program Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work-in-progress paper describes the development and initial deployment of a large language model (LLM)-based tool to assist in the evaluation of approximately 1,200 student Statements of Purpose (SoPs) for the SURF 2026 cycle at Purdue University. |
Varun Aggarwal; Kay Kobak; John Howarter; | arxiv-cs.CL | 2026-06-03 |
| 260 | AgenticRL: Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Deep reinforcement learning has shown strong potential for enabling autonomous robots to learn complex navigational tasks. However, its practical use still depends heavily on … |
Roohan Ahmed Khan; Yasheerah Yaqoot; A. Habel; Muhammad Ahsan Mustafa; D. Tsetserukou; | ArXiv | 2026-06-02 |
| 261 | A Unified Comparative Evaluation of Machine Learning, Deep Learning and GPT-2 for Suicide Ideation Detection from Social Media Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research examines various computer algorithms for detecting suicide-related content in social media text within a standardized experimental framework. |
Yasmeen Mohamed Saleh; Fahad Kamal Alsheref; Mahmoud Mohamed Bahloul; | Future Business Journal | 2026-06-02 |
| 262 | KITE: A Tri-Modal Transformer Integrating Text, Images, and Knowledge Graphs for Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we introduce KITE (Knowledge-Integrated Text-Image Encoder), a tri-modal fake news detection framework that jointly models textual, visual, and factual knowledge representations. |
Kevin Patel; Shashi Bhushan Jha; | arxiv-cs.LG | 2026-06-02 |
| 263 | Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We compare nine models across three paradigms — BART-MNLI, three Llama variants, three commercial frontier LLMs (Claude Haiku 4.5, Gemini Flash Lite 2.5, Claude Sonnet 4.6), and fine-tuned DistilBERT and RoBERTa — under universal and topic-specific label schemas. |
JooYoung Lee; Lin Tian; Angela Brillantes; Adriana-Simona Mihăiţă; Marian-Andrei Rizoiu; | arxiv-cs.CL | 2026-06-02 |
| 264 | Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Humanoid-GPT, a GPT-style Transformer with causal attention trained on a billion-scale motion corpus for whole-body control. |
ZEKUN QI et. al. | arxiv-cs.RO | 2026-06-02 |
| 265 | The Word and The Way: Strategies for Domain-Specific BERT Pre-Training in German Medical NLP Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present ChristBERT (Clinical- and Healthcare-Related Issues and Subjects Tuned BERT), a family of domain-specific German RoBERTa-based language models trained on a 13.5GB corpus of scientific publications, clinical texts, health-related web content, and translated clinical resources. |
Henry He; Johann Frei; Raphael Schmitt; | arxiv-cs.CL | 2026-06-02 |
| 266 | Leveraging BART to Assess CS1 C++ Programming Assignments Using Rubric-based Criteria Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, with the goal of producing grade predictions that better reflect instructor grading behavior than general-purpose LLMs. |
Kelsey Rainey; Jesse Roberts; | arxiv-cs.AI | 2026-06-02 |
| 267 | Transformer Models for Text Summarization: A Comparative Study of BART, BERT, and RoBERTa Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents a focused review of modern summarization techniques with an emphasis on transformer based models and large language models (LLMs), specifically BERT, RoBERTa and BART. |
Daisy Aptovska; Vinayak Elangovan; | arxiv-cs.CL | 2026-06-02 |
| 268 | DxPTA: An Architecture Design Space Exploration with Optical Dataflow-guided Strategy for HW/SW Co-Design of Photonic Transformer Accelerators Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Moreover, their manual design approach also requires huge design time to determine a suitable architecture for the targeted application, hence making this approach not scalable. To address these limitations, we propose DxPTA, a novel design space exploration methodology for enabling efficient hardware/software co-design of the appropriate PTA architecture that meets all constraints. |
Rachmad Vidya Wicaksana Putra; Solomon Micheal Serunjogi; Mahmoud Rasras; Muhammad Shafique; | arxiv-cs.AR | 2026-06-02 |
| 269 | Construction of A Comprehensive Dataset for Named Entity Recognition and Entity Linking in Algerian Dialectal Arabic Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a large-scale, multi-script dataset constructed through a novel hybrid methodology that integrates manual annotation of authentic texts, automated knowledge graph extraction from Wikidata, and rule-based synthetic generation. |
Wissem Bouarroudj; Adel Belbekri; Ilhem Djouablia; Hadjer Hanine Bouguettoucha; | Intelligent Data Analysis: An International Journal | 2026-06-01 |
| 270 | Fast Transformer Inference on ARM-Based HMPSoCs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we implement several new transformer kernels in ARM-CL to support native transformer execution. |
Hang Xu; Yixian Shen; Thanassis Giannetsos; Anuj Pathania; | arxiv-cs.AR | 2026-06-01 |
| 271 | Heterogeneous Mapping for Analog In-Memory Computing Accelerators: A Unified Workflow Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This letter classifies existing methods for partitioning DNN workloads across these resources by mapping granularity, optimization strategy, and model support, and distills them into a unified four-stage workflow. To demonstrate the workflow on a model class not yet addressed by existing methods, we apply its first two stages to GPT-2, producing the first AIMC-specific precision sensitivity profile for a decoder-only transformer. |
Corey Lammie; | arxiv-cs.AR | 2026-06-01 |
| 272 | Improving Hotel Review Rating Prediction with Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a novel dataset of 68,785 English hotel reviews from TripAdvisor (2014-2023) in Turkey. |
Ayhan Topçu; Mert Arda Asar; Günce Keziban Orman; | Sakarya University Journal of Computer and Information … | 2026-06-01 |
| 273 | PortBERT: Navigating The Depths of Portuguese Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In the present work, we introduce PortBERT, a family of RoBERTa-based language models for Portuguese, designed to balance performance and efficiency. |
Raphael Scheible-Schmitt; Henry He; Armando B. Mendes; | arxiv-cs.CL | 2026-06-01 |
| 274 | Scalable Sparse Transformer Accelerator With In-Memory Butterfly Zero Skipper and Local Attention Reusable Engine for Semi-Structured-Pruned NN Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Transformer-based large language models (LLMs) have achieved unprecedented advances across diverse AI tasks. However, their execution remains power-hungry, primarily due to the … |
Shiwei Liu; Jiangnan Yu; Peizhe Li; Feng Lin; Chixiao Chen; | IEEE Journal on Emerging and Selected Topics in Circuits … | 2026-06-01 |
| 275 | Context-Aware Sentiment Analysis Using Transformer-Driven Sarcasm Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed framework investigates advanced transformer architectures inclusive of BERT, RoBERTa, XLNet, DistilBERT, and contextual attention mechanisms to discover sarcastic expressions and improve sentiment prediction performance. |
Sudhir Kumar; | International Journal for Research in Applied Science and … | 2026-05-30 |
| 276 | Automated Fraud and Phishing Detection Through Natural Language Processing: A Deep Contextual Approach to Combating Generative Threats Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This failure is aggravated by the rise of consumer-grade generative artificial intelligence, which allows adversaries to launch highly sophisticated, grammatically perfect, and context-aware social engineering campaigns at an unprecedented scale. To bridge this defensive gap, this paper proposes an advanced, automated detection framework leveraging Natural Language Processing (NLP) paired with hybrid deep learning architectures. |
Ms. Meenu Verma; | International Journal for Research in Applied Science and … | 2026-05-30 |
| 277 | Artificial Intelligence-assisted Feedback in Pharmacology Education: A Pilot Evaluation of A Custom Generative Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods A prospective pilot study was conducted within a UK Physician Associate MSc programme between March and October 2025. |
Edward Stephenson; Katie Morrigan; Ayesha Irfan; Kate Bascombe; Michael Okorie; | BMC Medical Education | 2026-05-30 |
| 278 | AI-Based Resume Screening and Job Matching System Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research report examines AI-based resume screening and job matching systems that leverage Natural Language Processing, deep learning, and transformer models to automate and optimize talent acquisition.We review state-of-the-art architectures including CNN-Attention for resume topic segmentation, GA-LightGBM and Fuzzy NLP models for human-job matching, and LLM-based systems using GPT-4/GPT-5 embeddings. |
Vishal Junghare Vishal Junghare; Vaibhav Naydekar Vaibhav Naydekar; Aalok Kushwah Aalok Kushwah; Aayush Verma Aayush Verma; | International Journal of Creative and Open Research in … | 2026-05-30 |
| 279 | AI-Based Resume Screening and Job Matching System Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research report examines AI-based resume screening and job matching systems that leverage Natural Language Processing, deep learning, and transformer models to automate and optimize talent acquisition. |
Shailendra Singh Bhalla Shailendra Singh Bhalla; Vishal Junghare Vishal Junghare; Vaibhav Naydekar Vaibhav Naydekar; Aayush Verma Aayush Verma; Aalok Kushwaha Aalok Kushwaha; | International Journal of Creative and Open Research in … | 2026-05-29 |
| 280 | Positional Versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length Generalization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study the learning dynamics of attention heads in a controlled setting by training a decoder-only Transformer (GPT-J) on two structurally equivalent multi-hop reasoning tasks: a number task requiring positional reasoning and a letter task requiring symbolic reasoning. |
FELIPE URRUTIA et. al. | arxiv-cs.LG | 2026-05-29 |
| 281 | Interpreting User Opinions: A Multidimensional Approach Leveraging Explainable AI and Generative Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a multidimensional, explainable framework that combines LLM-based classification across latent dimensions (e.g., sentiment, topic, emotion), interpretable AI for identifying influential words, and generative AI for producing human-readable explanations. |
Cristian Cosentino; Merve Gunduz Cure; Fabrizio Marozzo; Sule Ozturk Birim; | Machine Learning | 2026-05-29 |
| 282 | Literature Review on Fake News Detection Using Machine Learning and Deep Learning Techniques Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive literature review of fake news detection methodologies, tracing the evolution from early rule-based and traditional machine learning approaches to modern deep learning architectures and transformer-based pre-trained language models. |
Gourav Yadav; Piyush Moghe; | Interdisciplinary Journal of AI, Machine Learning & … | 2026-05-29 |
| 283 | TRIMODAL FUSION FRAMEWORK FOR LAYER 2 NETWORK FORENSICS: INTEGRATING CNN, TRANSFORMER, AND BERT EMBEDDINGS FOR REAL-TIME MALWARE DETECTION IN RAW ETHERNET TRAFFIC Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a multimodal deep learning–based malware detection framework for Layer 2 traffic. |
Dantene Davis; | EPRA International Journal of Research & Development … | 2026-05-27 |
| 284 | CNM-BERT: A Drop-In Structural Embedding for Chinese Characters Via Ideographic Description Sequences Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose the Compositional Network Model (CNM), a lightweight augmentation that injects discrete compositional structure into Transformer encoders. |
Thomas Sing-wing Wu; Liqian Yan; | arxiv-cs.CL | 2026-05-27 |
| 285 | Enhancing Products Performance Evaluation Through Hybrid DistilRoBERTa and BiGRU Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research study aims to explore product complexity in customer reviews by utilizing transformer and recurrent neural network-based models. |
SHOUKAT ULLAH et. al. | PLOS One | 2026-05-26 |
| 286 | Retrieval-Augmented Transformer Architecture for Cross-Domain Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes an AI-driven Hybrid Transformer–Retrieval Architecture for robust cross-domain fake news detection. |
Jyothilakshmi Kava; Rajeshwari N; | International Journal For Multidisciplinary Research | 2026-05-26 |
| 287 | MechRL: Reinforcement Learning Agents Perform Circuit Discovery for Mechanistic Interpretability Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Mechanistic interpretability has identified small sets of attention heads that implement specific behaviours in transformer language models, but recovering these circuits … |
Barsat Khadka; | arxiv-cs.LG | 2026-05-25 |
| 288 | The Stability of Singular Distribution: A Spectral Perspective on The Two-Phase Dynamics of Language Model Pre-training Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We identify an underlying spectral phenomenon, Stability of Singular Distribution (SoSD), where the trace-normalized singular value spectrum stabilizes early, even as parameter matrices continue to evolve. |
Hongtao Zhang; Wenjie Zhou; Chenxi Jia; Wei Chen; Xueqi Cheng; | arxiv-cs.LG | 2026-05-25 |
| 289 | Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce the Memory Benefit Score (MBS) as a per-turn diagnostic metric. |
Ravi Kumar Tummalapenta; Suman Addanki; | arxiv-cs.CL | 2026-05-25 |
| 290 | Privacy Risks Associated with The Use of LLMs in Software Development Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Research Context: Large Language Models (LLMs), e.g., GPT-2 and GPT-4, are increasingly embedded in software development to assist with code generation, testing, and … |
Diego Menegazzi; Edna Dias Canedo; | Brazilian Symposium on Information Systems | 2026-05-25 |
| 291 | A Controlled Synthetic Benchmark for Educational Aspect-Based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a controlled synthetic benchmark for educational ABSA built from 10,000 synthetic course reviews with explicit train-validation-test splits and a 20-aspect pedagogical schema spanning instructional quality, assessment and course management, learning demand, learning environment, and engagement. |
Yehudit Aperstein; Alexander Apartsin; | arxiv-cs.CL | 2026-05-25 |
| 292 | Generic Interpretation Approach for Transformer Models Incorporating Heterogenous Attention Structures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In terms of method, we propose an interpretation method for Transformer models with heterogenous attention structures. |
Yongjin Cui; Xiaohui Fan; Huajun Chen; | arxiv-cs.CV | 2026-05-25 |
| 293 | Note-Level Phenotyping of Multiple-Sclerosis Notes By A Large Language Model Achieves Near Human-Level Agreement Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods: We analyzed 100 de-identified MS neurology progress notes from a single academic medical center. |
Daniel B. Hier; Pavankumar Y. Srinivasula; Michael D. Carrithers; | Journal of Clinical Medicine | 2026-05-25 |
| 294 | A Scalable Hybrid Framework for Sentiment Analysis of COVID-19 Tweets Using Transformer Embeddings and Lightweight Classifiers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study explores a simple and scalable way to analyze the sentiment of COVID-19 tweets without the heavy computational cost of full transformer fine-tuning. |
Zahra Rezaei; Sara Safi Samghabadi; Yaser Mike Banad; | PeerJ Computer Science | 2026-05-25 |
| 295 | Forgotten Words: Benchmarking NeoBERT for Dementia Detection in Low-Resource Conversational Filipino and English Speech Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the first systematic evaluation of transformer-based dementia detection in Filipino speech and the first assessment of NeoBERT in a clinical NLP setting. |
REZ SAMANTHA Z. FLORESCA et. al. | arxiv-cs.CL | 2026-05-25 |
| 296 | Multilingual Humour-Aware Retrieval with Dense and Re-Ranking Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, Team DUTH studies multilingual humour-aware information retrieval using the CLEF 2025 JOKER Task 1 benchmark, which evaluates humour retrieval in English and Portuguese. |
Georgios Arampatzis; Avi Arampatzis; | arxiv-cs.IR | 2026-05-24 |
| 297 | Continuous-Depth Field Theory for Transformer Patching and Mechanistic Interpretability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Mechanistic interpretability often uses activation patching, causal tracing, path patching, and steering directions to reveal behaviorally meaningful directions in Transformer activation space. This paper develops a field-theoretic framework for organizing and predicting such interventions. |
David N. Olivieri; Antonio F. Pérez Rodríguez; | arxiv-cs.LG | 2026-05-24 |
| 298 | Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a novel approach, Grammatically-Guided Sparse Attention, which constrains attention computations based on the grammatical roles of tokens. |
Spandan Pratyush; | arxiv-cs.CL | 2026-05-23 |
| 299 | Evaluating OpenAI’s Privacy Filter: Cross-Lingual, Cross-Domain PII Detection Across 42 Benchmarks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the first independent, systematic evaluation of OpenAI’s Privacy Filter (OPF), a 1.5B-parameter bidirectional PII detector, across 42 synthetic benchmarks spanning 22 languages and 5 domains. |
Rohith Uppala; | arxiv-cs.CL | 2026-05-23 |
| 300 | Sentiment Analysis Using BERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed model utilizes a pretrained BERT Transformerarchitecture fine-tuned on 50,000 labeled movie reviews. |
Rahul Nayak; P.Sakthi Murugan; | International Journal For Multidisciplinary Research | 2026-05-23 |
| 301 | Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs of Ukrainian Court Decisions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Legal NLP benchmarks evaluate models on randomly split data, implicitly assuming that legal language is stationary. |
Volodymyr Ovcharov; | arxiv-cs.CL | 2026-05-23 |
| 302 | Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for Political News Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a comparative study of RoBERTa-based sentiment analysis and an LLM-based multi-dimensional framing analysis platform applied to a corpus of 50 political news articles from 17 international media outlets. |
Maryam Fooladi; Federico Bottino; | arxiv-cs.CL | 2026-05-22 |
| 303 | Sentiment Analysis of Ukrainian-language Citizen Appeals: Classical Methods and Transformer Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents an expanded experimental comparison of the effectiveness of machine learning methods for the task of three-class sentiment classification (negative, positive, neutral). |
Myroslav Konyk; Andrii Chornyi; | Vìsnik Nacìonalʹnogo unìversitetu Lʹvìvsʹka polìtehnìka. … | 2026-05-22 |
| 304 | From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The objective of this paper is to examine and categorise movie reviews into positive and negative sentiments. |
Dip Biswas Shanto; Mitali Yadav; Prajwal Panth; Suresh Chandra Satapathy; | arxiv-cs.CL | 2026-05-21 |
| 305 | Check Your LLM’s Secret Dictionary! Five Lines of Code Reveal What Your LLM Learned (Including What It Shouldn’t Have) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce the Vocabulary Cluster Score (VCS) to quantify subspace coherence, and the Weighted Projection Score (WPS) as a static glitch token detector; applying WPS to GPT-OSS-120B recovers shokubutsu-hyakka-tsu (ID 137606), a well-known glitch token widely reported in the CJK language community, without any model inference. |
Hisashi Miyashita; | arxiv-cs.LG | 2026-05-21 |
| 306 | Reading Task Failure Off The Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We report a small, reproducible audit of which sparse-autoencoder (SAE) features of GPT-2 small fire differently on failed versus successful trials of the Indirect Object Identification (IOI) task. |
Mahdi Nasermoghadasi; | arxiv-cs.LG | 2026-05-21 |
| 307 | From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a five-stage methodology for causal feature analysis in transformer language models (probe design, feature extraction, causal validation, robustness testing, and deployment integration) and demonstrate it end-to-end on GPT-2 small performing the Indirect Object Identification (IOI) task. |
Caleb Munigety; | arxiv-cs.CL | 2026-05-21 |
| 308 | Towards Verifiable Transformers: Solver-Checkable Circuit Explanations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Verifiable Transformers, a framework for converting task-localized Transformer circuits into bounded, solver-checkable claims. |
Neel Somani; | arxiv-cs.LG | 2026-05-21 |
| 309 | Hidden-State Privacy Has An Empty Middle Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We prove a complementary Fisher-ball lower bound: every full-rank Gaussian release at $O(1)$ Fisher utility admits a direction whose Mahalanobis signal grows linearly in hidden width, ruling out uniform Gaussian safety in the class and matching the empirical empty middle. |
Alexander Okezue Bell; | arxiv-cs.LG | 2026-05-21 |
| 310 | The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language Modeling Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the first ablation-validated evidence that simplicial message passing improves language-model perplexity at the 306M-parameter scale on WikiText-103. |
Al Kari; | arxiv-cs.AI | 2026-05-21 |
| 311 | Multi-Domain Machine Learning Framework for Electric Vehicle Charging Prediction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, most existing studies rely on single-domain data, such as behavioral charging sessions or station metadata, which limits their ability to capture the joint effects of user behavior, charger characteristics, and market context. To address this gap, this study proposes a multi-domain machine learning framework for EV charger-type prediction by integrating behavioral, infrastructure, and market-level data. |
Hanan Thwany; Muhammad Alolaiwy; Mohamed Zohdy; | Vehicles | 2026-05-20 |
| 312 | SymbolicLight V1: Spike-Gated Dual-Path Language Modeling with High Activation Sparsity and Sub-Billion-Scale Pre-Training Evidence Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present SymbolicLight V1, a spike-gated dual-path language model that combines binary Leaky Integrate-and-Fire spike dynamics with a continuous residual stream. |
Ting Liu; | arxiv-cs.CL | 2026-05-20 |
| 313 | Memory-Efficient Partitioned DNN Inference on Resource-Constrained Android Crowds Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the DNN pipeline scheduling subsystem of CROWDio, which achieves practical ONNX inference across resource-constrained Android workers without model modification, by distributing memory pressure across devices via five mechanisms: JIT deferred partition loading, a single-partition-resident constraint, a 4-tier affinity scheduler, a zlib-compressed tensor transport, and a streaming 1:1 dependency model. |
Lakshani Manamperi; Disumi Pathirana; Thiwanka Pathirana; Nipun Premarathna; Kutila Gunasekera; | arxiv-cs.LG | 2026-05-20 |
| 314 | Comparing LLM and Fine-Tuned Model Performance on NVDRS Circumstance Extraction with Varying Prompt Complexity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We develop a “Complexity Score” algorithm that analyzes coding manual structure to predict when detailed prompts with full coding guidelines improve over name-only prompts. |
Geoffrey Martin; Xuan Zhong Feng; Yifan Peng; | arxiv-cs.CL | 2026-05-20 |
| 315 | Amplifying, Not Learning: Fine-Tuned AI Text Detectors Amplify A Pretrained Direction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: On raw encoders before any task supervision, projecting onto centroid(AI)-centroid(HC3) achieves NYT-vs-HC3 AUROC 0.806/0.944/0.834 across three architectures (86-106% of the fine-tuned discrimination ceiling: on RoBERTa-base, raw projection exceeds fine-tuning); on RoBERTa-base, full fine-tuning reduces discrimination below raw on both fluent-formal populations tested. |
Alexander Smirnov; | arxiv-cs.LG | 2026-05-20 |
| 316 | Post-Hoc Understanding of Metaphor Processing in Decoder-Only Language Models Via Conditional Scale Entropy Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce conditional scale entropy (CSE), a wavelet-derived measure of how broadly transformer computation engages across frequency scales at each layer position. |
LAWHORI CHAKRABARTI et. al. | arxiv-cs.CL | 2026-05-20 |
| 317 | Rethinking Cross-Layer Information Routing in Diffusion Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a systematic empirical analysis of cross-layer information flow in DiTs, jointly along depth and denoising timestep, and identify three concrete symptoms of traditional residual addition, namely monotonic forward magnitude inflation, sharp backward gradient decay, and pronounced block-wise redundancy. |
CHAO XU et. al. | arxiv-cs.CV | 2026-05-20 |
| 318 | Readability of Retina Patient Education Materials Generated With A Large Language Model Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Purpose: To determine whether large language models (LLMs) can be harnessed to improve the readability of educational material for retina patients. Methods: Forty-one documents … |
TURNER D. WIBBELSMAN et. al. | Journal of VitreoRetinal Diseases | 2026-05-20 |
| 319 | A Comprehensive Narrative Review of Extractive Text Summarization: Techniques, Transformer Models and Open Challenges (2010–2026) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Automatic text summarization is critical for managing information overload in the digital era. This comprehensive narrative review examines extractive text summarization research published between 2010 and 2026, synthesizing classical statistical methods, machine learning approaches, and contemporary deep learning and transformer-based architectures. |
Baraa A. Elhady; Osama E. Emam; Helal A. Suleiman; | Journal of Advances in Mathematics and Computer Science | 2026-05-20 |
| 320 | Findings of The Counter Turing Test: AI-Generated Text Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper provides a comprehensive analysis of state-of-the-art AI-generated text detection techniques and evaluates their effectiveness through the Counter Turing Test (CT2) shared tasks. |
RAJARSHI ROY et. al. | arxiv-cs.CL | 2026-05-20 |
| 321 | A Study of Word Embedding Models for Measuring Topic Coherence Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we thoroughly explore the application of embedded representations to evaluate the quality of topics. |
Manuel Couto; Javier Parapar; David E. Losada; | Knowledge and Information Systems | 2026-05-19 |
| 322 | Cross-Model Deepfake Text Detection with XLM-RoBERTa: A Strongly Generalizable Multi-LLM Training Strategy Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, a deep learning approach based on the XLM-RoBERTa architecture is proposed for detecting deepfake (DF) texts, with a focus on achieving strong generalization capability within the academic domain. |
İsmail Öner; Erdal Özbay; | Applied Sciences | 2026-05-19 |
| 323 | Fast Tensorization of Neural Networks Via Slice-wise Feature Distillation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a scalable tensorization framework for neural network compression based on slice-wise feature distillation. |
Safa Hamreras; Sukhbinder Singh; Román Orús; | arxiv-cs.LG | 2026-05-19 |
| 324 | Extracting Social Determinants of Health From Electronic Health Records: Development and Comparison of Rule-Based and Large Language Model Methods Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective This study develops and systematically evaluates cost-efficient methods for extracting SDoH information from unstructured clinical text using rule-based natural language processing (NLP) and large language model (LLM)–based approaches. |
Bo Wang; Dia Kabir; Cheryl Renee Clark; Karmel W Choi; Jordan W Smoller; | JMIR Medical Informatics | 2026-05-19 |
| 325 | The Performance of ChatGPT and Other Large Language Models on Multiple‐choice Questions in Biomedical Disciplines: A Meta‐analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract While large language models (LLMs) have shown promise as learning tools for medical education, their reported accuracy on multiple‐choice questions (MCQs) varies widely across studies, necessitating synthesis. |
COLLEEN M. CHEVERKO et. al. | Anatomical Sciences Education | 2026-05-19 |
| 326 | Inferring Causality Between Entities and Events from CTI Reports Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, most existing causal inference research primarily focuses on event-to-event causality, overlooking the relationships between entities and individual events. To bridge this gap, we constructed a specialized dataset designed to analyze causal relationships between entities and cyber threat events. |
CHANG-HONG JIANG et. al. | ACM Transactions on Internet Technology | 2026-05-18 |
| 327 | Tensor Language Model Enables Generative Scheduling for Efficient Tensor Compilation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The paper presents the Tensor Language Model (TLM), a generative framework of a compiler that redefines the optimisation of tensor programmes as a language modelling problem. |
SAJID MEHMOOD et. al. | Scientific Reports | 2026-05-18 |
| 328 | Fundamentos Para O Ensino De Inteligência Artificial Com Redes Neurais Explicáveis: Uma Abordagem Construcionista Mediada Pelo Simulador Robótico Open Roberta Lab Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Objetivo: Analisar os fundamentos para o ensino de Inteligência Artificial (IA) com Redes Neurais Explicáveis (XAI), a partir de uma abordagem construcionista, investigando o … |
Genarde M. Trindade; Jorge Mikael C. Alves; João da M. Libório Filho; Jhonathan A. Oliveira; Dayane R. de S. Trindade; | Review of Artificial Intelligence in Education | 2026-05-18 |
| 329 | Bug or Feature$^2$: Weight Drift, Activation Sparsity, and Spikes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We identify and analyze a negative weight drift induced by the interaction between standard losses and positively biased activation functions. |
EGOR SHVETSOV et. al. | arxiv-cs.LG | 2026-05-17 |
| 330 | Efficient and Responsible Transformer Based Conversational Agents for Emotionally Supportive Dialogue Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work presents a lightweight, domain-adapted dialogue generation system based on the T5-small architecture, fine-tuned on MentalChat16K, a curated corpus of real and synthetic emotional-support conversations. |
DIVYA SALEELA et. al. | Discover Artificial Intelligence | 2026-05-17 |
| 331 | MiniGPT: Rebuilding GPT from First Principles Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. |
Jibin Joseph; | arxiv-cs.CL | 2026-05-17 |
| 332 | Understanding Human Language Through Natural Language Processing: A Comprehensive Review of Models, Applications, Challenges, and Future Directions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review aims to synthesize recent developments in NLP, analyze major application areas, and identify key limitations and future research directions. |
Redeer Avdal Saleh; Ibrahim Mahmood Ibrahim; | Asian Journal of Advanced Research and Reports | 2026-05-16 |
| 333 | Transformer-like Inference from Optimal Control Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The framework is developed for two model classes: a nonlinear model of discrete-valued processes, directly motivated by the transformer, and a linear Gaussian model as a tractable baseline. |
Aditya Kudre; Heng-Sheng Chang; Prashant G. Mehta; | arxiv-cs.LG | 2026-05-15 |
| 334 | AI Interview Question Prediction System Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research proposes an AI-Powered Interview Question Prediction System that leverages Natural Language Processing (NLP) and a Local Large Language Model (LLM) to generate personalized interview questions based on user inputs such as resumes and job descriptions. |
Aditya Madhukar Sase; Deshmukh N. S.; | International Journal of Innovative Science and Research … | 2026-05-15 |
| 335 | ITGPT: Generative Pretraining on Irregular Timeseries Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we introduce ITGPT, an attention-based architecture designed for handling multimodal, irregularly sampled timeseries by allowing training with both SSL losses and GPT-like objectives. |
Antoine Honoré; Ming Xiao; | arxiv-cs.LG | 2026-05-15 |
| 336 | Comparative Performance of Three GPT Models on Japanese Dental Board-Style Multiple-Choice Questions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study compared two reasoning-optimized models, GPT-o3 and GPT-5T, with a general-purpose multimodal model, GPT-4o, using 399 Japanese dental board-style multiple-choice questions from 2018 to 2022. |
HIKARU FUKUDA et. al. | Computers | 2026-05-15 |
| 337 | Transformer-based NLP Approaches for Credit Risk Prediction: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Introduction This study systematically reviews transformer based Natural Language Processing (NLP) and Large Language Model (LLM) approaches for credit risk prediction, addressing limitations of traditional structured data credit scoring models. |
Pfarelo Raliphada; Seun Olukanmi; Micheal Olusanya; | Frontiers in Artificial Intelligence | 2026-05-15 |
| 338 | TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce TFGN, an architectural overlay for transformer language models that produces input-conditioned, parameter-efficient updates while leaving the rest of the transformer unchanged. |
Anurup Ganguli; | arxiv-cs.LG | 2026-05-14 |
| 339 | LLMs Applied to Web Scraping and Web Crawling: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Future work should explore SLM implementation, hybrid pipelines, and standardized evaluation benchmarks. |
Pablo Landeta-López; José María García; Cathy Guevara-Vega; Antonio Ruiz-Cortés; | Computing | 2026-05-14 |
| 340 | A Comprehensive Review on Multilingual News Recommender Systems and Their Challenges Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review paper provides a comprehensive analysis of existing techniques, datasets, and evaluation methods used in multilingual news recommendation, highlighting their strengths and limitations. |
Siddhant Siddhant; | International Journal of Creative and Open Research in … | 2026-05-14 |
| 341 | Automating Multi-label Crisis Detection in Psychological Support Hotlines with Pre-trained Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We adopted two strategies: deep learning classification with pre-trained models and Large Language Models (LLMs)-based prediction via prompt engineering (GPT-4 and DeepSeek series). |
SHUYING RAO et. al. | PLOS Digital Health | 2026-05-13 |
| 342 | Continual Learning with Multilingual Foundation Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a multi-stage framework for detecting reclaimed slurs in multilingual social media discourse. |
Barathi Ganesh HB; Michal Ptaszynski; Rene Melendez; Juuso Eronen; | arxiv-cs.CL | 2026-05-13 |
| 343 | Applying Bibliometrics and A RoBERTa Transformer in The Circular Bioeconomy: A PRISMA 2020 Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This exploratory methodological study demonstrates an integrated workflow that combines systematic evidence collection Preferred Reporting Items for Systematic reviews and Meta-Analyses (PRISMA 2020), bibliometric mapping, and Transformer-based natural language processing (RoBERTa) to generate multi-layer insights from Circular Economy-related scholarship, using circular bioeconomy literature as a domain case (2017–2025). |
GARY CHRISTIAM FARFÁN-CHILICAUS et. al. | Publications | 2026-05-13 |
| 344 | Mining Patient Narratives to Analyze Lifestyle–Blood Glucose Relationships: An LLM-Based Text Mining Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a natural language processing (NLP) framework that analyzes long-form illness blogs to identify lifestyle factors associated with elevated blood glucose levels. |
Kazuyuki Matsumoto; Minoru Yoshida; Chikaho Karino; | J | 2026-05-13 |
| 345 | Towards Intelligent Government Grievance Redressal Systems: A Survey of Computational Methods and System Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This survey examines 20 research papers from 2017 to 2026, tracing the full arc of AI-driven grievance management, from early Naive Bayes and SVM classifiers to transformer-based architectures, zero-shot LLM pipelines, Graph Neural Networks (GNNs) and blockchain-integrated multimodal frameworks. |
Navanidhi D J; | International Journal for Research in Applied Science and … | 2026-05-13 |
| 346 | Exploiting Pre-trained Encoder-Decoder Transformers for Sequence-to-Sequence Constituent Parsing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, the use of pre-trained encoder-decoder language models for constituency parsing has not been thoroughly explored. To bridge this gap, we extend the sequence-to-sequence framework by investigating parsers built on pre-trained encoder-decoder architectures, including BART, mBART, and T5. |
Daniel Fernández-González; Cristina Outeiriño Cid; | arxiv-cs.CL | 2026-05-13 |
| 347 | EduRL-GPT: A Reinforcement Learning Optimized Generative AI Framework for Intelligent Teaching Content Generation and Personalized Feedback Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, most existing generative AI-based educational systems rely on static prompting strategies and lack mechanisms to continuously optimize feedback quality according to students’ evolving learning states, which limits their effectiveness in personalized education scenarios. To address this limitation, this paper proposes EduRL-GPT, a reinforcement learning optimized generative AI framework for intelligent teaching content generation and personalized feedback. |
Xueqi Tang; Sitong Liu; | Journal of Computer Science and Frontier Technologies | 2026-05-12 |
| 348 | Explainable Fake News Detection Using Transformer Models for Multilingual Social Media Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The core contribution of this work is the integration of an Explainable AI (XAI) module utilizing techniques such as LIME and SHAP. |
Sudarshan J. Sikchi Sudarshan J. Sikchi; Nuzhat F. Shaikh Nuzhat F. Shaikh; | International Journal of Creative and Open Research in … | 2026-05-12 |
| 349 | Datasets, Models and NLP Techniques for Legal Contracts—A Survey Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This survey paper identifies gaps in clause relationship linkage and neuro‐symbolic approaches. This survey reviews the techniques and datasets employed in NLP for legal contract analysis, summarizing recent advancements in this field. |
Kapil Vuthoo; Sonia Khetarpaul; L. Venkata Subramaniam; | Expert Systems | 2026-05-12 |
| 350 | BERT: Advancements in Language Understanding for Different NLP Tasks: Challenges and Future Perspectives Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Compared to other language representation models such as ELMo and traditional transformer-based architectures, BERT demonstrates significant advancements in performance and understanding of human language. The purpose of this paper is to provide an in-depth discussion on the BERT model including its basic concept, architectural structure, method of training, and application in different NLP tasks in the real world, emphasizing its importance in enhancing NLP research. |
Md Saiful Islam; Li Xiangdong; Jubayer Ahmed; | Journal of Electrical Systems and Information Technology | 2026-05-12 |
| 351 | An Annotation Scheme and Classifier for Personal Facts in Dialogue Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present an extended annotation scheme for personal fact classification that addresses limitations in existing approaches, particularly PeaCoK. |
Konstantin Zaitsev; | arxiv-cs.CL | 2026-05-11 |
| 352 | Towards A Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We examine newly published LLVM methods, including CLIP and LLaVA neural network transformer architectures. |
David F. Ramirez; Tim L. Overman; Kristen Jaskie; Marv Kleine; Andreas Spanias; | arxiv-cs.CV | 2026-05-11 |
| 353 | Large Spectrum Models (LSMs): Decoder-Only Transformer-Powered Spectrum Activity Forecasting Via Tokenized RF Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, foundational large spectrum models (LSMs) are presented. |
Mohammad Mosiur Lunar; Mehmet C. Vuran; | arxiv-cs.NI | 2026-05-11 |
| 354 | Simply Stabilizing The Loop Via Fully Looped Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our analysis reveals that this instability stems from two sources: gradient oscillation and residual explosion. To address these two problems, we propose the Fully Looped Transformer, which introduces two parameter-free modifications: (1) Fully Looped Architecture, which distributes inter-loop signals across all layers to mitigate residual explosion; (2) Attention Injection, which reuses the existing attention block to suppress gradient oscillation. |
RAO FU et. al. | arxiv-cs.LG | 2026-05-11 |
| 355 | Transformer Interpretability from Perspective of Attention and Gradient Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: From the perspective of attention and gradient, we conduct an in-depth study of Transformer interpretation and propose a method to achieve it by guiding the gradient direction, or more precisely, the attention direction. |
Yongjin Cui; Xiaohui Fan; Huajun Chen; | arxiv-cs.AI | 2026-05-11 |
| 356 | Cantnlp@DravidianLangTech 2026: Organic Domain Adaptation Improves Multi-class Hope Speech Detection in Tulu Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents our systems and results for the Hope Speech Detection in Code-Mixed Tulu Language shared task at the Sixth Workshop on Speech, Vision, and Language Technologies for Dravidian Languages (DravidianLangTech-2026). |
Andrew Li; Sidney Wong; | arxiv-cs.CL | 2026-05-10 |
| 357 | Continuous Latent Contexts Enable Efficient Online Learning in Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Recently, continuous transformer architectures with latent chain of thought have shown promise for offline iterative tasks such as directed graph-reachability. Motivated by this, we study whether continuous latent context tokens equip transformers to more effectively realize online learning. |
Emile Anand; Abdullah Ateyeh; Xinyuan Cao; Max Dabagia; | arxiv-cs.LG | 2026-05-10 |
| 358 | Online Recruitment Fraud Detection Using Deep Learning Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a deep learning-based approach for detecting online recruitment fraud by analyzing job descriptions and related textual information. |
Ivaranjani. D; Sri Harini.G; | International Research Journal on Advanced Engineering and … | 2026-05-09 |
| 359 | Language-Conditioned Visual Grounding with CLIP Multilingual Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Multilingual vision-language models exhibit systematic performance gaps across languages, but the mechanism remains ambiguous: cross-language divergence could arise from the visual encoder, the text branch, or their interaction. We resolve this ambiguity through a dense multilingual CLIP probe in which the visual encoder is held identical across thirteen typologically diverse languages and only the XLM-RoBERTa text branch varies. |
J. de Curtò; Mauro Liz; I. de Zarzà; | arxiv-cs.CL | 2026-05-09 |
| 360 | Real-Time Sentiment Analysis of YouTube Comments Using Ensemble Deep Learning and VADER Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a hybrid ensemble system for real-time sentiment analysis of YouTube comments, integrating VADER lexicon-based analysis with RoBERTa transformer-based deep learning. |
Komal Mayukha Mamidi; | International Journal for Research in Applied Science and … | 2026-05-09 |
| 361 | Non-Monotonic Latency in Apple MPS Decoding: KV Cache Interactions and Execution Regimes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we identify unexpected non-monotonic latency behavior in the Apple MPS backend, where latency changes abruptly across nearby decoding configurations. |
Willy Fitra Hendria; | arxiv-cs.LG | 2026-05-09 |
| 362 | Going with The Mainstream: Exploring GPT Representation of Journalistic Culture Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study explores how large language models (LLMs), such as ChatGPT, can reflect journalistic value systems. |
TAEWOO KANG et. al. | The International Journal of Press/Politics | 2026-05-09 |
| 363 | Transformers Can Implement Preconditioned Richardson Iteration for In-Context Gaussian Kernel Regression Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: For nonlinear ICL, prior work has related softmax and kernelized attention to functional-gradient-type dynamics, but it remains unclear whether a standard transformer with softmax attention can implement a convergent solver with an end-to-end prediction-error guarantee. In this paper, we study in-context kernel ridge regression (KRR) with Gaussian kernels and show that a standard softmax-attention transformer can approximate the KRR predictor during its forward pass by implementing preconditioned Richardson iteration on the associated kernel linear system. |
Mingsong Yan; Dongyang Li; Charles Kulick; Sui Tang; | arxiv-cs.LG | 2026-05-08 |
| 364 | Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study focuses on contextual hallucination detection in summarization tasks. |
I. F. Atasoy; B. Mutlu; E. A. Sezer; A. Wahdan; | arxiv-cs.CL | 2026-05-08 |
| 365 | 100,000+ Movie Reviews from Kazakhstan: Russian, Kazakh, and Code-Switched Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a new publicly available corpus of 100,502 movie reviews from Kazakhstan collected from kino.kz, spanning 2001-2025 and covering 4,943 unique titles. |
Rustem Yeshpanov; | arxiv-cs.CL | 2026-05-08 |
| 366 | TextLDM: Language Modeling with Continuous Latent Diffusion Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose TextLDM, which transfers the visual latent diffusion recipe to text generation with minimal architectural modification. |
JIAXIU JIANG et. al. | arxiv-cs.CL | 2026-05-08 |
| 367 | MMEF-Net: Multimodal Emotion Feature Network with Contextual Enrichment and Dynamic Modality Weighting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents MMEF-Net, a Multimodal Emotion Feature Network that integrates audio, visual, and textual modalities through a hierarchical contextual enrichment strategy combined with a dynamic modality weighting mechanism. |
Harshita Dubey; Mohit Kadwal; | International Journal For Multidisciplinary Research | 2026-05-08 |
| 368 | Multimodal Emotion Detection in Low-Resource Languages Using Lightweight Transformer Architectures: A Dual-Level Fusion Framework Integrating DistilBERT, CNN-BiGRU, and MobileViT for Efficient Real-Time Urdu Affective Computing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose an efficiency-driven, lightweight multimodal framework for Urdu emotion detection integrating facial expressions, speech, and text. |
Muhammad Azhar; Adeen Amjad; Muhammad Arman; Deshinta Arrova Dewi; | Information | 2026-05-08 |
| 369 | Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mitigating LLMs’Hallucinations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Building on these, we propose EAACD, an expert-aware adaptive contrast decoding that uses expert differences in MoE’s higher layers to mitigate hallucinations on QA tasks. |
XINYUE FANG et. al. | arxiv-cs.CL | 2026-05-08 |
| 370 | Automated Coding of Content and Pedagogical Content Knowledge of Mathematics Using A Multi-agent Large Language Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we explored the potential of a multi-agent LLM, GradeOpt, which we developed to code mathematics content knowledge (CK) and PCK reliably, in comparison with commonly used automated coding approaches, including two natural language processing (NLP) models (RoBERTa and SBERT) and a widely used LLM (GPT-4o). |
Yasemin Copur-Gencturk; Kyle Moreno; Yucheng Chu; Hang Li; Jiliang Tang; | ZDM – Mathematics Education | 2026-05-07 |
| 371 | A Comparative Study of Feature-Based and Transformer-Based NLP Approaches for Multi-Label Movie Genre Prediction from Reviews with Genre Mapping Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates multi-label movie genre prediction from user-written reviews in which textual content is inherently subjective and the movies reviewed naturally belong to multiple genres. |
Anzhela Davityan; Arpine Janunts; Sachin Kumar; | Multimedia | 2026-05-07 |
| 372 | Weight-Decay Turns Transformer Loss Landscapes Villani: Functional-Analytic Foundations for Optimization and Generalization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To validate our theory, we introduce a scalable Villani diagnostic $Ψ_s(θ) = -Δ\mathcal{F} + s^{-1}\|\nabla \mathcal{F}\|^2$ and estimate it efficiently using Hutchinson trace probes in models with over 100M parameters. |
Abhijit Das; Sayantan Dutta; | arxiv-cs.LG | 2026-05-07 |
| 373 | SMolLM: Small Language Models Learn Small Molecular Grammar Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We train SMolLM, a 53K-parameter weight-shared transformer, to generate novel SMILES with 95% validity on the ZINC-250K drug-like-molecule benchmark, outperforming a standard GPT with 10 times more parameters. |
Akhil Jindal; Harang Ju; | arxiv-cs.LG | 2026-05-07 |
| 374 | From Token Lists to Graph Motifs: Weisfeiler-Lehman Analysis of Sparse Autoencoder Features Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a graph-structured representation in which each SAE feature is modelled as a token co-occurrence graph: nodes are the tokens most frequent near strong activations, and edges connect pairs that co-occur within local context windows. |
Ruben Fernandez-Boullon; Pablo Magariños-Docampo; Javier Perez-Robles; | arxiv-cs.AI | 2026-05-07 |
| 375 | Language-Similarity-Guided Transfer Fine-Tuning of Pre-trained Transformer Models for Sentiment Analysis Across 12 Indonesian Regional Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Sentiment analysis for Indonesian regional languages faces two persistent challenges: labeled training data is extremely limited for most regional varieties, and transformer models pre-trained on Bahasa Indonesia do not generalize reliably to languages with substantially different morphological structures. |
Brian Rizqi Paradisiaca Darnoto; Dony Bahtera Firmawan; | Journal of Computing Theories and Applications | 2026-05-07 |
| 376 | Integrating Thesaurus-Based Knowledge Into Transformer Models for Semantic Understanding of Domain-Specific Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a framework for incorporating knowledge from a military thesaurus of the Ground Forces, structured according to the XML Zthes standard, into pre-trained transformed language models, including KazBERT, multilingual BERT, and XLM-RoBERTA. |
BAYANGALI ABDYGALYM et. al. | Computers | 2026-05-07 |
| 377 | Cross Platform Political Sentimental Analysis Using Deep Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a cross-platform multilingual sentiment analysis system for analyzing political opinion related to Karnataka elections. |
Dr. Deepak K Sinha Dr. Deepak K Sinha; Pooja K V Pooja K V; | International Journal of Creative and Open Research in … | 2026-05-07 |
| 378 | Patch-Effect Graph Kernels for LLM Interpretability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a framework that reframes mechanistic analysis as a graph machine-learning problem by representing activation-patching profiles as patch-effect graphs over model components. |
Ruben Fernandez-Boullon; David N. Olivieri; | arxiv-cs.AI | 2026-05-07 |
| 379 | Spatiotemporal Expert GPT Network for Traffic Flow Prediction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The framework employs a Spatiotemporal Encoder to transform node observations, temporal context, and road-network topology into unified spatiotemporal token representations. |
Hongyan Wang; Linlong Chen; | Proceedings of the Institution of Mechanical Engineers, … | 2026-05-07 |
| 380 | The E$Δ$-MHC-Geo Transformer: Adaptive Geodesic Operations with Guaranteed Orthogonality Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the E$Δ$-MHC-Geo Transformer, a novel architecture that unifies Manifold-Constrained Hyper-Connections (mHC), Deep Delta Learning (DDL), and the Cayley transform to obtain input-adaptive, unconditionally orthogonal residual connections. |
Arash Shahmansoori; | arxiv-cs.LG | 2026-05-07 |
| 381 | Do LLMs Outperform Fine-tuned Transformers in Emotion Classification? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using a recently introduced multi-dataset emotion benchmark, we compare a Llama-based generative model with previously reported results from a fine-tuned RoBERTa classifier. |
Timothy Meinert; Anna Koufakou; | The International FLAIRS Conference Proceedings | 2026-05-06 |
| 382 | HYSARD: A Hybrid Feature-Fusion Model for Sarcasm Detection Using RoBERTa Embeddings and Linguistic Features Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Although transformer-based encoders such as RoBERTa capture contextual semantics effectively, sparse linguistic signals common in sarcastic user-generated text, such as exaggerated punctuation, elongated words, capitalization, and sentiment contrast, may not always remain explicitly accessible in the final sentence representation. To address this limitation, we propose HYSARD, a hybrid feature-fusion model that combines RoBERTa-based sentence embeddings with complementary linguistic features, including sentiment polarity, stylistic markers, syntactic patterns, and TF-IDF lexical cues. |
Ismail Jabri; Zine Eddine Louriga; Aziza El Ouaazizi; Abdelaziz Ahaitouf; | Big Data and Cognitive Computing | 2026-05-06 |
| 383 | On The (In-)Security of The Shuffling Defense in The Transformer Secure Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we show that the shuffling defense is not as robust as previously claimed. |
ZHENGYI LI et. al. | arxiv-cs.CR | 2026-05-06 |
| 384 | Shortcut Solutions Learned By Transformers Impair Continual Compositional Reasoning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using this continual LEGO experimental paradigm, we study the capability of feedforward and recurrent Transformer models to perform CL. We find that BERT, a canonical feedforward Transformer model, learns shortcut solutions that limits its ability to generalize and prevents strong forward transfer to new experiences. |
William T. Redman; Erik C. Johnson; Brian Robinson; | arxiv-cs.LG | 2026-05-06 |
| 385 | ThinkAI: A Natural Language Processing-based Intelligent Framework for Mental Health Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose ThinkAI helps individuals to better understand and manage their mental health. |
KASHISH ARA SHAKIL et. al. | Big Data | 2026-05-06 |
| 386 | Fake News Detection Using Machine Learning: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review paper presents a comprehensive survey of machine learning (ML) and deep learning (DL) approaches proposed for automated fake news detection. |
RAJNEESH SHRIVASTAVA RAJNEESH SHRIVASTAVA et. al. | International Journal of Creative and Open Research in … | 2026-05-06 |
| 387 | PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present our system for SemEval-2026 Task 9: Multilingual Polarization Detection, a binary classification task spanning 22 languages. |
Srikar Kashyap Pulipaka; | arxiv-cs.CL | 2026-05-06 |
| 388 | Optimal Attention Temperature Improves The Robustness of In-Context Learning Under Distribution Shift in High Dimensions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In a high-dimensional linear-regression framework, we analyze a Transformer with approximate softmax attention, which preserves softmax’s normalization and temperature-dependent selectivity while remaining tractable. |
Samet Demir; Zafer Dogan; | icml | 2026-05-05 |
| 389 | Using Large Language Models for Solving Tests in Anesthesiology and Intensive Care: A Comparative Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: MATERIALS AND METHODS: We conducted a comparative study of responses to 30 test items from the qualifying stage of the Professionals competition. |
ANDREI A. KLIMOV et. al. | Annals of Critical Care | 2026-05-05 |
| 390 | SimpleGPT: Improving GPT Via A Simple Normalization Strategy Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we revisit Transformer optimization through the lens of second-order geometry and establish a direct connection between architectural design, activation scale, the Hessian matrix, and the maximum tolerable learning rate. |
Marco Chen; Xianbiao Qi; Yelin He; Jiaquan Ye; Rong Xiao; | icml | 2026-05-05 |
| 391 | Rank-Aware Spectral Bounds on Attention Logits for Stable Low-Precision Training Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We derive a \emph{rank-aware concentration inequality}: when the interaction matrix $M = W^Q W^{K\top}$ has rank $r \ll d$, tail probabilities for $\max_{i,j}|S_{ij}|$ decay as $\exp(-d^2\alpha^2/r)$ rather than $\exp(-d\alpha^2)$, an improvement of $d/r$ in the exponent. |
Seyed Morteza Emadi; | icml | 2026-05-05 |
| 392 | RSTR: Reducing SpatioTemporal Redundancy in Diffusion Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present RSTR, the first framework to jointly reduce spatiotemporal redundancy in diffusion transformers. |
Ruitong Sun; Tianze Yang; Wei Niu; Jin Sun; | icml | 2026-05-05 |
| 393 | Rethinking The Rank Threshold for LoRA Fine-Tuning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The condition is stated for general output dimension $K$, so its sharpness in any particular regime, and its practical implication for the cross-entropy loss actually used in fine-tuning, are open. We give three results that together reduce the prescribed rank to $r = 1$ for binary classification in this regime. |
Juneyoung Park; | arxiv-cs.LG | 2026-05-05 |
| 394 | Universality, Function Composition, and Algorithm Emulation All In-Context Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study the in-context universal approximation and compositional generalization of softmax Transformers. |
Hong-Yu Chen; Po-Chiao Lin; Maojiang Su; Jerry Yao-Chieh Hu; Han Liu; | icml | 2026-05-05 |
| 395 | Adaptive Emotion-aware Chatbot for Mental Health Diagnosis Using Recurrent Reinforcement Learning and Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: There are many standard tests like GAD-7 for anxiety, PHQ-9 for depression, PSS-10 for stress, and many more openly available on the internet, but people might miss estimate their situation while answering these questionnaires, leading to wrong or inaccurate diagnoses. This paper focuses on integrating the questions of these three standard questionnaires and creating an emotion-aware chatbot with a dynamic questionnaire. |
Sonia Dessai; Sandhya Arora; Gayatri Joshi; | Frontiers in Artificial Intelligence | 2026-05-05 |
| 396 | Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with The Tajik Web Corpus Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: On a subsample of 10,000 documents, 17 configurations were benchmarked, covering autoregressive, encoder-decoder, and encoder-only models with three fine-tuning strategies: full fine-tuning, LoRA, and QLoRA (ranks 8 and 16). |
Mullosharaf K. Arabov; | arxiv-cs.CL | 2026-05-05 |
| 397 | Log-Normal Multiplicative Dynamics for Stable Low-Precision Deep Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a new algorithm enabling stable training under low-precision computations. |
Keigo Nishida; Eren Mehmet KIRAL; Kenichi Bannai; Mohammad Emtiyaz Khan; Thomas Moellenhoff; | icml | 2026-05-05 |
| 398 | CoFrGeNet: Continued Fraction Architectures for Language Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, inspired by continued fractions, we introduce a new function class for generative modeling. |
AMIT DHURANDHAR et. al. | icml | 2026-05-05 |
| 399 | Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: AI-generated text is nowadays produced at scale across domains and heterogeneous generation pipelines, making robustness to distribution shift a central requirement for supervised binary detectors. We train transformer-based detectors on HC3 PLUS and calibrate a single decision threshold by maximising balanced accuracy on held-out validation; this threshold is then kept fixed for all downstream test distributions, revealing domain- and generator-dependent error asymmetries under shift. |
Mohamed Mady; Johannes Reschke; Björn Schuller; | arxiv-cs.CL | 2026-05-05 |
| 400 | SVD As A Fast Interpretability Method for Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose an alternative, training-free interpretability framework that directly exploits the Singular Value Decomposition (SVD) of weight matrices in Transformer MLP sublayers. |
Min Xue; Artur Andrzejak; | icml | 2026-05-05 |
| 401 | A Capacity-Based Rationale for Multi-Head Attention Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce *Relational Graph Recognition*, where the key-query channel encodes a directed graph and, given a context (a subset of the vertices), must recover the neighbors of each vertex in the context. |
Micah Adler; | icml | 2026-05-05 |
| 402 | Step-resolved Data Attribution for Looped Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To make SDI practical at transformer scale, we propose a TensorSketch implementation that never materialises per-example gradients. |
Georgios Kaissis; David Mildenberger; Felipe Gomez; Martin Menten; Eleni Triantafillou; | icml | 2026-05-05 |
| 403 | Weights to Code: Extracting Interpretable Algorithms from The Discrete Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose the Discrete Transformer, an architecture explicitly designed to bridge the gap between continuous representations and discrete symbolic logic. |
YIFAN ZHANG et. al. | icml | 2026-05-05 |
| 404 | FairCareNLP: An AI-driven Patient Review Analyzer for Healthcare Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Materials and methods We designed a multi-component pipeline incorporating sentiment analysis, key theme extraction, clinical Named Entity Recognition (NER), and fairness modules. |
SAYYED MOHAMMAD POURYA MOMTAZ ESFAHANI et. al. | PLOS One | 2026-05-04 |
| 405 | A Hybrid Lexicon–Transformer Framework for Sentiment, Emotion, and Context Classification in Moroccan Darija (TriLex-Darija) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces TriLex-Darija , a large-scale affective lexicon suite and a hybrid lexicon–transformer framework for analyzing Moroccan Arabic (Darija) social media text across three complementary dimensions: sentiment, emotion, and pragmatic context. |
Sara El Ouahabi; Safâa El Ouahabi; El Wardani Dadi; | ACM Transactions on Asian and Low-Resource Language … | 2026-05-04 |
| 406 | Cascade Token Selection for Transformer Attention Acceleration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A method is presented for reducing the cost of representative token selection in transformer attention layers by exploiting the coherence of the representative set across depth. |
Stephen J. Thomas; | arxiv-cs.LG | 2026-05-04 |
| 407 | A Hybrid Transformer Architecture for Unsupervised Aspect‐based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose an unsupervised hybrid architecture called the BERT + GPT‐2 fusion model. |
Deepika Puvvula; Sireesha Rodda; | ETRI Journal | 2026-05-04 |
| 408 | Occlusion Aware Graph Transforemr for 3D Multi-Object Tracking Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, existing GNN based approaches often fail to capture occlusion relationships explicitly, which limits their tracking accuracy in cluttered environments. To address this challenge, we propose an occlusion-aware hybrid graph transformer (OAHGT), a novel graph transformer architecture that fuses the graph transformer with exponential decay and the edge-augmented graph transformer, enabling more expressive modeling of spatio-temporal dependencies in the tracking graph. |
J. Wei; H. Gao; J. Nie; Z. Cai; Y. Liu; | icassp | 2026-05-04 |
| 409 | Gated Subspace Inference for Transformer Acceleration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A method is presented for accelerating inference in transformer language models by exploiting the low effective rank of the token activation manifold at each layer. |
Stephen J. Thomas; | arxiv-cs.LG | 2026-05-04 |
| 410 | TriCon-Fair: Triplet Contrastive Learning for Mitigating Social Bias in Pre-Trained Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: An auxiliary language-modeling objective is trained jointly to mitigate forgetting and preserve general capability. |
C. Lyu; L. Li; S. Wu; J. Yuan; | icassp | 2026-05-04 |
| 411 | Physiologically Informed Attention: Tackling Class Imbalance in Myocardial Infarction Localization with An ST-Segment Guided Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Accurate localization of Myocardial Infarction (MI) from 12-lead electrocardiograms (ECGs) faces two key challenges: the subtlety of ischemic patterns, which makes them difficult to detect, and the severe class imbalance inherent in clinical data. To address these issues, we propose ST-Former, a novel dual-path deep learning framework. |
Y. Chen; S. Yang; Y. Liu; Q. Fu; | icassp | 2026-05-04 |
| 412 | Who’s Related? Fast and Accurate Family Relationship Detection in Conversations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a lightweight on-device model to classify the relationship of a dialogue pair into fine-grained family roles or none. |
A. Kushwaha; C. Sanchi; G. B. Rajeshbhai; | icassp | 2026-05-04 |
| 413 | 3-Key-Input: Exploring The Theoretical Minimum Keys for Text Entry Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: How far can we reduce the number of physical keys if we endow an ambiguous keyboard with modern language models? |
N. Kimura; | icassp | 2026-05-04 |
| 414 | Dynamic Semantic Path Routing with Learnable Priors for Image Captioning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing methods often rely on fixed or adaptive prompting, but remain constrained by holistic control or external labels, limiting fine-grained, multi-dimensional modeling and dynamic control. To overcome these issues, we propose a Vision-Routed Prior-driven Caption framework (VRPCap) that guides language generation via a lightweight learnable prior pool. |
W. Li; J. Yu; X. Li; Z. Li; X. Wang; | icassp | 2026-05-04 |
| 415 | Layer-Aware Early Fusion of Acoustic and Linguistic Embeddings for Cognitive Status Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Speech contains both acoustic and linguistic patterns that reflect cognitive decline, and therefore models describing only one domain cannot fully capture such complexity. |
K. Novotny; L. Moro-Velázquez; J. Mekyska; | icassp | 2026-05-04 |
| 416 | Transformer Image Quality Assessment with Multimodal Features Fusion Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Harsh conditions cause distortion in collected transformer images, requiring a quality assessment method to improve transformer fault detection accuracy. To solve this, this paper proposes a model called Multimodal Feature Fusion Transformer Image Quality Assessment(MFF-TIQA). |
W. Zhao; J. Wu; S. Yue; M. Li; | icassp | 2026-05-04 |
| 417 | Smart Review Summarizer and Sentiment Analyzer for Local Businesses in Hyderabad Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents the design and implementation of a machine-learning-based customer review analysis and sum-marization system optimized for local businesses in Hyderabad. |
M. YOGESHWARI M. YOGESHWARI et. al. | International Journal of Creative and Open Research in … | 2026-05-03 |
| 418 | So Many Opinions, So Many LLMs: Comparing Large Language Models to Traditional Machine Learning for Open- Ended Survey Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Building on previous work that employed traditional machine learning to classify text (So Many Responses, So Little Time: A Machine-Learning Approach to Analyzing Open-Ended Survey Data) [1], this study investigates how different large language models (LLMs) understand and analyze NSSE open-ended survey responses. |
Abdullah Akinde; Mariam Akinde; Rasheedat Emiola; Ahmed Akinsola; | arxiv-cs.CY | 2026-05-03 |
| 419 | Numerical Fragility in Transformers: A Layer-wise Theory for Risk Estimation and Selective Stabilization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We develop a first-order decomposition of output mismatch into layer-local attention, LayerNorm, and residual-transport terms, and derive from it a practical causal risk estimator and a budgeted controller, Bound-Guided Selective Stabilization (BGSS). |
Jinwoo Baek; | aistats | 2026-05-02 |
| 420 | Creating and Evaluating Figurative Language Dataset for Sindhi Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this article, we introduce SiNFluD, a novel benchmark dataset for Sindhi figurative language classification. |
Wazir Ali; Adeeb Noor; Saifullah Tumrani; | arxiv-cs.CL | 2026-05-02 |
| 421 | GraphTeacher: Transductive Fine-Tuning of Encoders Through Graph Neural Networks Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: We present GraphTeacher for fine-tuning transformer encoders by leveraging graph neural networks (GNNs) to effectively train models when fully labeled training data is … |
E. Koç; A. Aras; Tuna Alikaşifoğlu; Aykut Koç; | IEEE Transactions on Artificial Intelligence | 2026-05-01 |
| 422 | DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this article, we present DEFault++, a hierarchical learning-based diagnostic technique that operates at three level of abstraction: it detects whether a fault is present, classifies it into one of 12 transformer-specific fault categories (covering both attention-internal mechanisms and surrounding architectural components), and identifies the underlying root cause from up to 45 mechanisms. |
Sigma Jahan; Saurabh Singh Rajput; Tushar Sharma; Mohammad Masudur Rahman; | arxiv-cs.SE | 2026-04-30 |
| 423 | Sentiment Analysis of AI Adoption in Indonesian Higher Education Using Machine Learning and Transformer-Based Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study analyzes Indonesian student opinions on the adoption of artificial intelligence in higher education using two approaches: TF-IDF-based machine learning and Transformer-based deep learning. |
HAPPY SYAHRUL RAMADHAN et. al. | arxiv-cs.CL | 2026-04-30 |
| 424 | A Hybrid Deep Learning and Knowledge Graph Framework for Automated Safety Requirement Inference: Combining BERTopic, RoBERTa, and Graph-based Reasoning Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
KunYang Liu; YiMing Sun; YaXiao Chen; | Journal of King Saud University Computer and Information … | 2026-04-30 |
| 425 | Differentiating Translational English from Original English Using Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: More specifically, this study utilized transformer-based large language models to differentiate translational English and original English. |
Ruitao Hu; Gui Wang; Bin Shao; | Across Languages and Cultures | 2026-04-30 |
| 426 | The Mechanism Behind Understanding Anaphoric One and Noun Phrase Ellipsis: How Language Models Work Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we created manually annotated datasets (collectively referred to as AO-NPE datasets) from English Wikipedia, containing annotated instances of Anaphoric one and NPE. |
Nayoun Kim; JinYeong Bak; Yunyan Duan; Ziying Li; Kwangsu Kim; | Computational Linguistics | 2026-04-29 |
| 427 | HierBias: Context-Conditioned Hierarchical Media Bias Detection with Multi-Task Type Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present \textbf{HierBias}, a hierarchical context-conditioned media bias detector that formally models document context in bias prediction. |
Kaining Li; Ruichen Yan; Yuxin Dong; | arxiv-cs.CL | 2026-04-29 |
| 428 | Learning Rate Transfer in Normalized Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Through extensive empirical validation, we find $ν$GPT exhibits learning rate transfer across width, depth, and token horizon. |
Boris Shigida; Boris Hanin; Andrey Gromov; | arxiv-cs.LG | 2026-04-29 |
| 429 | SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose SpatialFusion, a novel framework that internalizes 3D geometric awareness into unified image generation models. |
HAIYI QIU et. al. | arxiv-cs.CV | 2026-04-29 |
| 430 | Naamah: A Large Scale Synthetic Sanskrit NER Corpus Via DBpedia Seeding and LLM Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce Naamah, a high quality silver standard Sanskrit NER dataset comprising 102,942 sentences. |
Akhil Rajeev P; Annarao Kulkarni; | arxiv-cs.CL | 2026-04-29 |
| 431 | SG-UniBuc-NLP at SemEval-2026 Task 6: Multi-Head RoBERTa with Chunking for Long-Context Evasion Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We describe our system for SemEval-2026 Task 6 (CLARITY: Unmasking Political Question Evasions), which classifies English political interview responses by coarse-grained clarity (3-way) and fine-grained evasion strategy (9-way). |
Gabriel Stefan; Sergiu Nisioi; | arxiv-cs.CL | 2026-04-29 |
| 432 | Optimisation of Decision Efficiency for Autonomous Driving at Unsignalised Intersections Based on DRL and GPT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This framework aims to combine the powerful sequence modelling capability of GPT with the goal-directed optimisation strength of DRL. |
Bojun LIU; | Promet – Traffic&Transportation | 2026-04-28 |
| 433 | Duygu-Turk: A Context-Aware Sentiment Analysis Framework for Turkish, Based on Plutchik’s Emotion Model Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This study presents Duygu-Turk, a novel deep learning-based sentiment analysis framework specifically designed for the Turkish language which is characterized by its agglutinative … |
Rabia Tintin; Sait Can Yücebaş; | J. Univers. Comput. Sci. | 2026-04-28 |
| 434 | A Study on Learners’ Emotion Classification Based on An Improved Convolutional Neural Network Algorithm in Online Teaching and Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, conventional sentiment analysis approaches often struggle with unstructured textual data, limiting their ability to discern the emotional inclinations embedded in student comments precisely. To address this challenge, this study introduces RoBERTa-BiLSTM-TextCNN Network (RBTCN-Net), a novel framework that integrates Robustly Optimized BERT Pretraining Approach (RoBERTa), a Convolutional Neural Network (CNN), a Bidirectional Long Short-Term Memory (Bi-LSTM), and an attention mechanism to classify sentiment in an online learning environment. |
Yiling Chen; Hui Chen; Zixuan Wang; | PeerJ Computer Science | 2026-04-28 |
| 435 | Fake News Classification Using Machine Learning and Deep Learning Technique Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This project focuses on detecting fake news using a hybrid approach combining machine learning and deep learning techniques. |
Somendr Kumar; Mahmood Adnan; | International Journal of Innovative Research in Engineering | 2026-04-28 |
| 436 | Spreadsheet Modeling Experiments Using GPTs on Small Problem Statements and The Wall Task Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We identify two central challenges – the problem of confidence and the problem of workflow – which highlight the need for skilled users to verify and adapt GPT-generated spreadsheets. |
Thomas A. Grossman; Yuan Chen; Sopiko Datuashvili; | arxiv-cs.SE | 2026-04-28 |
| 437 | Retraction Note: Towards Improved Fake News Detection Using A Hybrid RoBERTa and Metadata Enhanced XGBoost Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
ARMUGHAN ALI et. al. | Scientific Reports | 2026-04-28 |
| 438 | ImproBR: Bug Report Improver Using LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose ImproBR, an LLM-based pipeline that automatically detects and improves bug reports by addressing missing, incomplete, and ambiguous S2R, OB, and EB sections. |
Emre Furkan Akyol; Mehmet Dedeler; Eray Tüzün; | arxiv-cs.SE | 2026-04-28 |
| 439 | Pregnancy Meal Planning and Nutrition Recommendation Using BERT, FLAN-T5 Transformers and Gemini Tool Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional diet charts are often static and fail to consider individual factors such as trimester, dietary preferences, and medical conditions. This system adopts a hybrid AI approach by combining dataset-driven meal planning with transformer-based Natural Language Processing and generative AI. |
M. BHANU SRIDHAR et. al. | International Journal of Innovative Science and Research … | 2026-04-27 |
| 440 | AI-Driven Web-Based Video Conferencing System with Real-Time Meeting Intelligence and Analytics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents an AI-Driven Web-Based Video Conferencing System that integrates real-time communication with an embedded artificial intelligence analytics pipeline. |
Asina Begam A; Ramya K; Anushiya V; Dharshani A; Dharshini T; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-27 |
| 441 | Real-Time Fact-Checking Using Retrieval-Augmented Generation and Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract—The rapid spread of misinformation on online plat-forms has motivated extensive research in automated fake news detection. |
Riharika Riharika; Navya Sidarla; Shivakumar Jagadam; Abhilash Nadigottu; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-27 |
| 442 | ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block Decoding Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We develop ClusterFusion++, a CUDA-level extension that broadens fusion to the full Transformer decoder block for GPT-NeoX/Pythia models: LayerNorm -> QKV -> RoPE -> decode attention -> output projection -> Post-LN -> MLP -> residual. |
ChiHeng Jin; Hongche Yu; Xihui Chen; | arxiv-cs.DC | 2026-04-26 |
| 443 | Adaptive Swin Transformer Partitioning Over AI-RAN Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To address the large intermediate activations inherent to transformers, we introduce an efficient, accuracy-preserving activation compression pipeline that substantially reduces uplink payload. |
TAM THANH NGUYEN et. al. | arxiv-cs.NI | 2026-04-26 |
| 444 | Graph Memory Transformer (GMT) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We investigate whether the Feed-Forward Network (FFN) sublayer in a decoder-only transformer can be replaced by an explicit learned memory graph while preserving the surrounding autoregressive architecture. |
Nicola Zanarini; Niccolò Ferrari; | arxiv-cs.LG | 2026-04-26 |
| 445 | Konnect: A Distributed Microservices Architecture for Real-Time Chat with Multi-Task NLP and Context-Aware Toxicity Dampening Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents Konnect, a cloud-native group messaging platform built on six containerized microservices interconnected through Apache Kafka publish-subscribe pipelines. |
Daniel Odametey; Ankit Sinha; Shantanu Chhetri; Raipuriya Sakshi; Manan Gulati; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-25 |
| 446 | Analyzing Digital Consumer Insights Through RoBERTa LLM Based Sentiment Analysis and Topic Modeling Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Qi Shasha; | Scientific Reports | 2026-04-24 |
| 447 | HubRouter: A Pluggable Sub-Quadratic Routing Primitive for Hybrid Sequence Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce HubRouter, a pluggable module that replaces O(n^2) attention layers with O(nM) hub-mediated routing, where M << n is a small number of learned hub tokens. |
Abhinaba Basu; | arxiv-cs.LG | 2026-04-24 |
| 448 | EverydayGPT: Confidence-Gated Routing for Efficient and Safe Hybrid GPT-RAG Conversational QA Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce EverydayGPT, a lightweight conversational QA system built around a Confidence-Gated Routing (CGR) mechanism that formalises the routing decision as a joint policy over retrieval distance and extraction adequacy. |
Jaspreet Singh Nahal; | arxiv-cs.CL | 2026-04-24 |
| 449 | A Comparative Review of VADER and BERT for Sentiment Analysis: Methods, Performance, and Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive review of VADER (Valence Aware Dictionary and Sentiment Reasoner) and BERT (Bidirectional Encoder Representations from Transformers), focusing on their architectures, methodologies, advantages, limitations, and comparative performance. |
Pamarthi Maheswarini; Bandaru Leela Harika; Malisetti Amrutha; Malla Ramani ; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-24 |
| 450 | Geometric Monomial (GEM): A Family of Rational 2N-differentiable Activation Functions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work we propose a family of $C^{2N}$-smooth activation functions whose gate follows a log-logistic CDF, achieving ReLU-like performance with purely rational arithmetic. |
Eylon E. Krause; | arxiv-cs.LG | 2026-04-23 |
| 451 | MKJ at SemEval-2026 Task 9: A Comparative Study of Generalist, Specialist, and Ensemble Strategies for Multilingual Polarization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a systematic study of multilingual polarization detection across 22 languages for SemEval-2026 Task 9 (Subtask 1), contrasting multilingual generalists with language-specific specialists and hybrid ensembles. |
Maziar Kianimoghadam Jouneghani; | arxiv-cs.CL | 2026-04-23 |
| 452 | An Optimization-driven Framework for Sentiment Classification in Customer Reviews Addressing Linguistic and Rating Bias Via MOORA and Transformer-based Embeddings Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed framework aims to answer the following research questions: (i) how transformer-based feature extraction combined with optimization techniques can improve sentiment classification accuracy, (ii) whether the model can remain robust across multilingual datasets, and (iii) how rating-independent sentiment tagging can mitigate rating bias. |
Neha Punetha; Goonjan Jain; | Intelligent Data Analysis: An International Journal | 2026-04-23 |
| 453 | LC-ATSR: A Transformer-Based Approach for Semantic Retrieval in Lakehouse Data Platforms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a Lakehouse Collaborative Adaptive Transformer Semantic Representation Algorithm (LC-ATSR) to address the core challenge of semantic retrieval of unstructured data in a Lakehouse integrated data platform. |
Youfang Xu; | Informatica | 2026-04-23 |
| 454 | Learning Scientific Document Representations Via Triple-Source Automatic Supervision Without Annotations or Citations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a Triple-Source automatic supervision framework for learning document embeddings from scientific corpora. |
MUSSA TURDALYULY et. al. | Computers | 2026-04-23 |
| 455 | Prompt Analysis in Large Language Models Adversarial Prompt Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional rule-based defenses are insufficient since adversarial prompts exploit the semantic and contextual nature of language. This project proposes the development of an AIdriven detection system to identify and block adversarial prompts before they reach the LLM by leveraging transformer-based NLP models (BERT, RoBERTa, LLaMA) along with adversarial training techniques to build a classifier that distinguishes between safe, suspicious and malicious prompts. |
M. SRIRAM et. al. | Research Digest on Engineering Management and Social … | 2026-04-23 |
| 456 | Automated Scoring for Korean Essays Utilizing The Neural Pairwise Contrastive Regression Along with Korean NLP Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Introduction Automated Essay Scoring (AES) research has increasingly adopted deep learning and transformer-based language models. |
Ji Hoon Ryoo; Jeongheum Cho; Yeongjin Jo; | Frontiers in Education | 2026-04-22 |
| 457 | VALIANT: A Vision‐Authenticity Language Framework Through Integrated Experts and Aligned Numerical‐Textual Descriptors for Citri Reticulatae Pericarpium Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present VALIANT, a multimodal vision‐language framework for smartphone images that enables age classification (1, 6, 10, 15, and 20 years) and authenticity verification against standard operating procedure (SOP) mimics. |
SIMON C. K. CHAN et. al. | Advanced Intelligent Systems | 2026-04-22 |
| 458 | Multimodal and Explainable Deep Learning for Occupational Accident Classification Using Transformer-LSTM Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a multimodal Hybrid Transformer-LSTM framework for classifying occupational fatalities by jointly modeling unstructured narratives, cyclical temporal features, and regional spatial indicators. |
Esin Ayşe Zaimoğlu; | Buildings | 2026-04-22 |
| 459 | Working Memory Constraints Scaffold Learning in Transformers Under Data Scarcity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We investigate the integration of human-like working memory constraints into the Transformer architecture and implement several cognitively inspired attention variants, including fixed-width windows based and temporal decay based attention mechanisms. |
Pranava Madhyastha; Dagmar Adamcova; | arxiv-cs.CL | 2026-04-22 |
| 460 | Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a systematic empirical study of transformer compression through over 40 experiments on GPT-2 (124M parameters) and Mistral 7B (7.24B parameters). |
Samuel Salfati; | arxiv-cs.LG | 2026-04-22 |
| 461 | SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a novel architecture for trajectory-conditioned forecasting of future 3D scene occupancy. |
JIAYUAN DU et. al. | cvpr | 2026-04-21 |
| 462 | The Signal Is The Ceiling: Measurement Limits of LLM-predicted Experience Ratings from Open-ended Survey Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one point 67% of the time from open-ended survey text. |
Andrew Hong; Jason Potteiger; Luis E. Zapata; | arxiv-cs.CL | 2026-04-21 |
| 463 | BioLAMR: A Biomimetically Inspired Large Language Model Adaptation Framework for Automatic Modulation Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose BioLAMR, a GPT-2 adaptation framework for AMR inspired by the auditory system’s parallel time–frequency processing and cortical hierarchy. |
Yubo Mao; Wei Xu; Jijia Sang; Haoan Liu; | Biomimetics | 2026-04-21 |
| 464 | Humanoid Generative Pre-Training for Zero-Shot Motion Tracker Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Humanoid-GPT, the first GPT-style humanoid motion Transformer trained with causal attention on a billion-scale motion corpus for whole-body control. |
ZEKUN QI et. al. | cvpr | 2026-04-21 |
| 465 | Reviving ConvNeXt for Efficient Convolutional Diffusion Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here we introduce the fully convolutional diffusion model (FCDM), a ConvNeXt-inspired backbone redesigned for conditional diffusion modeling. |
TAESUNG KWON et. al. | cvpr | 2026-04-21 |
| 466 | Text Marker -Text Evaluation Tool for Marking AI and Real Knowledge-based Entries and Responses Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract – In today’s world, because of how quickly artificial intelligence writing tools are getting better, it’s now hard to tell whether something was written by AI or a person. This makes things complicated for students writing papers, for checking if information is true, and for making sure work is original. |
Asst. Prof.K. Priyanka; D. Sreya; K.Srivalli Varshini; K. Ashrith; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-21 |
| 467 | Multi‐modal Large Language Models for Paediatric Tele‐ophthalmology: A Blinded Real‐world Evaluation of Diagnostic Accuracy and Safety Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study systematically assessed the diagnostic accuracy, safety and communication quality of advanced LLMs in real‐world paediatric ophthalmology tele‐consultations, with a focus on comparing text‐only inputs to multi‐modal inputs (text paired with parent‐taken mobile photographs). |
Daohuan Kang; Lu Yuan; Jia Feng; Andrzej Grzybowski; Kai Jin; | Acta Ophthalmologica | 2026-04-21 |
| 468 | Guiding A Diffusion Transformer with The Internal Dynamics of Itself Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The diffusion model presents a powerful ability to obtian the entire (conditional) data distribution. |
Xingyu Zhou; Qifan Li; Xiaobin Hu; Hai Chen; Shuhang Gu; | cvpr | 2026-04-21 |
| 469 | Assessing Capabilities of Large Language Models in Social Media Analytics: A Multi-task Quest Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we present the first comprehensive evaluation of modern LLMs – including GPT-4, GPT-4o, GPT-3.5-Turbo, Gemini 1.5 Pro, DeepSeek-V3, Llama 3.2, and BERT – across three core social media analytics tasks on a Twitter (X) dataset: (I) Social Media Authorship Verification, (II) Social Media Post Generation, and (III) User Attribute Inference. |
Ramtin Davoudi; Kartik Thakkar; Nazanin Donyapour; Tyler Derr; Hamid Karimi; | arxiv-cs.CL | 2026-04-20 |
| 470 | Does Welsh Media Need A Review? Detecting Bias in Nation.Cymru’s Political Reporting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: I use a two-stage natural language processing (NLP) pipeline: (1) a robustly optimized BERT approach (RoBERTa) bias detector for efficient bias discovery and (2) a large language model (LLM) for target-attributed sentiment classification of bias labels from (1). |
Cai Parry-Jones; | arxiv-cs.CL | 2026-04-19 |
| 471 | Recurrent Reasoning on Symbolic Puzzles with Sequence Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce RecurrReason, a difficulty-controlled benchmark of four recurrent logic puzzles (Tower of Hanoi, River Crossing, Block World, and Checkers Jumping) with BFS-optimal trajectories and a single interpretable difficulty parameter $N \in \{1,\dots,10\}$, totalling 10{,}817 unique puzzles and 285{,}933 moves. |
Gowrav Mannem; Chowdhury Marzia Mahjabin; Jason Chen; Shivank Garg; Kevin Zhu; | arxiv-cs.AI | 2026-04-19 |
| 472 | Sentiment Analysis of Social Media Reviews: A Machine Learning and Deep Learning Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive sentiment analysis framework for social media reviews that leverages state-of-the-art natural language processing (NLP) and machine learning techniques. |
Janhvi Dave; Siddhkant Pathak; | International Journal of Creative and Open Research in … | 2026-04-18 |
| 473 | Enhancing Transformer Attention Mechanisms for Knowledge Retention in Fine-Tuned Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper suggests a better attention-based model to enhance the knowledge retention in fine-tuned transformer models. |
Navya Veginati; | International Journal of Scientific Research in Science and … | 2026-04-18 |
| 474 | When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study how informal surface forms degrade NLI accuracy in ELECTRA-small (14M) and RoBERTa-large (355M) across four transforms applied to SNLI and MultiNLI: slang substitution, emoji replacement, Gen-Z filler tokens, and their combination. |
Avinash Goutham Aluguvelly; | arxiv-cs.CL | 2026-04-17 |
| 475 | PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present PIIBench, a unified benchmark corpus for Personally Identifiable Information (PII) detection in natural language text. |
Pritesh Jha; | arxiv-cs.CL | 2026-04-17 |
| 476 | Evaluating ChatGPT’s Adherence to Medical Ethics: A Prerequisite for Artificial Intelligence in Medicine Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Background As artificial intelligence continues to play an expanding role in healthcare, ensuring its compliance with medical ethics is essential. However, the ethical performance … |
YING ZHANG et. al. | Health Care Science | 2026-04-16 |
| 477 | Domain Fine-Tuning FinBERT on Finnish Histopathological Reports: Train-Time Signals and Downstream Correlations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: (1) We describe our observations from fine-tuning the Finnish BERT model on Finnish medical text data. |
RAMI LUISTO et. al. | arxiv-cs.CL | 2026-04-16 |
| 478 | Capture-the-Phish Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents Capture the Phish, an integrated AI-powered framework that combines transformer-based natural language processing, explainable AI (XAI), and privacy-preserving federated learning. |
Prof.Pankaj Chandre; Kanak Lingwat; Prachi Katrodia; Tejas Vidhate; Swethanshi Patro; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-16 |
| 479 | AI Against Suicide: Real – Time Risk Detection from Social Media Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study is aimed at automated suicide risk detection on the dataset Suicide Post Detection on Social Media Articles published by Google. |
V V N S S Raghu Nath; Tejinder Thind; Uminder Kaur; | International Journal of Drug Delivery Technology | 2026-04-16 |
| 480 | Mapping The LLM Landscape: A Cross-Family Survey of Architectures, Alignment Methods, and Benchmark Performance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This survey provides a comprehensive review of major proprietary and open-source LLM families, including GPT, LLaMA 2, Gemini, Claude, DeepSeek, Falcon, and Qwen. |
Deepshikha Bhati; Fnu Neha; Devi Sri Bandaru; Matthew Weber; Ishan Dilipbhai Gajera; | AI | 2026-04-16 |
| 481 | Fake News Detection Through LLM-Driven Text Augmentation Across Media and Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a fake news detection framework based on LLM-driven, feature-guided text augmentation. |
Abdul Sittar; Mateja Smiljanic; Alenka Guček; Marko Grobelnik; | Machine Learning and Knowledge Extraction | 2026-04-15 |
| 482 | RG-Hybrid: Pure Late Fusion of RoBERTa and A Graph Transformer for Robust, Interpretable Sentiment Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce RG-Hybrid , a dual-branch architecture that combines a fine-tuned RoBERTa encoder with a graph transformer operating over dependency parses. |
Khaled Alahmadi; Sultan Alharbi; Xianzhi Wang; | International Journal of Data Science and Analytics | 2026-04-15 |
| 483 | Towards Benchmarking Transformer Models for Biomedical Text Simplification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Especially in biomedical texts, simplification can play an essential role in making scientific information understandable to patients and the general public. In this context, this study investigates the text simplification performance of pre-trained general-purpose and domain-specific language models (PLMs) for biomedical texts. |
Öykü Berfin Mercan; Mansur Alp Toçoğlu; Nezihe Turhan Turan; Aytuğ Onan; | Scientific Journal of Mehmet Akif Ersoy University | 2026-04-15 |
| 484 | Cognitive-Linguistic Indicators of Depression in Online Communities: Analysed By DistilBERT and Holographic Reduced Representation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates whether combining cognitively grounded linguistic features with transformer-based embeddings improves automated detection of depression in online text. |
Brian Van Steen; | arxiv-cs.CL | 2026-04-15 |
| 485 | SUSTAINABLE AI-NLP FRAMEWORK FOR EARLY DETECTION OF MENTAL HEALTH RISKS IN SOCIAL MEDIA POST Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we apply natural language processing (NLP) and machine learning (ML) techniques to predict mental health disorders from social media posts. |
Aravind Potharaju; Begum Jahanara; Hema Mannem; Rahul Rathla; Dhanunjay Vemula; | International Scientific Journal of Engineering and … | 2026-04-15 |
| 486 | Identifying and Mitigating Gender Cues in Academic Recommendation Letters: An Interpretability Case Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we investigate the extent to which Transformer-based encoder models as well as Large Language Models (LLMs) can infer the gender of applicants in academic LoRs submitted to an U.S. medical-residency program after explicit identifiers like names and pronouns are de-gendered. |
CHARLOTTE S. ALEXANDER et. al. | arxiv-cs.LG | 2026-04-14 |
| 487 | Classification of Cochrane Plain Language Summaries By Conclusiveness Using Transformer-Based Models and ChatGPT: Retrospective Observational Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We used a publicly available dataset containing 4405 Cochrane PLSs of systematic reviews published until 2019, already classified by humans according to 9 categories of conclusiveness regarding the intervention’s effectiveness or safety. |
ANTONIJA MIJATOVIĆ et. al. | JMIR Medical Informatics | 2026-04-14 |
| 488 | L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Learning to Defer for clinical text (L2D-Clinical), a framework that learns when a BERT classifier should defer to an LLM based on uncertainty signals and text characteristics. |
Rishik Kondadadi; John E. Ortega; | arxiv-cs.CL | 2026-04-14 |
| 489 | Towards Intelligent Mobile On-device Assistants: Low-level Text-to-actions with GPT LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents an approach to LLM inference that allows LLMs with billions of parameters to be executed directly on mobile devices without network connectivity. |
Samuel Carreira; Tomás Marques; José Ribeiro; Carlos Grilo; António Pereira; | Progress in Artificial Intelligence | 2026-04-14 |
| 490 | ARTIFICIAL INTELLIGENCE FOR EARLY DETECTION OF MENTAL HEALTH DISORDERS USING SOCIAL MEDIA DATA Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The review summarizes more recent “natural language processing” (NLP), deep learning systems, including BERT, RoBERTa, and Bidirectional LSTM networks, multimodal fusion models, and “Explainable AI” (XAI) models related to improving clinical interpretability. |
Sukhpreet Kaur; | International Journal of Engineering Technologies and … | 2026-04-14 |
| 491 | Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To validate against human-annotated data, we evaluate all models zero-shot on GoEmotions (English) and SemEval-2018 Task 1 E-c (English, Arabic, Spanish). |
Vadim Borisov; | arxiv-cs.CL | 2026-04-14 |
| 492 | Chatgpt And Beyond: A Study on Conversational Generative AI Systems Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This paper presents a comprehensive study of conversational generative artificial intelligence (AI) systems, with a primary focus on ChatGPT (GPT-4) and a systematic comparison … |
Thota Keerthini ; Kamunuri Ganapathi Babu; | International Scientific Journal of Engineering and … | 2026-04-14 |
| 493 | Fake News Detection Using Machine Learning and NLP Models: A Comparative Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a systematic comparative evaluation of machine learning (ML) and natural language processing (NLP) methodologies applied to the challenge of identifying fake news. |
T .POOJITH; SOUMYA E; | International Scientific Journal of Engineering and … | 2026-04-14 |
| 494 | Quantization Dominates Rank Reduction for KV-Cache Compression Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We compare two strategies for compressing the KV cache in transformer inference: rank reduction (discard dimensions) and quantization (keep all dimensions, reduce precision). |
Samuel Salfati; | arxiv-cs.LG | 2026-04-13 |
| 495 | A Contextual Similarity Analysis System for Detecting Original Authored and Machine Generated Scientific Abstract Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Ensuring academic integrity has become increasingly challenging due to the rapid growth of artificial intelligence, as machine-generated scientific abstracts closely resemble human-written content. To address this, this study presents a Contextual Similarity Analysis System designed to detect and classify original-authored and machine-generated scientific abstracts. |
Srilakshmi G; | International Journal for Research in Applied Science and … | 2026-04-13 |
| 496 | Efficient Emotion-Aware Iconic Gesture Prediction for Robot Co-Speech Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Most data-driven robot systems generate rhythmic beat-like motion, yet few integrate semantic emphasis. To address this, we propose a lightweight transformer that derives iconic gesture placement and intensity from text and emotion alone, requiring no audio input at inference time. |
EDWIN C. MONTIEL-VAZQUEZ et. al. | arxiv-cs.RO | 2026-04-13 |
| 497 | LLM-Enhanced Log Anomaly Detection: A Comprehensive Benchmark of Large Language Models for Automated System Diagnostics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a comprehensive benchmark study evaluating both LLM-based and traditional approaches for log anomaly detection across four widely-used public datasets: HDFS, BGL, Thunderbird, and Spirit. |
Disha Patel; | arxiv-cs.LG | 2026-04-13 |
| 498 | A Generative AI Framework for Marathi Grammar Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a generative AI model for learning the Marathi grammar which is further categorized into two parts. |
Prof. Gayatri Dharap; Dr. Gauri Deshpande; Prof. Vandana Sharma; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-13 |
| 499 | LLMs Struggle with Abstract Meaning Comprehension More Than Expected Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: (2) A proposed bidirectional attention classifier, inspired by human cognitive strategies, enhances fine-tuned models by dynamically attending to passages and options. |
Hamoud Alhazmi; Jiachen Jiang; | arxiv-cs.CL | 2026-04-13 |
| 500 | A Synthetic Conversational Smishing Dataset for Social Engineering Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Despite the existence of several datasets for single-message smishing detection, datasets capturing conversational smishing remain largely unavailable, limiting research on multi-turn attack detection. To address this gap, this paper presents a synthetically generated dataset of 3,201 labeled multi-round conversations designed to emulate realistic conversational smishing attacks. |
Carl Lochstampfor; Ayan Roy; | arxiv-cs.CR | 2026-04-13 |
| 501 | Sentiment Scope: Context-Aware Social Media Sentiment Analysis Using Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, traditional sentiment analysis techniques often struggle to capture contextual meaning, sarcasm, and complex linguistic patterns present in social media text. To address these challenges, this study proposes a context-aware sentiment analysis approach using transformer-based models. |
Prashanthi B; Hansika Dasari; | International Scientific Journal of Engineering and … | 2026-04-12 |
| 502 | INCRT: An Incremental Transformer That Determines Its Own Architecture Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Transformer architectures are designed by trial and error: the number of attention heads, the depth, and the head size are fixed before training begins, with no mathematical … |
Giansalvo Cirrincione; | arxiv-cs.LG | 2026-04-12 |
| 503 | Predicting Depression Via Transformer-Based Text Embeddings and XGBoost Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a machine learning–based framework to detect depressive tendencies from user-generated social media content through automated and data-driven analysis. |
Isha B; Drashti I; Mahek J; Shradha B; | International Scientific Journal of Engineering and … | 2026-04-12 |
| 504 | From GPT-3 to GPT-5: Mapping Their Capabilities, Scope, Limitations, and Consequences Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the progress of the GPT family from GPT-3 through GPT-3.5, GPT-4, GPT-4 Turbo, GPT-4o, GPT-4.1, and the GPT-5 family. |
Hina Afridi; Habib Ullah; Sultan Daud Khan; Mohib Ullah; | arxiv-cs.AI | 2026-04-11 |
| 505 | Batch Size Effects on Mid‐2025 State‐of‐the‐Art Large Language Model Performance in Automated Title and Abstract Screening Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We used a gold‐standard dataset of 790 references (93 considered relevant) from a published Cochrane Review on stem cell treatment for acute myocardial infarction. |
PETTER FAGERBERG et. al. | Cochrane Evidence Synthesis and Methods | 2026-04-11 |
| 506 | An AI-Driven Framework for Fake News Classification and Sentiment Prediction Using Transformer and RNN Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work presents an AI-based platform for fake news classification and sentiment analysis based on Transformer models (BERT, DistilBERT) combined with Recurrent Neural Networks (LSTM, Bi-LSTM). |
Parulekar Waman R.; Surve Supriya S.; Khanche Gousiya A.; Joshi Sanika C.; Naik Hemangi D.; | International Journal Of Recent Trends In Multidisciplinary … | 2026-04-11 |
| 507 | AI Based Fake News Detection Using Natural Language Processing (NLP) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work contributes to the development of scalable and efficient AI-based solutions for combating misinformation. |
TANMAY AKRE et. al. | International Scientific Journal of Engineering and … | 2026-04-11 |
| 508 | Study on ML Based Automated Paper Checking and Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents an ML-driven system designed to evaluate handwritten answer sheets without manual intervention. |
Akshat Andhale Akshat Andhale; Rohit Chaudhari Rohit Chaudhari; Pratik Derle Pratik Derle; Rohit Chaudhari Rohit Chaudhari; Kunal Ahire Kunal Ahire; | International Journal of Creative and Open Research in … | 2026-04-11 |
| 509 | A Robust Transformer Framework for Identifying Large-Scale Disease Spread Patterns Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work focuses on the problem of automatically classifying clinical text into appropriate medical specialties, which is crucial for enhancing patient care, optimizing healthcare resources, and supporting accurate clinical decision-making. |
Satish Chander Govindu; Chittimalla Saidulu; Akkapalli Saikrishna; Bachala Shravya; Elukapelly Arthika; | International Journal of Engineering Research and Science … | 2026-04-10 |
| 510 | EncFormer: Secure and Efficient Transformer Inference Over Encrypted Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present EncFormer, a two-party private Transformer inference framework that introduces Stage Compatible Patterns so that FHE kernels compose efficiently, reducing repacking and conversions. |
Yufan Zhu; Chao Jin; Khin Mi Mi Aung; Xiaokui Xiao; | arxiv-cs.CR | 2026-04-10 |
| 511 | Probabilistic Short-Term Sky Image Forecasting Using VQ-VAE and Transformer Models on Sky Camera Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents a deep learning framework that directly estimates cloud movement from ground-based all-sky camera images, rather than predicting future production from past power data. |
Chingiz Seyidbayli; Soheil Nezakat; Andreas Reinhardt; | Journal of Imaging | 2026-04-10 |
| 512 | An Optimized Transformer Ensemble Model for Context-Aware Speech Act Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research proposes a transformer-driven approach that utilizes eXtreme Language Network (XLNet) for feature extraction, combined with multiple machine learning models including Logistic Regression (LR), Random Forest (RF), Support Vector Machine (SVM), and a Boosting Fusion Model (BFM) integrating Light Gradient Boosting Machine (LGBM) and Categorical Boosting (CB). |
M. Ganesh; Mogulagani Ankitha; Komakula Sathwika; Kaniganti Chandu; Pogula Nagaraju; | International Journal of Data Science and IoT Management … | 2026-04-10 |
| 513 | Uncertainty-Aware Transformers: Conformal Prediction for Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents an uncertainty quantification framework for transformer-based language models. |
Abhiram Vellore; Niraj K. Jha; | arxiv-cs.LG | 2026-04-09 |
| 514 | Improving DNS Exfiltration Detection Via Transformer Pretraining Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: While previous work mostly examines fine-tuned generic Transformers, it does not aim to isolate the effect of pretraining on the downstream task of classification. To address this gap, we develop a controlled pipeline where we freeze operating points on validation and transfer them to the test set, thus enabling clean ablations across different label and pretraining budgets. |
Miloš Tomić; Aleksa Cvetanović; Predrag Tadić; | arxiv-cs.CR | 2026-04-09 |
| 515 | NCL-BU at SemEval-2026 Task 3: Fine-tuning XLM-RoBERTa for Multilingual Dimensional Sentiment Regression Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper describes a system developed for Track A – Subtask 1 (Dimensional Aspect Sentiment Regression), aiming to predict real-valued VA scores in the [1, 9] range for each given aspect in a text. |
Tong Wu; Nicolay Rusnachenko; Huizhi Liang; | arxiv-cs.CL | 2026-04-09 |
| 516 | GAEn/MGT: A Novel Framework for Bilingual Sentiment Analysis in Mental Health Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes GAEn/MGT (Generative Auto-encoder Network / Multi-generative Transformer), an integrated generative-transformer model used for robust sentiment analysis task for mental health discourse, which has bilingual and transliterated data. |
Rahul Dubey; Premnarayan Arya; Shiv Shakti Shrivastava; Rahul Kumar Chawda; | International Journal on Artificial Intelligence Tools | 2026-04-09 |
| 517 | Deep Learning-Based Phishing Email Detection for Cybersecurity Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The paper analytically reviews the typical datasets (Enron, Nazario, SpamAssassin), preprocessing techniques, architecture innovations and hybrid models and attention mechanisms, and metrics. |
Gunda Joshith Kumar; Ponnamanda Lahari; Rokam Likhita; B N.V Sai Durga; Peddinti Rajesh; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-09 |
| 518 | ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-generative Clinical Prediction Tasks Summary Related Papers Related Patents Related Grants Related Venues Related Experts Related Code View Save Abstract: Large Language Models (LLMs) are increasingly deployed in medicine. However, their utility for non-generative clinical prediction is under-evaluated, and they are often assumed to … |
YINGHAO ZHU et. al. | npj Digital Medicine | 2026-04-08 |
| 519 | TRAPTI: Time-Resolved Analysis for SRAM Banking and Power Gating Optimization in Embedded Transformer Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work presents TRAPTI, a two-stage methodology that combines cycle-level inference simulation with time-resolved analysis of on-chip memory occupancy to guide design decisions. |
Jan Klhufek; Alberto Marchisio; Vojtech Mrazek; Lukas Sekanina; Muhammad Shafique; | arxiv-cs.AR | 2026-04-08 |
| 520 | Trilinear Compute-in-Memory Architecture for Energy-Efficient Transformer Acceleration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present TrilinearCIM, a Double-Gate FeFET (DG-FeFET)-based architecture that uses back-gate modulation to realize a three-operand multiply-accumulate primitive for in-memory attention computation without dynamic ferroelectric reprogramming. |
Md Zesun Ahmed Mia; Jiahui Duan; Kai Ni; Abhronil Sengupta; | arxiv-cs.AR | 2026-04-08 |
| 521 | Detecting HIV-Related Stigma in Clinical Narratives Using Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to develop a large language model (LLM)-based tool for identifying HIV stigma from clinical notes. |
ZIYI CHEN et. al. | arxiv-cs.CL | 2026-04-08 |
| 522 | Phishing Detection Techniques: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a systematic review of phishing detection techniques which authorities developed between 2015 and 2025. |
Poorvi .; | International Journal for Research in Applied Science and … | 2026-04-08 |
| 523 | CINESENTIMENT: FILM REVIEW CLASSIFICATION USING DISTILBERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The core sentiment model is a fine-tuned machine learning model developed using the Transformers library within a Google Colab environment. |
Sai Vasanth M; Dharani K; Aravind Chowdary A; Vamsi Krishna J; Jessie David K; | International Scientific Journal of Engineering and … | 2026-04-08 |
| 524 | Inventory Management with Transformer: Automated Decision Making for Order Timing and Quantity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our study highlights the potential of Transformer-based neural networks for large-scale decision making in service operations. |
Mo Liu; Yumo Bai; Meng Qi; Zuo-Jun (Max) Shen; | Service Science | 2026-04-07 |
| 525 | Empirical Characterization of Rationale Stability Under Controlled Perturbations for Explainable Pattern Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a novel metric aimed at assessing the consistency of model explanations, ensuring that models consistently reflect the intended objectives and consistency under label-preserving perturbations. |
Abu Noman Md Sakib; Zhensen Wang; Merjulah Roby; Zijie Zhang; | arxiv-cs.AI | 2026-04-06 |
| 526 | DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using Multi-VALUE’s linguistically grounded transformations, we introduce D3 (Dialectal Disinformation Detection), a corpus of 195K samples derived from established disinformation benchmarks. |
JASON LUCAS et. al. | arxiv-cs.CL | 2026-04-06 |
| 527 | COBOLAssist: Analyzing and Fixing Compilation Errors for LLM-Powered COBOL Code Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, maintaining legacy COBOL systems is increasingly challenging due to a declining pool of skilled developers and the persistence of COBOL errors that require deep domain expertise to resolve. This paper investigates the challenges of COBOL compilation errors and introduces a framework leveraging large language models (LLMs) to address these issues. |
Anh T. V. Dau; Shin Hwei Tan; Jinqiu Yang; Nghi D. Q. Bui; Anh Tuan Nguyen; | arxiv-cs.SE | 2026-04-05 |
| 528 | BWTA: Accurate and Efficient Binarized Transformer By Algorithm-Hardware Co-design Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we analyze zero-point distortion in binarization and propose a Binary Weights & Ternary Activations (BWTA) quantization scheme, which projects tiny values to zero and preserves the accuracy of extremely low-bit models. |
Yifu Ding; Xianglong Liu; Shenghao Jin; Jinyang Guo; Jiwen Lu; | arxiv-cs.LG | 2026-04-05 |
| 529 | Extracting and Steering Emotion Representations in Small Language Models: A Methodological Comparison Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the first comparative analysis of emotion vector extraction methods for SLMs, evaluating 9 models across 5 architectural families (GPT-2, Gemma, Qwen, Llama, Mistral) using 20 emotions and two extraction methods (generation-based and comprehension-based). |
Jihoon Jeong; | arxiv-cs.CL | 2026-04-05 |
| 530 | Aletheia: Gradient-Guided Layer Selection for Efficient LoRA Fine-Tuning Across Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Aletheia, a gradient-guided layer selection method that identifies the most task-relevant layers via a lightweight gradient probe and applies LoRA adapters only to those layers with asymmetric rank allocation. |
Abdulmalek Saket; | arxiv-cs.LG | 2026-04-04 |
| 531 | AutoCompress: Critical Layer Isolation for Efficient Transformer Compression Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present AutoCompress, a transformer compression method motivated by an empirical finding: in small transformers, Layer 0 carries disproportionately high task-critical information, with an NTK-based importance score of 3.6 compared to a maximum of 0.054 for all other layers — a gap of over 60x. |
Archit Thorat; | arxiv-cs.LG | 2026-04-04 |
| 532 | Explainability-Guided Adversarial Attacks on Transformer-Based Malware Detectors Using Control Flow Graphs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a white-box adversarial evasion attack that leverages explainability mechanisms to identify and perturb most influential graph components. |
Andrew Wheeler; Kshitiz Aryal; Maanak Gupta; | arxiv-cs.CR | 2026-04-04 |
| 533 | AI-BASED INTRUSION DETECTION IN IOT SYSTEMS USING BERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive review of recent advancements in applying Natural Language Processing (NLP) techniques, particularly transformer-based Large Language Models such as BERT, GPT-2, and BART, for intrusion detection. |
Patil Sakshi Chandrakant; Musale Aarya Shantilal; Patil Siddhi Ajit; Patil Shivani Nagesh; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-04-04 |
| 534 | DiffSparse: Accelerating Diffusion Transformers with Learned Token Sparsity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, these methods fail to achieve superior acceleration performance in few-step diffusion transformer models due to inefficient feature caching strategies, manually designed sparsity allocation, and the practice of retaining complete forward computations in several steps in these token cache methods. To tackle these challenges, we propose a differentiable layer-wise sparsity optimization framework for diffusion transformer models, leveraging token caching to reduce token computation costs and enhance acceleration. |
HAOWEI ZHU et. al. | arxiv-cs.CV | 2026-04-04 |
| 535 | Abstract 2745: Patients Prefer ChatGPT to Physician Responses in Cancer Communication Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Introduction Large language models like ChatGPT (GPT) are increasingly utilized by patients and physicians yet it is unknown whether either group prefers the output of GPT compared to physician-authored content. |
AARON SEGURA et. al. | Cancer Research | 2026-04-03 |
| 536 | The Spectral Lifecycle of Transformer Training: Transient Compression Waves, Persistent Spectral Gradients, and The Q/K–V Asymmetry Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present the first systematic study of weight matrix singular value spectra \emph{during} transformer pretraining, tracking full SVD decompositions of every weight matrix at 25-step intervals across three model scales (30M–285M parameters). |
Yi Liu; | arxiv-cs.LG | 2026-04-03 |
| 537 | BioUNER: A Benchmark Dataset for Clinical Urdu Named Entity Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this article, we present a gold-standard benchmark dataset for Biomedical Urdu Named Entity Recognition (BioUNER), developed by crawling health-related articles from online Urdu news portals, medical prescriptions, and hospital health blogs and websites. |
Wazir Ali; Adeeb Noor; Sanaullah Mahar; Muhammad Mazhar Younas; | arxiv-cs.CL | 2026-04-03 |
| 538 | From Theory to Practice: Code Generation Using LLMs for CAPEC and CWE Frameworks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The existing software vulnerability datasets frequently fall short in providing comprehensive, detailed code snippets explicitly linked to specific vulnerability descriptions, reducing their utility for advanced research and hindering efforts to develop a deeper understanding of security vulnerabilities. To address this challenge, we present a novel dataset that provides examples of vulnerable code snippets corresponding to Common Attack Pattern Enumerations and Classifications (CAPEC) and Common Weakness Enumeration (CWE) descriptions. |
Murtuza Shahzad; Joseph Wilson; Ibrahim Al Azher; Hamed Alhoori; Mona Rahimi; | arxiv-cs.CR | 2026-04-02 |
| 539 | Erudit AI SaaS: An Artificial Intelligence Tool Based on RoBERTa for Classifying Employee Engagement Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Claudia García-Navarro; Manuel Pulido-Martos; | BMC Psychology | 2026-04-02 |
| 540 | Adversarially Resilient Detection of Business Email Compromise Using Hybrid Machine Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes an adversarially resilient hybrid detection framework that synergistically integrates Natural Language Processing (NLP), classical machine learning models (Support Vector Machines and Random Forest), and deep learning architectures, including Long Short-Term Memory (LSTM) networks and Bidirectional Encoder Representations from Transformers (BERT). |
ALIYU YAHAYA ZAKARI et. al. | Journal of High-Frequency Communication Technologies | 2026-04-02 |
| 541 | Research on Long Sequence Learning Behavior Modeling Based on Transformer-XL Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In view of the difficulties in modeling long-sequence dependence and the high computational complexity of online learning behavior data, this paper proposes a long sequence learning behavior modeling method based on Transformer-XL. |
Yuxiao Qin; | International Scientific Technical and Economic Research | 2026-04-02 |
| 542 | Public Sentiment on Indonesia’s Free Nutritious Meal Program: A Mixed-Methods NLP Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: More specifically, we aimed to identify the key emerging issues. |
Firmansyah Ibrahim; Didik Dwi Prasetya; Andi Baso Kaswar; Hardyanti Pratiwi; | Journal of Health and Nutrition Research | 2026-04-01 |
| 543 | What Are Adversaries Doing? Automating Tactics, Techniques, and Procedures Extraction: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The goal of this study is to aid security researchers in understanding the state of the art in extracting attack tactics, techniques, and procedures (TTPs) from unstructured text by analyzing relevant literature. |
MAHZABIN TAMANNA et. al. | arxiv-cs.SE | 2026-04-01 |
| 544 | From Raw Text to Fairseq RoBERTa: A Modular Snakemake-based Framework Enabling Language-specific BPE Tokenization Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
R. Schmitt; | Softw. Impacts | 2026-04-01 |
| 545 | ChatCM‐RAG: A Deep Learning‐based Natural Language Processing Pipeline for Analysing ChatGPT Applications in Medicine Using BERTopic and Transformer‐based Retrieval‐augmented Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective We developed ChatCM‐RAG, a deep learning pipeline integrating BERTopic with transformer‐based retrieval‐augmented generation to analyse ChatGPT applications in the Medicine literature. |
Cheng Zhang; Ning Wang; Guoming Chen; Yibin Feng; | Clinical and Translational Discovery | 2026-03-31 |
| 546 | A Transformer-Based Method for Bidirectional French–Lingala Machine Translation in Speech and Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose a deep neural network pipeline for bidirectional French–Lingala automatic translation, covering both text-to-text and voice-to-text scenarios, by integrating Long Short-Term Memory (LSTM) and Transformer models on a specialized parallel corpus. |
REAGAN E. MANDIYA et. al. | Applied Sciences | 2026-03-31 |
| 547 | LLMs Underperform on Classifying Anxiety and Depression Using Therapy Conversations: A First-Step Benchmark Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Early and accurate automated detection from naturalistic conversations (e.g., those recorded with a remote chatbot) could eventually improve screening and, in turn, access to timely care. As a first step towards this goal, we aim to evaluate the efficacy of both traditional machine learning and large language models (LLMs) in classifying anxiety and depression from psychotherapy sessions using labels derived from clinician-annotated session metadata reflecting the primary presenting psychiatric concerns. |
Junwei Sun; Siqi Ma; Yiran Fan; Peter Washington; | Applied Sciences | 2026-03-31 |
| 548 | Automatic Identification of Parallelizable Loops Using Transformer-Based Source Code Representations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a Transformer-based approach to classify the parallelization potential of source code, focusing on distinguishing independent (parallelizable) loops from undefined ones. |
Izavan dos S. Correia; Henrique C. T. Santos; Tiago A. E. Ferreira; | arxiv-cs.SE | 2026-03-31 |
| 549 | Fake Review Detection in Online Platforms: A Machine Learning and Deep Learning Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research proposes a hybrid fake review detection system that integrates Natural Language Processing (NLP), traditional machine learning models, and a lightweight transformer-based model, DistilBERT. |
Prakash Wagh; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-03-30 |
| 550 | From Reviews to Requirements: Can LLMs Generate Human-Like User Stories? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we evaluate how well large language models (LLMs) such as GPT-3.5 Turbo, Gemini 2.0 Flash, and Mistral 7B Instruct can generate usable user stories directly from raw app reviews. |
Shadman Sakib; Oishy Fatema Akhand; Tasnia Tasneem; Shohel Ahmed; | arxiv-cs.CL | 2026-03-30 |
| 551 | YouTube Transcript Summarizer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents an AI-driven tool that automatically extracts text from YouTube videos and generates concise, clear summaries using advanced Natural Language Processing (NLP) techniques. |
M Sharath Chandra; K Shanmukh Preetham; R Vishnu; Mr. M Hari Krishna; | INTERNATIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING … | 2026-03-30 |
| 552 | Efficient Domain Adaptation for Text Line Recognition Via Decoupled Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a modular detection-and-correction framework that achieves near-SOTA accuracy with single-GPU training. |
Arundhathi Dev; Justin Zhan; | arxiv-cs.CV | 2026-03-30 |
| 553 | Compressing Transformer Language Models Via Matrix Product Operator Decomposition: A Case Study on PicoGPT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Transformer-based language models achieve strong performance across NLP tasks, but their quadratic parameter scaling with hidden dimension makes deployment on resource-constrained hardware expensive. |
Younes Javanmard; Tanmoy Pandit; Masoud Mardani; | arxiv-cs.CL | 2026-03-30 |
| 554 | GPU-Accelerated Optimization of Transformer-Based Neural Networks for Real-Time Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents the design and evaluation of a GPU-accelerated inference pipeline for transformer models using NVIDIA TensorRT with mixed-precision optimization. |
Soutrik Mukherjee; Sangwhan Cha; | arxiv-cs.LG | 2026-03-30 |
| 555 | Decoding Minds Through Machines: A Transformer-Driven Deep Learning Framework for Mental Health Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this research, a transformer-based model-specifically BERT was used to predict whether individuals are likely to seek mental health treatment based on their survey responses. |
Sujata Patil; Vidya Shinde; | International Journal of Mathematics And Computer Research | 2026-03-30 |
| 556 | On The Applicability of LLMs and SLMs for Privacy-Preserving Named Entity Recognition in Financial Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work explores how deep learning models, with different numbers of parameters, can be effectively applied to detect personal data within unstructured text using Named Entity Recognition (NER) techniques. |
Evgenia Psarra; Kyriakos Stefanidis; | Applied Sciences | 2026-03-30 |
| 557 | Low-Rank Adaptation Reduces Catastrophic Forgetting in Sequential Transformer Encoder Fine-Tuning: Controlled Empirical Evidence and Frozen-Backbone Representation Probes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Sequential fine-tuning of pretrained language encoders often overwrites previously acquired capabilities, but the forgetting behavior of parameter-efficient updates remains under-characterized. We present a controlled empirical study of Low-Rank Adaptation (LoRA) in sequential transformer encoder fine-tuning with companion representation probes that test a frozen-backbone explanation of its robustness. |
Ashish Pandey; | arxiv-cs.LG | 2026-03-29 |
| 558 | Structural Stress and Learned Helplessness in Afghanistan: A Multi-Layer Analysis of The AFSTRESS Dari Corpus Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce AFSTRESS, the first multi-label corpus of self-reported stress narratives in Dari (Eastern Persian), comprising 737 responses collected from Afghan individuals during an ongoing humanitarian crisis. |
Jawid Ahmad Baktash; Mursal Dawodi; Nadira Ahmadi; | arxiv-cs.CL | 2026-03-28 |
| 559 | Clash of The Models: Comparing Performance of BERT-based Variants for Generic News Frame Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Developments in computation, particularly with the introduction of transformer architecture and more so with large language models (LLMs), have naturally prompted scholars to explore various novel computational approaches, especially for deductive frame detection, in recent years. |
Vihang Jumle; | arxiv-cs.CL | 2026-03-27 |
| 560 | Clinical Named Entity Recognition in The Portuguese Language: A Benchmark of Modern BERT Models and LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we aimed to evaluate BERT-based models and large language models (LLMs) for clinical NER in Portuguese and to test strategies for addressing multilabel imbalance. |
VINICIUS ANJOS DE ALMEIDA et. al. | arxiv-cs.CL | 2026-03-27 |
| 561 | EnTaCs: Analyzing The Relationship Between Sentiment and Language Choice in English-Tamil Code-Switching Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We apply a fine-tuned XLM-RoBERTa model for token-level language identification on 35,650 romanized YouTube comments from the DravidianCodeMix dataset, producing per-utterance measurements of English proportion and language switch frequency. |
Paul Bontempo; | arxiv-cs.CL | 2026-03-27 |
| 562 | CROSS-LINGUAL TRANSFORMER-BASED SCREENING OF POST-TRAUMATIC STRESS DISORDER BASED ON COMPARATIVE ANALYSIS OF BERT AND XLM-ROBERTA WITH MACHINE TRANSLATION ADAPTATION FOR UKRAINIAN LANGUAGE Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The paper presents a comprehensive study on automated screening of post-traumatic stress disorder (PTSD) using transformer-based natural language processing models in a cross-lingual setting. |
Andrii FEDORYCHKO; Victoria VYSOTSKA; Lyubomyr CHYRUN; | Computer systems and information technologies | 2026-03-26 |
| 563 | LLM-generated Text Detection: Enhancing Accuracy Using XLM-RoBERTa & DistilBERT Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, this advancement has also raised significant concerns about its potential misuse, particularly in academic and research contexts, such as the production of unoriginal or fabricated content. To address this challenge, we propose a robust approach for detecting artificial intelligence (AI)-generated text that emphasizes a deep understanding of linguistic context. |
AKABRA JAVED et. al. | PeerJ Computer Science | 2026-03-26 |
| 564 | Towards Explainable Graph Spectral Clustering for BERT Embeddings Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a novel theoretical methodology for explanation, based on the premise that document similarity in GSC is computed as cosine similarity of BERT embeddings of documents. |
Mieczysław Kłopotek; Sławomir T. Wierzchoń; Bartłomiej Starosta; Piotr Borkowski; Dariusz Czerski; | Journal of Automation, Mobile Robotics and Intelligent … | 2026-03-25 |
| 565 | Vision Transformer Architectures for High-Resolution Medical Image Segmentation and Disease Localization: Architectural Advances, Localization Strategies, and Benchmark Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A unified conceptual framework is proposed that combines patch-based tokenization, hierarchical transformer encoding, multi-scale feature fusion, and decoder-based reconstruction for dense segmentation, alongside an auxiliary localization module informed by attention and gradient-based interpretability techniques. |
Xin Nie; Mengmin Du; | Clinical Medicine And Health Research Journal | 2026-03-25 |
| 566 | Decoding Market Emotions in Cryptocurrency Tweets Via Predictive Statement Classification with Machine Learning and Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a novel classification framework for identifying predictive statements in cryptocurrency-related tweets, focusing on five popular cryptocurrencies: Cardano, Matic, Binance, Ripple, and Fantom. |
MOEIN SHAHIKI TASH et. al. | arxiv-cs.AI | 2026-03-25 |
| 567 | NLP Framework to Safeguard Youngsters Online Using Advanced Transformer-Based Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research study focuses on developing a natural language processing (NLP) framework designed to monitor online interactions and identify online harmful conversations. |
Hisham AbouGrad; Sankar Santhosh; Salem Alsaid; | Journal of Data Science and Intelligent Systems | 2026-03-25 |
| 568 | Leveraging GPT Models for Efficient Integrated Care Projects Planning: An Interactive Workshop Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Audience: This workshop is designed for healthcare professionals, project managers, program evaluators, quality improvement teams, and healthcare system planners who are directly involved in the design, implementation, and evaluation of healthcare interventions. |
Christina Png; Yeuk Fan Ng; | International Journal of Integrated Care | 2026-03-24 |
| 569 | Decoding AI Authorship: Can LLMs Truly Mimic Human Style Across Literature and Politics? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Amidst the rising capabilities of generative AI to mimic specific human styles, this study investigates the ability of state-of-the-art large language models (LLMs), including GPT-4o, Gemini 1.5 Pro, and Claude Sonnet 3.5, to emulate the authorial signatures of prominent literary and political figures: Walt Whitman, William Wordsworth, Donald Trump, and Barack Obama. |
Nasser A Alsadhan; | arxiv-cs.CL | 2026-03-24 |
| 570 | DariMis: Harm-Aware Modeling for Dari Misinformation Detection on YouTube Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We benchmark a Dari/Farsi-specialized model (ParsBERT) against XLM-RoBERTa-base; ParsBERT achieves the best test performance with accuracy of 76.60 percent and macro F1 of 72.77 percent. |
Jawid Ahmad Baktash; Mosa Ebrahimi; Mohammad Zarif Joya; Mursal Dawodi; | arxiv-cs.CL | 2026-03-24 |
| 571 | Depression Subtype Classification from Social Media Posts: Few-shot Prompting Vs. Fine-tuning of Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective We benchmarked few-shot, prompt-only LLMs against parameter-efficient fine-tuned encoders for identifying depression subtypes in posts on X (formerly Twitter). |
Rawan AlSaad; Sulaiman Alshakhs; Rajat Thomas; | Frontiers in Digital Health | 2026-03-23 |
| 572 | Transforming Patient Education on Retinal Detachment A Multilingual Voice Enabled Retrieval Augmented Generation Chatbot Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: GPT 4o outperformed Claude Opus and Gemini 1.5 Pro across all metrics (BLEU 0.56; ROUGE L 0.72; BERTScore F1 0.86). … |
Mohammad Hossein Amirhosseini; | Journal of Artificial Intelligence & Robotics | 2026-03-23 |
| 573 | Neural Network with Multi-Level Attention for Extraction of Clinical Relationships Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods In this work, we propose a deep learning framework that combines RoBERTa, a bidirectional gated recurrent units (Bi-GRU), a multi-level attention mechanism, and a conditional random field (CRF) layer for clinical relation extraction. |
Vijaya Madhavi Lakshmi Challa; Ramakrishnudu Tene; | Journal of Intelligent & Fuzzy Systems: Applications in … | 2026-03-23 |
| 574 | The Library Theorem: How External Organization Governs Agentic Reasoning Capacity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We formalize the transformer context window as an I/O page and prove that tool-augmented agents with indexed external memory achieve exponentially lower retrieval cost than agents restricted to sequential scanning: $O(\log_b N)$ versus $Ω(N)$ page reads per query, and $O(T \log_b T)$ versus $Θ(T^2)$ cumulative cost over $T$ reasoning steps — a gap that widens as deliberation deepens. |
Zachary F. Mainen; | arxiv-cs.AI | 2026-03-22 |
| 575 | Responsible Scaling of Deep Learning for Sustainable Apple Disease Prediction: An Ensemble Learning Approach Using LSTM, Transformer, and Temporal Fusion Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Ishana Attri; Lalit Kumar Awasthi; Teek Parval Sharma; | Theoretical and Applied Climatology | 2026-03-21 |
| 576 | A Context-Aware Model for Sentiment Analysis of Restaurant Reviews Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes an Aspect-Based Sentiment Analysis (ABSA) framework that uses transformer-based models to identify sentiment for each aspect in restaurant reviews. |
Shaveta Koundal; Sandeep Ranjan; | International Journal for Research in Applied Science and … | 2026-03-21 |
| 577 | Structural Sensitivity in Compressed Transformers: Error Propagation, Lyapunov Stability, and Formally Verified Bounds Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A single matrix out of 468 in GPT-2 Small can increase perplexity by 20,000x when compressed, revealing that transformer compression sensitivity spans five orders of magnitude. We map this sensitivity landscape across five architectures (117M-8B parameters), finding a consistent hierarchy: early-layer MLP up-projections are catastrophically sensitive while value projections compress nearly for free. |
Abhinaba Basu; | arxiv-cs.LG | 2026-03-21 |
| 578 | Exploring Data Augmentation and Resampling Strategies for Transformer-Based Models to Address Class Imbalance in AI Scoring of Scientific Explanations in NGSS Classroom Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates augmentation strategies to improve transformer-based text classification of student responses to a physical science assessment based on an NGSS-aligned learning progression. |
Prudence Djagba; Kevin Haudek; Clare G. C. Franovic; Leonora Kaldaras; | arxiv-cs.AI | 2026-03-21 |
| 579 | Trained Persistent Memory for Frozen Decoder-Only LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We adapt six methods — prefix, parallel cross-attention, KV extension, Hebbian memory, context-gated branch, and slot-based sparse write — to a frozen GPT-2, training only a small adapter $θ_{mem}$. |
Hong Jeong; | arxiv-cs.LG | 2026-03-20 |
| 580 | Collaborating with Large Language Models in Literature Screening for A Systematic Review of College Students’ GenAI Literacy Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study evaluated 12 GPT model–prompt configurations in an SLR of 1,616 publications. |
Wonchan Choi; Joyce Lee; Besiki Stvilia; Yan Zhang; Hyerin Bak; | Information Research an international electronic journal | 2026-03-20 |
| 581 | IoT-Driven Enhanced Transformer-Based Prediction of Rock Slope Stability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed IoT-Enhanced Transformer model provides a highly accurate, real-time solution for slope stability prediction, significantly outperforming conventional models. |
Ibrahim Haruna Umar; Hang Lin; Chaoyi Yang; Müge Elif Fırat; | Quarterly Journal of Engineering Geology and Hydrogeology | 2026-03-19 |
| 582 | Detecting Basic Values in A Noisy Russian Social Media Text Data: A Multi-Stage Classification Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a multi-stage classification framework for detecting human values in noisy Russian language social media, validated on a random sample of 7.5 million public text posts. |
Maria Milkova; Maksim Rudnev; | arxiv-cs.CL | 2026-03-19 |
| 583 | FitQT: A Fully Integrated and Trainable Quantum Transformer for Scalable Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces FitQT – a novel, fully integrated, and trainable quantum transformer that unifies three essential components: (1) a quantum multi-head attention mechanism employing a scaled Gaussian kernel, (2) a quantum feed-forward network, and (3) quantum residual connections. |
Lenh Phan Cong Pham; Anh The Le; | Engineering Research Express | 2026-03-19 |
| 584 | Application of BERT-Based Japanese Writing Intelligent Grading System in Blended Teaching Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Japanese writing instruction in foreign language education continues to face challenges such as low correction efficiency, limited error identification, and insufficient personalized feedback. This study examines the application of a BERT-based intelligent grading system within a blended teaching framework to address these issues. |
Ping Yan; | Journal of Advanced Computational Intelligence and … | 2026-03-19 |
| 585 | Computational Investigation of Amorphous Core Materials in Electromagnetic Induction Systems Using Power Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines electromagnetic iron losses in a three-phase power transformer, which serves as a representative electromagnetic induction system. |
RAYAN AJEEB et. al. | Proceedings of the Institution of Mechanical Engineers, … | 2026-03-19 |
| 586 | A Hybrid NER-Sentiment Model for Uzbek Texts: Integrating Lexical, Deep Learning, and Entity-Based Approaches Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This work proposes a hybrid Uzbek sentiment analysis model (sometimes referred to as tonality analysis in the local literature) that integrates contextual text representations … |
B. SAIDOV et. al. | Big Data Cogn. Comput. | 2026-03-19 |
| 587 | Ensemble Learning with Transformer Models for Sentiment Analysis on Cryptocurrency Tweets Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a comparative analysis of three transformer-based language models—RoBERTa, DeBERTa, and FinBERT—for multi-class sentiment classification of cryptocurrency tweets. |
Dr.Mohammed Iqbal Y.; Mohamed Fawaz S.; Dr.Rajakumar M; S Dr.Peerbasha; Dr.Mohamed Surputheen M; | International Scientific Journal of Engineering and … | 2026-03-19 |
| 588 | NeuroGame Transformer: Gibbs-Inspired Attention Driven By Game Theory and Statistical Physics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Standard attention mechanisms in transformers are limited by their pairwise formulation, which hinders the modeling of higher-order dependencies among tokens. We introduce the NeuroGame Transformer (NGT) to overcome this by reconceptualizing attention through a dual perspective: tokens are treated simultaneously as players in a cooperative game and as interacting spins in a statistical physics system. |
Djamel Bouchaffra; Fayçal Ykhlef; Hanene Azzag; Mustapha Lebbah; Bilal Faye; | arxiv-cs.AI | 2026-03-19 |
| 589 | DAPA: Distribution Aware Piecewise Activation Functions for On-Device Transformer Inference and Training Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose Distribution-Aware Piecewise Activation (DAPA), a differentiable and hardware-friendly activation function for Transformer architectures by exploiting the distribution of pre-activation data. |
Maoyang Xiang; Bo Wang; | arxiv-cs.LG | 2026-03-19 |
| 590 | Where Are The Hidden Gems? Applying Transformer Models for Design Discussion Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The goal of this work is to investigate the performance of transformer-based models (i.e., BERT, RoBERTa, XLNet, LaMini-Flan-T5-77M, and ChatGPT-4o-mini) for detecting design-related discussions. |
Lawrence Arkoh; Daniel Feitosa; Wesley K. G. Assunção; | arxiv-cs.SE | 2026-03-18 |
| 591 | Detecting The Machine: A Comprehensive Benchmark of AI-Generated Text Detectors Across Architectures, Domains, and Adversarial Conditions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a comprehensive benchmark evaluating diverse detection approaches across two corpora: HC3 (23,363 human-ChatGPT pairs) and ELI5 (15,000 human-Mistral-7B pairs). |
Madhav S. Baidya; S. S. Baidya; Chirag Chawla; | arxiv-cs.CL | 2026-03-18 |
| 592 | Beyond Muon: MUD (MomentUm Decorrelation) for Faster Transformer Training Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce MUD (MomentUm Decorrelation), a complementary whitening approach that replaces Muon’s polar update with a triangular (Cholesky-like) whitening surrogate inspired by classical Gram–Schmidt and Gauss-Seidel ideas. |
Ben S. Southworth; Stephen Thomas; | arxiv-cs.LG | 2026-03-18 |
| 593 | Humans and Transformer LMs: Abstraction Drives Language Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Categorization is a core component of human linguistic competence. We investigate how a transformer-based language model (LM) learns linguistic categories by comparing its behaviour over the course of training to behaviours which characterize abstract feature-based and concrete exemplar-based accounts of human language acquisition. |
Jasper Jian; Christopher D. Manning; | arxiv-cs.CL | 2026-03-18 |
| 594 | Graph Convolution-based Techniques for Pragmatic Arabic Figurative Language Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a graph-based embedding framework for figurative language classification that captures both syntactic dependencies and semantic relationships using heterogeneous graphs. |
ZOUHEIR BANOU et. al. | Frontiers in Artificial Intelligence | 2026-03-18 |
| 595 | Bhagavad Gita AI: An Intelligent Spiritual Guide Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The Bhagavad Gita AI project presents an innovative fusion of ancient spiritual philosophy and modern Artificial Intelligence. |
Dr. Rahul M. Dhokane; | International Journal for Research in Applied Science and … | 2026-03-18 |
| 596 | CodeT5-RNN: Reinforcing Contextual Embeddings for Enhanced Code Comprehension Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Consequently, how to further refine and enhance LLM embeddings for improved code understanding remains an open research question. To address this gap, we propose a hybrid LLM-RNN framework that reinforces LLM-generated contextual embeddings with a sequential RNN architecture. |
MD MOSTAFIZER RAHMAN et. al. | arxiv-cs.SE | 2026-03-18 |
| 597 | AutoScreen-FW: An LLM-based Framework for Resume Screening Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Moreover, since companies typically do not make resumes with evaluation results publicly available, it remains unclear which resume samples should be used during learning to improve an LLM’s judgment performance. To address these problems, we propose AutoScreen-FW, an LLM-based locally and automatically resume screening framework. |
Zhelin Xu; Shuhei Yamamoto; Atsuyuki Morishima; | arxiv-cs.CL | 2026-03-18 |
| 598 | Capability-Guided Compression: Toward Interpretability-Aware Budget Allocation for Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We term this the capability-blind compression problem and argue it is a root cause of two well-documented failures — the insensitivity of perplexity-based evaluation to reasoning capability loss, and the abrupt phase transitions in model performance recently characterized by Ma et al. (2026). We propose Capability-Guided Compression (CGC), a framework that addresses this by using Sparse Autoencoder (SAE)-derived capability density maps to allocate differential compression budgets across transformer components. |
Rishaank Gupta; | arxiv-cs.LG | 2026-03-17 |
| 599 | Cross-domain Sentiment Classification in Slovak: A Comparative Analysis of Transformer-based Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents the comprehensive evaluation of multi-domain sentiment classification for the Slovak language, addressing a significant gap in existing NLP research. |
Lívia Kelebercová; Dávid Držík; | PeerJ Computer Science | 2026-03-17 |
| 600 | Predictive Analytics for Stock Market Investment Using Advanced Large Language Models (LLMS) in Finance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper focuses on stock market prediction based on the historical stock data of Apple Inc., which consists of Open, High, Low, Adjusted Close, and Volume. |
Sandeep Gupta; | International Journal of Advanced Research in Science … | 2026-03-17 |
| 601 | AI-based Modeling of Treatment Decisions in Benign Prostatic Hyperplasia: A Transformer-based Comparative Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study employs advanced LLMs and deep learning models to predict whether BPH patients were managed with TURP or continued medical therapy using historical clinical data. |
Mohammad Alshraideh; Bahaaldeen Alshraideh; Abedalrahman Alshraideh; Bayan Alfayoumi; | BMC Medical Informatics and Decision Making | 2026-03-17 |
| 602 | Legal Decision Biases in GPT: A Comparison with Human Judgment Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Legal decision-making is expected to meet high standards of consistency and rationality, yet human judgments in this domain are known to be influenced by procedural factors such … |
Toscane F. Bessis; Andy J. Wills; Bartosz W. Wojciechowski; Lee C. White; Emmanuel M. Pothos; | Behavioral Sciences | 2026-03-17 |
| 603 | Tabular LLMs for Interpretable Few-Shot Alzheimer’s Disease Prediction with Multimodal Biomedical Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose TAP-GPT Tabular Alzheimer’s Prediction GPT, a domain-adapted tabular LLM framework built on TableGPT2 and fine-tuned for few-shot AD classification using tabular prompts rather than plain texts. |
SOPHIE KEARNEY et. al. | arxiv-cs.CL | 2026-03-17 |
| 604 | A Multi-Model Approach to English-Bangla Sentiment Classification of Government Mobile Banking App Reviews Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study analyzed 5,652 Google Play reviews in English and Bangla (filtered from 11,414 raw reviews) for four Bangladeshi government banking apps. |
MD. NAIM MOLLA et. al. | arxiv-cs.CL | 2026-03-17 |
| 605 | Indirect Question Answering in English, German and Bavarian: A Challenging Task for High- and Low-Resource Languages Alike Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present two multilingual corpora for IQA of varying quality that both cover English, Standard German and Bavarian, a German dialect without standard orthography: InQA+, a small high-quality evaluation dataset with hand-annotated labels, and GenIQA, a larger training dataset, that contains artificial data generated by GPT-4o-mini. |
Miriam Winkler; Verena Blaschke; Barbara Plank; | arxiv-cs.CL | 2026-03-16 |
| 606 | Fine-tuning RoBERTa for CVE-to-CWE Classification: A 125M Parameter Model Competitive with LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a fine-tuned RoBERTa-base classifier (125M parameters) for mapping Common Vulnerabilities and Exposures (CVE) descriptions to Common Weakness Enumeration (CWE) categories. |
Nikita Mosievskiy; | arxiv-cs.CR | 2026-03-16 |
| 607 | Financial Education in The Age of Artificial Intelligence: A Systematic Review with Text Mining and Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: From a methodological contribution perspective, the study combines bibliometric mapping with text mining and an NLP process that triangulates sentiment using lexicon-based approaches (VADER, TextBlob) and a multilingual transformer model (XLM-RoBERTa), producing continuous indicators (sentiment index) and reproducible research artifacts. |
Eveling Sussety Balcazar-Paiva; Alexander Fernando Haro-Sarango; Juan Amilcar Villanueva-Calderón; | International Journal of Financial Studies | 2026-03-16 |
| 608 | Enhancing Badminton Rule Learning Through A GPT -Integrated LINE Bot Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Over the past 2 years, the Generative Pre‐trained Transformer (GPT), a large language model (LLM) developed by OpenAI, has gained significant momentum across various educational … |
Kuo-Chin Lin; Hui-Chun Hung; Nian-Shing Chen; | J. Comput. Assist. Learn. | 2026-03-16 |
| 609 | Residual Stream Duality in Modern Transformer Architectures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Recent work has made clear that the residual pathway is not mere optimization plumbing; it is part of the model’s representational machinery. We agree, but argue that the cleanest way to organize this design space is through a two-axis view of the Transformer. |
Yifan Zhang; | arxiv-cs.LG | 2026-03-16 |
| 610 | Gait Transformer: End-to-End Transformer Backbone for Gait Recognition Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Gait recognition has emerged as a promising biometric technique for long-distance and non-intrusive human identification. While Transformers have revolutionized vision tasks, … |
SAIHUI HOU et. al. | AAAI Conference on Artificial Intelligence | 2026-03-14 |
| 611 | The GELATO Dataset for Legislative NER Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces GELATO (Government, Executive, Legislative, and Treaty Ontology), a dataset of U.S. House and Senate bills from the 118th Congress annotated using a novel two-level named entity recognition ontology designed for U.S. legislative texts. |
Matthew Flynn; Timothy Obiso; Sam Newman; | arxiv-cs.CL | 2026-03-14 |
| 612 | Generating Effective Ensembles for Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we introduce the hierarchical ensemble construction (HEC) algorithm, a novel greedy-based ensemble method that differs from traditional approaches (e.g., bagging, boosting, stacking) by iteratively building ensembles from scratch using simulated annealing to escape local optima. |
Itay Etelis; Avi Rosenfeld; Abraham Itzhak Weinberg; David Sarne; | International Journal of Data Science and Analytics | 2026-03-14 |
| 613 | Selective Fine-Tuning of GPT Architectures for Parameter-Efficient Clinical Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a parameter-efficient selective fine-tuning framework for adapting GPT-2 to clinical text classification tasks. |
Fariba Afrin Irany; Sampson Akwafuo; | arxiv-cs.CL | 2026-03-14 |
| 614 | Context-Aware Hate Speech Detection Using Transformer-based Models (BERT) for Social Media Text Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a machine learning-based framework for hate speech recognition, leveraging advanced algorithms and data preprocessing techniques to enhance detection accuracy. |
Apoorva Gugri; Srigouri M Gavai; Deeksha Nagesh; Raghavendra K; Shravan M; | International Journal of Scientific Research in Engineering … | 2026-03-14 |
| 615 | Spectral Edge Dynamics of Training Trajectories: Signal-Noise Geometry Across Scales Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Despite hundreds of millions of parameters, transformer training trajectories evolve within only a few coherent directions. We introduce Spectral Edge Dynamics (SED) to quantify … |
Yongzhong Xu; | ArXiv | 2026-03-14 |
| 616 | Spectral Edge Dynamics of Training Trajectories: Signal–Noise Geometry Across Scales Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Despite hundreds of millions of parameters, transformer training trajectories evolve within only a few coherent directions. We introduce \emph{Spectral Edge Dynamics} (SED) to measure this structure: rolling-window SVD of parameter updates reveals a sharp boundary — the \emph{spectral edge} — between coherent optimization directions and stochastic noise, identified by the maximum consecutive singular value ratio $σ_k/σ_{k+1}$. |
Yongzhong Xu; | arxiv-cs.LG | 2026-03-14 |
| 617 | Granularity Paradox: How Emotion Taxonomies Shape GPT-5’s Affective Cognition and Human-AI Alignment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Method This study systematically evaluates the impact of emotion taxonomy on GPT-5’s annotation behavior. |
Fa Zhang; Jian Chen; | Frontiers in Psychology | 2026-03-13 |
| 618 | Intelligent Paraphrase Recognition Using Advanced NLP Techniques Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents an Intelligent Paraphrase Recognition System that leverages advanced Natural Language Processing (NLP) techniques to accurately identify whether two sentences convey the same meaning despite differences in structure or vocabulary. |
D. Ruby; | International Journal for Research in Applied Science and … | 2026-03-13 |
| 619 | Privacy Preserving Topic-wise Sentiment Analysis of The Iran Israel USA Conflict Using Federated Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to analyze global public sentiment regarding the Iran Israel USA conflict by mining user-generated comments from YouTube news channels. |
Md Saiful Islam; Tanjim Taharat Aurpa; Sharad Hasan; Farzana Akter; | arxiv-cs.CL | 2026-03-13 |
| 620 | SOR-BDNet: Semantic-Optical Representation for Boundary-Aware Video Anomaly Detection with GPT-4o Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional reconstruction- and prediction-based methods, relying on motion or appearance patterns learned from normal data, often misclassify previously unseen yet semantically normal events as anomalies. To address this limitation, we propose SOR-BDNet (Semantic-Optical Representation with Boundary Detection Network), an annotation-free multimodal VAD framework that jointly leverages visual appearance and motion dynamics to generate interpretable semantic representations at the frame level. |
YI SUN et. al. | ACM Transactions on Multimedia Computing, Communications, … | 2026-03-13 |
| 621 | GLIEAM: A ViT–GPT Dual-Stream Syntax-Guided Fusion Model for Multimodal Sentiment Analysis and Semantic Communication Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Overall, GLIEAM proposes a noise-resilient, interpretable, and generalizable approach to cross-modal semantic understanding and provides groundwork for advanced applications in semantic communication and multimodal reasoning. |
Haoyang Fei; | International Journal of Computational Intelligence and … | 2026-03-13 |
| 622 | A Multi-AI Agent Framework for Interactive Neurosurgical Education and Evaluation: From Vignettes to Virtual Conversations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We aimed to develop and evaluate a multi–artificial intelligence (AI) agent conversation framework for neurosurgical case assessment that enables realistic clinical interactions through simulated patients and structured access to objective clinical data. |
KARL L. SANGWON et. al. | Neurosurgery Practice | 2026-03-13 |
| 623 | As Language Models Scale, Low-order Linear Depth Dynamics Emerge Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Large language models are often viewed as high-dimensional nonlinear systems and treated as black boxes. |
Buddhika Nettasinghe; Geethu Joseph; | arxiv-cs.LG | 2026-03-12 |
| 624 | Is The GPT Model Suitable for Sentiment Analysis? Testing for Geographical, Political and Gender Bias Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, the GPT-4o-mini model by OpenAI is tested for the presence of geographical, political and gender bias in the case of Polish economic news headlines. |
Agnieszka Choczyńska; | Statistics in Transition new series | 2026-03-12 |
| 625 | The AI-driven Decision-Making (AIDM) Framework: Integrating AHP and ChatGPT for Supplier Selection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a novel approach for supplier selection in manufacturing by combining the Analytic Hierarchy Process (AHP) with the Generative Pre-trained Transformer (GPT). |
Negar Sadeghi; Mohammad Dehghani; Nihan Kabadayi; | Annals of Operations Research | 2026-03-12 |
| 626 | Artificial Intelligence for Sentiment Analysis of Persian Poetry Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In the present work, we employ multiple Bidirectional encoder representations from transformers (BERT) and Generative Pre-trained Transformer (GPT) based language models to analyze the works of two prominent Persian poets: Jalal al-Din Muhammad Rumi (Rumi) and Parvin E’tesami. |
ARASH ZARGAR et. al. | arxiv-cs.CL | 2026-03-11 |
| 627 | The Discrete Charm of The MLP: Binary Routing of Continuous Signals in Transformer Feed-Forward Layers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose that the well-established piecewise-affine characterization of deep networks can be complemented by a routing characterization: along the natural data manifold, the piecewise boundaries implement binary decisions about which tokens need nonlinear processing, routing continuous signals through qualitatively different computational paths. |
Peter Balogh; | arxiv-cs.LG | 2026-03-11 |
| 628 | Evaluating GPT As Automated Analyzer for Detecting Students’ Erroneous Mental Models in Programming Education Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Purpose This study evaluates Large Language Models as automated analyzers for detecting students’ erroneous mental models in programming education, addressing the scalability barrier of labor-intensive manual analysis. |
Francisco J. Gallego-Durán; Patricia Compañ-Rosique; Carlos J. Villagrá-Arnedo; | Universal Access in the Information Society | 2026-03-10 |
| 629 | Reviving ConvNeXt for Efficient Convolutional Diffusion Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here we introduce the fully convolutional diffusion model (FCDM), a model having a backbone similar to ConvNeXt, but designed for conditional diffusion modeling. |
TAESUNG KWON et. al. | arxiv-cs.CV | 2026-03-10 |
| 630 | Maghrebi Dialects – Arabic Bidirectional Translation: An Improved Transformer with Transfer Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a hybrid approach for translating the Maghrebi dialects into/from Modern Standard Arabic (MSA). |
Jihad R’baiti; Youssef Hmamouche; Amal El Fallah Seghrouchni; | Natural Language Processing | 2026-03-10 |
| 631 | Efficient Information Extraction Using LLMs and Knowledge Distillation: A Study on HPV Health Communication Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a dataset consolidating 48 different DOH websites’ data targeted towards HPV and HPV vaccination. |
Saadat Hasan Khan; Kevin Lybarger; | PLOS Digital Health | 2026-03-10 |
| 632 | Detecting Abnormal User Feedback Patterns Through Temporal Sentiment Aggregation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional sentiment analysis methods focus on individual text classification, which is insufficient to capture collective behavioral shifts over time due to inherent noise and class imbalance in short user comments. In this work, we propose a temporal sentiment aggregation framework that leverages pretrained transformer-based language models to extract per-comment sentiment signals and aggregates them into time-window-level scores. |
Yalun Qi; Sichen Zhao; Zhiming Xue; Xianling Zeng; Zihan Yu; | arxiv-cs.CL | 2026-03-10 |
| 633 | Unveiling Patterns in Clinical Data: Exploring The Role of Large Language Models and Clustering Algorithms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective Large Language Models (LLMs) have shown exceptional performance in natural language processing, yet their utility in structured clinical data analysis remains relatively underexplored. This pilot study investigates whether LLM-generated embeddings can preserve the structural integrity of clinical datasets and enhance predictive modeling, particularly in resource-constrained settings. |
ABBAS S. ALI et. al. | Frontiers in Artificial Intelligence | 2026-03-09 |
| 634 | QuadAI at SemEval-2026 Task 3: Ensemble Learning of Hybrid RoBERTa and LLMs for Dimensional Aspect-Based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present our system for SemEval-2026 Task 3 on dimensional aspect-based sentiment regression. |
A. J. W. de Vink; Filippos Karolos Ventirozos; Natalia Amat-Lefort; Lifeng Han; | arxiv-cs.CL | 2026-03-08 |
| 635 | Stock Market Prediction Using Node Transformer Architecture Integrated with BERT Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents an integrated framework combining a node transformer architecture with BERT-based sentiment analysis for stock price forecasting. |
Mohammad Al Ridhawi; Mahtab Haj Ali; Hussein Al Osman; | arxiv-cs.LG | 2026-03-06 |
| 636 | IGLU: The Integrated Gaussian Linear Unit Activation Function Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce IGLU, a parametric activation function derived as a scale mixture of GELU gates under a half-normal mixing distribution. |
Mingi Kang; Zai Yang; Jeova Farias Sales Rocha Neto; | arxiv-cs.LG | 2026-03-06 |
| 637 | BiLAT: A BiLSTM Attention Transformer Model with Hyperparameter Optimization for Robust Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper includes a complete comparative examination of six deep learning models such as RNN, CNN, BERT, GRU, LSTM, and a suggested BiLSTM Attention Transformer (BiLAT) assessed across three benchmark datasets: Fake and Real News, FakeNewsNet, and ISOT Fabricated News. |
T .POORNIMA; | International Journal of Scientific Research in Engineering … | 2026-03-06 |
| 638 | Transparent AI for Mathematics: Transformer-Based Large Language Models for Mathematical Entity Relationship Extraction with XAI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: By combining transformer-based learning, a task-specific dataset, and explainable modeling, this work offers an effective and interpretable framework for MERE, supporting future applications in automated problem solving, knowledge graph construction, and intelligent educational systems. |
Tanjim Taharat Aurpa; | arxiv-cs.CL | 2026-03-06 |
| 639 | Transformer-Based Architectures: The Future of Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The history, design principles, and real-world uses of transformer-based models are all examined in this paper. |
Oluebube Nzube Ezenwankwo; Chime kosisochukwu Martina; | International Journal of Latest Technology in Engineering … | 2026-03-05 |
| 640 | MUTEX: Leveraging Multilingual Transformers and Conditional Random Fields for Enhanced Urdu Toxic Span Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this research, we propose MUTEX: a multilingual transformer combined with conditional random fields (CRF) for Urdu toxic span detection framework that uses manually annotated token-level toxic span dataset to improve performance and interpretability. |
Inayat Arshad; Fajar Saleem; Ijaz Hussain; | arxiv-cs.CL | 2026-03-05 |
| 641 | Structured Multidimensional Representation Learning for Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce a structured spectral factorization of the embedding space based on the L-product for third-order tensors. |
Alaa El Ichi; Khalide Jbilou; Mohamed El Guide; Franck Dufrenois; | arxiv-cs.CL | 2026-03-05 |
| 642 | NCTB-QA: A Large-Scale Bangla Educational Question Answering Dataset and Benchmarking Performance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: These systems tend to produce unreliable responses when correct answers are absent from context. To solve this problem, we introduce NCTB-QA, a large-scale Bangla question answering dataset comprising 87,805 question-answer pairs extracted from 50 textbooks published by Bangladesh’s National Curriculum and Textbook Board. |
Abrar Eyasir; Tahsin Ahmed; Muhammad Ibrahim; | arxiv-cs.CL | 2026-03-05 |
| 643 | Hate Speech Detection Using Large Language Models with Data Augmentation and Feature Enhancement Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper evaluates data augmentation and feature enhancement techniques for hate speech detection, comparing traditional classifiers, e.g., Delta Term Frequency-Inverse Document Frequency (Delta TF-IDF), with transformer-based models (DistilBERT, RoBERTa, DeBERTa, Gemma-7B, gpt-oss-20b) across diverse datasets. |
BRIAN JING HONG NGE et. al. | arxiv-cs.CL | 2026-03-04 |
| 644 | Exploring The Use of A Customised Chatbot on Time-limited Python Tasks in Undergraduate Physics Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Generative artificial intelligence (GenAI) is transforming how novices learn to programme, yet evidence of effective and equitable use in physics education remains limited. We … |
Arin Mizouri; Cristina Zambon; Charlotte Stevenson; | European Journal of Physics | 2026-03-04 |
| 645 | LISTA-Transformer Model Based on Sparse Coding and Attention Mechanism and Its Application in Fault Diagnosis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Specifically, CNN is limited by local receptive fields, while Transformer has shortcomings in effectively modeling local structures, and both face challenges of high model complexity and insufficient interpretability. In response to the above issues, we proposes the following innovative work: A sparse Transformer based on Learnable Iterative Shrinkage Threshold Algorithm (LISTA-Transformer) was designed, which deeply integrates LISTA sparse encoding with visual Transformer to construct a model architecture with adaptive local and global feature collaboration mechanism. |
Shuang Liu; Lina Zhao; Tian Wang; Huaqing Wang; | arxiv-cs.CV | 2026-03-04 |
| 646 | Multi-Label Legal Document Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This project proposes a Multi-Label Legal Document Classification System using Legal-BERT, a transformer model pretrained on large-scale legal corpora. |
SK. Feroz Pasha; R. Rajesh; Indhu Indhu; Swetha Swetha; | International Journal of Scientific Research in Engineering … | 2026-03-04 |
| 647 | Toward Total Recall: Enhancing Data FAIRness Through AI-driven Metadata Standardization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a method that combines Generative Pre-trained Transformer 4 (GPT-4) with structured metadata templates from the Center for Expanded Data Annotation and Retrieval (CEDAR) knowledge base to automatically standardize metadata and to ensure compliance with established standards. |
Sowmya S Sundaram; Rafael S Gonçalves; Mark A Musen; | GigaScience | 2026-03-03 |
| 648 | A Resource-Efficient Approach to Fine-Tuning A BERT-Base Model for Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces SCALE, a novel resource-efficient fine-tuning method that targets the most critical transformer layers, which reduces computational costs without sacrificing performance. |
ABDULLAH M. BASAHEL et. al. | Computers | 2026-03-03 |
| 649 | Half The Nonlinearity Is Wasted: Measuring and Reallocating The Transformer’s MLP Budget Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Through systematic investigation across six models (162M-2.8B parameters), two architectures, and three corpora, we establish that nonlinearity need cannot be predicted from token identity: cross-corpus correlation is zero ($r < 0.05$). |
Peter Balogh; | arxiv-cs.LG | 2026-03-03 |
| 650 | KGLMQA: Enhancing Medical Visual Question Answering with Knowledge Graphs and LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, existing models frequently encounter challenges such as restricted multimodal interaction, insufficient guidance from external medical knowledge, and a lack of rigorous diagnostic logic in their responses. To address these issues, we propose KGLMQA, a novel framework that integrates knowledge graphs with Large Language Models (LLMs). |
Wenhu Wang; Huina Liu; Changfa Wei; | PeerJ Computer Science | 2026-03-03 |
| 651 | Reproducing and Comparing Distillation Techniques for Cross-Encoders Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This lack of comprehensive studies in controlled environments makes it difficult to identify robust design choices. In this work, we reproduce \citet{schlattRankDistiLLMClosingEffectiveness2025} LLM-based distillation strategy and compare it to \citet{hofstatterImprovingEfficientNeural2020} approach based on an ensemble of cross-encoder teachers, as well as other supervised objectives, to fine-tune a large range of cross-encoders, from the original BERT and its follow-ups RoBERTa, ELECTRA and DeBERTa-v3, to the more recent ModernBERT. |
VICTOR MORAND et. al. | arxiv-cs.IR | 2026-03-03 |
| 652 | Divergent Patterns of Probabilistic Reasoning in Humans and GPT-5 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examined GPT‑5’s adherence to classical probability rules, focusing on conjunction fallacies, disjunction fallacies, and violations of binary complementarity. |
Pegah Imannezhad; Emmanuel M. Pothos; Andy J. Wills; | Frontiers in Psychology | 2026-03-03 |
| 653 | GPUTOK: GPU Accelerated Byte Level BPE Tokenization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We built a GPU-based byte-level BPE tokenizer that follows GPT-2’s merge rules. |
Venu Gopal Kadamba; Kanishkha Jaisankar; | arxiv-cs.CL | 2026-03-02 |
| 654 | EvoDropX:Evolutionary Optimization of Feature Corruption Sequences for Faithful Explanations of Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Through comprehensive experiments across multiple datasets (IMDB, Stanford Sentiment Treebank (SST-2), Amazon Polarity (AP)), multiple transformer models (BERT, roberta, distilbert), and multiple metrics (SRG, MIF, LIF, Counterfactual Conciseness (CFC)), we demonstrate that EvoDropX significantly outperforms all state-of-the-art (SOTA) xAI baselines including Attention-Aware Layer-Wise Relevance Propagation for Transformers (AttnLRP), SHapley Additive exPlanations (SHAP), and Local In |
Dhiraj Kumar Singh; Conor Ryan; | Algorithms | 2026-03-02 |
| 655 | Construction Accident Prediction Via Generative AI and AutoML Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To address this research gap, this study provides a structured comparative evaluation of AutoML and a fine-tuned Generative Pre-trained Transformer (GPT) model in terms of predictive performance, training efficiency, robustness under external validation, and operational usability. |
Sungchul Seo; Dahyun Oh; Kyubyung Kang; HyunJung Park; JungHo Jeon; | Applied Sciences | 2026-03-02 |
| 656 | A 3.51 TOPS/mm2 Transformer Accelerator Exploiting Bipolar Sparsity and Approximate Gating Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Transformer models excel at natural language processing tasks but are challenging to deploy on edge devices due to high memory and computation demands. To address this, we … |
ZHONGYU ZHAO et. al. | IEEE Transactions on Circuits and Systems I: Regular Papers | 2026-03-01 |
| 657 | COMET-3D: Compute-in-Memory-Based Transformer Accelerator With Optimized Pipeline and 3D Heterogeneous Integration Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Transformers have become the backbone of large language decoder and encoder models, but their compute- and memory-intensive nature makes them inefficient on traditional von … |
Ashish Reddy Bommana; F. Firouzi; Krishnendu Chakrabarty; | IEEE Transactions on Computer-Aided Design of Integrated … | 2026-03-01 |
| 658 | HypeLoRA: Hyper-Network-Generated LoRA Adapters for Calibrated Language Model Fine-Tuning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work investigates the calibration dynamics of LoRA: Low-Rank Adaptation and a novel hyper-network-based adaptation framework as parameter-efficient alternatives to full fine-tuning for RoBERTa. |
Bartosz Trojan; Filip Gębala; | arxiv-cs.CL | 2026-03-01 |
| 659 | Graph Attention Based Prioritization of Disease Responsible Genes from Multimodal Alzheimer’s Network Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose NETRA (Node Evaluation through Transformer-based Representation and Attention), a multimodal graph transformer framework that replaces heuristic centrality metrics with attention-driven relevance scoring. |
Binon Teji; Subhajit Bandyopadhyay; Swarup Roy; | arxiv-cs.LG | 2026-03-01 |
| 660 | Application of Knowledge Distillation in Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper is a review of the fundamental technical strategies in output layer distillation, feature layer distillation and multi-teacher assisted distillation. |
Xinyi Xu; | Science and Technology of Engineering, Chemistry and … | 2026-02-28 |
| 661 | An Empirical Comparison of BERT and Lightweight Variants for IMDb Sentiment Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper discusses the implications of findings for selecting sentiment analysis models under different application requirements. |
Xuanyu Wu; | Science and Technology of Engineering, Chemistry and … | 2026-02-28 |
| 662 | Evolution and Challenges of Natural Language Processing Technologies Based on Text Understanding and Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: With their powerful data learning capabilities and deep semantic understanding, they have greatly expanded the application boundaries of natural language processing, showing unprecedented performance in everything from machine translation to text generation. This paper systematically reviews the development of natural language processing technology, analyzes the evolution from early rule-based methods to modern neural networks and pre-trained models (such as RNN, Transformer, BERT, GPT, etc.), and discusses the current status and problems of these technologies in text understanding, generation and cross-modal applications. |
Junjie Li; | Science and Technology of Engineering, Chemistry and … | 2026-02-28 |
| 663 | Optimizing Regulatory Compliance with Machine Learning: Boosting Accuracy and Efficiency Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, I describe a new ML pipeline that includes an XGBoost classifier, BERT transformer for NLP, and Isolation Forest for anomaly detection using Snowflake’s ML functions and Apache Spark’s GPU clusters to allow for easy scalability. |
Mohan Kumar Sonne Gowda; | International Journal of Scientific Research in Engineering … | 2026-02-27 |
| 664 | Analysing AI Utilisation in Education Through Learner Question Types: A Constructivist Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigated the evolving role of artificial intelligence (AI) in higher education by analysing learner-generated questions through a constructivist framework. |
Hyunmin Lee; Amara Atif; Kyeong Kang; | Australasian Journal of Educational Technology | 2026-02-27 |
| 665 | Benchmarking BERT-based Models for Sentence-level Topic Classification in Nepali Language Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study benchmarks multilingual, Indic, Hindi, and Nepali BERT variants to evaluate their effectiveness in Nepali topic classification. |
Nischal Karki; Bipesh Subedi; Prakash Poudyal; Rupak Raj Ghimire; Bal Krishna Bal; | arxiv-cs.CL | 2026-02-27 |
| 666 | A Personalized Emotional Therapy System Driven By Generative Artificial Intelligence and Transformer Technology Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study takes 214 adult users aged 18-45 as the object, and collects their music behavior data and text evaluation in the past half year for experimental verification. |
Dongyun Chang; | Journal of Mechanics in Medicine and Biology | 2026-02-27 |
| 667 | Evaluating The Capability of Base and Large-scale Language Models for Multilingual Sarcasm Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study assesses the efficacy of base-scale pre-trained models (Bidirectional Encoder Representations from Transformers (BERT), Robustly Optimized BERT Pretraining Approach (RoBERTa), Unsupervised Cross-lingual Representation Learning at Scale (XLM-RoBERTa), and Distilled version of BERT (DistilBERT)) via task-specific fine-tuning, and large language models (LLMs) (GPT-4) in few-shot contexts, across three languages: English, Spanish, and Amharic. |
Girma Yohannis Bade; Olga Kolesnikova; Jose Luis Oropeza; Moein Shahiki Tash; | PeerJ Computer Science | 2026-02-27 |
| 668 | Tracing The Evolution of Word Embedding Techniques in Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our study covers four major embedding paradigms, statistical representation-based methods (one-hot encoding, bag-of-words, TF-IDF), static word embeddings (Word2Vec, GloVe, FastText), contextual word embeddings (ELMo, BERT, GPT), and sentence/document embeddings, critically discussing the strengths, limitations, and intellectual lineage connecting each category. |
Minh Anh Nguyen; Kuheli Sai; Minh Nguyen; | arxiv-cs.CY | 2026-02-26 |
| 669 | Artificial Intelligence Simplification of English and Spanish Surgical Consent Forms: 1-Size Does Not Fit All Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods: We evaluated three strategies to improve consent readability. |
MORGAN F PETTIGREW et. al. | Journal of the American College of Surgeons | 2026-02-26 |
| 670 | Residual Koopman Spectral Profiling for Predicting and Preventing Transformer Training Instability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: They therefore need an expected probability of failure for a transformer before training starts. Our study of Residual Koopman Spectral Profiling (RKSP) provides such an estimate. |
Bum Jun Kim; Shohei Taniguchi; Makoto Kawano; Yusuke Iwasawa; Yutaka Matsuo; | arxiv-cs.LG | 2026-02-26 |
| 671 | Transformers Converge to Invariant Algorithmic Cores Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Large language models exhibit sophisticated capabilities, yet understanding how they work internally remains a central challenge. A fundamental obstacle is that training selects … |
Joshua S. Schiffman; | arxiv-cs.LG | 2026-02-25 |
| 672 | Small Wins Big: Comparing Large Language Models and Domain Fine-Tuned Models for Sarcasm Detection in Code-Mixed Hinglish Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study compares four large language models, Llama 3.1, Mistral, Gemma 3, and Phi-4, with a fine-tuned DistilBERT model for sarcasm detection in code-mixed Hinglish text. |
Bitan Majumder; Anirban Sen; | arxiv-cs.CL | 2026-02-25 |
| 673 | The Evolution of The Brain From Analog to Digital: The Birth of Artificial Intelligence Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review outlines the evolution of machine learning from handcrafted feature extraction to deep representation learning, explains the fundamental principles of transformer architecture and probabilistic text generation, and summarizes current applications of LLMs within plastic and reconstructive surgery. |
George R. Nahass; Pravin K. Patel; | Journal of Craniofacial Surgery | 2026-02-25 |
| 674 | ANCHOLIK-NER: A Benchmark Dataset for Bangla Regional Named Entity Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: While NER systems for Standard Bangla have made progress, no existing resources or models specifically address the challenge of regional dialects such as Barishal, Chittagong, Mymensingh, Noakhali, and Sylhet, which exhibit unique linguistic features that existing models fail to handle effectively. To fill this gap, we introduce ANCHOLIK-NER, the first benchmark dataset for NER in Bangla regional dialects, comprising 17,405 sentences and 101,817 words annotated with 10 entity tags across 5 regions. |
BIDYARTHI PAUL et. al. | PLOS One | 2026-02-25 |
| 675 | Overton Pluralistic Reinforcement Learning for Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces OP-GRPO (Overton Pluralistic Group Relative Policy Optimization), a reinforcement learning framework for implicit Overton Pluralism that enables a single large language model to produce pluralistic responses without explicit prompting or modular orchestration. |
Yu Fu; Seongho Son; Ilija Bogunovic; | arxiv-cs.CL | 2026-02-24 |
| 676 | Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: They fail to fully capture orthographic similarities and morphological variations, especially in highly inflected and under-resource languages. To mitigate this problem, we propose to computes word vectors directly from character strings, integrating both semantic and syntactic information. |
Felix Schneider; Maria Gogolev; Sven Sickert; Joachim Denzler; | arxiv-cs.CL | 2026-02-24 |
| 677 | Explicit Grammar Semantic Feature Fusion for Robust Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing models capture features by learning from large corpora with transformer models, which are computationally intensive and unsuitable for resource-constrained environments. Therefore, our proposed study incorporates comprehensive grammatical rules alongside semantic information to build a robust, lightweight classification model without resorting to full parameterised transformer models or heavy deep learning architectures. |
Azrin Sultana; Firoz Ahmed; | arxiv-cs.CL | 2026-02-24 |
| 678 | Enhancing Multilingual Embeddings Via Multi-Way Parallel Text Alignment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we show that training standard pretrained models for cross-lingual alignment with a multi-way parallel corpus in a diverse pool of languages can substantially improve multilingual and cross-lingual representations for NLU tasks. |
Barah Fazili; Koustava Goswami; | arxiv-cs.CL | 2026-02-24 |
| 679 | Improving The Understandability of Clinical Guidelines: Development and Evaluation of A GPT-4–Based Pipeline Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods To align LLM revisions with research evidence and enable comparison with manual editing, the National Health Service Injectable Medicines Guide (IMG) was used as a case study, to which a GPT-4–based pipeline was applied, with prompts based on user testing–derived recommendations for IMG authors. |
Matthew D Jones; Melissa Torgbi; Harish Tayyar Madabushi; | Journal of Medical Internet Research | 2026-02-23 |
| 680 | A Novel Transformer-based Platform for The Prediction and Design of Biosynthetic Gene Clusters for (un)natural Products Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here, we present a transformer-based framework that models functional domains as linguistic units to capture and predict their positional relationships within genomes. |
Tomoki Kawano; Taro Shiraishi; Tomohisa Kuzuyama; Maiko Umemura; | PLOS Computational Biology | 2026-02-23 |
| 681 | Rethinking LoRA for Privacy-Preserving Federated Learning in Large Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our analysis reveals three previously underexplored challenges: (1) gradient coupling caused by the simultaneous update of two asymmetric low-rank matrices, (2) compounded noise amplification under differential privacy, and (3) sharpness of the global aggregated model in the parameter space. To address these issues, we propose LA-LoRA (\textbf{L}ocal \textbf{A}lternating \textbf{LoRA}), a novel approach that decouples gradient interactions and aligns update directions across clients to enhance robustness under stringent privacy constraints. |
Jin Liu; Yinbin Miao; Ning Xi; Junkang Liu; | arxiv-cs.LG | 2026-02-23 |
| 682 | Natural Language Processing Models for Robust Document Categorization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This article presents an evaluation of several machine learning methods applied to automated text classification, alongside the design of a demonstrative system for unbalanced document categorization and distribution. |
Radoslaw Roszczyk; Pawel Tecza; Maciej Stodolski; Krzysztof Siwek; | arxiv-cs.CL | 2026-02-23 |
| 683 | PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research introduces the first large-scale, well-balanced Persian social media text classification dataset, specifically designed to address the lack of comprehensive resources in this domain. |
Isun Chehreh; Ebrahim Ansari; | arxiv-cs.CL | 2026-02-22 |
| 684 | AI-Based Classification of IT Support Requests in Enterprise Service Management Systems Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: In modern organizations, IT Service Management (ITSM) relies on the efficient handling of large volumes of unstructured textual data, such as support tickets and incident reports. … |
Audrius Razma; Robertas Jurkus; | Syst. | 2026-02-21 |
| 685 | Analytic Framework for Imbalanced Traffic Crash Type Classification and Management Using A Hybrid Tabular Transformer Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, research on classifying and managing imbalanced traffic crash data is limited. Therefore, this study proposes an analytic framework for imbalanced traffic crash type classification and management using a hybrid tabular transformer approach. |
MINGZHU ZHENG et. al. | Transportation Research Record: Journal of the … | 2026-02-21 |
| 686 | A Hybrid BERT-ALBERT Model for Text Classification: Improving Accuracy in Document Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper explores a combination(Bidirectional Encoder Representations from Transformers) BERT+ALBERT (A Lite BERT) to enhanceclassification accuracy while reducing computational complexity. |
Xiaokui Liu; | Informatica | 2026-02-21 |
| 687 | Multilingual Sentiment Analysis of Summarized Texts: A Cross-language Study of Text Shortening Effects Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines the impact of extractive and abstractive summarization techniques on sentiment classification in eight typologically diverse languages: English (analytic), German (fusional), French, Spanish, Italian (moderately synthetic), Finnish and Hungarian (agglutinative), and Arabic (root-based inflectional). |
Mikhail Krasitskii; Grigori Sidorov; Olga Kolesnikova; Liliana Chanona-Hernandez; Alexander Gelbukh; | PeerJ Computer Science | 2026-02-20 |
| 688 | LLM-based Multi-agent Poetry Generation in Non-cooperative Environments Related Papers Related Patents Related Grants Related Venues Related Experts Related Code View Save Highlight: Our work argues for a paradigm shift in creative tasks such as automatic poetry generation to include social learning processes (via LLM-based agent modeling) similar to human interaction. |
Ran Zhang; Steffen Eger; | Journal of Language Modelling | 2026-02-20 |
| 689 | Fine-tuning Foundational Models to Code Diagnoses from Veterinary Health Records Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The results of this study contribute to the improvement of the quality of veterinary EHRs by investigating accessible methods for automated coding and support both animal and human health research by paving the way for more integrated and comprehensive health databases that span species and institutions. |
MAYLA R. BOGUSLAV et. al. | PLOS Digital Health | 2026-02-20 |
| 690 | Prediction of Major Solar Flares Using Interpretable Class-dependent Reward Framework with Active Region Magnetograms and Domain Knowledge Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract In this work, we develop, for the first time, a supervised classification framework with class-dependent rewards (CDR) to predict ≥M flares within 24 hr. |
ZIXIAN WU et. al. | Monthly Notices of the Royal Astronomical Society | 2026-02-19 |
| 691 | AI for Urban Well-being: Using NLP To assess User Sentiment in Dubai’s Historic Spaces Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study develops a context-aware natural language processing approach using bidirectional encoder representations from transformers (BERT) to analyse resident sentiment toward urban features in Dubai’s Al Ras district, a traditional commercial area experiencing rapid transformation. |
Sumana Hossain; Yasemin Nielsen; Harpreet Seth; Zyead Ahmad Alwsh; Aya Elshabshiri; | Smart and Sustainable Built Environment | 2026-02-19 |
| 692 | Analisis Sentimen Berbasis Aspek Pada Ulasan Hotel Menggunakan Metode BERT Untuk Rekomendasi Layanan Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Industri perhotelan menghadapi tantangan besar dalam mengelola reputasi di berbagai platform Online Travel Agent (OTA). Objek penelitian ini menghadapi kendala dalam menganalisis … |
Muhammad Tamim Shidqi; Yulistia Yulistia; | Jurnal Ekonomi Manajemen Sistem Informasi | 2026-02-19 |
| 693 | LORA-CRAFT: Cross-layer Rank Adaptation Via Frozen Tucker Decomposition of Pre-trained Attention Weights Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce CRAFT (Cross-layer Rank Adaptation via Frozen Tucker), a parameter-efficient fine-tuning (PEFT) method that applies Tucker tensor decomposition to pre-trained attention weight matrices stacked across transformer layers and trains only small square adaptation matrices on the resulting frozen Tucker factors. |
Kasun Dewage; Marianna Pensky; Suranadi De Silva; Shankadeep Mondal; | arxiv-cs.LG | 2026-02-19 |
| 694 | The Anxiety of Influence: Bloom Filters in Transformer Attention Heads Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Some transformer attention heads appear to function as membership testers, dedicating themselves to answering the question has this token appeared before in the context? We identify these heads across four language models (GPT-2 small, medium, and large; Pythia-160M) and show that they form a spectrum of membership-testing strategies. |
Peter Balogh; | arxiv-cs.LG | 2026-02-19 |
| 695 | Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: They also fail to capture how relevance evolves across layers and how structural components shape decision-making. To address these limitations, we proposed the \textbf{Context-Aware Layer-wise Integrated Gradients (CA-LIG) Framework}, a unified hierarchical attribution framework that computes layer-wise Integrated Gradients within each Transformer block and fuses these token-level attributions with class-specific attention gradients. |
Melkamu Abay Mersha; Jugal Kalita; | arxiv-cs.CL | 2026-02-18 |
| 696 | Video-GPT Via Next Clip Diffusion Related Papers Related Patents Related Grants Related Venues Related Experts Related Code View Save Highlight: Alternatively, the video sequence is good at capturing such details. Motivated by this fact, we propose a concise Video-GPT in this paper by treating video as new language for visual world modeling. |
SHAOBIN ZHUANG et. al. | iclr | 2026-02-17 |
| 697 | Rethinking LoRA for Privacy-Preserving Federated Learning in Large Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our analysis reveals three previously underexplored challenges: (1) gradient coupling caused by the simultaneous update of two asymmetric low-rank matrices, (2) compounded noise amplification under differential privacy, and (3) sharpness of the global aggregated model in the parameter space. To address these issues, we propose LA-LoRA (\textbf{L}ocal \textbf{A}lternating \textbf{LoRA}), a novel approach that decouples gradient interactions and aligns update directions across clients to enhance robustness under stringent privacy constraints. |
Jin Liu; Ning Xi; Yinbin Miao; Junkang Liu; | iclr | 2026-02-17 |
| 698 | NRGPT: An Energy-based Alternative for GPT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a minimal modification of the GPT setting to unify it with the EBM framework. |
NIMA DEHMAMY et. al. | iclr | 2026-02-17 |
| 699 | MOAI: Module-Optimizing Architecture for Non-Interactive Secure Transformer Inference IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, the high computational overhead of HE remains a major bottleneck. To address this challenge, we propose MOAI, an efficient HE-based, non-interactive framework for secure transformer inference. |
LINRU ZHANG et. al. | iclr | 2026-02-17 |
| 700 | LEGACY: A Lightweight Dynamic Gradient Compression Strategy for Distributed Deep Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a **L**ightweight **E**fficient **G**r**A**dient **C**ompression strategy**Y** or LEGACY, which, in theory, can work with any compression technique to produce a simple dynamic counterpart. |
Mostapha Essoullami; El houcine Bergou; Aritra Dutta; | iclr | 2026-02-17 |
| 701 | Adaptive Engram Memory System for Indonesian Language Model: Generative AI Based on TOBA LM for Batak and Minang Language Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents TOBA-LM, a trilingual language model based on GPT-2 architecture with 1.2 billion parameters, trained on a corpus encompassing Indonesian, Batak, and Minangkabau using syllabic-agglutinative tokenization. |
Hokky Situngkir; Kevin Siringoringo; Andhika Bernard Lumbantobing; | arxiv-cs.CL | 2026-02-17 |
| 702 | Transformers Don’t Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and Implications for Mechanistic Interpretability Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here we show that all LN layers can be removed from every GPT-2 model with only a small increase in validation loss (e.g. +0.03 cross-entropy loss for GPT-2 XL). |
Luca Baroni; Galvin Khara; Joachim Schaeffer; Marat Subkhankulov; Stefan Heimersheim; | iclr | 2026-02-17 |
| 703 | DNT: A Deeply Normalized Transformer That Can Be Trained By Momentum SGD Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Previous works show that it is mainly due to a heavy-tailed distribution of the gradients. In this paper, we introduce a Deeply Normalized Transformer (DNT), which is meticulously engineered to overcome this limitation enabling seamless training with vanilla mSGDW while yielding comparable performance to the Transformers trained via AdamW. |
XIANBIAO QI et. al. | iclr | 2026-02-17 |
| 704 | DIFFSPARSE: ACCELERATING DIFFUSION TRANSFORMERS WITH LEARNED TOKEN SPARSITY Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, these methods fail to achieve superior acceleration performance in few-step diffusion transformer models due to inefficient feature caching strategies, manually designed sparsity allocation, and the practice of retaining complete forward computations in several steps in these token cache methods. To tackle these challenges, we propose a differentiable layer-wise sparsity optimization framework for diffusion transformer models, leveraging token caching to reduce token computation costs and enhance acceleration. |
HAOWEI ZHU et. al. | iclr | 2026-02-17 |
| 705 | Error Notebook-Guided, Training-Free Part Retrieval in 3D CAD Assemblies Via Vision-Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose an inference-time adaptation framework that combines corrected Error Notebooks with RAG to substantially improve VLM-based part retrieval. |
Yunqing Liu; Nan Zhang; Zhiming Tan; | iclr | 2026-02-17 |
| 706 | In-Context Algorithm Emulation in Fixed-Weight Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our key idea is to construct prompts that encode an algorithm’s parameters into token representations, creating sharp dot-product gaps that force the softmax attention to follow the intended computation. |
Jerry Yao-Chieh Hu; Hude Liu; Jennifer Yuntong Zhang; Han Liu; | iclr | 2026-02-17 |
| 707 | Human-AI Collaboration in Large Language Model-Integrated Building Energy Management Systems: The Role of User Domain Knowledge and AI Literacy Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study provides foundational insights into human-AI collaboration dynamics and promising development directions in the context of LLM-integrated BEMS and contributes to realizing human-centric LLM-integrated energy systems. |
Wooyoung Jung; Kahyun Jeon; Prosper Babon-Ayeng; | arxiv-cs.HC | 2026-02-17 |
| 708 | Extracting Consumer Insight from Text: A Large Language Model Approach to Emotion and Evaluation Measurement Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces the Linguistic eXtractor (LX), a fine-tuned, large language model trained on consumer-authored text that also has been labeled with consumers’ self-reported ratings of 16 consumption-related emotions and four evaluation constructs: trust, commitment, recommendation, and sentiment. |
STEPHAN LUDWIG et. al. | arxiv-cs.CL | 2026-02-16 |
| 709 | Automated Structuring and Analysis of Unstructured Equipment Maintenance Text Data in Manufacturing Using Generative AI Models: A Comparative Study of Pre-Trained Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a practical generative AI-based framework for structured information extraction that automatically converts unstructured equipment maintenance texts into predefined semantic fields to support predictive maintenance in manufacturing environments. |
Yongju Cho; | Applied Sciences | 2026-02-16 |
| 710 | CGRA-DeBERTa Concept Guided Residual Augmentation Transformer for Theologically Islamic Understanding Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents Hadith QA systems that are efficient, interpretable, and accurate and that scale provide educational materials with necessary theological nuance. |
Tahir Hussain; Saddam Hussain Khan; | arxiv-cs.CL | 2026-02-16 |
| 711 | A Comparative Study of BERT-Based Models for Sarcasm Detection in Social Media Texts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper compares several BERT-based models—BERT, RoBERTa, ALBERT, DistilBERT, and DeBERTa—to assess their effectiveness in sarcasm detection on iSarcasmEval and Sarcasm Corpus V2 datasets. |
Rafael Jiménez Castro; Julia Patricia Sánchez Solís; Vicente García Jiménez; Gilberto Rivera Zárate; Rogelio Florencia Juárez; | International Journal of Combinatorial Optimization … | 2026-02-16 |
| 712 | Multimodal Hierarchical Transformer Enriched By Temporal Features and The CTC Loss Model: Depression Detection Model Based on Multimodal Hierarchical Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
XIAOPING YUE et. al. | Biomed. Signal Process. Control. | |
| 713 | Cyberbullying Detection Based on Hybrid Neural Networks and Multi-Feature Fusion Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, although a single pre-trained model demonstrates strong performance in contextual modeling, it still faces challenges including inadequate feature representation and limited generalization capability in classifying cyberbullying texts. This study proposes a cyberbullying detection model employing BERT-BiGRU-CNN (BBGC) to address this issue. |
Junkuo Cao; Yunpeng Xiong; Weiquan Wang; Guolian Chen; | Information | 2026-02-16 |
| 714 | Thin Keys, Full Values: Reducing KV Cache Via Low-Dimensional Attention Selection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Standard transformer attention uses identical dimensionality for queries, keys, and values ($d_q = d_k = d_v = \dmodel$). Our insight is that these components serve fundamentally different roles, and this symmetry is unnecessary. |
Hengshuai Yao; Guan Wang; | arxiv-cs.LG | 2026-02-16 |
| 715 | Evaluation of GRU with Attention and DistilBERT in Text Classification Tasks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper compares two deep learning models for multiclass text classification in the cyberbullying domain. |
Ana Laura Lezama Sánchez; Mireya Tovar Vidal; | International Journal of Combinatorial Optimization … | 2026-02-16 |
| 716 | Detecting LLM Hallucinations Via Embedding Cluster Geometry: A Three-Type Taxonomy with Measurable Signatures Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a geometric taxonomy of large language model hallucinations based on observable signatures in token embedding cluster structure. |
Matic Korun; | arxiv-cs.CL | 2026-02-15 |
| 717 | Empathy Is Not What Changed: Clinical Assessment of Psychological Safety Across GPT Model Generations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: What changed is the safety posture: crisis detection improved monotonically from GPT-4o to GPT-5-mini (H=13.88, p=0.001), while advice safety declined (H=16.63, p<0.001). Per-turn trajectory analysis — a novel methodological contribution — reveals these shifts are sharpest during mid-conversation crisis moments invisible to aggregate scoring. |
Michael Keeman; Anastasia Keeman; | arxiv-cs.CL | 2026-02-15 |
| 718 | Chemical Language Models for Natural Products: A State-Space Model Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using a dataset of about 1M NPs, we present the first systematic comparison of selective state-space models and transformers for NP-focused tasks, together with eight tokenization strategies including character-level, Atom-in-SMILES (AIS), byte-pair encoding (BPE), and NP-specific BPE. |
Ho-Hsuan Wang; Afnan Sultan; Andrea Volkamer; Dietrich Klakow; | arxiv-cs.LG | 2026-02-14 |
| 719 | Data Security and Privacy in GPT Models: Techniques and Challenges Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a systematic literature review conducted in accordance with the PRISMA 2020 guidelines, analyzing 60 peer-reviewed empirical studies published between 2020 and 2025 in Q1 and Q2 journals indexed in the Web of Science Core Collection. |
David Ghiurău; Daniela Elena Popescu; | Applied Sciences | 2026-02-13 |
| 720 | PoultryLeX-Net: Domain-Adaptive Dual-Stream Transformer Architecture for Large-Scale Poultry Stakeholder Modeling Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents PoultryLeX-Net, a lexicon-enhanced, domain-adaptive dual-stream transformer framework for fine-grained sentiment analysis in poultry-related text. |
STEPHEN AFRIFA et. al. | arxiv-cs.CL | 2026-02-13 |
| 721 | Using Transformer-based Models for Vietnamese Language Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a solution to the problem of detecting whether a sequence of text is Vietnamese based on its orthography and contextual features. |
Son Tran; Phuoc Tran; | PLOS One | 2026-02-13 |
| 722 | Beyond Generalist LLMs: Building and Validating Domain-specific Models with The SpAMCQA Benchmark Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Aim: General-purpose Large Language Models (LLMs) exhibit significant limitations in high-stakes clinical domains such as spondyloarthritis (SpA) diagnosis, yet the absence of specialized evaluation tools precludes the quantification of these failures. This study aims to break this critical evaluation impasse and rigorously test the hypothesis that domain specialization is a necessity for achieving expert-level performance in complex medical diagnostics. |
XIAOJIAN JI et. al. | Artificial Intelligence Surgery | 2026-02-12 |
| 723 | Generalized Discrete Diffusion with Self-Correction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a Self-Correcting Discrete Diffusion (SCDD) model to reformulate pretrained self-correction with explicit state transitions and learn directly in discrete time. |
LINXUAN WANG et. al. | arxiv-cs.LG | 2026-02-12 |
| 724 | Simplifying Radiology Reports with Large Language Models: Privacy-compliant Open- Versus Closed-weight Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study compares closed-weight and in-hospital deployed privacy-compliant open-weight LLMs in generating patient-friendly radiology reports. |
ANNEMARIE KATHARINA PROFF et. al. | European Radiology | 2026-02-12 |
| 725 | CONTEXT-AWARE SENTIMENT CLASSIFICATION OF SOCIAL MEDIA TEXT USING ATTENTION-BASED TRANSFORMER MODELS Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a Context-Aware Sentiment Classification Framework using attention-based transformer models to capture semantic relationships and contextual dependencies within social media text. |
Mr.M.N.Mallikarjuna Reddy; | American Journal of AI Cyber Computing Management | 2026-02-12 |
| 726 | An Explainable Deep Learning Framework for Biosensing Data Interpretation in Biomedical Engineering and Real-time Health Diagnostics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Introduction This work proposes an explainable deep learning framework to transform complex biosignal dynamics into interpretable health assessments. |
Zheng Yang; Weihong Huang; Heng Zhang; | Frontiers in Bioengineering and Biotechnology | 2026-02-12 |
| 727 | CLAS-Net: A Study on Cross-lingual Intelligent Sentiment Analysis Model Fusing Semantic Alignment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a novel cross-lingual sentiment analysis framework, CLAS-Net, designed to address the bottlenecks of current public opinion analysis systems in multilingual scenarios. |
Jia-Qi Wang; | PLOS One | 2026-02-11 |
| 728 | Step-resolved Data Attribution for Looped Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To make SDI practical at transformer scale, we propose a TensorSketch implementation that never materialises per-example gradients. |
Georgios Kaissis; David Mildenberger; Juan Felipe Gomez; Martin J. Menten; Eleni Triantafillou; | arxiv-cs.LG | 2026-02-10 |
| 729 | Training Objectives and Evaluation Metrics for Counterfactual Story Rewriting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Analogously, standard evaluation metrics that weigh all tokens equally may not be able to discriminate effectively between more correct and less correct predictions. For these reasons, in this paper we propose novel training objectives and evaluation metrics that mirror this task more closely, and train and evaluate two Flan-T5 transformer models accordingly. |
Amelie Girard; Inigo Jauregi Unanue; Massimo Piccardi; | ACM Transactions on Asian and Low-Resource Language … | 2026-02-10 |
| 730 | LLM-IARE: An Input-Aware Resilience Estimation Methodology for LLMs Under Hardware Transient Faults Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Traditional Fault Injection (FI) approaches are time consuming due to a large number of repeated executions, which limits their scalability to fast evaluate large-scale resilience. To address these challenges, we propose LLM-IARE, a novel Input-Aware Resilience Estimation Model for LLMs under hardware transient faults. |
Jiajia Jiao; Tainian Zhou; Ran Wen; Yulian Li; Jin Liu; | ACM Transactions on Design Automation of Electronic Systems | 2026-02-09 |
| 731 | A Small-Scale System for Autoregressive Program Synthesis Enabling Controlled Experimentation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a system called Cadmus which includes an integer virtual machine (VM), a dataset composed of true programs of diverse tasks, and an autoregressive transformer model that is trained for under \$200 of compute cost. |
Russ Webb; Jason Ramapuram; | arxiv-cs.AI | 2026-02-09 |
| 732 | Evaluating The Role of ChatGPT in Structured Radiology Reporting: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
SHAHAD ALALAWI et. al. | Medicine | 2026-02-06 |
| 733 | Automated Radiological Report Generation from Breast Ultrasound Images Using Vision and Language Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a multimodal Transformer-based framework for automatic breast ultrasound report generation that integrates visual and textual information through cross-attention mechanisms. |
Shaheen Khatoon; Azhar Mahmood; | Journal of Imaging | 2026-02-06 |
| 734 | Deep Research Capabilities in GPT‐5 Thinking and Gemini 2.5 Pro Improve Citation Integrity and Concordance with American Academy of Orthopaedic Surgeons Anterior Cruciate Ligament and Rotator Cuff Guidelines Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract Purpose To assess whether large language models (LLMs) with advanced reasoning and live web search (LWS) provide recommendations concordant with evidence‐based clinical practice guidelines (CPGs) developed by the American Academy of Orthopaedic Surgeons (AAOS) for anterior cruciate ligament (ACL) and rotator cuff (RC) injury management. |
Hilmi Burak Şengül; Barış Akın; Mahmut Enes Kayaalp; Erdem Aras Sezgin; | Knee Surgery, Sports Traumatology, Arthroscopy | 2026-02-06 |
| 735 | DARWIN: Dynamic Agentically Rewriting Self-Improving Network Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: DARWIN is an evolutionary GPT model, utilizing a genetic-algorithm like optimization structure with several independent GPT agents being trained individually using unique training … |
Henry Jiang; | arxiv-cs.NE | 2026-02-05 |
| 736 | Explainable Turkish E-Commerce Review Classification Using A Multi-Transformer Fusion Framework and SHAP Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to classify Turkish e-commerce reviews as either useful or useless, thereby highlighting high-quality content to support more informed consumer decisions. |
Sıla Çetin; Esin Ayşe Zaimoğlu; | Journal of Theoretical and Applied Electronic Commerce … | 2026-02-05 |
| 737 | Greedy-Gnorm: A Gradient Matrix Norm-Based Alternative to Attention Entropy for Head Pruning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose Greedy-Gradient norm (Greedy-Gnorm), a novel head pruning algorithm that dynamically recalculates head importance after each pruning step. |
Yuxi Guo; Paul Sheridan; | arxiv-cs.LG | 2026-02-04 |
| 738 | Can LLMs Capture Stable Human-generated Sentence Entropy Measures? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here, we address both issues using two large publicly available cloze datasets in German 1 and English 2. |
Estrella Pivel-Villanueva; Elisabeth Frederike Sterner; Franziska Knolle; | arxiv-cs.CL | 2026-02-04 |
| 739 | Alignment Drift in Multimodal LLMs: A Two-Phase, Longitudinal Evaluation of Harm Across Eight Model Releases Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a two-phase evaluation of MLLM harmlessness using a fixed benchmark of 726 adversarial prompts authored by 26 professional red teamers. |
Casey Ford; Madison Van Doren; Emily Dix; | arxiv-cs.CL | 2026-02-04 |
| 740 | Multiclass Hate Speech Detection with RoBERTa-OTA: Integrating Transformer Attention and Graph Convolutional Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose RoBERTa-OTA, which introduces ontology-guided attention mechanisms that process textual features alongside structured knowledge representations through enhanced Graph Convolutional Networks. |
Mahmoud Abusaqer; Jamil Saquer; | arxiv-cs.CL | 2026-02-03 |
| 741 | SOGPTSpotter: Detecting ChatGPT-Generated Answers on Stack Overflow Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a novel approach, SOGPTSpotter, that employs Siamese Neural Networks, leveraging the BigBird model and the Triplet loss, to detect ChatGPT-generated answers on Stack Overflow. |
Suyu Ma; Chunyang Chen; Hourieh Khalajzadeh; John Grundy; | arxiv-cs.SE | 2026-02-03 |
| 742 | A Novel Dual-Layer Deep Learning Architecture for Phishing and Spam Email Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a dual-layer deep learning architecture designed to enhance email security by improving the detection of phishing and spam messages. |
Sarmad Rashed; Caner Ozcan; | Electronics | 2026-02-02 |
| 743 | UAT-LITE: Inference-Time Uncertainty-Aware Attention for Pretrained Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose UAT-LITE, an inference-time framework that makes self-attention uncertainty-aware using approximate Bayesian inference via Monte Carlo dropout in pretrained transformer classifiers. |
ELIAS HOSSAIN et. al. | arxiv-cs.AI | 2026-02-02 |
| 744 | Semi-supervised CAPP Transformer Learning Via Pseudo-labeling Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a semi-supervised learning approach to improve transformer-based CAPP transformer models without manual labeling. |
DENNIS GROSS et. al. | arxiv-cs.LG | 2026-02-01 |
| 745 | MBTI Personality Prediction Using GPT-2 LLM Augmentation and Ensemble Machine Learning Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Devraj Patel; Sunita V. Dhavale; | Multimedia Tools and Applications | 2026-02-01 |
| 746 | Transformer-Based Model for Multilingual Hope Speech Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work various transformers have been implemented and evaluated for hope speech detection for English and Germany. |
Nsrin Ashraf; Mariam Labib; Hamada Nayel; | arxiv-cs.CL | 2026-01-31 |
| 747 | Hierarchical Shift Mixing — Beyond Dense Attention in Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Attempts have been made to replace it with less complex methods, at the cost of reduced performance in most cases. We introduce Hierarchical Shift Mixing (HSM), a general framework for token mixing that distributes pairwise token interactions across Transformer layers rather than computing them densely within each layer. |
Robert Forchheimer; | arxiv-cs.LG | 2026-01-30 |
| 748 | Quantifying Model Uniqueness in Heterogeneous AI Ecosystems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here, we introduce a statistical framework for auditing model uniqueness based on In-Silico Quasi-Experimental Design (ISQED). |
Lei You; | arxiv-cs.AI | 2026-01-30 |
| 749 | Robust and Multicenter Detection of Breast Malignancies Using Vision Transformers in Digital Breast Tomosynthesis Imaging Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This framework focuses on image‐level classification of normal, benign, and malignant cases, rather than lesion‐level detection, utilizing DBT images to enhance accuracy and generalizability. |
AMIT SHARMA et. al. | International Journal of Imaging Systems and Technology | 2026-01-30 |
| 750 | YuriiFormer: A Suite of Nesterov-Accelerated Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a variational framework that interprets transformer layers as iterations of an optimization algorithm acting on token embeddings. |
Aleksandr Zimin; Yury Polyanskiy; Philippe Rigollet; | arxiv-cs.LG | 2026-01-30 |
| 751 | Advanced of LLM Transformers and Zero-shot XGBoost for Accurate Arabic Text Insights and Profit Predictions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract This research proposes an innovative Arabic financial forecasting model that integrates linguistic relation extraction with advanced deep learning and machine learning techniques. |
Sally Mohamed Ali Elmorsy; | Journal of Electrical Systems and Information Technology | 2026-01-30 |
| 752 | MiNER: A Two-Stage Pipeline for Metadata Extraction from Municipal Meeting Minutes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose a two-stage pipeline for metadata extraction from municipal minutes. |
RODRIGO BATISTA et. al. | arxiv-cs.CL | 2026-01-30 |
| 753 | From Generative Modeling to Clinical Classification: A GPT-Based Architecture for EHR Notes Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a GPT-based architecture for clinical text classification that adapts a pretrained decoder-only Transformer using a selective fine-tuning strategy. |
Fariba Afrin Irany; | arxiv-cs.CL | 2026-01-29 |
| 754 | CoFrGeNet: Continued Fraction Architectures for Language Generation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, inspired by continued fractions, we introduce a new function class for generative modeling. |
AMIT DHURANDHAR et. al. | arxiv-cs.CL | 2026-01-29 |
| 755 | Abstract DP042: Explainable Natural Language Processing (NLP) Models to Predict 90-day Mortality of Different Stroke Types from Clinical Note Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods: We used the Medical Information Mart for Intensive Care (MIMIC-IV) database (>40,000 ICU patients; Beth Israel Deaconess Medical Center, 2008–2019) and identified 7,511 patients with acute ischemic stroke, spontaneous intracerebral hemorrhage (ICH), non-traumatic subarachnoid hemorrhage (SAH), and traumatic SAH. |
ANH TUAN TRAN et. al. | Stroke | 2026-01-29 |
| 756 | Metadata Driven Malicious URL Detection Using RoBERTa Large and Multi Source Network Threat Intelligence Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Lina Chen; Liang Meng; | Scientific Reports | 2026-01-29 |
| 757 | Tell Me What I Missed: Interacting with GPT During Recalling of One-Time Witnessed Events Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: LLM-assisted technologies are increasingly used to support cognitive processing and information interpretation, yet their role in aiding memory recall—and how people choose to … |
Suifang Zhou; Qi Gong; Xi Shen; Ray Lc; | Proceedings of the 2026 CHI Conference on Human Factors in … | 2026-01-29 |
| 758 | LAMP: Look-Ahead Mixed-Precision Inference of Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Based on the rounding error analysis of a composition $f(g(\mathrm{x}))$, we provide an adaptive strategy that selects a small subset of components of $g(\mathrm{x})$ to be computed more accurately while all other computations can be carried out with lower accuracy. |
STANISLAV BUDZINSKIY et. al. | arxiv-cs.LG | 2026-01-29 |
| 759 | Tell Me What I Missed: Tell Me What I Missed: Interacting with GPT During Recalling of One-Time Witnessed Events Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We studied participants who watched a short robbery video (approximating a one-time eyewitness scenario) and composed recall statements using either a default GPT or a guided GPT prompted with a standardized eyewitness protocol. |
Suifang Zhou; Qi Gong; Ximing Shen; RAY LC; | arxiv-cs.HC | 2026-01-29 |
| 760 | Strategic Energy Project Investment Decisions Using RoBERTa: A Framework for Efficient Infrastructure Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates the application of the robustly optimized BERT model (RoBERTa) for identifying high-value energy infrastructure projects. |
Recep Özkan; Fatemeh Mostofi; Fethi Kadıoğlu; Vedat Toğan; Onur Behzat Tokdemir; | Buildings | 2026-01-28 |
| 761 | Investigating Transformer Models for Textual Bias Detection in Model, Data, and Dataspace Cards Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Partial fine-tuning (zero-shot) evaluations of models trained only on BABE or synthetic data revealed substantial performance drops on this real-world set. To mitigate this cross-domain loss, we introduce a cascaded, full fine-tuning (few-shot) pipeline in which Transformer models are sequentially fine-tuned on BABE, synthetic text, and a subset of the Hugging Face corpus. |
ANDY DONALD et. al. | AI and Ethics | 2026-01-28 |
| 762 | Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans, GPT-4, and GPT-4o Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Study 1 comprised four experiments (1a, 1b, 2a, 2b) with 588 human participants and 680 GPT-4 outputs; Study 2 included two experiments (3a, 3b) with 751 human participants and 1,080 GPT-4o outputs. |
Lydia Uhler; Verena Jordan; Jürgen Buder; Markus Huff; Frank Papenmeier; | Communications Psychology | 2026-01-28 |
| 763 | Rethinking Intelligence: Brain-like Neuron Network Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: From the perspective of neuroscience, we rethink the formation and evolution of intelligence and proposes a new neural network paradigm, Brain-like Neural Network (BNN). |
Weifeng Liu; | arxiv-cs.NE | 2026-01-27 |
| 764 | A GPT-reinforced Social Robot for Patient Communication: A Pilot Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The aim of this study is to develop and test the feasibility of using a social robot that can convincingly provide health information in patient dialogues within clinical practice, to support patient communication and information exchange. |
JAN-WILLEM J. R. VAN ‘T KLOOSTER et. al. | Frontiers in Digital Health | 2026-01-27 |
| 765 | Transformer-Based Hate Speech Detection in Online Content Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research demonstrates the efficiency of deep learning in detecting hate speech and underscores the need for continued study in order to create fair, reliable, and flexible detection systems. |
P B Jishnu; C Vishnu Mohan; | International Research Journal on Advanced Engineering Hub … | 2026-01-27 |
| 766 | Fine-Grained Emotion Detection on GoEmotions: Experimental Comparison of Classical Machine Learning, BiLSTM, and Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we benchmark three modeling families on the GoEmotions dataset: a TF-IDF-based logistic regression system trained with binary relevance, a BiLSTM with attention, and a BERT model fine-tuned for multi-label classification. |
Ani Harutyunyan; Sachin Kumar; | arxiv-cs.CL | 2026-01-26 |
| 767 | Capsule-enhanced RoBERTa for Hierarchical Sentiment Analysis on Social Media Texts IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Ritu Gauraha; Ayush Kumar Agrawal; Parul Dubey; | Discover Artificial Intelligence | 2026-01-25 |
| 768 | Sentiment Analysis in Social Networks: A Case Study on Student Feedback Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods: This study uses several sentiment analysis approaches, including Support Vector Machine (SVM) along with TF-IDF features, Long Short-Term Memory (LSTM) networks, a hybrid TF-IDF–LSTM model, and transformer-based models like BERT and RoBERTa. |
Manoj Kumar Srivastav; Somsubhra Gupta; | Indian Journal Of Science And Technology | 2026-01-24 |
| 769 | Survey on Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we explore and compare multiple approaches for fake news detection using a common dataset. |
Soham Sathe, Amaan Shaikh, Sahil Shrotri; A. A. Chandorkar Vaishnavi Thakur; | International Journal of Advanced Research in Science … | 2026-01-24 |
| 770 | Seq2Turk: Turkish Spelling Error Correction Using Context-Dependent Sequence-to-Sequence Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we propose a context-dependent spelling correction model for Turkish, which is an agglutinative language with a rich morphology and a complex grammatical structure. |
Burak Aytan; C. Okan Sakar; | ACM Transactions on Asian and Low-Resource Language … | 2026-01-24 |
| 771 | Comparative Analysis of Multimodal Large Language Models GPT-4o and O1 Versus Clinicians in Clinical Case Challenge Questions: Retrospective Cross-sectional Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Jaewon Jung; Hyunjae Kim; SungA Bae; Jin Young Park; | Medicine | 2026-01-23 |
| 772 | Research on The State of Charge Estimation of Electric Forklift Batteries Based on An Improved Transformer Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The mean absolute error (MAE) and root mean square error (RMSE) of the improved Transformer model obtained using the Kalman filter – PCA method are reduced by 26.32% and 27.73% respectively, compared to the single Kalman method. |
Jia Wang; Shenglong Zhang; Xia Hu; | Batteries | 2026-01-23 |
| 773 | A Multimodal Ensemble-Based Framework for Detecting Fake News Using Visual and Textual Features Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To enhance accuracy in fake news detection, this article introduces an ensemble-based framework that integrates textual and visual data using ViLBERT’s two-stream architecture, incorporates VADER sentiment analysis to detect emotional language, and uses Image–Text Contextual Similarity to identify mismatches between visual and textual elements. |
MUHAMMAD ABDULLAH et. al. | Mathematics | 2026-01-21 |
| 774 | TrackletGPT: A Language-like GPT Framework for White Matter Tract Segmentation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This task remains complex, as tracts differ among themselves, across subjects and conditions, yet have similar 3D structure across hemispheres and subjects. To address these challenges, we propose TrackletGPT, a language-like GPT framework which reintroduces sequential information in tokens using tracklets. |
ANOUSHKRIT GOEL et. al. | arxiv-cs.CV | 2026-01-20 |
| 775 | Leveraging DistilBERT-Multilingual for Robust and Efficient AI-Based Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this research, a machine learning–based framework for fake news detection using the DistilBERT-base-multilingual-cased Transformer model is proposed. |
Arpita Kulkarni; Vaishnavi Pawade; Shilpa Mangshetty; | International Research Journal on Advanced Engineering Hub … | 2026-01-20 |
| 776 | Context-aware Graph Meta-learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, the generalization performance of the former is limited by the representation capability of GNNs, while the latter faces the challenge of LLMs understanding graph structures. Therefore, we propose CAGML, a context-aware graph meta-learning model, which learns to generalize to cross-domain and cross-granularity graph tasks using a meta-trained Transformer. |
NINGBO HUANG et. al. | aaai | 2026-01-20 |
| 777 | NucEL: Single-Nucleotide ELECTRA-Style Genomic Pre-training for Efficient and Interpretable Representations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: While masked language modeling (MLM)-based approaches, such as DNABERT and Nucleotide Transformer (NT), achieve strong performance, they are hindered by inefficiencies due to partial token supervision, pre-training/fine-tuning mismatches, and high computational costs. We introduce NucEL, the first ELECTRA-style pre-training framework for genomic foundation models, which overcomes these challenges. |
Ke Ding; Brian Parker; Jiayu Wen; | aaai | 2026-01-20 |
| 778 | PrefixGPT: Prefix Adder Optimization By A Generative Pre-trained Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce PrefixGPT, a generative pre-trained Transformer (GPT) that directly generates optimized prefix adders from scratch. |
Ruogu Ding; Xin Ning; Ulf Schlichtmann; Weikang Qian; | aaai | 2026-01-20 |
| 779 | LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present LLMOrbit, a comprehensive circular taxonomy navigating the landscape of large language models spanning 2019-2025. |
Badri N. Patro; Vijay S. Agneeswaran; | arxiv-cs.LG | 2026-01-20 |
| 780 | EMAformer: Enhancing Transformer Through Embedding Armor for Time Series Forecasting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We attribute this performance gap to unstable inter-channel relationships. To bridge this gap, we propose EMAformer, a simple yet effective model that enhances the Transformer with an auxiliary embedding suite, akin to armor that reinforces its ability. |
Zhiwei Zhang; Xinyi Du; Xuanchi Guo; Weihao Wang; Wenjuan Han; | aaai | 2026-01-20 |
| 781 | Training-Free ANN-to-SNN Conversion for High-Performance Spiking Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, existing methods still suffer from notable limitations, failing to effectively handle nonlinear operations in Transformer architectures and requiring additional fine-tuning processes for pre-trained ANNs. To address these issues, we propose a high-performance and training-free ANN-to-SNN conversion framework tailored for Transformer architectures. |
JINGYA WANG et. al. | aaai | 2026-01-20 |
| 782 | Hate Speech Detection in Nepali Social Media: A Comparative Analysis of Machine Learning and Transformer-based Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study demonstrates that lexicon-enhanced featureengineering significantly improves hate speech detection in lowresourcelanguages and provides practical recommendationsfor developing content moderation systems for Nepali-speakingcommunities. |
Palisha Shakya; Suraj Chand; Lalit Buda Pal; Manil Vaidhya; | Proceedings of International Conference on Innovation in … | 2026-01-19 |
| 783 | Neural Organ Transplantation (NOT): Checkpoint-Based Modular Adaptation for Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce Neural Organ Transplantation (NOT), a modular adaptation framework that enables trained transformer layers to function as reusable transferable checkpoints for domain adaptation. |
Ahmad Al-Zuraiqi; | arxiv-cs.LG | 2026-01-19 |
| 784 | A Novel Hybrid Model for Emotion Detection in Text Through Sequential and Transformer-based Approaches: LSTM Enhanced RoBERTa (LER) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The results show that the hybrid approach efficiently detects intricate emotional cues, thus improving the state of emotion detection for real-world, context-sensitive applications. |
Bilal Khan; Muhammad Usman; Muhammad Binsawad; | Scientific Reports | 2026-01-19 |
| 785 | Early Prediction of Type 2 Diabetes Using Multimodal Data and Tabular Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a novel approach for early Type 2 Diabetes Mellitus (T2DM) risk prediction using a tabular transformer (TabTrans) architecture to analyze longitudinal patient data. |
Sulaiman Khan; Md. Rafiul Biswas; Zubair Shah; | arxiv-cs.CV | 2026-01-19 |
| 786 | Capability-Aware Early-Stage Research Idea Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our approach integrates author information, (inferred) capability presentation, and research ideas through a three-way transformer architecture with flexible fusion mechanisms. |
Renlong Jie; Chen Chu; Zhen Wang; | arxiv-cs.CL | 2026-01-18 |
| 787 | Predictive Prototyping: Evaluating Design Concepts with ChatGPT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a retrieval-augmented generation (RAG) method to emulate design feedback using OpenAI GPT-4o, grounded in prototyping data scraped from Instructables.com to increase access to relevant precedent. |
Hilsann Yong; Bradley A. Camburn; | arxiv-cs.HC | 2026-01-18 |
| 788 | Tolerance Principle and Small Language Model Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We trained BabyBERTa (Huebner et al. 2021), a transformer model optimized for small datasets, on artificial grammars. |
Adam E. Friedman; Stevan Harnad; Rushen Shi; | arxiv-cs.CL | 2026-01-17 |
| 789 | Using Large Language Models to Detect and Debunk Climate Change Misinformation Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The rapid spread of climate change misinformation across digital platforms undermines scientific literacy, public trust, and evidence-based policy action. Advances in Natural … |
Zeinab Shahbazi; Sara Behnamian; | Big Data Cogn. Comput. | 2026-01-17 |
| 790 | NLP-ROPCare: Predicting Retinopathy of Prematurity with Admission Notes Using Natural Language Processing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our study aimed to develop a new prediction model for ROP occurrence and severity, named NLP-ROPCare, using natural language processing (NLP). |
YULIN ZHANG et. al. | BMJ Open Ophthalmology | 2026-01-16 |
| 791 | Efficient Multilingual Name Type Classification Using Convolutional Networks Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a convolutional neural network approach for classifying proper names by language and entity type. |
Davor Lauc; | arxiv-cs.CL | 2026-01-16 |
| 792 | Human-AI Collaborative Inductive Thematic Analysis: AI Guided Analysis and Human Interpretive Authority Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Guided by a Human-Artificial Intelligence Collaborative Inductive Thematic Analysis (HACITA) framework, the study focuses on analytic process rather than substantive findings. |
MATTHEW NYAABA et. al. | arxiv-cs.AI | 2026-01-16 |
| 793 | AI-Driven Objective Structured Clinical Examination Generation in Digital Health Education: Comparative Analysis of Three GPT-4o Configurations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Objective This study aims to evaluate 3 GPT-4o configurations for generating OSCE stations in digital health: (1) standard GPT with a simple prompt and OSCE guidelines; (2) personalized GPT with a simple prompt, OSCE guidelines, and a reference book in digital health; and (3) simulated-agents GPT with a structured prompt simulating specialized OSCE agents and the digital health reference book. |
ZINEB ZOUAKIA et. al. | JMIR Medical Education | 2026-01-15 |
| 794 | LOOKAT: Lookup-Optimized Key-Attention for Memory-Efficient Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose LOOKAT, which applies product quantization and asymmetric distance computation, to transformer architecture by decomposing key vectors into subspaces, learning codebooks and computing attention tables via lookup tables. |
Aryan Karmore; | arxiv-cs.LG | 2026-01-15 |
| 795 | Transformer-Based Cognitive Radio: Adaptive Modulation Strategies Using Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work investigates the application of Transformer models, specifically the GPT-2 architecture, to generate novel modulation schemes for wireless communications. |
Andrea Melis; Andrea Piroddi; Roberto Girau; | arxiv-cs.LG | 2026-01-15 |
| 796 | The OB-GYN Take on GPT: Objective Assessment of Artificial Intelligence Models in Patient Education Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Lindsey Burleson; Ella Boardley; Alexandra LaShell; Anthony Shanks; | Cureus | 2026-01-15 |
| 797 | Advancing Cyberbullying Detection in Low-resource Languages: A Transformer- Stacking Framework for Bengali Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Many existing approaches overlook essential language-specific preprocessing, neglect the integration of advanced transformer-based models, and do not adequately address model validation, scalability, and adaptability. To address these limitations, this study introduces three Bengali-specific preprocessing strategies to enhance feature representation. |
Md. Nesarul Hoque; Rudra Pratap Deb Nath; Abu Nowshed Chy; Debasish Ghose; Md Hanif Seddiqui; | Frontiers in Artificial Intelligence | 2026-01-13 |
| 798 | Knowledge-based Learning in Text-RAG and Image-RAG Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this research, we utilized the NIH Chest X-ray image to train the model and compared it in image-based RAG, text-based RAG, and baseline. |
Alexander Shim; Khalil Saieh; Samuel Clarke; | arxiv-cs.CV | 2026-01-13 |
| 799 | A One-step Generation Model with A Single-Layer Transformer: Layer Number Re-distillation of FreeFlow Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We observe that the 28-layer Transformer architecture of FreeFlow can be characterized as an Euler discretization scheme for an ODE along the depth axis, where the layer index serves as the discrete time step. Therefore, we distill the number of layers of the FreeFlow model, following the same derivation logic as FreeFlow, and propose SLT (Single-Layer Transformer), which uses a single shared DiT block to approximate the depth-wise feature evolution of the 28-layer teacher. |
HAONAN WEI et. al. | arxiv-cs.CV | 2026-01-13 |
| 800 | Temporal Feature Mixed Inverted Transformer: An Inverted Transformer for Effective Real-time Electricity Price Forecasting Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Baichun Wang; Baoxian Huang; Qinglun Zhang; Yan Shi; Hong Men; | Engineering Applications of Artificial Intelligence | 2026-01-13 |
| 801 | An Under-Explored Application for Explainable Multimodal Misogyny Detection in Code-mixed Hindi-English Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a multi-modal and explainable web application for detecting misogyny in text and memes in code-mixed Hindi and English. |
Sargam Yadav; Abhishek Kaushik; Kevin Mc Daid; | arxiv-cs.AI | 2026-01-13 |
| 802 | Contrastive Bi-Encoder Models for Multi-Label Skill Extraction: Enhancing ESCO Ontology Matching with BERT and Attention Mechanisms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a zero-shot skill extraction framework that eliminates the need for manually labeled job-ad training data. |
Yongming Sun; | arxiv-cs.CL | 2026-01-13 |
| 803 | Large Language Model Performance in Clinical Cardiology Multiple Choice Questions; Has Reasoning Improved Performance? Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Purpose This study aims to compare the ability of Chat GPT-4o, GPT-4.5, GPT-o1, DeepSeek, and DeepSeek R1 to accurately respond to cardiology multiple choice questions (MCQs) from a commonly used UK cardiology textbook. |
R Crichton; B Liu; S Hothi; | European Heart Journal – Digital Health | 2026-01-12 |
| 804 | Dynamical Systems Analysis Reveals Functional Regimes in Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Most interpretability approaches emphasise static representations or causal interventions, leaving temporal structure largely unexplored. Drawing on neuroscience, where temporal integration and metastability are core markers of neural organisation, we adapt these concepts to transformer models and discuss a composite dynamical metric, computed from activation time-series during autoregressive generation. |
Hassan Ugail; Newton Howard; | arxiv-cs.AI | 2026-01-11 |
| 805 | Transformer-Based Ensemble Model for Classification of Documents Based on English Vocabulary Words Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: While traditional classifiers and individual deep learning models achieve moderate success, they often struggle to capture both local lexical patterns and long-range contextual dependencies. To overcome these limitations, we propose a Transformer-based Ensemble Model that integrates BERT with Convolutional Neural Networks (CNNs) and Bidirectional Long Short-Term Memory (BiLSTMs) networks. |
Nisar Kangoo; Nisar Wani; | International Journal For Multidisciplinary Research | 2026-01-11 |
| 806 | A Transformer-Based Multi-Task Learning Framework for Sentiment, Emotion, and Sarcasm Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract – A major challenge in natural language processing is determining the emotional intent of text because sentiment, emotion, and sarcasm all work together. |
Rohan Rajendra Chimbaikar; | International Journal of Scientific Research in Engineering … | 2026-01-09 |
| 807 | Enhancing E-Commerce Recommender System Inputs Using Transformer-Based Aspect-Based Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Such aggregation often obscures feature-level preferences and limits the usefulness of sentiment outputs for personalization and decision support. To address this limitation, this study proposes a transformer-based ABSA pipeline that integrates KeyBERT for unsupervised aspect extraction with Bidirectional Encoder Representations from Transformers (BERT) for contextual sentiment classification at the aspect level. |
Dr. Nazima Khanam; Dr Karen Robinson; | International Journal of Latest Technology in Engineering … | 2026-01-09 |
| 808 | Vibe Coding An LLM-powered Theorem Prover Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present Isabellm, an LLM-powered theorem prover for Isabelle/HOL that performs fully automatic proof synthesis. |
Zhe Hou; | arxiv-cs.AI | 2026-01-08 |
| 809 | AI-driven Risk Estimation: A GPT-based Approach to News Monitoring for Manufacturing Resilience Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Our research demonstrates the tool’s potential to enhance proactive risk management in supply chains, validated through testing on both real and augmented datasets. |
Adrian Jacob; Anas Ben Achour; Uwe Teicher; | The International Journal of Advanced Manufacturing … | 2026-01-08 |
| 810 | Cyber Threat Detection and Vulnerability Assessment System Using Generative AI and Large Language Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The proposed RoBERTa model achieved better results than the existing BERT model in terms of accuracy (0.99), recall (0.91), and precision (0.89) respectively. |
Keerthi Kumar. M; Swarun Kumar Joginpelly; Sunil Khemka; Lakshmi. S R; Navin Chhibber; | arxiv-cs.CR | 2026-01-08 |
| 811 | Beyond The Truth: Investigating Election Rumors on Truth Social During The 2024 Election Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Large language models (LLMs) offer unprecedented opportunities for analyzing social phenomena at scale. |
Etienne Casanova; R. Michael Alvarez; | arxiv-cs.AI | 2026-01-08 |
| 812 | A Pilot Study on Multilingual Detection of Irregular Migration Discourse on X and Telegram Using Transformer-Based Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents an exploratory multilingual natural language processing (NLP) framework for detecting irregular migration discourse across five languages. |
Dimitrios Taranis; Gerasimos Razis; Ioannis Anagnostopoulos; | Electronics | 2026-01-08 |
| 813 | Explainable Transformer-Based Modelling for Pathogen-Oriented Food Safety Inspection Grade Prediction Using New York State Open Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The objective of this study is to develop an explainable transformer-based framework for predicting food safety inspection grades using multimodal inspection data. |
Omer Faruk Sari; Mohamed Bader-El-Den; Volkan Ince; | Foods | 2026-01-08 |
| 814 | Large Language Models for Detecting Cyberattacks on Smart Grid Protective Relays Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a large language model (LLM)-based framework for detecting cyberattacks on transformer current differential relays (TCDRs), which, if undetected, may trigger false tripping of critical transformers. |
AHMAD MOHAMMAD SABER et. al. | arxiv-cs.CR | 2026-01-07 |
| 815 | Enhancing Aspect Category Sentiment Analysis Via Prompt-Based Prediction with Semantic Augmentation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, existing methodologies continue to face challenges of data scarcity and insufficient contextual understanding in short text sentiment analysis, with notable accuracy decline when processing implicit aspect expressions. This paper proposes a prompt-based RoBERTa-semantic model (PRSM) to address these complex issues. |
Ziwei Xiao; Junfeng Shen; | Physica Scripta | 2026-01-06 |
| 816 | A Novel Unified Approach to Deepfake Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, a novel architecture for Deepfake detection in images and videos is presented. |
Lord Sen; Shyamapada Mukherjee; | arxiv-cs.CV | 2026-01-06 |
| 817 | Designing Conversational Intelligence: Effect of Large Language Models (GPT-Driven) Platforms for Precision Maternal and Newborn Health Engagement: A Systematic Review Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Large language models (LLM/GPTs), particularly GPT-driven conversational agents, have emerged as scalable, versatile digital health tools capable of delivering evidence-based information, mental health support, and risk stratification for complications such as preeclampsia, gestational diabetes, and preterm birth. Objective This systematic review aimed to synthesize global evidence on the design, implementation, and effectiveness of LLM/GPT-powered chatbots for precision maternal and newborn health engagement. |
Robab rasoli; Fahimeh Ebrahimisadrabadi; Zahra Khedri; Solmaz Sohrabei; | Oxford Open Digital Health | 2026-01-06 |
| 818 | Empirical Comparison of Encoder-Based Language Models and Feature-Based Supervised Machine Learning Approaches to Automated Scoring of Long Essays Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study trained several commonly used encoder-based language models for automated scoring of long essays. |
Kuo Wang; Haowei Hua; Pengfei Yan; Hong Jiao; Dan Song; | arxiv-cs.CL | 2026-01-05 |
| 819 | Is Sanskrit The Most Token-efficient Language? A Quantitative Study Using GPT, Gemini, and SentencePiece Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We use a dataset of 701 parallel verses of the Bhagavad Gita, which comprises three languages-Sanskrit, English, and Hindi along with transliteration of Sanskrit into English. |
Anshul Kumar; | arxiv-cs.CL | 2026-01-05 |
| 820 | Automating The Classification of Economic Activities in Official Statistics: A Comparative Study of Neural Networks and Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper explores the process of automation of the classification of open-ended questions regarding the economic activities of enterprises, in official statistics. |
Helda Curma; Valentina Sinaj; | WSEAS TRANSACTIONS ON COMPUTER RESEARCH | 2026-01-05 |
| 821 | Boosting Accuracy and Interpretability in Multilingual Hate Speech Detection Through Layer Freezing and Explainable AI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we examine the performance of three transformer-based models: BERT-base-multilingual-cased, RoBERTa-base, and XLM-RoBERTa-base with the first eight layers frozen, for multilingual sentiment analysis and hate speech detection. |
Meysam Shirdel Bilehsavar; Negin Mahmoudi; Mohammad Jalili Torkamani; Kiana Kiashemshaki; | arxiv-cs.CL | 2026-01-05 |
| 822 | Power-of-Two Quantization-Aware-Training (PoT-QAT) in Large Language Models (LLMs) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we investigate compressing weights with a special quantization that limits numbers to only power-of-two (PoT). |
Mahmoud Elgenedy; | arxiv-cs.CL | 2026-01-05 |
| 823 | Adversarial Question Answering Robustness: A Multi-Level Error Analysis and Mitigation Study Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We perform comprehensive multi-level error analysis using five complementary categorization schemes, identifying negation confusion and entity substitution as the primary failure modes. |
Agniv Roy Choudhury; Vignesh Ponselvan Rajasingh; | arxiv-cs.CL | 2026-01-05 |
| 824 | Lightweight Transformer Architectures for Edge Devices in Real-Time Applications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We establish real-time performance boundaries and provide a practical 6-step deployment pipeline achieving 8-12x size reduction with less than 2% accuracy degradation. |
Hema Hariharan Samson; | arxiv-cs.LG | 2026-01-04 |
| 825 | Adapting Feature Attenuation to NLP Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Concretely, we adapt the COSTARR framework–originally designed for classification in computer vision–to two modest language models (BERT (base) and GPT-2) trained to label 176 arXiv subject areas. |
Tianshuo Yang; Ryan Rabinowitz; Terrance E. Boult; Jugal Kalita; | arxiv-cs.LG | 2026-01-02 |
| 826 | A Comparative Study of Deep Learning and Transformer Models for Twitter Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we investigate sentiment classification on Twitter using the Sentiment140 dataset and compare traditional deep learning approaches, including CNN, BiLSTM, and GRU, with a lightweight transfer learning model, DistilBERT. |
FATIMA HAFEEZ et. al. | International Journal of Combinatorial Optimization … | 2026-01-02 |
| 827 | AI Chatbot As IFRS Advisory Tool: GPT-4 Experimental Design Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The complexity of International Financial Reporting Standards (IFRS) challenges accounting professionals to navigate intricate judgment calls and estimations. This paper tackles a … |
Todor Tocev; Atanasko Atanasovski; | Intell. Syst. Account. Finance Manag. | 2026-01-02 |
| 828 | Sentiment Analysis of TikTok User Comments on Student Proposal Hearing Videos Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to analyze the sentiment and linguistic patterns of TikTok user comments on a student proposal hearing video to understand audience responses to academic content in digital media. |
Mugi Lestari; | International Journal of Linguistics, Communication, and … | 2026-01-02 |
| 829 | Detecting Hope in Social Media Discourse Using Machine and Deep Learning Classifiers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, hope speech detection has received comparatively limited attention in social media discourse analysis when contrasted with tasks such as hate speech detection. This study addresses this gap by conducting both binary and multiclass classification of hope speech in two languages: (i) English and (ii) Spanish. |
Ahmad Imam Amjad; Hamza Imam Amjad; Grigori Sidorov; | International Journal of Combinatorial Optimization … | 2026-01-02 |
| 830 | AraBART-based Arabic Lemmatization Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we introduce AraBART, the first Arabic model to feature an end-to-end pre-trained encoder-decoder, leveraging the BART architecture. |
Soumia Afartass; Fadoua Ataa Allah; Khalid Minaoui; | WSEAS TRANSACTIONS ON INFORMATION SCIENCE AND APPLICATIONS | 2026-01-02 |
| 831 | Improving Router Security Using BERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we demonstrate that using a high-fidelity eBPF-based system call sensor, together with contrastive augmented learning (which introduces controlled mutations of negative samples), improves detection performance at a low false positive rate. |
John Carter; Spiros Mancoridis; Pavlos Protopapas; Brian Mitchell; Benji Lilley; | arxiv-cs.CR | 2026-01-02 |
| 832 | Comparative Efficiency Analysis of Lightweight Transformer Models: A Multi-Domain Empirical Benchmark for Enterprise NLP Deployment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study conducts a comparative analysis of three prominent lightweight Transformer models – DistilBERT, MiniLM, and ALBERT – across three distinct domains: customer sentiment classification, news topic classification, and toxicity and hate speech detection. |
Muhammad Shahmeer Khan; | arxiv-cs.CL | 2026-01-01 |
| 833 | Mechanistic Interpretability of Animacy Effects on Structure Choice in GPT-2 Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Yue Li; Yan Cong; Elaine J. Francis; | Conference on Computational Natural Language Learning | 2026-01-01 |
| 834 | Zhangpeng at SemEval-2026 Task 10: PsyCoMark – Psycholinguistic Conspiracy Marker Extraction and Detection Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: We describe our system for SemEval-2026 Task 10 on psycholinguistic conspiracy marker extraction and conspiracy detection from English texts (Ghosh et al., 2026). The shared task … |
Peng Zhang; Gehao Lu; | SemEval@ACL | 2026-01-01 |
| 835 | Transformer-Based Dynamic Resource Allocation for Multi-Carrier NOMA Systems Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: We present an attention-based transformer learning approach for dynamic resource allocation in multi-carrier non-orthogonal multiple access (NOMA) downlink systems. We propose … |
Liang Dong; Jun Huang; Robert W. Heath; | IEEE Transactions on Cognitive Communications and Networking | 2026-01-01 |
| 836 | Large Language Models for Beam Prediction Under The MmWave Massive MIMO Hybrid-Field Channels Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Millimeter-wave (mmWave) technology is a key enabling technology for next-generation wireless communications. However, beam prediction (BP) in millimeter-wave massive … |
JIE YANG et. al. | IEEE Communications Letters | 2026-01-01 |
| 837 | Token Titans at BEA 2026 Shared Task 1: Multilingual Lexical Complexity Prediction Via Fine-Tuned XLM-RoBERTa with Ensemble Decoding Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Anubhab Parashar; Sandeep Albert Mathias; | Workshop on Innovative Use of NLP for Building Educational … | 2026-01-01 |
| 838 | Spatial–Spectral Transformer With Patch-Local Mixed-Axis 2-D Rotary Position Embedding for Hyperspectral Image Classification Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Hyperspectral image classification is a critical task in remote sensing, where recent transformer-based methods have shown promising improvements in spatial–spectral feature … |
Zirak Khan; Noyon Dey; K. Kathiravan; Seung-Chul Yoon; Suchendra Bhandarkar; | IEEE Journal of Selected Topics in Applied Earth … | 2026-01-01 |
| 839 | Wangkongqiang at SemEval-2026 Task 10: PsyCoMark- Psycholinguistic Conspiracy Marker Extraction and Detection Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This paper presents our system developed for the SemEval-2026 Task 10: PsyCoMark – Psycholinguistic Conspiracy Marker Extraction and Detection. on Subtask 1: Conspiracy Marker … |
Kongqiang Wang; Qing Tan; | SemEval@ACL | 2026-01-01 |
| 840 | Psy Detectives at SemEval-2026 Task 10: PsyCoMark – Psycholinguistic Conspiracy Marker Extraction and Detection Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Psycholinguistic markers provide interpretable signals for identifying conspiratorial reasoning in online discourse. SemEval-2026 Task 10 (PsyCoMark) couples document-level … |
Roxana Carabas; Anamaria Nacu; Lucian Isac; Daniela Gîfu; | SemEval@ACL | 2026-01-01 |
| 841 | Enhancing Machine Learning Models for Mental Health Classification Through Iterative Training and Text-Based Augmentation Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Machine learning (ML) models are frequently used to classify mental health information from textual data, but their practical use is constrained by their poor interpretability and … |
Suparna Das; K. Khondakar; Hirak Mazumdar; A. Kaushik; Sunil Kumar Singh; | Int. J. Intell. Syst. | 2026-01-01 |
| 842 | BERT-OTA: Enhancing Hate Speech Detection With Ontology-Guided Transformer Attention Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The proliferation of hate speech on social media platforms presents a significant challenge for content moderation, requiring sophisticated detection methods that can understand … |
Mahmoud Abusaqer; Jamil Saquer; Mukulika Ghosh; | IEEE Access | 2026-01-01 |
| 843 | From Raw to Synthetic: Evaluating LLM-Based Data Augmentation for Sentiment Analysis Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This study examines the impact of large language model-based data augmentation on Turkish sentiment analysis by analyzing how synthetic data influences model performance across … |
Busra Tekinay; Mansur Alp Tocoglu; | IEEE Access | 2026-01-01 |
| 844 | Temporal Windowed and Internal Feature (TWIF) Transformer for Attack Detection in Robotics Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Ensuring cybersecurity in robotic systems is critically important, as successful attacks can not only disrupt operations but also cause significant physical damage and safety … |
E. N. Yolaçan; Hande Çavşi Zaim; | IEEE Access | 2026-01-01 |
| 845 | Pfr821 at SemEval-2026 Task 9: Multilingual Polarization Detection Via Hybrid XLM-RoBERTa with Targeted Data Augmentation and Imbalance-Aware Training Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Antoine Durand; Rémi Hamon; Matthieu Pereira; Nathan Boucneau; P. Cintra; | SemEval@ACL | 2026-01-01 |
| 846 | PEU Lab at SemEval-2026 Task 4: Pairwise Text Comparison Using RoBERTa and Ranking Loss Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
HANGCHAO MA et. al. | SemEval@ACL | 2026-01-01 |
| 847 | VARH-AI at SemEval-2026 Task 10: Exploiting Architectural Diversity with Transformer-SSM Ensembles and Confidence-Based Iterative Refinement for Conspiracy Detection IF:3 Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This paper describes our system for SemEval-2026 Task 10 (PsyCoMark), focusing on Sub-task 2: binary conspiracy classification in Reddit submission statements. We present a … |
Hritav Solanki; Shubham Sharma; Manish Prasad; Rakhi Agrawal; Yashvardhan Sharma; | SemEval@ACL | 2026-01-01 |
| 848 | LocoGPT: GPT-Based Multi-Humanoid-Task Policy for Humanoid Locomotion Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Recent studies have explored developing controllers that can generalize across several humanoid robots that differ in shape and size. However, limited studies have developed … |
Siddharth Padmanabhan; Kazuki Miyazawa; Takato Horii; | IEEE Access | 2026-01-01 |
| 849 | Near-Field Channel Estimation for XL-MIMO Via IDiT-Based Variance Exploding SDE Generator Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Extremely large-scale MIMO (XL-MIMO) is regarded as a pivotal enabler for achieving ultra-high spectral efficiency in 6G communications. Near-field channel models, which integrate … |
YING FANG et. al. | IEEE Transactions on Wireless Communications | 2026-01-01 |
| 850 | Khaleesiyali at SemEval-2026 Task 2: Lexicon-Augmented RoBERTa for Valence-Arousal Regression on Ecological Essays Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
E. Tee; | SemEval@ACL | 2026-01-01 |
| 851 | Semantic Vectors at SemEval-2026 Task 9: Robust Multilingual Polarization Detection Via Dual-Encoder Fusion and Expert Ensembling IF:3 Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: We present S EMANTIC V ECTORS , our sys-tem for POLAR@SemEval-2026 Task 9 on multilingual online polarization detection across 22 typologically diverse languages. Polarization is … |
A. Dash; Priyanshu Mittal; Piyush Prashant; Sunil Saumya; | SemEval@ACL | 2026-01-01 |
| 852 | From COCOMO to GPT: A Comprehensive Evaluation of LLM-Based Software Effort Estimation Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Accurate software effort estimation remains a critical yet challenging task in software engineering. While recent advances in Large Language Models (LLMs) have demonstrated … |
Feisal Alaswad; Ieee E. POOVAMMAL Senior Member; And Batoul Aljaddouh; | IEEE Access | 2026-01-01 |
| 853 | MedFuseT: A Transformer-Based Model for Advancing Medical Visual Question Answering Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: As medicine and healthcare continue to evolve, Visual Question Answering (VQA) has emerged as an important application of artificial intelligence. In the medical domain, VQA … |
Hamza Mbarek; O. Elharrouss; Hela Mahersia; N. Litayem; | IEEE Access | 2026-01-01 |
| 854 | Modeling Language As A Sequence of Thoughts Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: On the other hand, cognitive science shows that human comprehension involves converting the input linguistic stream into compact, event-like representations that persist in memory while verbatim form is short-lived. Motivated by this view, we introduce Thought Gestalt (TG) model, a recurrent Transformer that models language at two levels of abstraction – tokens and sentence-level thought states. |
Nasim Borazjanizadeh; James McClelland; | arxiv-cs.CL | 2025-12-31 |
| 855 | WISE: Web Information Satire and Fakeness Evaluation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study develops WISE (Web Information Satire and Fakeness Evaluation) framework which benchmarks eight lightweight transformer models alongside two baseline models on a balanced dataset of 20,000 samples from Fakeddit, annotated as either fake news or satire. |
Gaurab Chhetri; Subasish Das; Tausif Islam Chowdhury; | arxiv-cs.CL | 2025-12-30 |
| 856 | INTEGRASI ALGORITMA NLP UNTUK PENINGKATAN KECERDASAN CHATBOT: KAJIAN LITERATUR DAN ANALISIS PERKEMBANGAN TERKINI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study aims to map NLP algorithm trends, application domains, and future research directions. |
Razan Muhammad Rizqi; Devano Agastya Harshavardana; Sulthan Valeri Osmond R; | Jurnal Riset Teknik Komputer | 2025-12-30 |
| 857 | Detection and Classification of Ideological Texts in The Kazakh Language Using Machine Learning and Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper examines deep learning methods and transformers models for the automatic classification of ideologically charged texts in the Kazakh language. |
Milana Bolatbek; Shynar Mussiraliyeva; Kymbat Baisylbayeva; | Research in Language | 2025-12-30 |
| 858 | Enhancing English Language Learning Through Moral Dilemmas: A Comparative Study of GPT and Human-written Stories Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates the effects of generative pre-trained transformer (GPT) generated vs. human-written moral dilemma stories on English as a foreign language (EFL) learners’ speaking skills, storytelling ability, and behavioral regulation, framed within the theoretical context of embodied cognition. |
Jiaqi Wang; Chengliang Wang; Tong Xiao; Xinyu Zhang; | Language Teaching Research | 2025-12-30 |
| 859 | Weakly‐Aligned Region‐Language Transformer for Real‐Time Artistic Content Detection in SAGIN Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: These challenges include limited on‐board computational capacity, fluctuating bandwidth, and the requirement for fine‐grained visual‐semantic reasoning under weak supervision. To overcome these limitations, we propose the Weakly‐Aligned Region‐Language Transformer (WARL‐Transformer), a novel framework designed for robust AI‐generated content detection under realistic SAGIN constraints. |
Jiayue Yu; Sudip Kumar Sahana; | Transactions on Emerging Telecommunications Technologies | 2025-12-30 |
| 860 | IELTS Writing Revision Platform with Automated Essay Scoring and Adaptive Feedback Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents the design, development, and evaluation of a proposed revision platform assisting candidates for the International English Language Testing System (IELTS) writing exam. |
Titas Ramancauskas; Kotryna Ramancauske; | arxiv-cs.CL | 2025-12-30 |
| 861 | Evaluating The Appropriateness and Safety of Generative AI in Delivering Lifestyle Guidance for Atrial Fibrillation Patients Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study assessed the clinical utility of three Large Language Models (LLMs) for delivering accurate and personalized lifestyle guidance: (1) GPT-4o, (2) a retrieval-augmented model using a curated Q&A database (DB GPT), and (3) a modular RAG model retrieving evidence from PubMed (PubMed GPT). |
Masahiro Makino; Wan Jou She; Panote Siriaraya; Satoaki Matoba; Keitaro Senoo; | Scientific Reports | 2025-12-29 |
| 862 | How Emotional Content in Tweets Drive Funding Success? – A Study on Indian Startups Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Purpose This study aims to explore how emotional expressions in startups’ social media posts influence funding outcomes, with a focus on multiple dimensions of emotions using the Circumplex model. |
Nidhi Singhal; Neelmani Gupta; Deepak Kapur; | Journal of Entrepreneurship in Emerging Economies | 2025-12-29 |
| 863 | Splitwise: Collaborative Edge-Cloud Inference for LLMs Via Lyapunov-Assisted DRL Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose Splitwise, a novel Lyapunov-assisted deep reinforcement learning (DRL) framework for fine-grained, adaptive partitioning of LLMs across edge and cloud environments. |
ABOLFAZL YOUNESI et. al. | arxiv-cs.LG | 2025-12-29 |
| 864 | StressRoBERTa: Cross-Condition Transfer Learning from Depression, Anxiety, and PTSD to Stress Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces StressRoBERTa, a cross-condition transfer learning approach for automatic detection of self-reported chronic stress in English tweets. |
Amal Alqahtani; Efsun Kayi; Mona Diab; | arxiv-cs.CL | 2025-12-29 |
| 865 | Explaining News Bias Detection: A Comparative SHAP Analysis of Transformer Model Decision Mechanisms Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we present a comparative interpretability study of two transformer-based bias detection models: a bias detector fine-tuned on the BABE dataset and a domain-adapted pre-trained RoBERTa model fine-tuned on the BABE dataset, using SHAP-based explanations. |
Himel Ghosh; | arxiv-cs.CL | 2025-12-29 |
| 866 | A Review on Fake News Detection and Personalized Recommendation on Social Media Using BERT Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive review of transformer-based fake news detection approaches, particularly those utilizing Bidirectional Encoder Representations from Transformers (BERT), along with Feder- ated Learning techniques for secure and decentralized personalization. |
Adithya A A; | International Journal for Research in Applied Science and … | 2025-12-28 |
| 867 | HELM-BERT: A Transformer for Medium-sized Peptide Property Prediction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here, we propose HELM-BERT, the first encoder-based peptide language model trained on HELM notation. |
Seungeon Lee; Takuto Koyama; Itsuki Maeda; Shigeyuki Matsumoto; Yasushi Okuno; | arxiv-cs.LG | 2025-12-28 |
| 868 | Enhancing Academic Writing Efficiency with ChatGPT: A Natural Language Processing Framework for Innovation, Opportunities, and Challenges Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The rapid development of artificial intelligence (AI) has opened up new avenues for improving the efficiency and quality of academic writing. This paper presents ChatGPT, an … |
Wen Zhao; | J. ICT Stand. | 2025-12-28 |
| 869 | Fake News Classification in Urdu: A Domain Adaptation Approach for A Low-Resource Language Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We evaluate two widely used multilingual models, XLM-RoBERTa and mBERT, and apply domain-adaptive pretraining using a publicly available Urdu news corpus. |
Muhammad Zain Ali; Bernhard Pfahringer; Tony Smith; | arxiv-cs.CL | 2025-12-27 |
| 870 | GHaLIB: A Multilingual Framework for Hope Speech Detection in Low-Resource Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a multilingual framework for hope speech detection with a focus on Urdu. |
Ahmed Abdullah; Sana Fatima; Haroon Mahmood; | arxiv-cs.CL | 2025-12-27 |
| 871 | Advanced Cross-Validation Framework for Mental Health AI: BERT and Neural Networks Achieve High Accuracy on Mental Chat16K Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a comprehensive analysis of the MentalChat16K dataset, which contains 16,084 mental health conversation pairs (6,338 real clinical interviews and 9,746 synthetic dialogues), using modern deep learning architectures. |
Irfan Ali; | Indian Journal of Artificial Intelligence and Neural … | 2025-12-27 |
| 872 | FedEnsemble: Federated Learning Model for Efficient Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, deploying state-of-the-art transformer-based models in real-world applications poses two key challenges: preserving user data privacy and mitigating the computational overhead associated with large-scale models. This study introduces FedEnsemble, a novel federated learning framework that addresses these challenges through three core innovations: (i) a heterogeneous ensemble of BERT, RoBERTa, and DistilBERT to enhance classification robustness; (ii) an entropy-based attention stacking mechanism that adaptively fuses model outputs according to predictive confidence; and (iii) Quantization-Aware Training (QAT) to compress models while maintaining high accuracy and communication efficiency. |
Hesham Ayman; Shaimaa Haridy; Yasmine M. Afify; Walaa Gad; | Computing | 2025-12-26 |
| 873 | Detecting AI-Generated Paraphrases in Bengali: A Comparative Study of Zero-Shot and Fine-Tuned Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates five transformer-based models: XLMRoBERTa-Large, mDeBERTaV3-Base, BanglaBERT-Base, IndicBERT-Base and MultilingualBERT-Base. |
Md. Rakibul Islam; Most. Sharmin Sultana Samu; Md. Zahid Hossain; Farhad Uz Zaman; Md. Kamrozzaman Bhuiyan; | arxiv-cs.CL | 2025-12-25 |
| 874 | Automatic Replication of LLM Mistakes in Medical Conversations Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Yet, replicating specific mistakes in other LLM models is not straightforward and often requires manual effort. We introduce MedMistake, an automatic pipeline that extracts mistakes LLMs make in patient-doctor conversations and converts them into a benchmark of single-shot QA pairs. |
Oleksii Proniakin; Diego Fajardo; Ruslan Nazarenko; Razvan Marinescu; | arxiv-cs.CL | 2025-12-24 |
| 875 | SentXFormer: A Transformer-enhanced Hybrid Deep Learning Framework for Cross-domain Sentiment Analysis of Customer Reviews IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a deep learning model named SentXFormer, which is a transformer-based hybrid framework that enhances the sentiment classification in heterogeneous domains. |
Ajeet Kumar; Kumar Abhishek; Ahamed Shafeeq B M; | Scientific Reports | 2025-12-24 |
| 876 | Computer Assisted Verbal Autopsy: Comparing Large Language Models to Physicians for Assigning Causes to 6939 Deaths in Sierra Leone from 2019–2022 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We analyzed 6,939 VA records from a random sample of deaths in Sierra Leone (2019–2022) to compare five models: three LLMs (GPT-3.5, GPT-4, GPT-5) and two based on symptom algorithms (InterVA-5, InSilicoVA), against physician-assigned CODs. |
RICHARD WEN et. al. | BMC Medicine | 2025-12-24 |
| 877 | SMART SLM: Structured Memory and Reasoning Transformer, A Small Language Model for Accurate Document Assistance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The user of Engineering Manuals (EM) finds it difficult to read EM s because they are long, have a dense format which includes written documents, step by step procedures, and standard parameter lists for engineering equipment. |
Divij Dudeja; Mayukha Pal; | arxiv-cs.CL | 2025-12-24 |
| 878 | From LSTM to GPT-2: Recurrent and Transformer-Based Deep Learning Architectures for Multivariate High-Liquidity Cryptocurrency Price Forecasting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces a unified and methodologically symmetric comparative framework for multivariate cryptocurrency forecasting, addressing long-standing inconsistencies in prior research where model families, feature sets, and preprocessing pipelines differ across studies. |
Erçin Dinçer; Zeynep Hilal Kilimci; | Symmetry | 2025-12-24 |
| 879 | Performance of Large Language Models in Lung Cancer Clinical Decision-Making: A Comparative Analysis Based on DeepSeek, Grok, and GPT Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Yuyang Zhang; Dandan Yang; Yifan Shi; Ying Liu; | Cureus | 2025-12-22 |
| 880 | Transformer Reconstructed with Dynamic Value Attention Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: I propose a method to decide a value for each query dynamically, which could cut down all the redundant heads, keeping only one. |
Xiaowei Wang; | arxiv-cs.LG | 2025-12-21 |
| 881 | InstructNet: A Novel Approach for Multi-Label Instruction Classification Through Advanced Deep Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study uses the How To articles to determine the multi-label instruction category. |
Tanjim Taharat Aurpa; Md Shoaib Ahmed; Md Mahbubur Rahman; Md. Golam Moazzam; | arxiv-cs.CL | 2025-12-20 |
| 882 | A Review on Handwritten Malayalam to English Digitization and Translation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The review covers a range of methodologies, including classical Optical Character Recognition (OCR) techniques, statistical machine translation (SMT), neural machine translation (NMT), and new Vision Language Models (VLMs). |
Mohammed Farhan; | International Journal for Research in Applied Science and … | 2025-12-20 |
| 883 | BERT-MED Chatbot for Healthcare Assistance Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces BERT-MED, an AI-driven healthcare assistance system that follows templates. |
Champa M S; Prajwal D R; Puneeth H K; Pola Manoj Kumar; Shreyank T N; | International Journal of Scientific Research in Engineering … | 2025-12-20 |
| 884 | Bangla MedER: Multi-BERT Ensemble Approach for The Recognition of Bangla Medical Entity Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A major challenge in MedER for low-resource languages is the lack of annotated datasets. To address this issue, we developed a high-quality dataset tailored for the Bangla MedER task. |
TANJIM TAHARAT AURPA et. al. | arxiv-cs.CL | 2025-12-19 |
| 885 | ESummarizer AI Service- Document Summarization Model Using BART Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research presents eSummarizer AI Service, a custom Transformer-based machine learning model designed specifically for summarizing Indian government documents such as policy papers, circulars, legislative texts, and departmental reports. |
Rongdeep Pathak; Mriganka Mohan Bora; Nelson R Varte; | International Journal of Latest Technology in Engineering … | 2025-12-19 |
| 886 | ScoutGPT: Capturing Player Impact from Team Action Sequences Using GPT-Based Framework Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Existing evaluation practices often rely on static summary statistics or post-hoc value models, which fail to capture how a player’s contribution adapts to a new tactical environment or different teammates. To address this gap, we introduce EventGPT, a player-conditioned, value-aware next-event prediction model built on a GPT-style autoregressive transformer. |
MIRU HONG et. al. | arxiv-cs.AI | 2025-12-19 |
| 887 | Advances and Challenges in Semantic Textual Similarity: A Comprehensive Survey Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This survey reviews progress across six key areas: transformer-based models, contrastive learning, domain-focused solutions, multi-modal methods, graph-based approaches, and knowledge-enhanced techniques. |
Lokendra Kumar; Neelesh S. Upadhye; Kannan Piedy; | arxiv-cs.CL | 2025-12-19 |
| 888 | Confidence-Credibility Aware Weighted Ensembles of Small LLMs Outperform Large LLMs in Emotion Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces a confidence-weighted, credibility-aware ensemble framework for text-based emotion detection, inspired by Condorcet’s Jury Theorem (CJT). |
Menna Elgabry; Ali Hamdi; | arxiv-cs.CL | 2025-12-19 |
| 889 | A Review of Sentiment Analysis Research Based on BERT and Its Improved Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: A series of improved models such as RoBERTa-wwm-ext, ERNIE, ALBERT-zh, and MacBERT continued to refresh performance records in Chinese sentiment analysis tasks. This paper systematically reviews the research progress of sentiment analysis based on BERT and its improved models in recent years. |
Jingxuan Chen; | Science and Technology of Engineering, Chemistry and … | 2025-12-19 |
| 890 | A Hybrid Deep Learning Model Based on Local and Global Features for Amazon Product Reviews: An Optimal ALBERT-Cascade CNN Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: To address these challenges, in this study, the researchers first performed a series of ablation experiments on 14 models derived from various variations in Deep Learning (DL) methods, including A Lite BERT (ALBERT) together with Convolutional Neural Networks (CNNs), Long Short-Term Memory (LSTM), Bidirectional LSTM (BiLSTM), Max Pooling layer, and attention mechanism. Subsequently, they proposed an ALBERT-cascaded CNN hybrid model as an effective method to overcome the related challenges by evaluating the performance results obtained from these models. |
Israa Mustafa Abbas; İsmail Atacak; Sinan Toklu; Necaattin Barışçı; İbrahim Alper Doğru; | Applied Sciences | 2025-12-19 |
| 891 | PhishGuard: AI-Driven Graph-Based Analysis for Smarter Email Security Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This project proposes a dual-model solution: a RoBERTa-based transformer is used to classify the email body content, while a Neo4j-powered graph model analyses sender-receiver domain relationships using graph metrics such as PageRank, ArticleRank, and Degree Centrality. |
Harchana Ramesh; Noris Ismail; ; Nor Azlina Abd Rahman; ; Aitizaz Ali; | STAP Journal of Security Risk Management | 2025-12-18 |
| 892 | Predictive Modeling of Maritime Radar Data Using Transformer Architecture Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This survey systematically reviews predictive modeling approaches relevant to maritime radar, with emphasis on transformer architectures for spatiotemporal sequence forecasting, where existing representative methods are analyzed according to data type, architecture, and prediction horizon. |
Bjorna Qesaraku; Jan Steckel; | arxiv-cs.CV | 2025-12-18 |
| 893 | Radiology Report Generation with Layer-Wise Anatomical Attention Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce a compact image-to-text architecture that generates the Findings section of chest X-ray reports from a single frontal image. |
EMMANUEL D. MUÑIZ-DE-LEÓN et. al. | arxiv-cs.CV | 2025-12-18 |
| 894 | LLMCache: Layer-Wise Caching Strategies for Accelerated Reuse in Transformer Inference Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present LLMCache, a novel layer-wise caching framework that accelerates transformer inference by reusing intermediate activations based on semantic similarity of input sequences. |
Harsh Vardhan Bansal; | arxiv-cs.CL | 2025-12-18 |
| 895 | Cutting-edge Technologies for Analyzing Student Feedback to Inform Institutional Decision-making in Higher Education Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a multitask learning framework to analyze student evaluations of teaching (SET) by extracting and classifying opinions on specific aspects of teaching performance. |
Sabur Butt; Sandra Dennis Núñez Daruich; Joanna Alvarado-Uribe; Hector G. Ceballos; | Foresight and STI Governance | 2025-12-17 |
| 896 | When A Nation Speaks: Machine Learning and NLP in People’s Sentiment Analysis During Bangladesh’s 2024 Mass Uprising Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Through Latent Dirichlet Allocation (LDA), we identified prevalent themes like political corruption and public protests, and analyzed how events such as internet blackouts shaped sentiment patterns. |
Md. Samiul Alim; Mahir Shahriar Tamim; Maisha Rahman; Tanvir Ahmed Khan; Md Mushfique Anwar; | arxiv-cs.CL | 2025-12-17 |
| 897 | Performance of GPT‐5 in The Interpretation of IBD Histopathology Reports Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We analyzed 100 real‐life histological reports from ileo‐colonoscopies, equally representing CD, UC, IBD‐U, and NIBDC, collected across five Italian healthcare centers, including both IBD‐specialized and non‐specialized hospitals. |
MARCELLO MAIDA et. al. | United European Gastroenterology Journal | 2025-12-17 |
| 898 | Adaptive Cache Pollution Control for Large Language Model Inference Workloads Using Temporal CNN-Based Prediction and Priority-Aware Replacement Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Large Language Models (LLMs), such as GPT and LLaMA, introduce unique memory access characteristics during inference due to frequent token sequence lookups and embedding vector retrievals. |
Songze Liu; Hongkun Du; Shaowen Wang; | arxiv-cs.AR | 2025-12-16 |
| 899 | Prompt Repetition Improves Non-Reasoning LLMs IF:3 Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: When not using reasoning, repeating the input prompt improves performance for popular models (Gemini, GPT, Claude, and Deepseek) without increasing the number of generated tokens … |
Yaniv Leviathan; Matan Kalman; Yossi Matias; | arxiv-cs.LG | 2025-12-16 |
| 900 | OUSAC: Optimized Guidance Scheduling with Adaptive Caching for DiT Acceleration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present OUSAC (Optimized gUidance Scheduling with Adaptive Caching), a framework that accelerates diffusion transformers (DiT) through systematic optimization. |
Ruitong Sun; Tianze Yang; Wei Niu; Jin Sun; | arxiv-cs.CV | 2025-12-16 |
| 901 | Inflation Attitudes of Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates the ability of Large Language Models (LLMs), specifically GPT-3.5-turbo (GPT), to form inflation perceptions and expectations based on macroeconomic price signals. |
Nikoleta Anesti; Edward Hill; Andreas Joseph; | arxiv-cs.CL | 2025-12-16 |
| 902 | Towards Nepali-language LLMs: Efficient GPT Training with A Nepali BPE Tokenizer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a GPT-2-based Nepali language model trained using several training strategies inspired by GPT-3, including optimized learning rate schedules, batch scaling, and architectural refinements. |
Adarsha Shrestha; Basanta Pokharel; Binit Shrestha; Smriti Adhikari; Dinesh Gothe; | arxiv-cs.CL | 2025-12-16 |
| 903 | Fake News Detection Using Albert-base-v2 Transformer and CNN-BiLSTM Architectures: A Comparative Analysis of Transformer-Based and Deep Learning Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
Chi Zhang; | Informatica | 2025-12-15 |
| 904 | Adapter‐Regularised Continual Learning for Dynamic Financial Sentiment Encoding in Multi‐Modal Market Fusion Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: ABSTRACT We propose an adapter‐regularised continual learning framework for dynamic financial sentiment encoding, addressing the dual challenge of retaining long‐term domain knowledge while adapting to transient market sentiment patterns. |
Zihe Song; Renke Huang; Aiqi Li; Aoran Shen; Heng Chen; | Expert Systems | 2025-12-15 |
| 905 | Detecting Emotion Drift in Mental Health Text Using Pre-Trained Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates emotion drift: the change in emotional state across a single text, within mental health-related messages. |
Shibani Sankpal; | arxiv-cs.CL | 2025-12-15 |
| 906 | Unveiling User Perceptions in The Generative AI Era: A Sentiment-Driven Evaluation of AI Educational Apps’ Role in Digital Transformation of E-Teaching Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study performs a sentiment-driven evaluation of user reviews from top AI ed-apps on the Google Play Store to assess efficacy, challenges, and pedagogical implications. |
Adeleh Mazaherian; Erfan Nourbakhsh; | arxiv-cs.CY | 2025-12-12 |
| 907 | Surveillance Video-Based Traffic Accident Detection Using Transformer Architecture Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Utilizing the curated dataset, we propose an accident detection model based on a transformer architecture using pre-extracted spatial video features. |
Tanu Singh; Pranamesh Chakraborty; Long T. Truong; | arxiv-cs.CV | 2025-12-12 |
| 908 | How Much Data in Low-resource Indian Languages Is Sufficient’ for Transfer Learning: A Comparative Study for POS Annotation Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study is conducted with Hindi as the high-resource language and the three related languages – Magahi, Bhojpuri and Braj – as extremely low-resource languages. |
Mohit Raj; Ritesh Kumar; | ACM Transactions on Asian and Low-Resource Language … | 2025-12-12 |
| 909 | Integrating Artificial Intelligence and Extended Reality for Enhanced Cultural Heritage Preservation: A Neurophysiological and Computational Approach Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: This study presents a framework integrating artificial intelligence, augmented reality, and neurophysiological assessment for preserving and digitizing Chinese Keju cultural … |
Xinmin Jin; Limin Zhang; Shuyi Zhang; Jian Teng; | Int. J. Gaming Comput. Mediat. Simulations | 2025-12-12 |
| 910 | An AI-Driven Product Recommendation Framework Integrating Collaborative Filtering and BERT-Based NLP Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Furthermore, current deep learning-based models like Bert4Rec or composite methods like J-NCFc continue to have high prediction errors and poor ranking accuracy. To fill these voids, this paper introduces a new CF+BERT model that incorporates collaborative filtering with contextualized review representations obtained from BERT. |
Ashrf Althbiti; | Journal of Multiscale Modelling | 2025-12-11 |
| 911 | LabelFusion: Learning to Fuse LLMs and Transformer Classifiers for Robust Text Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The package provides a simple high-level interface (AutoFusionClassifier) that trains the full pipeline end-to-end with minimal configuration, and a flexible API for advanced users. |
Michael Schlee; Christoph Weisser; Timo Kivimäki; Melchizedek Mashiku; Benjamin Saefken; | arxiv-cs.CL | 2025-12-11 |
| 912 | UrbanAI 2025 Challenge: Linear Vs Transformer Models for Long-Horizon Exogenous Temperature Forecasting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study long-horizon exogenous-only temperature forecasting – a challenging univariate setting where only the past values of the indoor temperature are used for prediction – using linear and Transformer-family models. |
Ruslan Gokhman; | arxiv-cs.LG | 2025-12-11 |
| 913 | LLM-Based Support for Diabetes Diagnosis: Opportunities, Scenarios, and Challenges with GPT-5 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study evaluates GPT-5, the latest generative pretrained transformer, using a simulation framework built entirely on synthetic cases aligned with ADA Standards of Care 2025 and inspired by public datasets including NHANES, Pima Indians, EyePACS, and MIMIC-IV. |
Gaurav Kumar Gupta; Nirajan Acharya; Pranal Pande; | International Journal of Modern Developments in Engineering … | 2025-12-11 |
| 914 | Graph-augmented Transformer Ensemble Framework for Robust and Scalable Fake News Detection in Social Media Ecosystems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present a new hybrid model named Graph-Augmented Transformer Ensemble (GETE) for efficient and scalable fake news detection. |
CHANCHAL KUMAR et. al. | Scientific Reports | 2025-12-11 |
| 915 | Semantic Similarity for Drug Slang Identification: A Comparative Analysis of Word2Vec and BERT Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: The rapid emergence of novel psychoactive substances and evolving slang presents ongoing challenges for effective drug surveillance and public health intervention. Traditional … |
Srikar Reddy Gadusu; H. Mcginty; | Proceedings of the 13th Knowledge Capture Conference 2025 | 2025-12-10 |
| 916 | AI‐Driven Intelligent Feedback System for Enhancing Self‐Assessment Accuracy in Higher Education Writing Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: With the rapid advancement of generative artificial intelligence, large language models (LLMs) have become increasingly integrated into education, particularly for automated … |
Shih-Yeh Chen; Wei-Cheng Chen; | Expert Systems | 2025-12-10 |
| 917 | Explainable Multilingual and Multimodal Fake-news Detection: Toward Robust and Trustworthy AI for Combating Misinformation IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study introduces two key innovations: (i) a new multilingual–multimodal dataset of 74,000 news articles in Hindi, Gujarati, Marathi, Telugu, and English with paired images, and (ii) Hybrid Explainable Multimodal Transformer Fake (HEMT-Fake) that integrates text, image, and relational signals with hierarchical explainability. |
ROHINI JADHAV et. al. | Frontiers in Artificial Intelligence | 2025-12-10 |
| 918 | A Comprehensive Review of Machine Learning and Deep Learning Approaches for Fake News Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This review paper presents a comprehensive survey of machine learning (ML), deep learning (DL), and transformer-based approaches designed for identifying misleading or deceptive content across digital platforms. |
Prof. Sarwesh Site; Shahbaz Akhtar; | International Journal of Scientific Research in Engineering … | 2025-12-10 |
| 919 | GThinker – A Chatbot Using AI Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Abstract—This paper presents the development of GThinker, an intelligent chatbot designed using Natural Language Processing (NLP) and advanced AI models to deliver human-like, context-aware, and emotionally adaptive interactions. |
R Harshith Raj; Dr. Sheethal Aji Mani; Aditya N; Darshan V; Madhu S; | International Journal of Scientific Research in Engineering … | 2025-12-09 |
| 920 | LLMs for Analog Circuit Design Continuum (ACDC) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we investigate the applicability and consistency of LLMs for analog circuit design — a task requiring domain-specific reasoning, adherence to physical constraints, and structured representations — focusing on AI-assisted design where humans remain in the loop. |
Yasaman Esfandiari; Jocelyn Rego; Austin Meyer; Jonathan Gallagher; Mia Levy; | arxiv-cs.LG | 2025-12-09 |
| 921 | Integrating Multimodal Clinical Data to Predict Intravenous (IV) Fluid Utilization: A Comparative Analysis of Natural Language Processing Techniques Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Methods We analyzed a large dataset from the National Hospital Ambulatory Medical Care Survey—Emergency Department (NHAMCS-ED, n = 13,115), comprising both structured patient demographics and clinical variables, alongside unstructured chief complaints. |
Hairong Wang; Haipeng Ling; Xingyu Zhang; | PeerJ Computer Science | 2025-12-09 |
| 922 | GNN-ATIVE: An AI-native, Graph-based Orchestrator for Next-Generation Wireless Networks Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Traditional rule-based or static management approaches struggle to cope with the dynamic, multi-layered nature of 5G/6G networks, creating a strong motivation for AI-native … |
VARUN GOWTHAM et. al. | GLOBECOM 2025 – 2025 IEEE Global Communications Conference | 2025-12-08 |
| 923 | Classifying Human Vs. AI Text with Machine Learning and Explainable Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a comprehensive framework for distinguishing between human-written and GPT-generated text using a combination of machine learning, sequential deep learning, and transformer-based models. |
ADVEN MASIH et. al. | Scientific Reports | 2025-12-08 |
| 924 | Aligning Pre-Trained LLMs for Enhanced UAV Power Consumption Forecasting Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Unmanned Aerial Vehicles (UAVs) are expanding beyond military use into sectors such as logistics, communication, and transportation. However, their dependence on high-power … |
Aroosa Hameed; Syed Muhammad Danish; Aris Leivadeas; | GLOBECOM 2025 – 2025 IEEE Global Communications Conference | 2025-12-08 |
| 925 | Mechanistic Interpretability of GPT-2: Lexical and Contextual Layers in Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present a mechanistic interpretability study of GPT-2 that causally examines how sentiment information is processed across its transformer layers. |
Amartya Hatua; | arxiv-cs.CL | 2025-12-07 |
| 926 | KV-CAR: KV Cache Compression Using Autoencoders and KV Reuse in Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The KV cache grows with sequence length and embedding dimension, often exceeding the memory footprint of the model itself and limiting achievable batch sizes and context windows. To address this challenge, we present KV CAR, a unified and architecture agnostic framework that significantly reduces KV cache storage while maintaining model fidelity. |
Sourjya Roy; Shrihari Sridharan; Surya Selvam; Anand Raghunathan; | arxiv-cs.LG | 2025-12-07 |
| 927 | Deep Reinforcement Learning for Phishing Detection with Transformer-Based Semantic Features Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study proposes a Quantile Regression Deep Q-Network (QR-DQN) approach that integrates RoBERTa semantic embeddings with handcrafted lexical features to enhance phishing detection while accounting for uncertainties. |
Aseer Al Faisal; | arxiv-cs.LG | 2025-12-07 |
| 928 | Chemistry Integrated Language Model Using Hierarchical Molecular Representation for Polymer Informatics Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce CI-LLM (Chemically Informed Language Model), a framework combining HAPPY (Hierarchically Abstracted rePeat unit of PolYmer), which encodes chemical substructures as tokens, with numerical descriptors within transformer architectures. |
Jihun Ahn; Gabriella Pasya Irianti; Vikram Thapar; Su-Mi Hur; | arxiv-cs.LG | 2025-12-06 |
| 929 | Transformer-Based Sentiment Analysis on Amazon Reviews Using Kaggle Dataset Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comprehensive study of sentiment classification on a large Amazon Reviews dataset from Kaggle, containing millions of reviews with star ratings, by fine-tuning state-of-the-art transformer models. |
Shwetaba B. Chauhan; Japan M. Mavani; | International Journal of Innovative Science and Research … | 2025-12-06 |
| 930 | Enhancing Dementia and Cognitive Decline Detection with Large Language Models and Speech Representation Learning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study describes our submission to the PROCESS Signal Processing Grand Challenge (ICASSP 2025), which tasked participants with predicting cognitive decline from speech samples. |
Karol Chlasta; Piotr Struzik; Grzegorz M. Wójcik; | Frontiers in Neuroinformatics | 2025-12-05 |
| 931 | AGF-HAM: Adaptive Gated Fusion Hierarchical Attention Model for Explainable Sentiment Analysis Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research paper presents a new hybrid model, HAM (Hybrid Attention-based Model), a Transformer-based contextual embedding model combined with deep sequential modeling and multi-layer explainability. |
Mahander Kumar; Lal Khan; Mohammad Zubair Khan; Amel Ali Alhussan; | Mathematics | 2025-12-05 |
| 932 | Leveraging Large Language Models to Detect Academic Anxiety in Indonesian English for Specific Purposes Students Through Reflective Writing Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study investigates the capacity of Large Language Models to identify academic anxiety in reflective writing produced by English for Specific Purposes students from Indonesia. |
Khoirul Anwar; Bambang Harmanto; | International Journal of Learning, Teaching and Educational … | 2025-12-05 |
| 933 | Robust Sentiment Analysis Through Bayesian Dropout-Enhanced RoBERTa-LSTM Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces RoBERTa-LSTM-Drop, a hybrid architecture that combines RoBERTa embeddings with bidirectional LSTM layers and integrates Bayesian Dropout to capture uncertainty through Monte Carlo sampling while acting as an effective regulariser. |
Soufien Jaffali; | BRAIN. Broad Research in Artificial Intelligence and … | 2025-12-05 |
| 934 | BERTO: An Adaptive BERT-based Network Time Series Predictor with Operator Preferences in Natural Language Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce BERTO, a BERT-based framework for traffic prediction and energy optimization in cellular networks. |
Nitin Priyadarshini Shankar; Vaibhav Singh; Sheetal Kalyani; Christian Maciocco; | arxiv-cs.LG | 2025-12-05 |
| 935 | Automated Identification of Incidentalomas Requiring Follow-Up: A Multi-Anatomy Evaluation of LLM-Based and Supervised Approaches Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduced a novel inference strategy using lesion-tagged inputs and anatomy-aware prompting to ground model reasoning. |
NAMU PARK et. al. | arxiv-cs.CL | 2025-12-05 |
| 936 | Fusing Semantic and Structural Features for Code Error Detection Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Nevertheless, their efficacy could be further improved by addressing the inherent weakness in handling structural code dependencies. In response to this, we introduce a novel model that integrates the semantic comprehension power of RoBERTa with the structural learning strength of Graph Neural Networks. |
Yiwen Zhang; Wei Liu; Fazhong Jiang; Jiquan Ma; Jingtai Cao; | Entropy | 2025-12-04 |
| 937 | KV Cache Recycling to Expand Usable Context Capacity in Low Parameter LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Whether attention key value (KV) states computed for one prompt for a small LLM can be reused to accelerate inference on a new similar prompt, giving an increase to the space to its context memory using an approach called token recycling. |
Prashant Pandey; | arxiv-cs.LG | 2025-12-04 |
| 938 | Towards Improved Fake News Detection Using A Hybrid RoBERTa and Metadata Enhanced XGBoost Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
ARMUGHAN ALI et. al. | Scientific Reports | 2025-12-04 |
| 939 | Multiclass Hate Speech Detection: Evaluating 303 Model Configurations Across Traditional Machine Learning, Deep Learning, and Transformer Approaches Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Multiclass hate speech detection across demographic categories remains challenging due to implicit targeting strategies and linguistic variability in social media content. This … |
Mahmoud Abusaqer; K. Hasan; Jamil Saquer; | 2025 International Conference on Machine Learning and … | 2025-12-03 |
| 940 | GRASP: GRouped Activation Shared Parameterization for Parameter-Efficient Fine-Tuning and Robust Inference of Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce GRASP – GRouped Activation Shared Parameterization – a lightweight PEFT framework that partitions the D-dimensional token representations of selected layers into K << D groups and learns a shared scaling and shifting vector for each group. |
Malyaban Bal; Abhronil Sengupta; | arxiv-cs.LG | 2025-12-03 |
| 941 | Hamisfera: Sistem Rekomendasi Progresi Chord Berbasis Sentimen Lirik Melalui Studi Komparatif Arsitektur Transformer Dan Mixture of Experts Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Abstrak – Proses penciptaan lagu sering kali memerlukan penyelarasan antara nuansa emosional lirik dengan harmoni musik yang tepat. Namun, penerjemahan sentimen lirik menjadi … |
Fara Daud Ibra; Muhammad Fachrie; | Jurnal Informatika dan Multimedia | 2025-12-03 |
| 942 | Multi-Modal Opinion Integration for Financial Sentiment Analysis Using Cross-Modal Attention Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes an end-to-end deep learning framework that integrates two distinct modalities of financial opinions: recency modality (timely opinions) and popularity modality (trending opinions), through a novel cross-modal attention mechanism specifically designed for financial sentiment analysis. |
Yujing Liu; Chen Yang; | arxiv-cs.LG | 2025-12-03 |
| 943 | Dual LoRA: Enhancing LoRA with Magnitude and Direction Updates Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we propose a novel method called Dual LoRA to improve the performance by incorporating an inductive bias into the original LoRA. |
YIXING XU et. al. | arxiv-cs.CL | 2025-12-02 |
| 944 | SPECTRE: Computational Methods in Natural Language Processing for Automated Hardware Trojan Insertion Using Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Experts Stealthy Processor Exploitation and Concealment Through Reconfigurable Elements (SPECTRE) is the new framework proposed in this paper to use the computational methods of Natural Language Processing (NLP) to automate the addition of Hardware Trojans (HTs) to the complex hardware design. |
Moneer Alshaikh; Rashid Amin; Sajid Mehmood; Faisal S. Alsubaei; | Contemporary Mathematics | 2025-12-02 |
| 945 | Artificial Intelligence for Employee Engagement and Well-Being: A Review of Digital Tools, Psychometric Measures and Workforce Sentiment Datasets in Modern HR Systems Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper assesses empirical and conceptual evidence from 2015–2025 across three interconnected domains of modern HR analytics: AI-driven digital engagement and well-being tools, psychometric measures embedded in AI systems, and real-world workforce sentiment datasets used for model development and validation. |
Francis Dumbili; Onyinye Uzoka; Seun Adeniran; Mercy Afreh; | World Journal of Advanced Research and Reviews | 2025-12-02 |
| 946 | Bangla Hate Speech Classification with Fine-tuned Transformer Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we study Subtask 1A and Subtask 1B of the BLP 2025 Shared Task on hate speech detection. |
Yalda Keivan Jafari; Krishno Dey; | arxiv-cs.CL | 2025-12-02 |
| 947 | Enhancing Recommendation Systems with Autoencoder-SVD and Transformer-Based Summarization: A Sentiment-Aware Approach Using GPT-2 and VADER Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper proposes a novel approach to improve recommendation accuracy by integrating Autoencoder-SVD with language model-based summarization. |
Muhi Saadi Rahdi; Farsad Zamani Boroujeni; Aladdin Abdulhassan; Mehdi Akbari Kopayei; Keyvan Mohebbi; | Qubahan Academic Journal | 2025-12-02 |
| 948 | What Signals Really Matter for Misinformation Tasks? Evaluating Fake-News Detection and Virality Prediction Under Real-World Constraints Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We present an evaluation-driven study of two practical tasks regarding online misinformation: (i) fake-news detection and (ii) virality prediction in the context of operational settings, with the necessity for rapid reaction. |
Francesco Paolo Savatteri; Chahan Vidal-Gorène; Florian Cafiero; | arxiv-cs.CL | 2025-12-02 |
| 949 | Idea-Gated Transformers: Enforcing Semantic Coherence Via Differentiable Vocabulary Pruning Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we introduce the Idea-Gated Transformer, a novel architecture that separates semantic planning from syntactic generation. |
Darshan Fofadiya; | arxiv-cs.CL | 2025-12-02 |
| 950 | Enhancing Task Prioritization in Software Development Issues Tracking System Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates the potential of automated issue priority classification using state‐of‐the‐art Transformer models to alleviate this burden. We evaluate the performance of models like BERT, DeBERTa, and ModernBERT, comparing them against general large language models (LLMs) such as GPT‐3.5, Qwen2.5‐3B and Llama‐3.2‐3B, using curated datasets derived from public Jira and GitHub repositories. |
Karthik Shivashankar; Kristian Marison Haugerud; Antonio Martini; | Journal of Software: Evolution and Process | 2025-12-01 |
| 951 | Application of A Dual-Branch Recursive Time-Series Joint Model Based on A Multiscale Transformer and Multimodal Fusion to Fault Diagnosis of Fixed-Wing UAV Actuators Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Fixed-wing UAV actuators are prone to compound failures under complex multimodal dynamic environments, posing challenges to fault diagnosis accuracy and robustness. To address … |
Wenqi Zhang; Zhenbao Liu; Zhen Jia; Sheng-Long Wang; Xiao Wang; | IEEE Transactions on Aerospace and Electronic Systems | 2025-12-01 |
| 952 | Handwritten Text Recognition for Low Resource Languages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a ViT-Transformer Decoder-LM architecture for handwritten text recognition, where a Vision Transformer (ViT) extracts visual features, a Transformer decoder generates text sequences, and a pre-trained language model (LM) refines the output to improve accuracy, fluency, and coherence. |
Sayantan Dey; Alireza Alaei; Partha Pratim Roy; | arxiv-cs.CV | 2025-12-01 |
| 953 | A Novel Multimodal Deep Learning Framework for Conversational AI: Integrating Vision, Text, and Speech With Knowledge‐Augmented Attention Mechanisms Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Existing conversational AI models lack effective multimodal fusion, leading to poor contextual understanding and factual inconsistencies. This research proposes a novel multimodal … |
T. V. PAI et. al. | Computational Intelligence | 2025-12-01 |
| 954 | Testing Transformer Learnability on The Arithmetic Sequence of Rooted Trees Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We study whether a Large Language Model can learn the deterministic sequence of trees generated by the iterated prime factorization of the natural numbers. |
Alessandro Breccia; Federica Gerace; Marco Lippi; Gabriele Sicuro; Pierluigi Contucci; | arxiv-cs.AI | 2025-12-01 |
| 955 | Splitwise: Collaborative Edge–Cloud Inference for LLMs Via Lyapunov-Assisted DRL Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Deploying large language models (LLMs) on edge devices is challenging due to their limited memory and power resources. Cloud-only inference reduces device burden but introduces … |
ABOLFAZL YOUNESI et. al. | Proceedings of the 18th IEEE/ACM International Conference … | 2025-12-01 |
| 956 | MicroProbe: Efficient Reliability Assessment for Foundation Models with Minimal Data Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We introduce microprobe, a novel approach that achieves comprehensive reliability assessment using only 100 strategically selected probe examples. |
Aayam Bansal; Ishaan Gangwani; | arxiv-cs.AI | 2025-11-30 |
| 957 | Generalized Graph Transformer Variational Autoencoder Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose the Generalized Graph Transformer Variational Autoencoder (GGT-VAE). |
Siddhant Karki; | arxiv-cs.LG | 2025-11-29 |
| 958 | Financial Text Classification Based On RLoRA Finetuning On Qwen3-8B Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we assess the performance of the large language model Qwen3-8B on both tasks. |
Zhiming Lian; | arxiv-cs.LG | 2025-11-29 |
| 959 | MCP Vs RAG Vs NLWeb Vs HTML: A Comparison of The Effectiveness and Efficiency of Different Agent Interfaces to The Web (Technical Report) Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, no prior work has compared these four architectures within a single controlled environment using identical tasks. To address this gap, we introduce a testbed consisting of four simulated e-shops, each offering its products via HTML, MCP, and NLWeb interfaces. |
Aaron Steiner; Ralph Peeters; Christian Bizer; | arxiv-cs.CL | 2025-11-28 |
| 960 | Transformer-Driven Triple Fusion Framework for Enhanced Multimodal Author Intent Classification in Low-Resource Bangla Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Recognizing limitations in previous unimodal approaches, we systematically benchmark transformer-based language models (mBERT, DistilBERT, XLM-RoBERTa) and vision architectures (ViT, Swin, SwiftFormer, ResNet, DenseNet, MobileNet), utilizing the Uddessho dataset of 3,048 posts spanning six practical intent categories. We introduce a novel intermediate fusion strategy that significantly outperforms early and late fusion on this task. |
Ariful Islam; Tanvir Mahmud; Md Rifat Hossen; | arxiv-cs.LG | 2025-11-28 |
| 961 | Pooling Attention: Evaluating Pretrained Transformer Embeddings for Deception Classification Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper investigates fake news detection as a downstream evaluation of Transformer representations, benchmarking encoder-only and decoder-only pre-trained models (BERT, GPT-2, Transformer-XL) as frozen embedders paired with lightweight classifiers. |
Sumit Mamtani; Abhijeet Bhure; | arxiv-cs.CL | 2025-11-28 |
| 962 | Challenges of Heterogeneity in Big Data: A Comparative Study of Classification in Large-Scale Structured and Unstructured Domains Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work provides a unified framework for algorithm selection based on data nature and infrastructure constraints. |
González Trigueros Jesús Eduardo; Alonso Sánchez Alejandro; Muñoz Rivera Emilio; Peñarán Prieto Mariana Jaqueline; Mendoza González Camila Natalia; | arxiv-cs.LG | 2025-11-28 |
| 963 | Tourism Question Answer System in Indian Language Using Domain-Adapted Foundation Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, a dataset comprising 7,715 Hindi QA pairs pertaining to Varanasi tourism was constructed and subsequently augmented with 27,455 pairs generated via Llama zero-shot prompting. |
Praveen Gatla; Nikita Kanwar; Gouri Sahoo; Rajesh Kumar Mundotiya; | arxiv-cs.CL | 2025-11-28 |
| 964 | Tree Matching Networks for Natural Language Inference: Parameter-Efficient Semantic Understanding Via Dependency Parse Trees Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Explicit structural representations significantly outperform sequence-based models at comparable scales, but current aggregation methods limit scalability. We propose multi-headed attention aggregation to address this limitation. |
Jason Lunder; | arxiv-cs.CL | 2025-11-28 |
| 965 | Exploring The Predictive Performance of Deep Learning for Fracturing Fluid Flowback and Shale Gas Production Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Here, we developed CNN-Transformer, a production and fluid flowback predicted system using deep learning, by integrating a convolutional neural network (CNN) and a Transformer network. |
SHASHA SUN et. al. | Scientific Reports | 2025-11-28 |
| 966 | Transformer and Pre-Transformer Model-Based Sentiment Prediction with Various Embeddings: A Case Study on Amazon Reviews Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study contributes to both sentiment analysis and sustainable AI by offering a scalable, entropy-aware evaluation framework that supports informed, context-sensitive model selection for practical applications. |
Ismail Duru; Ayşe Saliha Sunar; | Entropy | 2025-11-27 |
| 967 | Sentiment Analysis Of Shopee Product Reviews Using Distilbert Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study examines the use of DistilBERT, a lightweight transformer-based deep learning model, for sentiment classification on Shopee product reviews. |
Zahri Aksa Dautd; Aviv Yuniar Rahman; | arxiv-cs.CL | 2025-11-27 |
| 968 | BERT-Based Approaches for Web Service Selection and Recommendation: A Systematic Review with A Focus on QoS Prediction Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: Effective web service selection and recommendation are critical for ensuring high-quality performance in distributed and service-oriented systems. Recent research has increasingly … |
Vijayalakshmi Mahanra Rao; R. Ramasamy; M. Sayeed; | Future Internet | 2025-11-27 |
| 969 | Efficient Adaptation of Large Language Models for Sentiment Analysis: A Fine-Tuning Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a systematic comparative analysis of sentiment classification on financial news headlines using two transformer architectures, Mistral-7B and GPT-2, fine-tuned with advanced adaptation techniques—Quantized Low-Rank Adaptation (QLoRA) and Low-Rank Adaptation (LoRA). |
Seda Bayat Toksöz; Gültekin Işık; | Iğdır Üniversitesi Fen Bilimleri Enstitüsü Dergisi | 2025-11-27 |
| 970 | Contextual Gating Within The Transformer Stack: Synergistic Feature Modulation for Enhanced Lyrical Classification and Calibration Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: I propose the SFL Transformer, a novel deep learning model that utilizes a Contextual Gating mechanism (an Intermediate SFL) to modulate the sequence of hidden states within the BERT encoder stack, rather than fusing features at the final output layer. |
M. A. Gameiro; | arxiv-cs.LG | 2025-11-27 |
| 971 | LC4-DViT: Land-cover Creation for Land-cover Classification with Deformable Vision Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose LC4-DViT (Land-cover Creation for Land-cover Classification with Deformable Vision Transformer), a framework that combines generative data creation with a deformation-aware Vision Transformer. |
KAI WANG et. al. | arxiv-cs.CV | 2025-11-27 |
| 972 | A Theoretically Grounded Hybrid Ensemble for Reliable Detection of LLM-Generated Text Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose a theoretically grounded hybrid ensemble that systematically fuses three complementary detection paradigms: (i) a RoBERTa-based transformer classifier for deep semantic feature extraction, (ii) a GPT-2-based probabilistic detector using perturbation-induced likelihood curvature, and (iii) a statistical linguistic feature analyzer capturing stylometric patterns. |
Sepyan Purnama Kristanto; Lutfi Hakim; | arxiv-cs.CL | 2025-11-27 |
| 973 | Using Text-Based Life Trajectories from Swedish Register Data to Predict Residential Mobility with Pretrained Transformers Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We transform large-scale Swedish register data into textual life trajectories to address two long-standing challenges in data analysis: high cardinality of categorical variables and inconsistencies in coding schemes over time. Leveraging this uniquely comprehensive population register, we convert register data from 6.9 million individuals (2001-2013) into semantically rich texts and predict individuals’ residential mobility in later years (2013-2017). |
Philipp Stark; Alexandros Sopasakis; Ola Hall; Markus Grillitsch; | arxiv-cs.LG | 2025-11-26 |
| 974 | Evaluating The Effectiveness of AI-Assisted Emotional Metadata in Enhancing The Discoverability of Literary Texts in Digital Libraries Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The research shows that discovery in digital libraries, user interaction, and search can be boosted by adding emotional metadata in digital libraries. |
Bushara Iqbal; | Social Sciences & Humanity Research Review | 2025-11-26 |
| 975 | Visualizing LLM Latent Space Geometry Through Dimensionality Reduction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we extract, process, and visualize latent state geometries in Transformer-based language models through dimensionality reduction. |
Alex Ning; Vainateya Rangaraju; | arxiv-cs.LG | 2025-11-26 |
| 976 | BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this study, we present BanglaMM-Disaster, an end-to-end deep learning-based multimodal framework for disaster classification in Bangla, using both textual and visual data from social media. |
Ariful Islam; Md Rifat Hossen; Md. Mahmudul Arif; Abdullah Al Noman; Md Arifur Rahman; | arxiv-cs.LG | 2025-11-26 |
| 977 | DPATransLLM: Detection of Pronominal Anaphora in Turkish Sentences Using Transformer-Based, Large Language Models and Hybrid Ensemble Approach Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, fine-tuning was performed on Transformer-based language models pre-trained on Turkish data, such as BERT and RoBERTa. |
Engin Demir; Metin Bilgin; | Applied Sciences | 2025-11-25 |
| 978 | ChatGpt Content Detection: A New Approach Using Xlm-roberta Alignment Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The challenge of separating AI-generated text from human-authored content is becoming more urgent as generative AI technologies like ChatGPT become more widely available. In this work, we address this issue by looking at both the detection of content that has been entirely generated by AI and the identification of human text that has been reworded by AI. |
MD TASNIN TANVIR et. al. | arxiv-cs.LG | 2025-11-25 |
| 979 | HHFT: Hierarchical Heterogeneous Feature Transformer for Recommendation Systems IF:3 Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: We propose HHFT (Hierarchical Heterogeneous Feature Transformer), a Transformer-based architecture tailored for industrial CTR prediction. |
Liren Yu; Wenming Zhang; Silu Zhou; Zhixuan Zhang; Dan Ou; | arxiv-cs.IR | 2025-11-25 |
| 980 | Building A Foundation Model for Trajectory from Scratch Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Through a concise, step-by-step, code-driven process, we demonstrate adapting GPT-2 for spatiotemporal data. |
Gaspard Merten; Mahmoud Sakr; Gilles Dejaegere; | arxiv-cs.AI | 2025-11-25 |
| 981 | Directional Optimization Asymmetry in Transformers: A Synthetic Stress Test Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Using random string mappings with tunable branching factor K, we construct forward tasks with zero conditional entropy and inverse tasks with analytically determined entropy floors. |
Mihir Sahasrabudhe; | arxiv-cs.CL | 2025-11-25 |
| 982 | On The Role of Hidden States of Modern Hopfield Network in Transformer Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: It has been pointed out that the state update rule of the modern Hopfield network (MHN) in the adiabatic approximation is in agreement with the self-attention layer of Transformer. In this paper, we go beyond this approximation and investigate the relationship between MHN and self-attention. |
Tsubasa Masumura; Masato Taki; | arxiv-cs.LG | 2025-11-24 |
| 983 | Dissecting The Ledger: Locating and Suppressing Liar Circuits in Financial Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we propose a mechanistic approach to intrinsic hallucination detection. |
Soham Mirajkar; | arxiv-cs.CL | 2025-11-24 |
| 984 | Assessing The Efficacy of Ortho GPT: A Comparative Study with Medical Students and General LLMs on Orthopedic Examination Questions Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The domain-specific approach enables performance matching or exceeding top general LLMs in orthopedics, emphasizing the importance of domain specialization for reliable, curriculum-aligned support in medical education. |
PHILIPPE FABIAN POHLMANN et. al. | Bioengineering | 2025-11-24 |
| 985 | Intelligent Sustainability: Evaluating Transformers for Cryptocurrency Environmental Claims Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Employing design science research (DSR) methodology, we develop and empirically evaluate a novel framework comparing five state-of-the-art transformer models across multiple performance dimensions. |
Parisa Bouzari; Maria Fekete-Farkas; Zsigmond Gábor Szalay; | Information | 2025-11-24 |
| 986 | PeriodNet: Boosting The Potential of Attention Mechanism for Time Series Forecasting Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this paper, we present PeriodNet with a brand new structure to forecast univariate and multivariate time series. |
BOWEN ZHAO et. al. | arxiv-cs.LG | 2025-11-23 |
| 987 | From Reviewers’ Lens: Understanding Bug Bounty Report Invalid Reasons with LLMs Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: In this work, we conduct an empirical study with the purpose of helping bug hunters understand the validity of reports. |
Jiangrui Zheng; Yingming Zhou; Ali Abdullah Ahmad; Hanqing Yao; Xueqing Liu; | arxiv-cs.SE | 2025-11-23 |
| 988 | AGI Team at SHROOM-CAP: Data-Centric Approach to Multilingual Hallucination Detection Using XLM-RoBERTa Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper describes our submission to the SHROOM-CAP 2025 shared task on scientific hallucination detection across 9 languages. |
Harsh Rathva; Pruthwik Mishra; Shrikant Malviya; | arxiv-cs.CL | 2025-11-23 |
| 989 | CATEGORY-BASED SENTIMENT ANALYSIS OF SINDHI NEWS HEADLINES USING MACHINE LEARNING DEEP LEARNING AND TRANSFORMER MODELS Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This work provides a foundational resource for NLP researchers seeking to advance computational methods for Sindhi and similar underrepresented languages. |
Dr. RAKESH; | International Journal of Data Science and IoT Management … | 2025-11-22 |
| 990 | NX-CGRA: A Programmable Hardware Accelerator for Core Transformer Algorithms on Edge Devices Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper introduces NX-CGRA, a programmable hardware accelerator designed to support a range of transformer inference algorithms, including both linear and non-linear functions. |
Rohit Prasad; | arxiv-cs.AR | 2025-11-21 |
| 991 | Mining Emotions ‘A Comprehensive Study on Sentiment Analysis of Social Media’ Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This study presents a comprehensive study of how SA techniques are applied to social media data to “mine emotions” effectively. |
Dr. Rupali Kalekar; Vyas More; Omkar Nalawade; | International Journal of Scientific Research in Engineering … | 2025-11-21 |
| 992 | Pier: Efficient Large Language Model Pretraining with Relaxed Global Communication Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Global communication, such as all-reduce and allgather, is the prominent performance bottleneck in large language model (LLM) pretraining. To address this issue, we present Pier, an efficient and scalable optimizer with relaxed global communication. |
Shuyuan Fan; Zhao Zhang; | arxiv-cs.DC | 2025-11-21 |
| 993 | A Scientometric Survey of BERT and Transformer-based Research: An Analysis of 200 Highly-cited Scopus Publications Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study discloses a quick evolution from foundational architectural innovations and pre-training paradigms to widespread domain adaptation, rigorous model optimization for efficiency, and critical examination of model capabilities and societal effects. |
Ayman Mahgoub; | Scientometrica | 2025-11-21 |
| 994 | Enhancing Quranic Learning: A Multimodal Deep Learning Approach for Arabic Phoneme Recognition Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: Accurate pronunciation detection remains a key challenge in Arabic, particularly in the context of Quranic recitation, where subtle phonetic differences can alter meaning. Addressing this challenge, the present study proposes a transformer-based multimodal framework for Arabic phoneme mispronunciation detection that combines acoustic and textual representations to achieve higher precision and robustness. |
Ayhan Kucukmanisa; Derya Gelmez; Sukru Selim Calik; Zeynep Hilal Kilimci; | arxiv-cs.SD | 2025-11-21 |
| 995 | A Cloud-Based Cross-Modal Transformer for Emotion Recognition and Adaptive Human-Computer Interaction Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: However, existing systems often rely on single-modality analysis such as facial expressions, speech tone, or textual sentiment, resulting in limited robustness and poor generalization in real-world environments. To address these challenges, this study proposes a Cloud-Based Cross-Modal Transformer (CMT) framework for multimodal emotion recognition and adaptive human-computer interaction. |
Ziwen Zhong; Zhitao Shu; Yue Zhao; | arxiv-cs.CV | 2025-11-21 |
| 996 | Deep Learning Approaches for Multi-Class Classification of Phishing Text Messages Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This research proposes a chain transformer model that integrates GPT-2 for synthetic data generation and BERT for embeddings to detect Smishing within a multiclass dataset, including minority smishing variants. |
Miriam L. Munoz; Muhammad F. Islam; | Journal of Cybersecurity and Privacy | 2025-11-21 |
| 997 | Swin‐Decision Transformer: A Transformer‐Based Hybrid Protocol for Adaptive Clustering and Energy‐Efficient Routing in Large‐Scale WSNs Summary Related Papers Related Patents Related Grants Related Venues Related Experts View Save Abstract: In this era, large‐scale Wireless Sensor Networks (WSNs) provide high Quality of Service (QoS) with energy awareness and scalability. Specifically, existing clustering protocols … |
Basavaraj S. Mathapati; Nagaratna P. Hegde; S. P. Paramesh; Padmavathi Vurubindi; Subhra Chakraborty; | Internet Technology Letters | 2025-11-21 |
| 998 | Leveraging Transformer-based Models for Classification and Large Language Models for Information Extraction from Medical Claims Documents Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: The study explores the use of advanced AI models for document classification and information extraction in Indonesian medical claims documents. |
Dian Prambini; Kusworo Adi; Komang Budi Aryasa; | SISFORMA | 2025-11-21 |
| 999 | Evaluating Adversarial Vulnerabilities in Modern Large Language Models Related Papers Related Patents Related Grants Related Venues Related Experts View Save Highlight: This paper presents a comparative analysis of the susceptibility to jailbreak attacks for two leading publicly available LLMs, Google’s Gemini 2.5 Flash and OpenAI’s GPT-4 (specifically the GPT-4o mini model accessible in the free tier). |
Tom Perel; | arxiv-cs.CR | 2025-11-20 |
| 1000 | Multilingual Sentiment Analysis in E-commerce Customer Reviews Using GPT and Deep Learning-based Weighted-ensemble Model Related Papers Related Patents Related Grants Related Venues Related Experts View Save |
NUZHAT NOOR ISLAM PROVA et. al. | International Journal of Cognitive Computing in Engineering | 2025-11-20 |