e-ISSN: Pending

Browse the failure-mode index

744 real negative results, null findings, and replication failures in Computer Science. Search the index →

WASTE indexes published research — it does not host or republish full papers. Each entry is a metadata record compiled from open scholarly databases; the abstract is shown in full only where the paper is openly licensed, otherwise a short excerpt under fair use. Classifications are automated and approximate.

Negative / Null Result ReportOpen accessComputer Science

The added value for MRI radiomics and deep-learning for glioblastoma prognostication compared to clinical and molecular information

D. Abler, O. Pusterla, A. Joye-Kühnis et al. · 2025 · arXiv

Background: Radiomics shows promise in characterizing glioblastoma, but its added value over clinical and molecular predictors has yet to be proven. This study assessed the added value of conventional radiomics (CR) and deep learning (DL) MRI radiomics for glioblastoma prognosis ( 6 months survival) on a large multi-center dataset. Methods: After patient selection, our curated dataset gathers 1152 glioblastoma (WHO 2016) patients from five Swiss centers and one public source. It included clinical (age, gender), molecular (MGMT, IDH), and baseline MRI data (T1, T1 contrast, FLAIR, T2)

View details →
Negative / Null Result ReportOpen accessComputer Science

Shaping Ability of Different Single Niti Files: A Cbct Assessment

B. Selivany, A. Yousif · 2020 · The Journal Of Duhok University

Aim : The study aimed to compare Shaping ability of WaveOne Gold, Reciproc Blue and 2Shape NiTi systems having different design and metallurgic properties. Method : Forty five extracted human single root with 10 mm root length were…

View details →
Negative / Null Result ReportOpen accessComputer Science

Utilizing the Blackboard Learning Management System at the University of Ha’il from the Perspective of Faculty Members using Internet Of Things (IOT)

A. Ali, Khaled Mohammad Abu Sheirah · 2021 · International Journal of Education and Information Technologies

The study aims to investigate the perceptions of faculty members in the preparatory year at the University of Ha’il concerning the use of the Blackboard learning management system, and to identify the impact of the study variables (gender,…

View details →
Negative / Null Result ReportComputer Science

Can Modern NLP Systems Reliably Annotate Chest Radiography Exams? A Pre-Purchase Evaluation and Comparative Study of Solutions from AWS, Google, Azure, John Snow Labs, and Open-Source Models on an Independent Pediatric Dataset

Shruti Hegde, Mabon Ninan, Jonathan R. Dillman et al. · 2025 · arXiv.org

General-purpose clinical natural language processing (NLP) tools are increasingly used for the automatic labeling of clinical reports. However, independent evaluations for specific tasks, such as pediatric chest radiograph (CXR) report…

View details →
Negative / Null Result ReportOpen accessComputer Science

Evaluating Large Language Models for Time Series Anomaly Detection in Aerospace Software

Yang Liu, Yixing Luo, Xiaofeng Li et al. · 2026 · arXiv

Time series anomaly detection (TSAD) is essential for ensuring the safety and reliability of aerospace software systems. Although large language models (LLMs) provide a promising training-free alternative to unsupervised approaches, their effectiveness in aerospace settings remains under-examined because of complex telemetry, misaligned evaluation metrics, and the absence of domain knowledge. To address this gap, we introduce ATSADBench, the first benchmark for aerospace TSAD. ATSADBench comprises nine tasks that combine three pattern-wise anomaly types, univariate and multivariate signals, an

View details →
Negative / Null Result ReportOpen accessComputer Science

Domain Decomposition Based High Performance Parallel Computing

Mandhapati P. Raju, Siddhartha Khaitan · 2009 · arXiv

The study deals with the parallelization of finite element based Navier-Stokes codes using domain decomposition and state-ofart sparse direct solvers. There has been significant improvement in the performance of sparse direct solvers. Parallel sparse direct solvers are not found to exhibit good scalability. Hence, the parallelization of sparse direct solvers is done using domain decomposition techniques. A highly efficient sparse direct solver PARDISO is used in this study. The scalability of both Newton and modified Newton algorithms are tested.

View details →
Negative / Null Result ReportOpen accessComputer Science

Supersymmetry versus precision experiments revisited

Gi-Chol Cho, Kaoru Hagiwara · 1999 · arXiv

We study constraints on the supersymmetric standard model from the updated electroweak precision measurements --- the Z-pole experiments and the $W$-boson mass measurements. The supersymmetric-particle contributions to the universal gauge-boson-propagator corrections are parametrized by the three oblique parameters Sz, Tz and mw. The oblique corrections, the Zqq and Zll vertex corrections, and the vertex and box corrections to the μ-decay width are separately studied in detail. We first study individual contribution from the four sectors of the model, the squarks, the sleptons, the supersymmet

View details →
Negative / Null Result ReportOpen accessComputer Science

Probing neutron skin and symmetry energy with relativistic isobar collisions

Hao-jie Xu · 2023 · arXiv

In these proceedings, we present the three proposed observables to probe the neutron skin and symmetry energy with relativistic isobar collisions, namely, the isobar ratios of the produced hadron multiplicities ($N_{\rm ch}$), the mean transverse momenta ($\langle p_{\perp} \rangle$), and the net charge multiplicities ($ΔQ$). Our findings suggest potentially significant improvement to neutron skin and symmetry energy determination over traditional low energy methods.

View details →
Failed Experiment ReportOpen accessComputer Science

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

Sofia Morgado, Filipa Valdeira, Niklas Sander et al. · 2026 · arXiv

Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured natural language, limiting its use in research. We investigate the use of Large Language Models (LLMs) to automatically extract these values, along with their anatomical locations, from Portuguese CAG reports. To our knowledge, this study is the first addressing physiology indexes extraction from a large (1342 reports) corpus of CAG reports, and one of the few focusing on CAG or Portuguese clinical text. We explore local privacy-preserving gen

View details →
Negative / Null Result ReportOpen accessComputer Science

Positive pion absorption on 3He using modern trinucleon wave functions

S. Schneider, J. Haidenbauer, C. Hanhart et al. · 2002 · arXiv

We study pion absorption on 3He employing trinucleon wave functions calculated from modern realistic NN interactions (Paris, CD Bonn). Even though the use of the new wave functions leads to a significant improvement over older calculations with regard to both cross section and polarization data, there are hints that polarization data with quasifree kinematics cannot be described by just two-nucleon absorption mechanisms.

View details →
Negative / Null Result ReportOpen accessComputer Science

Hydrodynamic behaviour of Lattice Boltzmann and Lattice BGK models

O. Behrend, R. Harris, P. Warren · 1993 · arXiv

We present a numerical analysis of the validity of classical and generalized hydrodynamics for Lattice Boltzmann Equation (LBE) and Lattice BGK methods in two and three dimensions, as a function of the collision parameters of these models. Our analysis is based on the wave-number dependence of the evolution operator. Good ranges of validity are found for BGK models as long as the relaxation time is chosen smaller than or equal to unity. The additional freedom in the choice of collision parameters for LBE models does not seem to give significant improvement.

View details →
Negative / Null Result ReportOpen accessComputer Science

Magnitude Matters: a Superior Class of Similarity Metrics for Holistic Semantic Understanding

V. S. Raghu Parupudi · 2025 · arXiv

Vector comparison in high dimensions is a fundamental task in NLP, yet it is dominated by two baselines: the raw dot product, which is unbounded and sensitive to vector norms, and the cosine similarity, which discards magnitude information entirely. This paper challenges both standards by proposing and rigorously evaluating a new class of parameter-free, magnitude-aware similarity metrics. I introduce two such functions, Overlap Similarity (OS) and Hyperbolic Tangent Similarity (HTS), designed to integrate vector magnitude and alignment in a more principled manner. To ensure that my findings a

View details →
Failed Experiment ReportOpen accessComputer Science

Fragments in Gaussian Wave-Packet Dynamics with and without correlations

D. Kiderlen, P. Danielewicz · 1996 · arXiv

Generalization of Gaussian trial wave functions in quantum molecular dynamics models is introduced, which allows for long-range correlations characteristic for composite nuclear fragments. We demonstrate a significant improvement in the description of light fragments with correlations. Utilizing either type of Gaussian wave functions, with or without correlations, however, we find that we cannot describe fragment formation in a dynamic situation. Composite fragments are only produced in simulations if they are present as clusters in the substructure of original nuclei. The difficulty is traced

View details →
Negative / Null Result ReportOpen accessComputer Science

Do Language Models Understand Measurements?

Sungjin Park, Seungwoo Ryu, Edward Choi · 2022 · arXiv

Recent success of pre-trained language models (PLMs) has stimulated interest in their ability to understand and work with numbers. Yet, the numerical reasoning over measurements has not been formally studied despite their importance. In this study, we show that PLMs lack the capability required for reasoning over measurements. Furthermore, we find that a language model trained on a measurement-rich corpus shows better performance on understanding measurements. We propose a simple embedding strategy to better distinguish between numbers and units, which leads to a significant improvement in the

View details →
Negative / Null Result ReportOpen accessComputer Science

The Effect of Diversity in Meta-Learning

Ramnath Kumar, Tristan Deleu, Yoshua Bengio · 2022 · arXiv

Recent studies show that task distribution plays a vital role in the meta-learner's performance. Conventional wisdom is that task diversity should improve the performance of meta-learning. In this work, we find evidence to the contrary; (i) our experiments draw into question the efficacy of our learned models: similar manifolds can be learned with a subset of the data (lower task diversity). This finding questions the advantage of providing more data to the model, and (ii) adding diversity to the task distribution (higher task diversity) sometimes hinders the model and does not lead to a signi

View details →
Failed Experiment ReportOpen accessComputer Science

Fragments in Gaussian Wave-Packet Dynamics with and without Correlations

P. Danielewicz, D. Kiderlen · 1997 · arXiv

Generalization of Gaussian trial wave functions in quantum molecular dynamics models is introduced, which allows for long-range correlations characteristic for composite nuclear fragments. We demonstrate a significant improvement in the description of light fragments with the correlations. Utilizing either type of Gaussian wave functions, with or without correlations, however, we find that we cannot describe fragment formation in a dynamic situation. Composite fragments are only produced in simulations if these fragments are present as clusters in the substructure of original nuclei. The diffi

View details →
Replication FailureOpen accessComputer Science

Spin Transfer Torque Driven Coupled Oscillators for Self-Oscillating RF Mixers

Supriyo Maji · 2017 · arXiv

Spin transfer torque oscillators (STOs) based on magnetic tunnel junction (MTJ) devices are emerging as a possible replacement for complementary metal-oxide semiconductors for radio-frequency (RF) signal generation. Advantages include low power consumption, small device area, and large frequency tunability. But such a single device cannot achieve the necessary noise performance for RF applications. It has been reported lately that a network of globally coupled STOs achieves significant improvement in phase noise. The study here is to propose use of such coupled STOs as self-oscillating RF mixe

View details →
Negative / Null Result ReportOpen accessComputer Science

Hadronic Contribution to (g-2)_{mu}

Andreas Hocker · 2001 · arXiv

The recent precise measurement of the muon magnetic anomaly (g-2)_{mu} at BNL opens a window into possible new physics, provided the contribution from hadronic vacuum polarization is well understood. This talk summarizes the development in the evaluation of the leading order hadronic contributions. Significant improvement has been achieved in a series of analyses which is presented historically in three steps: (1), use of tau spectral functions in addition to e+e- cross sections, (2), extended use of perturbative QCD and (3), application of QCD sum rule techniques. The uncertainties, in partic

View details →
Failed Experiment ReportOpen accessComputer Science

Improved Channel Coding Performance Through Cost Variability

Adeel Mahmood, Aaron B. Wagner · 2024 · arXiv

Channel coding for discrete memoryless channels (DMCs) with mean and variance cost constraints has been recently introduced. We show that there is an improvement in coding performance due to cost variability, both with and without feedback. We demonstrate this improvement over the traditional almost-sure (per-codeword) cost constraint that prohibits any cost variation above a fixed threshold. Our result simultaneously shows that feedback does not improve the second-order coding rate of simple-dispersion DMCs under the almost-sure cost constraint. This finding parallels similar results for unco

View details →
Negative / Null Result ReportOpen accessComputer Science

Trading Inference-Time Compute for Adversarial Robustness

Wojciech Zaremba, Evgenia Nitishinskaya, Boaz Barak et al. · 2025 · arXiv

We conduct experiments on the impact of increasing inference-time compute in reasoning models (specifically OpenAI o1-preview and o1-mini) on their robustness to adversarial attacks. We find that across a variety of attacks, increased inference-time compute leads to improved robustness. In many cases (with important exceptions), the fraction of model samples where the attack succeeds tends to zero as the amount of test-time compute grows. We perform no adversarial training for the tasks we study, and we increase inference-time compute by simply allowing the models to spend more compute on reas

View details →
Negative / Null Result ReportOpen accessComputer Science

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection

Hongxu Zhou · 2026 · arXiv

Intrinsic self-correction in Large Language Models (LLMs) frequently fails in open-ended reasoning tasks due to ``hallucination snowballing,'' a phenomenon in which models recursively justify early errors during free-text reflection. While structured feedback can mitigate this issue, existing approaches often rely on externally trained critics or symbolic tools, reducing agent autonomy. This study investigates whether enforcing structured reflection purely through Outlines-based constrained decoding can disrupt error propagation without additional training. Evaluating an 8-billion-parameter mo

View details →
Negative / Null Result ReportOpen accessComputer Science

Durkheim Project Data Analysis Report

Linas Vepstas · 2013 · arXiv

This report describes the suicidality prediction models created under the DARPA DCAPS program in association with the Durkheim Project [http://durkheimproject.org/]. The models were built primarily from unstructured text (free-format clinician notes) for several hundred patient records obtained from the Veterans Health Administration (VHA). The models were constructed using a genetic programming algorithm applied to bag-of-words and bag-of-phrases datasets. The influence of additional structured data was explored but was found to be minor. Given the small dataset size, classification between c

View details →
Negative / Null Result ReportOpen accessComputer Science

Good Data, Large Data, or No Data? Comparing Three Approaches in Developing Research Aspect Classifiers for Biomedical Papers

Shreya Chandrasekhar, Chieh-Yang Huang, Ting-Hao 'Kenneth' Huang · 2023 · arXiv

The rapid growth of scientific publications, particularly during the COVID-19 pandemic, emphasizes the need for tools to help researchers efficiently comprehend the latest advancements. One essential part of understanding scientific literature is research aspect classification, which categorizes sentences in abstracts to Background, Purpose, Method, and Finding. In this study, we investigate the impact of different datasets on model performance for the crowd-annotated CODA-19 research aspect classification task. Specifically, we explore the potential benefits of using the large, automatically

View details →
Negative / Null Result ReportOpen accessComputer Science

Data filtering methods for training language models

Egor Shevchenko, Elena Bruches · 2026 · arXiv

Data quality is a critical factor in the effectiveness of machine learning models. Label errors, present even in widely used benchmarks, introduce noise into training data and reduce model generalization. In this work, we conduct a comparative analysis of two automatic label error detection methods - Confident Learning and Dataset Cartography - on three Russian text classification corpora of varying size, number of classes, and domain: ru_emotion_e-culture (49,123 examples, emotion classification), RuCoLA (8,524 examples, linguistic acceptability), and TERRa (2,337 examples, textual entailment

View details →
Negative / Null Result ReportOpen accessComputer Science

Corrective In-Context Learning: Evaluating Self-Correction in Large Language Models

Mario Sanz-Guerrero, Katharina von der Wense · 2025 · arXiv

In-context learning (ICL) has transformed the use of large language models (LLMs) for NLP tasks, enabling few-shot learning by conditioning on labeled examples without finetuning. Despite its effectiveness, ICL is prone to errors, especially for challenging examples. With the goal of improving the performance of ICL, we propose corrective in-context learning (CICL), an approach that incorporates a model's incorrect predictions alongside ground truth corrections into the prompt, aiming to enhance classification accuracy through self-correction. However, contrary to our hypothesis, extensive exp

View details →
Negative / Null Result ReportOpen accessComputer Science

Topic Level Disambiguation for Weak Queries

Hui Zhang, Kiduk Yang, Elin Jacob · 2015 · arXiv

Despite limited success, information retrieval (IR) systems today are not intelligent or reliable. IR systems return poor search results when users formulate their information needs into incomplete or ambiguous queries (i.e., weak queries). Therefore, one of the main challenges in modern IR research is to provide consistent results across all queries by improving the performance on weak queries. However, existing IR approaches such as query expansion are not overly effective because they make little effort to analyze and exploit the meanings of the queries. Furthermore, word sense disambiguati

View details →