Library

Subject
Tags

1,793 matches · Software Testing and Debugging Techniques

#
001Provably Lossless Acceleration of DNN Mutation Testing via MemoizationOpenAlexPaperAli Ghanbari, Ben Greenman et al.Jul 20
002Testing Methods for Machine Learning Systems: From Data Validation to Model EvaluationOpenAlexPaperDmitrii KochetovJul 14
003Testing Methods for Machine Learning Systems: From Data Validation to Model EvaluationOpenAlexPaperDmitrii KochetovJul 14
004Research on Fuzzing Mutation Strategy Based on Particle Swarm OptimizationOpenAlexPaperYuting Xiao, Jun Li et al.Jul 9
005<p>Candidate ranges and final hyperparameter settings for RF, XGBoost, and CatBoost under the combined 5D + non-5D feature group.</p>OpenAlexPaperWenfang Li (111747), Xingchen Zhang (3218931) et al.Jul 7
006Collaborative Multi-Agent Testing for Emergent Failure Discovery in Autonomous Driving SystemsOpenAlexPaperRuizhen Gu, Konstantinos Koufos et al.Jul 7
007Pre-registration - Recourse-robustness audit: benchmarking robust-by-design recourse under model retraining (Experiments 1–3)OpenAlexPaperA AnonymousJul 6
008Pre-registration - Recourse-robustness audit: benchmarking robust-by-design recourse under model retraining (Experiments 1–3)OpenAlexPaperA AnonymousJul 6
009JunoBench: A Benchmark Dataset of Crashes in Python Machine Learning Jupyter NotebooksOpenAlexPaperYiran Wang, José Antonio Hernández López et al.Jul 5
010Fault Detection and Explainable Classification in Automotive HIL Validation via Denoising Autoencoders and In-Context Large Language ModelsOpenAlexPaperMohammad Abboush, Hamza Ouarrad et al.Jul 4
011ShardFlow ML: Design and Empirical Evaluation of a Deterministic Manifest, Planning, and Checkpoint Layer for Reproducible Machine-Learning Data PipelinesOpenAlexPaperV D. GUPTAJul 4
012ShardFlow ML: Design and Empirical Evaluation of a Deterministic Manifest, Planning, and Checkpoint Layer for Reproducible Machine-Learning Data PipelinesOpenAlexPaperV D. GUPTAJul 4
013Adaptive Multi‐Metric Test Case Selection for Deep Neural Networks Based on Genetic AlgorithmOpenAlexPaperWeiwei Wang, Qingshuai Chen et al.Jul 1
014LLM SECURITY PENTESTING A Systematic Methodology for Adversarial Testing of Large Language Model SystemsOpenAlexPaperMohd. Arslan NasirJul 1
015PRTS: Test Sample Selection Based on Category Probability Repair for DNN TestingOpenAlexPaperFeifan Gao, Zhiyi Zhang et al.Jul 1
016Active Learning of Symbolic Automata for Reactive Programs via Dynamic Symbolic MapperOpenAlexPaperY Y Kim, Yunja ChoiJun 30
017DECODE: Dynamic Exploration for Constraint-Guided Vulnerability Discovery in Deep Learning OperatorsOpenAlexPaperHaotong Liu, Zhi Wang et al.Jun 30
018Empirical Insights of Test Selection Metrics under Multiple Testing Objectives and Distribution ShiftsOpenAlexPaperJingyu Zhang, Fan Wang et al.Jun 30
019Fault Diversity in Reinforcement Learning Policy TestingOpenAlexPaperQuentin Mazouni, Arnaud Gotlieb et al.Jun 29
020How students use generative AI for software testing: An observational studyOpenAlexPaperBarış Ardıç, Quentin Le Dilavrec et al.Jun 20
021A quality-preserving model for test reduction in electronics productionOpenAlexPaperEinav Peretz-Andersson, Noufa Haneefa et al.Jun 17
022Optimized Sequential Testing for Binary Ensemble ClassifiersOpenAlexPaperJoseph Kalman, Amit MoscovichJun 13
023Intent-to-Silicon: A Deterministic Ambiguity-Reduction Framework for Natural Language Software Specification GenerationOpenAlexPaperAyush Kumar MishraJun 12
024Intent-to-Silicon: A Deterministic Ambiguity-Reduction Framework for Natural Language Software Specification GenerationOpenAlexPaperAyush Kumar MishraJun 12
025Dynamically attributed grammatical evolution: function-based attribute grammars for flexible and context-aware mappingOpenAlexPaperDhiraj Kumar Singh, Rajkumar Sarma et al.Jun 10
026Dynamically attributed grammatical evolution: function-based attribute grammars for flexible and context-aware mappingOpenAlexPaperD Singh, Rajkumar Sarma et al.Jun 10
027An Empirical Comparison of General Context-Free ParsersOpenAlexPaperHuan Vo, Danushka Liyanage et al.Jun 7
028Developer Farm: An Architectural Approach to Goodhart-Proof AI Code Generation via Strict Layer IsolationOpenAlexPaperIllia RochevJun 5
029Developer Farm: An Architectural Approach to Goodhart-Proof AI Code Generation via Strict Layer IsolationOpenAlexPaperIllia RochevJun 5
030Token-Efficient Machine Code Representations for Large Language Models: A Complete Research Journey from Hypothesis to Performance ValidationOpenAlexPaperSushanth TiruvaipatiJun 3
031Token-Efficient Machine Code Representations for Large Language Models: A Complete Research Journey from Hypothesis to Performance ValidationOpenAlexPaperSushanth TiruvaipatiJun 3
032COAPT: Bridging Semantic-Operational Divide in Autonomous Penetration Testing Through LLM-Driven Cognitive PlanningOpenAlexPaperHaonan Zhang, Hui Lu et al.May 26
033Self-Evolving Multi-Agent Fuzzing for Industrial IoT with Knowledge-Driven Cognitive ReasoningOpenAlexPaperBowei Ning, Xuejun Zong et al.May 25
034Understanding and Mitigating Multilingual Bias in LLM-Driven Verilog Code Generation via Hard-Example In-Context LearningOpenAlexPaperGuang YangMay 25
035Causal Inference Testing in MVL SignalsOpenAlexPaperEvelyn Tadlock, Allison Reed et al.May 19
036From Active Learning to Inference-Time Governance for LLM-Assisted Code Generation: A Systematic Literature Review and Design FrameworkOpenAlexPaperKetut Adnyana, Andreas SchwungMay 14
037From Active Learning to Inference-Time Governance for LLM-Assisted Code Generation: A Systematic Literature Review and Design FrameworkOpenAlexPaperKetut Adnyana, Andreas SchwungMay 14
038Autonomous Evaluation Architectures: Multi-Agent LLM Pipelines, Browser-Grounded Testing: Programmatic Alignment via DSPy, and Adversarial Robustness in Production Orchestration SystemsOpenAlexPaperVenkata Chandra Sekhar Sastry ChilkuriMay 12
039NeuroFlake: A Neuro-Symbolic LLM Framework for Flaky Test ClassificationOpenAlexPaperKhondaker Tasnia Hoque, Toukir AhammedMay 12
040GLITCH_AI: A Hybrid Framework for Automated Penetration Testing with LLM-Driven Adaptation and ReportingOpenAlexPaperMaria Gonzalez Herrero, Amit Kumar Singh et al.May 8
041GraphRAG-Driven Compliance Validation of Software Requirements: Semantic, Content, and Data Compliance MetricsOpenAlexPaperMatas Čenys, Asta SlotkienėMay 8
042A process-centric review of large language models in graphical user interface testing: architectures, lifecycle impact, and challengesOpenAlexPaperMai Phu Trong, Ngo Tung Son et al.May 6
043ARIADNE: Agentic Reward-Informed Adaptive Decision Exploration via Blackboard-Driven MCTS for Competitive Program GenerationOpenAlexPaperMinnan Wei, Xiang Chen et al.May 4
044Система типов встроенного языка Basic Syntax Language платформы 1С: Предприятие и верификация допустимости преобразований между типами данныхOpenAlexPaperАлександр Владимирович ЛеоновMay 4
045Система типов встроенного языка Basic Syntax Language платформы 1С: Предприятие и верификация допустимости преобразований между типами данныхOpenAlexPaperАлександр Владимирович ЛеоновMay 4
046Using large language models to test voice user interfacesOpenAlexPaperEmanuela Guglielmi, Angelica Spina et al.May 2
047A Review on Testing Approaches for Autonomous Driving Systems Based on Metamorphic Testing and Fuzzing TestingOpenAlexPaperCun Sun, Jinfu Chen et al.May 1
048AI-Driven Approaches to System Requirements and Test Case Generation: A New Paradigm in Software EngineeringOpenAlexPaperZiad Salem, Luay Tahat et al.Apr 25
049Empirical Insights of Test Selection Metrics under Multiple Testing Objectives and Distribution ShiftsOpenAlexPaperJingyu Zhang, Fan Wang et al.Apr 25
050RedOps AI: Multi-Agent Orchestration for Autonomous Penetration TestingOpenAlexPaperSamarth C Dinesh, Sheba Selvam et al.Apr 24
051Test Design and Review Argumentation in AI-Assisted Test GenerationOpenAlexPaperEduard Paul Enoiu, Robert FeldtApr 24
052Styxx v6.0.0 — Three Calibrated Cognometric Instruments + Phase-Transition AblationOpenAlexPaperAlexander RodabaughApr 23
053Traceable Metamorphic Test Cases for Robust Safety‐Critical Systems: A Deep Learning LiDAR Object Detector ExampleOpenAlexPaperSimon Speth, Tobias Springer et al.Apr 23
054DDSA: Dual-Domain Strategic Attack for Spatial-Temporal Efficiency in Adversarial Robustness TestingOpenAlexPaperJinwei Hu, Shiyuan Meng et al.Apr 21
055Model Equality Testing of Black-Box LLM APIs Via Prefix Tree StatisticsOpenAlexPaperKai-Yuan Shi, Fang-Qi Li et al.Apr 21
056PenPlan-PDDL: A Multi-Agent Framework for Automated Penetration Testing Planning With PDDL-Based VerificationOpenAlexPaperYu Qiao, Lun Li et al.Apr 21
057Characterizing and Refactoring Table-Driven Tests in GoOpenAlexPaperMax Green, Lu Xiao et al.Apr 12
058Enhancing Automated Video Game Regression Testing through Behavior-Driven Development and Imitation LearningOpenAlexPaperVincent Mastain, Fábio PetrilloApr 12
059ITEP4SDC at the SBFT 2026 Tool Competition: Self-Driving Car Testing TrackOpenAlexPaperAli İhsan Güllü, Faiz Shah et al.Apr 12
060Improving a Parallel C++ Intel SSE SIMD Linear Genetic Programming InterpreterOpenAlexPaperw langdon, Carol HannaApr 12

Showing 60 of 1,793 documents · scroll for more