Conceptio
›
software-testing
Topic
software-testing
Knowledge-graph topic
· documents ABOUT software-testing across the archive
17
Documents about software-testing
Documents about software-testing
BONSAI: A Mixed-Initiative Workspace for Human-AI Co-Development of Visual Analytics Applications
#124133
arXiv CS
SpecSyn: LLM-based Synthesis and Refinement of Formal Specifications for Real-world Program Verification
#126584
arXiv CS
mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code
#126589
arXiv CS
UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems
#139017
arXiv CS
AI as Consumer and Participant: A Co-Design Agenda for MBSE Substrates and Methodology
#141557
arXiv CS
Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora
#141572
arXiv CS
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
#155356
arXiv CS
How Compliant Are GitHub Actions Workflows? A Checklist-Based Study with LLM-Assisted Auditing
#155399
arXiv CS
Correct Code, Vulnerable Dependencies: A Large Scale Measurement Study of LLM-Specified Library Versions
#168376
arXiv CS
MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System
#168385
arXiv CS
AICoFe: Implementation and Deployment of an AI-Based Collaborative Feedback System for Higher Education
#168401
arXiv CS
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code
#168405
arXiv CS
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
#168408
arXiv CS
When Context Hurts: The Crossover Effect of Knowledge Transfer on Multi-Agent Design Exploration
#168414
arXiv CS
Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution
#175326
arXiv CS
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
#175344
arXiv CS
Minimalistic Terminal Editor for Julia Programming -- MinTEJ: A Friendly Approach for a Scientific Programmer
#178934
arXiv CS
It's Not the Size: Harness Design Determines Operational Stability in Small Language Models
#178939
arXiv CS
Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought
#187371
arXiv CS
LogRouter: Adaptive Two-Level LLM Routing for Log Question Answering in Big Data Systems
#200548
arXiv CS
LLM-Based Static Verification of Code Against Natural-Language Requirements: An Industrial Experience Report
#200552
arXiv CS
EGI: A Multimodal Emotional AI Framework for Enhancing Scrum Master Real-time Self-Awareness
#200553
arXiv CS
Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review
#200559
arXiv CS
Stop Starving or Stuffing Me: Boosting Firmware Fuzzing Efficiency with On-demand Input Delivery
#200578
arXiv CS
LLM-based vs. Search-based Merge Conflict Resolution: An Empirical Study of Competing Paradigms
#200582
arXiv CS
Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization
#204841
arXiv CS
When to Answer and When to Defer: A Decision Framework for Reliable Code Predictions
#204851
arXiv CS
Sibyl-AutoResearch: Autonomous Research Needs Self-Evolving Trial-and-Error Harnesses, Not Paper Generators
#216888
arXiv CS
The 2nd Workshop on Agile Practice & Research: A Summary and Call For Research
#216894
arXiv CS
SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository Setup?
#229578
arXiv CS
A Universal Cliff and a Design Fingerprint: Cross-Section Defect Detection Under LLM Orchestration
#229583
arXiv CS
Leveraging Language Models for Log Statement Generation in Multilingual Scenarios: How Far Are We?
#229584
arXiv CS
Gamified Requirement Elicitation for a Multi-Modal Decision Support System. The Case of SYNCHROMODE
#229590
arXiv CS
Confident Learning-based Network for Detecting Bug-Inducing Commits on SZZ with Noisy Labels
#241584
arXiv CS
AndroidDaily: A Verifiable Benchmark for Mobile GUI Agents on Real-World Closed-Source Applications
#241585
arXiv CS
Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems
#241590
arXiv CS
When Surface Form Changes Moderation Decisions: A Paired Study of Code-Mixed Workflow Instability
#259544
arXiv CS
Context-as-AI-Service: Surfacing Cross-File Dependency Chains for LLM-Generated Developer Documentation
#259569
arXiv CS
Architecturally Significant MLOps Guidelines for ML Model Integration and Deployment: a Gray Literature Review
#266255
arXiv CS
Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey
#267724
arXiv CS
Minimum Complete MR Subsets under Semantic-Mutation Fault Models: A Support-Set Domination Boundary
#267729
arXiv CS
DD-GEPA: Prompt Optimization for Dialogue Disentanglement Focusing on Task Instruction and Utterance Representation
#267734
arXiv CS
Toward Instructions-as-Code: Understanding the Impact of Instruction Files on Agentic Pull Requests
#271911
arXiv CS
Mind your key: An Empirical Study of LLM API Credential Leakage in iOS Apps
#271921
arXiv CS
Repository-Level Solidity Code Generation with Large Language Models: From Prompting to Fine-Tuning
#290646
arXiv CS
Long-Range Correlation in Code Commit Dynamics as a Novel Indicator of Software Product Stability: A Detrended Fluctuation Analysis Study
#155364
arXiv CS
AdaLoad-JMeter: CSV file of trainings and evaluations
#221740
Figshare
Shedding Light onto Safety Integrity Level and Basic Software Constraints in a Real-World Automotive Application: Case Study with Driverator Framework
#168396
arXiv CS
Can LLMs Deobfuscate Binary Code? A Systematic Analysis of Large Language Models into Pseudocode Deobfuscation
#2670
arXiv CS
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
#2682
arXiv CS
← Previous
Page 17 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.