FSE 2023
205 papers
- "We Feel Like We're Winging It: " A Study on Navigating Open-Source Dependency Abandonment
- A Case Study of Developer Bots: Motivations, Perceptions, and Challenges
- A Data Set of Extracted Rationale from Linux Kernel Commit Messages
- A Four-Year Study of Student Contributions to OSS vs. OSS4SG with a Lightweight Intervention
- A Generative and Mutational Approach for Synthesizing Bug-Exposing Test Cases to Guide Compiler Fuzzing
- A Highly Scalable, Hybrid, Cross-Platform Timing Analysis Framework Providing Accurate Differential Throughput Estimation via Instruction-Level Tracing
- A Language Model of Java Methods with Train/Test Deduplication
- A Large-Scale Empirical Review of Patch Correctness Checking Approaches
- A Multidimensional Analysis of Bug Density in SAP HANA
- A Practical Human Labeling Method for Online Just-in-Time Software Defect Prediction
- A Unified Framework for Mini-game Testing: Experience on WeChat
- A Vision on Intentions in Software Engineering
- AG3: Automated Game GUI Text Glitch Detection Based on Computer Vision
- API-Knowledge Aware Search-Based Software Testing: Where, What, and How
- Accelerating Continuous Integration with Parallel Batch Testing
- Ad Hoc Syntax-Guided Program Reduction
- Adapting Performance Analytic Techniques in a Real-World Database-Centric System: An Industrial Experience Report
- AdaptivePaste: Intelligent Copy-Paste in IDE
- An Automated Approach to Extracting Local Variables
- An Extensive Study on Adversarial Attack against Pre-trained Models of Code
- Analyzing Microservice Connectivity with Kubesonde
- Appaction: Automatic GUI Interaction for Mobile Apps via Holistic Widget Perception
- Assess and Summarize: Improve Outage Understanding with Large Language Models
- Assisting Static Analysis with Large Language Models: A ChatGPT Experiment
- Automata-Based Trace Analysis for Aiding Diagnosing GUI Testing Tools for Android
- Automated Test Generation for Medical Rules Web Services: A Case Study at the Cancer Registry of Norway
- Automated Testing and Improvement of Named Entity Recognition Systems
- Automated and Context-Aware Repair of Color-Related Accessibility Issues for Android Apps
- Automatically Resolving Dependency-Conflict Building Failures via Behavior-Consistent Loosening of Library Version Constraints
- BFSig: Leveraging File Significance in Bus Factor Estimation
- Baldur: Whole-Proof Generation and Repair with Large Language Models
- Benchmarking Robustness of AI-Enabled Multi-sensor Fusion Systems: Challenges and Opportunities
- Beyond Sharing: Conflict-Aware Multivariate Time Series Anomaly Detection
- BiasAsker: Measuring the Bias in Conversational AI System
- BigDataflow: A Distributed Interprocedural Dataflow Analysis Framework
- Building and Sustaining Ethnically, Racially, and Gender Diverse Software Engineering Teams: A Study at Google
- CAmpactor: A Novel and Effective Local Search Algorithm for Optimizing Pairwise Covering Arrays
- CCT5: A Code-Change-Oriented Pre-trained Model
- CONAN: Statically Detecting Connectivity Issues in Android Applications
- Can Machine Learning Pipelines Be Better Configured?
- Co-dependence Aware Fuzzing for Dataflow-Based Big Data Analytics
- Code Coverage Criteria for Asynchronous Programs
- CodeMark: Imperceptible Watermarking for Code Datasets against Neural Code Completion Models
- Commit-Level, Neural Vulnerability Detection and Assessment
- Comparison and Evaluation on Static Application Security Testing (SAST) Tools for Java
- Compatibility Issues in Deep Learning Systems: Problems and Opportunities
- Compositional Taint Analysis for Enforcing Security Policies at Scale
- Contextual Predictive Mutation Testing
- Contribution-Based Firing of Developers?
- Copiloting the Copilots: Fusing Large Language Models with Completion Engines for Automated Program Repair
- Crystallizer: A Hybrid Path Analysis Framework to Aid in Uncovering Deserialization Vulnerabilities
- C³: Code Clone-Based Identification of Duplicated Components
- D2S2: Drag 'n' Drop Mobile App Screen Search
- DENT: A Tool for Tagging Stack Overflow Posts with Deep Learning Energy Patterns
- DeMinify: Neural Variable Name Recovery and Type Inference
- Dead Code Removal at Meta: Automatically Deleting Millions of Lines of Code and Petabytes of Deprecated Data
- DecompoVision: Reliability Analysis of Machine Vision Components through Decomposition and Reuse
- Deep Learning Based Feature Envy Detection Boosted by Real-World Examples
- DeepDebugger: An Interactive Time-Travelling Debugging Approach for Deep Classifiers
- DeepInfer: Deep Type Inference from Smart Contract Bytecode
- DeepRover: A Query-Efficient Blackbox Attack for Deep Neural Networks
- Deeper Notions of Correctness in Image-Based DNNs: Lifting Properties from Pixel to Entities
- Demystifying Dependency Bugs in Deep Learning Stack
- Demystifying the Composition and Code Reuse in Solidity Smart Contracts
- Design by Contract for Deep Learning APIs
- Detecting Atomicity Violations in Interrupt-Driven Programs via Interruption Points Selecting and Delayed ISR-Triggering
- Detecting Overfitting of Machine Learning Techniques for Automatic Vulnerability Detection
- Detection Is Better Than Cure: A Cloud Incidents Perspective
- Detection of Optimizations Missed by the Compiler
- DiagConfig: Configuration Diagnosis of Performance Violations in Configurable Software Systems
- Diffusion-Based Time Series Data Imputation for Cloud Failure Prediction at Microsoft 365
- Discovering Parallelisms in Python Programs
- DistXplore: Distribution-Guided Testing for Evaluating and Enhancing Deep Learning Systems
- Distinguishing Look-Alike Innocent and Vulnerable Code by Subtle Semantic Representation Learning and Explanation
- Do All Software Projects Die When Not Maintained? Analyzing Developer Maintenance to Predict OSS Usage
- Do CONTRIBUTING Files Provide Information about OSS Newcomers' Onboarding Barriers?
- Dynamic Data Fault Localization for Deep Neural Networks
- Dynamic Prediction of Delays in Software Projects using Delay Patterns and Bayesian Modeling
- Efficient Text-to-Code Retrieval with Cascaded Fast and Slow Transformer Models
- Engineering a Formally Verified Automated Bug Finder
- Enhancing Coverage-Guided Fuzzing via Phantom Program
- EtherDiffer: Differential Testing on RPC Services of Ethereum Nodes
- EvaCRC: Evaluating Code Review Comments
- Evaluating Transfer Learning for Simplifying GitHub READMEs
- EvoCLINICAL: Evolving Cyber-Cyber Digital Twin with Active Transfer Learning for Automated Cancer Registry System
- Exploring Moral Principles Exhibited in OSS: A Case Study on GitHub Heated Issues
- Fix Fairness, Don't Ruin Accuracy: Performance Aware Fairness Repair using AutoML
- Flow Experience in Software Engineering
- From Leaks to Fixes: Automated Repairs for Resource Leak Warnings
- From Point-wise to Group-wise: A Fast and Accurate Microservice Trace Anomaly Detection Approach
- FunProbe: Probing Functions from Binary Code through Probabilistic Analysis
- Getting Outside the Bug Boxes (Keynote)
- Getting pwn'd by AI: Penetration Testing with Large Language Models
- Gitor: Scalable Code Clone Detection by Building Global Sample Graph
- Grace: Language Models Meet Code Edits
- Helion: Enabling Natural Testing of Smart Homes
- Heterogeneous Testing for Coverage Profilers Empowered with Debugging Support
- How Early Participation Determines Long-Term Sustained Activity in GitHub Projects?
- How Practitioners Expect Code Completion?
- Hue: A User-Adaptive Parser for Hybrid Logs
- HyperDiff: Computing Source Code Diffs at Scale
- Incrementalizing Production CodeQL Analyses
- InferFix: End-to-End Program Repair with LLMs
- Inferring Complexity Bounds from Recurrence Relations
- Input-Driven Dynamic Program Debloating for Code-Reuse Attack Mitigation
- IoPV: On Inconsistent Option Performance Variations
- Issue Report Validation in an Industrial Context
- KDDT: Knowledge Distillation-Empowered Digital Twin for Anomaly Detection
- KG4CraSolver: Recommending Crash Solutions via Knowledge Graph
- Keeping Mutation Test Suites Consistent and Relevant with Long-Standing Mutants
- Knowledge-Based Version Incompatibility Detection for Deep Learning
- LExecutor: Learning-Guided Execution
- LLM-Based Code Generation Method for Golang Compiler Testing
- Last Diff Analyzer: Multi-language Automated Approver for Behavior-Preserving Code Revisions
- LazyCow: A Lightweight Crowdsourced Testing Tool for Taming Android Fragmentation
- Learning Program Semantics for Vulnerability Detection via Vulnerability-Specific Inter-procedural Slicing
- Lessons from the Long Tail: Analysing Unsafe Dependency Updates across Software Ecosystems
- Leveraging Hardware Probes and Optimizations for Accelerating Fuzz Testing of Heterogeneous Applications
- LibKit: Detecting Third-Party Libraries in iOS Apps
- LightF3: A Lightweight Fully-Process Formal Framework for Automated Verifying Railway Interlocking Systems
- Log Parsing with Generalization Ability under New Log Types
- MASC: A Tool for Mutation-Based Evaluation of Static Crypto-API Misuse Detectors
- Matching Skills, Past Collaboration, and Limited Competition: Modeling When Open-Source Projects Attract Contributors
- Mate! Are You Really Aware? An Explainability-Guided Testing Framework for Robustness of Malware Detectors
- Metamong: Detecting Render-Update Bugs in Web Browsers through Fuzzing
- Mining Resource-Operation Knowledge to Support Resource Leak Detection
- Modeling the Centrality of Developer Output with Software Supply Chains
- MuRS: Mutant Ranking and Suppression using Identifier Templates
- Multilingual Code Co-evolution using Large Language Models
- NaNofuzz: A Usable Tool for Automatic Test Generation
- Natural Language to Code: How Far Are We?
- NeuRI: Diversifying DNN Generation via Inductive Rule Inference
- Neural-Based Test Oracle Generation: A Large-Scale Evaluation and Lessons Learned
- Nezha: Interpretable Fine-Grained Root Causes Analysis for Microservices on Multi-modal Observability Data
- OOM-Guard: Towards Improving the Ergonomics of Rust OOM Handling via a Reservation-Based Approach
- On Using Information Retrieval to Recommend Machine Learning Good Practices for Software Engineers
- On the Dual Nature of Necessity in Use of Rust Unsafe Code
- On the Relationship between Code Verifiability and Understandability
- On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code
- On-Premise AIOps Infrastructure for a Software Editor SME: An Experience Report
- Outage-Watch: Early Prediction of Outages using Extreme Event Regularizer
- Ownership in the Hands of Accountability at Brightsquid: A Case Study and a Developer Survey
- P4b: A Translator from P4 Programs to Boogie
- PEM: Representing Binary Program Semantics for Similarity Analysis via a Probabilistic Execution Model
- PPR: Pairwise Program Reduction
- Pitfalls in Experiments with DNN4SE: An Analysis of the State of the Practice
- Practical Inference of Nullability Types
- Pre-training Code Representation with Semantic Flow Graph for Effective Bug Localization
- Predicting Software Performance with Divide-and-Learn
- Prioritizing Natural Language Test Cases Based on Highly-Used Game Features
- Privacy-Centric Log Parsing for Timely, Proactive Personal Data Protection
- Program Repair Guided by Datalog-Defined Static Analysis
- PropProof: Free Model-Checking Harnesses from PBT
- Property-Based Fuzzing for Finding Data Manipulation Errors in Android Apps
- RAP-Gen: Retrieval-Augmented Patch Generation with CodeT5 for Automatic Program Repair
- Recommending Analogical APIs via Knowledge Graph Embedding
- Reflecting on the Use of the Policy-Process-Product Theory in Empirical Software Engineering
- Revisiting Neural Program Smoothing for Fuzzing
- Rotten Green Tests in Google Test
- SJFuzz: Seed and Mutator Scheduling for JVM Fuzzing
- STEAM: Observability-Preserving Trace Sampling
- STraceBERT: Source Code Retrieval using Semantic Application Traces
- Scalable Program Clone Search through Spectral Analysis
- Self-Supervised Query Reformulation for Code Search
- Semantic Debugging
- Semantic Test Repair for Web Applications
- SmartFix: Fixing Vulnerable Smart Contracts by Accelerating Generate-and-Verify Repair using Statistical Models
- Software Architecture Recovery with Information Fusion
- Software Architecture in Practice: Challenges and Opportunities
- Software Composition Analysis for Vulnerability Detection: An Empirical Study on Java Projects
- Speeding up SMT Solving via Compiler Optimization
- State Merging with Quantifiers in Symbolic Execution
- Statfier: Automated Testing of Static Analyzers via Semantic-Preserving Program Transformations
- Statistical Reachability Analysis
- Statistical Type Inference for Incomplete Programs
- Test Case Generation for Drivability Requirements of an Automotive Cruise Controller: An Experience with an Industrial Simulator
- Testing Coreference Resolution Systems without Labeled Test Sets
- Testing Real-World Healthcare IoT Application: Experiences and Lessons Learned
- The Call Graph Chronicles: Unleashing the Power Within
- The EarlyBIRD Catches the Bug: On Exploiting Early Layers of Encoder Models for More Efficient Code Classification
- The Most Agile Teams Are the Most Disciplined: On Scaling out Agile Development
- The State of Survival in OSS: The Impact of Diversity
- Towards AI-Driven Software Development: Challenges and Lessons from the Field (Keynote)
- Towards Automated Detection of Unethical Behavior in Open-Source Software Projects
- Towards Efficient Record and Replay: A Case Study in WeChat
- Towards Feature-Based Analysis of the Machine Learning Development Lifecycle
- Towards Greener Yet Powerful Code Generation via Quantization: An Empirical Study
- Towards Strengthening Formal Specifications with Mutation Model Checking
- Towards Top-Down Automated Development in Limited Scopes: A Neuro-Symbolic Framework from Expressibles to Executables
- Towards Understanding Emotions in Informal Developer Interactions: A Gitter Chat Study
- TraceDiag: Adaptive, Interpretable, and Efficient Root Cause Analysis on Large-Scale Microservice Systems
- TransMap: Pinpointing Mistakes in Neural Code Translation
- TransRacer: Function Dependence-Guided Transaction Race Detection for Smart Contracts
- Triggering Modes in Spectrum-Based Multi-location Fault Localization
- Tritor: Detecting Semantic Code Clones by Building Social Network-Based Triads Model
- Understanding Hackers' Work: An Empirical Study of Offensive Security Practitioners
- Understanding Solidity Event Logging Practices in the Wild
- Understanding the Bug Characteristics and Fix Strategies of Federated Learning Systems
- Understanding the Topics and Challenges of GPU Programming by Classifying and Analyzing Stack Overflow Posts
- ViaLin: Path-Aware Dynamic Taint Analysis for Android
- When Function Inlining Meets WebAssembly: Counterintuitive Impacts on Runtime Performance
- llvm2CryptoLine: Verifying Arithmetic in Cryptographic C Programs
- npm-follower: A Complete Dataset Tracking the NPM Ecosystem
- xASTNN: Improved Code Representations for Industrial Practice
- µAkka: Mutation Testing for Actor Concurrency in Akka using Real-World Bugs