build scalable backend services

Designs and implements high‑concurrency backend services and server‑side architectures (often in languages like Python and Node.js), selecting and applying concurrency primitives, patterns, and control mechanisms to achieve scalable throughput and low latency. Builds and analyzes concurrency management, debugging, and optimization solutions — diagnosing and fixing race conditions and deadlocks, tuning multithreaded or asynchronous code, and architecting systems for high concurrency and scalability.

buildscalablebackendservices

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
-0.6
Oct 01, 2026Oct 01, 2026
Career
Value
No comparison yet
$202K/year
Oct 01, 2026Oct 01, 2026

Must-Read Papers

Most classic and influential ideas
View more

MicroRacer: Detecting Concurrency Bugs for Cloud Service Systems

Dec 05, 2025
ZD
Zhiling Deng
🏛️ Sun Yat-sen University

Distributed concurrent execution of microservices in cloud environments frequently induces race conditions; however, existing detection techniques suffer from high intrusiveness and poor adaptability to cross-service interactions. Method: This paper proposes a non-intrusive, automated concurrency vulnerability detection framework that leverages runtime library-level dynamic instrumentation and distributed tracing to model cross-service happened-before relationships and resource access patterns, coupled with a three-stage validation mechanism for precise, code-modification-free identification and localization of race conditions. Contribution/Results: Evaluated on an industrial-grade open-source microservice benchmark, our approach achieves significantly higher detection rates and substantially lower false-positive rates compared to state-of-the-art methods, while demonstrating strong architectural adaptability across diverse microservice deployments.

Analyzes runtime traces to identify suspicious concurrent operationsDetects concurrency bugs in microservice-based cloud systemsValidates bugs through a non-intrusive, automated three-stage process

Deep Learning Based Concurrency Bug Detection and Localization

Aug 28, 2025
ZF
Zuocheng Feng
🏛️ Tongji University | Beijing Institute of Control Engineering | Hong Kong University of Science and Technology | Hong Kong University

Concurrency bugs—arising from improper synchronization of shared resources—pose severe reliability threats to multithreaded and distributed systems, yet remain notoriously difficult to detect and localize precisely. This paper introduces the first deep learning framework for concurrent bug detection and code-level precise localization. First, we construct a large-scale, diverse, and domain-specific dataset of concurrency bugs. Second, we design a concurrency-aware Code Property Graph (CCPG) and an associated heterogeneous Graph Neural Network (GNN) that explicitly models critical semantic features, including thread interactions, lock ordering, and data races. Third, we integrate SubgraphX, a model-agnostic explainability technique, to enable end-to-end transition from binary classification to line-level fault localization. Experimental evaluation demonstrates that our approach achieves average improvements of 10% in accuracy and precision, and 26% in recall over state-of-the-art methods.

Addressing limitations in semantic representation and datasetsDetecting concurrency bugs in multi-threaded systemsLocalizing bugs to specific source code lines

DR.FIX: Automatically Fixing Data Races at Industry Scale

Apr 22, 2025
FB
Farnaz Behrang
🏛️ Uber Technologies | Aarhus University

Automatically repairing data races in industrial-scale shared-memory concurrent programs—particularly Go-based microservices—remains highly challenging due to complex concurrency semantics and context-sensitive interference patterns. Method: This paper introduces the first end-to-end repair framework that tightly integrates large language models (LLMs) with precise program analysis. It combines static analysis, interprocedural data-flow tracking, concurrency-aware LLM prompting, and domain-informed code generation to accurately identify and safely fix intricate data race patterns. Contribution/Results: Deployed at Uber for 18 months, the framework automatically generated 224 patches addressing 55% of known data races; 86% were approved by developers and merged into the mainline. This significantly improves reliability and repair efficiency in high-concurrency production systems.

Addressing diverse racy patterns in complex concurrent programming contextsAutomatically fixing prevalent data races in large-scale industrial codebasesIntegrating LLM-based fixes into developer workflows for practical adoption

Concurrency Model of BDI Programming Frameworks: Why Should We Control It?

Apr 16, 2024
MB
Martina Baiardi
🏛️ University of Bologna

BDI programming systems suffer from inconsistent concurrency support and poor customizability, hindering framework selection and extensibility. This paper introduces the first unified taxonomy of concurrency models for BDI frameworks, formally defining a multidimensional assessment framework for customization capabilities. Through an empirical comparative study, we systematically model and evaluate the concurrency mechanisms of prominent frameworks—including Jason, 2APL, and Jadex—identifying critical design trade-offs and limitations. We propose a reusable classification framework that exposes common deficiencies across these systems, particularly in scheduling granularity, intervention depth, and configuration flexibility. Our analysis provides both theoretical foundations and practical guidelines for designing highly controllable, configurable, and concurrency-aware BDI agent architectures. The findings enable principled framework evaluation, informed customization, and targeted enhancement of concurrency support in intelligent agent systems.

BDI Programming SystemsConcurrent ModeCustomization Difficulty

Latest Papers

What's happening recently
View more

This work addresses the challenges students face in understanding and debugging nondeterministic concurrency bugs—such as deadlocks and race conditions—when learning parallel programming. The authors propose ParaView, an educational tool that integrates execution trace visualization with large language model (LLM) analysis, uniquely combining program execution logs, visual representations of parallel behavior, and LLM-driven error explanations and repair suggestions for concurrent programming instruction. In an evaluation with 17 students, the use of ParaView led to significantly higher success rates in both debugging and implementation tasks. Most participants reported that ParaView effectively supported their learning, and the LLM accurately identified common concurrency errors and interpreted execution traces, though its repair suggestions remained limited in complex synchronization scenarios.

concurrency bugsconcurrent programmingdebugging

Hot Scholars

MG

Minyi Guo

IEEE Fellow, Chair Professor, Shanghai Jiao Tong University
Parallel ComputingCompiler OptimizationCloud ComputingNetworking
YH

Yao Hu

浙江大学
Machine Learning
JZ

Jidong Zhai

Tsinghua University
Parallel ComputingCompilerProgramming ModelGPU
YS

Yihan Sun

Assistant Professor, University of California, Riverside
Parallel Algorithms
YZ

Yuheng Zhao

Fudan University
Data VisualizationVisual AnalyticsHuman-AI Collaboration