AO
Back
Cursor Common Problems

Design a Distributed Cron / Job Scheduler

System DesigneasyLast reported September 2026
By AceOffer · Updated September 2026 · Reported 2× across 14 reports

Understanding the Problem

A Cursor system design round, two reports (August and September 2026). In one it was introduced as "design CI/CD", but the candidate says it turned out to be a job scheduler: the interviewer said CI/CD is just the workload. The interviewer's focus was how scheduled jobs are triggered, the task queue and how work is scheduled onto workers, and how worker failures are handled. The other report lists the round as a cron job scheduler. No report gives scale numbers, an API or a data model, so those are yours to propose and ask about. The candidate who described the focus passed this round and went on to the next ones.

Functional Requirements

Structured requirements coming soon. For now, see the full problem statement above and the deep-dive prompts below.

Non-Functional Requirements

Latency, throughput, availability, consistency targets — being authored.

The Set Up

Defining the Core Entities

Core entities (Request, Batch, Worker, Cache, etc.) — being authored.

The API

POST /endpoint → describe request shape GET /endpoint → describe response shape (API spec being authored)

High-Level Design

Component diagram + walkthrough mapping each functional requirement to a system flow — being authored.

Potential Deep Dives

These are the directions the interviewer is likely to push you. Each one has multiple valid solutions at different quality tiers.

1)How are scheduled jobs triggered on time? (when: the core design)

Bad

Naive approach with serious trade-off — being authored.

Good

Solid baseline with reasonable trade-offs — being authored.

Great

Production-grade approach with explicit trade-off rationale — being authored.

2)How do the task queue and the workers share out the work? (when: after the trigger)

Bad

Naive approach with serious trade-off — being authored.

Good

Solid baseline with reasonable trade-offs — being authored.

Great

Production-grade approach with explicit trade-off rationale — being authored.

3)What happens when a worker fails in the middle of a job? (when: after the queue)

Bad

Naive approach with serious trade-off — being authored.

Good

Solid baseline with reasonable trade-offs — being authored.

Great

Production-grade approach with explicit trade-off rationale — being authored.

What is Expected at Each Level?

L4 / Mid-level

Cover happy path. Clarify scope. Identify the obvious bottleneck. Pick a reasonable storage and reasonable scaling approach.

L5 / SeniorTarget

All of the above plus: explicit failure handling, durability vs latency trade-offs, choose the right batching/caching strategy, articulate why.

L6 / Staff+

All of the above plus: organizational concerns (rollout, migration, on-call), quantitative analysis, multi-region considerations, what could go wrong with the proposed solution at 10x scale.

Insider Notes

Interviewer hints: The interviewer said CI/CD is only the workload; the focus is triggering scheduled jobs, the task queue and worker scheduling, and handling worker failure (one report).

Cursor · System Design · Last reported September 2026
Is this helpful?