Expert-built · updated quarterly
Python — Professional
Used in engineering pipelines · three coding tasks + one debugging exercise
- Duration
- 60 min
- Difficulty
- Advanced
- Format
- Coding + MCQ
- Integrity tier
- Verified
- Credential
- 24 months
What it measures
- Correctness under constraints — three tasks written and submitted as complete work samples.
- Debugging — one broken module with a realistic production failure.
- Code quality — naming, structure and idiom.
- Roadmap: automated execution and rubric scoring aren't live yet — see the note below.
Sample question · multiple choice
What does dict.setdefault(k, []) return when k is already present?
How it's scored
This weighting reflects our intended rubric. Roadmap: automated scoring against it isn't live yet — today your solution is saved exactly as written, as a real work sample.
Integrity: Verified tier — tab-switch monitoring today (identity check and full-screen enforcement are on our roadmap). No webcam recording. Scope is disclosed in the pre-check and flags are always human-reviewed. Details