Coding agents and vibe coding

How to work with agents that write code without fooling yourself about the result: a spec instead of an assignment, rules files and skills, hooks as deterministic guarantees, subagents, independent verification and review.

 

Loading…

Take the whole "Coding agents and vibe coding" topic

Every question of the topic — one per fact, from easy to hard. An honest check of the whole topic rather than of a single section.

#QUIZQUESTIONSDIFFICULTYSTATUS
5.1The harness around the modelWhat belongs to the harness and what to the environment, The evolution of engineering approaches, The five elements of the harness
5.2Specs and task briefsWhat a spec answers, The path of a task, A requirement that describes the implementation
5.3Rules files and skillsWhat to write in the rules file, A superfluous line in the rules file, Checking that the rules arrived
5.4Hooks and deterministic guaranteesA hook or a rule in the prompt, Why a hook beats a rule, A first hook in fifteen minutes
5.5Subagents and delegationWhy delegate to a subagent, What gets lost in the handover, Which work to give to subagents
5.6Independent verification and reviewWhy you cannot judge yourself, Review questions that work, The review loop
5.7Autonomy and intent driftA request for an overnight run, A mistake in the overnight request, The day before an overnight run

Ranking needs a sign-in

A ranked attempt is available after signing in with Telegram: that way one person is one participant in the ranking.

How the quizzes and the fair ranking work

01Modes

Practice
No timer and no limit on the number of attempts. After every answer you see the correct option and an explanation. It does not affect the ranking. Available without signing in.
Ranked attempt
Time for each question is limited; correctness and explanations come only after the last question. The result goes into the ranking. Requires signing in with Telegram.
Diagnostic
32 questions, about 20 minutes: for each of the eight topics one easy, two medium and one hard question, new every time. It is a check for yourself: no sign-in and no ranking, and at the end — a knowledge map and a list of what is worth brushing up on. Topics and their sections count toward the ranking.

02Time and answers

  • The server keeps the time, not the browser: from 30 to 90 seconds per question depending on difficulty and format. Reloading the page does not reset the timer — the question comes back with the same time.
  • Questions come in different formats: pick one option, mark all correct ones, put steps in order, find the faulty line.
  • If time runs out, the question counts as skipped. A small allowance for network delay is built in.
  • You cannot go back to a previous question. Sending an answer again changes nothing — the first one counts.
  • Number keys select elements, Enter sends the answer. In practice you can open a hint.
  • The "Not sure about the answer" mark does not affect the points, but a correct answer with it goes into the review as a topic worth brushing up on.
  • You can change the site language at any time, even in the middle of an attempt: the question and the timer stay the same.

03Limits on ranked attempts

  • One ranked attempt per quiz per day. Days are counted in UTC — a new day starts at 00:00 UTC.
  • Switching to another tab or minimizing the window during a ranked question counts it as skipped.
  • Leaving the quiz early: the remaining questions count as skipped, and you can retry this quiz the next day.

04How points are counted

  • Every question has its own difficulty rating, refined by the answers of all participants. A correct answer to a hard question gives more points, a mistake on an easy one costs more.
  • A skip and "time is up" count as mistakes. The "not sure" mark does not affect the points.
  • Every participant has an overall rating and a separate rating for each topic. The starting rating is 1000.
  • The ranking is the same for the Russian and English versions of the site: the questions and their difficulty are shared.
  • The "Week" and "Month" tables show how many points a participant gained or lost over the calendar week or month. "All time" is the current rating.
  • Sometimes a control question is added to a ranked attempt. It does not count toward the score.

05A fair ranking

  • One Telegram account is one participant. When you sign in, we receive only your name and username.
  • An attempt clearly taken by a program rather than a person does not count toward the ranking.
  • Fast answers alone are not penalised: strong participants answer fast too.