Assessment
AI engineer assessment by conversation, not multiple choice
Aimed at the engineer being interviewed rather than the employer doing the screening: you explain AI and ML systems aloud, and the follow-ups go wherever the reasoning thins.
15 free minutes every week
- About 15 minutes
- No pass mark
- Evidence-backed
Coverage
3 of 6 areas checked
- Evaluation designChecked in this session
- Retrieval and groundingChecked in this session
- Serving latencyChecked in this session
- Drift monitoringNot yet tested
- Cost under loadNot yet tested
- Incident triageNot yet tested
Most results for this search are hiring tests. This is not one.
Search the phrase and what comes back is mostly pre-employment screening — a fixed question bank, a cut-off score, a report that goes to a recruiter rather than to you. This sits on the other side of that transaction. Same subject matter, no result anyone else receives, and nothing to pass. What it produces is a map of which areas the conversation actually reached and the single one most worth your next study block.
Fifteen minutes later, you have three things.
- A map of what was actually checked
- A gap traced back to what you said
- One thing to study tonight
What the assessment covers
The ground an AI engineer is expected to hold, taken from the five roles the diagnostic carries questions for: ML Engineer, LLM Engineer, Applied AI Engineer, AI Infrastructure Engineer, and MLOps. General software engineering sits outside it, and so does anything the question bank cannot support — an assessment that stretches past its evidence is guessing with a confident tone.
- Model training, evaluation design, and choosing the metric that matches the decision
- LLM application ground: retrieval, grounding, and how those fail in production
- Serving, latency, and cost trade-offs under real traffic
- Monitoring, drift, and triaging a system that worked last quarter
What it opens with, and why there is no question list
It starts somewhere ordinary for your role — how you would evaluate a retrieval step on its own, why a model that scored well offline degraded after release, what you would check first when serving latency doubled. Those are openings, not the assessment. What is being measured is the second and third question, which are written from your answer and therefore cannot be printed in advance.
Explanation as the measurement
You are asked to explain a concept in your own words. There are no options to select between, nothing to eliminate, and no way to pattern-match to a familiar-looking answer. If the explanation thins out under a follow-up, that is the signal.
- No answer options to recognise or eliminate
- Follow-ups probe the specific claim you just made
- Depth is assessed categorically, never as a percentage
What the fifteen minutes look like
You pick a role, you talk, and Veda follows. Clean explanations are marked and left behind; a fluent but hollow one gets pressed until it either holds or gives way. There is no timer to beat and no counter telling you how far through you are, because the path length depends on where your answers lead. When the session closes, the map is waiting.
Honest coverage reporting
The result separates three states that a single score would blur together: concepts you explained clearly, concepts where the explanation ran out, and concepts the session never reached. The third category is reported as untested rather than being folded into an aggregate that would misrepresent it.
Where this parts company with a scored screening test
A screening test exists to rank one candidate against others for someone who is hiring, so it needs a fixed instrument and a comparable number at the end. This has neither, and the absence is deliberate rather than a missing feature.
- Nothing is sent to an employer, because no employer is on the other end
- No cut-off, percentile, or comparison against other engineers
- Every gap it names shows the reasoning behind it, and you can dispute it
Who should skip it
If you are hiring and need a scored instrument to compare applicants, this is the wrong tool and will frustrate you within a minute. Same if you want a bank of questions to memorise, or if your interviews are general backend, DSA, or behavioural. It is worth your time when you have prepared for AI/ML rounds already and cannot tell which part of that preparation is thinner than it feels.
Nothing to install. Talk it through, keep the map.
Before you start
- Is this a test with a pass mark?
- No. There is no pass mark, score, or ranking against other candidates. The output is a coverage map, one possible gap with its evidence, and one next action.
- Can employers use this to screen AI engineers?
- It is not built for that. Nothing is scored, nothing is comparable between two people, and no result leaves the account that produced it — so there would be nothing to screen on. Training providers running cohort diagnostics are a different case and have their own surface.
- How is this different from a multiple-choice assessment?
- Multiple choice measures whether you recognise a correct answer. This requires you to produce the explanation yourself and then defend it under a follow-up, which is what interviews and production incidents actually demand.
- Does "AI engineer" mean something narrower than ML engineer here?
- The five roles overlap and the diagnostic treats them as one body of ground with different centres of gravity. Pick the title closest to the work you are interviewing for and the questions weight toward it; picking Applied AI over ML Engineer changes the emphasis, not whether a concept is reachable.
- Which skills are in scope?
- Those belonging to ML Engineer, LLM Engineer, Applied AI Engineer, AI Infrastructure Engineer, and MLOps work. It is not a general engineering skills assessment, and for a role it has no questions for it returns no result rather than a weak one.
- Is the assessment free?
- Accounts receive 15 minutes weekly on a renewing cycle, no card involved, which covers one assessment start to finish. Assessing more often than weekly costs $29/month — 60% off the regular $72.50 price.
Find the concept your prep has not tested.
15 free minutes every week. No card. One gap map, one next action.