Originally published on AI Tech Connect.
Almost everyone can call an API. Almost nobody can tell you if it got better Sit on the other side of an interview loop for a few months and a pattern emerges that is almost boring in its consistency. Candidates can wire up a model. They can build retrieval. Many of them can put an agent loop together, handle tool calls, stream tokens to a front end and deploy the whole thing behind a gateway. Then you ask one question — how did you know the change you made was an improvement? — and the room goes quiet. The honest answers are variations on the same theme. "It looked better." "The demo worked." "We tried a few prompts and picked the one we liked." Occasionally someone mentions a benchmark score from a model card, which is a measurement of somebody else's task, not theirs. Very rarely,…
Top comments (0)