The score was keyword counting, even after last round's fix. A job listing heavy with AI words scored high whatever the work was.
Now the agent reads the job listing and scores it, 0–100 for fit with the person's profile, with a one-line reason on the card. Keywords stay as flags. It happens in the agent's own session, with no paid API: a first version built on one was reverted the same day, because the choice offered to the person hadn't said so in its label.
When it reads: with its triage vote (one read, so the vote isn't anchored to an earlier number) and each morning for the day's new job listings. Triage prompts still carry no score, so the other models' votes stay blind.
First morning: 55 new job listings, every plausible one read in full; a job listing the agent hasn't read yet keeps its keyword score, labeled as an estimate.
Fixed on the way: "a hybrid of strategy consultant and account owner" had marked remote job listings hybrid.