Blog
How to Use Jev for Cold Email Agents [Step-by-Step Guide]

How to Use Jev for Cold Email Agents [Step-by-Step Guide]

Summarize with

Letting an AI agent find leads is easy. Letting it choose who receives an email, what it says, and how to answer replies is harder. One poor decision still goes out under your name. 

This guide covers that handover. I’ve already written about using Jev for sales prospecting, cold email outreach, and LinkedIn outreach. 

Those guides focused on running campaigns. This one focuses on the agent itself: which decisions it makes, what Jev handles, what follows fixed rules, and what still comes to me.

If you’re considering Jev for email agents, ask how much your agent can safely do without supervision. The steps below explain how I decide.

Table of Contents

How to Use Jev in a Cold Email Agent

While building this workflow, I kept returning to one question: What can my agent do without asking me? Each step answers part of that question. 

1. Map Every Decision Your Email Agent Makes

A cold email agent makes many small decisions during one campaign. Before adding Jev, I listed each decision and assigned an owner.

Decision Who Owns It
Researching a lead's company and recent news The agent
Writing the email The agent
Does this company fit my ICP? Jev
Does this signal matter for this lead? Jev
Contact now, nurture, or don't contact? Jev
What kind of reply is this? Jev
Stopping the sequence after an unsubscribe or a "no" A fixed rule
Continuing after an out-of-office A fixed rule
Personalizing when there's no strong signal A fixed rule: don't
Launching the campaign Me
Any call Jev isn't sure about Me

The split follows a simple principle. Jev handles questions with a right answer and enough evidence to find it. Research and writing have no single correct answer, so the agent handles them. Anything that cannot be undone, such as emailing someone who asked to be removed, follows a fixed rule.

These decisions sit on top of the rules my agent already follows, including never guessing an email address and asking me before launching anything.

Email agent architecture showing where Jev, Grok, and the Forge Stack tools each act
This image shows the Email agent architecture showing where Jev, Grok, and the Forge Stack tools each act

2. Give the Agent the TypeSafe Skill and API Key

Do you need to write the code that calls Jev? No. My agent writes it, while the TypeSafe skill shows it how.

TypeSafe’s documentation lists agents inventing request or response fields as a common problem. That is why I want my agent to use the skill rather than rely on memory. 

Mine uses the skill whenever it writes code for TypeSafe and checks TypeSafe’s current documentation each time.

TypeSafe agent skill page, including its list of common agent issues
This image shows the TypeSafe agent skill page, including its list of common agent issues

Install the TypeSafe skill instructions with:

npx skills add typesafe-ai/skills --skill typesafe-ai

You can also ask your agent to save the official SKILL.md as a skill, which is what mine did.

Grok agent confirming the TypeSafe skill is saved and explaining when it will use it
This image shows the Grok agent confirming the TypeSafe skill is saved and explaining when it will use it

Next, create an API key in the TypeSafe console and give it to the agent, which stores it securely.

If your agent is not connected to Leadsforge and Salesforge, connect them first. Both work through Salesforge’s Forge MCP using an API key.

TypeSafe console with the API key created for Grok Bot
This image shows the TypeSafe console with the API key created for Grok Bot

3. Turn Each Jev Decision Into a Typed Question

My agent is already a language model, so why ask another model? When an agent reasons in prose, it must read its own paragraph and decide what the answer means. An unsupervised agent should not have to interpret itself.

Jev returns one of three answer types: a pick from a list, a score, or a probability. The agent can act on each result without interpreting it.

The work lies in turning a broad thought into a narrow question.

What I'd Think What I Ask Jev Answer Shape
Is this lead any good? How well does this company match my ICP? Score from 0 to 3
Who is this person? Which of my personas is this, or not a target? Pick
Should the email be personal? Is there enough here to personalize? Probability
Do I email them? Contact now, nurture, or don't contact? Pick

That is the main lesson for anyone using Jev with an email agent. One broad question produces one vague answer. Four narrow questions give your agent four answers it can use.

Using the TypeSafe skill, turn these into Jev questions for every lead: ICP fit as a score from 0 to 3, persona as a pick from my list, enough information to personalize as a probability, and contact now, nurture, or don’t contact as a pick. 

Test them on 20 leads and show me every answer before using them elsewhere.

Grok agent turning the lead checks into Jev questions and testing them on 20 leads
This image shows the Grok agent turning the lead checks into Jev questions and testing them on 20 leads

In my first test, Jev answered the “who” questions confidently but sent all 20 leads to nurture. The agent had not failed. It simply didn't give Jev enough information to judge timing. 

Jev's score, persona and decision for each lead in the first test
This image shows the Jev's score, persona and decision for each lead in the first test

4. Make the Agent Bring Evidence Before Jev Decides

Jev’s answer is only as good as the information it receives. Before any lead reaches Jev, my agent attaches sourced signals from Leadsforge:

  • Job changes
    ‍
  • Funding
    ‍
  • Acquisitions
    ‍
  • Investor activity

Once the same 20 leads had signals, Jev began sending them to all three outcomes instead of putting everyone in nurture.

Jev decisions changing once the agent attached sourced signals
This image shows the Jev decisions changing once the agent attached sourced signals

I gave my agent three evidence rules:

  • A signal without a source does not count.
    ‍
  • Anything it cannot verify gets removed instead of being sent to Jev.
    ‍
  • Job changes flagged during the first pass get checked manually because most appear only in LinkedIn posts. The agent suggested this rule itself.

Before asking Jev about a lead, attach their Leadsforge signals from the past six months and include a source for each one. Remove anything you cannot verify. 

Flag job changes for manual review. If the lead has no signal, tell Jev there is no signal.

Agent flow from finding a signal to Jev's contact decision
This image shows the Agent flow from finding a signal to Jev's contact decision

5. Set an Unclear Route for Low-Confidence Answers

What happens when Jev is unsure? This is the step I would least want an agent to mishandle. To an agent, “contact now, unsure” can look too similar to “contact now.”

Two leads in my test produced that result. One had a new role at a strong company. The other had a new role at a company that barely matched my ICP, and Jev showed very low confidence. Neither should receive an email without a review.

My agent now treats every low-confidence answer as unclear. Unclear leads never enter a campaign. They come to me with their signal and scores, so I can decide quickly without repeating the research.

If any Jev answer has low confidence, mark the lead unclear. Do not add unclear leads to a campaign. Send them to me with the signal, its source, and Jev’s scores.

6. Let the Agent Act on Approved Leads, but Only as Drafts

After Jev approves a lead, my agent handles the work I least want to do manually. It researches what changed at the company, connects that change to what I sell, chooses a sourced angle, and builds the Salesforge campaign.

Grok agent researching approved leads and connecting each angle to a source
This image shows the Grok agent researching approved leads and connecting each angle to a source

The campaign still remains a draft. During my test, the agent enrolled five leads in Salesforge but sent nothing.

Salesforge campaign draft created by the agent with five leads enrolled and no emails sent
This image shows the Salesforge campaign draft created by the agent with five leads enrolled and no emails sent

If I still approve every campaign, what have I automated? Everything before the final approval. I only need to check three things:

  • Are these the leads Jev approved, with nobody from the unclear list?
    ‍
  • Does every personal line include a source?
    ‍
  • Does the sequence follow my instructions?

7. Put Fixed Rules Above Jev for Replies

Reply handling is where a poor decision becomes hardest to reverse. My agent asks Jev five questions about each reply:

  • What category does it belong to?
    ‍
  • Do I need to see it?
    ‍
  • Should the sequence stop?
    ‍
  • Is there clear meeting intent?
    ‍
  • Should the agent draft a response?
Jev’s reply classifier choosing a category and next action for each test reply
This image shows the Jev’s reply classifier choosing a category and next action for each test reply

Jev did not give consistent next steps for negative replies and unsubscribes, so fixed rules handle both. Each stops the sequence, and neither receives a drafted response. 

Out-of-office replies never stop the sequence, while unclear replies always come to me.

The classifier also has a minimum confidence level. Anything below it becomes unclear. Mine is currently 0.5. 

The agent suggested raising it to around 0.7, but only after testing it on real replies. So far, I have only used replies I wrote myself.

8. Accept the Agent's Weekly Run and Turn On Notifications

An agent that only works when I ask does not save much time. Mine offered to check all enrolled leads every Monday for new signals, run them through Jev, and send me a sourced shortlist of people to contact that week.

Grok agent offering three options for a Monday signal check
This image shows the Grok agent offering three options for a Monday signal check

It gave me three options: run weekly on Mondays, run once across every lead, or wait. I would begin with one run. The agent could not tell me the weekly Jev cost, and a single run reveals the volume before you commit.

Turn on notifications in the agent’s Grok Bot settings. You will receive an alert when it finishes or needs your input, so you do not have to keep checking its progress.

Grok Bot agent settings with notifications enabled
This image shows the Grok Bot agent settings with notifications enabled

9. Give the Agent More Autonomy Only After It Earns It

I do not plan to give my agent full autonomy at once. Each part receives more freedom only after it performs reliably in testing. 

  1. Reply routing: Before letting the agent act on replies without me, I want to compare a few hundred real replies with my own labels. Written test replies are cleaner than real ones. 
    ‍
  2. Lead decisions: My next test will label every sent email by opener, signal, CTA, and personalization level. I will compare those labels with replies and meetings to see which Jev decisions deserve more trust. 
    ‍
  3. Cost: Every Jev check costs money. Once I know the cost per lead, I will replace any check that a simple rule can handle. 

For the next 30 days, record every Jev decision beside my decision for the same lead or reply. Group and show every disagreement by question. Do not automate any decision until I approve it. 

Is Jev Worth Using for an Email Agent?

It depends on how much you want the agent to handle without you.

If you review every lead and reply yourself, Jev’s fixed answers add little value because you already make each decision.

Jev becomes useful when your agent handles more than you can review manually, such as hundreds of leads, several signals per company, and clear ICP rules. At that point, the agent needs direct answers it can act on. A paragraph it must interpret does not help.

Do not use Jev for open-ended work such as writing emails, researching companies, or planning campaigns. Leave those tasks with the agent.

Three parts of my setup remain unproven:

  • I have only tested the reply rules on replies I wrote.
    ‍
  • Job-change data still requires manual checks.
    ‍
  • I have not confirmed TypeSafe’s pricing.

I would compare all three with the extra qualified leads and meetings they generate before letting the agent work independently.

Personalized Outbound Strategy

Get The Right Outbound Strategy In Minutes

Enter your email to get a custom plan & stack recommendation for your business

It's being carefully crafted by AI

Please check your mailbox in 5 minutes

Summary

You do not have to trust an email agent all at once. I assigned every decision an owner. The agent researches and writes. Jev answers narrow questions using evidence. Fixed rules cover replies that cannot be mishandled. I still control launches and unclear decisions.

Every step stays with me until it performs reliably in testing. Once it does, the agent earns a little more freedom.

Book more meetings on autopilot

Set up in minutes. No per-seat pricing. No commitment.
Try
free
4.6 rating on G2
Add as a preferred
source on Google