2026-10-06

How we made Partyline the fastest AI assistant on earth

Friends cheering as a model rocket launches off a dry lakebed at golden hour
Launch day, dry lakebedTry it on Partyline

AI assistants live or die by speed. An assistant has to get things done for you, but on iMessage it has very few ways to show you that it is working on your request: a tapback, a typing bubble, a short reply.

Partyline is a special kind of AI assistant in your iMessage. We're not focused on triaging your email: our opinion is that as a human you shouldn't have been on email in the first place. Our goal is to get you outside, with other people, having a good time. Our product metric is the amount of life our users have that's worth living.

For an assistant, however, that means competing in a very, very busy world that is doing its best to tire out our users. Make them depressed and unhappy. But most of all, compete for their attention.

Because people text from chaotic environments, we have to be fast. We only get a moment with our user before they put their phone away (as they should!), and we don't want them looking at the screen for a moment longer than they have to.

When we built Partyline, our replies were too slow. Finding a place or an event took 29 to 46 seconds. That is a long time when you are out with friends. If you are the one who pulls out your phone to find the next spot, you are putting your social capital on the line: you are the person who knows something cool is going on. At 46 seconds, someone else has already suggested a place and taken over the conversation.

How we compare

We compared Partyline with two other personal assistants on iMessage, Tomo and Instinct. Instinct was the slowest: its answers to searches took 42 to 82 seconds, which is too slow to use while you are out with friends. Tomo is excellent at conversation. Its chat replies arrive within a few seconds, and its search answers take 28 to 37 seconds, with a tapback at about 2 seconds and a first text at about 5.

Time to a full answer
PartylineTomoInstinct
0 s10 s20 s30 s40 s50 s60 s70 s80 s90 sPartyline before29–46 sInstinct42–82 sTomo28–37 sPartyline now2–3 s
Seconds from sending a place or event search to the last message of the answer. Each bar spans the range we measured; the dot is a typical run. Partyline before: production, October 5. Partyline now: production, October 6. Tomo and Instinct: the same kinds of questions on our test phone, October 5–6.
AssistantFirst sign it's workingFull search answer
Partyline now🔍 tapback at about 1 s2–3 s
TomoTapback at about 2 s, first text at about 5 s28–37 s
Partyline before🔍 and a first line at 1–2 s29–46 s
Instinct👀 at 4–6 s42–82 s

From 46 seconds to 3

Tomo set the question for us: how could Partyline reply as quickly as Tomo does in conversation, but for bigger tasks like finding places and events? We started at 29 to 46 seconds. Today a search takes 2 to 3 seconds from send to reply.

What arrives when
FIRST 10 SECONDSWHOLE EXCHANGE, 0–60 S02468100102030405060Partyline now🔍 · 1 spicks · 3 sTomo💭 · 2 sframing line · 5 squestion · 37 sPartyline before🔍 + first line · 2 sanswer · 42 sInstinct👀 · 4 sanswer · 42 slink cards · 49 s
One real exchange per assistant, timed from the thread. Tomo: West Village bars. Instinct: a jazz bar near SoHo. Partyline before: “what’s happening in SF tonight?” Partyline now: steak near Union Square.

Two changes made most of the difference.

First, we index events in advance for the places people ask about most. During a16z's Tech Week in San Francisco our inventory holds about 1,700 events for the week, and a search narrows them to a short list in under a tenth of a second. Where an area isn't indexed, a web search runs alongside and answers a few seconds later.

Second, we put Jev classifiers at several steps of our message pipeline. Jev is TypeSafe's classifier model. In one call of about 0.2 seconds it answers many questions about a message: what kind of request it is, whether it continues the last search, whether it is sensitive. Jev also ranks the candidate events and places, and checks each line of a reply against the data it describes. When a large model made those decisions, each one took about 4 seconds.

Where the time went
OCT 5 · FULL AGENT TURN · 46 SClassify, 🔍, first line1 sWaiting for the agent turn4 sModel step: search the inventory?4 sInventory search: Jev ranks 1,749 events20 sModel step: search the web?5 sWeb search and page reads7 sModel step: write the reply4 sSend1 sOCT 6 · FAST LANE · 2 SJev classifies0.20 sQwen extracts the search0.50 sSQL short list: 1,667 → 600.05 sJev ranks 60 events0.25 sQwen picks, Jev checks lines0.40 sSend and deliver0.60 s0 s5 s10 s15 s20 s25 s30 s35 s40 s45 s
Top: the production turn for “any AI events in SF this week?” on October 5 (46 seconds). Bottom: the same question on October 6 through the fast path: Jev classifies, Qwen extracts the search, SQL narrows 1,667 events to 60, Jev ranks them, Qwen picks three, Jev checks each line, and the answer is sent.

What's next

Partyline now finishes a search answer before Tomo sends even its first text, and it can hold many rounds of casual conversation before Instinct replies once. The downside is that some answers read less like a conversation and more like a fast response. We're working on it. Our next job is to tune the system toward a middle ground that still feels near-instant but reads like a conversation.

We're also building tooling to keep inference costs affordable as we grow, so we don't have to trade speed or quality for cost.

Today, Partyline is the fastest AI iMessage assistant in the world, especially for event search. If you are at a16z's Tech Week in San Francisco this week, text Partyline for the fastest event search you have seen. Try it on Partyline

Questions and answers

How fast is Partyline?

Partyline answers an event or place search on iMessage in about 2 to 3 seconds from send to reply, and a quick question in about a second. In production, event searches took a median of 1.85 seconds from the message being sent to the answer being delivered.

What is Jev?

Jev is TypeSafe's classifier model. In one call of about 0.2 seconds it answers many typed questions about a message. Partyline uses it to route each message, rank candidate events and places, and check each line of a reply against the data it describes.

How does Partyline compare with Tomo and Instinct?

In our tests on October 5 and 6, 2026, a full search answer took 2 to 3 seconds on Partyline, 28 to 37 seconds on Tomo and 42 to 82 seconds on Instinct.

Where does Partyline find events?

From its own event inventory, indexed in advance (about 1,700 events for San Francisco's Tech Week), and from a live web search that runs alongside it.

Do I need an app?

No. You text Partyline on iMessage. There is no smartphone app to download.

How we measured

Partyline's times come from our production logs and test phone on October 5 and 6, 2026. Tomo's and Instinct's come from our own threads with them on the same days. Each figure is a handful of runs, not a benchmark. Logos belong to their companies.

Related: Going out and meeting people: how the assistants compare.

Text to get started.

Text Partyline to get started