by Nodical
Research · 2027-2028

ASDIC. The model that listens to recognise.

On the phone, everything happens in seconds. Off-the-shelf models were built for read speech, in English. We are building our own, for phone conversation, in French, Spanish and English, trained and hosted in Europe.

In development · research project 2027-2028

Pulse sentEcho: a humanEcho: voicemail or noise
Today
Target
  1. T+0.0 sanswered
    T+0.0 sanswered
  2. T+0.4 s“Hello?”
    T+0.4 s“Hello?”
  3. T+8 svoicemail recognised (median on our campaigns)
    T+2 svoicemail recognised, under 3% error
  4. T+1.3 sresponse delay, 9 times out of 10
    T+0.5 sresponse delay, without cutting in
  5. T+135 shandover to an adviser (median)
    right momenthandover with a two-line summary

“Today” values are measured on our September 2026 calls. Targets are targets: they will be published with the measurements, not before.

The problem

Half of answered calls last ten seconds or less.

Speech recognition models learned from read speech, in wide band, mostly in English. On the phone the sound is narrow, people cut each other off, and a decision must land in under a second. Every second of understanding lost costs a contact.

60,287calls answered by a human, September 2026
10 smedian length of an answered call
8 in 10calls shorter than twenty seconds
8 sbefore a voicemail is recognised today

Read from Nodical's production database, 14 campaigns, without reading the content of conversations.

What it does

Five decisions, during the call.

  1. from the pick-up

    Human or voicemail

    ASDIC says who the agent is talking to, with a confidence score. Today the decision lands after eight seconds (median). The target: under two seconds.

  2. while they speak

    What the person wants

    Interested, not interested, call back later, wrong number, asks for a human, does not want to be contacted again: the intent is recognised live, along with the answers to eligibility questions. Matched against the sales the customer actually closed, these signals refine each contact's chances of winning: the team knows who to focus on.

  3. every turn

    When they have finished speaking

    Neither cutting them off nor leaving a gap. Waiting for silence is not enough: it is the weak point of every voice agent, and the heart of our work.

  4. at the right moment

    Handover to a human

    When an adviser should take over, ASDIC says so, with a two-line summary so nobody arrives cold.

  5. everywhere

    Three languages

    French, Spanish, English: the same model, the same performance, without rebuilding one model per language.

Starting point

What we measure, and what we aim for.

ChallengeToday (measured)Target
Human or voicemaildecision at 8 s (median), 14 s for 9 calls in 10under 2 s, under 3% error
Response delay0.14 s median, 1.32 s 9 times out of 10, excluding end-of-speech detectionunder 0.5 s, cutting in less than 5% of the time
Handover1.19% of answered calls, at 135 s (median)more than 9 handovers in 10 judged relevant
Intent and eligibilityno measurement without an annotated corpusset after the baseline
Chances of winningcomputed in Nodical since October 2026 from the call outcome and duration, based on each customer's sales; no published measurement yetrank better who signs, using signals from the conversation; target set after a one-month pilot
Sovereignty0% own model100% for the understanding model

Targets proposed from the September 2026 measurements and the state of the art; they will be validated with the team we hire.

The name

Why ASDIC.

19172027Europe
1917

The first machine that listened to recognise

ASDIC is the ancestor of sonar, developed during the First World War with the decisive contribution of French physicist Paul Langevin. It sends a sound, listens to the echo and decides within seconds whether it hears a real target or just noise.

2027

The same job, on a conversation

Our model does the same thing on the phone: it listens, tells a human from a voicemail at once, spots a real intent in the middle of the conversation, and raises its hand at the right moment.

Europe

A technology we own

The model, its weights and its corpus belong to Nodical. It is trained on European computing power and hosted in the European Union. No single vendor can switch it off.

What exists

What published measurements say.

×5 to ×6

The open reference model for speech recognition gets 2.7% of words wrong in read speech, but 13.8 to 17.6% in telephone conversation.

OpenAI, Whisper paper (table 8)
≈ 4 s

The average time Twilio's answering-machine detector announces with default settings, “close to 100%” accuracy in the United States, lower internationally. No figure published for French or Spanish voicemails.

Twilio, AMD documentation (FAQ)
55.6%

How often you cut someone off by answering after 300 ms of silence. To get down to 5%, you have to wait 1.6 seconds.

LiveKit, eot-bench benchmark

No public reference measurement exists for French telephone speech. Building that corpus is part of the research.

Programme

Twenty-four months, six steps.

  1. Months 1 to 8

    Data

    A corpus of pseudonymised calls, annotated in French, then Spanish and English. None exists in open access for French telephone speech.

  2. Months 2 to 4

    Baseline

    Off-the-shelf models measured on this corpus: the quantified starting point.

  3. Months 4 to 14

    French model

    Design and training of the first four functions.

  4. Months 12 to 20

    Multilingual

    Spanish and English, accents included.

  5. Months 10 to 22

    Real time

    Integration into Nodical and trials in real conditions, on live campaigns.

  6. Months 20 to 24

    Evaluation

    Final measurements, published; protection of the model; roll-out.

With the support of a public research laboratory in the Paris region specialised in speech processing, currently being selected. Training on European computing power.

What it does not do

Three lines we will not cross.

No voice imitated

ASDIC listens and understands. It never produces a voice and never learns to imitate one.

No decision about the person

It helps the agent lead the conversation, and the customer's team choose the order of its callbacks. It makes no decision imposed on the person called, who never sees its result.

No contact details in the data

Calls used for training are pseudonymised: names, numbers and addresses are removed before any annotation. Anyone can object by a simple email.

Who it is for

Nodical first. Then any software that has to understand a call.

Inside Nodical

The agent's ears and judgement

ASDIC replaces the “listen and judge” part of Nodical's voice agents. The voice is still produced by speech synthesis.

As an API

For other software

Phone reception, healthcare, public services, emergency lines: any system that receives or places calls in French, Spanish or English.

In Europe

Without depending on a vendor

A model hosted in the European Union, with its weights and its corpus, for organisations that cannot send their conversations elsewhere.

Team

We are hiring two speech and language AI engineers.

The project is led by Douglas Demart, president of Nodical. Two engineers join the team in Paris: the first in January 2027 (French model), the second in 2027 (multilingual and real time). A real corpus of more than 60,000 calls a month, European computing, a partner laboratory, the possibility to publish.

  1. January 2027 · Paris

    AI engineer — speech and language

  2. 2027 · Paris

    AI engineer — multilingual and real time

Write to apply

You have calls in French, Spanish or English and want to follow the project, or contribute to it?

An email is enough. We answer ourselves, and we publish nothing that has not been measured.