heya Contact the team

Heya Labs.

Where the Heya team ships experiments in Australian AI.

Talk to the team
Research below

Research

Papers on conversational analysis, verifiable rewards, and voice.

A working record of the research underneath the product. Each paper ships with a public reproduction kit so the numbers can be checked, not just cited.

20 Jul 2026 · Research paper · 30 pp · 845 KB

Analysis of Conversation Trajectory Representations

Jacob Sussmilch · Heya Enterprises Pty Ltd

The first public-data benchmark of TRACE's hand-crafted conversational geometry, measured against a low-order orthogonal projection of the same turn-embedding trajectory. The projection leads by 5–12 points across three satisfaction corpora, and splitting the trajectory by speaker — a per-role mean plus drift, no training and no feature selection — is never distinguishably worse than any measured alternative in either regime. The boundary is reported as a result: on short, first-person-rated chats trajectory shape collapses and curated scalars become a genuine complement.

Trajectory geometry TRACE benchmark USS-SGD · USS-MWOZ · PRISM Reproduction kit
Read PDF
20 Jul 2026 · Method paper · 17 pp · 347 KB

Gram Projections for Conversation Trajectories

Jacob Sussmilch · Heya Enterprises Pty Ltd

Introduces the Gram (discrete orthogonal polynomial) projection as a polynomial equivalent of the DCT for turning short, aperiodic embedding sequences into a fixed-size vector. The coefficients have stable semantics — degree 0 is the mean, degree 1 is the least-squares drift, and the drift coefficient reports the same end-to-end drift at every sequence length. On three conversation-satisfaction corpora the bases tie almost everywhere; the exception favours the polynomial basis on populations dominated by very short per-role sequences, by a measured 2.3 points.

Orthogonal polynomials Embedding pooling USS-SGD · USS-MWOZ · PRISM Reproduction kit
Read PDF
23 Jul 2026 · Research paper · 13 pp · 294 KB

Manipulability of Trajectory-Geometric Conversation Rewards

Jacob Sussmilch · Heya Enterprises Pty Ltd

A dense reward that scores a conversation on every turn is what reinforcement learning wants where the outcome is sparse and late — and what Goodhart's law warns will be gamed. We show it need not be. Exploitability is a property of which role the adversary can write, not of content. Score the fixed evidence — user turns, tool calls, tool results — admit the model's reasoning and response only as its consistency with that evidence, and condition on the task: the construction recovers 97–99% of the predictor's accuracy while remaining un-gameable by the selection adversary that breaks the naive proxy.

Reward hacking Verifiable rewards ToolBench · ABCD · SpokenWOZ Pre-registered Reproduction kit
Read PDF
Reproduction. Each research paper ships as a self-contained kit on GitHub — a portable pipeline spec plus the generator scripts, regenerating every artifact from public data with one command. Details in Appendix A of each paper.
© Heya Enterprises, 2026